MiniMax

최신 AI 뉴스, 모델 및 출시 정보 MiniMax. ['Abigail', 'Hailuo Minimax', 'M2.1', 'M2.5', 'MiniMax', 'MiniMax-M1', 'MiniMax-M2', 'MiniMax M2.1', 'MiniMax M2.5', 'MiniMax M2.7', 'MiniMax-M2.7', 'MiniMax M3', 'MiniMax-W']

연구 🇺🇸

Beyond Standard LLMs: Linear Attention Hybrids, Text Diffusion Models, Code World Models, and Small Recurrent Transformers

The article explores alternatives to standard autoregressive transformer-based LLMs: linear attention hybrids (MiniMax-M1, Qwen3-Next, DeepSeek V3.2, Kimi Linear), text diffusion models, code world models, and small recurrent transformers. The author notes a return to classical attention in MiniMax-M2 and the complexity of linear attention in production.

Moonshot AIMoonshot AI MiniMaxMiniMax Alibaba/QwenAlibaba/Qwen DeepSeekDeepSeek Google/DeepMindGoogle/DeepMind MistralMistral MetaMeta Hugging FaceHugging Face Allen Institute for AIAllen Institute for AI xAIxAI OpenAIOpenAI IBMIBM NVIDIANVIDIA
Sebastian Raschka27.07 · 17:04
연구 🇺🇸

Survey of LLM Research Papers for the First Half of 2026

Sebastian Raschka published a curated list of key scientific papers on large language models (LLMs) from January to May 2026. The selection highlights trends: hybrid architectures (alternating attention and state-space layers), efficient inference, agentic systems, and diffusion language models. Special attention is given to Nvidia's Nemotron 3 with hybrid design and Qwen3.6 with Gated DeltaNet.

NVIDIANVIDIA Alibaba/QwenAlibaba/Qwen MistralMistral BaiduBaidu CohereCohere Technology Innovation InstituteTechnology Innovation Institute MiniMaxMiniMax
Sebastian Raschka27.07 · 12:03
새로운 뉴스