Tag
Evidence of an upcoming Qwen3.7 open weights release, with a flash variant (likely small MoE) listed on OpenRouter featuring a 1M context window and cheaper pricing than Qwen3.6 flash.
Poolside releases Laguna S 2.1, a 118B MoE model with 8B activated parameters per token, optimized for agentic coding. It claims to outperform DeepSeek V4 Pro while being cheaper than DeepSeek V4 Flash, with a 1M context window and open-source license.
Kimi K3, a new AI model with 2.8 trillion parameters and 1 million context length, has been released on web and app, featuring leading capabilities in coding, agentic tasks, reasoning, vision, and agent swarms.
Thinking Machines AI releases Inkling, an open-weights Mixture-of-Experts model with 975B total parameters (41B active), supporting text, images, and audio over a 1M token context window. It is a broad foundation model designed for fine-tuning via their Tinker platform, with a smaller Inkling-Small variant also previewed.
MiniMax has open-sourced its native multimodal large model M3, with approximately 428B total parameters (~23B active), supporting 1M context length, and introducing MiniMax Sparse Attention (MSA) technology to improve long-context efficiency.
LongCat-2.0 model released with trillion parameters and 1M ultra-long context, supporting native tool calling and multi-step reasoning, with outstanding coding capabilities. Also introduces Token resource packs and pay-as-you-go API billing service.
Empero released Qwythos-9B-Claude-Mythos-5, a full-parameter reasoning model fine-tuned with 1M context, based on synthetic chain-of-thought data from Fable-5 and Mythos-5 session logs.
Z.AI introduces GLM-5.2, a flagship model designed for long-horizon tasks with a solid 1M-token context, improved coding capabilities, and an MIT open-source license, showing competitive performance against leading models like Opus 4.8 and GPT-5.5.
GLM-5.2 has been released with open weights under MIT license, featuring a 1M context window and two reasoning effort modes. Early benchmarks show it performing strongly in coding tasks, making it worth testing beyond benchmark screenshots.
GLM 5.2 is released as a 753B parameter open-source model with 1M context length, MIT license, and achieves 99.2 on AIME 2026, outperforming GPT-5.5, Gemini 3.1 Pro, and Claude Opus 4.8.
GLM-5.2, a new flagship coding model with 1M-context support and enhanced reasoning, is now available to GLM Coding Plan users and will be open-sourced under MIT License next week.
Modular's kernel team is optimizing serving for MiniMax M3's 1M-token context and native multimodality, with open weights dropping soon for immediate deployment on Modular.
MiniMax released M3, a model with a 1M-token context window and native multimodal input, via API. The company promises open-weight release and a technical report within 10 days.
MiniMax releases M3, an open-weight model with frontier coding, agentic, 1M context, and native multimodal capabilities, achieving top benchmarks on coding and agentic tasks with autonomous task decomposition and long-context support.