1m-context

Tag

Cards List
#1m-context

First evidence of a pending qwen3.7 open weights release. Qwen3.7-flash is on open router. They referred to Qwen3.6-35b-a3b as Qwen3.6 flash so this is likely a small MoE. The prices are substantially cheaper than 3.6 flash with a native 1M context window.

Reddit r/LocalLLaMA · 20h ago Cached

Evidence of an upcoming Qwen3.7 open weights release, with a flash variant (likely small MoE) listed on OpenRouter featuring a 1M context window and cheaper pricing than Qwen3.6 flash.

0 favorites 0 likes
#1m-context

Laguna S 2.1 Released: Cheaper than Deepseek v4 Flash, Better than V4 Pro

Reddit r/LocalLLaMA · 2026-07-21 Cached

Poolside releases Laguna S 2.1, a 118B MoE model with 8B activated parameters per token, optimized for agentic coding. It claims to outperform DeepSeek V4 Pro while being cheaper than DeepSeek V4 Flash, with a 1M context window and open-source license.

0 favorites 0 likes
#1m-context

Kimi K3 released on web and app

Reddit r/LocalLLaMA · 2026-07-16

Kimi K3, a new AI model with 2.8 trillion parameters and 1 million context length, has been released on web and app, featuring leading capabilities in coding, agentic tasks, reasoning, vision, and agent swarms.

0 favorites 0 likes
#1m-context

@yifanzhang_: https://thinkingmachines.ai/news/introducing-inkling/… RoPE is dead, Long live GRAPE!

X AI KOLs Timeline · 2026-07-15 Cached

Thinking Machines AI releases Inkling, an open-weights Mixture-of-Experts model with 975B total parameters (41B active), supporting text, images, and audio over a 1M token context window. It is a broad foundation model designed for fine-tuning via their Tinker platform, with a smaller Inkling-Small variant also previewed.

0 favorites 0 likes
#1m-context

@seclink: MiniMax M3 is now open source. Generally, those willing to open source are outdated, with undisclosed firepower still unreleased. Hugging Face main repository (recommended): https://huggingface.co/MiniMaxAI/MiniMax-M3… Here provides the complete model weights...

X AI KOLs Timeline · 2026-07-14 Cached

MiniMax has open-sourced its native multimodal large model M3, with approximately 428B total parameters (~23B active), supporting 1M context length, and introducing MiniMax Sparse Attention (MSA) technology to improve long-context efficiency.

0 favorites 0 likes
#1m-context

@seclink: LongCat-2.0 Released & New Billing Service Launched ​ Core features of LongCat-2.0: Trillion parameters, 1M ultra-long context: native tool calling and multi-step reasoning, reliably handling long-context Agent tasks. Outstanding coding capabilities: in code generation, code understanding, and automatic...

X AI KOLs Following · 2026-07-01 Cached

LongCat-2.0 model released with trillion parameters and 1M ultra-long context, supporting native tool calling and multi-step reasoning, with outstanding coding capabilities. Also introduces Token resource packs and pay-as-you-go API billing service.

0 favorites 0 likes
#1m-context

@uzairansar: Qwythos-9B-Claude-Mythos-5 Fine Tune with 1M Context released! Empero just released their Claude Mythos Fine Tune based…

X AI KOLs Timeline · 2026-06-21 Cached

Empero released Qwythos-9B-Claude-Mythos-5, a full-parameter reasoning model fine-tuned with 1M context, based on synthetic chain-of-thought data from Fable-5 and Mythos-5 session logs.

0 favorites 0 likes
#1m-context

GLM-5.2: Built for Long-Horizon Tasks

Hugging Face Blog · 2026-06-17 Cached

Z.AI introduces GLM-5.2, a flagship model designed for long-horizon tasks with a solid 1M-token context, improved coding capabilities, and an MIT open-source license, showing competitive performance against leading models like Opus 4.8 and GPT-5.5.

0 favorites 0 likes
#1m-context

GLM-5.2 just dropped open weights and it already looks weirdly strong for coding

Reddit r/LocalLLaMA · 2026-06-16

GLM-5.2 has been released with open weights under MIT license, featuring a 1M context window and two reasoning effort modes. Early benchmarks show it performing strongly in coding tasks, making it worth testing beyond benchmark screenshots.

0 favorites 0 likes
#1m-context

@AdinaYakup: GLM 5.2 is here 753B ( smaller than you expect? ) 1M context MIT license GLM IndexShare: reuses the indexer across laye…

X AI KOLs Following · 2026-06-16 Cached

GLM 5.2 is released as a 753B parameter open-source model with 1M context length, MIT license, and achieves 99.2 on AIME 2026, outperforming GPT-5.5, Gemini 3.1 Pro, and Claude Opus 4.8.

0 favorites 0 likes
#1m-context

GLM-5.2 (1 minute read)

TLDR AI · 2026-06-15 Cached

GLM-5.2, a new flagship coding model with 1M-context support and enhanced reasoning, is now available to GLM Coding Plan users and will be open-sourced under MIT License next week.

0 favorites 0 likes
#1m-context

@Modular: Our kernel team has been deep in MiniMax M3 all week. The 1M-token context and native multimodality make it a hard mode…

X AI KOLs Following · 2026-06-09 Cached

Modular's kernel team is optimizing serving for MiniMax M3's 1M-token context and native multimodality, with open weights dropping soon for immediate deployment on Modular.

0 favorites 0 likes
#1m-context

MiniMax promises M3 weights after 1M-context model launch (2 minute read)

TLDR AI · 2026-06-03 Cached

MiniMax released M3, a model with a 1M-token context window and native multimodal input, via API. The company promises open-weight release and a technical report within 10 days.

0 favorites 0 likes
#1m-context

MiniMax M3 - Coding & Agentic Frontier, 1M Context, Multimodal

Reddit r/LocalLLaMA · 2026-06-01 Cached

MiniMax releases M3, an open-weight model with frontier coding, agentic, 1M context, and native multimodal capabilities, achieving top benchmarks on coding and agentic tasks with autonomous task decomposition and long-context support.

0 favorites 0 likes
← Back to home

Submit Feedback