@Meituan_LongCat: Introducing LongCat-2.0 1.6T parameters · MoE with ~48B active · 1M context The full model behind Owl Alpha on @OpenRou…
Summary
Meituan introduces LongCat-2.0, a 1.6T parameter MoE model with ~48B active parameters and 1M context, featuring novel architectures like LongCat Sparse Attention and Zero-Compute Experts, achieving strong benchmark scores on coding and reasoning tasks.
View Cached Full Text
Cached at: 06/30/26, 03:34 AM
Introducing LongCat-2.0 1.6T parameters · MoE with ~48B active · 1M context The full model behind Owl Alpha on @OpenRouter — now available.
Built for agentic coding from the ground up: ◆ LongCat Sparse Attention (LSA) — scales efficiently for 1M-context tokens ◆ Zero-Compute Experts — dynamic activation 33B–56B per token, zero wasted compute ◆ MOPD — three specialized expert groups (Agent / Reasoning / Interaction), gate-routed per task
How it stacks up: → Terminal-Bench 2.1: 70.8 → SWE-bench Pro: 59.5 (GPT-5.5: 58.6) → SWE-bench Multilingual: 77.3 → FORTE: 73.2 · RWSearch: 78.8 · BrowseComp: 79.9
Tech Blog: https://longcat.chat/blog/longcat-2.0/… Try it across different scenarios
Similar Articles
meituan-longcat/LongCat-2.0
LongCat-2.0 is a large-scale MoE language model with 1.6 trillion total parameters and ~48B activated per token, trained on AI ASIC superpods with 1M-context data. It achieves strong performance on coding and agentic tasks.
Meituan launches LongCat-2.0 1.6T parameter model on APIs (2 minute read)
Meituan launched LongCat-2.0, a 1.6 trillion-parameter Mixture-of-Experts model with a 1 million-token context window, available via API for agentic coding, tool use, and complex workflows.
LongCat-2.0, a large-scale MoE model with 1.6T total and 48B Active
LongCat-2.0 is a large-scale Mixture-of-Experts (MoE) model with 1.6 trillion total parameters and 48 billion active parameters.
@eliebakouch: the new sparse attention method introduced with this model is basically a combination of components from existing ones.…
Meituan introduces LongCat-2.0, a 1.6T parameter MoE model with 48B active parameters and 1M context length, featuring a new LongCat Sparse Attention (LSA) method that combines components from existing sparse attention techniques.
@sheriyuo: The industry's first trillion-parameter model to complete end-to-end training and inference on a 50,000-GPU Chinese com…
Meituan released LongCat-2.0, a 1.6T-parameter MoE model with 1M context, claimed as the first to train on a 50,000-GPU Chinese cluster, now available on OpenRouter for agentic coding.