meituan-longcat/LongCat-2.0
Summary
LongCat-2.0 is a large-scale MoE language model with 1.6 trillion total parameters and ~48B activated per token, trained on AI ASIC superpods with 1M-context data. It achieves strong performance on coding and agentic tasks.
View Cached Full Text
Cached at: 06/30/26, 11:30 PM
meituan-longcat/LongCat-2.0 · Hugging Face
Source: https://huggingface.co/meituan-longcat/LongCat-2.0
https://huggingface.co/meituan-longcat/LongCat-2.0#model-introductionModel Introduction
We introduce LongCat-2.0, a large-scale MoE language model with1.6 trillion total parametersand ~48 billion activated per token — a substantial step up from previous LongCat models, accompanied by several architectural improvements.
Both the full training run and the large-scale deployment are built entirely onAI ASIC superpods. Pretraining spans millions of accelerator-hours across more than 35 trillion tokens, with no rollbacks or irrecoverable loss spikes — demonstrating that we have the capability to conduct frontier-scale training on alternative hardware platforms.
To strengthen the model on long-horizon tasks, we introduce LongCat Sparse Attention and train LongCat-2.0 on hundreds of billions of tokens of1M-contextdata. Together with dedicated post-training, this gives LongCat-2.0 strong performance on coding and agentic tasks.
🏋️Model weights coming soon— stay tuned!
Similar Articles
@Meituan_LongCat: Introducing LongCat-2.0 1.6T parameters · MoE with ~48B active · 1M context The full model behind Owl Alpha on @OpenRou…
Meituan introduces LongCat-2.0, a 1.6T parameter MoE model with ~48B active parameters and 1M context, featuring novel architectures like LongCat Sparse Attention and Zero-Compute Experts, achieving strong benchmark scores on coding and reasoning tasks.
LongCat-2.0, a large-scale MoE model with 1.6T total and 48B Active
LongCat-2.0 is a large-scale Mixture-of-Experts (MoE) model with 1.6 trillion total parameters and 48 billion active parameters.
Meituan launches LongCat-2.0 1.6T parameter model on APIs (2 minute read)
Meituan launched LongCat-2.0, a 1.6 trillion-parameter Mixture-of-Experts model with a 1 million-token context window, available via API for agentic coding, tool use, and complex workflows.
LongCat-2.0
LongCat-2.0 is a 1.6 trillion parameter mixture-of-experts model trained entirely on custom AI ASICs.
@sheriyuo: The industry's first trillion-parameter model to complete end-to-end training and inference on a 50,000-GPU Chinese com…
Meituan released LongCat-2.0, a 1.6T-parameter MoE model with 1M context, claimed as the first to train on a 50,000-GPU Chinese cluster, now available on OpenRouter for agentic coding.