ByteDance is at an early stage of training a model with as many as 10 trillion parameters
Summary
ByteDance is in the early stages of training a large language model with up to 10 trillion parameters, signaling a massive scale-up in AI development.
Similar Articles
ByteDance trains a 10-trillion-parameter AI model, aiming for global leadership (1 minute read)
据金融时报报道,字节跳动正在训练一个估算参数量达10万亿级别的AI大模型,规模接近Anthropic的先进系统,旨在缩小与美国顶级AI实验室的差距。
ByteDance trains massive AI model in bid to rival Anthropic
ByteDance is training a massive AI model, aiming to rival Anthropic, with an independent development approach led by its Seed team under former Google DeepMind scientist Wu Yonghui, while its Doubao model leads China with 324 million MAU.
ByteDance vows to avoid AI distillation, develop new model its own way
ByteDance vows to develop its AI models independently, avoiding the practice of distillation from other models.
A 4b model is now beating 30b ones at web research and the reason is not size
A 4 billion parameter open model from the Apodex family outperforms 30 billion parameter models on web research benchmarks, attributed to careful training data and self-verification techniques rather than raw scale, suggesting a more democratic trajectory for AI capability.
Minimax plans to release a 2.7-trillion parameter model.
Minimax is planning to release a massive 2.7-trillion parameter AI model, which would be one of the largest ever.