@MiniMax_AI: Congrats to our long-term partner SGLang/RadixArk on the launch of Miles v0.1! From M-Series to H3 and Music 3, we’ve b…

X AI KOLs Timeline Tools

Summary

RadixArk launches Miles v0.1, an open-source reinforcement learning framework for large language models and multimodal models, aimed at simplifying and scaling RL training.

Congrats to our long-term partner SGLang/RadixArk on the launch of Miles v0.1! From M-Series to H3 and Music 3, we’ve been closely building and pushing the OSS ecosystem forward together. Nothing beats the feeling of seeing things we built together keep growing @radixark @MiniMax_AI
Original Article
View Cached Full Text

Cached at: 08/18/26, 08:33 PM

Congrats to our long-term partner SGLang/RadixArk on the launch of Miles v0.1!

From M-Series to H3 and Music 3, we’ve been closely building and pushing the OSS ecosystem forward together.

Nothing beats the feeling of seeing things we built together keep growing @radixark @MiniMax_AI

RadixArk (@radixark): Today we’re launching Miles v0.1, an open-source RL framework for LLMs and multimodal models.

RL training is easy to start and hard to debug. Miles helps you ensure your run is correct, use hardware efficiently, and keep RL running at scale.

Over the past 9 months, 72

Similar Articles

The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence

Hugging Face Daily Papers

The MiniMax-M2 series introduces Mixture-of-Experts language models that achieve high performance on agentic tasks with minimal activated parameters (9.8B per token out of 229.9B total), leveraging agent-driven data pipelines, a scalable RL system called Forge, and a checkpoint that takes early steps toward self-evolution.