@LiTianleli: Incredibly proud of the team. After countless late nights, Inkling is out, and I especially want to highlight the post-…
Summary
Thinking Machines releases Inkling, an open-source multi-modal reasoning model with innovations in post-training RL, achieving stable scaling to 30M+ rollouts and controllable thinking effort. The model exhibits compressed reasoning and will be available soon for fine-tuning.
View Cached Full Text
Cached at: 07/16/26, 04:21 PM
Incredibly proud of the team. After countless late nights, Inkling is out, and I especially want to highlight the post-training stack and RL recipes behind it.
A few of my favorite details:
We scaled our largest RL run to 30M+ rollouts and thousands of continuous training steps—with no collapse, no restarts, and stable KL and entropy throughout. Reasoning performance improved log-linearly from the SFT initialization all the way to the released checkpoint. A number of innovations under the hood made this possible, and the result is a strong testament to our post-training technology.
We trained controllable thinking effort directly through RL. By varying the system message and per-token cost, the model learned to trade off tokens and performance on demand.
We also saw an emergent shift in reasoning style: as RL progressed, the chain of thought became increasingly compressed, shedding grammatical overhead. Inkling reasons like a caveman mathematician—a distinctive style unlike that of other open-source models.
We’re also previewing Inkling-small today and plan to release it very soon. It is exceptionally capable for its size, and we expect the community will find it broadly useful.
Building a simple, stable, and scalable RL stack in such a short time was something few thought possible. This team proved otherwise.
Frrrr
Thank you Clare!
Miss you
Similar Articles
@levie: Fantastic to see more open weights innovation happening right now, especially coming from a US Lab. The future of AI is…
Thinking Machines released Inkling, a multimodal AI model with open weights, capable of reasoning across text, image, and audio modalities. The model is available for fine-tuning on Tinker and via the Inkling Playground.
@LiorOnAI: New model from Thinking Machines: - Full weights available - Native text, image, and audio reasoning - 975B total param…
Thinking Machines releases Inkling, an MoE model with 975B total / 41B active parameters, supporting native text, image, and audio reasoning, up to 1M-token context, and full weights availability.
Thinking Machines Lab Drops Its First Model
Thinking Machines Lab, founded by ex-OpenAI executives, releases its first open-weight AI model, Inkling, a 975-billion-parameter model capable of reasoning, coding, and processing audio, video, and text.
The Benchmarks of Thinking Machine's first open-source model Inkling
Thinking Machine has released its first open-source model, Inkling, alongside benchmark results demonstrating its performance.
@yifanzhang_: https://thinkingmachines.ai/news/introducing-inkling/… RoPE is dead, Long live GRAPE!
Thinking Machines AI releases Inkling, an open-weights Mixture-of-Experts model with 975B total parameters (41B active), supporting text, images, and audio over a 1M token context window. It is a broad foundation model designed for fine-tuning via their Tinker platform, with a smaller Inkling-Small variant also previewed.