Inkling
Summary
Inkling is an open-weights 975B multimodal AI model designed for fine-tuning.
Similar Articles
Inkling: Our Open-Weights Model
Thinking Machines AI releases Inkling, a new open-weights mixture-of-experts multimodal foundation model with 975B total parameters and 41B active, supporting text, images, audio, and video, along with a preview of Inkling-Small.
thinkingmachines/Inkling
Inkling is a large open-weights multimodal model (975B total, 41B active parameters) using a sparse MoE architecture, accepting text, image, and audio inputs and generating text outputs, intended for agentic systems, coding assistants, and chatbots.
Welcome Inkling by Thinking Machines
Inkling by Thinking Machines is a large open multimodal LLM with ~1T parameters, 1M context, and native support for image, audio, and text. It uses a Mixture-of-Experts architecture and is available on Hugging Face with day-0 inference support.
thinkingmachines/Inkling-NVFP4
Inkling is a 975B-parameter sparse mixture-of-experts multimodal model accepting text, image and audio inputs and generating text outputs. Released with open weights for research, fine-tuning, and integration.
Inkling-Small (4 minute read)
Thinking Machines released Inkling-Small, an efficient open-weights Mixture-of-Experts model with 276B total and 12B active parameters, achieving comparable performance to its larger sibling Inkling at a quarter of the size. It features native reasoning over audio and images, variable thinking effort, and a 1M-token context window.