@AYi_AInotes: Everyone is raving about Japan's Fugu beating GPT on benchmarks, but I bet 99% of people haven't understood what really makes it mind-blowing. First off, this isn't some giant monolithic model at all—it has only 0.6B parameters and essentially works as an AI project manager. It handles simple tasks on its own, automatically splits complex ones, and selects the most suitable models from a global pool of top-tier models...
Summary
Sakana AI releases Fugu, a multi-agent orchestration system with only 0.6B parameters. By intelligently splitting tasks and coordinating multiple models, it achieves state-of-the-art performance while bypassing traditional parameter scaling. This marks the transition of multi-agent orchestration from a lab curiosity to a practical productivity tool.
View Cached Full Text
Cached at: 06/23/26, 02:09 PM
Everyone online is hyping that Japan’s Fugu beats GPT in benchmarks, but I bet 99% of people don’t get what’s truly game-changing here.
First, this thing isn’t a giant monolithic model at all.
It only has 0.6B parameters, and its core job is basically an AI project manager.
Simple tasks it handles itself; complex tasks it automatically breaks down, picks the best model from a global pool of top-tier models, assigns three roles—thinking, execution, verification—iterates in multi-agent collaboration, and finally synthesizes the answer.
Calling it is no different from using a regular model—just one line of API.
But the orchestration strategy behind it is trained, not crafted by hand-written prompts or manually tuned routing. It can discover collaboration patterns that humans would never think of.
What I find most impressive isn’t that it beats Claude and GPT on benchmarks. It’s that it directly bypasses the scaling law arms race.
No need to stack trillion parameters, no need for massive supercomputing centers. By relying on a smarter collaboration mechanism, it reaches the ceiling of frontier models. For the first time, AI competition has shifted from piling on parameters to managing intelligence.
Of course, it’s no silver bullet. There’s opacity in the black box, higher latency for complex tasks, and it’s actually more expensive for simple queries.
But the signal here is a hundred times more important than the benchmark numbers. It means multi-agent orchestration has officially gone from a lab toy to a usable productivity tool.
The new track—the orchestration layer—has officially started today.
Sakana AI (@SakanaAILabs):
Introducing Sakana Fugu: A full multi-agent orchestration system accessible via a single model API.Our ‘Fugu Ultra’ model matches the performance of Fable and Mythos, delivering frontier capability without the risk of export controls.
Try it: https://t.co/aDEFyySWlS 🐡
Similar Articles
@mylifcc: This is not an ordinary large model, but a Multi-Agent Orchestration System—a small model itself that intelligently and dynamically coordinates multiple cutting-edge models such as GPT, Claude, and Gemini, autonomously assigning roles, decomposing tasks, and completing comp...
Sakana AI has released a Multi-Agent Orchestration System that uses a small model to intelligently coordinate cutting-edge large models like GPT, Claude, and Gemini to autonomously assign tasks and handle complex workloads.
@DeRonin_: HOLY SH*T, got released Fable-class model in public from Japan by coding and research benchmarks it's literally equival…
Sakana AI released Fugu Ultra, a multi-agent orchestration system accessible via a single model API, achieving performance competitive with Fable and Mythos models.
@berryxia: Small model, big wisdom? It's now real! A 7B small model now acts as the boss of top large models like GPT-5, Claude Sonnet 4, Gemini 2.5 Pro. A new paper shows an RL-trained 7B model learned to write natural language subtasks, assign them to different models, precisely...
A new paper proposes training a 7B small model via reinforcement learning as a task scheduler, automatically decomposing subtasks and assigning them to top models like GPT-5 and Claude. It surpasses individual frontier models on several hard benchmarks, demonstrating that end-to-end reward learning can effectively replace manual prompt engineering and multi-agent pipeline design.
@sashimikun_void: @serenaa_ge Deepswe benchmark pls
Sakana AI announced Sakana Fugu, a multi-agent orchestration system accessible via a single model API, with the Fugu Ultra model matching frontier performance without export control risks.
@omarsar0: OMG! Fugu Ultra is ridiculously good at these 3D renders.
Sakana AI announces Fugu Ultra, a multi-agent orchestration model that matches frontier performance of Fable and Mythos while avoiding export controls.