@TheAhmadOsman: 600M that beats a 397B and Sonnet 4.5 Small and specialized models FTW
Summary
A 600M parameter reasoning model trained using SYNTH reportedly outperforms a 397B model and Sonnet 4.5 in an industrial application for the Paris subway, highlighting the effectiveness of small, specialized models.
View Cached Full Text
Cached at: 06/20/26, 06:20 PM
600M that beats a 397B and Sonnet 4.5
Small and specialized models FTW https://t.co/AvJf3ykNHa
Alexander Doria (@Dorialexander): Announcing the first industrial application of SYNTH: we trained a 600m reasoning model for one of the largest infrastructure in the world, the subway of Paris.
Similar Articles
A 4b model is now beating 30b ones at web research and the reason is not size
A 4 billion parameter open model from the Apodex family outperforms 30 billion parameter models on web research benchmarks, attributed to careful training data and self-verification techniques rather than raw scale, suggesting a more democratic trajectory for AI capability.
@svpino: For the first time, I feel open-weight models are impossible to ignore. We are at a point where these models are compet…
Santiago (@svpino) highlights MiniMax-M2.7, a 230B open-weight model that rivals top proprietary models like Opus 4.6 and GPT-5.4, achieving 440+ tokens/s inference on SambaNova at low cost.
@natolambert: Thinky with a ~1T param, 41B active, apache-2 model Benchmarks are a clear step up from Nemotron Ultra (55B active), ne…
Thinky is a ~1T parameter mixture-of-experts model with 41B active parameters, released under Apache-2 license. It achieves new best results among American models, benchmark improvements over Nemotron Ultra, with omni-modal input.
@rohanpaul_ai: Can a smaller model purpose-built for one domain beat a frontier general model that's 100× its size? A recent paper sho…
PolyAI's Raven 3.5, a smaller specialist model, outperforms GPT-5 and Claude Sonnet 4.6 on all customer service benchmarks with under 300ms latency. The company also launches ADK and PolyPhone to accelerate enterprise voice AI deployment.
[MASSIVE TINY RELEASE] - Supra2-Medium-Base - a tiny 25M parameters model competing heavily with our previous 50M model!
Supra2-Medium-Base is a new 25M parameter AI model based on qwen3 architecture, trained from scratch, that benchmarks competitively against a larger 50M parameter model.