Damn muse 1.1 now better than fable 5 for medical and legal use ??
Summary
Damn muse 1.1 is claimed to outperform fable 5 in medical and legal use cases.
Similar Articles
@_jasonwei: Muse Spark 1.1 outperforms GPT-5.6 Sol and Gemini 3.1 on Radiology's Last Exam. We don't beat Fable (yet). And Humans a…
Muse Spark 1.1 outperforms GPT-5.6 Sol and Gemini 3.1 on Radiology's Last Exam 2.0, a new visual reasoning benchmark for autonomous AI diagnosis in healthcare, though it still lags behind Fable and human radiologists.
Muse Spark 1.1 is not far from Fable 5 in Text Arena ranking now
Muse Spark 1.1 has improved its Text Arena ranking, now close to Fable 5.
@_jasonwei: In addition to agents and coding, Muse Spark 1.1 is also really strong at answering health questions, a steadily growin…
Muse Spark 1.1, a new agentic and coding model from Meta, achieves +5% improvement on HealthBench-Pro, outperforming all competitors except Fable and Mythos.
@cline: Muse Spark 1.1 just launched and it's their most capable coding agent model yet. On Terminal-Bench 2.1 it scores 80.0%,…
Meta launches Muse Spark 1.1, an upgraded coding agent model scoring 80.0% on Terminal-Bench 2.1, alongside a public preview of the Meta Model API.
Introducing Muse Spark 1.1
Meta released Muse Spark 1.1, the first Spark model with an API, featuring improvements in agentic tool calling and computer use.