@poetiq_ai: Poetiq's Meta-System built its own coding harness from scratch. It got SOTA on LiveCodeBench Pro. No fine-tuning, no sp…
Summary
Poetiq's Meta-System achieved state-of-the-art results on LiveCodeBench Pro by autonomously building a coding harness using standard APIs and Gemini 3.1 Pro, without fine-tuning or special model access.
View Cached Full Text
Cached at: 05/14/26, 06:42 PM
Poetiq’s Meta-System built its own coding harness from scratch. It got SOTA on LiveCodeBench Pro.
No fine-tuning, no special model access. Just standard APIs. Using Gemini 3.1 Pro, it made a harness that beat all frontier models we tested. https://t.co/v575oUYJeH
Similar Articles
Poetiq: Recursive Self-Improvement Delivers New SOTA Coding Performance
Poetiq's Meta-System, using recursive self-improvement via standard API access without fine-tuning, achieves new state-of-the-art results on the LiveCodeBench Pro coding benchmark, outperforming leading models like GPT 5.5.
New SOTA: Poetiq uses self-optimizing harness to surpass e.g. Opus 4.7 with Gemini 3 Flash
Poetiq claims new state-of-the-art coding performance using a self-optimizing harness with Gemini 3 Flash, surpassing Opus 4.7.
@cline: Muse Spark 1.1 just launched and it's their most capable coding agent model yet. On Terminal-Bench 2.1 it scores 80.0%,…
Meta launches Muse Spark 1.1, an upgraded coding agent model scoring 80.0% on Terminal-Bench 2.1, alongside a public preview of the Meta Model API.
favorite Agentic Coding Harness
The author compares several agentic coding harnesses (Codex CLI, Claude Code, Gemini CLI, OpenCode, Pi) and finds Pi the leanest and best for local models, praising its simplicity and compatibility with Qwen 27B-MXFP8.
Meta says its new AI model is ready to compete on coding
Meta releases Muse Spark 1.1, a new AI model for advanced coding tasks including bug detection and agentic workflows, available via the Meta Model API.