Tag
Fusion, a new AI model from a Singapore lab, demonstrates strong performance in SVG generation with accurate structural connections and significant cost savings compared to Opus 5 from Claude.
A 7.9B MoE model runs at 152 tokens per second on an M4 Pro with 64GB unified memory, enabling offline processing of sensitive contract data and demonstrating the practical use of local AI.
This article benchmarks various engines for running the Qwen 3.8 27B model on macOS, comparing their speed and performance in agentic coding tasks, and recommends MTPLX or llama.cpp with MTP for best results.
RunAnywhere releases an on-device AI benchmark app that lets users measure model performance directly on their own phones, replacing posted numbers with real-time testing.