Tag
Kimi K3 is a 2.8T parameter open model from Moonshot AI, showing strong benchmark performance but likely over-optimized and lagging behind top closed models by months. It is distilled from Claude and its release may precede an IPO.
NVIDIA released Cosmos 3 Edge, a 4-billion-parameter open world model for edge devices that helps robots and vision AI agents understand surroundings, reason in real time, and generate actions. It achieves best-in-class throughput and accuracy among similar-sized models.
Google announced Gemma 4 E2B optimized for the Pixel 10's TPU, enabling on-device multimodal AI capabilities like offline chat, image recognition, and audio transcription.
An open model that predicts a robot's actions from a control signal, raising questions about whether it constitutes a world model or just a video generator.
This paper analyzes open language model adoption, finding that Chinese models led by Qwen now dominate downloads, surpassing US models by March 2026. Qwen's lead comes from a diverse range of model sizes, while DeepSeek leads in very large models, and some US models still show strong momentum.
Introducing Leanstral 1.5, a 119B parameter (6B active) open model for formal proof engineering in Lean 4, achieving 100% on miniF2F, state-of-the-art scores on PutnamBench and FATE benchmarks, and discovering previously unknown bugs in open-source repositories.
GLM 5.2, an open AI model, is now available in the Cursor coding tool via a partnership with Fireworks.
NVIDIA released Nemotron-TwoTower-30B-A3B-Base-BF16, a diffusion-based language model that uses block-wise autoregressive diffusion to generate text by iterative denoising of token blocks, achieving 2.42× the generation throughput of the autoregressive baseline while retaining 98.7% of benchmark quality.
GLM-5.2 is a new open-source AI model that sets a high bar for open models, though it still trails proprietary frontier models and lacks some features like vision.
NVIDIA released the Nemotron 3 open model, offering three sizes: Nano, Super, and Ultra. It optimizes hardware efficiency through architectural innovations such as hybrid Mamba Transformer, latent MoE, and multi-token prediction, and adopts the Open MDW 1.1 open license.
NVIDIA optimizes Google DeepMind's DiffusionGemma, an open model that generates text in parallel 256-token blocks, achieving up to 4x faster performance on local RTX GPUs, DGX Spark, and DGX Station systems.
Google introduces DiffusionGemma, an experimental 26B MoE open model that achieves up to 4x faster text generation on GPUs using text diffusion, targeting speed-critical interactive local workflows.
Nex AGI releases Nex-N2, an open-source agentic model series for coding, tool use, deep research, and long-horizon workflows, with state-of-the-art benchmarks and Apache 2.0 license.
NVIDIA announces Alpamayo 2 Super, a 32B open reasoning model for Level 4 robotaxis, featuring 360-degree perception, meta-actions, and a full stack including AlpaGym simulation and OmniDreams scenario generation.
The article observes that Tencent's Hy3 Preview open model performs surprisingly well in evaluations, narrowing the gap with top closed models, yet remains underdiscussed compared to Western AI labs.
NVIDIA announces Nemotron 3 Nano Omni, an open multimodal model that unifies vision, audio, and language processing to enable faster and more efficient AI agents, achieving up to 9x higher throughput compared to other open omni models.
Kimi K2.6, a highly-regarded open model, is available free for 24h via Nous Portal using Vercel’s AI Gateway.
Andrew Ng discusses the nuanced impact of AI on the job market, noting that while widespread layoffs are overhyped, AI skills are becoming crucial. The newsletter also covers news about OpenClaw, Kimi's open model, Ministral distilled, and Wikipedia's partners.