LensVLM-9B is a 9B-parameter Vision Language Model from Apple that scans compressed images of text and selectively expands relevant pages using learned tools, with paper, code, and usage instructions provided.
Black Forest Labs has released FLUX 3 Action, a collection of 7B AI models for robotics, including base models and policies for SO-101 and DROID, available on Hugging Face.
Google introduces Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, new text-to-speech models that enable expressive and customizable audio generation for creators and developers.
DrivingBench evaluates whether frontier AI models like GPT-6 Astra can drive a real car by controlling a Toyota Corolla on a predefined course, tracking metrics such as progress, distance, and token cost.
A new stealth multimodal AI model named Space Bunny has been released in OpenCode, available for free trial, with the releasing company unknown.
Release of high-quality quantized versions of the Qwen 3.8 27B AI model, claiming to outperform ISTA and Unsloth quants in quality tests.
MiMo-V3 is being updated with a new architecture centered on HySparse2, which has been released as a preprint on arXiv.
Opus 5.5, an AI model, generated content with a single prompt.
A tweet predicts that an AI model named Opus 5.5 will reach a specified intelligence level and be available from open-source labs in approximately six months.
NVIDIA releases Nemotron 3 Diarization, an open-weight 100M-parameter model that achieves state-of-the-art speaker diarization with a 14.72% error rate, supporting real-time and offline processing for up to eight speakers.
GPT-6 Astra achieves a massive leap on ZeroBench, an extremely difficult vision benchmark, surpassing the human baseline across all three metrics.
GPT-6 Sol and Claude Opus 5.5 achieve near-frontier performance at a fraction of the cost and speed of previous generations, highlighting major efficiency gains.
Open-Jev-27B-v1.1 is an open-source AI model with a LoRA adapter, achieving 85.28% accuracy on JevBench and featuring interactive demos for various tasks.
The latest Remote Labor Index results show that the GPT-6 Astra model can automate 20.8% of randomly chosen remote projects, a significant increase from 2.5% in October, highlighting rapid progress in AI automation.
The GGUF quantized version of the Qwen-Image-2.1 AI model is released, featuring Dynamic 2.0 technology for efficient 4-bit quantization, supporting text-to-image and transparent image generation in a 4.2GB size suitable for Mac users.
This article provides an update on the Engram model, detailing its 2.6b parameter architecture with a large Engram table and initial training progress at 100m tokens, showing improved completions with context-aware data offloading.
Addy Osmani praises Opus 5.5 as a cheaper and faster AI model, declaring it his new daily driver for most work tasks.
The title references 'opus 5.5', likely referring to Anthropic's AI model, suggesting a version update or release.
Someone gained early access to Opus 5.5 and used it to create game projects, including a Dark Souls-like project, with results exceeding expectations.
Claude Opus 5.5 has generated a Minecraft-like 3D world in the browser, featuring procedural terrain, dynamic lighting, water, and building elements.