Tag
Fable 5.5 was prompted to 'build a voxel version of yourself' and unexpectedly produced a fully explorable 3D voxel world, showcasing surprising emergent capabilities in spatial and world-generation tasks.
WorldAuditBench is a benchmark of 213 anomaly-auditing tasks across 13 interactive 3D environments (Unreal Engine 5 / Three.js) that evaluates whether multimodal agents can couple action with visual reasoning. Frontier models achieve only 6.6–42.3% success versus 83.4% human performance, highlighting current limitations in evidence gathering and interpretation.
GPT-6 Astra, an AI model, is viral for its diverse applications in gaming, protein design, sports tracking, and 3D world building, highlighted by 10 stunning cases.
Anthropic has released Claude Fable 5.1, an AI model that users are rapidly using to generate games, build 3D worlds, and create simulations.
Figma Weave integrates Marble from @theworldlabs to enable the creation and exploration of 3D worlds from text, images, or video inputs.
A new benchmark for evaluating multimodal agents in building 3D open worlds from natural language shows GPT-5.5 and Qwen3.8-Max scoring below 60%, while an open 30B model achieves the best performance through reinforcement learning.
This paper proposes VibeWorlding, a framework for benchmarking and training multimodal agents to construct 3D open worlds from user queries, showing that reinforcement learning improves open-source models to compete with closed-source frontiers.