spatial-reasoning

Tag

Cards List
#spatial-reasoning

Soft Spatial Reasoning

Hugging Face Daily Papers ↗ · yesterday Cached

The paper proposes Soft Spatial Reasoning, a post-training framework that introduces adaptive "soft thinking" for spatial tasks in Large Vision-Language Models, where intermediate reasoning steps mix token embeddings instead of committing to a single discrete token. An AdaptSoft controller dynamically adjusts softness based on hidden state and predictive uncertainty, outperforming hard and fixed-soft CoT baselines on spatial benchmarks.

0 favorites 0 likes
#spatial-reasoning

SpatialCORE: Confidence-Aware Grounded Spatial Reasoning in Large Vision--Language Models

Hugging Face Daily Papers ↗ · yesterday Cached

SpatialCORE is a post-training framework that uses a model's confidence in its generated bounding-box grounding as a learning signal for spatial reasoning in large vision-language models, achieving state-of-the-art results among open-source models on spatial reasoning benchmarks.

0 favorites 0 likes
#spatial-reasoning

LeRF: Learning Reference Coordinate Frames for Perspective Taking Reasoning

Hugging Face Daily Papers ↗ · 3d ago Cached

LeRF introduces a method to enhance perspective-taking reasoning in vision-language models by learning reference coordinate frames, improving performance on benchmarks through supervised fine-tuning and reinforcement learning.

0 favorites 0 likes
#spatial-reasoning

SpatialSpeak: QA-Native Reconstruction with Local and Global Context for Spatial Chain-of-Thought Reasoning

Hugging Face Daily Papers ↗ · 4d ago Cached

SpatialSpeak introduces a two-stage framework for spatial reasoning in vision-language models, using QA-native reconstruction pretraining and spatial chain-of-thought learning to achieve state-of-the-art performance on benchmarks.

0 favorites 0 likes
#spatial-reasoning

Astra leads in IKEA furniture assembly

Reddit r/singularity ↗ · 5d ago Cached

Epoch AI created the Furniture Assembly Benchmark (FAB) to test AI models' visual reasoning in spotting mistakes in IKEA assembly photos. OpenAI's GPT-6 Astra leads with 80% accuracy, showing major improvements over previous models.

0 favorites 0 likes
#spatial-reasoning

Spatial-Interactor: Learning Spatial Reasoning through Interaction with the Observable Physical World

arXiv cs.AI ↗ · 2026-09-23 Cached

Spatial-Interactor is a framework that trains vision-language models to enhance spatial reasoning through interaction with the physical world, employing a three-level curriculum and two-stage training strategy to improve state transition modeling and long-horizon integration.

0 favorites 0 likes
#spatial-reasoning

HarnessVLN: Unifying Training-Free Embodied Navigation through an Agent Harness

Hugging Face Daily Papers ↗ · 2026-09-14 Cached

HarnessVLN is a zero-shot, training-free framework for embodied navigation that unifies perception, retrieval, grounding, navigation, recovery, and termination through a unified tool interface, achieving state-of-the-art results on benchmarks like R2R and RxR.

0 favorites 0 likes
#spatial-reasoning

Astra scores the highest on Blueprint-bench2

Reddit r/singularity ↗ · 2026-09-07

Astra achieved the highest score on Blueprint-Bench 2, a benchmark testing AI agents' ability to convert apartment photos into accurate 2D floor plans using spatial reasoning and cross-apartment learning.

0 favorites 0 likes
#spatial-reasoning

@reach_vb: Astra is SoTA on MazeBench by a huge margin:

X AI KOLs Timeline ↗ · 2026-09-07 Cached

Astra achieves state-of-the-art performance on MazeBench, significantly outperforming GPT-6 in a 3D open world spatial reasoning evaluation.

0 favorites 0 likes
#spatial-reasoning

GPT-6 Astra's 3D modeling capabilities are way beyond what I expected

Reddit r/ArtificialInteligence ↗ · 2026-09-06

The author discusses demos of GPT-6 Astra demonstrating advanced iterative 3D modeling capabilities, including reasoning about geometry and progressively improving results, and asks what technical changes are driving this leap.

0 favorites 0 likes
#spatial-reasoning

@BenjaminDEKR: Astra understands spatial / 3D relationships better than other leading models. This is why it's so good at CAD, models,…

X AI KOLs Timeline ↗ · 2026-09-05 Cached

Astra reportedly surpasses other leading models in spatial and 3D understanding, achieving first place on the VoxelBench benchmark with an Elo rating exceeding 2600 and a lead of over 300 points.

0 favorites 0 likes
#spatial-reasoning

TaichuAI/ZDTaichu5.0-9B

Hugging Face Models Trending ↗ · 2026-09-04 Cached

ZDTaichu5.0-9B is a multimodal foundation model that combines a Qwen3.5-9B language backbone with a C-RADIOv4-H vision encoder, excelling in general visual understanding, spatial reasoning, and agent tasks among 10B-scale VLMs.

0 favorites 0 likes
#spatial-reasoning

RoboSPA: Can VLA Models Go Beyond Simple Scenes and Short-Horizon Tasks?

Hugging Face Daily Papers ↗ · 2026-09-04 Cached

The paper introduces RoboSPA, a large-scale benchmark for evaluating vision-language-action models on fine-grained spatial reasoning and long-horizon procedural planning in robotic manipulation.

0 favorites 0 likes
#spatial-reasoning

Playco cut manual fixes 50% prototyping games with GPT-6 Astra

OpenAI Blog ↗ · 2026-09-03 Cached

Playco used GPT-6 Astra to build Playbot, an AI-powered IDE for game developers, cutting manual fixes by 50% and improving spatial reasoning and prototyping efficiency.

0 favorites 0 likes
#spatial-reasoning

Unfold The World: Factorize 4D Properties in Reinforcing Spatial Reasoning

Hugging Face Daily Papers ↗ · 2026-09-03

This paper introduces FactoSR, a factorized reinforcement learning framework that enhances spatial reasoning in Vision-Language Models by decomposing 4D properties into orthogonal sub-objectives, achieving significant performance boosts on multi-view and video benchmarks.

0 favorites 0 likes
#spatial-reasoning

Autoregressive Mosaics: Probing 2D Spatial Reasoning in Text-Only Language Models

Hugging Face Daily Papers ↗ · 2026-09-01 Cached

Introduces Autoregressive Mosaics (AM-Bench), a benchmark to evaluate whether text-only LLMs have genuine 2D spatial reasoning abilities, distinct from code generation. Findings show spatial reasoning varies among models and is influenced by output medium like SVG vs. code.

0 favorites 0 likes
#spatial-reasoning

UrbanGround: From Local Perception to Spatial Agency in a Real-Scale City

Hugging Face Daily Papers ↗ · 2026-08-27 Cached

UrbanGround evaluates whether multimodal language model agents can sustain reliable navigation and spatial reasoning in a realistic 3D city replica, revealing that local perceptual skills fail to compose into extended goal-directed behavior.

0 favorites 0 likes
#spatial-reasoning

AI models flub these intelligence tests. Can you fare any better?

MIT Technology Review ↗ · 2026-08-26 Cached

The article discusses how AI models struggle with intelligence tests like spatial reasoning and memory puzzles, highlighting gaps compared to human cognition and inviting readers to test their own skills.

0 favorites 0 likes
#spatial-reasoning

GUI-Primitives: Diagnosing Spatial Reasoning Failures in Vision-Language GUI Grounding

arXiv cs.CL ↗ · 2026-08-25 Cached

The paper introduces GUI-Primitives, a benchmark of 994 contrastive instruction pairs to diagnose spatial reasoning failures in vision-language models for GUI grounding, revealing that most failures stem from candidate localization rather than relation understanding.

0 favorites 0 likes
#spatial-reasoning

StateSight: Benchmarking Latent Spatial-State Reconstruction in Vision-Language Models

arXiv cs.AI ↗ · 2026-08-24 Cached

StateSight is a new benchmark for evaluating spatial-state reconstruction in vision-language models, showing that models like GPT-5.5 and Claude Sonnet 5 struggle with spatial reasoning tasks compared to human performance.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback