latent-world-models

Tag

Cards List
#latent-world-models

Reinforced Planning with Latent World Models

arXiv cs.LG · 4d ago Cached

The paper introduces Reinforced Planning, a method that learns to improve multi-step plans using latent world models, achieving near-perfect success in tasks like visual navigation and robotic manipulation with significantly higher efficiency than hand-designed algorithms.

0 favorites 0 likes
#latent-world-models

Decision-Metric Alignment in Latent World Models: Diagnostics and Action-Conditioned Objectives for MPC Planning

Hugging Face Daily Papers · 5d ago Cached

This paper introduces decision-metric alignment diagnostics and action-conditioned objectives to improve latent world models for model-predictive control. The proposed DA-LeWM method enhances convergence and success rates in planning tasks.

0 favorites 0 likes
#latent-world-models

The Objective Is the Bottleneck: Latent World Models Encode What Their Planners Cannot Use

arXiv cs.LG · 2026-08-14 Cached

The paper investigates why latent world models fail at long-horizon planning and finds the bottleneck is the planning objective (squared latent distance), not the predictor's accuracy; replacing the objective with a learned cost dramatically improves planning performance.

0 favorites 0 likes
#latent-world-models

Predicting Closed-Loop Performance of Latent World Models: Offline Checkpoint Selection for MPC and Model-Based RL Under Non-Markovian Rewards in LunarLander

arXiv cs.LG · 2026-07-03 Cached

This paper addresses objective mismatch in model-based RL by proposing offline diagnostics to predict closed-loop performance of latent world models. On LunarLander-v3, the Reward Observability Fraction (ROF) and a Composite score (CROF) enable selecting checkpoints that yield strong MPC and model-based RL policies with far fewer real-environment interactions.

0 favorites 0 likes
← Back to home

Submit Feedback