ExplorerArtificial IntelligenceAI
Research PaperResearchia:202608.25063

Correcting a learned physical invariant improves world-model rollouts

Richard Bao

Abstract

World models can predict video without learning dynamics that they reliably preserve. We test whether a frozen DreamerV3 trained only on pendulum video learns a scalar that its own latent transition treats as approximately conserved. A label-free search recovers the same energy-like invariant across independently trained conservative models, while the same procedure finds no comparable invariant in matched damped models. During autonomous rollouts, this quantity drifts. Projecting the latent sta...

Submitted: August 25, 2026Subjects: AI; Artificial Intelligence

Description / Details

World models can predict video without learning dynamics that they reliably preserve. We test whether a frozen DreamerV3 trained only on pendulum video learns a scalar that its own latent transition treats as approximately conserved. A label-free search recovers the same energy-like invariant across independently trained conservative models, while the same procedure finds no comparable invariant in matched damped models. During autonomous rollouts, this quantity drifts. Projecting the latent state back toward its initial level set reduces rollout error in all three conservative models, whereas matched random constraints usually increase it. These results distinguish a dynamically meaningful invariant from a merely decodable correlate and reveal a concrete failure mode: a world model can learn a physical constraint from pixels yet violate that constraint when it imagines forward.


Source: arXiv:2608.23526v1 - http://arxiv.org/abs/2608.23526v1 PDF: https://arxiv.org/pdf/2608.23526v1 Original Link: http://arxiv.org/abs/2608.23526v1

Please sign in to join the discussion.

No comments yet. Be the first to share your thoughts!

Access Paper
View Source PDF
Submission Info
Date:
Aug 25, 2026
Topic:
Artificial Intelligence
Area:
AI
Comments:
0
Bookmark