ExplorerComputer VisionComputer Vision
Research PaperResearchia:202609.23007

DreamStream: Towards Policy-Oriented Generative Simulation for End-to-End Driving

Ziyang Leng

Abstract

Faithfully evaluating end-to-end driving policies in simulation requires observations that are not merely photo-realistic, but preserve the scene features a policy relies on to make decisions. Existing platforms, however, exhibit a sim-to-real visual gap that corrupts policy perception, undermining their ability to assess a policy's closed-loop decision-making. To this end, we propose DreamStream, a generative, closed-loop simulator that achieves policy-oriented fidelity using a simulator-ground...

Submitted: September 23, 2026Subjects: Computer Vision; Computer Vision

Description / Details

Faithfully evaluating end-to-end driving policies in simulation requires observations that are not merely photo-realistic, but preserve the scene features a policy relies on to make decisions. Existing platforms, however, exhibit a sim-to-real visual gap that corrupts policy perception, undermining their ability to assess a policy's closed-loop decision-making. To this end, we propose DreamStream, a generative, closed-loop simulator that achieves policy-oriented fidelity using a simulator-grounded autoregressive video model. Our video model is distilled from a large pretrained video model via traffic layout guidance, varying visual appearance while preserving policy-relevant features such as scenario layout and the temporal consistency of dynamic objects. We further observe that perceptual metrics like FID misrank how well these features are preserved. To tackle this, we introduce FDππ, a new multi-representation metric that measures the sim-to-real gap as the Fréchet distance over scene-context features from public E2E policies. Under FDππ, DreamStream improves over the strongest prior closed-loop simulator by 1.6×1.6\times on nuScenes and 4.7×4.7\times on NAVSIM, and induces the least perturbation to policy's perceptual observability. Based on DreamStream, we construct Navhard-CL benchmark, which turns non-reactive real-world benchmark NAVSIM into interactive testing environments with adversarial driving behaviors and weather variations. This benchmark exposes many failure modes of driving policies, such as scorer bias and lack of recovery behaviors, that prior closed-loop benchmarks overlook. Code and data are available at https://github.com/VAIL-UCLA/DreamStream.


Source: arXiv:2609.26792v1 - http://arxiv.org/abs/2609.26792v1 PDF: https://arxiv.org/pdf/2609.26792v1 Original Link: http://arxiv.org/abs/2609.26792v1

Please sign in to join the discussion.

No comments yet. Be the first to share your thoughts!

Access Paper
View Source PDF
Submission Info
Date:
Sep 23, 2026
Topic:
Computer Vision
Area:
Computer Vision
Comments:
0
Bookmark
DreamStream: Towards Policy-Oriented Generative Simulation for End-to-End Driving | Researchia