Skip to content
arXiv Robotics — research abstracts· Xurui Song, Shuo Huai, JingJing Jiang, Jiayi Kong, Jun Luo·· 1 days agoEditorial score58

More Than Meets the Eye? Uncovering the Reasoning-Planning Disconnect in Training Vision-Language Driving Models

More Than Meets the Eye? Uncovering the Reasoning-Planning Disconnect in Training Vision-Language Driving Models

Summary

This research explores whether reasoning in vision-language driving models causally influences planning. Using a new dataset, DriveMind, the authors find that planning is primarily driven by priors rather than reasoning, proposing a decoupling hypothesis and introducing a diagnostic tool for model evaluation.

Source: arXiv Robotics — research abstracts · Read original article ↗

Loading article text…

Source:arXiv Robotics — research abstracts · arxiv.org

Timezone · UTC

Article dates follow your selected timezone. Briefing editions use Hong Kong time (UTC+8).