Skip to content
arXiv Robotics — research abstracts· Yuqi Ye, Shangkun Sun, Junhong Lin, Jiayi Zhao, Changhao Peng, Wei Zheng, Guoqing Liu, Tiesong Zhao, Wei Gao·· 1 days agoEditorial score66

Sometimes You Gotta Run Before You Can Walk: Run-then-Walk Scheduling Strategy for VLM Autonomous Driving

Sometimes You Gotta Run Before You Can Walk: Run-then-Walk Scheduling Strategy for VLM Autonomous Driving

Summary

This research proposes a two-stage reward scheduling strategy called 'Run-then-Walk' for VLM-based autonomous driving planners. It addresses the limitations of existing GRPO-style reinforcement learning methods by separating progress discovery and safety repair. The approach achieves better performance and faster convergence, validated on multiple benchmarks with reduced training epochs.

Source: arXiv Robotics — research abstracts · Read original article ↗

Loading article text…

Source:arXiv Robotics — research abstracts · arxiv.org

Timezone · UTC

Article dates follow your selected timezone. Briefing editions use Hong Kong time (UTC+8).