Qwen-Drive-1.0: A Vision-Language Foundation Model for Autonomous Driving

Qwen · hf · 2026-09-02

Qwen-Drive-1.0 is a vision-language foundation model for autonomous driving that unifies 3D perception, visual question answering, and motion planning. It utilizes shared representations and staged training to improve comprehensive understanding and decision-making.

Original post →

More from Embodied

Embodied channel →