Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving
This story is from 2026-09-02. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.00111v1 Announce Type: new Abstract: We present Qwen-Drive-1.0, an initial step towards a vision-language foundation model for autonomous driving. Qwen-Drive-1.0 retains the architecture of the pretrained vision-language model (VLM) and integrates 3D perception, visual question answering…
Read the full story at arXiv cs.CV ↗
Timeline · 1 report
- 2026-09-02 04:00 · arXiv cs.CV
Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving