๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

ModelScope
@ModelScope2022
Driving innovations with open communities. ๐Ÿ’ฌ Join our Discord:
๊ฐ€์ž… April 2024
183 ํŒ”๋กœ์ž‰ ์ค‘    16K ํŒฌ
Qwen just stepped into autonomous driving! ๐Ÿš— Qwen-Drive-1.0-4B is a vision-language foundation model that handles 3D perception, driving VQA, and motion planning in one framework, with the Qwen3.5-4B backbone left fully unmodified. Apache 2.0. ๐Ÿค– โš™๏ธ Two plug-in modules do the driving: a BEV head for 3D perception, and a flow matching Planning Expert for trajectories. ๐Ÿ“Š Leads driving VQA across the board: 77.8 on LingoQA, lowest Ego3D distance error, and 41.3 on causal reasoning where others score under 5. ๐Ÿ The RL planner hits 90.7 PDMS on NAVSIM, ahead of AutoVLA and SpanVLA. ๐Ÿง  No catastrophic forgetting: general benchmarks stay on par with the base model.
๋” ๋ณด๊ธฐ