返回全部动态
Qwen-Drive-1.0:迈向自动驾驶视觉语言基础模型
原标题:Paper page - Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving
AI 摘要
Qwen-Drive-1.0 是阿里推出的自动驾驶视觉语言基础模型,基于 Qwen3.5-4B,通过外部 BEV 感知头和规划专家模块,统一了 3D 感知、视觉问答和运动规划。实验表明其在保持通用视觉语言能力的同时,显著提升了自动驾驶性能,为驾驶场景适配提供了新基础。
以上摘要由 AI 生成,可能存在误差。事实请以原文为准。
正文节选
Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving Abstract Qwen-Drive-1.0 is a vision-language foundation model for autonomous driving that unifies 3D perception, visual question answering, and motion planning via shared representations and staged training. We present Qwen-Drive-1.0, an initial step towards a vision-language foundation model for autonomous driving. Qwen-Drive-1.0 retains the architecture of the pretrained vision-language model (VLM
发布时间:—
抓取时间:2026-09-02 22:07
来源机构:Hugging Face