Back to News
Qwen Team & Huazhong University of Science and Technology•September 2, 2026
Qwen-Drive-1.0
Open Source
Autonomous Driving
Vision-Language Model
Motion Planning
3D Perception
### TL;DR
Qwen-Drive-1.0 is a vision-language foundation model for autonomous driving that integrates 3D perception, visual question answering, and motion planning into a unified framework. It leverages the Qwen3.5-4B vision-language model as a base, attaching specialized heads for BEV perception and trajectory planning to enable comprehensive driving scene understanding and ego-vehicle control.
Key Insights & Metrics
Pricing
Open Source (Apache 2.0 License); Weights available on Hugging Face
Cost structure
Version
1.0
Current release version
Hardware
GPU with 24 GB+ of memory recommended
Compute requirements
Category
Open Source
Licensing model
Region
China
Primary region
Key Features
- Unified framework for 3D perception, visual question answering, and motion planning
- BEV Perception Head for 3D object detection, semantic occupancy, and map segmentation
- Planning Expert utilizing flow matching for future ego trajectory generation
- Maintains general-purpose vision-language capabilities of the Qwen3.5-4B base
- Staged training strategy for domain-specific adaptation and consistency
Ad

Master AI Marketing: Work Smarter, Not Harder
Unlock the power of AI to automate workflows and scale your results. Join 1.5M+ professionals and stay ahead of the curve with the latest AI insights.
Sponsored
Discussion
0
Upvotes
0
Downvotes
0 reviews
Sign in to leave a review
Reviews
No reviews yet. Be the first to review!