现在不做VLA和世界模型的公司,还有哪些?
Ranking
No observed public metrics; popularity remains neutral/archived.
Merged summary
TL;DR - A WeChat industry roundup (from 自动驾驶之心, ending in a promotion for a 14-week paid research course) arguing that VLA (Vision-Language-Action) plus world models have become the default technical route for essentially every major autonomous-driving player, with differentiation now in the entry point rather than the direction.
- Companies with publicly disclosed VLA work: Li Auto (Mind-VLA), XPeng (VLA 2.0), DeepRoute (元戎), Xiaomi; internationally Tesla (world model fused with FSD), Waymo, NVIDIA. BYD and Changan are named as recent entrants to VLA + world model.
- World-model-leaning players cited: NIO, Huawei, Momenta, Pony.ai, WeRide — with few public papers, so the routing is inferred by the industry rather than confirmed.
- Holdouts named (Geely, GAC, SAIC, Chery) are argued to be doing internal pre-research; the author's claim is that silence reflects KPIs tied to mass-production milestones, not absence of work. This is opinion, not verified.
- Open technical problems the piece flags: coupling VLA reasoning chains with world-model future prediction, shared representations between action generation and physics modeling, on-vehicle stability, VLM real-time latency, language hallucination, and long-horizon spatiotemporal coherence. Referenced baselines/datasets: UniAD, VAD, DiffusionDrive, OpenDriveVLA, Senna; nuScenes, Waymo, Argoverse, Bench2Drive.
Sources (1)
现在不做VLA和世界模型的公司,还有哪些?
TL;DR - A WeChat industry roundup (from 自动驾驶之心, ending in a promotion for a 14-week paid research course) arguing that VLA (Vision-Language-Action) plus world models have become the default technical route for essentially every major autonomous-driving player, with differentiation now in the entry point rather than the direction.
- Companies with publicly disclosed VLA work: Li Auto (Mind-VLA), XPeng (VLA 2.0), DeepRoute (元戎), Xiaomi; internationally Tesla (world model fused with FSD), Waymo, NVIDIA. BYD and Changan are named as recent entrants to VLA + world model.
- World-model-leaning players cited: NIO, Huawei, Momenta, Pony.ai, WeRide — with few public papers, so the routing is inferred by the industry rather than confirmed.
- Holdouts named (Geely, GAC, SAIC, Chery) are argued to be doing internal pre-research; the author's claim is that silence reflects KPIs tied to mass-production milestones, not absence of work. This is opinion, not verified.
- Open technical problems the piece flags: coupling VLA reasoning chains with world-model future prediction, shared representations between action generation and physics modeling, on-vehicle stability, VLM real-time latency, language hallucination, and long-horizon spatiotemporal coherence. Referenced baselines/datasets: UniAD, VAD, DiffusionDrive, OpenDriveVLA, Senna; nuScenes, Waymo, Argoverse, Bench2Drive.