🛰️ Daily AI Frontier
‹ back to 2026-08-12

现在不做VLA和世界模型的公司,还有哪些?

WeChat: 自动驾驶之心 Autonomous Driving VLA 2026-08-10
Representative image for 现在不做VLA和世界模型的公司,还有哪些?

TL;DR - A WeChat industry roundup (from 自动驾驶之心, ending in a promotion for a 14-week paid research course) arguing that VLA (Vision-Language-Action) plus world models have become the default technical route for essentially every major autonomous-driving player, with differentiation now in the entry point rather than the direction.

  • Companies with publicly disclosed VLA work: Li Auto (Mind-VLA), XPeng (VLA 2.0), DeepRoute (元戎), Xiaomi; internationally Tesla (world model fused with FSD), Waymo, NVIDIA. BYD and Changan are named as recent entrants to VLA + world model.
  • World-model-leaning players cited: NIO, Huawei, Momenta, Pony.ai, WeRide — with few public papers, so the routing is inferred by the industry rather than confirmed.
  • Holdouts named (Geely, GAC, SAIC, Chery) are argued to be doing internal pre-research; the author's claim is that silence reflects KPIs tied to mass-production milestones, not absence of work. This is opinion, not verified.
  • Open technical problems the piece flags: coupling VLA reasoning chains with world-model future prediction, shared representations between action generation and physics modeling, on-vehicle stability, VLM real-time latency, language hallucination, and long-horizon spatiotemporal coherence. Referenced baselines/datasets: UniAD, VAD, DiffusionDrive, OpenDriveVLA, Senna; nuScenes, Waymo, Argoverse, Bench2Drive.

view merged work →