🛰️ Daily AI Frontier
‹ back to 2026-08-17

国内第一视角数据最早押注者,北大卢宗青:隐空间才是具身的路

Industry & News Embodied AI

Ranking

Overall 71
Content 80
Popularity N/A

No observed public metrics; popularity remains neutral/archived.

Representative image for 国内第一视角数据最早押注者,北大卢宗青:隐空间才是具身的路

Merged summary

TL;DR - BeingBeyond launched Being-H0.8, described as the first latent world-action model to jointly encode vision, touch, actions, and future state changes for robot control. Its latent-space approach targets real-time deployment at roughly 1% of pixel-based video-model training cost.

  • Being-H0.8 adds tactile signals to large-scale pretraining, aiming to model physical interactions rather than merely visual observations.
  • The company has curated over 500,000 hours of first-person human video, arguing it offers greater scale and diversity than robot-collected or simulated data.
  • Its models predict actions and world responses directly in embedding space, avoiding costly frame generation and supporting faster inference.
  • Founder Lu Zongqing cautions that embodied AI still lacks a proven paradigm comparable to next-token prediction for LLMs; even world models may not be the final answer.

Sources (1)

国内第一视角数据最早押注者,北大卢宗青:隐空间才是具身的路

雷峰网 (AI科技评论) 2026-08-17
Public signals N/A
Providers: Hugging Face · N/A OpenAlex · N/A Publisher · N/A Semantic Scholar · N/A X · N/A Fetched 2026-09-16 14:20:05.052027 UTC

TL;DR - BeingBeyond launched Being-H0.8, described as the first latent world-action model to jointly encode vision, touch, actions, and future state changes for robot control. Its latent-space approach targets real-time deployment at roughly 1% of pixel-based video-model training cost.

  • Being-H0.8 adds tactile signals to large-scale pretraining, aiming to model physical interactions rather than merely visual observations.
  • The company has curated over 500,000 hours of first-person human video, arguing it offers greater scale and diversity than robot-collected or simulated data.
  • Its models predict actions and world responses directly in embedding space, avoiding costly frame generation and supporting faster inference.
  • Founder Lu Zongqing cautions that embodied AI still lacks a proven paradigm comparable to next-token prediction for LLMs; even world models may not be the final answer.
item →