AI news story

Ant Group’s Robbyant Unveils LingBot-VA 2.0: A Causal Video-Action Model Built Natively for Physical AI

Ant Group's Robbyant has released the LingBot-VA 2.0 technical report — a Physical AI video-action foundation model b…

  • Generative
  • Source: MarkTechPost
  • Published: 2026-07-11

Editor's take

Ant Group's Robbyant has introduced LingBot-VA 2.0, a novel foundation model designed specifically for physical AI tasks, capable of predicting future states for robotic action planning.

This development is significant because it moves beyond adapting existing generative video models, like those from OpenAI or Google DeepMind, to the embodied domain. By building natively for physical interaction, LingBot-VA 2.0 targets the critical challenge of real-world robotic control, where precise prediction and reaction are paramount, distinguishing it from models primarily focused on visual synthesis.

Future developments to monitor include the model's performance on complex, multi-step manipulation tasks in unstructured environments, and whether it can achieve comparable or superior dexterity to robots controlled by traditional, hand-engineered planning algorithms. Observing its integration into real-world robotic platforms and its ability to generalize across different hardware will be key indicators of its practical impact.