Loading market data...
ai

Ant Group’s Robbyant Unveils LingBot-VA 2.0: A Causal Video-Action Model Built Natively for Physical AI

MarkTechPost
Read Full Article at MarkTechPost
Share:PostShare
Ant Group’s Robbyant Unveils LingBot-VA 2.0: A Causal Video-Action Model Built Natively for Physical AI
Ad Slot — In-Article (728x90)

Ant Group's Robbyant has released the LingBot-VA 2. 0 technical report — a Physical AI video-action foundation model built from scratch for embodiment rather than fine-tuned from a video generator.

It predicts future states ahead of execution through Foresight Reasoning, re-grounds on every real observation, and reaches 225 Hz asynchronous control. We break down the causal DiT, the sparse-MoE video stream, the semantic visual-action tokenizer, and where the paper's own numbers don't line up.

This is a summary. For the full story, read the original article at MarkTechPost.

Original source: MarkTechPost

Ad Slot — Below Article (300x250)