Tagged articles

latent action

3 articles · Page 1 of 1
Amap Tech
Amap Tech
Jul 10, 2026 · Artificial Intelligence

ABot-M0.5: The First Unified World Action Model for Mobile Manipulation

ABot-M0.5 introduces a unified world action model that aligns video prediction, intermediate latent actions, and low‑level physical control for mobile manipulation, achieving state‑of‑the‑art long‑horizon success rates and fine‑grained precision across benchmarks such as RoboCasa365, RoboTwin, and LIBERO, while detailing novel architectural components and a three‑stage progressive training regime.

benchmark resultsdream forcingdual-level transformer
0 likes · 14 min read
ABot-M0.5: The First Unified World Action Model for Mobile Manipulation
Machine Heart
Machine Heart
Jun 29, 2026 · Artificial Intelligence

How MWA™'s Long‑Sequence Bidirectional Physical Causal Chain Sets a New Record in Embodied AI

The article presents MWA™, the first long‑sequence bidirectional physical causal chain hidden‑space world model, details its bidirectional dynamics, latent‑action pre‑training, three‑gradient constraints and AnyPhys negative‑sample system, and shows it achieved a 75.2% success rate on the RoboCasa GR1 TableTop benchmark, surpassing leading competitors.

AnyPhysEmbodied AIRoboCasa benchmark
0 likes · 14 min read
How MWA™'s Long‑Sequence Bidirectional Physical Causal Chain Sets a New Record in Embodied AI
Machine Heart
Machine Heart
May 27, 2026 · Artificial Intelligence

CVPR 2026: Learning Camera Pose from 10M Unlabeled Driving Videos

LA‑Pose shows that a model can acquire accurate camera pose estimation for autonomous driving by self‑supervised pretraining on roughly ten million unlabeled driving video clips and fine‑tuning with only a small amount of high‑quality 3D annotations, achieving over 10% accuracy gains while drastically reducing labeling cost.

CVPR 2026LA-Poseautonomous driving
0 likes · 8 min read
CVPR 2026: Learning Camera Pose from 10M Unlabeled Driving Videos