| AHA-WAM:Asynchronous Horizon-Adaptive
World-Action Modeling with Observation-Guided
Context Routing |
AHA-WAM |
2026 |
2606.09811 |
|
arXiv
Code
Opt
|
|
| Being-H0.7: A Latent World-Action Model
from Egocentric Videos |
Being-H0.7 |
2026 |
2605.00078 |
|
arXiv
Code
Opt
|
|
| Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning |
cosmos-policy |
2026 |
2601.16163 |
manipulation |
arXiv
Code
Opt
|
|
| FACT: Failure-Aware Causal Training for World-Action Models |
FACT |
2026 |
2608.10232 |
|
arXiv
Code
Opt
|
|
| Fast-WAM: Do World Action Models Need Test-time Future Imagination? |
Fast-WAM |
2026 |
2603.16666 |
manipulation |
arXiv
Code
Opt
|
|
| From Imagined Futures to Executable Actions: Mixture of Latent Actions forRobot Manipulation |
MoLA |
2026 |
2605.12167 |
|
arXiv
Code
Opt
|
|
| MotionWAM: Towards Foundation World Action
Models for Real-Time Humanoid Loco-Manipulation |
MotionWAM |
2026 |
2606.09215 |
|
arXiv
Code
Opt
|
|
| ViVa: A Video-Generative Value Model for Robot
Reinforcement Learning |
ViVa |
2026 |
2604.08168 |
|
arXiv
Code
Opt
|
|
| WALL-WM: Carving World Action Modeling
at the Event Joints |
WALL-WM |
2026 |
2606.01955 |
|
arXiv
Code
Opt
|
|
| When to Trust Imagination: Adaptive Action
Execution for World Action Models |
FFDC |
2026 |
2605.06222 |
|
arXiv
Code
Opt
|
|