AI GlossaryㅇTechnical words in the news
World Action Model
An AI modeling approach in which a robot first predicts physical changes visually before deciding on an action, rather than relying on language.
In plain words
A World Action Model is a way for a robot to understand the world not through photos and text, but by watching video and first picturing in its 'mind' how things will physically change when touched or moved — before it actually moves. It's like the model has already learned from video how a cup gets dented depending on finger shape when gripped, or how a towel folds when grabbed. Then, when the actual robot arm moves, it decides its next action based on this internal prediction.
Traditional robot 'brains' were built by taking a model trained to understand images and text together, and then bolting on an arm-control function afterward. These models are good at describing a scene in words, but weak at foreseeing how an object will actually deform once touched. A World Action Model swaps out that weakness for a prediction ability learned directly from video.
How it shows up in the news
In the article, it appears like this: NVIDIA introduced the World Action Model (WAM) as a new foundation for robot policies. A common misunderstanding is treating this as the name of a single specific product, but it actually functions more as an umbrella term for an entire class of models that decide actions by predicting physical changes from video.
See also
Stories using this term
- NVIDIA proposes robot policy trained on video instead of vision-language modelsAI · 2026.08.09
- World Labs Turns One Robot Task Into Thousands of Simulated RunsAI · 2026.08.15
- Google Unveils Gemini Robotics 2 for Full-Body Robot ControlAI · 2026.08.04
- Alibaba unveils Wan-Animate-2, adds real-time streaming generationAI · 2026.08.11
- Sakana AI Brings Recursive Self-Improvement to 'Physical AI' for Robots That Learn on Their OwnAI · 2026.08.11
- Adding a 'mind' variable to world models boosted accuracy from 63 to 88AI · 2026.08.24
