AIIC AI Intelligence Centre

SOURCE-LINKED INTELLIGENCE

AffordanceWAM: Affordance-Aware Joint World-Action Modeling for Robot Manipulation

arXiv · AI, language, vision and robotics · article · Sep 16, 2026 · UTC

Generalizable robot manipulation requires predicting how a scene will evolve, identifying where interactions are feasible, and determining how to act. Action-labeled robot videos directly supervise control but are costly and limited in diversity, whereas egocentric human videos capture diverse interactions but lack robot actions and differ in embodiment and appearance. We introduce AffordanceWAM, an affordance-aware generative World Action Model that represents object-centric spatiotemporal affordance through Scalar Affordance and Affordance Heatmap, within the generated future World. This rep

Read original source ↗ Open in workspace

recordType
paper
region
Global

Evidence & attribution

First collected: 2026-09-23T18:11:26.115Z. This is not the publication date.