AIIC AI Intelligence Centre

SOURCE-LINKED INTELLIGENCE

Generative Tutorial: Towards Live Contextualized Visual Instructions for Physical Tasks

arXiv · AI, language, vision and robotics · article · Sep 21, 2026 · UTC

Visual instructions for physical tasks are typically authored in one context and followed in another, requiring users to translate demonstrated tools, materials, and spatial relationships into their own environment. We introduce Generative Tutorial, a conceptual framework for live visual instruction that depicts intended outcomes and actions within the user's environment and task flow. A formative evaluation of state-of-the-art image and video generation identifies failures and potential benefits across 15 physical tasks. Drawing on these findings, we build an augmented-reality prototype syste

Read original source ↗ Open in workspace

recordType
paper
region
Global

Evidence & attribution

First collected: 2026-09-23T06:11:12.848Z. This is not the publication date.