Guided Reality: Generating Visually-Enriched AR Task Guidance with LLMs and Vision Models

Guided Reality: Generating Visually-Enriched AR Task Guidance with LLMs and Vision Models

Ada Yi Zhao, Aditya Gunturu, Ellen Yi-Luen Do, Ryo Suzuki  

The ACM Symposium on User Interface Software and Technology (UIST 2025)

Links:     PDF   Video   ACM DL   arXiv


Abstract

Guided Reality automatically generates visually enriched, spatially situated AR instructions. The system combines large language models and vision models to create task steps, select suitable visual guidance, locate interaction points, and embed dynamic guidance into the physical environment.

Publication

Ada Yi Zhao, Aditya Gunturu, Ellen Yi-Luen Do, and Ryo Suzuki. 2025. Guided Reality: Generating Visually-Enriched AR Task Guidance with LLMs and Vision Models. Proceedings of the 38th Annual ACM Symposium on User Interface Software and Technology. ACM, New York, NY, USA.
DOI: https://doi.org/10.1145/3746059.3747784


Slide

coming soon