Guided Reality automatically generates visually enriched, spatially situated AR instructions. The system combines large language models and vision models to create task steps, select suitable visual guidance, locate interaction points, and embed dynamic guidance into the physical environment.
Ada Yi Zhao, Aditya Gunturu, Ellen Yi-Luen Do, and Ryo Suzuki. 2025. Guided Reality: Generating Visually-Enriched AR Task Guidance with LLMs and Vision Models. Proceedings of the 38th Annual ACM Symposium on User Interface Software and Technology. ACM, New York, NY, USA.
DOI: https://doi.org/10.1145/3746059.3747784
coming soon