Enhancing the Travel Experience for People with Visual Impairments through Multimodal Interaction: NaviGPT, A Real-Time AI-Driven Mobile Navigation System

He Zhang; Nicholas J Falletta; Jingyi Xie; Rui Yu; Sooyeon Lee; Syed Masum Billah; John M Carroll

doi:10.1145/3688828.3699636

Enhancing the Travel Experience for People with Visual Impairments through Multimodal Interaction: NaviGPT, A Real-Time AI-Driven Mobile Navigation System

GROUP ACM SIGCHI Int Conf Support Group Work. 2025;2025(Companion):29-35. doi: 10.1145/3688828.3699636. Epub 2025 Jan 12.

Authors

He Zhang¹, Nicholas J Falletta¹, Jingyi Xie¹, Rui Yu², Sooyeon Lee³, Syed Masum Billah¹, John M Carroll¹

Affiliations

¹ College of Information Sciences and Technology, The Pennsylvania State University, University Park, Pennsylvania, USA.
² Department of Computer Science and Engineering, University of Louisville, Louisville, KY, USA.
³ Ying Wu College of Computing, New Jersey Institute of Technology Newark, NJ, USA.

Abstract

Assistive technologies for people with visual impairments (PVI) have made significant advancements, particularly with the integration of artificial intelligence (AI) and real-time sensor technologies. However, current solutions often require PVI to switch between multiple apps and tools for tasks like image recognition, navigation, and obstacle detection, which can hinder a seamless and efficient user experience. In this paper, we present NaviGPT, a high-fidelity prototype that integrates LiDAR-based obstacle detection, vibration feedback, and large language model (LLM) responses to provide a comprehensive and real-time navigation aid for PVI. Unlike existing applications such as Be My AI and Seeing AI, NaviGPT combines image recognition and contextual navigation guidance into a single system, offering continuous feedback on the user's surroundings without the need for app-switching. Meanwhile, NaviGPT compensates for the response delays of LLM by using location and sensor data, aiming to provide practical and efficient navigation support for PVI in dynamic environments.

Keywords: AI-assisted tool; People with visual impairments; accessibility; disability; llm; mobile application; multimodal interaction; navigation; prototype.