An Integrated Mobile Vision System for Enhancing the Interaction of Blind and Low Vision Users with Their Surroundings [An Integrated Mobile Vision System for Enhancing the Interaction of Blind and Low Vision Users with Their Surroundings]

Chen, Jin; Ramnath, Satesh; Samaroo, Tyron; Maksakuli, Fani; Ruci, Arber; Sturdivant, E’edresha; Zhu, Zhigang

doi:10.5220/0011984400003497

Citation Details

An Integrated Mobile Vision System for Enhancing the Interaction of Blind and Low Vision Users with Their Surroundings [An Integrated Mobile Vision System for Enhancing the Interaction of Blind and Low Vision Users with Their Surroundings]

This paper presents a mobile-based solution that integrates 3D vision and voice interaction to assist people who are blind or have low vision to explore and interact with their surroundings. The key components of the system are the two 3D vision modules: the 3D object detection module integrates a deep-learning based 2D object detector with ARKit-based point cloud generation, and an interest direction recognition module integrates hand/finger recognition and ARKit-based 3D direction estimation. The integrated system consists of a voice interface, a task scheduler, and an instruction generator. The voice interface contains a customized user request mapping module that maps the user’s input voice into one of the four primary system operation modes (exploration, search, navigation, and settings adjustment). The task scheduler coordinates with two web services that host the two vision modules to allocate resources for computation based on the user request and network connectivity strength. Finally, the instruction generator computes the corresponding instructions based on the user request and results from the two vision modules. The system is capable of running in real time on mobile devices. We have shown preliminary experimental results on the performance of the voice to user request mapping module and the two vision modules. more »

Award ID(s):: 2131186 1827505 1737533

NSF-PAR ID:: 10440684

Author(s) / Creator(s):: Chen, Jin; Ramnath, Satesh; Samaroo, Tyron; Maksakuli, Fani; Ruci, Arber; Sturdivant, E’edresha; Zhu, Zhigang

Date Published:: 2023-01-01

Journal Name:: Proceedings of the 3rd International Conference on Image Processing and Vision Engineering

Page Range / eLocation ID:: 180 to 187

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
https://doi.org/10.5220/0011984400003497

More Like this