Unifying terrain awareness for the visually impaired through real-time semantic segmentation

Navigational assistance aims to help visually-impaired people to ambulate the environment safely and independently. This topic becomes challenging as it requires detecting a wide variety of scenes to provide higher level assistive awareness. Vision-based technologies with monocular detectors or dept...

ver descrição completa

Detalhes bibliográficos
Autores: Yang, Kailun, Wang, Kaiwei, Bergasa Pascual, Luis Miguel|||0000-0002-0087-3077, Romera Carmena, Eduardo|||0000-0001-6250-6160, Hu, Weijian, Sun, Dongming, Sun, Junwei, Cheng, Ruiqi, Chen, TIanxue, López Guillén, María Elena
Formato: artículo
Fecha de publicación:2018
País:España
Recursos:Universidad de Alcalá (UAH)
Repositorio:e_Buah Biblioteca Digital Universidad de Alcalá
Idioma:inglés
OAI Identifier:oai:ebuah.uah.es:10017/43213
Acesso em linha:http://hdl.handle.net/10017/43213
https://dx.doi.org/10.3390/s18051506
Access Level:acceso abierto
Palavra-chave:Navigation assistance
Semantic segmentation
Traversability awareness
Obstacle avoidance
RGB-D sensor
Visually-impaired people
Electrónica
Electronics
Descrição
Resumo:Navigational assistance aims to help visually-impaired people to ambulate the environment safely and independently. This topic becomes challenging as it requires detecting a wide variety of scenes to provide higher level assistive awareness. Vision-based technologies with monocular detectors or depth sensors have sprung up within several years of research. These separate approaches have achieved remarkable results with relatively low processing time and have improved the mobility of impaired people to a large extent. However, running all detectors jointly increases the latency and burdens the computational resources. In this paper, we put forward seizing pixel-wise semantic segmentation to cover navigation-related perception needs in a unified way. This is critical not only for the terrain awareness regarding traversable areas, sidewalks, stairs and water hazards, but also for the avoidance of short-range obstacles, fast-approaching pedestrians and vehicles. The core of our unification proposal is a deep architecture, aimed at attaining efficient semantic understanding. We have integrated the approach in a wearable navigation system by incorporating robust depth segmentation. A comprehensive set of experiments prove the qualified accuracy over state-of-the-art methods while maintaining real-time speed. We also present a closed-loop field test involving real visually-impaired users, demonstrating the effectivity and versatility of the assistive framework.