Papers by Chieko Nishimura
Text360Nav: 360-Degree Image Captioning Dataset for Urban Pedestrians Navigation (2024.lrec-main)
Copied to clipboard
| Challenge: | Existing image captioning datasets focus on the overall image description and lack detailed scene descriptions, overlooking features for pedestrians walking on urban streets. |
| Approach: | They develop a dataset to provide textual feedback from 360-degree camera images to visually impaired pedestrians . they generate meaningful captions focusing on obstacles on the streets . |
| Outcome: | The proposed dataset provides textual feedback from machinery visual perception to visually impaired individuals and distracted pedestrians . the results show that the models trained with the dataset can generate meaningful captions focusing on street objects and obstacles in urban scenes . |