Papers by Chieko Nishimura

1 papers
Text360Nav: 360-Degree Image Captioning Dataset for Urban Pedestrians Navigation (2024.lrec-main)

Copied to clipboard

Challenge: Existing image captioning datasets focus on the overall image description and lack detailed scene descriptions, overlooking features for pedestrians walking on urban streets.
Approach: They develop a dataset to provide textual feedback from 360-degree camera images to visually impaired pedestrians . they generate meaningful captions focusing on obstacles on the streets .
Outcome: The proposed dataset provides textual feedback from machinery visual perception to visually impaired individuals and distracted pedestrians . the results show that the models trained with the dataset can generate meaningful captions focusing on street objects and obstacles in urban scenes .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations