COLING 2024main2 citations

Text360Nav: 360-Degree Image Captioning Dataset for Urban Pedestrians Navigation

Chieko Nishimura, Shuhei Kurita, Yohei Seki

Abstract

Text feedback from urban scenes is a crucial tool for pedestrians to understand surroundings, obstacles, and safe pathways. However, existing image captioning datasets often concentrate on the overall image description and lack detailed scene descriptions, overlooking features for pedestrians walking on urban streets. We developed a new dataset to assist pedestrians in urban scenes using 360-degree camera images. Through our dataset of Text360Nav, we aim to provide textual feedback from machinery visual perception such as 360-degree cameras to visually impaired individuals and distracted pedestrians navigating urban streets, including those engrossed in their smartphones while walking. In experiments, we combined our dataset with multimodal generative models and observed that models trained with our dataset can generate textual descriptions focusing on street objects and obstacles that are meaningful in urban scenes in both quantitative and qualitative analyses, thus supporting the effectiveness of our dataset for urban pedestrian navigation.

BibTeX
@inproceedings{nishimura-etal-2024-text360nav,
    title = "{T}ext360{N}av: 360-Degree Image Captioning Dataset for Urban Pedestrians Navigation",
    author = "Nishimura, Chieko  and
      Kurita, Shuhei  and
      Seki, Yohei",
    editor = "Calzolari, Nicoletta  and
      Kan, Min-Yen  and
      Hoste, Veronique  and
      Lenci, Alessandro  and
      Sakti, Sakriani  and
      Xue, Nianwen",
    booktitle = "Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024)",
    month = may,
    year = "2024",
    address = "Torino, Italia",
    publisher = "ELRA and ICCL",
    url = "https://aclanthology.org/2024.lrec-main.1371/",
    pages = "15783--15788"
}
Text360Nav: 360-Degree Image Captioning Dataset for Urban Pedestrians Navigation · COLING 2024