Describing Image Focused in Cognitive and Visual Details for Visually Impaired People: An Approach to Generating Inclusive Paragraphs

Daniel L. Fernandes, Marcos H. F. Ribeiro, Fabio R. Cerqueira, Fabio R. Cerqueira, Michel M. Silva

2022

Abstract

Several services for people with visual disabilities have emerged recently due to achievements in Assistive Technologies and Artificial Intelligence areas. Despite the growth in assistive systems availability, there is a lack of services that support specific tasks, such as understanding the image context presented in online content, e.g., webinars. Image captioning techniques and their variants are limited as Assistive Technologies as they do not match the needs of visually impaired people when generating specific descriptions. We propose an approach for generating context of webinar images combining a dense captioning technique with a set of filters, to fit the captions in our domain, and a language model for the abstractive summary task. The results demonstrated that we can produce descriptions with higher interpretability and focused on the relevant information for that group of people by combining image analysis methods and neural language models.

Download


Paper Citation


in Harvard Style

Fernandes D., Ribeiro M., Cerqueira F. and Silva M. (2022). Describing Image Focused in Cognitive and Visual Details for Visually Impaired People: An Approach to Generating Inclusive Paragraphs. In Proceedings of the 17th International Joint Conference on Computer Vision, Imaging and Computer Graphics Theory and Applications (VISIGRAPP 2022) - Volume 5: VISAPP; ISBN 978-989-758-555-5, SciTePress, pages 526-534. DOI: 10.5220/0010845700003124


in Bibtex Style

@conference{visapp22,
author={Daniel L. Fernandes and Marcos H. F. Ribeiro and Fabio R. Cerqueira and Michel M. Silva},
title={Describing Image Focused in Cognitive and Visual Details for Visually Impaired People: An Approach to Generating Inclusive Paragraphs},
booktitle={Proceedings of the 17th International Joint Conference on Computer Vision, Imaging and Computer Graphics Theory and Applications (VISIGRAPP 2022) - Volume 5: VISAPP},
year={2022},
pages={526-534},
publisher={SciTePress},
organization={INSTICC},
doi={10.5220/0010845700003124},
isbn={978-989-758-555-5},
}


in EndNote Style

TY - CONF

JO - Proceedings of the 17th International Joint Conference on Computer Vision, Imaging and Computer Graphics Theory and Applications (VISIGRAPP 2022) - Volume 5: VISAPP
TI - Describing Image Focused in Cognitive and Visual Details for Visually Impaired People: An Approach to Generating Inclusive Paragraphs
SN - 978-989-758-555-5
AU - Fernandes D.
AU - Ribeiro M.
AU - Cerqueira F.
AU - Silva M.
PY - 2022
SP - 526
EP - 534
DO - 10.5220/0010845700003124
PB - SciTePress