A PHP Error was encountered

Severity: Warning

Message: file_get_contents(https://...@pubfacts.com&api_key=b8daa3ad693db53b1410957c26c9a51b4908&a=1): Failed to open stream: HTTP request failed! HTTP/1.1 429 Too Many Requests

Filename: helpers/my_audit_helper.php

Line Number: 176

Backtrace:

File: /var/www/html/application/helpers/my_audit_helper.php
Line: 176
Function: file_get_contents

File: /var/www/html/application/helpers/my_audit_helper.php
Line: 250
Function: simplexml_load_file_from_url

File: /var/www/html/application/helpers/my_audit_helper.php
Line: 3122
Function: getPubMedXML

File: /var/www/html/application/controllers/Detail.php
Line: 575
Function: pubMedSearch_Global

File: /var/www/html/application/controllers/Detail.php
Line: 489
Function: pubMedGetRelatedKeyword

File: /var/www/html/index.php
Line: 316
Function: require_once

VIIDA and InViDe: computational approaches for generating and evaluating inclusive image paragraphs for the visually impaired. | LitMetric

VIIDA and InViDe: computational approaches for generating and evaluating inclusive image paragraphs for the visually impaired.

Disabil Rehabil Assist Technol

Department of Informatics, Universidade Federal de Viçosa - UFV, Viçosa, Brazil.

Published: December 2024

AI Article Synopsis

  • Existing image description methods for blind or low vision individuals are often inadequate, either oversimplifying visuals into short captions or overwhelming users with lengthy descriptions.
  • VIIDA is introduced as a new procedure to enhance image description specifically for webinar scenes, along with InViDe, a metric for evaluating these descriptions based on accessibility for BLV people.
  • Utilizing advanced tech like a multimodal Visual Question Answering model and Natural Language Processing, VIIDA effectively creates descriptions closely matching image content, while InViDe provides insights into the effectiveness of various methods, fostering further development in Assistive Technologies.

Article Abstract

Background: Existing image description methods when used as Assistive Technologies often fall short in meeting the needs of blind or low vision (BLV) individuals. They tend to either compress all visual elements into brief captions, create disjointed sentences for each image region, or provide extensive descriptions.

Purpose: To address these limitations, we introduce VIIDA, a procedure aimed at the Visually Impaired which implements an Image Description Approach, focusing on webinar scenes. We also propose InViDe, an Inclusive Visual Description metric, a novel approach for evaluating image descriptions targeting BLV people.

Methods: We reviewed existing methods and developed VIIDA by integrating a multimodal Visual Question Answering model with Natural Language Processing (NLP) filters. A scene graph-based algorithm was then applied to structure final paragraphs. By employing NLP tools, InViDe conducts a multicriteria analysis based on accessibility standards and guidelines.

Results: Experiments statistically demonstrate that VIIDA generates descriptions closely aligned with image content as well as human-written linguistic features, and that suit BLV needs. InViDe offers valuable insights into the behaviour of the compared methods - among them, state-of-the-art methods based on Large Language Models - across diverse criteria.

Conclusion: VIIDA and InViDe emerge as efficient Assistive Technologies, combining Artificial Intelligence models and computational/mathematical techniques to generate and evaluate image descriptions for the visually impaired with low computational costs. This work is anticipated to inspire further research and application development in the domain of Assistive Technologies. Our codes are publicly available at: https://github.com/daniellf/VIIDA-and-InViDe.

Download full-text PDF

Source
http://dx.doi.org/10.1080/17483107.2024.2437567DOI Listing

Publication Analysis

Top Keywords

visually impaired
12
assistive technologies
12
viida invide
8
image description
8
image descriptions
8
image
7
viida
5
invide computational
4
computational approaches
4
approaches generating
4

Similar Publications

Want AI Summaries of new PubMed Abstracts delivered to your In-box?

Enter search terms and have AI summaries delivered each week - change queries or unsubscribe any time!