Alternative descriptions of digital images have always been an accessibility issue for screen reader users. Over time, numerous guidelines have been proposed in the literature, but the problem still exists. Recently, artificial intelligence (AI) has been introduced in digital applications to support visually impaired people in getting information about the world around them. In this way, such applications become a digital assistant for people with visual impairments. Increasingly, generative AI is being exploited to create accessible content for visually impaired people. In the education field, image description can play a crucial role in understanding even scientific content. For this reason, alternative descriptions should be accurate and educational-oriented. In this work, we investigate whether existing AI-based tools on the market are mature for describing images related to scientific content. Five AI-based tools were used to test the generated descriptions of four STEM images chosen for this preliminary study. Results indicate that answers are prompt and context dependent, and this technology can certainly support blind people in everyday tasks; but for STEM educational content more effort is required for delivering accessible and effective descriptions, supporting students in satisfying and accurate image exploration.

Is Generative AI Mature for Alternative Image Descriptions of STEM Content?

Buzzi, Marina;Galesi, Giulio;Leporini, Barbara;Nicotera, Annalisa
2024-01-01

Abstract

Alternative descriptions of digital images have always been an accessibility issue for screen reader users. Over time, numerous guidelines have been proposed in the literature, but the problem still exists. Recently, artificial intelligence (AI) has been introduced in digital applications to support visually impaired people in getting information about the world around them. In this way, such applications become a digital assistant for people with visual impairments. Increasingly, generative AI is being exploited to create accessible content for visually impaired people. In the education field, image description can play a crucial role in understanding even scientific content. For this reason, alternative descriptions should be accurate and educational-oriented. In this work, we investigate whether existing AI-based tools on the market are mature for describing images related to scientific content. Five AI-based tools were used to test the generated descriptions of four STEM images chosen for this preliminary study. Results indicate that answers are prompt and context dependent, and this technology can certainly support blind people in everyday tasks; but for STEM educational content more effort is required for delivering accessible and effective descriptions, supporting students in satisfying and accurate image exploration.
2024
978-989-758-718-4
File in questo prodotto:
File Dimensione Formato  
Leporini+et+al_WEBIST_2024_preprint.pdf

accesso aperto

Tipologia: Documento in Pre-print
Licenza: Creative commons
Dimensione 348.66 kB
Formato Adobe PDF
348.66 kB Adobe PDF Visualizza/Apri

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11568/1365971
Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus 4
  • ???jsp.display-item.citation.isi??? ND
social impact