Skip to Main Content (Press Enter)

Logo UNIME
  • ×
  • Home
  • Corsi
  • Insegnamenti
  • Professioni
  • Persone
  • Pubblicazioni
  • Strutture
  • Terza Missione
  • Competenze

Competenze e Professionalità
Logo UNIME

|

UNIFIND - Competenze e Professionalità

unime.it
  • ×
  • Home
  • Corsi
  • Insegnamenti
  • Professioni
  • Persone
  • Pubblicazioni
  • Strutture
  • Terza Missione
  • Competenze
  1. Pubblicazioni

A Voice User Interface on the Edge for People with Speech Impairments

Articolo
Data di Pubblicazione:
2024
Abstract:
Nowadays, fine-tuning has emerged as a powerful technique in machine learning, enabling models to adapt to a specific domain by leveraging pre-trained knowledge. One such application domain is automatic speech recognition (ASR), where fine-tuning plays a crucial role in addressing data scarcity, especially for languages with limited resources. In this study, we applied fine-tuning in the context of atypical speech recognition, focusing on Italian speakers with speech impairments, e.g., dysarthria. Our objective was to build a speaker-dependent voice user interface (VUI) tailored to their unique needs. To achieve this, we harnessed a pre-trained OpenAI's Whisper model, which has been exposed to vast amounts of general speech data. However, to adapt it specifically for disordered speech, we fine-tuned it using our private corpus including 65 K voice recordings contributed by 208 speech-impaired individuals globally. We exploited three variants of the Whisper model (small, base, tiny), and by evaluating their relative performance, we aimed to identify the most accurate configuration for handling disordered speech patterns. Furthermore, our study dealt with the local deployment of the trained models on edge computing nodes, with the aim to realize custom VUIs for persons with impaired speech.
Tipologia CRIS:
14.a.1 Articolo su rivista
Keywords:
automatic speech recognition; whisper; dysarthria; atypical speech; transformer; edge; assistive technology; AI
Elenco autori:
Mulfari, Davide; Villari, Massimo
Autori di Ateneo:
MULFARI Davide
VILLARI Massimo
Link alla scheda completa:
https://iris.unime.it/handle/11570/3311550
Pubblicato in:
ELECTRONICS
Journal
  • Informazioni
  • Assistenza
  • Accessibilità
  • Privacy
  • Utilizzo dei cookie
  • Note legali

Realizzato con VIVO | Designed by Cineca | 25.10.4.0