Skip to Main Content (Press Enter)

Logo UNIMORE
  • ×
  • Home
  • Corsi
  • Insegnamenti
  • Professioni
  • Persone
  • Pubblicazioni
  • Strutture
  • Terza Missione
  • Attività
  • Competenze

UNI-FIND
Logo UNIMORE

|

UNI-FIND

unimore.it
  • ×
  • Home
  • Corsi
  • Insegnamenti
  • Professioni
  • Persone
  • Pubblicazioni
  • Strutture
  • Terza Missione
  • Attività
  • Competenze
  1. Pubblicazioni

An automatic caption alignment mechanism for off-the-shelf speech recognition technologies

Articolo
Data di Pubblicazione:
2014
Citazione:
An automatic caption alignment mechanism for off-the-shelf speech recognition technologies / Federico, Maria; Furini, Marco. - In: MULTIMEDIA TOOLS AND APPLICATIONS. - ISSN 1380-7501. - STAMPA. - 72:1(2014), pp. 21-40. [10.1007/s11042-012-1318-3]
Abstract:
With a growing number of online videos, many producers feel the need to use video captions in order to expand content accessibility and face two main issues: production and alignment of the textual transcript. Both activities are expensive either for the high labor of human resources or for the employment of dedicated software. In this paper, we focus on caption alignment and we propose a novel, automatic, simple and low-cost mechanism that does not require human transcriptions or special dedicated software to align captions. Our mechanism uses a unique audio markup and intelligently introduces copies of it into the audio stream before giving it to an off-the-shelf automatic speech recognition (ASR) application; then it transforms the plain transcript produced by the ASR application into a timecoded transcript, which allows video players to know when to display every single caption while playing out the video. The experimental study evaluation shows that our proposal is effective in producing timecoded transcripts and therefore it can be helpful to expand video content accessibility.
Tipologia CRIS:
Articolo su rivista
Keywords:
Automatic caption alignment; speech recognition
Elenco autori:
Federico, Maria; Furini, Marco
Autori di Ateneo:
FURINI Marco
Link alla scheda completa:
https://iris.unimore.it/handle/11380/909690
Pubblicato in:
MULTIMEDIA TOOLS AND APPLICATIONS
Journal
  • Utilizzo dei cookie

Realizzato con VIVO | Designed by Cineca | 26.5.0.0