Amb motiu del tancament d'estiu, la validació de documents es reprendrà a partir del 28 d'agost de 2026. Disculpeu les molèsties.
Con motivo del cierre de verano, la validación de documentos se reanudará a partir del 28 de agosto de 2026. Disculpad las molestias
Due to the summer closure, document validation will resume starting August 28, 2026. We apologize for any inconvenience.

Document type

Article

Version

Accepted version

Publication date

All rights reserved

Please use this identifier to cite or link to this item: https://hdl.handle.net/2445/182546

Corpora compilation for prosody-informed speech processing

Journal Title

Director/Tutor

Journal ISSN

Volume Title

Abstract

Research on speech technologies necessitates spoken data, which is usually obtained through read recorded speech, and specifically adapted to the research needs. When the aim is to deal with the prosody involved in speech, the available data must reflect natural and conversational speech, which is usually costly and difficult to get. This paper presents a machine learning-oriented toolkit for collecting, handling, and visualization of speech data, using prosodic heuristic. We present two corpora resulting from these methodologies: PANTED corpus, containing 250 h of English speech from TED Talks, and Heroes corpus containing 8 h of parallel English and Spanish movie speech. We demonstrate their use in two deep learning-based applications: punctuation restoration and machine translation. The presented corpora are freely available to the research community.

Citation

Citation

ÖKTEM, Alp, FARRÚS, Mireia and BONAFONTE, Antonio. Corpora compilation for prosody-informed speech processing. Language Resources And Evaluation. 2021. Vol. 55, num. 4, pags. 925-946. ISSN 1574-020X. [consulted: 11 of August of 2026]. Available at: https://hdl.handle.net/2445/182546

Export metadata

JSON - METS

Share record