Conference contribution
(Conference Contribution)


LMELectures: a Multimedia Corpus of Acedemic Spoken English


Publication Details
Author(s): Riedhammer K, Gropp M, Bocklet T, Hönig F, Nöth E, Steidl S
Editor(s): ISCA SIG on Speech and Language in Multimedia IEEE SIG on Audio and Speech Processing in Multimedia
Publisher: CEUR-WS
Publication year: 2013
Conference Proceedings Title: Proceedings of the First Workshop on Speech, Language and Audio in Multimedia
Pages range: 102-107

Event details
Event: SLAM 2013 - First Workshop on Speech, Language and Audio in Multimedia
Event location: Marseille
Start date of the event: 22/08/2013
End date of the event: 23/08/2013

Abstract

This paper describes the acquisition, transcription and annotation of a multi-media corpus of academic spoken English, the LMELectures. It consists of two lecture series that were read in the summer term 2009 at the computer science department of the University of Erlangen- Nuremberg, covering topics in pattern analysis, machine learning and interventional medical image processing. In total, about 40 hours of high-definition audio and video of a single speaker was acquired in a constant recording environment. In addition to the recordings, the presentation slides are available in machine readable (PDF) format. The manual annotations include a suggested segmentation into speech turns and a complete manual transcription that was done using BLITZSCRIBE2, a new tool for the rapid transcription. For one lecture series, the lecturer assigned key words to each recordings; one recording of that series was further annotated with a list of ranked key phrases by five human annotators each. The corpus is available for non-commercial purpose upon request.



How to cite
APA: Riedhammer, K., Gropp, M., Bocklet, T., Hönig, F., Nöth, E., & Steidl, S. (2013). LMELectures: a Multimedia Corpus of Acedemic Spoken English. In ISCA SIG on Speech and Language in Multimedia IEEE SIG on Audio and Speech Processing in Multimedia (Eds.), Proceedings of the First Workshop on Speech, Language and Audio in Multimedia (pp. 102-107). CEUR-WS.

MLA: Riedhammer, Korbinian, et al. "LMELectures: a Multimedia Corpus of Acedemic Spoken English." Proceedings of the SLAM 2013 - First Workshop on Speech, Language and Audio in Multimedia, Marseille Ed. ISCA SIG on Speech and Language in Multimedia IEEE SIG on Audio and Speech Processing in Multimedia, CEUR-WS, 2013. 102-107.

BibTeX: Download
Share link
Last updated on 2017-06-27 at 02:31
PDF downloaded successfully