LibriVoxDeEn: a corpus for German-to-English speech translation and speech recognition

We present a corpus of sentence-aligned triples of German audio, German text, and English translation, based on German audio books. The corpus consists of over 100 hours of audio material and over 50k parallel sentences. The audio data is read speech and thus low in disfluencies. The quality of audi...

Full description

Saved in:
Bibliographic Details
Main Authors: Beilharz, Benjamin (Author) , Sun, Xin (Author) , Karimova, Sariya (Author) , Riezler, Stefan (Author)
Format: Article (Journal) Chapter/Article
Language:English
Published: 18 Oct 2018
In: Arxiv

Online Access:Verlag, Volltext: http://arxiv.org/abs/1910.07924
Get full text
Author Notes:Benjamin Beilharz, Xin Sun, Sariya Karimova, Stefan Riezler
Search Result 1

LibriVoxDeEn: a corpus for German-to-English speech translation and speech recognition by Beilharz, Benjamin (Author) , Sun, Xin (Author) ,

Universität 2019-10-21

Get full text
Database Research Data Online Resource