LibriVoxDeEn: a corpus for German-to-English speech translation and speech recognition
We present a corpus of sentence-aligned triples of German audio, German text, and English translation, based on German audio books. The corpus consists of over 100 hours of audio material and over 50k parallel sentences. The audio data is read speech and thus low in disfluencies. The quality of audi...
Saved in:
| Main Authors: | , , , |
|---|---|
| Format: | Article (Journal) Chapter/Article |
| Language: | English |
| Published: |
18 Oct 2018
|
| In: |
Arxiv
|
| Online Access: | Verlag, Volltext: http://arxiv.org/abs/1910.07924 |
| Author Notes: | Benjamin Beilharz, Xin Sun, Sariya Karimova, Stefan Riezler |
Search Result 1
LibriVoxDeEn: a corpus for German-to-English speech translation and speech recognition
Universität 2019-10-21
Database
Research Data
Online Resource