Refine
Year of publication
- 2015 (3) (remove)
Document Type
- Book (1)
- Part of a Book (1)
- Conference Proceeding (1)
Keywords
- Deutsch (3)
- Aussprache (2)
- Sprachvariante (2)
- Akustische Phonetik (1)
- Annotation (1)
- Gesprochene Sprache (1)
- Korpus <Linguistik> (1)
- Metadaten (1)
- Regionalsprache (1)
- Sprachatlas (1)
Publicationstate
Reviewstate
- Peer-Review (1)
Publisher
Ph@ttSessionz and Deutsch heute are two large German speech databases. They were created for different purposes: Ph@ttSessionz to test Internet-based recordings and to adapt speech recognizers to the voices of adolescent speakers, Deutsch heute to document regional variation of German. The databases differ in their recording technique, the selection of recording locations and speakers, elicitation mode, and data processing.
In this paper, we outline how the recordings were performed, how the data was processed and annotated, and how the two databases were imported into a single relational database system. We present acoustical measurements on the digit items of both databases. Our results confirm that the elicitation technique affects the speech produced, that f0 is quite comparable despite different recording procedures, and that large speech technology databases with suitable metadata may well be used for the analysis of regional variation of speech.