OPUS 4 | Sprache im 20. Jahrhundert. Gegenwartssprache

Sprache im 20. Jahrhundert. Gegenwartssprache

42 search hits

1 to 10

Sort by

Das Bild von der 'Sprache der DDR' in der alten Bundesrepublik oder: Haben sie so gesprochen? (2004)

Hellmann, Manfred W.

Thema erledigt - oder doch noch nicht? Was bleibt zu tun bei der Erforschung des DDR-Sprachgebrauchs? (2004)

Hellmann, Manfred W.

Deutsch-türkische Kontaktvarietäten. Am Beispiel der Sprache von deutsch-türkischen Jugendlichen (2004)

Kallmeyer, Werner ; Keim, Inken

Kommunikativer Umgang mit sozialen Grenzziehungen. Zur Analyse von Sprachstilen aus soziolinguistischer Perspektive (2004)

Kallmeyer, Werner

Multi-dimensional annotation of linguistic corpora for investigating information structure (2004)

Baumann, Stefan ; Brinckmann, Caren ; Hansen-Schirra, Silvia ; Kruijff, Geert-Jan ; Kruijff-Korbayová, Ivana ; Neumann, Stella ; Teich, Elke

We present the annotation of information structure in the MULI project. To learn more about the information structuring means in prosody, syntax and discourse, theory- independent features were defined for each level. We describe the features and illustrate them on an example sentence. To investigate the interplay of features, the representation has to allow for inspecting all three layers at the same time. This is realised by a stand-off XML mark-up with the word as the basic unit. The theory-neutral XML stand-off annotation allows integrating this resource with other linguistic resources such as the Tiger Treebank for German or the Penn treebank for English.

The MULI Project: Annotation and Analysis of Information Structure in German and English (2004)

Baumann, Stefan ; Brinckmann, Caren ; Hansen-Schirra, Silvia ; Kruijff, Geert-Jan ; Kruijff-Korbayová, Ivana ; Neumann, Stella ; Steiner, Erich ; Teich, Elke ; Uszkoreit, Hans

The goal of the MULI (MUltiLingual Information structure) project is to empirically analyse information structure in German and English newspaper texts. In contrast to other projects in which information structure is annotated and investigated (e.g. in the Prague Dependency Treebank, which mirrors the basic information about the topic-focus articulation of the sentence), we do not annotate theory-biased categories like topic-focus or theme-rheme. Trying to be as theory-independent as possible, we annotate those features which are relevant to information structure and on the basis of which typical patterns, co-occurrences or correlations can be determined. We distinguish between three annotation levels: syntax, discourse and prosody. The data is based on the TIGER Corpus for German and the Penn Treebank for English, since the existing information on part-of-speech and syntactic structure can be re-used for our purposes. The actual annotation of an English example sequence illustrates our choice of categories on each level. Their combination offers the possibility to investigate how information structure is realised and can be interpreted.

A Multilingual Phonological Resource Toolkit for Ubiquitous Speech Technology (2004)

Aioanei, Daniel ; Carson-Berndsen, Julie ; Geumann, Anja ; Kelly, Robert ; Neugebauer, Moritz ; Wilson, Stephen

This paper outlines the generation process of a specifi computational linguistic representation termed the Multilingual Time Map, conceptually a multi-tape finit state transducer encoding linguistic data at different levels of granularity. The fi st component acquires phonological data from syllable labeled speech data, the second component define feature profiles the third component generates feature hierarchies and augments the acquired data with the define feature profiles and the fourth component displays the Multilingual Time Map as a graph.

Das heutige Deutsch – Tendenzen und Wertungen (2004)

Stickel, Gerhard

Towards a new level of annotation detail of multilingual speech corpora (2004)

Geumann, Anja

The aim of this paper is to highlight the actual need for corpora that have been annotated based on acoustic information. The acoustic information should be coded in features or properties and is needed to inform further processing systems, i.e. to present a basis for a speech recognition system using linguistic information. Feature annotation of existing corpora in combination with segmental annotation can provide a powerful training material for speech recognition systems, but will as well challenge the further processing of features to segments and syllables. We present here the theoretical preliminaries for our multilingual feature extraction system, that we are currently working on.

Vorbemerkung (2004)

Keim, Inken

1 to 10

Open Access

Sprache im 20. Jahrhundert. Gegenwartssprache

Refine

Author

Year of publication

Document Type

Language

Has Fulltext

Is part of the Bibliography

Keywords

Publicationstate

Reviewstate

Publisher

42 search hits