Refine
Document Type
Language
- English (3) (remove)
Has Fulltext
- yes (3)
Keywords
- Annotation (1)
- Digital Humanities (1)
- Hypertext (1)
- Informationsstruktur (1)
- Korpus <Linguistik> (1)
- Sprachtypologie (1)
- Standardisierung (1)
- Strukturbaum (1)
- Syntax (1)
- Texttechnologie (1)
Publicationstate
- Postprint (1)
- Veröffentlichungsversion (1)
- Zweitveröffentlichung (1)
Reviewstate
- Peer-Review (2)
Publisher
- Universität Tübingen (3) (remove)
In 2010, ISO published a standard for syntactic annotation, ISO 24615:2010 (SynAF). Back then, the document specified a comprehensive reference model for the representation of syntactic annotations, but no accompanying XML serialisation. ISO’s subcommittee on language resource management (ISO TC 37/SC 4) is working on making the SynAF serialisation ISOTiger an additional part of the standard. This contribution addresses the current state of development of ISOTiger, along with a number of open issues on which we are seeking community feedback in order to ensure that ISOTiger becomes a useful extension to the SynAF reference model.
The motivation for this article is to describe a methodology for interrelating and analyzing language and theory-specific corpus data from various languages. As an example phenomeon we use information structure (IS, see [3]) in treebanks from three languages: Spanish, Korean and Japanese. Korean and Japanese are typologically close, while both are typologically different from Spanish. Therefore, the problem of annotating IS is that there are diverging language-specific formal linguistic means for the realization of IS-functions (like “topicalization / contrast”) on various levels like prosody, morphology and word-order. Hence, it is necessary to describe the relations between language-specific formal means and functional views on IS, and how to operationalize these relations for corpus analysis.