Refine
Year of publication
- 2012 (118) (remove)
Document Type
- Part of a Book (57)
- Article (38)
- Part of Periodical (10)
- Book (7)
- Conference Proceeding (3)
- Other (2)
- Review (1)
Has Fulltext
- yes (118)
Keywords
- Deutsch (118) (remove)
Publicationstate
- Veröffentlichungsversion (31)
- Postprint (4)
- Zweitveröffentlichung (3)
Reviewstate
- (Verlags)-Lektorat (33)
- Peer-Review (5)
- Review-Status-unbekannt (1)
Publisher
- Institut für Deutsche Sprache (31)
- de Gruyter (24)
- Narr (7)
- Buske (3)
- De Gruyter (3)
- Akademie Verlag (2)
- Dudenverlag (2)
- Hempen (2)
- Lang (2)
- Stauffenburg (2)
In meiner 2010 erschienenen Dissertation „Migration, Sprache und Rassismus“ habe ich mit ethnografischen, gesprächsanalytischen und -rhetorischen Methoden den Kommunikationsstil von zwei akademischen Migrantenmilieus(„emanzipatorische Migranten“ und „akademische Europatürken“) in Deutschland untersucht. Die Studie war Teil des Projekts „Deutschtürkische Sprachvariation und die Herausbildung kommunikativer Stile in dominant türkischen Migrantengruppen“, das am Institut für Deutsche Sprache durchgeführt wurde.
The paper presents an XML schema for the representation of genres of computer-mediated communication (CMC) that is compliant with the encoding framework defined by the TEI. It was designed for the annotation of CMC documents in the project Deutsches Referenzkorpus zur internetbasierten Kommunikation (DeRiK), which aims at building a corpus on language use in the most popular CMC genres on the German-speaking Internet. The focus of the schema is on those CMC genres which are written and dialogic―such as forums, bulletin boards, chats, instant messaging, wiki and weblog discussions, microblogging on Twitter, and conversation on “social network” sites.
The schema provides a representation format for the main structural features of CMC discourse as well as elements for the annotation of those units regarded as “typical” for language use on the Internet. The schema introduces an element <posting>, which describes stretches of text that are sent to the server by a user at a certain point in time. Postings are the main constituting elements of threads and logfiles, which, in our schema, are the two main types of CMC macrostructures. For the microlevel of CMC documents (that is, the structure of the <posting> content), the schema introduces elements for selected features of Internet jargon such as emoticons, interaction words and addressing terms. It allows for easy anonymization of CMC data for purposes in which the annotated data are made publicly available and includes metadata which are necessary for referencing random excerpts from the data as references in dictionary entries or as results of corpus queries.
Documentation of the schema as well as encoding examples can be retrieved from the web at http://www.empirikom.net/bin/view/Themen/CmcTEI. The schema is meant to be a core model for representing CMC that can be modified and extended by others according to their own specific perspectives on CMC data. It could be a first step towards an integration of features for the representation of CMC genres into a future new version of the TEI Guidelines.
Conversation Analysis (CA) and Discursive Psychology (DP) reject the view that assumptions
about cognitive processes should be used to account for discursive phenomena. Instead, cognitive
issues are respecified as discursive phenomena. Discursive psychologists do this by
studying discursive practices of talking about mental phenomena and using mental predicates.
This approach is exemplified by a study of the use of constructions with German verstehen
(‘to understand’) in conversation. Some conversation analysts take another approach,
namely, inquiring into how participants display mental states in talk-in-interaction. This is
exemplified by a study of how grammatical constructions are used to display different types
of inferences drawn from a partner’s prior turn. It will be argued that the constructivist, antiessentialist
stance which CA and DP take with regard to cognition is a prosperous line of
research, which has much in its favor from a methodological point of view. However, it
can be shown that tacit assumptions about cognitive processes are still inevitable when
doing CA and DP. As a conclusion, the paper pleads for an enhanced awareness of how cognitive
processes come into play when analysing talk-in-interaction and it advocates the integration
of a more explicit cognitive perspective into research on talk-in-interaction.
Although most of the relevant dictionary productions of the recent past have relied on digital data and methods, there is little consensus on formats and standards. The Institute for Corpus Linguistics and Text Technology (ICLTT) of the Austrian Academy of Sciences has been conducting a number of varied lexicographic projects, both digitising print dictionaries and working on the creation of genuinely digital lexicographic data. This data was designed to serve varying purposes: machine-readability was only one. A second goal was interoperability with digital NLP tools. To achieve this end, a uniform encoding system applicable across all the projects was developed. The paper describes the constraints imposed on the content models of the various elements of the TEI dictionary module and provides arguments in favour of TEI P5 as an encoding system not only being used to represent digitised print dictionaries but also for NLP purposes.
This paper describes work in progress on I5, a TEI-based document grammar for the corpus holdings of the Institut für Deutsche Sprache (IDS) in Mannheim and the text model used by IDS in its work. The paper begins with background information on the nature and purposes of the corpora collected at IDS and the motivation for the I5 project (section 1). It continues with a description of the origin and history of the IDS text model (section 2), and a description (section 3) of the techniques used to automate, as far as possible, the preparation of the ODD file documenting the IDS text model. It ends with some concluding remarks (section 4). A survey of the additional features of the IDS-XCES realization of the IDS text model is given in an appendix.