Refine
Year of publication
- 2016 (5) (remove)
Document Type
- Conference Proceeding (3)
- Article (1)
- Part of a Book (1)
Has Fulltext
- yes (5)
Is part of the Bibliography
- no (5)
Keywords
- Korpus <Linguistik> (5)
- Computerlinguistik (2)
- Textlinguistik (2)
- Aufsatzsammlung (1)
- Deutsch (1)
- Institut für Deutsche Sprache <Mannheim> (1)
- Kontrastive Linguistik (1)
- Korpusanalyseplattform (KorAP) (1)
- Rechtschreibung (1)
- Rumänisch (1)
Publicationstate
Reviewstate
- (Verlags)-Lektorat (2)
- Peer-Review (1)
Constructing a Corpus
(2016)
This paper introduces the recently started DRuKoLA-project that aims at providing mechanisms to flexibly draw virtual comparable corpora from the German Reference Corpus DeReKo and the Reference Corpus of Contemporary Romanian Language CoRoLa in order to use these virtual corpora as empirical basis for contrastive linguistic research.
Editorial
(2016)
KorAP is a corpus search and analysis platform, developed at the Institute for the German Language (IDS). It supports very large corpora with multiple annotation layers, multiple query languages, and complex licensing scenarios. KorAP’s design aims to be scalable, flexible, and sustainable to serve the German Reference Corpus DEREKO for at least the next decade. To meet these requirements, we have adopted a highly modular microservice-based architecture. This paper outlines our approach: An architecture consisting of small components that are easy to extend, replace, and maintain. The components include a search backend, a user and corpus license management system, and a web-based user frontend. We also describe a general corpus query protocol used by all microservices for internal communications. KorAP is open source, licensed under BSD-2, and available on GitHub.