OPUS 4 | Search

14 search hits

1 to 10

Sort by

The New IDS Corpus Analysis Platform: Challenges and Prospects (2012)

Bański, Piotr ; Fischer, Peter M. ; Frick, Elena ; Ketzan, Erik ; Kupietz, Marc ; Schnober, Carsten ; Schonefeld, Oliver ; Witt, Andreas

The present article describes the first stage of the KorAP project, launched recently at the Institut für Deutsche Sprache (IDS) in Mannheim, Germany. The aim of this project is to develop an innovative corpus analysis platform to tackle the increasing demands of modern linguistic research. The platform will facilitate new linguistic findings by making it possible to manage and analyse primary data and annotations in the petabyte range, while at the same time allowing an undistorted view of the primary linguistic data, and thus fully satisfying the demands of a scientific tool. An additional important aim of the project is to make corpus data as openly accessible as possible in light of unavoidable legal restrictions, for instance through support for distributed virtual corpora, user-defined annotations and adaptable user interfaces, as well as interfaces and sandboxes for user-supplied analysis applications. We discuss our motivation for undertaking this endeavour and the challenges that face it. Next, we outline our software implementation plan and describe development to-date.

Pynchon Nods: Proust in Gravity’s Rainbow (2012)

Ketzan, Erik

This paper argues that Pynchon may allude to Marcel Proust through the character Marcel in Part 4 of Gravity's Rainbow and, if so, what that could mean. I trace the textual clues that relate to Proust and analyze what Pynchon may be saying about a fellow great experimental writer.

Les autorisations tacites - une révolution silencieuse en droit d'auteur numérique. Perspectives étasunienne, Allemand et Francaise (2013)

Beurskens, Michael ; Kamocki, Paweł ; Ketzan, Erik

Creative commons and language resources: general issues and what's new in CC 4.0 (2014)

Kamocki, Paweł ; Ketzan, Erik

Lizenzauswahlwerkzeuge für die digitalen Geisteswissenschaften (2016)

Kamocki, Paweł ; Ketzan, Erik ; Witt, Andreas

Guidelines for Building Language Corpora Under German Law. Guidelines by the DFG Review Board on Linguistics (2017)

Ketzan, Erik ; Wildgans, Julia ; Weitzmann, John

The possibilities of re-use and archiving of spoken and written corpora are affected by personality rights (depending on legal tradition also called: the right of publicity), copyright law and data protection / privacy laws. These recommendations include information about legal aspects which should be considered while creating corpora to ensure the greatest archivability and re-usability possible in compliance with current laws. The information compiled here shall serve researchers who plan to create corpora or who are involved in evaluation of such measures as a guideline. This information is not exhaustive or to be considered as legal advice. Researchers should consult institutional legal departments and management before making legally relevant decisions. That said, further legal expertise should be sought if possible as early as project planning phases.

New exceptions for Text and Data Mining and their possible impact on the CLARIN infrastructure (2018)

Kamocki, Pawel ; Ketzan, Erik ; Wildgans, Julia ; Witt, Andreas

The proposed paper discusses new exceptions for Text and Data Mining that have recently been adopted in some EU Member States, and probably will soon be adopted also at the EU level. These exceptions are of great significance for language scientists, as they exempt those who compile corpora from the obligation to obtain authorisation from rightholders. However, corpora compiled on the basis of such exceptions cannot be freely shared, which in a long run may have serious consequences for Open Science and the functioning of research infrastructure such as CLARIN ERIC.

Toward a CLARIN Data Protection Code of Conduct (2018)

Kamocki, Pawel ; Ketzan, Erik ; Wildgans, Julia ; Witt, Andreas

This abstract discusses the possibility to adopt a CLARIN Data Protection Code of Conduct pursuant art. 40 of the General Data Protection Regulation. Such a code of conduct would have important benefits for the entire language research community. The final section of this abstract proposes a roadmap to the CLARIN Data Protection Code of Conduct, listing various stages of its drafting and approval procedures.

Das neue "Gesetz zur Angleichung des Urheberrechts an die aktuellen Erfordernisse der Wissensgesellschaft" und seine Auswirkungen für Digital Humanities (2018)

Kamocki, Pawel ; Ketzan, Erik ; Wildgans, Julia ; Witt, Andreas

CLARIN Legal Information Plattformen und Legal Helpdesk (2018)

Kamocki, Pawel ; Ketzan, Erik ; Wildgans, Julia ; Witt, Andreas