Publikationsserver des Instituts für Deutsche Sprache
Not a member yet
    11061 research outputs found

    basic level

    No full text

    Rediscovering a German Creole

    No full text
    „Unserdeutsch”, a creole spoken in a former German South Pacific colony, and what is now Papua New Guinea, is being extensively documented and studied by linguists for the first time. There is no time to lose, because after a chequered history the world's only German-based creole – long ignored – is facing extinction

    Unserdeutsch: an (a)typical creole?

    Get PDF
    Dieser Aufsatz diskutiert die Frage, inwieweit Unserdeutsch sich aus soziohistorischer und sprachstruktureller Perspektive in die Kategorie Kreolsprache einfügt. Als tertium comparationis dienen dabei Merkmale, die in der einschlägigen Literatur prominent als charakteristisch für Kreolsprachen angenommen werden. Es zeigt sich, dass Unserdeutsch trotz einer Reihe atypischer Entstehungsumstände, die auf den ersten Blick eine große strukturelle Nähe zum deutschen Superstrat, damit ein relativ akrolektales Kreol erwarten ließen, verhältnismäßig gut mit dem Muster eines Average Creole, wie es sich etwa aufgrund der Daten des „Atlas of Pidgin and Creole Language Structures“ (Michaelis et al. 2013) abzeichnet, harmoniert. Eine mögliche Erklärung findet diese augenfällige Diskrepanz in der primären Funktion von Unserdeutsch als Identitätsmarker und der linguistischen Struktur seiner Substratsprache Tok Pisin.This paper deals with the question to what extent Unserdeutsch fits into the category of creole languages based on a sociohistorical and structural perspective. For that purpose, features which are considered as typical for creole languages in the literature serve as tertium comparationis. Despite certain atypical circumstances in the genesis of Unserdeutsch, which lead to the assumption that Unserdeutsch could be considered as a relatively acrolectal creole due to the unlimited access of the speakers to the German superstrate, it will be shown that Unserdeutsch accords quite well with the pattern of an “Average Creole” compared to the data of the “Atlas of Pidgin and Creole Language Structures” (Michaelis et al. 2013). Possible explanations of this obvious discrepancy could be found in the primary function of Unserdeutsch as identity marker as well as in the linguistic structure of its substrate Tok Pisin

    megageil, mega geil, and voll mega: Intensification in YouTube comments

    Get PDF
    This paper analyses intensification in German digitally-mediated communication (DMC) using a corpus of YouTube comments written by young people (the NottDeuYTSch corpus). Research on intensification in written language has traditionally focused on two grammatical aspects: syntactic intensification, i.e. the use of particles and other lexical items and morphological intensification, i.e. the use of compounding. Using a wide variety og examples from the corpus, the paper identifies novel ways that have been used for intensification in DMC, and suggests a new taxonomy of classification for future analysis of intensification

    The IVK-Ler corpus of adolescent foreign-language learners of German

    Get PDF
    This paper presents the IVK-Ler corpus, a longitudinal, annotated learner corpus of weekly writings produced by a group of 18 adolescents in a preparatory class. The corpus consists of 117 student texts collected between 2020 and 2021 and has a structure layered by student and text number. It includes metadata that enables researchers to analyze and track individual student progress in terms of syntactic competence and literacy. The annotation schema, manual and automatic annotation processes, and corpus representation are described in detail. The corpus currently includes target hypotheses and gold standard part-of-speech tags. Future work could include additional annotation layers for topological fields and dependency relations, as well as semantic and discourse annotations to make the corpus usable for tasks beyond syntactic evaluations.Dieser Artikel präsentiert das IVK-Ler Korpus, ein longitudinal annotiertes Lernkorpus von wöchentlichen Aufsätzen, produziert von einer Gruppe von 18 Jugendlichen in einer Vorbereitungsklasse. Das Korpus besteht aus 117 Schülertexten, die zwischen 2020 und 2021 gesammelt wurden und hat eine Struktur, die nach Schüler und Textnummer geordnet ist. Es enthält Metadaten, die Forscher ermöglichen, den individuellen Fortschritt der Schüler hinsichtlich syntaktischer Kompetenz und Literacy zu analysieren und zu verfolgen. Das Annotation-Schema, die manuellen und automatischen Annotation-Prozesse sowie die Korpus-Darstellung werden detailliert beschrieben. Das Korpus enthält derzeit Zielhypothesen und Goldstandard-POS-Tags. Zukünftige Erweiterungen könnten zusätzliche Annotation-Schichten für topologische Felder und Abhängigkeitsbeziehungen sowie semantische und Diskurs-Annotationen beinhalten, um das Korpus für Aufgaben jenseits syntaktischer Bewertungen nutzbar zu machen

    Datensatz attributive dass-Sätze und zu-Infinitive

    Get PDF
    Der Datensatz enthält 10.113 Korpusbelege für Konstruktionen, in denen ein Substantiv mit einem dass-Satz oder einem zu-Infinitiv auftritt (das Versprechen, dass man sich irgendwann wiedersieht vs. das Versprechen, sich irgendwann wiederzusehen). Die Daten wurden erhoben aus: 1. dem Korpusgrammatik-Untersuchungskorpus (Bubenhofer et al. 2014), basierend auf dem Deutschen Referenzkorpus DeReKo (Kupietz et al. 2010, 2018), Release 2017-II. 2. dem Subkorpus “Forum” des DECOW16B-Webkorpus (Schäfer & Bildhauer 2012)

    Online comprehension of conditionals in context: A self-paced reading study on wenn (‘if’) versus nur wenn (‘only if’) in German

    No full text
    Comprehending conditional statements is fundamental for hypothetical reasoning about situations. However, the online comprehension of conditional statements containing different conditional connectives is still debated. We report two self-paced reading experiments on German conditionals presenting the conditional connectives wenn (‘if’) and nur wenn (‘only if’) in identical discourse contexts. In Experiment 1, participants read a conditional sentence followed by the confirmed antecedent p and the confirmed or negated consequent q. The final, critical sentence was presented word by word and contained a positive or negative quantifier (ein/kein ‘one/no’). Reading times of the two quantifiers did not differ between the two conditional connectives. In Experiment 2, presenting a negated antecedent, reading times for the critical positive quantifier (ein) did not differ between conditional connectives, while reading times for the negative quantifier (kein) were shorter for nur wenn than for wenn. The results show that comprehenders form distinct predictions about discourse continuations due to differences in the lexical semantics of the tested conditional connectives, shedding light on the role of conditional connectives in the online interpretation of conditionals in general

    Proceedings of the 12th Web as Corpus Workshop (ACL SIGWAC). Language Resources and Evaluation Conference (LREC 2020), Marseille, 11–16 May 2020

    No full text
    The 12th Web as Corpus workshop (WAC-XII) looks at the past, present, and future of web corpora given the fact that large web corpora are nowadays provided mostly by a few major initiatives and companies, and the diversity of the early years appears to have faded slightly. Also, we acknowledge the fact that alternative sources of data (such as data from Twitter and similar platforms) have emerged, some of them only available to large companies and their affiliates, such as linguistic data from social media and other forms of the deep web. At the same time, gathering interesting and relevant web data (web crawling) is becoming an ever more intricate task as the nature of the data offered on the web changes (for example the death of forums in favour of more closed platforms)

    New opportunities for researching digital youth language: The NottDeuYTSch corpus

    No full text
    This article details the process of creating the Nottinghamer Korpus deutscher YouTube-Sprache ('The Nottingham German YouTube Language Corpus' - or NottDeuYTSch corpus) and outlines potential research opportunities. The corpus was compiled to analyse the online language produced by young German-speakers and offers significant opportunity for in-depth research across several linguistic fields including lexis, morphology, syntax, orthography, and conversational and discursive analysis. The NottDeuYTSch corpus contains over 33 million words taken from approximately 3 million YouTube comments from videos published between 2008 to 2018 targeted at a young, German-speaking demographic and represent an authentic language snapshot of young German speakers. The corpus was proportionally sampled based on video category and year from a database of 112 popular German-speaking YouTube channels in the DACH region for optimal representativeness and balance and contains a considerable amount of associated metadata for each comment that enable further longitudinal cross-sectional analyses. The NottDeuYTSch corpus is available for analysis as part of the German Reference Corpus (DeReKo)

    IDS aktuell. Neues aus dem Leibniz-Institut für Deutsche Sprache in Mannheim. Jg. 2023, Heft 2

    No full text

    9,397

    full texts

    11,061

    metadata records
    Updated in last 30 days.
    Publikationsserver des Instituts für Deutsche Sprache is based in Germany
    Access Repository Dashboard
    Do you manage Open Research Online? Become a CORE Member to access insider analytics, issue reports and manage access to outputs from your repository in the CORE Repository Dashboard! 👇