Publikationsserver des Instituts für Deutsche Sprache
Not a member yet
    11061 research outputs found

    DAS SONGKORPUS. Perspektiven einer korpuslinguistischen Nutzung deutschsprachiger Popmusik für die Fremd- und Zweitsprachenvermittlung

    Get PDF
    Vorgestellt wird das Korpus deutschsprachiger Songtexte als innovative Sprachdatenquelle für interdisziplinäre Untersuchungsszenarien und speziell für den Einsatz im Fremd- und Zweitsprachenunterricht. Die Ressource dokumentiert Eigenschaften konzeptioneller Schriftlichkeit und konzeptioneller Mündlichkeit und erlaubt empirisch begründete Analysen sprachlicher Phänomene bzw. Tendenzen in den Texten moderner Popmusik. Vorgestellt werden Design, Annotationen und Anwendungsbeispiele des in thematische und autorenspezifische Archive stratifizierten Korpus.We present the Corpus of German Song Lyrics as an innovative language resource for interdisciplinary research scenarios and especially for use in foreign and second language teaching. The resource documents characteristics of conceptual literacy and conceptual orality, and allows for empirical analyzes of linguistic phenomena and tendencies in modern pop lyrics. Design, annotations and application examples of the corpus, which is stratified into thematic and author-specific archives, are presented

    Nachhaltige Dokumentation virtueller Forschungsumgebungen

    Get PDF
    In den letzten Jahren werden immer mehr virtuelle Forschungsumgebungen für die maschinelle Sprachverarbeitung zur Verfügung gestellt. Diese sollten zum einen nachhaltig und zum anderen für potenzielle Nutzer vergleichbar dokumentiert werden. In diesem Beitrag werden daher Bedingungen für die Nachhaltigkeit insbesondere von NLP- (Natural Language Processing) Werk-zeugen beschrieben: Die Dokumentation sollte nicht nur die Software, son-dern auch ihre Evaluierung anhand einer – ebenfalls gut dokumentierten – Testsuite umfassen. Im Beitrag werden auch Möglichkeiten dargestellt, den Dokumentationsvorgang selbst anhand von DocBook XML zu automatisieren.hroughout the last years, an increasing number of virtual research environ-ments have been offered in the field of Natural Language Processing (NLP). These should be documented in a sustainable way that also guarantees com-parability for potential users. This paper thus describes constraints for the sustainability of NLP-environments: the documentation must describe not only the software from the developer’s view, but also its evaluation accor-ding to a testsuite, which is itself to be documented comprehensively. The paper also describes the possibility of automating the documentation proc-esses by utilizing DocBook XML

    Joint utterance formulation from a cross-linguistic perspective. Co-constructions in Czech and German

    Get PDF
    This presentation deals with collaborative turn-sequences (Lerner 2004), a syntactically coherent unit of talk that is jointly formulated by at least two speakers, in Czech and German everyday conversations. Based on conversation analysis (e.g., Schegloff 2007) and a multimodal approach to social interaction (e.g., Deppermann/Streeck 2018), we aim at comparing recurrent patterns and action types within co-constructional sequences in both languages. The practice of co-constructing turns-at-talk has been described for typologically different languages, especially for English (e.g., Lerner 1996, 2004), but also for languages such as Japanese (Hayashi 2003) or Finnish (Helasvuo 2004). For German, various forms and functions of co-constructions have already been investigated (e.g., Brenning 2015); for Czech, a detailed, interactionally based description is still pending (but see some initial observations in, e.g., Hoffmannová/Homoláč/Mrázková (eds.) 2019). Although the existence of co-constructions in different languages points to a cross-linguistic conversational practice, few explicitly comparative studies exist (see, e.g., Lerner/Takagi 1999, for English and Japanese). The language pair Czech-German has mainly been studied with respect to language contact and without specifically considering spoken language or complex conversational sequences (e.g., Nekula/Šichová/Valdrová 2013). Therefore, our second aim is to sketch out a first comparison of co-constructional sequences in German and Czech, thereby contributing to the growing field of comparative and cross-linguistic studies within conversation analysis (e.g., Betz et al. (eds.) 2021; Dingemanse/Enfield 2015; Sidnell (ed.) 2009). More specifically, we will present three main sequential patterns of co-constructional sequences, focusing on the type of action a second speaker carries out by completing a first speaker’s possibly incomplete turn-at-talk, and on how the initial speaker then responds to this suggested completion (Lerner 2004). Excerpts from video recordings of Czech and German ordinary conversations will illustrate these recurrent co-constructional sequence types, i.e., offering help during word searches (see example 1 above), displaying understanding, or claiming independent knowledge. The third objective of this paper is to underline the participants’ orientation to similar interactional problems, solved by specific syntactic and/or lexical formats in Czech and German. Considering the more recent focus on the embodied dimension of co-constructional practices (e.g., Dressel 2020), we will also investigate the multimodal formatting of a started utterance as more or less “permeable” (Lerner 1996) for co-participant completion, the participants’ mutual embodied orientation, and possible embodied responses to others’ turn-completions (such as head nods or eyebrow flashes, cf. De Stefani 2021). More generally, this contribution reflects on the possibilities and challenges of a cross-linguistic comparison of complex multimodal sequences

    Funktionen auf dem Weg zu Formen, oder: Warum die Oberfläche doch recht hat

    Get PDF
    Komposition als Element nominaler Integration passt zum Sprachtyp des Deutschen. Diese Technik wird in verschiedenen Texttypen in unterschiedlicher Weise genutzt und funktional ausdifferenziert. Zweigliedrige Komposita prägen den alltäglichen Wortschatz. Die Erfahrung damit und ihre formale Offenheit bilden den Grund für spezifische Ausweitungen des Gebrauchs. Das wird gezeigt an der die Öffnung der Muster im literarischen Bereich, dann an der Interaktion von Kompositionstypen im Hinblick auf größtmögliche Explizitheit in juristischen Texten und letztlich an der Mischung von alltäglicher Klassifikation in gängigen Komposita und textfunktionaler Kondensierung in einem Sachtext

    A computational implementation of the Northern Sotho infinitive

    Get PDF
    The aim of this article is to describe the infinitive in Northern Sotho based on corpus data and the respective literature; so far, all share the same view: The infinitive is a noun (of class 15) and a verb at the same time—‘it manifests both nominal as well as verbal features’ (Poulos & Louwrens, 1994:42). When implementing these constellations in a parser, however, a new perspective is found: to achieve its successful implementation, the infinitive must be defined as a verb on the one hand and as a noun of class 15 on the other, derived from this verb through nominalization (transposition). Instead of a subject concord, the verb stem in the infinitive is preceded by the respective class prefix

    News from the International Comparable Corpus. First launch of ICC written

    No full text
    The International Comparable Corpus (ICC) (Kirk/Čermáková 2017; Čermáková et al. 2021) is an open initiative which aims to improve the empirical basis for contrastive linguistics by compiling comparable corpora for many languages and making them as freely available as possible as well as providing tools with which they can easily be queried and analysed. In this contribution we present the first release of written language parts of the ICC which includes corpora for Chinese, Czech, English, German, Irish (partly), and Norwegian. Each of the released corpora contains 400k words distributed over 14 different text categories according to the ICC specifications. Our poster covers the design basics of the ICC, its TEI encoding, a demonstration of using the ICC via different query tools, and an outlook on future plans. Similar to the European Reference Corpus EuReCo (Kupietz et al. 2020), ICC follows the approach of reusing existing linguistic resources wherever possible in order to cover as many languages as possible with realistic effort in as short a time as possible. In contrast to EuReCo, however, comparable corpus pairs are not defined dynamically in the usage phase, but the compositions of the corpora are fixed in the ICC design. The approaches are thus complementary in this respect. The design principles and composition of the ICC are based on those of the International Corpus of English (ICE) (Greenbaum (ed.) 1996), with the deviation that the ICC includes the additional text category blog post and excludes spoken legal texts (see Čermáková et al. 2021 for details). ICC’s fixed-design approach has the advantage that all single-language corpora in the ICC have the same composition with respect to the selected text types and that this guarantees that the selected broad spectrum of potential influencing variables for linguistic variation is always represented. The disadvantage, however, is that this can only be achieved for quite small corpora and that the generalisability of comparative findings based on the ICC corpora will often need to be checked on larger monolingual corpora or translation corpora (Čermáková/Ebeling/Oksefjell Ebeling forthcoming). Arguing that such issues with comparability and representativeness are inevitable, in one way or the other, and need to be dealt with, our poster will discuss and exemplify the text selections in more detail

    Collecting and analysing multi-source video data. Grasping the opacity of smartphone use in face-to-face encounters

    No full text
    The ubiquity of smartphones has been recognised within conversation analysis as having an impact on conversational structures and on the participants’ interactional involvement. However, most of the previous studies have relied exclusively on video recordings of overall encounters and have not systematically considered what is taking place on the device. Due to the personal nature of smartphones and their small displays, onscreen activities are of limited visibility and are thus potentially opaque for both the co-present participants (“participant opacity”) and the researchers (“analytical opacity”). While opacity can be an inherent feature of smartphones in general, analytical opacity might not be desirable for research purposes. This chapter discusses how a recording set-up consisting of static cameras, wearable cameras and dynamic screen captures allowed us to address the analytical opacity of mobile devices. Excerpts from multi-source video data of everyday encounters will illustrate how the combination of multiple perspectives can increase the visibility of interactional phenomena, reveal new analytical objects and improve analytical granularity. More specifically, these examples will emphasise the analytical advantages and challenges of a combined recording set-up with regard to smartphone use as multiactivity, the role of the affordances of the mobile device, and the prototypicality and “naturalness” of the recorded practices

    Когнітивно-дискурсивна реконструкція комунікативних девіацій в українсько- і німецькомовних відеоінтерв’ю

    No full text
    The thesis presents a new theoretical and methodological concept for the performance of a cognitive and discursive reconstruction of communicative deviations in Ukraine and German video interviews from the position of cognitive comparative linguistics and contrastive typological linguistics. The research paper substantiates the status of video interviews as an integrated speech genre, which includes television interviews and special interviews stored on YouTube video hosting on the Internet. The definition of the notion of “communicative deviation” as a dynamic and complex cognitive and discursive phenomenon is clarified against related terms and notions.У дисертації розроблено нову теоретико-методологічну концепцію для виконання когнітивно-дискурсивної реконструкції комунікативних девіацій в українсько- і німецькомовних відеоінтерв’ю. Обґрунтовано статус відеоінтерв’ю як інтегрованого мовленнєвого жанру, який включає теле- і спеціальні інтерв’ю, збережені на відеохостингу YouTube в мережі Інтернет. Уточнено визначення поняття “комунікативної девіації” на тлі суміжних термінів як динамічного й складного когнітивно-дискурсивного явища. Реконструйовано причини виникнення комунікативних девіацій і побудовано їхню модель, характерну для українсько- та німецькомовних відеоінтерв’ю. Результати дисертації можна застосовувати у зіставно-типологічних дослідженнях, у дослідженнях із теорії мови, психо- і соціолінгвістики, лінгвопрагматики, когнітивної і комунікативної лінгвістик, методології мовознавства, у курсах зіставного мовознавства, загального мовознавства, теоретичної граматики німецької мови, сучасної української літературної мови, а також у викладанні відповідних навчальних дисциплін. Підсумки дослідження можуть бути також корисними для представників мас-медійної сфери, фахівців, які спеціалізуються у галузі теорії комунікації, а також представників сфери соціальних комунікацій, дипломатичних служб, державних і приватних інституцій різного профілю з метою запобігання конфліктним ситуаціям і покращення соціальної та міжкультурної комунікації

    Sprachpolitik der Parteien in den Wahlprogrammen zur Bundestagswahl 2021

    No full text
    Sprachpolitik war in der Bundesrepublik Deutschland seit 1949 nie ein größeres Thema in Wahlkämpfen. Seit der Bundestagswahl 2017 hat sich dies jedoch geändert. Damals waren unter dem Eindruck des großen Migrationsandrangs im Jahr 2016 von einigen Parteien Positionen zu sprachlicher Integration in die Wahlprogramme aufgenommen worden. Unter Positionen sei hier der explizite sprachliche Ausdruck einer Haltung zu einem politischen Thema bzw. Themenbereich zu verstehen, der unter anderem im Rahmen von parteilichen Grundsatz- und Wahlprogrammen Orientierung hinsichtlich des (zukünftig zu erwartenden) politischen Handelns parteilicher Akteur/-innen bieten soll. Und auch die zunehmende Diversität der deutschen Gesellschaft führte schon bei der Wahl im Jahr 2017 zu einer Berücksichtigung von Themen der sprachlichen Bildung in der Programmatik der Parteien. Dieser Beitrag untersucht somit die Grundsatz- und Wahlprogramme der größten Parteien anhand der sprachpolitischen Ausdrucksweise

    9,397

    full texts

    11,061

    metadata records
    Updated in last 30 days.
    Publikationsserver des Instituts für Deutsche Sprache is based in Germany
    Access Repository Dashboard
    Do you manage Open Research Online? Become a CORE Member to access insider analytics, issue reports and manage access to outputs from your repository in the CORE Repository Dashboard! 👇