Publikationsserver des Instituts für Deutsche Sprache
Not a member yet
11061 research outputs found
Sort by
Prepositional object clauses in West Germanic. Experimental evidence from wh-movement
The issue: We discuss (declarative) prepositional object clauses (PO-clauses) in the West Germanic languages Dutch (NL), German (DE), and English (EN). In Dutch and German, PO-clauses occur with a prepositional proform (=PPF, Dutch: ervan, erover, etc.; German: drauf/darauf, drüber/darüber, etc.). This proform is optional with some verbs (1). In English, by contrast, P embeds a clausal complement in the case of gerunds or indirect questions (2), however, P is obligatorily absent when the embedded CP is a that-clause in its base positionv(3a). However, when the that-clause is passivized or topicalized, the stranded P is obligatory (3b). Given this scenario, we will address the following questions: i) Are there structural differences between PO-clauses with a P/PPF and those in which the P/PPF is optionally or obligatorily omitted? ii) In particular, do PO-clauses without P/PPF structurally coincide with direct object (=DO) clauses? iii) To what extent are case and nominal properties of clauses relevant? We use wh-extraction as a relevant test for such differences.
Previous research: Based on pronominalization and topicalization data in German and Dutch, PO-clauses are different from DO-clauses independent of the presence of the PPF (see, e.g., Breindl 1989; Zifonun/Hoffmann/Strecker 1997; Berman 2003; Broekhuis/Corver 2015 and references therein) (4,5). English pronominalization and topicalization data (3b) appear to point in the same direction (Fischer 1997; Berman 2003; Delicado Cantero 2013). However, the obligatory absence of P before that-clauses in base position indicates a convergence with DO-clauses.
Experimental evidence: To provide further evidence to these questions we tested PO-clauses in all three languages for long wh-extraction, which is usually possible for DO-clauses in English and Dutch, and in German for southern regional varieties. For German and Dutch we conducted rating studies using the thermometer method (Featherston 2008). Each study contained two sets of sentences: the first set tested long wh-extraction with regular DO-clauses (6). The second set tested wh-extraction from PO-clauses with and without PPFs (7), respectively. The results show no significant difference in extraction with PO-clauses whether or not the PPF was present even for those speakers who otherwise accept long-distance extraction in German. This supports a uniform analysis of PO-clauses with and without the PPF in contrast to DO-clauses. For English we tested extraction with verbs that select for PP-objects in two configurations: V+that-clause and V+P-gerund (8) in comparison to sentences without extraction. Participants rated sentences on a scale of 1 (unnatural) to 7 (natural). We included the gerund for English as this is a regular alternative for such objects. The results show that extraction is licit in both configurations. This suggests that English PO-clauses are different from German and Dutch PO-clauses: They rather behave as DO-clauses allowing for extraction. Note though, that the availability of extraction from P+gerund also shows that PPs are not islands for extraction in English. Overall, this shows that there is a split between English vs. German/Dutch PO-clauses when the P/PPF is absent. While these clauses behave like PO-clauses in the latter languages, extraction does not show a difference between DO- and PO-clauses in English. We will discuss the results in relation to the questions i)–iii) above
Human languages trade off complexity against efficiency
A central goal of linguistics is to understand the diverse ways in which human language can be organized (Gibson et al. 2019; Lupyan/Dale 2016). In our contribution, we present results of a large scale cross-linguistic analysis of the statistical structure of written language (Koplenig/Wolfer/Meyer 2023) we approach this question from an information-theoretic perspective. To this end, we conduct a large scale quantitative cross-linguistic analysis of written language by training a language model on more than 6,500 different documents as represented in 41 multilingual text collections, so-called corpora, consisting of ~3.5 billion words or ~9.0 billion characters and covering 2,069 different languages that are spoken as a native language by more than 90% of the world population. We statistically infer the entropy of each language model as an index of un. To this end, we have trained a language model on more than 6,500 different documents as represented in 41 parallel/multilingual corpora consisting of ~3.5 billion words or ~9.0 billion characters and covering 2,069 different languages that are spoken as a native language by more than 90% of the world population or ~46% of all languages that have a standardized written representation. Figure 1 shows that our database covers a large variety of different text types, e.g. religious texts, legalese texts, subtitles for various movies and talks, newspaper texts, web crawls, Wikipedia articles, or translated example sentences from a free collaborative online database. Furthermore, we use word frequency information from the Crúbadán project that aims at creating text corpora for a large number of (especially under-resourced) languages (Scannell 2007). We statistically infer the entropy rate of each language model as an information-theoretic index of (un)predictability/complexity (Schürmann/Grassberger 1996; Takahira/Tanaka-Ishii/Dębowski 2016). Equipped with this database and information-theoretic estimation framework, we first evaluate the so-called ‘equi-complexity hypothesis’, the idea that all languages are equally complex (Sampson 2009). We compare complexity rankings across corpora and show that a language that tends to be more complex than another language in one corpus also tends to be more complex in another corpus. This constitutes evidence against the equi-complexity hypothesis from an information-theoretic perspective. We then present, discuss and evaluate evidence for a complexity-efficiency trade-off that unexpectedly emerged when we analysed our database: high-entropy languages tend to need fewer symbols to encode messages and vice versa. Given that, from an information theoretic point of view, the message length quantifies efficiency – the shorter the encoded message the higher the efficiency (Gibson et al. 2019) – this indicates that human languages trade off efficiency against complexity. More explicitly, a higher average amount of choice/uncertainty per produced/received symbol is compensated by a shorter average message length. Finally, we present results that could point toward the idea that the absolute amount of information in parallel texts is invariant across different languages
Sprachliche Zweifelsfälle. Lexikalisch-semantische, flexivische und wortbildungsbedingte Zweifelsfälle
Sprachliche Zweifelsfälle kommen auf allen linguistischen Ebenen vor. Ihre Einordnung erfolgt zumeist nach Systemebene, nach Entstehungsursache oder nach lexematischer Struktur. Sprachlicher Zweifel kann auch nach intra- und interlingualen Aspekten unterschieden werden. Stehen zwei oder mehrere lexikalische Varianten zur Verfügung, kann es zu Unsicherheiten bezüglich des angemessenen Gebrauchs kommen. Nicht nur Muttersprachler*innen sind mit Schwierigkeiten konfrontiert, Zweifelsfälle stellen auch ein Problem bei der Fremdsprachenproduktion dar.
Dieser Band beschränkt sich auf lexikalisch-semantische, flexivische und wortbildungsbedingte Zweifelsfälle und führt interessierte Leser*innen in Fachliteratur und Nachschlagewerke ein. Er streift Fragen der Sprachdidaktik, der Fehler- und Variationslinguistik, denn die Auseinandersetzung mit typischen Zweifelsfällen zeigt auch das Spannungsfeld zwischen allgemeinem Usus und kodifizierter Norm, zwischen Gegenwart und Wandel, zwischen Dynamik, sprachlichem Reichtum und erlernter Bildungstradition
Winter conference in summer: report on the International Conference on Conversation Analysis (ICCA) from June 26th to July 2nd 2023 in Brisbane/Meanjin (Australia)
From June 26th to July 2nd 2023 the International Conference on Conversation Analysis (ICCA) took place in Brisbane/Meanjin, Australia – after a long pause due to the Covid-pandemic and for the first time in the southern hemisphere. About 350 participants from about 50 different countries attended the conference. This year’s ICCA came up with 36 panels and about 300 papers that were presented. Four plenary speakers have been invited and 24 pre-conference workshops took place. On Wednesday evening Ilana Mushin, in her role as conference chair, officially opened ICCA. The President of the International Society of Conversation Analysis (ISCA), Tanya Stivers, also welcomed all participants. To get acquainted with the indigenous culture of Queensland, the opening ceremony was enriched with a highly impressive dance performance by First Nations people. After the official inauguration the international community met at the Welcome Reception to look forward together to the days ahead with many opportunities for exchange and networking.
As it will become clear throughout this report, the research topics revolved around not only classic CA concepts, but also importantly concerned embodiment, which continued the line of past conferences (Dix 2019). Another aspect that has been highlighted was conflict and social norms. Due to personal capacities, we can only present a selection of presentations within the scope of this conference report. The selection was influenced by the personal interest of the authors and should not be understood as rating in any sense
The FAIR Index of CMC Corpora
In this article, we examine the current situation of data dissemination and provision for CMC corpora. By that we aim to give a guiding grid for future projects that will improve the transparency and replicability of research results as well as the reusability of the created resources. Based on the FAIR guiding principles for research data management, we evaluate the 20 European CMC corpora listed in the CLARIN CMC Resource family, individuate successful strategies among the existing corpora and establish best practices for future projects. We give an overview of existing approaches to data referencing, dissemination and provision in European CMC corpora, and discuss the methods, formats and strategies used. Furthermore, we discuss the need for community standards and offer recommendations for best practices when creating a new CMC corpus
Einsatz von EDV und Mikrocomputer in Lehrveranstaltungen zur digitalen Lexikografie
Gerd Hentschel gehört zu den Pionieren der heutigen Computerlexikografie und der IT-gestützten Korpuserschließung. Eine seiner ersten Zeitschriftenpublikationen, mit dem Titel Einsatz von EDV und Mikrocomputer in einem lexikographischen Forschungsprojekt zum deutschen Lehnwort im Polnischen (Hentschel 1983), befasst sich mit der Frage, wie - unter den damaligen technischen Vorzeichen - Forschungs- und Dokumentationsarbeiten zu polnischen Germanismen sinnvoll durch die Verwendung von Computern unterstützt werden können. Die besagten Arbeiten mündeten später in die Online-Publikation des Wörterbuchs der deutschen Lehnwörter in der polnischen Schrift- und Standardsprache (WDLP). Es ist aus heutiger Sicht bemerkenswert, mit welchen Beschränkungen die Arbeit mit dem Computer noch vor 40 Jahren zu kämpfen hatte. Aus gegebenem Anlass sei es gestattet, diesen Punkt etwas ausführlicher zu illustrieren
Let's talk European! Politolinguistische Überlegungen zur europäischen Integration anhand der deutschen Berichterstattung im Sommer 2015
Im folgenden Beitrag, der im Bereich der Politolinguistik und der Diskursanalyse angesiedelt ist, wird auf der Grundlage der deutschen Berichterstattung des Sommers 2015 die brisante Problematik der griechischen Euro-Währungskrise, die das ganze Europa wochenlang in Atem hält, unter die Lupe genommen. Die Debatte über die bis dahin „schwerste Krise der europäischen Integration" verläuft als äußerst emotional geführter gesamteuropäischer Meinungsaustausch. Obwohl man annehmen könnte, dass die nervenaufreibenden Auseinandersetzungen über die Euro-Währungskrise eigentlich nur auf Staaten der Euro-Zone begrenzt sein sollten, beweist die europäische Berichterstattung, dass man in der heutigen EU nicht mehr aus der Beobachter-, sondern eigentlich aus der Teilnehmerperspektive berichtet, weil die Probleme eines Landes genauso Schwierigkeiten für andere, die sogar selbst nicht unbedingt in der Euro-Zone sein müssen, bedeuten können. Im Jahr 2015 wird die griechische Euro-Krise zum Auslöser für Fragen nach der Zukunft Europas. Sie betreffen in erster Linie die Problematik der weiteren Integration und der europäischen Identität
Hermeneutische Betrachtungen zum Stellenwert von “Wort” in der christlichen Mystik
Das Wort als das wichtigste und ureigenste Element des Sprachsystems wird in der modernen Linguistik als linguistische Einheit unter phonetisch/phonologischem, orthographischem, morphologischem, syntaktischem und semantischem Kriterium untersucht und beschrieben. Wenn man jedoch der Frage nachgeht, wie ein Wort rein physikalisch entsteht und besteht, wird man feststellen, dass alles auf die Energie zurückgeht. Unter diesen Gesichtspunkten wäre es daher angebracht, unseren linguistischen Blickwinkel zu ändern und das Wort nicht nur als linguistische Einheit sondern auch als Energieträger aufzufassen und die energietragende Funktion des Wortes bzw. der Sprache im Zusammenhang mit der heutigen Wissenschaft und den religiösen und mystischen Betrachtungen zu untersuchen
20 Jahre danach: Soziale Veränderung und sprachliche Verbreitung. Verkaufsgespräch bei Japanern in Düsseldorf
Fast 20 Jahre sind vergangen, seit ich für meine Dissertation Untersuchungen über Ein- und Verkaufsgespräche von Deutschen und Japanern in Deutschland und Japan durchführte. Dort wurden konkrete verbale und nonverbale Handlungen zwischen deutschen bzw. japanischen Verkäufern und deutschen bzw. japanischen Kunden beim Ein- und Verkaufen untersucht. Untersuchungsorte waren dabei Düsseldorf, wo die meisten Japaner in Deutschland ansässig sind, Tokio, wo die meisten Deutschen in Japan ansässig sind, Heidelberg, das von vielen japanischen Touristen besucht wird, und Nagano, wo deutsche Touristen damals bei der Olympiade waren. Anlässlich dieser Festschrift für meinen Doktorvater Prof. Dr. Gerhard Stickel versuchte ich, eine kleine Untersuchung durchzuführen, um sprachliche Veränderungen im Verlauf der Zeit und der sozialen Veränderung zu beobachten. In dieser Abhandlung werden die Veränderung der Gesellschaft und ihr Einfluss auf die Sprache behandelt. Im folgenden zweiten Abschnitt werden soziale Veränderungen in Düsseldorf thematisiert, im dritten Abschnitt werden die Ergebnisse der zwei Befragungen analysiert und zum Schluss wird eine Möglichkeit der Sprachverbreitung im Zusammenhang mit der heutigen Gesellschaft dargestellt
Multimodalität der Kooperation im Lehr-Lern-Diskurs. Wie Ideen für Filme entstehen
Die Untersuchung präsentiert die multimodale Struktur und Komplexität eines besonderen Kooperationstyps, dem »Pitching«. Das Pitching ist eine Mischform aus Arbeits- und Lehr-Lern-Diskurs, bei der vier Studierende gemeinsam mit zwei Dozenten Filmideen entwickeln. Als empirische Grundlage dient ein Datenkorpus von 72 Stunden Videoaufnahmen, das methodisch mit einer Kombination aus ethnographischer Gesprächsanalyse, ethnomethodologischer Konversationsanalyse und deren Erweiterung um eine multimodale Analyseperspektive untersucht wird. Dabei wird detailliert der komplexe Gesamtzusammenhang von Verbalität, Mimik, Gestik, Körperpositur und anderen körperlichen Ausdruckformen in seiner Bedeutung für die gemeinsame Arbeit ersichtlich. Basierend auf den beiden zentralen Konzepten »Kooperation« und »Handlungsschema« werden die spezifischen Situationsmerkmale des Pitchings und die typischen Aufgaben und Probleme rekonstruiert, die von den Interaktionsbeteiligten durch unterschiedliche Verfahren bearbeitet werden. Aufgrund einer longitudinalen Perspektive gibt die Untersuchung zudem Einblicke in die Professionalisierung der Studierenden im Studienverlauf