1,720,997 research outputs found

    No explanation required: Entrenchment and perception of Māori loanwords in a diachronic newspaper corpus of New Zealand English

    Get PDF
    Māori loanwords are uncontestably the most notable feature of New Zealand English (Gordon, 2005; Macalister, 2006b). Although frequencies of Māori loanwords are reported to be increasing in recent years (Macalister, 2006b), and analyses strongly indicate a skew in loanwords across certain topics (Davies & Maclagan, 2006; de Bres, 2006), no studies have yet addressed whether Māori loanword behaviour differs in a corpus restricted to a single topic (with the exception of Calude, Miller, Harper, & Whaanga, Forthcoming). Nor has a thorough diachronic analysis been made on the subject since Davies and Maclagan (2006) and Macalister (2006b). This thesis is concerned with frequency of use and perception of Māori loanwords in a diachronic corpus restricted to a singular theme: te Wiki o te Reo Māori, Māori Language Week. Using a quantitative methodology in a corpus constructed from New Zealand regional and national newspaper articles spanning 2008-2017, average Māori loanword frequencies are found to be almost 5 times that of the most recent comparable diachronic study (Macalister, 2006b), and demonstrate a statistically significant diachronic increase in use. In contrast, the markedness (or translation) of Māori loanwords shows a statistically significant decreasing diachronic trend over the 10-year period. However, practice of marking appears to be affected by user perceptions and intention to educate, and therefore cannot be relied upon as a gauge of loanword entrenchment. Other conventional measures of loanword entrenchment such as frequency, listedness and morphological assimilation are also shown to be unreliable when considered independently from one another, and in cases such as that of New Zealand English where a change from above influenced by social and cultural factors (Māori loanwords functioning as an expression of political attitude and stance) appears to be interfering in loanword use. Measurement of Māori loanword entrenchment cannot be made using the same criteria as loanwords from non-threatened languages, and suggestions are made for the revised criteria for studying loanword entrenchment which takes into account the sociolinguistic context in which the loanwords are used

    Exhausted men: Making fatigue visible in The Metamorphosis and At Swim-Two-Birds

    Get PDF
    Though not always immediately recognizable, fatigue and tiredness play a central role in modernist literature. As society placed increasing value on productive, strong, energetic bodies, the growing presence of bodies tired from the effects of overwork, rapid social progress, illness and post-war malaise posed a threat both to society and ideas of the able-bodied self. Recumbent bodies challenge expectations of productivity; they remindus of our own bodily fallibility and blur the line between life and death as a visual representation of our own mortality. Further, the fatigued body at rest is predominantly associated with the feminine, as has been reflected in much prior academic engagement with neurasthenic women in modernist literature. What then are the implications for fatigued men? This thesis traces the previously overlooked depictions and uses of fatigue in two texts that have been the subject of significant analysis: The Metamorphosis by Franz Kafka and At Swim-Two-Birds by Flann O’Brien. Though neither of these texts are inherently medical in nature, I show how fatigue nonetheless plays a central role in each, albeit in quite different ways. By unpacking these contrasting engagements with fatigue and the simultaneous lack of critical interest in fatigue in these texts, I uncover stigmatizing social perceptions of fatigue as deviant, burdensome, and indicative of emasculating weakness. Further, such depictions seem to go unquestioned, suggesting that these beliefs are largely naturalized within Western society. In the current time of pandemic, as society is again increasingly confronted with very visible exhaustion in the form of Long Covid amongst other fatiguing conditions, modernist fiction from the early-twentieth century offers valuable insights into how we engage with fatigue

    Visualising categorical data: Linguistic case studies from te Reo Māori and New Zealand English

    No full text
    Categorical variables are prevalent in real-world datasets across numerous domains, yet few visualisation techniques accommodate them effectively. This is especially true of datasets comprising three or more categorical variables, termed multivariate categorical data. Visualising such data is challenging due to the lack of inherent ordering of nominal categories, the so-called ‘curse of dimensionality’, and the potential variability in the number of categories per variable. Corpus linguistics, which involves the study of large digital collections of naturally occurring language, serves as the primary application domain in this thesis. This domain was chosen because it is rich in multivariate categorical data and, at the same time, is often visualised using only basic techniques. This thesis contributes to the area of categorical data visualisation in several ways. First, we propose a taxonomy of techniques for visualising categorical data, highlight limitations of existing solutions, and identify relevant analysis tasks. Building on this foundation, the thesis introduces novel techniques and enhancements for visualising datasets involving multiple categorical variables. We focus on adapting the layout and interactive capabilities of an existing technique that uses a matrix of heatmaps to represent pairwise category intersections. These modifications show that directly visualising statistical test results for categorical data can be beneficial for exploring bivariate patterns and associations. Furthermore, we contribute the design, implementation and evaluation of a novel technique called MultiCat, which is not restricted to pairwise intersections but rather facilitates analysis of relationships among multiple variables simultaneously. Both these techniques are interactive and offer greater scalability than existing alternatives, thereby affording new possibilities for analysing multivariate categorical data. However, since categorical variables can occur within more complex data structures, we also consider their presence in networks and hypergraphs, which require specialised methods. To demonstrate the application of these techniques, we draw on two linguistic case studies that focus on languages of special significance in Aotearoa New Zealand. Addressing the low-resource status of Māori, the country’s Indigenous language, we first contribute two related Twitter datasets—a monolingual Māori corpus and a mixed–language Māori–English corpus—together with an architecture for differentiating Māori and English words. Our initial case study uses the monolingual Māori corpus and proposed visualisation techniques to investigate grammatical possession in Māori, offering fresh insights into the linguistic practices of contemporary speakers. The second case study uses networks and hypergraphs with categorical attributes to explore Māori loanword co-occurrence in New Zealand English newspaper articles. We find that loanwords tend not to occur in isolation and that New Zealanders are still importing new (unlisted) borrowings from Māori. Ultimately, the techniques developed in this thesis have broad applications both within and beyond the corpus linguistics community. By enabling more effective visualisation and analysis of multivariate categorical data, this research has the potential to facilitate deeper insights into domains as diverse as education, healthcare, business and science

    Going Beyond Counting First Authors in Author Co-citation Analysis

    Get PDF
    The present study examines one of the fundamental aspects of author co-citation analysis (ACA) - the way co-citation counts are defined. Co-citation counting provides the data on which all subsequent statistical analyses and mappings are based, and we compare ACA results based on two different types of co-citation counting - the traditional type that only counts the first one among a cited work's authors on the one hand and a non-traditional type that takes into account the first 5 authors of a cited work on the other hand. Results indicate that the picture produced through this non-traditional author co-citation counting contains more coherent author groups and is therefore considerably clearer. However, this picture represents fewer specialties in the research field being studied than that produced through the traditional first-author co-citation counting when the same number of top-ranked authors is selected and analyzed. Reasons for these effects are discussed

    Variations on the Author

    Get PDF
    “Variations on the Author” discusses two of Eduardo Coutinho’s recent films (Um Dia na Vida, from 2010, and Últimas Conversas, posthumously released in 2015) and their contribution to the general question of documentary authorship. The director’s filmography is characterized by a consistent yet self-effacing form of authorial self-inscription: Coutinho often features as an interviewer that rather than express opinions propels discourses; an interviewer that is good at listening. This mode of self-inscription characterizes him as an author who is not expressive but who is nonetheless markedly present on the screen. In Um Dia na Vida, however, Coutinho is completely absent form the image, while Últimas Conversas, on the contrary, includes a confessional prologue that moves the director from the margins to the center of his films. This article examines the ways in which these works stand out in the filmography of a director who offers new insights into the notion of cinematic authorship

    Appropriate Similarity Measures for Author Cocitation Analysis

    Get PDF
    We provide a number of new insights into the methodological discussion about author cocitation analysis. We first argue that the use of the Pearson correlation for measuring the similarity between authors’ cocitation profiles is not very satisfactory. We then discuss what kind of similarity measures may be used as an alternative to the Pearson correlation. We consider three similarity measures in particular. One is the well-known cosine. The other two similarity measures have not been used before in the bibliometric literature. Finally, we show by means of an example that our findings have a high practical relevance.information science;Pearson correlation;cosine;similarity measure;author cocitation analysis

    Automating vocabulary tests and enriching online courses for language learners

    Get PDF
    The past decade has seen a massive growth in online academic courses, most of which are offered in the English language. However, although more people speak English as their second language than as their first, online course providers do not offer language assistance. Providing online learners with language resources would allow them to both learn about a subject through a foreign language and learn the foreign language through the subject. This is referred to as “content-based language learning”. Supporting content-based language learning using online courses raises several challenges, three of which are addressed in this thesis. First, courses teach subjects in particular domains, but supporting domain-specific language requires knowledge of specialised vocabulary. This thesis develops an automated approach to generating domain-specific corpora and wordlists, extracting domain-specific vocabulary in a way that can be applied to any online course. Second, acquiring and measuring language come hand-in-hand. Tools that help learners acquire new language should also include methods for testing. This thesis takes an existing vocabulary test and automates it. This has two main advantages: it requires no assumed knowledge of the language, allowing automatic generation of vocabulary tests; and the tests reflect the wordlists used to create them, allowing them to be targeted toward a particular domain. Finally, for content-based language learning to be used successfully, the language components must be smoothly integrated into courses without disturbing the original content. Furthermore, vocabulary support should include multi-word lexical items as well as single words. The thesis describes a tool that enhances online course content, via a browser extension. It is completely automated, though would also lend itself to selective teacher intervention. It is illustrated here with reference to courses offered by the FutureLearn MOOC consortium

    Dispelling the Myths Behind First-author Citation Counts

    Get PDF
    We conducted a full-scale evaluative citation analysis study of scholars in the XML research field to explore just how different from each other author rankings resulting from different citation counting methods actually are, and to demonstrate the capability of emerging data and tools on the Web in supporting more realistic citation counting methods. Our results contest some common arguments for the continued use of first-author citation counts in the evaluation of scholars, such as high correlations between author rankings by first-author citation counts and other citation counting methods, and high costs of using more realistic citation counting methods that are not well-supported by the ISI databases. It is argued that increasingly available digital full text research papers make it possible for citation analysis studies to go beyond what the ISI databases have directly supported and to employ more sophisticated methods

    Author Index

    No full text
    Nao informado
    corecore