1,720,960 research outputs found

    Exploring BERT's Capabilities to Detect English Preposition Errors

    Get PDF
    Preposition errors are some of the most common errors created by L2 speakers. In addition, improving error correction and detection methods remains an open issue in the realm of Natural Language Processing (NLP). This research investigates whether the bidirectional encoder representations from transformers model (BERT) has the potential to correct preposition errors accurately enough to be useful in error correction software. I used an open-source BERT model and over three hundred thousand edited sentences from Wikipedia, tagged for part of speech, where only a preposition edit had occurred. A method of error detection was devised using multi-level masking to generate suggestions based on sentence context for every prepositional environment in the test data. These suggestions were compared with the original errors in the data and their known corrections to evaluate BERT’s performance. This research finds that BERT performs strongly when the scope of its error correction is limited to preposition choice

    Machine Learning of Inflection

    No full text

    Transderivational relations and paradigm gaps in Russian verbs

    No full text
    In this paper I argue that the notorious case of paradigm gaps, the first person singular gaps of Russian verbs, are not synchronically arbitrary as is often assumed (Graudina et al. 1976; Daland et al. 2007; Baerman 2008), but are predictable and connected to the opaque morphophonological alternations affecting stem-final consonants in the 1st person singular (1sg) present tense. I present evidence for a new empirical generalization showing that the problematic alternations in verbs are subject to a 'lexical conservatism 'effect (Steriade 1997). Namely, stems that appear in other derivationally or inflectionally related forms with the same alternation as the one expected in 1sg generally do not have gaps, while stems that have no attested related forms with alternations do. Overall, a larger set of verbs are problematic for the speakers than indicated in dictionaries, and there are degrees of “gappiness” with a lot of variation across speakers. Additionally, I consider how different theoretical proposals for handling ineffability fare in accounting for these findings. I propose to augment the framework of Harmonic Grammar (Legendre et al. 1990) with an additional post-competition step during which outputs can be compared to each other based on their Harmony scores. This proposal is not tied to violations of specific constraint and it has potential to account for both paradigm gaps and gradient grammaticality judgments

    Grounding Systematic Syncretism in Learning

    No full text
    It is commonly assumed that patterns of syncretism in inflectional paradigms are restricted in some way. In this article, I show how such restrictions can reflect cognitive constraints on language learning. Namely, I construct a learning algorithm that is biased toward certain types of affix distributions in paradigms, thereby rendering them systematic. In developing this algorithm, I rely on the traditional notions of underspecification and blocking, but recast them in terms of learners' biases toward generalization strategies based on cross-situational intersections and default reasoning. This algorithm allows us to test claims about systematicity of syncretism using typological data and language acquisition studies. In the last part of the article, I present a crosslinguistic survey of verbal agreement paradigms that supports the algorithm's predictions. </jats:p

    Comparing learners for Boolean partitions

    No full text

    Going Beyond Counting First Authors in Author Co-citation Analysis

    Get PDF
    The present study examines one of the fundamental aspects of author co-citation analysis (ACA) - the way co-citation counts are defined. Co-citation counting provides the data on which all subsequent statistical analyses and mappings are based, and we compare ACA results based on two different types of co-citation counting - the traditional type that only counts the first one among a cited work's authors on the one hand and a non-traditional type that takes into account the first 5 authors of a cited work on the other hand. Results indicate that the picture produced through this non-traditional author co-citation counting contains more coherent author groups and is therefore considerably clearer. However, this picture represents fewer specialties in the research field being studied than that produced through the traditional first-author co-citation counting when the same number of top-ranked authors is selected and analyzed. Reasons for these effects are discussed

    Variations on the Author

    Get PDF
    “Variations on the Author” discusses two of Eduardo Coutinho’s recent films (Um Dia na Vida, from 2010, and Últimas Conversas, posthumously released in 2015) and their contribution to the general question of documentary authorship. The director’s filmography is characterized by a consistent yet self-effacing form of authorial self-inscription: Coutinho often features as an interviewer that rather than express opinions propels discourses; an interviewer that is good at listening. This mode of self-inscription characterizes him as an author who is not expressive but who is nonetheless markedly present on the screen. In Um Dia na Vida, however, Coutinho is completely absent form the image, while Últimas Conversas, on the contrary, includes a confessional prologue that moves the director from the margins to the center of his films. This article examines the ways in which these works stand out in the filmography of a director who offers new insights into the notion of cinematic authorship

    Appropriate Similarity Measures for Author Cocitation Analysis

    Get PDF
    We provide a number of new insights into the methodological discussion about author cocitation analysis. We first argue that the use of the Pearson correlation for measuring the similarity between authors’ cocitation profiles is not very satisfactory. We then discuss what kind of similarity measures may be used as an alternative to the Pearson correlation. We consider three similarity measures in particular. One is the well-known cosine. The other two similarity measures have not been used before in the bibliometric literature. Finally, we show by means of an example that our findings have a high practical relevance.information science;Pearson correlation;cosine;similarity measure;author cocitation analysis

    Dispelling the Myths Behind First-author Citation Counts

    Get PDF
    We conducted a full-scale evaluative citation analysis study of scholars in the XML research field to explore just how different from each other author rankings resulting from different citation counting methods actually are, and to demonstrate the capability of emerging data and tools on the Web in supporting more realistic citation counting methods. Our results contest some common arguments for the continued use of first-author citation counts in the evaluation of scholars, such as high correlations between author rankings by first-author citation counts and other citation counting methods, and high costs of using more realistic citation counting methods that are not well-supported by the ISI databases. It is argued that increasingly available digital full text research papers make it possible for citation analysis studies to go beyond what the ISI databases have directly supported and to employ more sophisticated methods
    corecore