1,720,956 research outputs found
Plume humaine ou plume de ChatGPT ? Comparaison entre traduction, post-édition et post-édition par ChatGPT
Ce poster décrit le projet de thèse en cours, qui repose sur le croisement entre deux pans de l’étude contrastive de la traduction et de la post-édition, ainsi que la méthodologie prévue. D'une part, les études comparant traduction humaine (TH) et post-édition humaine (PEH) en termes de qualité concluent généralement à des scores de fidélité et de fluidité similaires. Or, quand les évaluateurs se voient demander leur préférence à l’aveugle, une faible (Bowker, 2009 ; Martikainen et Kübler, 2016) ou grande majorité (Fiederer et O’Brien, 2009 ; Jia et al., 2019) d’entre eux préfèrent la version traduite à la version post-éditée, sans que ces études ne mettent en évidence une raison expliquant cette préférence. D’autre part, les études comparant ces deux modes de traduction en termes de différences linguistiques (notamment phénomène de post-editese) (Farrell, 2018 ; Castilho et al., 2019 ; Toral, 2019 ; Martikainen & Mestivier, 2020 ; Castilho et Resende, 2022 ; Volkart et Bouillon, 2022 ; 2023 ; 2024 ; Schumacher, 2025) semblent souvent indiquer que, par rapport aux TH, les PEH présentent i) une variété lexicale moindre ; ii) un nombre inférieur de solutions de traduction ; iii) un degré d’équivalence syntaxique, ou calque, plus élevé.
L'hypothèse de travail est que cette préférence inexpliquée pour la TH par rapport à la PEH peut être due à la variété lexicale inférieure et au degré de calque plus élevé généralement observés dans les PEH. Le nouveau paradigme de la post-édition automatique à l’aide d’un grand modèle de langage (LLM) s’ajoute en tant que point de comparaison. Quelques études ont montré le potentiel des LLM quand il s’agit de rendre plus fluides ou naturelles des traductions automatiques (TA) à l’aide de prompts assimilables à une étape de post-édition (Chen et al., 2024 ; Li et al., 2025), et ils semblent avoir tendance à augmenter la variété lexicale et syntaxique par rapport aux traductions automatiques (Macken, 2024), voire aux PEH (Farrell, 2023)
Effet de la post-édition automatique et des stratégies de prompting sur les caractéristiques linguistiques de traductions d'éditoriaux de l'anglais vers le français
peer reviewedThis article consists of a comparison of the raw machine translation (MT) and automatic post-editing (APE) of English editorials into French using three systems for MT (DeepL, Google Translate and GPT-4o) and three prompts with different levels of instruction specificity for APE. According to linguistic metrics, lexical and syntactic diversity increase across the board in APE as compared to MT, with the least constrained prompt generally leading to higher gains than the most constrained one. Meanwhile, quality estimation scores follow an opposite trend: less constrained prompts achieve lower COMETKiwi scores. A qualitative comparison of some machine translated and automatically post-edited excerpts shows that APE can correct MT errors, reduce disfluencies or calques and lead to more natural translations overall, but may also introduce new errors.El presente artículo consiste en una comparación entre traducción automática (TA) en bruto y posedición automática (PA). Se recupera la TA de tres sistemas (DeepL, Google Translate y GPT-4o) y se usan tres prompts que presentan instrucciones con distintos niveles de precisión para obtener las versiones poseditadas automáticamente. En cuanto a las métricas lingüísticas, el grado de diversidad léxica y sintáctica aumenta sistemáticamente en la PA en comparación con la TA en bruto: el prompt que presenta el menor grado de restricción desemboca generalmente en mejoras más marcadas que el de mayor grado de restricción. Sin embargo, las métricas de evaluación de la calidad producen resultados contrarios: los prompts con menor grado de restricción obtienen puntuaciones COMETKiwi inferiores. La comparación cualitativa de unos fragmentos de texto traducidos automáticamente con fragmentos poseditados automáticamente demuestra que la PA puede corregir los errores producidos por la TA, reducir la falte de fluidez y los calcos, y proponer traducciones más naturales, aunque también pueda llega a introducir errores nuevos.Aquest article consisteix a comparar una traducció automàtica (TA) en brut i una postedició automàtica (PA). Es recupera la TA de tres sistemes (DeepL, Google Translate i GPT-4o) i s’utilitzen tres prompts que presenten instruccions amb diversos nivells de precisió per obtenir les versions posteditades automàticament. Quant a les mètriques lingüístiques, el grau de diversitat lèxica i sintàctica augmenta sistemàticament a la PA en comparació amb la TA en brut: el prompt que presenta el menor grau de restricció desemboca generalment en millores más marcades que el de major grau de restricció. Tanmateix, les mètriques d’avaluació de la qualitat produeixen resultats contraris: els prompts amb menor grau de restricció obtenen puntuacions COMETKiwi inferiors. La comparació qualitativa d’uns fragments de text traduïts automàticament amb fragments posteditats automàticament demostra que la PA pot corregir els errors produïts per la TA, reduir la falta de fluidesa i els calcs, i proposar traduccions més naturals, tot i que també pugui arribar a introduir errors nous.Cet article consiste en une comparaison de la traduction automatique brute (TA) et de la post-édition automatique (PEA) vers le français d'éditoriaux en anglais. Trois systèmes (DeepL, Google Translate et GPT-4o) et trois prompts aux niveaux de spécificité variables ont été utilisés pour la TA et le PEA, respectivement. Dans l'ensemble, les métriques linguistiques montrent une augmentation de la diversité lexicale et syntaxique dans la PEA par rapport à la TA. Le prompt le moins spécifique amène généralement des gains plus élevés que le prompt le plus détaillé. La métrique d'estimation de la qualité COMETKiwi suit une tendance inverse : les prompts moins spécifiques entraînent des scores plus bas. La comparaison qualitative d'extraits traduits et post-édités automatiquement indique que la PEA peut corriger des erreurs présentes dans la TA, réduire le nombre de formulations peu fluides ou calquées et générer des traductions globalement plus naturelles, mais cette étape peut aussi introduire de nouvelles erreurs
Travail de comparaison de traductions : traductions de The Call of Cthulhu, The Whisperer in Darkness et The Shadow Out of Time de H.P. Lovecraft par Jacques Papy et François Bon
Les nouvelles de Howard Phillips Lovecraft ont marqué l’histoire de la littérature de l’horreur et du surnaturel à la fois par leur style et leurs thèmes novateurs. Elles ont notamment été traduites de nombreuses fois en français à plusieurs années d’intervalle. Dans le présent travail, la première et la dernière traduction de The Call of Cthulhu, The Whisperer in Darkness et The Shadow Out of Time, données respectivement par Jacques Papy et François Bon, seront comparées à l’aune du respect du sens et des caractéristiques stylistiques de Lovecraft. La comparaison sera précédée d’une courte biographie de l’auteur et des traducteurs, d’un résumé des nouvelles en question et d’une brève analyse du style de Lovecraft afin de fournir au lecteur toutes les clés permettant d’appréhender la comparaison des traductions. L’objectif sera de rendre compte des stratégies de traduction de Jacques Papy et François Bon, de déterminer leurs conséquences et effets pour ensuite les confronter à l’hypothèse de la retraduction de Bensimon et Berman
Calque, Lexical Variety and Style in Translation: How Do Human and ChatGPT Post-Editing Compare to Human Translation?
This research compares the style of human translations (HT) with that of post-edited machine translations. It examines the impact of human post-editing (HPE) and ChatGPT-4 post-editing (GPT-PE) on the stylistic quality of press articles translations from English to French, with a focus on lexical variety and calques. The presentation will be aiming at further describing pre-existing research, our methodology, and the expected results.
It has been shown that HPE features more interference and less lexical variety than HT (Toral, 2019: 279; Vanmassenhove et al., 2019: 9; Martikainen & Mestivier, 2020). However, in another recent small-scale preliminary study, lexical variety was higher in the GPT-PE than in the HT (Farrell, 2023: 112). In other studies exploring the differences between HT and HPE, evaluators had a slight (Martikainen & Kübler, 2016: 12-13) or a clear (Jia et al., 2019: 76) preference for HT over HPE, although the measured accuracy and fluency were generally similar. In this context, the objective is to determine a) whether the aforementioned findings also apply in this research and in the English-French language pair, where applicable, and b) if so, whether lexical variety and calques correlate with the perceived style and translation preferences, since accuracy and fluency do not appear to be decisive factors.
Methodologically, the investigation is at the crossroads of previous research conducted by Martikainen and Kübler (2016), Loock (2018), and Farrell (2023). It will involve the compilation of a corpus of press articles in English, their human translations, and their post-edited versions, both by humans and by ChatGPT-4. A quantitative analysis will be conducted using mesures such as MATTR and ASTrED to identify calques and measure lexical variety. A survey will also be carried out among both professionals and non-professionals to assess the perceived stylistic quality of the translations
Going Beyond Counting First Authors in Author Co-citation Analysis
The present study examines one of the fundamental aspects of author co-citation analysis (ACA) - the way co-citation
counts are defined. Co-citation counting provides the data on which all subsequent statistical analyses and mappings
are based, and we compare ACA results based on two different types of co-citation counting - the traditional type that
only counts the first one among a cited work's authors on the one hand and a non-traditional type that takes into
account the first 5 authors of a cited work on the other hand. Results indicate that the picture produced through this non-traditional author co-citation counting contains more coherent author groups and is therefore considerably clearer. However, this picture represents fewer specialties in the research field being studied than that produced through the traditional first-author co-citation counting when the same number of top-ranked authors is selected and analyzed. Reasons for these effects are discussed
Variations on the Author
“Variations on the Author” discusses two of Eduardo Coutinho’s recent films (Um Dia na Vida, from 2010, and Últimas Conversas, posthumously released in 2015) and their contribution to the general question of documentary authorship. The director’s filmography is characterized by a consistent yet self-effacing form of authorial self-inscription: Coutinho often features as an interviewer that rather than express opinions propels discourses; an interviewer that is good at listening. This mode of self-inscription characterizes him as an author who is not expressive but who is nonetheless markedly present on the screen. In Um Dia na Vida, however, Coutinho is completely absent form the image, while Últimas Conversas, on the contrary, includes a confessional prologue that moves the director from the margins to the center of his films. This article examines the ways in which these works stand out in the filmography of a director who offers new insights into the notion of cinematic authorship
Appropriate Similarity Measures for Author Cocitation Analysis
We provide a number of new insights into the methodological discussion about author cocitation analysis. We first argue that the use of the Pearson correlation for measuring the similarity between authors’ cocitation profiles is not very satisfactory. We then discuss what kind of similarity measures may be used as an alternative to the Pearson correlation. We consider three similarity measures in particular. One is the well-known cosine. The other two similarity measures have not been used before in the bibliometric literature. Finally, we show by means of an example that our findings have a high practical relevance.information science;Pearson correlation;cosine;similarity measure;author cocitation analysis
Dispelling the Myths Behind First-author Citation Counts
We conducted a full-scale evaluative citation analysis study of scholars in the XML research field to explore just how different from each other author rankings resulting from different citation counting methods actually are, and to demonstrate the capability of emerging data and tools on the Web in supporting more realistic citation counting methods. Our results contest some common arguments for the continued
use of first-author citation counts in the evaluation of scholars, such as high correlations between author rankings by first-author citation counts and other citation
counting methods, and high costs of using more realistic citation counting methods that are not well-supported by the ISI databases. It is argued that increasingly available digital full text research papers make it possible for citation analysis studies to go beyond what the ISI databases have directly supported and to employ more
sophisticated methods
- …
