1,720,986 research outputs found
Ranke.2 - A Teaching Platform for Digital Source Criticism
Abstract and poster of paper 0623 presented at the Digital Humanities Conference 2019 (DH2019), Utrecht , the Netherlands 9-12 July, 2019
L’infrastruttura CLARIN e il servizio di trascrizione multilingue T-Chain
Il contributo mira a presentare alla comunità italiana degli oralisti
l’infrastruttura di ricerca europea CLARIN e il prototipo di
un servizio di trascrizione multilingue chiamato T-chain. Questo
servizio è stato sviluppato con il supporto di CLARIN da una rete
internazionale di esperti provenienti da diversi campi tecnologici
e umanistici di cui le autrici fanno parte, con l’obiettivo di rafforzare
l’integrazione degli strumenti digitali nella ricerca umanistica,
storica e sociale. Le attività del gruppo di ricerca sono descritte
sul sito .
La descrizione di CLARIN e della T-chain è accompagnata da
riflessioni sui prerequisiti per una collaborazione interdisciplinare,
onde attuare una integrazione consapevole della tecnologia all’interno
della ricerca umanistica. In più, vengono spiegate le funzioni
elementari dei sistemi di riconoscimento vocale e le condizioni
per arrivare ad ottenere i migliori risultati. Sono descritti anche
i rischi e i limiti associati alla dipendenza dal settore privato
per questo tipo di tecnologia. Il contributo si conclude con un esempio
diretto: come usare il servizio di trascrizione partendo da una
intervista condotta in italiano dalla prima autrice
CLARIN Resource Families for Oral History
The CLARIN Resource Families (CRF) initiative provides manually curated overviews of prominent language resources and technologies deposited across the distributed CLARIN infrastructure (Lenardič and Fišer 2022). The main aim of CRF is to support other core services of CLARIN from the perspective of the FAIR principles (Wilkinson et al. 2016). CRF enhances the findability and accessibility of CLARIN resources by collating them under their most common typological characteristic. The initiative facilitates re-use by providing comprehensive descriptions tailored to the unique technical features of each of the families, as well as their qualitative characteristics. Furthermore, CRF provides a funding instrument for external projects to contribute new overviews.
Though originally focused on written corpora (e.g., corpora of parliamentary proceedings, corpora of academic texts), in 2022, CRF was expanded to include corpora of oral history. At present one collection is currently featured – the Ravensbrück corpora (Calamai et al. 2022a) – whose creation was supported by the aforementioned CRF funding instrument. This corpus family contains 8 collections of recorded interviews with survivors of the female concentration camp Ravensbrück, conducted in different languages, such as English, German, Hebrew, and French. See https://www.clarin.eu/resource-families/oral-history-corpora. One collection is available for download (Collection Bruzzone; see Bruzzone and Beccaria Rolfi 1976) while the others can be streamed online.
The inclusion of the Ravensbrück corpora in CRF represents an illustrative example of how the CLARIN infrastructure incorporates and provides documentation for complex objects like oral history sources whose provenance and metadata documentation widely differ from standard written corpora and even from contemporary interviews born digitally. The team working on the Ravensbrück resource family (see Calamai et al. 2022b) availed themselves of CLARIN’s Component Metadata Infrastructure (CMDI), which is a framework for metadata description that “supports flexible definitions of metadata structure and semantics” by allowing researchers to “create and use their own [metadata] schema tailored specifically towards the requirements of [their] project” (Windhouwer and Goosen 2022: 194 and 199). All the 8 collections within the Ravensbrück family are accompanied by extensive CMDI metadata, prepared by Calamai et al. (2022a,b).
The peculiarity of the interviews in the Ravensbrück family is that they were mostly recorded on an analogue carrier (i.e., audio cassettes), so a new CMDI metadata profile was created that is tailored to such legacy interviews not born digitally. This metadata profile has additional components describing “information about the context in which the interviews were conducted” as well as “information about the process of digitisation” (Calamai et al. 2022a: 3). Being thus digitised, comprehensively described, and carefully curated, the Ravensbrück corpora present a unique opportunity to study and compare these historical interviews. To facilitate their use in research, CLARIN offers through its Speech data and Technology network (Draxler et al. 2020) an open-source web application called TranscriptionPortal (https://speechandtech.eu/transcription-portal), where certain audio recordings (e.g., Collection Bruzzone, United States Holocaust Memorial Museum) can be uploaded and then orthographically transcribed on the fly, with manual phonetic and word alignment for a variety of languages
Ravensbrück interviews: how to curate legacy data to make it CLARIN compliant
This paper describes the preparatory phase of a CLARIN-funded project called ‘Voices from Ravensbrück’, which aims to introduce a new type of corpus in the CLARIN resource family called ‘Oral Histories’. The first task consisted in curating and transcribing a set of interviews conducted by the Italian author A.M. Bruzzone with five Italian survivors of the Ravensbrück concentration camp back in 1977. This posed considerable challenges inherent in integrating legacy data from the pre-digital era in the CLARIN infrastructure. The second task was ex-ploring the potential of automatic speech transcription for this type of oral history data. The third element of this exploratory phase was identifying potential partners and suitable data for creating a multilingual collection of existing oral history interviews with survivors of concen-tration camp Ravensbrück. These preparatory steps were necessary to move to the final phase of our project and realise our overall objective of creating a resource family compliant with CLARIN standards and enabling scholars to analyse interviews from a comparative multilin-gual and multidisciplinary perspectiv
A Multidisciplinary Approach To The Use Of Technology In Research: The Case Of Interview Data
Abstract of paper 1063 presented at the Digital Humanities Conference 2019 (DH2019), Utrecht , the Netherlands 9-12 July, 2019
Going Beyond Counting First Authors in Author Co-citation Analysis
The present study examines one of the fundamental aspects of author co-citation analysis (ACA) - the way co-citation
counts are defined. Co-citation counting provides the data on which all subsequent statistical analyses and mappings
are based, and we compare ACA results based on two different types of co-citation counting - the traditional type that
only counts the first one among a cited work's authors on the one hand and a non-traditional type that takes into
account the first 5 authors of a cited work on the other hand. Results indicate that the picture produced through this non-traditional author co-citation counting contains more coherent author groups and is therefore considerably clearer. However, this picture represents fewer specialties in the research field being studied than that produced through the traditional first-author co-citation counting when the same number of top-ranked authors is selected and analyzed. Reasons for these effects are discussed
Variations on the Author
“Variations on the Author” discusses two of Eduardo Coutinho’s recent films (Um Dia na Vida, from 2010, and Últimas Conversas, posthumously released in 2015) and their contribution to the general question of documentary authorship. The director’s filmography is characterized by a consistent yet self-effacing form of authorial self-inscription: Coutinho often features as an interviewer that rather than express opinions propels discourses; an interviewer that is good at listening. This mode of self-inscription characterizes him as an author who is not expressive but who is nonetheless markedly present on the screen. In Um Dia na Vida, however, Coutinho is completely absent form the image, while Últimas Conversas, on the contrary, includes a confessional prologue that moves the director from the margins to the center of his films. This article examines the ways in which these works stand out in the filmography of a director who offers new insights into the notion of cinematic authorship
Appropriate Similarity Measures for Author Cocitation Analysis
We provide a number of new insights into the methodological discussion about author cocitation analysis. We first argue that the use of the Pearson correlation for measuring the similarity between authors’ cocitation profiles is not very satisfactory. We then discuss what kind of similarity measures may be used as an alternative to the Pearson correlation. We consider three similarity measures in particular. One is the well-known cosine. The other two similarity measures have not been used before in the bibliometric literature. Finally, we show by means of an example that our findings have a high practical relevance.information science;Pearson correlation;cosine;similarity measure;author cocitation analysis
Dispelling the Myths Behind First-author Citation Counts
We conducted a full-scale evaluative citation analysis study of scholars in the XML research field to explore just how different from each other author rankings resulting from different citation counting methods actually are, and to demonstrate the capability of emerging data and tools on the Web in supporting more realistic citation counting methods. Our results contest some common arguments for the continued
use of first-author citation counts in the evaluation of scholars, such as high correlations between author rankings by first-author citation counts and other citation
counting methods, and high costs of using more realistic citation counting methods that are not well-supported by the ISI databases. It is argued that increasingly available digital full text research papers make it possible for citation analysis studies to go beyond what the ISI databases have directly supported and to employ more
sophisticated methods
- …
