International Journal of Digital Curation
Not a member yet
605 research outputs found
Sort by
Citation and Peer Review of Data: Moving Towards Formal Data Publication
This paper discusses many of the issues associated with formally publishing data in academia, focusing primarily on the structures that need to be put in place for peer review and formal citation of datasets. Data publication is becoming increasingly important to the scientific community, as it will provide a mechanism for those who create data to receive academic credit for their work and will allow the conclusions arising from an analysis to be more readily verifiable, thus promoting transparency in the scientific process. Peer review of data will also provide a mechanism for ensuring the quality of datasets, and we provide suggestions on the types of activities one expects to see in the peer review of data. A simple taxonomy of data publication methodologies is presented and evaluated, and the paper concludes with a discussion of dataset granularity, transience and semantics, along with a recommended human-readable citation syntax
Editorial
Kevin Ashley, Chief Editor, introduces Volume 6, Issue 1 (2011) of the International Journal of Digital Curation
Use and Impact of UK Research Data Centres
UK data centres are an important part of efforts to gain maximum value from research data. However, if they are to operate effectively, the services that they provide must be based upon an understanding of researchers\u27 practices and needs. Furthermore, in order to build a case for ongoing funding, data centres must be able to demonstrate their value to researchers work and, increasingly, their contribution to wider political "impact" agendas. This paper presents the findings of a survey of users of five UK data centres. It suggests that research data centres are highly valued by their users. Benefits appear to be particularly strong around improving research efficiency, especially access to data. Data centres are less important in terms of stimulating novel research questions. Despite a few interesting cases of observable impact, in the main it remains difficult to understand the wider reach of research which draws upon data centre resources
The Human Face of Digital Preservation: Organizational and Staff Challenges, and Initiatives at the Bibliothèque nationale de France
The process of setting up a digital preservation repository in compliance with the OAIS model is not only a technical challenge: libraries also need to develop and maintain appropriate skills and organizations. Digital activities, including digital preservation, are nowadays moving into the mainstream activity of the Library and are integrated in its workflows.The Bibliothèque nationale de France (BnF) has been working on the definition of digital preservation activities since 2003. This paper aims at presenting the organizational and human resources challenges that have been faced by the library in this context, and those that are still awaiting us.The library has been facing these challenges through a variety of actions at different levels: organizational changes, training sessions, dedicated working group and task forces, analysis of skills and processes, etc. The results of these actions provide insights on how a national library is going digital, and what is needed to reach this longstanding goal
TextGrid – Virtual Research Environment for the Humanities
The TextGrid research group, a consortium of 10 research institutions in Germany, is developing a virtual research environment for researchers in the arts and humanities that provides services and tools for the analysis of text data and supports the curation of research data by means of grid technology. The TextGrid virtual research environment consists of two main components: the TextGrid Laboratory (TextGridLab), which serves as the entry point to the virtual research environment, and the TextGrid Repository (TextGridRep), which is a long-term humanities data archive ensuring sustainability, interoperability and long-term access to research data. To support all stages of the research lifecycle, preserve and maintain research data, and ensure its long-term usefulness, existing research practices must be supported. Therefore the TextGridLab provides common functionalities in a sustainable environment to intensify re-use of data, tools, and services, and the TextGridRep enables researchers to publish and share their data in a way that supports long-term availability and re-usability
Making Sense: Talking Data Management with Researchers
Incremental is one of eight projects in the JISC Managing Research Data programme funded to identify institutional requirements for digital research data management and pilot relevant infrastructure. Our findings concur with those of other Managing Research Data projects, as well as with several previous studies. We found that many researchers: (i) organise their data in an ad hoc fashion, posing difficulties with retrieval and re-use; (ii) store their data on all kinds of media without always considering security and back-up; (iii) are positive about data sharing in principle though reluctant in practice; (iv) believe back-up is equivalent to preservation. The key difference between our approach and that of other Managing Research Data projects is the type of infrastructure we are piloting. While the majority of these projects focus on developing technical solutions, we are focusing on the need for ‘soft’ infrastructure, such as one-to-one tailored support, training, and easy-to-find, concise guidance that breaks down some of the barriers information professionals have unintentionally built with their use of specialist terminology.We are employing a bottom-up approach as we feel that to support the step-by-step development of sound research data management practices, you must first understand researchers’ needs and perspectives. Over the life of the project, Incremental staff will act as mediators, assisting researchers and local support staff to understand the data management requirements within which they are expect to work, and will determine how these can be addressed within research workflows and the existing technical infrastructure.Our primary goal is to build data management capacity within the Universities of Cambridge and Glasgow by raising awareness of basic principles so everyone can manage their data to a certain extent. We will ensure our lessons can be picked up and used by other institutions. Our affiliation with the Digital Curation Centre and Digital Preservation Coalition will assist in this and all outputs will be released under a Creative Commons licence
Curating Scientific Research Data for the Long Term: A Preservation Analysis Method in Context
The challenge of digital preservation of scientific data lies in the need to preserve not only the dataset itself but also the ability it has to deliver knowledge to a future user community. A true scientific research asset allows future users to reanalyze the data within new contexts. Thus, in order to carry out meaningful preservation we need to ensure that future users are equipped with the necessary information to re-use the data. This paper presents an overview of a preservation analysis methodology which was developed in response to that need on the CASPAR and Digital Curation Centre SCARP projects. We intend to place it in relation to other digital preservation practices, discussing how they can interact to provide archives caring for scientific data sets with the full arsenal of tools and techniques necessary to rise to this challenge
The Milieu and the MESSAGE: Talking to Researchers about Data Curation Issues in a Large and Diverse e-Science Project
MESSAGE (Mobile Environmental Sensing System Across Grid Environments) was an ambitious, multi-partner, interdisciplinary e-Science research project, jointly funded by the Engineering and Physical Sciences Research Council (EPSRC) and the UK Department for Transport (DfT) between 2006 and 2009. It aimed to develop and demonstrate the potential of diverse, low cost sensors to provide heterogeneous data for the planning, management and control of the environmental impacts of transport activity at urban, regional and national level. During the last year of the project, the Digital Curation Centre (DCC) interviewed and observed members of the project team in order to identify and analyse key aspects of their data-related activities, recording attitudes towards the data that they create and/or re-use. This paper describes the major issues identified over the course of the case study, which are presented in parallel with the perspectives of the project team in order to demonstrate the multiplicity of views that may be projected onto a single dataset. It concludes with a contextualisation of the case study\u27s themes with those of a number of contemporary reports
Assessing the Preservation Condition of Large and Heterogeneous Electronic Records Collections with Visualization
As collections become larger in size, more complex in structure and increasingly diverse in composition, new approaches are needed to help curators assess digital files and make decisions about their long-term preservation. We present research on the use of interactive visualization to analyze file characterization information for the purpose of assessing the preservation condition of a vast collection of complex electronic records. The case study collection contains over 1,000,000 files of diverse formats arranged in varied record structures and record groups. The visualization application uses tree maps and a relational database management system (RDBMS) to represent the collection\u27s arrangement and to show available characterization information at different levels of aggregation, classification and abstraction. Through this visualization interface curators can interact dynamically with the collections\u27 characterization information to discover trends, as well as compare and contrast various file characteristics across the collection. Curators may select and weight the variables that they want to analyze. They can pursue analysis workflows that go from a high-level overview of the collection\u27s preservation condition based on file format risks, to obtaining more detailed results about the condition of record groups and individual records. While there are various digital preservation planning tools available, to our knowledge none have been designed specifically to visually present assessment information across vast and complex collections. We present research to address the need for such a tool
Making Digital Curation a Systematic Institutional Function
Over the past decade, a rich body of research and practice has emerged under the rubrics of electronic records, digital preservation and digital curation. Most of this work has taken place as research activity (often financed by government agencies) within libraries and information/computer science departments. Many projects focus on one format of information, such as research publications or data, potentially de-contextualizing individual records. Meanwhile, most institutional archives and manuscript repositories, which possess a rich theoretical and practical framework for preserving context among mixed analog materials, have failed to extend their capabilities to digital records. As a result, relatively few institutions have implemented systematic methods to capture, preserve and provide access to the complete range of documentation that end users need to understand and interpret past human activity.The Practical E-Records Method attempts to address this problem by providing easy-to-implement software reviews, guidance/policy templates, and program recommendations that blend digital curation research findings with traditional archival processes and workflows. Using the method discussed in this paper, archives and manuscript repositories can use existing resources to incrementally develop digital curation skills, building a collaborative, expanding program in the process. Archival programs that make digital curation a systematic institutional function will systematically gather, preserve, and provide access to genres of documentation that are contextually-rich and highly susceptible to loss, complementing efforts undertaken by librarians, information scientists and external service providers. Over the next year, the suggested techniques will be tested and refined at the University of Illinois Archives and possibly elsewhere