ELPUB Digital Repository
Not a member yet
    774 research outputs found

    Encoding and Querying Multi-Structured Documents

    No full text
    This paper concerns the document multi-structuring issue. For various use objectives, many distinct structures may be defined simultaneously for the same original document. For example, a document may have a first structure for logical content organisation (logical structure), and a second structure to express a set of content formatting rules (physical structure). We have already proposed a generic model, called MSDM, for the multi-structured documents, in which several important features were established. In this paper, we address the encoding problem of this kind of documents. We present a new formalism, called MultiX, which allows encoding the multi-structured documents efficiently. This formalism is based on the MSDM model and uses XML syntax

    Publishing Multilingual Ontologies: A Quick Way of Obtaining Feedback

    No full text
    Dictionaries and Thesauri are valuable resources for Natural Language Processing but are not as widely available as one would hope, especially for languages other than English, and most can only be used for querying online. Our main goal with T2O - Thesaurus to Ontology framework - is to create a multilingual ontology: -- freely available online and for downloading; with a computer-readable format; with a good API; with a structure as rich as possible; reusing all the structured information we can get

    Webreview: The Evolution of the Algerian E-Journals

    No full text
    Scientific journals have always been an important tool in scientific research, and the advent of new technologies has fostered their development into e-journals that has brought about new editorial techniques and methods. Thus, e-journals have become the easiest and fastest means to meet the needs of researchers in their research works as the Internet and its services represent a tremendous opportunity for communication, edition and information retrieval. The Algerian \u27savoir-faire\u27 in this field has led to the setting up of an e-journals repository with a double experience that has paved the way to using the open source editorial system SPIP. This paper deals with the Algerian experience starting from 1999 by creating a national database for scientific journals that had to be accessible online for the researchers\u27 community. Even if it was relatively underestimated, it encouraged our team to go ahead and look for new tools to enhance the data base, its content and also the web site. After many tests on a few electronic publishing software such as LODEL and GREENSTONE, SPIP appeared to be the best one to meet our needs especially that it includes Arabic. Furthermore, we are keen to start the process of adhering to interoperability international standards in order to make our contents more accessible. In addition, a study has been initiated on the archiving and long term preservation issue, particularly with an XML solution

    Open Access in Sweden 2002-2005

    No full text
    The paper gives an overview of recent developments in Sweden concerning the communication and reception of the Open Access concept and the growth and co-ordination of digital academic repositories. The Open Access concept was primarily introduced in Sweden from library circles which also coloured the arguments used. University leaders started to show an interest from 2002 and their association, the SUHF, issued a report favorable to Open Access in 2003. In the consecutive years both the SUHF and the Swedish Research Council signed the Berlin declaration. The relatively swift change of policy towards a support of Open Access by these two important stakeholders was obviously influenced by the international discussion but also by the way the issue was raised in Sweden and by the absence of a national publisher lobby arguing against Open Access. The development of e-publishing within Swedish higher education started to gather momentum in the years following 2000. Universities either chose to particicipate in the DiVA consortium or to implement available Open Source software. In 2003 the national SVEP project was launched to co-ordinate and scale up the development of e-publishing within higher education. Recommendations were issued for metadata descriptions of e-publications, for publication databases and for subject categories. The SUHF gave support to the metadata recommendations for publication databases (local registers of academic publications) which clears the way for a co-ordination between publication databases and freely available full-text material in repositories. Important building blocks of a generalized archiving workflow between a local repository and a national archive were implemented. A resolution service for the use of permanent identifiers and a service for administrating name spaces in URN:NBN were established. A national search service for undergraduate theses was developped and put into operation at the LIBRIS website. The service uses the OAI-PMH to harvest local repositories who have to implement a metada model based on simple Dublin Core. A new development programme to support Open Access in Sweden is now being initiated. It will engage the major stakeholders in Sweden and will have a clear focus on promoting the growth of the volume and diversity of content in academic repositories

    Open Access Publishing in Finland: Discipline Specific Publishing Patterns in Biomedicine and Economics

    No full text
    Open access publishing strategies have traditionally been directed towards what has been regarded as a homogenous scientific community of universities, researchers and libraries. However, discipline specific practices in communication and publishing strategies are prevailing in different scientific areas. In this study we, argue that discipline specific publishing patterns may affect the ways that open access strategies can be adopted in different scientific areas. We characterise and identify incentives for publishing open access into factors depending mostly on the social environment and factors mostly depending on personal factors of the researcher. In the case study comparing the field of biomedicine and economics and business administration we were able to find out figures on the proportion, type and channel of open access publishing of scientific articles by Finnish researchers in economics and medicine

    Identifying Trends in Accessible Content Processing

    No full text
    The European Accessible Information Network (EUAIN) is currently examining issues relating to accessible content processing. With the support of key European publishers, it has been possible to begin to identify key trends in accessible content processing which are likely to be of some importance in the coming years. This paper also describes some related standards activities and identifies a need for accessibility to be embedded within content creation and production processes at the earliest stages

    Living Reviews - Innovative Resources for Scholarly Communication Bridging Diverse Spheres of Disciplines and Organisational Structures

    No full text
    This contribution presents the concept and analyses the path of diffusion of an innovative publishing idea that originated in one speciality in physics and is now about to spread into other fields, including the social sciences. We discuss the conditions of success and fostering as well as hampering factors on the road for what we consider to be a truly revolutionary concept of presenting the state of knowledge in potentially any given academic discipline. First, we will present the overall idea of Living Reviews (www.livingreviews.org), which are open access online journals featuring an innovative editorial concept for the publication of high-quality scientific content. Second, we discuss the proliferation of this unique concept. Finally, we shall aim at summarising lessons learned from our shared experience in establishing these publishing ventures in our respective fields. Our conclusions give hints for potential future ?cyber-activists? who are considering setting up further Living Review projects

    Transcoding for Web Accessibility for the Blind: Semantics from Structure

    No full text
    True accessibility requires minimizing the scanning time to find a particular piece of information. Sequentially reading web pages do not provide this type of accessibility, for instance before the user gets to the actual text content of the page it has to go through a lot of menus and headers. However if the user could navigate a web page based through semantically classified blocks then the user could jump faster to the actual content of the page, skipping all the menus and other parts of the page. We propose a transcoding engine that tackles accessibility at two distinct, yet complementary, levels: for specific known sites and general unknown sites. We present a tool for building customized scripts for known sites that turns this process in an extremely simple task, which can be performed by anyone, without any expertise. For general unknown sites, our approach relies on statistical analysis of the structural blocks that define a web page to infer a semantic for the block

    WORLDWIDE "COMMUNITARIAN" ONLINE PUBLISHING: AN EXERCISE IN WISHFUL THINKING

    No full text
    Regretfully, this paper provides more questions than answers. Its main issue is how to balance the enthusiasm of the general public for web tools (wikies, blogs etc.) to provide content with the "authority" of the established content-providing (and content-preserving) institutions (i.e., universities, libraries, museums), in order to get the best possible worldwide resources. The central problem of the code of behaviour will be discussed - mainly for the reference resources -, including the related issues of editorial control and neutral point of view. Also the intellectual rights in a collaborative community and the problems of anonymity and "loose" pseudonyms will be considered. In the case of a public reference resource, the right balance between the institutionally controlled core and the (moderated) contributions of the public has to be carefully considered. How to motivate voluntary contributors? How to balance the goal of building a useful, reliable resource with the - less intellectually profitable, but culturally significant - effort to stimulate the people to research, document and write? A discussion will follow on how to publish online "problematic" resources (e.g., xenophobic texts by important authors), i.e., do we publish well online "critical editions"? The keynote speaker is not aware of really satisfactory solutions. In a "physical" book we can wrap conveniently the problematic text in a "critical envelope" (such as an explanatory introduction, critical footnotes etc.). Online, this is more problematic: the lack of "physical" boundaries on the web makes the critical apparatus less visible. Novel ways of presentation are required and some suggestions of making the critical annotations sharing the same page with the text (vs. sharing the same volume, in the print world) will be tentatively offered. The keynote address will also try to contribute to the reflection on the selection of materials for the future European Digital Library. Besides it will try to emphasize the distinction between a simple archiving on the web and the specific republishing (i.e., production of specific manifestations - in FRBR terms) in a genuine digital library. Finally, the Romanian project of an online "wish list" of works to be digitised and its mechanism of setting the priorities will be presented. Some "political" considerations on resource allocation for the European digital libraries will conclude the paper

    Text Parsing of a Complex Genre

    No full text
    A text parsing component designed to be part of a system that assists students in academic reading an writing is presented. The parser can automatically add a relational discourse structure annotation to a scientific article that a user wants to explore. The discourse structure employed is defined in an XML format and is based the Rhetorical Structure Theory. The architecture of the parser comprises preprocessing components which provide an input text with XML annotations on different linguistic and structural layers. In the first version these are syntactic tagging, lexical discourse marker tagging, logical document structure, and segmentation into elementary discourse segments. The algorithm is based on the shift-reduce parser by Marcu (2000) and is controlled by reduce operations that are constrained by linguistic conditions derived from an XML-encoded discourse marker lexicon. The constraints are formulated over multiple annotation layers of the same text

    0

    full texts

    774

    metadata records
    Updated in last 30 days.
    ELPUB Digital Repository
    Access Repository Dashboard
    Do you manage Open Research Online? Become a CORE Member to access insider analytics, issue reports and manage access to outputs from your repository in the CORE Repository Dashboard! 👇