ELPUB Digital Repository
Not a member yet
774 research outputs found
Sort by
A CONCEPTUAL ANALYSIS OF FUNCTIONS, PROCESSES, AND PRODUCTS IN THE SCHOLARLY COMMUNICATION CHAIN
This paper summarizes the user needs, functions and information processes in the scholarly communication chain, with an emphasis on the publisher?s role. Furthermore, it presents how product functionality, core benefits, attributes and features relate to each other and how a publisher can seek a competitive edge by differentiation of products and in processes executed. It shows how the functions and information processes are interrelated with the user needs and how an analytical approach can be
adopted towards these needs when developing a product
RE-ENGINEERING THE SCIENTIFIC PUBLISHING PROCESS FOR THE "INTERNETWORKED" GLOBAL ACADEMIC COMMUNITY
The SciX (Open, self organising repository for scientific information exchange) project is funded by the European Commission in order to demonstrate the feasibility of new alternative models of scientific publishing made possible by the Internet. The project builds upon the previous
experience of some of the partners in running an electronic peer reviewed journal and in setting up an e-prints archive. The project includes both theoretical work in making a formal model of the scientific publishing process, to be used as a basis for studying the life-cycle costs of alternative business models, and a demonstrator of a functioning e-prints archive
USE OF TEXT MINING METHODS IN A DIGITAL LIBRARY
The article deals with use of Itemsets classifier based on inductive machine learning in the context of digital library environment. We
provide a brief description of a real-world digital library implemented at a power utility. Its implementation and operating experience have motivated our research in inductive machine learning methods for text mining described in the paper.
Being inspired by mining of association rules, we have developed a new categorization method named ?Itemsets classifier?. By performing various
experiments we have proved its ability to surpass some well-known categorization methods, both in terms of precision/recall and efficiency. As the
task of classification is closely related to clustering, we have integrated the principles of Itemsets method into a new document-clustering algorithm as well. We are also presenting other Itemsets classifier applications in unsolicited
mail filtering and enhancement of the Naive Bayes classifier. Main ideas and experimental results are presented in the paper
NEW PUBLISHING MODELS FOR THE SLAVONIC MEDIAEVAL MANUSCRIPTS
The paper presents the experience of our team in using methods for electronic presentation of mediaeval manuscripts in Bulgaria.
The advantages of the electronic edition are:
- providing high quality of the images
- facilitating and widening the access (Slavonic manuscripts and early printed books are spread in different libraries in Europe and former
USSR. Because of the scattered holdings of the manuscripts, danger of additional harm, and specific conservation conditions of the fragile
and valuable manuscripts, the access to them is strictly limited and sometimes even impossible for specialists).
- providing presentation which can be used for comparative research and analysis
- presenting material for distance learning
Therefore the electronic presentation of the cultural heritage is the only way to make it accessible from everywhere without being harmed. We hope that the experience from this project will will boost future endeavours
ELECTRONIC DOCUMENTS - NEW ROLES AND VALUE-ADDED CONTENTS
The progressive research and development in computing and information processing technologies have stimulated the emergence of a new range of electronic documents that are expected to fully exploit the electronic medium?s basic properties of added interactivity and malleability.
Coupled with the a?liation with metadata and additional layers of e-contents, e-documents are engaging a new paradigm of users who are
actively seeking, selecting and constructing information more than they are ?receiving? information contained within e-documents and collections of such documents. This paper reports on the pertinent findings of a research that intentionally took an immense leap in the design and empirical evaluation of an information environment for a new generation of e-documents
that is endowed with a set of innovative interactions and electronic enhancements.
Apart from presenting a list of essential features of future edocuments that will serve as useful design guidelines, the research advocates
that the way forward is not to promote e-documents either as replacements or as alternatives to their print counterparts, but rather to continually strive to focus on the needs of users and on the tasks they will likely seek to accomplish. Future work on designing e-documents should thus focus on designing for media complementarity, and aim to further develop and incorporate tools and features that are most helpful to the user community
IMPROVING INFORMATION RETRIEVAL IN DIGITAL THESES USING METADATA
In this paper, we present an approach to improve information retrieval in digital scientific theses. This approach consists in defining and
using ?metadata? to help the users to find relevant information during search sessions. This research is one part of the CITHER project (Consultation en Texte Integral des THEses En Reseau) developed by the INSA of Lyon (France). CITHER concerns the online publishing of the INSA?s scientific theses. In a first step, CITHER has permitted the distribution of the theses in PDF format (Portable Document Format), via a server of documents. However, this system does not permit to select only the pertinent parts of the theses during a search session. It is necessary to read the entire document to find them.
In the first part of this paper, we present the initial structure of a thesis stored in the CITHER?s server. Then, we describe our method to define a new structure of document based on ?semantic metadata?. We propose to introduce the concept of ontology to define the semantic ?metadata? and to use XML (eXtended Markup Language) to structure the document. To define the semantic ?metadata?, we extract the main concepts such as ?model?, ?theorem?, ?method?, ?tool?, etc. found in almost all the theses. We formalize these concepts according to an ontology. Therefore, the new model of document includes ?classical metadata? (Dublin Core?s, etc) and ?semantic metadata?, which are both included in the document by the way of XML tags. This model of document is based on a logical and semantic structure. The information retrieval, in our new format of theses, takes place by the use of XML tags.
In the second part of this paper, we will present our prototype and some examples to illustrate our proposition and the results. We will finish this
paper with our conclusion and the future propositions to improve the prototype
FLEXML - AN INTEGRATED DISTANCE LEARNING TOOL TO IMPROVE COGNITIVE FLEXIBILITY
FleXml is an integrated distance learning tool grounded on Cognitive Flexibility Theory principles. We emphasized the knowledge deconstruction process in FleXml. This web learning tool is based on XML,XSL, DOM and ASPs. It has three types of users: the administrator, the readers and the authors. It has two kinds of modules ? author and reader ? that are explained in this paper as well as other functions and its architecture
OPEN WIDE: USING OPEN SANDARDS IN ACADEMIC WEB SERVICES
Despite enormous strides recently, it remains difficult for researchers to find relevant resources that are not ?flat? html files and may
be hidden within web services. This paper will draw upon recent experiences at the UK?s largest academic data centre providing web services to
the education community. Since May 2001 a project has been underway attempting to make the resources more visible by exploiting a number of
relevant open standards and initiatives to ensure interoperability, including:
Dublin Core, XML, Z39.50, Open Archives Initiative, OpenURL and Collection
Descriptions. The aim of the project being to increase the visibility and accessibility of ?appropriate? resources. This principally requires focusing on machine-to-machine metadata interchange. This paper documents some of the realities faced when implementing these potentially valuable, though sometimes ?over-hyped? technologies
TIME-SENSITIVE LINKING MECHANISMS ON THE FUTURE WEB
In our paper we will investigate the possibility of applying the
XML Linking Language (XLink) to time-sensitive applications. In timesensitive
applications fragments of documents are to be composed on fly
to form a time-dependent document assembly. The multiple linking mechanism
specified by the W3C?s XML Linking Language is one way to approach
the time-sensitivity problem in wired and wireless applications. Links with
multiple endpoints connect not only two but also a set of related nodes.
When a user initiates a link with multiple endpoints, he/she can be requested
to choose between the available options or automatically shown
the most decent destination defined by the user?s time profile. We also introduce
scenarios for time-sensitive applications
TOWARDS SCIENTIFIC INFORMATION DISCLOSURE THROUGH CONCEPT HIERARCHIES
We report on an ongoing project aimed at providing an exemplary
architecture for an electronic dissemination environment for scientific
handbooks. We focus on our way of facilitating navigation through and
access to electronic handbooks by using a WordNet-like concept hierarchy
consisting of synsets that are connected to each other and to external
sources by semantic relations for navigational purposes