1,721,021 research outputs found
VQ-BASED WRITTEN LANGUAGE IDENTIFICATION
Humans can recognize different types of written languages by their grammars and vocabularies. However, computers see everything as numbers. We present a computational algorithm for machine classification of written languages using the method of vector quantization. For a language document, each word is converted to a sequence of numbers and forms as a vector of numerical values according to its characters. This collection of vectors is then represented by a codebook that contains a number of template vectors for classification. The proposed method is more effective for machine learning than the n-gram based method, which has been widely used for written language identification. Experimental results of classifying a set of five closely roman-typed scripts show the promising application of the proposed method.Griffith Sciences, School of Information and Communication TechnologyNo Full Tex
Going Beyond Counting First Authors in Author Co-citation Analysis
The present study examines one of the fundamental aspects of author co-citation analysis (ACA) - the way co-citation
counts are defined. Co-citation counting provides the data on which all subsequent statistical analyses and mappings
are based, and we compare ACA results based on two different types of co-citation counting - the traditional type that
only counts the first one among a cited work's authors on the one hand and a non-traditional type that takes into
account the first 5 authors of a cited work on the other hand. Results indicate that the picture produced through this non-traditional author co-citation counting contains more coherent author groups and is therefore considerably clearer. However, this picture represents fewer specialties in the research field being studied than that produced through the traditional first-author co-citation counting when the same number of top-ranked authors is selected and analyzed. Reasons for these effects are discussed
Variations on the Author
“Variations on the Author” discusses two of Eduardo Coutinho’s recent films (Um Dia na Vida, from 2010, and Últimas Conversas, posthumously released in 2015) and their contribution to the general question of documentary authorship. The director’s filmography is characterized by a consistent yet self-effacing form of authorial self-inscription: Coutinho often features as an interviewer that rather than express opinions propels discourses; an interviewer that is good at listening. This mode of self-inscription characterizes him as an author who is not expressive but who is nonetheless markedly present on the screen. In Um Dia na Vida, however, Coutinho is completely absent form the image, while Últimas Conversas, on the contrary, includes a confessional prologue that moves the director from the margins to the center of his films. This article examines the ways in which these works stand out in the filmography of a director who offers new insights into the notion of cinematic authorship
Appropriate Similarity Measures for Author Cocitation Analysis
We provide a number of new insights into the methodological discussion about author cocitation analysis. We first argue that the use of the Pearson correlation for measuring the similarity between authors’ cocitation profiles is not very satisfactory. We then discuss what kind of similarity measures may be used as an alternative to the Pearson correlation. We consider three similarity measures in particular. One is the well-known cosine. The other two similarity measures have not been used before in the bibliometric literature. Finally, we show by means of an example that our findings have a high practical relevance.information science;Pearson correlation;cosine;similarity measure;author cocitation analysis
Nouvelle technique de localisation permettant d'atténuer l'impact des erreurs de NLoS sur le positionnement du mobile
Nous proposons une nouvelle technique de localisation de mobile utilisant les mesures de RTT "Round Trip Time" en UMTS-FDD [1]. Le RTT représente le temps qu'effectue un signal pour faire un aller-retour entre le mobile et la station de base (SB). Quand certaines des mesures correpondent à des trajets non directs (Non-Line-of-Sight "NLoS"), les erreurs de localisation peuvent être très grandes. Dans le cas où le nombre de mesures RTT est supérieur à 3, nous proposons une méthode qui atténue l'effet NLoS grâce à un critère de selection qui mesure la meilleure cohérence entre les RTT estimés. Cette méthode ne dépend pas d'une distribution particulière de l'erreur de NLoS et permet au mobile de choisir les trois mesures les plus fiables parmi la totalité des mesures de RTT disponibles. Ceci permet aussi de sélectionner les RTT les moins bruités quand tous les trajets sont directs
Recommended from our members
Blind identification of multi-input multi-output system using minimum noise subspace
Multi-Stage Reduced-Rank Adaptive Filter With Flexible Structure
Publication in the conference proceedings of EUSIPCO, Toulouse, France, 200
Estimation du propagateur au quatrième ordre
Cet article introduit, dans le contexte de la localisation de sources, une nouvelle technique d'estimation du Propagateur à partir des statistiques d'ordre quatre. Le Propagateur est un opérateur linéaire qui dépend des paramètres de propagation et qui permet la détermination du sous-espace bruit sans aucune décomposition propre de la matrice interspectrale des signaux reçus. A la différence des techniques classiques d'estimation du Propagateur à partir des statistiques d'ordre deux, cette technique fournit un estimateur asymptotiquement non biaisé. Les performances asymptotiques delà méthode proposée ont été développées. Des simulations numériques illustrent la validité de la méthode dans des contextes difficiles
- …
