1,720,969 research outputs found

    Perspektiv på prediktiv- och etikettosäkerhet i probabilistisk maskininlärning

    No full text
    Machine learning models are, just as us humans, exposed to the uncertainty of the world. Following the complexity of real-world events, these models are often employed for prediction tasks where there is no single, ground-truth answer, meaning that it may be impossible to determine the precise outcome of the predicted event beforehand. This aleatoric uncertainty is potentially, but not necessarily, a result of the event in question being part of a larger system, where some information remains undisclosed.  Moreover, machine learning models are data-driven and typically learn everything they know from data, called training data. The quality of the training data is vital in deter-mining the extent of a machine learning model’s knowledge and, consequently, how well the model performs on a given task. For instance, when training data is limited, this can result in uncertainty originating from a lack knowledge, often referred to as epistemic uncertainty. Furthermore, collected through observation, or measurements, of real-world events, the training data naturally incorporates the uncertainty inherent to these events. Some-times, additional uncertainty is integrated through the processes used to acquire the data, following, for instance, measurement error or human error. One such type of uncertainty is in this thesis termed annotation uncertainty, and relates to the collection of annotations for training models through supervised learning.  The focus of this thesis lies on probabilistic predictive machine learning models, as an approach to representing different sources of so-called predictive uncertainty, including aleatoric, epistemic and annotation uncertainty. Special attention is given to annotation uncertainty, beginning with an exploration of possible negative effects of this type of uncertainty on the performance of probabilistic predictive models. We analyse how annotation uncertainty, or noise, affects the properties of asymptotic risk minimisers when training models with two different classes of loss functions: strictly proper and a group of previously proposed robust loss functions. The analysis emphasises the importance of considering a model’s ability to accurately estimate predictive uncertainty, also referred to as the model’s reliability, when developing training algorithms robust to annotation noise.  However, under the umbrella of weak supervision, we also provide two examples of when annotation uncertainty can be allowed, to instead benefit model performance. In the first example, we use ensemble models to generate annotations for the training data, with the aim to teach individual probabilistic models to estimate both aleatoric and epistemic uncertainty in their predictions. Having this ability is beneficial in many applications, one of them being active learning, and, notably, the active learning algorithm constituting the second example. This specific active learning algorithm acquires data samples based on high epistemic uncertainty, believed to represent samples for which there is much gain to be made in terms of model performance. The contribution does not lie in the particular approach to acquiring data samples, but instead in introducing the possibility to make a trade-off between annotation costs and quality of annotations, as part of the active learning algorithm. Such a trade-off has the potential to lead to an improved model performance under a fixed annotation budget. The thesis also explores topics beyond annotation uncertainty. First, in the context of learning probabilistic machine learning models, we focus on unnormalised probabilistic models, with energy-based models among them. We establish a link between two groups of important methods used for estimating unnormalised models, namely noise-contrastive estimation and approximate maximum likelihood methods. This link provides an improved under-standing of noise-contrastive estimation and serves to create a more coherent framework for the estimation of unnormalised models. Second, for deeper insights into the generalisation behaviour of machine learning models trained using gradient-based learning, we study the epoch-wise double descent phenomenon in two-layer linear neural networks. With this, we identify additional factors contributing to epoch-wise double descent that has not been observed for the simpler linear regression model, which is commonly central to theoretical studies. Although not specific to probabilistic models, these insights could potentially be extended to such models in the future and used to further explore the interplay between annotation uncertainty and model performance.Funding: This research was financially supported by the Wallenberg AI, Autonomous Systems and Software Program (WASP) funded by the Knut and Alice Wallenberg Foundation, the Excellence Center at Linköping-Lund in Information Technology (ELLIIT), and the Swedish Research Council.2024-09-19: The thesis was first published online. The online published version reflects the printed version.2024-11-19: The thesis was updated with an errata list which is also downloadable from the DOI landing page. Before this date the PDF has been downloaded 119 times.</p

    Going Beyond Counting First Authors in Author Co-citation Analysis

    Get PDF
    The present study examines one of the fundamental aspects of author co-citation analysis (ACA) - the way co-citation counts are defined. Co-citation counting provides the data on which all subsequent statistical analyses and mappings are based, and we compare ACA results based on two different types of co-citation counting - the traditional type that only counts the first one among a cited work's authors on the one hand and a non-traditional type that takes into account the first 5 authors of a cited work on the other hand. Results indicate that the picture produced through this non-traditional author co-citation counting contains more coherent author groups and is therefore considerably clearer. However, this picture represents fewer specialties in the research field being studied than that produced through the traditional first-author co-citation counting when the same number of top-ranked authors is selected and analyzed. Reasons for these effects are discussed

    Validation of a Damage Accumulation Model of Replicative Ageing in S.cerevisiae.

    No full text
    Age-related diseases and conditions give rise to societal challenges and pose a threat to healthy ageing. At the same time, the more recent evolutionary theories of ageing hypothesise that the process of ageing is a consequence of living rather than an evolutionary strategy. Consequently, it is implied that ageing is not as inevitable as many might believe and, as a consequence, it is of interest to study this biological process and its underlying mechanisms. On a cellular level, accumulation of damage is often regarded as the main cause of ageing. Since the basic properties of ageing between unicellular and multicellular organisms are similar on this level, it is common to use the unicellular yeast Saccharomyces cerevisiae as a model organism in the field of ageing research. The aim of this project is to validate a mathematical damage accumulation model of replicative ageing in yeast. The model represents a cell by intact protein and damage and describes how these quantities change as the cell grows. In addition to cell growth, the model takes asymmetric division, retention and cell death into account. For the purpose of validating the model of replicative ageing, structural and numerical identifiability methods are applied and continuous optimisation is performed using single-cell area data. The model is fit to experimental data obtained for wildtype yeast and the two deletion strains sir2 and fob1. Moreover, replicative lifespan data of 4,698 single-gene deletion strains is analysed and, in conjunction to this, it is investigated how the model parameters affect the replicative lifespan of the simulations. The results show that the parameters in the model of replicative ageing that describes the rate of change of intact protein and damage in the cell, are structurally identifiable. In spite of this, they are not numerically identifiable based on the experimental data available; the parameter estimates obtained have high variances and are moderately or highly correlated with each other. Likewise, it is possible to generate parameter sets that make the mathematical model reproduce the replicative lifespans of the investigated strains, if a replicative lifespan constraint is inferred on the optimisation. For future work, it is suggested that new experimental data is generated as to fit the model of replicative ageing to growth curves belonging to cells of later life stages. Ultimately, the data should be sufficient enough for the optimisation to generate parameter sets that make the model adapt to the characteristics of the investigated strains, without having additional constraints added to the objective function

    Variations on the Author

    Get PDF
    “Variations on the Author” discusses two of Eduardo Coutinho’s recent films (Um Dia na Vida, from 2010, and Últimas Conversas, posthumously released in 2015) and their contribution to the general question of documentary authorship. The director’s filmography is characterized by a consistent yet self-effacing form of authorial self-inscription: Coutinho often features as an interviewer that rather than express opinions propels discourses; an interviewer that is good at listening. This mode of self-inscription characterizes him as an author who is not expressive but who is nonetheless markedly present on the screen. In Um Dia na Vida, however, Coutinho is completely absent form the image, while Últimas Conversas, on the contrary, includes a confessional prologue that moves the director from the margins to the center of his films. This article examines the ways in which these works stand out in the filmography of a director who offers new insights into the notion of cinematic authorship

    Appropriate Similarity Measures for Author Cocitation Analysis

    Get PDF
    We provide a number of new insights into the methodological discussion about author cocitation analysis. We first argue that the use of the Pearson correlation for measuring the similarity between authors’ cocitation profiles is not very satisfactory. We then discuss what kind of similarity measures may be used as an alternative to the Pearson correlation. We consider three similarity measures in particular. One is the well-known cosine. The other two similarity measures have not been used before in the bibliometric literature. Finally, we show by means of an example that our findings have a high practical relevance.information science;Pearson correlation;cosine;similarity measure;author cocitation analysis

    On Uncertainty Quantification in Neural Networks: Ensemble Distillation and Weak Supervision

    No full text
    Machine learning models are employed in several aspects of society, ranging from autonomous cars to justice systems. They affect your everyday life, for instance through recommendations on your streaming service and by informing decisions in healthcare, and are expected to have even more influence in society in the future. Among these machine learning models, we find neural networks which have had a wave of success within a wide range of fields in recent years. The success of neural networks are partly attributed to the very flexible model structure and, what it seems, endless possibilities in terms of extensions. While neural networks come with great flexibility, they are so called black-box models and therefore offer little in terms of interpretability. In other words, it is seldom possible to explain or even understand why a neural network makes a certain decision. On top of this, these models are known to be overconfident, which means that they attribute low uncertainty to their predictions, even when uncertainty is, in reality, high. Previous work has demonstrated how this issue can be alleviated with the help of ensembles, i.e. by weighing the opinion of multiple models in prediction. In Paper I, we investigate this possibility further by creating a general framework for ensemble distribution distillation, developed for the purpose of preserving the performance benefits of ensembles while reducing computational costs. Specifically, we extend ensemble distribution distillation to make it applicable to tasks beyond classification and demonstrate the usefulness of the framework in, for example, out-of-distribution detection. Another obstacle in the use of neural networks, especially deep neural networks, is that supervised training of these models can require a large amount of labelled data. The process of annotating a large amount of data is costly, time-consuming and also prone to errors. Specifically, there is a risk of incorporating label noise in the data. In Paper II, we investigate the effect of label noise on model performance. In particular, under an input-dependent noise model, we analyse the properties of the asymptotic risk minimisers of strictly proper and a set of previously proposed, robust loss functions. The results demonstrate that reliability, in terms of a model’s uncertainty estimates, is an important aspect to consider also in weak supervision and, particularly, when developing noise-robust training algorithms. Related to annotation costs in supervised learning, is the use of active learning to optimise model performance under budget constraints. The goal of active learning, in this context, is to identify and annotate the observations that are most useful for the model’s performance. In Paper III, we propose an approach for taking advantage of intentionally weak annotations in active learning. What is proposed, more specifically, is to incorporate the possibility to collect cheaper, but noisy, annotations in the active learning algorithm. Thus, the same annotation budget is enough to annotate more data points for training. In turn, the model gets to explore a larger part of the input space. We demonstrate empirically how this can lead to gains in model performance.Maskininlärningsmodeller används i flera delar av samhället, från autonoma fordon till rättssystem. De påverkar redan nu din vardag, exempelvis via personliga rekommendationer i din direktuppspelningstjänst (”streaming service”) och genom att agera beslutsstöd i vården, och förväntas ha än mer påverkan i samhället i framtiden. Bland dessa maskininlärningsmodeller, finner vi neurala nätverk som har haft stor framgång inom flera fält under det senaste årtiondet. Framgången beror delvis på neurala nätverks flexibla modellstruktur och, vad det verkar, oändliga utvecklingsmöjligheter. Neurala nätverk erbjuder stor flexibilitet, men har en nackdel i att de är så kallade black-box-modeller. Detta innebär att det sällan går att förklara eller ens förstå varför ett neuralt nätverk tar ett visst beslut. Dessutom, så har den här typen av modeller en tendens att vara överdrivet självsäkra, vilket betyder att de rapporterar låg osäkerhet i sina beslut, även när osäkerheten i själva verket är hög. För ett självkörande fordon skulle detta till exempel kunna innebära att fordonet bedömer en vänstersväng som mycket säker, när sikten över det mötande körfältet är skymd och ett mötande fordon mycket väl kan finnas just bakom krönet. Tidigare forskning har demonstrerat hur denna typ av problem kan avhjälpas genom att använda flera neurala nätverk som samspelar för att prediktera eller ta ett beslut. På detta sätt fås en modell som är mer korrekt och som också är mer pålitlig när det kommer till att ge en uppskattning av den egna osäkerheten. I denna avhandling undersöker vi vidare hur vi kan lära ett enskilt neuralt nätverk att efterlikna flera samspelande modeller, för att minska de kostnader som kommer med att ha flera samspelande modeller i bruk. En annan begränsande faktor när det kommer till neurala nätverk är att de kan behöva en stor mängd insamlad data med tillhörande etiketter för att lära sig den uppgift som de är ämnade för. Att införskaffa etiketter för en stor mängd datapunkter är både kostsamt och tidskrävande och det finns en risk att det blir fel i annoteringsprocessen. Mer specifikt så kan felaktiga etiketter, så kallat etikettbrus, inkluderas i datan. Detta i sin tur kan skada modellens förmåga att ta korrekta beslut. Vi undersöker hur denna effekt tar sig form och finner att etikettbrus inte bara kan ha en negativ effekt på nämnda förmåga att ta korrekta beslut, utan även på förmågan att skatta den egna osäkerheten. Relaterat till annoteringskostnader, föreslår vi till sist ett tillvägagångssätt för att utnyttja brusiga etiketter i aktiv inlärning. Målet med aktiv inlärning, i denna kontext, är att identifiera och annotera de observationer som kommer att vara mest hjälpsamma i modellens inlärningprocess. Förslaget är att, i aktiv inlärning, ge möjligheten att samla in billigare, men brusiga, etiketter. På så sätt kan en begränsad annoteringsbudget räcka till att annotera fler datapunkter, vilket i sin tur kan leda till en bättre modell. Det senare är något som påvisas experimentellt.Funding agencies: This research was financially supported by the Wallenberg AI, Autonomous Systems and Software Program (WASP) funded by the Knut and Alice Wallenberg Foundation.</p

    Dispelling the Myths Behind First-author Citation Counts

    Get PDF
    We conducted a full-scale evaluative citation analysis study of scholars in the XML research field to explore just how different from each other author rankings resulting from different citation counting methods actually are, and to demonstrate the capability of emerging data and tools on the Web in supporting more realistic citation counting methods. Our results contest some common arguments for the continued use of first-author citation counts in the evaluation of scholars, such as high correlations between author rankings by first-author citation counts and other citation counting methods, and high costs of using more realistic citation counting methods that are not well-supported by the ISI databases. It is argued that increasingly available digital full text research papers make it possible for citation analysis studies to go beyond what the ISI databases have directly supported and to employ more sophisticated methods

    Author Index

    No full text
    Nao informado

    koamabayili/VECTRON-author-checklist: VECTRON author checklist

    No full text
    We have done our best to complete the author checklist relating to the use of animals in the hut study. Note that the objective for the hut study was to evaluate the IRS treatment applications for residual efficacy against Anopheles mosquitoes, including the local An. coluzzii mosquito population. Cows were only used to attract mosquitoes into the huts and no tests were carried out directly on the cows. The author checklist is intended for use with studies where experiments are carried out on animals, which is why we have had such difficulty in completing this for the hut study, as many of the questions do not relate to how the cows were used
    corecore