DataverseNO
Not a member yet
2167 research outputs found
Sort by
A Database of Dead Sea Scroll Exhibitions in the 20th and 21st Centuries
The 20th and 21st centuries have seen more than 160 different exhibitions featuring the Dead Sea Scrolls. Since the first exhibition at the Library of Congress in Washington (DC) in 1949, these scrolls—stemming from the Essenes living northwest of the Dead Sea 2,000 years ago—have visited almost every continent on Earth. Most of the exhibitions have been blockbusters, drawing hundreds of thousands of visitors. Thus, the scroll exhibitions have had a significant religious and political impact.
This database of Dead Sea Scroll exhibitions features information about every one of them: Where they were held, when they were open, who the curators were, et cetera
Replication Data for: Zooming in on the semantics of French ingressives: a collostructional analysis
Dataset abstract:
The dataset includes an annotated corpus sample of N = 2000 French sentences with se mettre à or commencer à (1000 observations of each verb). The sample was drawn from the literary corpus Frantext and the journalistic corpus Le Monde (1000 observations from both corpora). The sample is balanced for verb as well as corpus, so we have 500 observations for each Verb-Corpus combination. The data is annotated for 3 variables: Source (corpus), Verb, collexeme.
Article abstract:
This paper examines the semantic value of the infinitive in the ingressive constructions se mettre à (SMA) and commencer à (COMA) using a distinctive collexeme analysis. We find that the collexemes significant for the construction SMA are fairly homogeneous across the different corpora and can be grouped into the general category of expressive collexemes. The collexemes significant for COMA are more heterogeneous and belong to the category of cognitive collexemes and to semantic fields of sensory and creative acts. The results are compatible with the hypothesis put forward by Verroens and De Cuypere (2023) stating that the overall meaning of the SMA construction is intrinsically punctual. The punctual value of SMA is not only compatible with expressive collexemes, but, moreover, emphasizes their unforeseen and unintentional meaning. Conversely, the incremental value of COMA is consistent with the gradual onset of cognitive and sensory collexemes.
Verroens, F., & De Cuypere, L. (2023). French ingressives and (phasal) aspect: A frame-semantic corpus-based analysis. Canadian Journal of Linguistics/Revue Canadienne de Linguistique, 68(3), 435-461. doi:10.1017/cnj.2023.1
Supporting Data for: Nickel Catalyzed Carbonylative Cross Coupling for Direct Access to Isotopically Labeled Alkyl Aryl Ketones
Introduction:
This Dataverse entry contains supporting data for our journal article “Nickel Catalyzed Carbonylative Cross Coupling for Direct Access to Isotopically Labeled Alkyl Aryl Ketones” submitted for review. The dataset contains additional experimental and computational work performed to support the chemical process, which was not included in the manuscript. The dataset consists of two files: 1) 'Additional experimental and computational results.pdf', and 2) 'optimized_geometries_Ni.xyz'. The first file explains the experimental and computational techniques employed and the analysis to support the chemical reaction. The second file contains the XYZ structures of the computationally optimized geometries in the work. Abstract from article:
Here we present an effective nickel-catalyzed carbonylative cross-coupling for direct access to alkyl aryl ketones from readily accessible redox-activated tetrachlorophthalimide esters and aryl boronic acids. The methodology, which is run employing only 2.5 equivalents of CO and simple Ni(II) salts as the metal source, exhibits a broad substrate scope under mild conditions. Furthermore, this carbonylation chemistry provides an easy switch between isotopologues for stable (13CO) and radioactive (14CO) isotope labeling, allowing its adaptation to the late-stage isotope labeling of pharmaceutically relevant compounds. Based on DFT calculations as well as experimental evidence, a catalytic cycle is proposed involving a carbon-centered radical formed via nickel(I)-induced outer-sphere decarboxylative fragmentation of the redox-active ester
Replication Data for: "The category of throw verbs as productive source of the Spanish inchoative construction."
The dataset contains the quantitative data used to create the tables and graphics in the article "The category of throw verbs as productive source of the Spanish inchoative construction."
The data from the 21th century originates from the Spanish Web Corpus (esTenTen18), accessed via Sketch Engine. Only the subcorpus for European Spanish Data was selected. After downloading, the samples were manually cleaned. In the dataset, maximally 500 tokens were retained per auxiliary. For the earlier centuries, the data was extracted from the Corpus Diacrónico del Español (Corde). See Spanish_ThrowVerbs_Inchoatives_queries_20230413.txt for the specific corpus queries that were used.
The data were annotated for the infinitive observed after the preposition 'a' and for the semantic class to which this infinitive belongs, following the existing ADESSE classification (see below), besides other criteria that are not taken into account for this study. Concretely, the variables 'Century', 'INF' (infinitive) and 'Class' were used as input for the analysis (see data-specific sections below for more information about the variables).
The empirical analysis is based on the downloaded data from the Spanish Web corpus (esTenTen18) (Kilgariff & Renau 2013). The Spanish Web corpus contains 20.3 billion words, from which 3.5 billion belong to the European Spanish domain. This corpus contains internet data, with observations originating from fora, blogs, Wikipedia, etc. Only the subcorpus with European Spanish data was consulted. The search syntax that was used to detect the inchoative construction was the following: “[lemma="echar"] [tag="R.*"]{0,3}"a"[tag="V.*"] within ” (consult Spanish_ThrowVerbs_Inchoatives_queries_20230413.txt for all corpus queries). After downloading, all the observations were manually cleaned. In total, the dataset contains, after the removal of false positives, 5514 tokens with a maximum of 500 tokens per auxiliary. False positive tokens were, for example, tagging errors wrongly coding nouns, such as Superman, Pokémon, Irán, among others, as infinitives, and also observations in which the auxiliary in combination with the infinitive did not express the inchoative value but its orginal semantic meaning, such as "saltar a nadar", for example, which means “to jump to swim” and not “to start to swim”. Of the auxiliaries with less than 500 relevant tokens in the esTenTen corpus, all tokens in the dataset were retained; for the auxiliaries with more than 500 tokens in the esTenTen corpus, only the first 500 were selected.
For this specific study on the throw verbs, only the following auxilaries were retained: arrojar, disparar, echar, lanzar and tirar.
For the diachronic data, the Corpus Diacrónico del Español (CORDE) was consulted. See Spanish_ThrowVerbs_Inchoatives_queries_20230413.txt for the specific queries that were used to retrieve the data in CORDE.</p
Carl Paul Caspari, Publications
The dataset contains bibliographic information about articles and books about the theologian Carl Paul Caspari (1814-1892). The language of the content is Norwegian, Danish, and German. This collection of references was collected by researcher Oskar Skarsaune and used when making an uncompleted biography about Caspari with the title: En Lærd af Guds Naade»: Carl Paul Caspari 1814–1892: En påbegynt biografi som bidrag til norsk teologi- og lærdomshistorie. The article and books date from 1835 to 1979. MF has archived physical copies of the content
System identification campaign - Skywalker X8 UAV
The data was collected by the NTNU UAV lab as part of the system identification campaign for the Skywalker X8 UAV in May of 2023 at Agdenes Airfield, Norway. The dataset includes 17 maneuvers, where each maneuver is approximately 10 seconds long. The maneuvers are split into a training and a validation set. There was a strong north-west wind during the experiments with significant vertical component
Background data for: Pre-hospital identification of infection focus in sepsis and timely empirical antibiotic therapy in a rural ambulance service: A prospective cohort study
This dataset is extracted from an ambulance quality registry of patients with suspected sepsis managed by the ambulance department of the University hospital of Northern Norway. Data was collected from patients with suspected sepsis who were given pre-hospital intravenous antibiotics and transported to hospital by the University of Northern Norways ambulance service from May 2018 to August 2022. The dataset was extracted to conduct a study on whether paramedics with or without assistance of general practitioners are able to identify the infection focus in sepsis patients and administer timely intravenous antibiotic treatment. The dataset contains demographic and clinical data, patient trajectories, treatment given in the prehospital environment, patient status at hospital arrival and at discharge, and 30-day all-cause mortality.Abstract from corresponding article:
Background:
Early diagnosis and initiation of intravenous antibiotic therapy in patients with sepsis reduce both morbidity and mortality, thus management of sepsis in the pre-hospital setting is likely to affect patient outcomes. A clear description of pre-hospital sepsis management with emphasis on trajectory and identification of source of infection may contribute to timely and more targeted pre-hospital antibiotic therapy. The aim of this study was to investigate whether paramedics with or without assistance of general practitioners are able to identify the infection focus in sepsis patients and administer timely intravenous antibiotic treatment.
Methods:
We conducted a cohort study of patients with suspected sepsis who were given pre-hospital intravenous antibiotics and transported to hospital. The setting was mainly rural with long average distance to hospital. Patients received targeted antibiotic treatment after assessment based on clinical work-up supported by scoring systems. Patients were prospectively included from May 2018 to August 2022. Data were registered in a sepsis management ambulance quality registry. Results are presented as median or absolute values. Chi-square tests were used to compare categorised data of source of infection and presence of general practitioners.
Results:
The study group consisted of 328 patients. Median age was 76 years (IQR 64, 83) and 30-days all-cause mortality was 10.4 %. Antibiotic treatment was initiated at a median of 44 minutes after arrival of ambulance, and median transportation time from place of incident to hospital was 69 minutes. In cases where a suspected source of infection was determined, hospital discharge papers confirmed the pre-hospital diagnosis of infection focus in 195 cases (79.3 %). The presence of a general practitioner during the pre-hospital assessment increased the rate of correctly identified source of infection from 72.6% to 86.1 % (p=0.009). Concordance between pre-hospital identification of a tentative focus and discharge diagnosis was highest for lower respiratory tract (p=0.02) and urinary tract infections (p=0.03).
Conclusions:
Ambulance personnel are able to identify focus of infection, and start intravenous antibiotics quickly. This is probably of particular value in areas with long transportation times. Collaboration with primary care physicians increases level of diagnostic accuracy.</p
Synthetic version of anonymized Norway Registry data containing prescriptions and hospitalization of the patients
This dataset represents synthetic data derived from anonymized Norwegian Registry Data of pa aged 65 and above from 2011 to 2013. It includes the Norwegian Patient Registry (NPR), which contains hospitalization details, and the Norwegian Prescription Database (NorPD), which contains prescription details. The NPR and NorPD datasets are combined into a single CSV file. This real dataset was part of a project to study medication use in the elderly and its association with hospitalization. The project has ethical approval from the Regional Committees for Medical and Health Research Ethics in Norway (REK-Nord number: 2014/2182). The dataset was anonymized to ensure that the synthetic version could not reasonably be identical to any real-life individuals. The anonymization process was done as follows: first, only relevant information was kept from the original data set. Second, individuals' birth year and gender were replaced with randomly generated values within a plausible range of values. And last, all dates were replaced with randomly generated dates. This dataset was sufficiently scrambled to generate a synthetic dataset and was only used for the current study. The dataset has details related to Patient, Prescriber, Hospitalization, Diagnosis, Location, Medications, Prescriptions, and Prescriptions dispatched. A publication using this data to create a machine learning model for predicting hospitalization risk is under review
UNN-LC High-Resolution Histopathological Lung Tissue Patch Dataset
The UNN-LC High-Resolution Histopathological Lung Tissue Patch Dataset is a collection of image patches designed for computational prognostic evaluation of lung cancer. Compiled from a subset of 194 whole-slide images (WSIs) from the University Hospital of North Norway, this dataset provides a comprehensive representation of various lung tissue conditions. Each 768 x 768 pixel patch contributes to a detailed analysis of tissue morphology.
The dataset was annotated by an oncologist (Thomas Kilvær) and a pathologist (Stig Dalen) with a concerted effort to minimize selection and labeling biases. Specifically, patches with predominantly cancer cells, including tumor-infiltrating lymphocytes, were annotated by Stig Dalen. Thomas Kilvær provided annotations for patches representing normal lung tissue. The combined efforts of Stig Dalen and Thomas Kilvær resulted in the annotations for the reactive stroma with tertiary lymphoid structures and necrosis areas data. Annotations were acquired using QuPath software and a custom-developed annotation tool.
The dataset categorizes patches into four classes: necrosis, tumor, stroma_tls, and normal_lung. The necrosis class includes patches of tissue associated with tumor regions, while the normal lung class represents areas of healthy lung tissue, inclusive of stromal components. The stroma_tls class is characterized by patches of reactive stroma with dense tissue and lymphocyte aggregates. The tumor tissue class comprises patches with a predominant presence of tumor content and may also include areas with tumor-infiltrating lymphocytes (TILs).
For those interested in further expanding the scope and improving the balance of classes within the dataset, additional patches from the LC25000 dataset can be integrated for a more diverse representation of tissue conditions. This approach can enhance the robustness of computational models developed using this data.
The dataset is divided into training and testing sets to facilitate and promote reproducibility in the development and validation of vision models. The training set includes a selection of patches from each class, while the testing set is composed of the remaining patches to ensure a comprehensive assessment of model performance.</p
Replication Data for: Performance investigation of refrigeration systems for freezing tunnels
This dataset contains temperature measurements from one of the batch blast freezers at Pelagia Kalvåg, during freezing of atlantic herring