DataverseNO
Not a member yet
2167 research outputs found
Sort by
Replication Data for: Shortening mechanisms in Construction Morphology: The Russian spec-N construction
This database contains Russian nouns beginning with the string spec from the Russian National Corpus (www.ruscorpora.ru), such as specoperacija ‘special operation’. The database was created in the following way. First, we searched for all lemmas beginning with spec in the part of the main corpus containing non-fiction (nexudozhestvennaja literatura) and created a list of all the head words (e.g., operacija). Second, we identified the number of attestations for each word and included this information in the list. Third, we searched for the full adjectives special’nyj and specializirovannyj immediately followed by the head words (e.g., special’naja operacija and specializirovannaja operacija). The number of attestations for each combination of adjective and noun was included in the list. The corpus searches were carried out in 2022. In the database, the words are given in Cyrillic.Abstract from article: This study presents an in-depth analysis of Russian stub compounds in spec ‘special’ and their competition with the corresponding full adjective special’nyj ‘special’ followed by a noun. Couched in Construction Morphology the corpus-based analysis addresses four understudied areas in theoretical and Russian morphology: shortening mechanisms, competition between morphological words and multiword expressions, blocking, and compounding in Russian. It is argued that shortening mechanisms create words that are more than stylistic variants of the corresponding longer constructions, although full synonymy may occur under specific conditions. The diachronic and synchronic motivation of the shortening mechanism under scrutiny is analyzed in terms of economy, extravagance and expressiveness. Blocking is demonstrated to be statistical (involving tendencies rather than categorical rules) and bidirectional, whereby a morphological construction may be favored over a syntactic construction and vice versa. The proposed analysis adds to the knowledge of stub compounds in Russian and demonstrates how a wide variety of generalizations can be adequately accounted for in Construction Morphology
GNSS Total Electron Content Data (60 s) at Hopen in 2023
This data set contains Total Electron Content data at 60 seconds time resolution at Hopen, Svalbard.
The measurements were collected by the University of Bergen using a NovAtel GPStation-6 global navigation satellite system receiver. The measurements include signals from GPS, GLONASS, and GALILEO at different frequencies. These data are used for research on space weather disturbances in the polar ionosphere.
A detailed description of the data structure and format is gathered in the documentation data set: Oksavik, Kjellmar, 2020, "Documentation of GNSS Total Electron Content and Scintillation Data (60 s) at Svalbard", DataverseNO, https://doi.org/10.18710/EA5BYX
This data set is part of a larger collection: Oksavik, Kjellmar, 2020. "The University of Bergen Global Navigation Satellite System Data Collection". DataverseNO. https://doi.org/10.18710/AJ4S-X394.
</p
Search strategies for a review on the state of knowledge of vaginal breech birth outcomes related to birth attendants profession
DATASET MIGRATED FROM FIGSHARE: Search strategies for a review article on state of knowledge of vaginal breech birth outcomes related to birth attendants profession are described. Contains strategies for the databases Maternity and Infant care, Medline, Cinahl and Cochrane Library. See the readme.txt for further documentation.</p
Search strategies for the systematic review "What is the current knowledge-base on sustainability within advanced practice nursing?"
DATASET MIGRATED FROM FIGSHARE: This dataset consists of the search documentation for a review on sustainability within advanced practice nursing. It contains strategies for the databases Cinahl, Medline, PsycInfo, ERIC, Web of Science and Scopus. See the readme.txt for further documentation.</p
Search strategies for a scoping review article on virtual reality simulation as a method of learning, for nursing students in a mental health care context
DATASET MIGRATED FROM FIGSHARE: Search strategies for a scoping review article on virtual reality simulation as a method of learning, for nursing students in a mental health care context are described.</p
Background data for: Negation as a predictor of clausal complement choice in World Englishes
This dataset contains tabular files recording occurrences of the verb REGRET complemented by a finite or non-finite complement clause (CC) in the GloWbE corpus. Tokens were retrieved using the online interface (https://www.english-corpora.org/glowbe/) and manually annotated for several syntactic and semantic variables (variety, L1 vs L2, CC_pattern, finiteness, negative marker, negation_yes/no, and temporal relation). See ReadMe file for more details. Related publication: Romasanta, Raquel P. 2021. Negation as a predictor of clausal complement choice in World Englishes. English Language and Linguistics 26(2): 307-32
Do DOM quality and origin affect the upake and accumulation of a lipid soluble contaminant in a filter feeding ascidian species (Ciona) that can target small particle size classes?
This is an experimental dataset within the field ecotoxicology. The target of the experiment was to find out whether terrestrial derived dissolved organic matter is a better contaminant vector for lipophilic contaminant to filter feeders than other types of dissolved organic matter or no dissolved organic matter. The contaminant was Teflubenzuron (in acetone), the filter feeder was Ciona intestinalis. For a detailed description of methods, see paper (submitted). For exact description of the variables measured, see S1_supplementary_Info_Data
Background data for: Advancing our understanding of dispersion measures in corpus research
Dataset description
This dataset contains background data and supplementary material for Sönning (forthcoming), a study that looks at the behavior of dispersion measures when applied to text-level frequency data. For the literature survey reported in that study, which examines how dispersion measures are used in corpus-based work, it includes tabular files listing the 730 research articles that were examined as well as annotations for those studies that measured dispersion in the corpus-linguistic (and lexicographic) sense. As for the corpus data that were used to train the statistical model parameters underlying the simulation study reported in that paper, the dataset contains a term-document matrix for the 49,604 unique word forms (after conversion to lower-case) that occur in the Brown Corpus. Further, R scripts are included that document in detail how the Brown Corpus XML files, which are available from the Natural Language Toolkit (Bird et al. 2009; https://www.nltk.org/), were processed to produce this data arrangement.Abstract: Related publication
This paper offers a survey of recent corpus-based work, which shows that dispersion is typically measured across the text files in a corpus. Systematic insights into the behavior of measures in such distributional settings are currently lacking, however. After a thorough discussion of six prominent indices, we investigate their behavior on relevant frequency distributions, which are designed to mimic actual corpus data. Our evaluation considers different distributional settings, i.e. various combinations of frequency and dispersion values. The primary focus is on the response of measures to relatively high and low sub-frequencies, i.e. texts in which the item or structure of interest is over- or underrepresented (if not absent). We develop a simple method for constructing sensitivity profiles, which allow us to draw instructive comparisons among measures. We observe that these profiles vary considerably across distributional settings. While D and DP appear to show the most balanced response contours, our findings suggest that much work remains to be done to understand the performance of measures on items with normalized frequencies below 100 per million words.</p
EcoSens: Enhanced Ocean Colour Remote Sensing for Optically Complex Waters
This dataset contains optical (absorption, scattering and radiance) as well as environmental data (temperature, salinity, concentrations of suspended particulate matter (SPM) and Chlorophyll) collected in optically complex Norwegian fjords (Hardangerfjord, Gaupnefjord and Lurefjord) as part of the EcoSens project.
Hardangerfjord was sampled during a coccolithophore bloom in 2022, and again during the following year without a visible bloom. Gaupnefjord was visited during a period of high glacial meltwater influx. Lurefjord is a fjord with high concentrations of colored dissolved organic matter (CDOM).
This dataset is divided into separate folders containing metadata, CDOM absorption, particulate scattering and absorption, phase functions, radiance, and irradiance spectra, and should be viewed in tree folder structure
Time Series of Oceanographic Data offshore Prins Karls Forland
This time series show oceanographic data collected from an seafloor ocean observatory monitoring a seabed methane seep site offshore West Spitsbergen at around 90m depth. Two datasets are included here: a first deployment from July 2015 to May 2016 and a second from October 2016 to July 2017. Here only temperature, salinity and pressure are given, as well as calculated conservative temperature and absolute salinity. The salinity during the first deployment drifted and a correction is proposed. Temperature, salinity and pressure from the first dataset has been previously published in Dølven et al. 2022 (https://doi.org/10.18710/CEIA1U