DataverseNO
Not a member yet
2167 research outputs found
Sort by
Replication data for: Salience-simplification strategy for markedness of causal subordinators: “because” and “since” in argumentative essays
The dataset supports the research article "Salience-simplification strategy to markedness of causal subordinators: The case of “because” and “since” in argumentative essays". In total, the dataset marks features of 976 causal adverbial subordinations retrieved from student argumentative essays.Data points were extracted from three corpora. Specifically, all essays in NESSIE (Native English Speakers’ Similarly or Identically-prompted Essays, created by Xu Jiajin, 781 essays; 291,911 tokens) and argumentative essays in LOCNESS (the Louvain Corpus of Native English Essays, created by Granger, 323 essays; 230,138 tokens) were selected. Native argumentative essays from BAWE’s (British Academic Written English, created by Hilary Nesi) Arts and Humanities disciplinary group were chosen (512 essays; 1,360,932 tokens). In total, 1,616 essays comprising 1,882,981 tokens were examined.
The dataset comprises 976 datapoints of causal subordinations conjoined by "because" and "since" in students' argumentative essays--488 data points of all "since" subordinations, and 488 randomly selected "because" subordinations. On these data points, ten contextual features that are potential predictors of people's choices between causal subordinators "because" and "since" were annotated.
The ten contextual features annotated are "position", "separation", "embeddedness", "initial adverbials", "sub-clause", "de-ranking", "clause-length ratio", "hedging terms", "clausal relationship", and "bridging".
Overall fourteen variables including ten contetual features are annotated:
(1) "No." is the ID of each data point(this is one ID marker);
(2) "subordinator" marks the logical subordinators (this categorical variable has two values: "because" and "since");
(3) "position" marks the logical adverbial clause positions compared with the main clause (this categorical variable has two values: "preposed" or "postposed");
(4) "sep" indicates whether a separating punctuation mark exists between the subordinate and main clauses(this categorical variable has two values: "YES" or "NO");
(5) "embeddedness" indicates whether a complex sentence is embedded in a larger comlex sentence(this categorical variable has two values: "YES" or "NO");
(6) "ini.adv" denotes whether an initial adverbial exists in the causal subordination(this categorical variable has two values: "YES" or "NO");
(7) "sub-clau" indicates whether the causal subordinate contains sub-clauses of any type(this categorical variable has two values: "YES" or "NO");
(8) "deranking" indicates whether the predicate of the subordinate clause is complete(this categorical variable has two values: "YES" or "NO");
(9) "sub.main.ratio" is the length ratio of the subordinate and main clauses in terms of word count (this numerical variable is converted into ln value for better interpretation);
(10) "hedging" indicates whether a hedging term exists in the subordinate clause(this categorical variable has two values: "YES" or "NO");
(11) "clau.rel" denotes the interclausal relationships on the general level(this categorical variable has two values: "direct" or "indirect");
(12) "spc.clau.rel2" denotes the interclausal relationships on the secondary level(this categorical variable has five values: "im", "rm", "asst", "inpr", and "sugg");
(13) "bridging" indicates whether the subordinate clause contains any information referring back to the preceding clause(this categorical variable has two values: "YES" or "NO");
(14) "source" shows specific corpora the data points come from (this categorical variable has three values: "NESSIE", "LOCNESS", or "BAWE") ;
This dataset was constructed to explore contextual features that discriminate between causal subordinators of "because" and "since" and to rank the effective features.</p
Genome-wide meta-analysis of iron status biomarkers and the effect of iron on all-cause mortality in HUNT
This dataset contains GWAS meta-analysis summary statistics for serum iron, serum ferritin, transferrin saturation and total iron binding capacity. Iron is essential for many biological processes, but iron levels must be tightly regulated to avoid harmful effects of both iron deficiency and overload. Here, we perform genome-wide association studies on four iron related biomarkers (serum iron, serum ferritin, transferrin saturation, total iron binding capacity) in the Trøndelag Health Study (HUNT), the Michigan Genomics Initiative (MGI) and the SardiNIA study, followed by their meta-analysis with publicly available summary statistics, analyzing up to 257,953 individuals
Replication Data for: A Three-Year Mixed Methods Study of Undergraduates’ Information Literacy Development: Knowing, Doing, and Feeling
This data set contains the replication data and supplements for the article "Knowing, Doing, and Feeling: A three-year, mixed-methods study of undergraduates’ information literacy development." The survey data is from two samples:
- cross-sectional sample (different students at the same point in time)
- longitudinal sample (the same students and different points in time)Surveys were distributed via Qualtrics during the students' first and sixth semesters. Quantitative and qualitative data were collected and used to describe students' IL development over 3 years. Statistics from the quantitative data were analyzed in SPSS. The qualitative data was coded and analyzed thematically in NVivo.
The qualitative, textual data is from semi-structured interviews with sixth-semester students in psychology at UiT, both focus groups and individual interviews. All data were collected as part of the contact author's PhD research on information literacy (IL) at UiT. The following files are included in this data set: 1. A README file which explains the quantitative data files. (2 file formats: .txt, .pdf)2. The consent form for participants (in Norwegian). (2 file formats: .txt, .pdf)3. Six data files with survey results from UiT psychology undergraduate students for the cross-sectional (n=209) and longitudinal (n=56) samples, in 3 formats (.dat, .csv, .sav). The data was collected in Qualtrics from fall 2019 to fall 2022. 4. Interview guide for 3 focus group interviews. File format: .txt5. Interview guides for 7 individual interviews - first round (n=4) and second round (n=3). File format: .txt 6. The 21-item IL test (Tromsø Information Literacy Test = TILT), in English and Norwegian. TILT is used for assessing students' knowledge of three aspects of IL: evaluating sources, using sources, and seeking information. The test is multiple choice, with four alternative answers for each item. This test is a "KNOW-measure," intended to measure what students know about information literacy. (2 file formats: .txt, .pdf)7. Survey questions related to interest - specifically students' interest in being or becoming information literate - in 3 parts (all in English and Norwegian): a) information and questions about the 4 phases of interest; b) interest questionnaire with 26 items in 7 subscales (Tromsø Interest Questionnaire - TRIQ); c) Survey questions about IL and interest, need, and intent. (2 file formats: .txt, .pdf)8. Information about the assignment-based measures used to measure what students do in practice when evaluating and using sources. Students were evaluated with these measures in their first and sixth semesters. (2 file formats: .txt, .pdf)9. The Norwegain Centre for Research Data's (NSD) 2019 assessment of the notification form for personal data for the PhD research project. In Norwegian. (Format: .pdf
Supporting Data for: Cavity-free continuum solvation: implementation and parametrization in a multiwavelet framework
Supplementary material to an article submitted for review, about the PCM implementation in the MRChem software (https://github.com/MRChemSoft/mrchem), entitled "Cavity-free continuum solvation: implementation and parametrization in a multiwavelet framework" (2022-11-03).
We present a multiwavelet-based implementation of a quantum/classical polariz-
able continuum model. The solvent model uses a diffuse solute-solvent boundary and
a position-dependent permittivity, lifting the sharp-boundary assumption underlying
many existing continuum solvation models. We are able to include both surface and
volume polarization effects in the quantum/classical coupling, with guaranteed pre-
cision, due to the adaptive refinement strategies of our multiwavelet implementation.
The model can account for complex solvent environments and does not need a pos-
teriori corrections for volume polarization effects. We validate our results against a
sharp-boundary continuum model and find very good correlation of the polarization
energies computed for the Minnesota solvation database
Replication Data for: Er politiet sikker eller sikre? Adjektivkongruens ved kollektiver i norsk
The “Adjektivkongruens” project investigates the incidence of singular vs. plural endings on adjectives that agree with collective nouns in Norwegian, as in: Politiet var sikker (singular) vs. sikre (plural) på at mannen var død ‘The police was/were sure that the man was dead’. The possible predictor factors that are investigated are the distance between the noun and the adjective, the semantic category of the substantive (geopolitical, political party, police, etc.), the status of the adjective as an ordinary qualitative adjective or a participle, and whether or not the adjective is part of a larger grammatical construction (for example with a prepositional phrase)
Replication data for "Tectonic evolution of the Indio Hills segment of the San Andreas fault in southern California"
Field photographs and structural measurements from fieldwork along the San Andreas Fault Zone in the Indio Hills, southern California, in February-March 2017. The text files containing the structural data include measurements of strike in the first column and dip in the second column.
Transpressional uplift domains of inverted Miocene–Pliocene basin fill along the San Andreas fault zone in Coachella Valley, southern California, are characterized by fault linkage and segmentation and deformation partitioning. The Indio Hills wedge-shaped uplift block is located in between two boundary fault strands, the Indio Hills fault to the northeast and the Banning fault to the southwest, which merge to the southeast. Uplift commenced about 2.2–0.76 million years ago and involved either three separate and/or progressive fold and faulting stages caused by a change from distributed strain, via partly partitioned to fully partitioned right-slip and reverse displacement on the bounding faults when approaching the fault junction, or single-stage partly partitioned transpression. Major fold structures in the study area include oblique, right-stepping, partly overturned en echelon macro-folds that tighten and bend into parallelism with the Indio Hills fault to the east and become more open towards the Banning fault to the west, thus indicating a close relationship of the macro-folds with the Indio Hills fault and a late initiation of the Banning fault. Sets of strike-slip to reverse step-over and right- and left-lateral cross faults and conjugate kink bands affect the entire uplifted area, and locally offset the en echelon macro-folds. Comparison with the Mecca Hills and Durmid Hills uplifts farther southeast in Coachella Valley reveals notable similarities, but also differences in fault architectures, spatial and temporal evolution, and deformation mechanisms, indicating that the Indio Hills uplift block is at an early stage of evolution of a ladder structure, in contrast to the more evolved Durmid Hills.</p
Replication Data for: Conformational tuning improves the stability of spirocyclic nitroxides with long paramagnetic relaxation times
Nitroxides are widely used as probes and polarization transfer agents in spectroscopy and imaging.These applications require high stability towards reducing biological environments, as well as beneficial relaxation properties. Closed spirocyclohexyl nitroxides exhibit a dramatically improved stability towards reduction by ascorbate, while maintaining long relaxation times in EPR spectroscopy. This dataset contains raw files and analysis reports of 1HNMR, 13CNMR, HRMS, FTIR of all compounds synthesized for and used in work: Conformational tuning improves the stability of spirocyclic nitroxides with long paramagnetic relaxation times. Moreover, for nitroxides raw files of X-band CW-EPR spectra, kinetic runs based on CW-EPR spectra and Q-band EPR relaxation measurements are included. Dataset is completed with X-ray crystallography CIF files and reports, as well as all optimized geometries for discussed conformer
Replication Data for 'Subject pro-drop and past-tense auxiliary clitics in South-Western Ukrainian'
These are the replication data for part of a journal article on 'Subject pro-drop and past-tense auxiliary clitics in South-Western Ukrainian'. The abstract of the article is below. The data consist of annotated strings of continuous text in Standard Ukrainian (=ProDrop_SU.txt) and in South-Western Ukrainian Dialect (=ProDrop_SWU.txt), with 460 parsed predications each for Standard Ukrainian and South-Western Ukrainian dialect. The sources and the annotation are detailed in the accompanying description of the data (=00_readme_file_for_ProDrop.txt). The aim is to identify the predications that allow for subject pro-drop, to correlate its occurrence with the main morpho-syntactic features of the predicate (tense, person, number), and to compare its frequency in South-Western Ukrainian Dialect versus Standard Ukrainian.Here is the abstract of the article: South-Western Ukrainian dialects have retained the option of auxiliary clitics in the formation of the past tense. At the same time, they have past-tense forms without auxiliary clitics as in Northern Ukrainian dialects, and in Standard Ukrainian based on South-Eastern dialects. A sample corpus study suggests that South-Western Ukrainian also shows a higher frequency of subject pro-drop than Standard Ukrainian. The South-Western Ukrainian pattern presents the precise mirror image of the same two features in South-Eastern ‘Borderland’ Polish. Here, the dialect adopted the option of past-tense forms without auxiliary clitics, next to those with them as in Standard Polish. At the same time, it shows a higher frequency of non-pro-drop than Standard Polish. I argue that these matching facts are the result of long-standing language contact that worked simultaneously in two directions: the increase in the use of an existing dialectal Ukrainian pattern under Polish influence, as well as the increase in the use of an existing dialectal Polish pattern under Ukrainian influence. As a result, both dialects show the same variation between past-tense forms with auxiliary clitics and without them, and they have mutually converging tendencies in subject pro-drop – the Ukrainian dialect adapting towards Polish pro-drop, and the Polish dialect towards Ukrainian non-pro-drop. The bi-directionality of influence in the SWU dialectal areal goes beyond ‘classical’, i.e. unidirectional language-contact scenarios
Replication Data for: The effect of bird droppings on the corrosion of steels and aluminium used in offshore applications
Bird dropping (bird faeces) in contact with metals may affect the corrosion of these metals when exposed to a certain environment. This dataset contains the experimental characterisation of the effect chicken excrements have on a carbon steel, AISI316 stainless steel, duplex stainless steel, super duplex stainless steel, a carbon steel coated with thermally sprayed aluminium (TSA) and aluminium EN AW6082. The following experimental results are included in this dataset: (i) results of electrochemical experiments (open circuit potential measurements, polarisation resistance measurements, cathodic and anodic polarisation curves in the presence and absence of bird droppings) in an electrolyte close to synthetic seawater but without some minor components, at pH 8.9; (ii) photos of samples before and after different exposures in a salt spray chamber; (iii) images of samples after exposure to a salt spray test for given times up to 692 h (29 days); (iv) Raman spectra recorded in a confocal microscope after exposure to a salt spray test; and (v) infrared (IR) spectra recorded after exposure to a salt spray test. In addition, (vi) the study also contains mass loss (weight loss) data, which is in full included in the supporting information of the accompanying article, because of its simplicity, and is not included here. The dataset contains ASCII files and images in JPEG and PNG formats with descriptive filenames and is structured in subfolders indicating the respective methods. A detailed interpretation of the results, and the experimental details associated with the dataset is available in the associated article
Replication Data for: French ingressives and (phasal) aspect. A frame-semantic corpus-based analysis
Dataset abstract
The dataset includes an annotated corpus sample of N = 2000 French sentences with se mettre à or commencer à (1000 observations of each verb). The sample was drawn from the literary corpus Frantext (FT) and the journalistic corpus Le Monde (1000 observations from both corpora). The sample is balanced for verb as well as corpus, so we have 500 observations for each Verb-Corpus combination. The data is annotated for 8 variables: Source (corpus), Verb, Mood & Tense, Event type, Adverb presence, Adverb token, and Adverb type.
Article abstract
This article compares the usage of commencer à ‘to begin’+Vinf. and se mettre à ‘to start’ + Vinf. in modern French. Using a corpus sample of 2000 observations, we examined the effect of Adverbial complementation, Event type (aspect), Tense. Based on a mixed-effects logistic regression analysis, we found evidence for Event type – se mettre à is associated with activities – and Tense – se mettre à seems to be associated with Passé Simple, Futur proche and Subjonctif présent, whereas commencer à with Plus-que-parfait and Indicatif Imparfait. We discuss the results in the frame-semantic model of Croft (2012). We make the case that commencer à can have the profile of an achievement or that of an accomplishment while se mettre à manifests only one profile, i.e. that of an achievement. Our results support a one-component approach to aspect in which the result of the interaction between grammatical aspect and lexical aspect can be attributed to the same aspectual contour.
References
Croft, William (2012). Verbs : aspect and causal structure. Oxford: Oxford University Press.</p