Gonzaga University Repository
Not a member yet
2718 research outputs found
Sort by
NOW 2016-12 December
Corpus of News on the Web data for December 2016.
The TAR folder contains linguistic data in three formats: Database: This is the format allows for the most robust searches and allows for powerful JOINs across corpus, lexicon, and source tables but requires knowledge of SQL. See Full-text corpus data for more information on how to use the database format. Linear Text: This format provides a textID for each text, and then the entire text on the same line. In this format, words are not annotated for part of speech or lemma. In addition, contracted words like can\u27t are separated into two parts (ca n\u27t) and punctuation is separated from words (eye level . As her). Word, Lemma, Part of Speech: Texts are separated by a line with ## and the textID
NOW 2016-09 September
Corpus of News on the Web data for September 2016.
The TAR folder contains linguistic data in three formats: Database: This is the format allows for the most robust searches and allows for powerful JOINs across corpus, lexicon, and source tables but requires knowledge of SQL. See Full-text corpus data for more information on how to use the database format. Linear Text: This format provides a textID for each text, and then the entire text on the same line. In this format, words are not annotated for part of speech or lemma. In addition, contracted words like can\u27t are separated into two parts (ca n\u27t) and punctuation is separated from words (eye level . As her). Word, Lemma, Part of Speech: Texts are separated by a line with ## and the textID
NOW 2016-11 November
Corpus of News on the Web data for November 2016.
The TAR folder contains linguistic data in three formats: Database: This is the format allows for the most robust searches and allows for powerful JOINs across corpus, lexicon, and source tables but requires knowledge of SQL. See Full-text corpus data for more information on how to use the database format. Linear Text: This format provides a textID for each text, and then the entire text on the same line. In this format, words are not annotated for part of speech or lemma. In addition, contracted words like can\u27t are separated into two parts (ca n\u27t) and punctuation is separated from words (eye level . As her). Word, Lemma, Part of Speech: Texts are separated by a line with ## and the textID
NOW 2016-08 August
Corpus of News on the Web data for August 2016.
The TAR folder contains linguistic data in three formats: Database: This is the format allows for the most robust searches and allows for powerful JOINs across corpus, lexicon, and source tables but requires knowledge of SQL. See Full-text corpus data for more information on how to use the database format. Linear Text: This format provides a textID for each text, and then the entire text on the same line. In this format, words are not annotated for part of speech or lemma. In addition, contracted words like can\u27t are separated into two parts (ca n\u27t) and punctuation is separated from words (eye level . As her). Word, Lemma, Part of Speech: Texts are separated by a line with ## and the textID
What to Do When Your Heritage is Hateful
In a National Public Radio piece aired on July 14, 2015, Gene Demby describe the “awkward mental gymnastics” involved in certain cultural preferences. A person’s musical tastes might run toward gangsta rap or outlaw country, for example, both of which can be misogynistic and reactionary, but the listener might paradoxically consider himself to be a supporter of women’s rights. Such inherent contradiction might seem hypocritical to some, but there is a certain elasticity of symbols, as cyphers meaning different things to different people. For some, the Confederate flag is a sign of racial Neanderthalism, the trademark of unreconstructed segregationists and rednecks. For others, the flag is a happy reminder of Tom Petty’s 1985 “Southern Accents” tour. Who is correct? Whose interpretation wins
The Implications of Doing Gender in the Digital Dating World
This study investigates digital dating and how men and women function within the phenomena in terms of gendered communication. The theory of semiotics, as described by CK Ogden and Roland Barthes, was used as a frame to analyze photographs, text, and actions among men and women particularly on the application Tinder to understand the construction of gender within a digital dating world. Original research was conducted using 15 participants, 11 female and 4 male, with 15-50 minute interviews, investigating their experiences with symbols and exploring gendered communication and symbols that perpetuate ideas about gender roles within dating in online settings. Results indicated that despite the ever-changing landscape of technology and the “on demand” culture users live in, the symbols and interpretation used on these websites propagated traditional gender roles. The results of this study should be the basis for further research to be done in more diverse online dating contexts
Grace and Human Action: Distinctively Christian Action in Schleiermacher\u27s Christian Ethics
Corpus del Español Sources
Corpus del Español source list. Zipped folder contains a single .txt file
Word, Lemma, and Part of Speech (Dominican Republic)
Corpus del Español word, lemma, and part of speech data format.
Zip folder contains 20 .txt files of linguistic data from the Dominican Republic split into two categories: General (g) and Blogs (b).
Texts are separated by a line with ## and the textID