National Institute for Japanese Language and Linguistics
Academic Repository of the National Institute for Japanese Language and Linguistics / 国立国語研究所学術情報リポジトリNot a member yet
3427 research outputs found
Sort by
「現代日本語書き言葉均衡コーパス」の漢字と表記
application/pdf国立国語研究所会議名: 国立国語研究所オープンハウス2022, 開催地: オンライン, 会期: 2022年9月9日conference objec
Application of Transkribus for a Digitization of the Japanese Christian Contemptus Mundi
application/pdf日本学術振興会 / ルール大学ボーフム国立国語研究所JSPS / Ruhr University BochumNational Institute for Japanese Language and Linguistics2017年にドイツのヘルツォーク・アウグスト図書館(HAB)で発見された新出キリシタン資料ローマ字本日本語訳『コンテムツス・ムンヂ』の翻刻プロジェクトで用いられた技術と公開方法について論じる。本論文著者両名による日独共同研究プロジェクトにおいて、ルール大学ボーフムのオースタカンプ・スヱン教授の指導のもと、本文献のデジタル化を目標としたデジタル技術活用に関する議論と実践がなされた。そこでは、機械学習による自動翻刻ソフトウェアTranskribusを用いて、そのHTR(Handwritten Text Recognition)モデルに学習させ、自動および手動修正で文献のレイアウト・補助記号などを忠実に再現したデジタル翻刻が行われた。本プロジェクトは、このデジタル翻刻に基づいて、本キリシタン版をデジタルアーカイブ化し、可能な研究対象として公開し、国際的な研究に活用できるようにするモデルを提示する。journal articl
日本の消滅危機言語・方言のためのデジタルアーカイブ開発
application/pdf国立国語研究所会議名: 国立国語研究所オープンハウス2022, 開催地: オンライン, 会期: 2022年9月9日conference objec
Relationship Between Part-of-Speech-Based Stylistic Indicators and Subjective Impressions: Psychological Study of MVR and Part-of-Speech Composition
大正大学法政大学大学院ライフスキル教育研究所日本大学法政大学国立国語研究所Taisho UniversityLife Skill Education Institute in the Graduate School of Hosei UniversityNihon UniversityHosei UniversityNational Institute for Japanese Language and Linguisticsjournal articl
危機言語としての地域のことば
国立国語研究所本ファイルは『ユリイカ2022年8月号 特集=現代語の世界』に掲載された「危機言語としての地域のことば」の校正原稿です。(2024.1)articl
Accent Data of Adjectives in the Tanohata Dialect, Iwate Prefecture : Part 2
東京大学名誉教授Emeritus Professor, The University of Tokyo岩手県田野畑村方言の形容詞につき,5モーラ語から9モーラ語までの資料を提示し,分析をする。前稿の2~4モーラ語と合わせた形容詞のアクセント体系は次の特徴を持つ。長さを問わず,次末核型は常にある。それに対して,無核型は少数派で3~7モーラ語にしかなく,しかも5モーラ語以上では語構造に偏りがある。この2つの基本型に加えて,3~7モーラ語には語頭核型もあり,5モーラ以上の例は強いマイナス評価の意味と連動しているという特徴を持つ。This paper presents and analyzes accent data of adjectives in the Tanohata dialect, Iwate Prefecture. The accent system of adjectives with two to nine morae has the following characteristics. There is always a penultimate kernel pattern regardless of length. However, the kernelless (i.e., unaccented) pattern constitutes a minority in that it appears only in words with three to seven morae, and in five-mora or longer words there is also a restriction in word structure. Aside from these two basic patterns, there is also an initial kernel pattern in adjectives with three to seven morae that is linked to strongly negative connotations in five-mora or longer words.application/pdfdepartmental bulletin pape
Reading Time and Vocabulary Rating in the Japanese Language : Large-Scale Reading Time Data Collection Using Crowdsourcing
application/pdfNational Institute for Japanese Language and Linguistics / Tokyo University of Foreign StudiesThis study examined the effect of the differences in human vocabulary on reading time. This study conducted a word familiarity survey and applied a generalised linear mixed model to the participant ratings, assuming vocabulary to be a random effect of the participants. Following this, the participants took part in a self-paced reading task, and their reading times were recorded. The results clarified the effect of vocabulary differences on reading time.conference pape