National Institute for Japanese Language and Linguistics

Academic Repository of the National Institute for Japanese Language and Linguistics / 国立国語研究所学術情報リポジトリ
Not a member yet
    3427 research outputs found

    A Quantitative Study Using Corpora on the Use of Contractions of Conditional Forms

    Get PDF
    同志社大学大学院 博士後期課程同志社大学Graduate Student, Doshisha UniversityDoshisya University音融合とは,二つ以上の単位が音の転訛などによって一つになったもので,話し言葉によく見られる特徴である。本研究では,『名大会話コーパス(名大C)』『日本語話し言葉コーパス(CSJ)』『日本語日常会話コーパスモニター公開版(CEJC)』を用いて,仮定形における音融合の使用状況について調査し,場面別・性別・年代別に音融合の使用状況にどのような差が見られるかについて,計量的に明らかにする。また,『CSJ』に付与されている印象評定データを用いて,各講演における音融合生起率と印象評定の関連について分析する。 その結果,改まった場面では音融合生起率は低いこと,性別に見ると,男性のほうが女性よりも音融合生起率が高いこと,年代別に見ると,動詞仮定形音融合の生起率は50代以上において高いことがわかった。さらに,仮定形音融合使用と『CSJ』の印象評定データ項目の「講演の自発性」と「発話スタイル」に密接な関係があることもわかった。Contractions, in which two or more words are shortened by omitting or combining some sounds, are frequently used in spoken language. In this study, through the quantitative analysis of contractions of conditional forms using corpora such as the Nagoya University Conversation Corpus, Corpus of Spontaneous Japanese (CSJ), and the Corpus of Everyday Japanese Conversation, we clarify the differences in the use of contractions according to scene, gender, or generation. We also analyze the relationships between the occurrence of contractions in each lecture speech and the impression rating score data of the CSJ. The results prove that the usage of contractions is high in an informal setting, higher in men than in women, and the usage of conditional contractions of verbs is highest among those in their 50s. In addition, we demonstrate that the occurrence of contractions in each lecture speech and the impression rating scores are correlated.application/pdfdepartmental bulletin pape

    The Use of Parallel and Sequential Conjunctions in JFL Learners' Writing : An Analysis of the Longitudinal Corpus of Chinese Native Speakers

    Get PDF
    一橋大学大学院 博士後期課程Graduate Student, Hitotsubashi University本稿は,中国語母語話者の縦断作文コーパスを用いて,並列・継起の接続表現の使用実態を分析し,学習歴に即した習得過程の解明を目指すものである。1年次後半,2年次前半,2年次後半の作文を分析した結果,次のことがわかった。 (1)1年次後半では,テ形の使用も見られるが,「そして」に過剰に依存している。 (2)2年次前半では,テ形が定着するため,接続詞の使用率の低下と,接続助詞の使用率の上昇が見られ,同時に並列と継起の形態上の区別が可能になる。 (3)2年次後半では,接続表現のバリエーションは増加しないが,動詞連用形中止法の使用が見られ,作文が書き言葉らしい文体となる。また,段落の冒頭で接続詞を使えるようになり,使用位置に配慮しながら接続表現を選択できるようになる。This study aims to reveal the acquisition process of parallel and sequential conjunctions in the longitudinal writing corpus of Chinese JFL learners. By analyzing three writing data sets from the second half of the first year, the first half of the second year, and the second half of the second year, the following conclusions were reached. (1) In the second half of the first year, the conjunctive adverb soshite is overused, though the conjunctive particle te is generally used. (2) In the first half of the second year, because the use of te form is stabilized, the usage rate of conjunctive adverbs descends, and the usage rate of conjunctive particles increases. Simultaneously, the proper distinction between parallel and sequential conjunctions is observed. (3) In the second half of the second year, the conjunctive variations do not change, ut the suspended forms are used as the written style of the te form; moreover, the conjunctive adverbs are used at the beginning of the paragraph from the point of view of global cohesion.application/pdfdepartmental bulletin pape

    The Possibility of Research on Japanese Language or Linguistic Landscape Using 'Hankyu Culture Archives'

    Get PDF
    会議名: 言語資源活用ワークショップ2020, 開催地: オンライン, 会期: 2020年9月8日−9日, 主催: 国立国語研究所 コーパス開発センター2017年4月1日,公益財団法人阪急文化財団は,財団が所有する各種資料をインターネット上で検索・閲覧できる「阪急文化アーカイブズ」を公開した。「阪急文化アーカイブズ」で検索・閲覧できる資料の中でも,1910年の開業以来阪急電鉄が手がけた事業に関する掲示物や,阪急沿線のイベントを告知する掲示物である「阪急・宝塚ポスター」類は,日本語研究,中でも言語景観研究の貴重な資料となり得る可能性を秘めていると思われる。 本稿では,「阪急文化アーカイブズ」の概要を紹介したうえで,「阪急文化アーカイブズ」を利用した日本語研究,中でも言語景観研究の簡単な実践例を示す。そのうえで,「阪急文化アーカイブズ」を利用した日本語研究,中でも言語景観研究の可能性と限界を考える。application/pdf新潟大学阪急文化財団Niigata UniversityHankyu Culture Foundationconference pape

    Short Unit Word Annotation for the Corpus of Everyday Japanese Conversation : Procedures and Evaluation

    Get PDF
    会議名: 言語資源活用ワークショップ2020, 開催地: オンライン, 会期: 2020年9月8日−9日, 主催: 国立国語研究所 コーパス開発センター『日本語日常会話コーパス』(CEJC)の短単位情報付与作業では、以下のような作業工程を踏んでいる:(i) 転記をMeCab(解析器)+ UniDic(解析辞書)で自動解析、(ii) 音声を聴取しながら、付加情報の一つである「発音形」のみを人手修正、(iii) 人手修正された発音形を尊重しつつ再び自動解析、(iv) 短単位情報(境界情報、発音形以外の付加情報)を人手修正。この作業工程の妥当性を検証するため、人手修正済みデータを対象に、複数の版の現代話し言葉UniDic(Ver2.2.0, 2.3.0, 3.0.1)で自動解析をしなおし、出力を比較した。その結果、どの版のUniDicを使っても、人手修正された発音形の情報を用いる方が、そうでない場合に比べ、短単位情報の精度向上を見込めることがわかった。特に、古い版のUniDic (Ver2.2.0)ではそれが顕著であった(境界+品詞+語彙素(F値):0.944→0.962)。一方で、最新版のUniDic (Ver3.0.1)では効果は限定的である(同:0.976→0.979)。application/pdf国立国語研究所国立国語研究所National Institute for Japanese Language and LinguisticsNational Institute for Japanese Language and Linguisticsconference pape

    An Attempt to Qualitative Expansion to the "Word List by Semantic Principles"

    Get PDF
    会議名: 言語資源活用ワークショップ2020, 開催地: オンライン, 会期: 2020年9月8日−9日, 主催: 国立国語研究所 コーパス開発センター『分類語彙表』は初版の刊行以来,日本語研究に利用されてきた。しかし,2004年に増補改訂版が刊行されて以来,さらなる増補は行われていない。本稿は,『分類語彙表』を研究に利用する上で,もっとも重要な課題の一つである,不採録語を減らすという観点から,語彙の拡充の方法を分類体系の見直しを中心に検討し,試案を提示するものである。語彙の拡充の候補は以下のとおりである。(1)助詞・助動詞などの機能語(2)固有名詞(固有表現)(3)外国語(4)メタ言語(5)句読点などの記号類(6)語断片(7)未知語。(1)~(3)は,意味の付与が可能なもの,(4)以降は,意味付与が可能でない(必要が無い)ものである。助詞・助動詞などの機能語は品詞相当と考え,0番台を与える(例えば格助詞「が」に分類語彙表番号0.1000を与えるなど)。固有名詞(固有表現)は,現在の分類体系をできるだけ維持するのであれば,内包的表現の所属する分類項目に位置付けるのが妥当であろう(「アカデミー賞」「グラミー賞」は「1.3682 賞罰」に置くなど)。メタ言語的用法は意味分類には反映させず,「用法」という別フィールドで属性を記述する。また,句読点,語断片,未知語は意味付与が不要という属性を与えて区別することを考えている。以上のような拡張で,ほぼ全ての語に何らかの分類語彙表番号を与えることが可能となる。application/pdf国立国語研究所National Institute for Japanese Language and Linguisticsconference pape

    Relationships between Fundamental Frequencies and Conversation Situations Based on the Corpus of Everyday Japanese Conversation

    Get PDF
    会議名: 言語資源活用ワークショップ2020, 開催地: オンライン, 会期: 2020年9月8日−9日, 主催: 国立国語研究所 コーパス開発センター自発音声ではパラ言語情報や感情の影響によりピッチが様々に変動することが知られているが、日常生活の多様な状況を反映した音声データの不足により、自発音声のピッチの多様性について大規模な定量的分析を行うことが困難であった。国立国語研究所では,2016年より多様な種類の日常会話をバランス良く収録した大規模な日常会話コーパスとして『日本語日常会話コーパス』(CEJC)の構築を進めている。CEJCに収録される自発音声のうち50時間のデータを基に様々な会話場面における声の高さの違いを調べたところ、子どもや配偶者、父母といった家族に対しては低く、取引先や客など丁寧さが必要な相手や友人には高い声で話していることが示された。また、発話の直接の向け先だけではなく、会話場面に同席している参与者の属性によっても声の高さが変わることが観察された。application/pdf国立国語研究所National Institute for Japanese Language and Linguisticsconference pape

    KOTONOHA Contest 2020 Excellence Award Winner 1

    Get PDF
    会議名: 言語資源活用ワークショップ2020, 開催地: オンライン, 会期: 2020年9月8日−9日, 主催: 国立国語研究所 コーパス開発センターapplication/pdf筑波大学University of Tsukubaconference pape

    Dynamically Updating Event Representations for Temporal Relation Classification with Multi-category Learning

    Get PDF
    application/pdfKyoto UniversityNational Institute for Japanese Language and LinguisticsOchanomizu UniversityKyoto UniversityTemporal relation classification is a pair-wise task for identifying the relation of a temporal link (TLINK) between two mentions, i.e. event, time and document creation time (DCT). It leads to two crucial limits: 1) Two TLINKs involving a common mention do not share information. 2) Existing models with independent classifiers for each TLINK category (E2E, E2T and E2D) hinder from using the whole data. This paper presents an event centric model that allows to manage dynamic event representations across multiple TLINKs. Our model deals with three TLINK categories with multi-task learning to leverage the full size of data. The experimental results show that our proposal outperforms state-of-the-art models and two transfer learning baselines on both the English and Japanese data.journal articl

    Research Report on Mutsu Dialect : Endangered Languages and Dialects in Japan

    Get PDF
    application/pdf国立国語研究所東京外国語大学アジア・アフリカ言語文化研究所オークランド大学国立国語研究所名古屋大学白百合女子大学国立歴史民俗博物館research repor

    概要

    Get PDF
    application/pdfresearch repor

    3,205

    full texts

    3,427

    metadata records
    Updated in last 30 days.
    Academic Repository of the National Institute for Japanese Language and Linguistics / 国立国語研究所学術情報リポジトリ
    Access Repository Dashboard
    Do you manage Open Research Online? Become a CORE Member to access insider analytics, issue reports and manage access to outputs from your repository in the CORE Repository Dashboard! 👇