National Institute for Japanese Language and Linguistics

Academic Repository of the National Institute for Japanese Language and Linguistics / 国立国語研究所学術情報リポジトリ
Not a member yet
    3427 research outputs found

    Can a story be generated from "the wisdom of crowds"? : Study on the possibility of story generation by social listening

    Get PDF
    会議名: 言語資源活用ワークショップ2021, 開催地: オンライン, 会期: 2021年9月13日-14日, 主催: 国立国語研究所 コーパス開発センターSNS上の投稿を、人々がその対象に対して抱く感情的な側面(サイコグラフィック変数)の表出と考え、それらを収集、分析するソーシャルリスニング手法は、マーケティング調査の重要な手段である。本研究では、ソーシャルリスニングを通した、特定のサービスや商品などに関する、物語性を持ったコンテンツ生成の可能性を検証する。物語を用いた情報伝達には、受け手側の認知や共感を高めるといった効果がある。しかし対象からどういう物語を生成、記述するかといったことに関しては、人間の感性、創造性に依存する部分があり、自動化、一般化することは難しい。そのため物語生成は、AI研究の応用としても、研究がなされている。本研究では、特定の対象に関する、SNS(Twitter)での表現に着目する。人々のサイコグラフィック変数を集約、分析することで、多くの人々による、対象に纏わる物語性が出現するのではないだろうか。こうした仮説に基づき、特定の対象についての、Twitter上での発信の語彙分析を行い、物語記述のための静的構造(時間、場所、行動主体、行動客体)の抽出と、物語を生成する可能性について、試行し検証する。これにより、特にコンテンツマーケティングへの貢献を目指すものである。application/pdfフェリス女学院大学上智大学FERRIS UniversitySophia Universityconference pape

    Orthographical Redundancy in Japanese Personal Names : A Relational Morphology Account

    Get PDF
    名古屋大学Nagoya University近年の日本人名には,漢字を冗長に用いるものが散見される。「水月/mizuki/」は/zu/の部分が,「咲希/saki/」は/ki/の部分が,「花奏/kanade/」は/ka/の部分が,それぞれ2つの漢字の読みの一部となっているように見える。本稿では,これらの非規範的な名前を関係形態論の観点から分析する。「水月」型の名前は,composium(compose + symposium)のような,構成要素どうしが音韻的に一部重なる混成語と同様に扱える。一方,「咲希」型と「花奏」型については,(The) Beatles(beat + beetles)のような包含型の混成語に似ているものの,構成要素の表記をともに残している点で異なる。この独自性は,漢字を用いた命名法の創造性の一端を示している。Numerous Japanese personal names in recent times have started including phoneme sequences that are dually represented by two kanji characters. For example, 水月 /mizuki/ is a female name whose components—水 /mizu/ ‘water’ and 月 /zuki/ ‘moon’—phonologically overlap with each other. 咲希 /saki/ is a female name in which 咲 ‘to flower’ alone can be pronounced /saki/ but is followed by the extra character 希 /ki/ ‘hope’. 花奏 /kanade/ is a female name that illustrates the opposite: 花 /ka/ ‘flower’ + 奏 /kanade/ ‘to play (a musical instrument)’. This paper analyzes these noncanonical names within the framework of Relational Morphology. It demonstrates that 水月-type names are parallel to blends with phonological overlap, such as composium (compose + symposium); 咲希- and 花奏-type names are similar to blends with phonological inclusion, such as (The) Beatles (beat + beetles), but retain their components' orthography, suggesting the creativity of logographic naming.application/pdfdepartmental bulletin pape

    Scheme for a Structured Description of an Old Japanese Dictionary : The Case of Wamyō-Ruijushō

    Get PDF
    京都府立大学韓国国立釜慶大学 大学院生国立国語研究所 研究系 言語変化研究領域Kyoto Prefectural UniversityGraduate Student, Pukyong National UniversityLanguage Change Division, Research Department, NINJAL古代の日本の辞書には,様々な構造を持つものがあり,各辞書の構成や仕様を理解していなければ解読が困難な面があった。また注文から必要な情報を抽出するためには,隈なく目視で捜索する必要があった。順不同に入り組んだ注文の情報から,効率的に目的の情報に到達するためには,注文に存在する要素の属性が,それぞれ可能な限り定義づけられているべきである。本稿では,平安時代の代表的な漢和辞書である『和名類聚抄』を例として,いかにその構造を記述することが可能か,検討し,『和名類聚抄』の内容に適したタグを設計した。There were many kinds of dictionaries in ancient Japan. Therefore, without highly specialized knowledge, it is quite difficult for users to interpret their contents. Furthermore, users have been compelled to seek the necessary information based on visual observation of the dictionaries. To effectively find the required information within a complex description, a strict definition of the contents is required.This paper thus proposes a model for such a structured description and discloses the tagged data by using Wamyō-Ruijushō, a representative dictionary of the Heian Era in Japan.application/pdfdepartmental bulletin pape

    How to Correctly Morphologically Analyze Text Containing a Mixture of Old- and New-Style Kanji Scripts

    Get PDF
    神戸松蔭女子学院大学Kobe Shoin Women’s University旧字体と新字体の混在するテキストは,形態素解析において誤解析の原因となることが多く,その対策としては形態素解析辞書の記載に異体字を加える方法,そして予め漢字を新字体に置換しておく方法,また複数の辞書を使い分けるといった方法が考えられる。本稿では字体置換6通りと,辞書の使い分け3通りを掛け合わせた18組の組み合わせで國/国,會/会,關/関3対の旧/新字体の対を含んだテキストの形態素解析を行うことで,目的とする漢字を含む形態素がどれほど正確に切り出せるのかを検討した。データとして第1~10回までの国会会議録を用いた。結果は,漢字置換で隣接する漢字が旧字体の場合に旧字体に置換し,隣接しない場合は新字体とするという置換法(デフォルトを新字体とする日和見置換)と,すべてについて近代文語UniDicを用いるか,1949年の当用漢字字体表告示を境として,それ以前では近代文語UniDicを用い,それ以後では現代語書き言葉UniDicを用いる方法が,もっとも正確に当該漢字を含む短単位形態素を切り出せるというものであった。形態素解析辞書の記載に異体字を加える方法には,異体字が記載されていない形態素が出現した場合に対応ができないという欠点があるのに対して,漢字置換と辞書の使い分けを活用する方法は,そうした場合にも柔軟に対応が可能であるという利点があることを主張した。Japanese texts containing a mixture of old-(kyūjitai) and new-(shinjitai) style kanji scripts pose a serious problem for an automatic morphological analyzer. However, recent developments in various dictionaries by era, undertaken by the corpora project at NINJAL, have brought about a new opportunity to solve this problem. Another promising solution is to replace the script in the text in some way, so that the analyzer can correctly identify the characters/morphemes. We designed an experiment with three dictionary selection methods and six replacement methods using three pairs of old/new kanji scripts (國/国, 會/会 and 關/関) to determine which combination would result in the most precise analysis. An analysis of the text data from the Minutes of the National Diet between 1947 and 1951 demonstrated that, of the 18 combinations, two dictionaries gave the best results. These were, the Contemporary Written Japanese UniDic dictionary up to the public notification of the Table of Script Styles of Jōyō Kanji on April 29, 1949, and The Modern Literary UniDic. With these, we coupled a replacement of a kanji script with an old counterpart when its immediate neighbor was also an old one, and with a new one when it was not. Although the addition of the different scripts to the dictionary entries would be another viable solution, our method is more desirable in that it is applicable to a wider range of texts without dictionary entry modifications.application/pdfdepartmental bulletin pape

    Nonrestrictive Adnominal Clauses and Text Genre in Japanese

    Get PDF
    実践女子大学Jissen Women's University本稿では,統語情報付きコーパスであるNPCMJを用いて,日本語の非制限的連体修飾構造に見られるテキストジャンル間の分布の差異を明らかにする。形態的情報に基づく従来のコーパスでは,任意の統語的環境を指定して連体修飾構造を量的に検索することは困難であった。一方で,統語・意味情報付きコーパスであるNPCMJを用いれば,主節環境や被修飾名詞の語彙的性質を指定したうえで連体修飾構造を検索することが可能になる。本研究の調査の結果,新聞などのいわゆる説明的文章においては「主名詞に対する情報付加」を行う非制限的連体修飾節が多く見られる一方で,小説などのいわゆる文学的文章においては,「主節に対する情報付加」を行う非制限的連体修飾節が相対的に多く見られた。このような分布は,新出の固有名詞が頻出するという説明的文章のテキストジャンル的特徴によって説明することが可能になると考えられる。This paper clarifies the differences in the distribution of nonrestrictive relative clauses among different text genres in Japanese by using a parsed corpus with syntactic and semantic information, the NINJAL Parsed Corpus of Modern Japanese (NPCMJ). Conventional corpora based on morphological information are difficult to quantify by specifying an arbitrary syntactic environment. In contrast, the NPCMJ, a corpus with syntactic and semantic information, can be used to specify the syntactic environment of the clause and the lexical properties of the head noun and to search for adnominal constructions. The results of this survey demonstrate that nonrestrictive relative clauses in Japanese occur more frequently in expository texts than in literary texts, and in the former text genre, they tend to be associated with the function of adding information to the head noun (rather than to the main clause as a whole). This distribution is explained by the characteristics of expository texts: the frequent occurrence of discourse-new proper nouns.application/pdfdepartmental bulletin pape

    [Relevant Data] Prosodic materials of the fusioned verbal forms of Tarama dialect

    No full text
    本データベースは南琉球宮古語多良間方言の動詞の進行融合形のアクセント資料を収納する。各アクセント資料に対して、「話者」「対象語」「枠文」「調」「書き起こし」「備考」「音声ファイル名」などの情報が収録されている。application/vnd.openxmlformats-officedocument.spreadsheetml.sheet国立国語研究所 研究系 言語変異研究領域Language Variation Division, Research Department, NINJALThis database contains the prosodic materials bearing on the fusioned verbal forms of Tarama dialect, a dialect of Southern Ryukyuan Miyako. For each prosodic material, the information for speaker, target word, frame sentence, allotone, transcription, comments and name of the corresponding sound file is given.datase

    戦中期のアメリカにおける日本語教育

    No full text
    国立国語研究所journal articl

    ビジネス文書における「カッコ」の使われ方

    Get PDF
    application/pdf会議名: 国立国語研究所オープンハウス2021, 開催地: オンライン, 会期: 2021年9月10日conference objec

    方言コーパスを使ってみよう

    Get PDF
    application/pdf会議名: 国立国語研究所オープンハウス2021, 開催地: オンライン, 会期: 2021年9月10日conference objec

    100年前の言語関係新聞記事から : 「キラキラネーム」「女学生語」「間違い言葉」

    Get PDF
    application/pdf会議名: 国立国語研究所オープンハウス2021, 開催地: オンライン, 会期: 2021年9月10日conference objec

    3,205

    full texts

    3,427

    metadata records
    Updated in last 30 days.
    Academic Repository of the National Institute for Japanese Language and Linguistics / 国立国語研究所学術情報リポジトリ
    Access Repository Dashboard
    Do you manage Open Research Online? Become a CORE Member to access insider analytics, issue reports and manage access to outputs from your repository in the CORE Repository Dashboard! 👇