National Institute for Japanese Language and Linguistics

Academic Repository of the National Institute for Japanese Language and Linguistics / 国立国語研究所学術情報リポジトリ
Not a member yet
    3427 research outputs found

    Design and Evaluation of the Corpus of Everyday Japanese Conversation

    Get PDF
    application/pdfNational Institute for Japanese Language and LinguisticsNational Institute for Japanese Language and LinguisticsGraduate School of Humanities, Chiba UniversityNational Institute for Japanese Language and LinguisticsNational Institute for Japanese Language and LinguisticsNational Institute for Japanese Language and LinguisticsNational Institute for Japanese Language and LinguisticsNational Institute for Japanese Language and LinguisticsNational Institute for Japanese Language and LinguisticsNational Institute for Japanese Language and LinguisticsNational Institute for Japanese Language and LinguisticsWe have constructed the Corpus of Everyday Japanese Conversation (CEJC) and published it in March 2022. The CEJC is designed to contain various kinds of everyday conversations in a balanced manner to capture their diversity. The CEJC features not only audio but also video data to facilitate precise understanding of the mechanism of real-life social behavior. The publication of a large-scale corpus of everyday conversations that includes video data is a new approach. The CEJC contains 200 hours of speech, 577 conversations, about 2.4 million words, and a total of 1675 conversants. In this paper, we present an overview of the corpus, including the recording method and devices, structure of the corpus, formats of video and audio files, transcription, and annotations. We then report some results of the evaluation of the CEJC in terms of conversant and conversation attributes. We show that the CEJC includes a good balance of adult conversants in terms of gender and age, as well as a variety of conversations in terms of conversation forms, places, activities, and numbers of conversants.conference pape

    Corpus of Japanese Telephone Conversation at Hiroshima University : Design and Current Status

    Get PDF
    国立国語研究所 研究系 言語変異研究領域Language Variation Division, Research Department, NINJAL『広島大学日本語電話会話コーパス』(COTCO-H)は現在開発中の大規模音声データベースである。COTCO-Hは,広島大学の日本語非標準変種の母語話者である50名の学生が2つのレジスター(出身地の友人との会話,キャンパスの友人との会話)で発話した電話会話を格納している。本コーパスには,約11万語(22時間)の音声信号に加えて,その転記および品詞や活用などの形態論情報が付与されている。分節音情報付与作業は現在進行中である。COTCO-Hにはさらに補助データとして同じ話者による読み上げ音声も含まれている。COTCO-Hは,地域や発話スタイル,自発性などの違いによる言語変異に興味を持つ研究者のコミュニティに貢献するものとなるだろう。The Corpus of Japanese Telephone Conversation at Hiroshima University (COTCO-H) is a large-scale speech database that is currently under development. COTCO-H contains spontaneous telephone conversations in two different registers (conversations with a local friend and with a campus friend) produced by 50 Hiroshima University students who are native speakers of nonstandard varieties of Japanese. The corpus consists of speech signals and transcriptions for approximately 110,000 words (22 hours), along with morphological annotations such as parts of speech and conjugations. Segmental labeling is currently in progress. COTCO-H also contains different types of read speech produced by the same speakers as auxiliary data. The corpus will contribute to a community of researchers interested in variations across different regions, speech styles, and spontaneity.application/pdfdepartmental bulletin pape

    鹿児島県大島郡龍郷町浦

    Get PDF
    application/pdf広島経済大学 / 奄美看護福祉専門学校広島大学research repor

    On Factors Contributing to Accent Shift

    Get PDF
    application/pdfORIGINAL PAPER 秋永一枝, 1957, 「アクセント推移の要因について」, 『国語学』31, 国語学会(現日本語学会), pp. 17-27. AKINAGA Kazue, 1957, Akusento-sui'i no yōin ni tsuite, Kokugogaku 31, Kokugo Gakkai [Society for the Study of Japanese Language], pp. 17-27. Translated by Wayne Lawrence (The University of Auckland) Proofed by Kikuo Maekawa (NINJAL)journal articl

    [Relevant Data] Prosodic materials of the Southern Ryukyuan Yaeyama Miyara dialect

    No full text
    本エクセルファイルは南琉球八重山語宮良方言の720語の名詞に関するアクセント資料を収納する。各語に対して、「ID」「仮名表記」「IPA」「型」「意味」の情報が収録されている。application/vnd.openxmlformats-officedocument.spreadsheetml.sheet国立国語研究所 研究系 言語変異研究領域名桜大学東京大学Language Variation Division, Research Department, NINJALMeio UniversityThe University of TokyoThis excel file contains the prosodic materials bearing on 720 nouns of the Miyara dialect, a dialect of Southern Ryukyuan Yaeyama. For each prosodic material, the information for ID, Kana orthography, IPA, tone class and meaning is given.datase

    2020 NINJAL Yearbook

    Get PDF
    application/pdfothe

    編集後記

    Get PDF
    application/pdf国立国語研究所National Institute for Japanese Language and Linguistics会議名: Evidence-based Linguistics Workshop 2022, 開催地: 国立国語研究所, 会期: 2022/09/05-06, 主催: 国立国語研究所、神戸大学人文学研究科othe

    Digital data of the glossary of the Minna dialect (31st March 2022 edition)

    No full text
    関連資料:セリック・ケナン、大浦辰夫(2022)『みんなふつ語彙集』国立国語研究所言語変異研究領域application/vnd.openxmlformats-officedocument.spreadsheetml.sheetdatase

    Infinite SCAN: An Infinite Model of Diachronic Semantic Change

    No full text
    Tokyo Metropolitan UniversityTokyo Metropolitan UniversityThe National Institute for Japanese Language and LinguisticsThe National Institute of Advanced Industrial Science and TechnologyThe Institute of Statistical MathematicsIn this study, we propose a Bayesian model that can jointly estimate the number of senses of words and their changes through time.The model combines a dynamic topic model on Gaussian Markov random fields with a logistic stick-breaking process that realizes Dirichlet process. In the experiments, we evaluated the proposed model in terms of interpretability, accuracy in estimating the number of senses, and tracking their changes using both artificial data and real data.We quantitatively verified that the model behaves as expected through evaluation using artificial data.Using the CCOHA corpus, we showed that our model outperforms the baseline model and investigated the semantic changes of several well-known target words.journal articl

    『比喩表現の理論と分類』データベース版

    No full text
    国立国語研究所報告57 『比喩表現の理論と分類』(http://doi.org/10.15084/00001254)の指標比喩・結合比喩例文などをMicrosoft Excel 形式にし、付加情報を付与したものapplication/zip国立国語研究所National Institute for Japanese Language and Linguisticsdatase

    3,205

    full texts

    3,427

    metadata records
    Updated in last 30 days.
    Academic Repository of the National Institute for Japanese Language and Linguistics / 国立国語研究所学術情報リポジトリ
    Access Repository Dashboard
    Do you manage Open Research Online? Become a CORE Member to access insider analytics, issue reports and manage access to outputs from your repository in the CORE Repository Dashboard! 👇