National Institute of Informatics, Center of Dataset Sharing and Collaborative Research: NII DSC Reference Portal / 国立情報学研究所
Not a member yet
    119 research outputs found

    Tokyo Institute of Technology Multilingual Speech Corpus - Indonesian (TITML-IDN)

    No full text
    音声認識用の学習データとして収集した,音素バランス文の読み上げ音声データベース.音素バランス文は,新聞記事や雑誌を基にしたテキストコーパスから選択した文章に,音素バランスを考慮して数十文を追加して作成した343文からなり,全話者が同一のリストを読み上げている.The Indonesian Phonetically Balanced Speech Corpus was developed for training the acoustic models of an automatic speech recognition system.This database contains Bahasa Indonesia speech data from 20 Indonesian speakers. Each speaker was asked to read 343 phonetically balanced sentences most of which selected from a text corpus

    慶應義塾大学 研究用感情音声データベース(Keio-ESD)

    No full text
    感情を含んだ合成音声の作成を目的として,モーラ数(2~6モーラ)とアクセント型の組み合わせごとに選んだ20文節語それぞれについて,47通りの感情で発声した音声を収集したデータベース. 感情の例:怒り,喜び,嫌悪,侮り,おかしい,心配,優しい,安堵,憤慨,羞恥,...(全47種類)目視にて付与した各音節区切り情報,上記収録音声を基に作成した12感情の合成音声(11語分)もあり.A set of human speech with vocal emotion spoken by a Japanese male speaker and a set of artificial speech that were synthesized by a system that had been developed using the subset of this database for training. Exsample of emotions: Angry, joyful, disgusting, downgrading, funny, worried, gentle, relief, indignation, shameful, etc. (47 emotions)Syllable bundary labels and synthesized speech of 12 emotions (11 words) created based on the recorded speech are attached

    東北大-松下 単語音声データベース(TMW)

    No full text
    Set 1 : 単語音声 音韻バランス 212語Set 2 : 単語音声 鉄道駅名・線名 3285語Set 1: Phonetically balanced 212 wordsSet 2: Railway station and line names 3285 word

    マルチモーダル音声認識評価環境(CENSREC-1-AV)

    No full text
    音声と口唇動画像を用いたバイモーダル音声認識用データ ・発話内容はCENSREC-1に準拠(連続数字1~7桁の読み上げ) ・音声とともにカラー映像と近赤外線映像を収録し,ムービーを時系列画像に分解して口唇付近のみ切り出した画像データを含む ・学習データ - オフィス環境で収録したクリーン音声・画像データ - 計3,234発話 ・テストデータ - 学習データと同じ環境の音声・画像データ - 同梱スクリプトにより生成される音声・画像データ 音声:乗用車走行雑音を重畳(雑音2種類,SNR6種類) 画像:走行中の乗用車内を明度値のガンマ補正によりシミュレート - 計1,963発話上記音声・画像データを対象とした音声認識実験を評価するための評価ツールCommon platform for evaluating independently speech recognition accuracy and speech interval detection under noisy environment.An evaluation corpus for audio-visual speech recognition of continuously spoken single digits in Japanese.The digit sequence of each utterance and the pronunciation of Japanese digits are the same as the CENSREC-1 (AURORA-2J) database.Including color and infrared mouth images, which were recorded simultaneously with speech. *Training data - Clean audio and visual data in an office environment - 3234 utterances in total *Test data - Clean audio and visual data in the office environment - Noisy data generated by baseline scripts > Audio: added two types of noise, driving on city streets and driving on expressway, at six different SNR levels > Image: simulated driving conditions using Gamma transform - 1963 utterances in tota

    身体情報付き男・女・子どもの母音音声データベース(JVPD)

    No full text
    日本語音声の標準的な科学資料としての公開を目的に作成された母音データベース.6歳から56歳まで(主に17歳以下),男女幅広い年齢層にわたる東京方言(共通語)話者の収録音声を「はー,ひー,ふー,へー,ほー」という音声ファイルに編集.大半の話者については,性別・年齢の他に,身長・体重の資料も収集している.This corpus has been developed in order to make the standard scientific material of spoken Japanese.The speech data of men, women, and children ranging between 6 and 56 years of age were edited into files containing /haa, hii, huu, hee, hoo/.The corpus also includes basic physical data on the speakers, such as their height and weight

    鶴岡調査音声データベース91-92(Tsuruoka91-92)

    No full text
    山形県鶴岡市を対象に実施された第3回共通語化調査(1991〜1992年)で録音された音声資料.調査員による面接形式(調査票に従った質問-回答形式)で実施.発音,アクセント,語彙に注目した78項目(91年調査),72項目(92年調査)を収録.1話者あたり約45分収録(調査員の音声も含む).Speech material recorded in the Third Investigation (1991-1992) in Tsuruoka, Yamagata.Each investigator interviewed the informant according to the investigation forms with question - answer mode.Answers to 78 questions regarding pronunciation, accent, and vocabulary were recorded.About 45 minutes speech is recorded for each speaker (including the speech of the investigator)

    音声研究用X線フィルムデータベース(X-Ray)

    No full text
    現在では困難なX線フィルムの高速撮影により,発声時の声道や舌の動きを鮮明に映像化したもの.話者ごとに異なる短文リスト約30文(1話者は無意味単語も含む)を読み上げ.This corpus offers the movies from high quality x-ray films compiled on CAV laserdisk.Each speaker read phonetically contrastive sentences (about 30 sentences per speaker).The movies yield the best dynamic view of the entire vocal tract and the complex movements of the tongue

    楽天GORAデータ

    No full text
    楽天GORAのサイトに掲載されたゴルフ施設データ (1,669施設)とレビューデータ (約32万レビュー)。楽天データセットの一部。詳細は「収録データセット」を参照。Golf facility data (1,669 facilities) and review data (320,000 reviews) published on Rakuten GORA site. A part of Rakuten Dataset. See the "Containing dataset" for the details

    特定領域研究「韻律と音声処理」日本語MULTEXT韻律コーパス(MULTEXT-J)

    No full text
    ヨーロッパで作成されたMultilingual Text Tools and Corpora(MULTEXT)の日本語版.1つが5~6文で構成される40の原稿を,次の2つの発声スタイルにて収録. 1. 朗読音声 2. 原稿の設定状況に従い,感情を込めたように演技した音声(模擬自発発話)The Japanese version of Multilingual Text Tools and Corpora (MULTEXT).The speakers were asked to read aloud the 40 passages (each passage includes 5-6 sentences) in following two speaking styles. 1. Reading-style 2. Spontaneous-style (instructed to perform with different emotional attitudes according to the text of each situation

    中国語MULTEXTコーパス(MULTEXT-C)

    No full text
    ヨーロッパで作成されたMultilingual Text Tools and Corpora(MULTEXT)の中国語版.1つが5〜6文で構成される40の原稿を,できるだけ自然に話すように指示して収録.1名は全40文章,残り9名は15文章を読み上げ(1文章あたり4〜5名).The Chinese version of Multilingual Text Tools and Corpora (MULTEXT).The speakers were asked to read aloud the 40 passages (each passage includes 5-6 sentences) as naturally as possible.1 speaker read all 40 passages, and each of the other 9 speakers read 15 passages (each passage was read by 4 or 5 speakers)

    0

    full texts

    0

    metadata records
    Updated in last 30 days.
    National Institute of Informatics, Center of Dataset Sharing and Collaborative Research: NII DSC Reference Portal / 国立情報学研究所
    Access Repository Dashboard
    Do you manage Open Research Online? Become a CORE Member to access insider analytics, issue reports and manage access to outputs from your repository in the CORE Repository Dashboard! 👇