13197 research outputs found
Sort by
Japanese Neural Incremental Text-to-Speech Synthesis Framework With an Accent Phrase Input
Work in the development of neural incremental text-to-speech (iTTS), which is attracting increasing attention, has recently pursued low-latency processing by generating speech on the fly before reading complete sentences. Most current state-of-the-art iTTS systems use a prefix-to-prefix neural iTTS framework with look-ahead of 1-2 unit segments (i.e., phonemes or words). However, since the Japanese language is based on accent phrase units that are longer than words, using a prefix-to-prefix neural iTTS with a look-ahead approach increases latency. Here, we propose an alternative to the end-to-end neural iTTS architecture that does not apply look-ahead input when synthesizing speech chunks. We further propose a method to use information from the previous time step by connecting the synthesized vector and the model’s internal state to the current time step. We experimentally investigated the latency of various iTTS systems with different modeling and synthesis chunks. The experimental results show that, for Japanese, the proposed iTTS is able to synthesize better speech quality, with a similar latency range, than the conventional baseline prefix-to-prefix neural iTTS with word units. Moreover, we found that our proposed approach improved the prosodic naturalness among synthesized units in the Japanese language. Subjective evaluations also revealed that the proposed approach with an incremental unit of two accent phrases achieved the best scores in Japanese iTTS systems.journal articl
ホウソウキョク オ オウダン スル ダイキボ テレビ シチョウ リレキ データ ノ トウゴウ シュホウ ノ テイアン ト ジッセン
近年,各テレビ放送局において,個人を特定しない形式で,インターネット接続されたテレビから視聴開始時刻や視聴終了時刻等を含む非特定視聴履歴データを収集し,利活用する取り組みが進められている.しかし,各放送局は自局の非特定視聴履歴データしか利用できないため,膨大なデータを蓄積しているにも関わらず,有用な知見を得るまでに至っていないのが現状である.さらに,非特定視聴履歴データの収集方式やデータ粒度は,各社各様となっており,各局が蓄積したデータを統合し,利用することもできていない.そこで本稿では,各局が独自の方式で取得している非特定視聴履歴データを放送局間でマッチングする手法を提案し,データ統合を行う.提案手法では,視聴履歴データ収集時に集めているIPアドレス・郵便番号・メーカID・ブラウザメジャーバージョン・ブラウザマイナーバージョンの5項目とチャンネル遷移タイミングが一致するテレビを同一テレビと推定する.在阪放送局4社にて,放送局間での非特定視聴履歴データ連携が技術的に可能か検証した「テレビ視聴データ連携に関する共同技術検証実験」において本手法を適用した結果,各放送局で取得された約376万台分のデータのうち約267万台分(約71.0%)のテレビをマッチングできることを確認した.journal articl
CISO ノ ノウリョク シュウトク ノ タメ ノ ケイエイ ギジュツ ノ リョウメン オ イシキシタ サイバー セキュリティ エンシュウ ノ コンテンツ ニカンスル ケンキュウ
奈良先端科学技術大学院大学修士(工学)master thesi
A Data Collection Protocol, Tool, and Analysis of Multimodal Data at Different Speech Voice Levels for Avatar Facial Animation
奈良先端科学技術大学院大学修士(工学)master thesi
The University of Trento. Opportunities for research and teaching collaborations
video/mp4In this seminar, I briefly present the University of Trento followed by a more detailed presentation of the Department of Information Engineering and Computer Science. In particular, I will present the degree programs and our Ph.D. school highlighting the opportunities we have to collaborate with international institutions. To finish I will provide an overview of the main results and research carried out by the different groups that are present in the department.講演日: 2023年6月26日 4限講演場所: 情報科学棟中講義室(L2), IS_L2vide
Special Speech in Start-up Session, NAIST STELLA Program 2023
video/mp4講演日:2023年8月25日講演場所: 情報科学棟 エーアイ大講義室(L1)講演者所属:UN Women civil society and private sectorrepresentative from Papua New Guineavide
A Framework to Evaluate the Reliability of Obfuscating Transformations in Program Code
奈良先端科学技術大学院大学修士(工学)master thesi