| 講演抄録/キーワード |
| 講演名 |
2008-03-20 15:15
[ポスター講演]Prosody Reconstruction by Rescaling Fundamental Frequency Contours in Order to Synthesize Communicative Speech ○Jinfu Ni・Shinsuke Sakai・Satoshi Nakamura(NICT/ATR) SP2007-193 |
| 抄録 |
(和) |
This paper presents a method of prosody reconstruction that can be used to synthesize conversational speech. In our method, we use a conventional text-to-speech engine to initially generate reading-style prosody for input text. We then use a frequency modulation technique to rescale the fundamental frequency (F0) contours to add the communicative functions of intonation to the synthesized speech. The frequency modulation technique is based on a functional F0 model, and the transformation scales are modeled by combining simple piecewise-linear patterns according to input tags. We conducted two experiments to evaluate our method: modulating the F0 range of reading-style prosody when synthesizing Japanese speech to convey "good news" and "bad news" (Experiment 1), and making a narrow focus when synthesizing Chinese dialog to convey emphasis (Experiment 2). The results of Experiment 1 showed that the listeners judged 94% of samples modulated with "bad news" F0 ranges as "bad news" and 78% of samples with "good news" F0 ranges as "good news." They are comparable with those obtained by style-specified corpora in our previous work. In Experiment 2, 90% of the samples with narrow focuses were identified. The results showed that proposed method could use paralinguistic information to achieve specific communicative purposes. |
| (英) |
This paper presents a method of prosody reconstruction that can be used to synthesize conversational speech. In our method, we use a conventional text-to-speech engine to initially generate reading-style prosody for input text. We then use a frequency modulation technique to rescale the fundamental frequency (F0) contours to add the communicative functions of intonation to the synthesized speech. The frequency modulation technique is based on a functional F0 model, and the transformation scales are modeled by combining simple piecewise-linear patterns according to input tags. We conducted two experiments to evaluate our method: modulating the F0 range of reading-style prosody when synthesizing Japanese speech to convey "good news" and "bad news" (Experiment 1), and making a narrow focus when synthesizing Chinese dialog to convey emphasis (Experiment 2). The results of Experiment 1 showed that the listeners judged 94% of samples modulated with "bad news" F0 ranges as "bad news" and 78% of samples with "good news" F0 ranges as "good news." They are comparable with those obtained by style-specified corpora in our previous work. In Experiment 2, 90% of the samples with narrow focuses were identified. The results showed that proposed method could use paralinguistic information to achieve specific communicative purposes. |
| キーワード |
(和) |
Conversational speech synthesis / prosody reconstruction / intonation synthesis / fundamental frequency control / text-to-speech synthesis / / / |
| (英) |
Conversational speech synthesis / prosody reconstruction / intonation synthesis / fundamental frequency control / text-to-speech synthesis / / / |
| 文献情報 |
信学技報, vol. 107, no. 551, SP2007-193, pp. 39-44, 2008年3月. |
| 資料番号 |
SP2007-193 |
| 発行日 |
2008-03-13 (SP) |
| ISSN |
Print edition: ISSN 0913-5685 Online edition: ISSN 2432-6380 |
著作権に ついて |
技術研究報告に掲載された論文の著作権は電子情報通信学会に帰属します.(許諾番号:10GA0019/12GB0052/13GB0056/17GB0034/18GB0034) |
| PDFダウンロード |
SP2007-193 |
| 研究会情報 |
| 研究会 |
SP |
| 開催期間 |
2008-03-20 - 2008-03-21 |
| 開催地(和) |
東大 |
| 開催地(英) |
Univ. Tokyo |
| テーマ(和) |
国際ワークショップ(20日),研究会(21日) 「聴覚・音声・言語とその障害,一般」 |
| テーマ(英) |
International Workshop (Mar 20), Speech Production, Speech Perception, Hearing and Speech, etc. (Mar 21) |
| 講演論文情報の詳細 |
| 申込み研究会 |
SP |
| 会議コード |
2008-03-SP |
| 本文の言語 |
英語 |
| タイトル(和) |
|
| サブタイトル(和) |
|
| タイトル(英) |
Prosody Reconstruction by Rescaling Fundamental Frequency Contours in Order to Synthesize Communicative Speech |
| サブタイトル(英) |
|
| キーワード(1)(和/英) |
Conversational speech synthesis / Conversational speech synthesis |
| キーワード(2)(和/英) |
prosody reconstruction / prosody reconstruction |
| キーワード(3)(和/英) |
intonation synthesis / intonation synthesis |
| キーワード(4)(和/英) |
fundamental frequency control / fundamental frequency control |
| キーワード(5)(和/英) |
text-to-speech synthesis / text-to-speech synthesis |
| キーワード(6)(和/英) |
/ |
| キーワード(7)(和/英) |
/ |
| キーワード(8)(和/英) |
/ |
| 第1著者 氏名(和/英/ヨミ) |
Jinfu Ni / Jinfu Ni / |
| 第1著者 所属(和/英) |
NICT/ATR (略称: NICT/ATR)
NICT/ATR (略称: NICT/ATR) |
| 第2著者 氏名(和/英/ヨミ) |
Shinsuke Sakai / Shinsuke Sakai / |
| 第2著者 所属(和/英) |
NICT/ATR (略称: NICT/ATR)
NICT/ATR (略称: NICT/ATR) |
| 第3著者 氏名(和/英/ヨミ) |
Satoshi Nakamura / Satoshi Nakamura / |
| 第3著者 所属(和/英) |
NICT/ATR (略称: NICT/ATR)
NICT/ATR (略称: NICT/ATR) |
| 第4著者 氏名(和/英/ヨミ) |
/ / |
| 第4著者 所属(和/英) |
(略称: )
(略称: ) |
| 第5著者 氏名(和/英/ヨミ) |
/ / |
| 第5著者 所属(和/英) |
(略称: )
(略称: ) |
| 第6著者 氏名(和/英/ヨミ) |
/ / |
| 第6著者 所属(和/英) |
(略称: )
(略称: ) |
| 第7著者 氏名(和/英/ヨミ) |
/ / |
| 第7著者 所属(和/英) |
(略称: )
(略称: ) |
| 第8著者 氏名(和/英/ヨミ) |
/ / |
| 第8著者 所属(和/英) |
(略称: )
(略称: ) |
| 第9著者 氏名(和/英/ヨミ) |
/ / |
| 第9著者 所属(和/英) |
(略称: )
(略称: ) |
| 第10著者 氏名(和/英/ヨミ) |
/ / |
| 第10著者 所属(和/英) |
(略称: )
(略称: ) |
| 第11著者 氏名(和/英/ヨミ) |
/ / |
| 第11著者 所属(和/英) |
(略称: )
(略称: ) |
| 第12著者 氏名(和/英/ヨミ) |
/ / |
| 第12著者 所属(和/英) |
(略称: )
(略称: ) |
| 第13著者 氏名(和/英/ヨミ) |
/ / |
| 第13著者 所属(和/英) |
(略称: )
(略称: ) |
| 第14著者 氏名(和/英/ヨミ) |
/ / |
| 第14著者 所属(和/英) |
(略称: )
(略称: ) |
| 第15著者 氏名(和/英/ヨミ) |
/ / |
| 第15著者 所属(和/英) |
(略称: )
(略称: ) |
| 第16著者 氏名(和/英/ヨミ) |
/ / |
| 第16著者 所属(和/英) |
(略称: )
(略称: ) |
| 第17著者 氏名(和/英/ヨミ) |
/ / |
| 第17著者 所属(和/英) |
(略称: )
(略称: ) |
| 第18著者 氏名(和/英/ヨミ) |
/ / |
| 第18著者 所属(和/英) |
(略称: )
(略称: ) |
| 第19著者 氏名(和/英/ヨミ) |
/ / |
| 第19著者 所属(和/英) |
(略称: )
(略称: ) |
| 第20著者 氏名(和/英/ヨミ) |
/ / |
| 第20著者 所属(和/英) |
(略称: )
(略称: ) |
| 第21著者 氏名(和/英/ヨミ) |
/ / |
| 第21著者 所属(和/英) |
(略称: )
(略称: ) |
| 第22著者 氏名(和/英/ヨミ) |
/ / |
| 第22著者 所属(和/英) |
(略称: )
(略称: ) |
| 第23著者 氏名(和/英/ヨミ) |
/ / |
| 第23著者 所属(和/英) |
(略称: )
(略称: ) |
| 第24著者 氏名(和/英/ヨミ) |
/ / |
| 第24著者 所属(和/英) |
(略称: )
(略称: ) |
| 第25著者 氏名(和/英/ヨミ) |
/ / |
| 第25著者 所属(和/英) |
(略称: )
(略称: ) |
| 第26著者 氏名(和/英/ヨミ) |
/ / |
| 第26著者 所属(和/英) |
(略称: )
(略称: ) |
| 第27著者 氏名(和/英/ヨミ) |
/ / |
| 第27著者 所属(和/英) |
(略称: )
(略称: ) |
| 第28著者 氏名(和/英/ヨミ) |
/ / |
| 第28著者 所属(和/英) |
(略称: )
(略称: ) |
| 第29著者 氏名(和/英/ヨミ) |
/ / |
| 第29著者 所属(和/英) |
(略称: )
(略称: ) |
| 第30著者 氏名(和/英/ヨミ) |
/ / |
| 第30著者 所属(和/英) |
(略称: )
(略称: ) |
| 第31著者 氏名(和/英/ヨミ) |
/ / |
| 第31著者 所属(和/英) |
(略称: )
(略称: ) |
| 第32著者 氏名(和/英/ヨミ) |
/ / |
| 第32著者 所属(和/英) |
(略称: )
(略称: ) |
| 第33著者 氏名(和/英/ヨミ) |
/ / |
| 第33著者 所属(和/英) |
(略称: )
(略称: ) |
| 第34著者 氏名(和/英/ヨミ) |
/ / |
| 第34著者 所属(和/英) |
(略称: )
(略称: ) |
| 第35著者 氏名(和/英/ヨミ) |
/ / |
| 第35著者 所属(和/英) |
(略称: )
(略称: ) |
| 第36著者 氏名(和/英/ヨミ) |
/ / |
| 第36著者 所属(和/英) |
(略称: )
(略称: ) |
| 講演者 |
第1著者 |
| 発表日時 |
2008-03-20 15:15:00 |
| 発表時間 |
90分 |
| 申込先研究会 |
SP |
| 資料番号 |
SP2007-193 |
| 巻番号(vol) |
vol.107 |
| 号番号(no) |
no.551 |
| ページ範囲 |
pp.39-44 |
| ページ数 |
6 |
| 発行日 |
2008-03-13 (SP) |
|