| Paper Abstract and Keywords |
| Presentation |
2011-07-21 13:30
One-model Speech Recognition and Synthesis System Based on Articulatory Masashi Kimura, Takayuki Onoda, Yurie Iribe, Kouichi Katsurada, Tsuneo Nitta (Toyohashi Tech) SP2011-41 |
| Abstract |
(in Japanese) |
(See Japanese page) |
| (in English) |
Speech recognition (SR) and speech synthesis (SS) based on one-model of articulatory movement HMMs that are commonly applied to both an SR module and an SS module are described. The SR module has an articulatory feature (AF) extractor with multi-layer neural networks (MLNs) that output an AF sequence to HMMs. In the SS module, the speaker-invariant HMMs are applied to generate an articulatory feature (AF) sequence, and then, after converting AFs into vocal tract parameters by using a multi-layer neural network (MLN), a speech signal is synthesized by an LSP (Line Spectrum Pairs) digital filter. CELP coding technique is applied to improve sound quality when generating voice source from embedded codes in the corresponding state of HMMs. The proposed speech synthesis system separate phonetic information and speaker individuality. Therefore, target speaker’s voice can be synthesized with a small amount of speech data. The experimental results show that the proposed system can produce good quality speech with only two-sentences. |
| Keyword |
(in Japanese) |
(See Japanese page) |
| (in English) |
Articulatory Features / One-model Speech Recognition and Synthesis / CELP codebook / LSP / / / / |
| Reference Info. |
IEICE Tech. Rep., vol. 111, no. 153, SP2011-41, pp. 1-6, July 2011. |
| Paper # |
SP2011-41 |
| Date of Issue |
2011-07-14 (SP) |
| ISSN |
Print edition: ISSN 0913-5685 Online edition: ISSN 2432-6380 |
Copyright and reproduction |
All rights are reserved and no part of this publication may be reproduced or transmitted in any form or by any means, electronic or mechanical, including photocopy, recording, or any information storage and retrieval system, without permission in writing from the publisher. Notwithstanding, instructors are permitted to photocopy isolated articles for noncommercial classroom use without fee. (License No.: 10GA0019/12GB0052/13GB0056/17GB0034/18GB0034) |
| Download PDF |
SP2011-41 |
| Conference Information |
| Committee |
SP |
| Conference Date |
2011-07-21 - 2011-07-23 |
| Place (in Japanese) |
(See Japanese page) |
| Place (in English) |
Jozankei Grand Hotel |
| Topics (in Japanese) |
(See Japanese page) |
| Topics (in English) |
|
| Paper Information |
| Registration To |
SP |
| Conference Code |
2011-07-SP |
| Language |
Japanese |
| Title (in Japanese) |
(See Japanese page) |
| Sub Title (in Japanese) |
(See Japanese page) |
| Title (in English) |
One-model Speech Recognition and Synthesis System Based on Articulatory |
| Sub Title (in English) |
|
| Keyword(1) |
Articulatory Features |
| Keyword(2) |
One-model Speech Recognition and Synthesis |
| Keyword(3) |
CELP codebook |
| Keyword(4) |
LSP |
| Keyword(5) |
|
| Keyword(6) |
|
| Keyword(7) |
|
| Keyword(8) |
|
| 1st Author's Name |
Masashi Kimura |
| 1st Author's Affiliation |
Toyohashi University of Technology (Toyohashi Tech) |
| 2nd Author's Name |
Takayuki Onoda |
| 2nd Author's Affiliation |
Toyohashi University of Technology (Toyohashi Tech) |
| 3rd Author's Name |
Yurie Iribe |
| 3rd Author's Affiliation |
Toyohashi University of Technology (Toyohashi Tech) |
| 4th Author's Name |
Kouichi Katsurada |
| 4th Author's Affiliation |
Toyohashi University of Technology (Toyohashi Tech) |
| 5th Author's Name |
Tsuneo Nitta |
| 5th Author's Affiliation |
Toyohashi University of Technology (Toyohashi Tech) |
| 6th Author's Name |
|
| 6th Author's Affiliation |
() |
| 7th Author's Name |
|
| 7th Author's Affiliation |
() |
| 8th Author's Name |
|
| 8th Author's Affiliation |
() |
| 9th Author's Name |
|
| 9th Author's Affiliation |
() |
| 10th Author's Name |
|
| 10th Author's Affiliation |
() |
| 11th Author's Name |
|
| 11th Author's Affiliation |
() |
| 12th Author's Name |
|
| 12th Author's Affiliation |
() |
| 13th Author's Name |
|
| 13th Author's Affiliation |
() |
| 14th Author's Name |
|
| 14th Author's Affiliation |
() |
| 15th Author's Name |
|
| 15th Author's Affiliation |
() |
| 16th Author's Name |
|
| 16th Author's Affiliation |
() |
| 17th Author's Name |
|
| 17th Author's Affiliation |
() |
| 18th Author's Name |
|
| 18th Author's Affiliation |
() |
| 19th Author's Name |
|
| 19th Author's Affiliation |
() |
| 20th Author's Name |
|
| 20th Author's Affiliation |
() |
| 21st Author's Name |
|
| 21st Author's Affiliation |
() |
| 22nd Author's Name |
|
| 22nd Author's Affiliation |
() |
| 23rd Author's Name |
|
| 23rd Author's Affiliation |
() |
| 24th Author's Name |
|
| 24th Author's Affiliation |
() |
| 25th Author's Name |
|
| 25th Author's Affiliation |
() |
| 26th Author's Name |
/ / |
| 26th Author's Affiliation |
()
() |
| 27th Author's Name |
/ / |
| 27th Author's Affiliation |
()
() |
| 28th Author's Name |
/ / |
| 28th Author's Affiliation |
()
() |
| 29th Author's Name |
/ / |
| 29th Author's Affiliation |
()
() |
| 30th Author's Name |
/ / |
| 30th Author's Affiliation |
()
() |
| 31st Author's Name |
/ / |
| 31st Author's Affiliation |
()
() |
| 32nd Author's Name |
/ / |
| 32nd Author's Affiliation |
()
() |
| 33rd Author's Name |
/ / |
| 33rd Author's Affiliation |
()
() |
| 34th Author's Name |
/ / |
| 34th Author's Affiliation |
()
() |
| 35th Author's Name |
/ / |
| 35th Author's Affiliation |
()
() |
| 36th Author's Name |
/ / |
| 36th Author's Affiliation |
()
() |
| Speaker |
Author-1 |
| Date Time |
2011-07-21 13:30:00 |
| Presentation Time |
25 minutes |
| Registration for |
SP |
| Paper # |
SP2011-41 |
| Volume (vol) |
vol.111 |
| Number (no) |
no.153 |
| Page |
pp.1-6 |
| #Pages |
6 |
| Date of Issue |
2011-07-14 (SP) |