| Paper Abstract and Keywords |
| Presentation |
2013-01-31 14:45
A Study on Style Control Based on Multiple-Regression HSMM for Synthesizing Singing Voices with Various Expressivity Takashi Nose, Misa Kanemoto, Tomoki Koriyama, Takao Kobayashi (Tokyo Inst. of Tech.) SP2012-111 |
| Abstract |
(in Japanese) |
(See Japanese page) |
| (in English) |
This paper proposes a style control technique based on multiple regression HSMM (MRHSMM)
for changing styles and their intensities appearing in synthetic singing voices.
In the proposed technique, styles and their intensities are represented by
low-dimensional vectors called style vectors and are modeled by
an assumption that mean parameters of acoustic models
are given as multiple regressions of the style vectors.
When synthesizing speech, we can weaken or emphasize the intensity of each style
by setting a desired style vector.
In addition, the idea of pitch adaptive training is introduced into the MRHSMM
to improve the modeling accuracy of F0 associated with musical notes.
The novel vibrato modeling technique is also presented to extract
vibrato parameters from singing voices that sometimes have unclear vibrato expressions.
Subjective evaluations show that we can intuitively contorol styles and their intensities
while keeping naturalness of synthetic speech. |
| Keyword |
(in Japanese) |
(See Japanese page) |
| (in English) |
HMM-based singing voice synthesis / HMM-based speech synthesis / style control / multiple-regression HSMM / pitch adaptive training / vibrato modeling / / |
| Reference Info. |
IEICE Tech. Rep., vol. 112, no. 422, SP2012-111, pp. 79-84, Jan. 2013. |
| Paper # |
SP2012-111 |
| Date of Issue |
2013-01-23 (SP) |
| ISSN |
Print edition: ISSN 0913-5685 Online edition: ISSN 2432-6380 |
Copyright and reproduction |
All rights are reserved and no part of this publication may be reproduced or transmitted in any form or by any means, electronic or mechanical, including photocopy, recording, or any information storage and retrieval system, without permission in writing from the publisher. Notwithstanding, instructors are permitted to photocopy isolated articles for noncommercial classroom use without fee. (License No.: 10GA0019/12GB0052/13GB0056/17GB0034/18GB0034) |
| Download PDF |
SP2012-111 |
| Conference Information |
| Committee |
SP |
| Conference Date |
2013-01-30 - 2013-01-31 |
| Place (in Japanese) |
(See Japanese page) |
| Place (in English) |
Doshisha Univ. |
| Topics (in Japanese) |
(See Japanese page) |
| Topics (in English) |
Speech, Language, and Dialogue, etc. |
| Paper Information |
| Registration To |
SP |
| Conference Code |
2013-01-SP |
| Language |
Japanese |
| Title (in Japanese) |
(See Japanese page) |
| Sub Title (in Japanese) |
(See Japanese page) |
| Title (in English) |
A Study on Style Control Based on Multiple-Regression HSMM for Synthesizing Singing Voices with Various Expressivity |
| Sub Title (in English) |
|
| Keyword(1) |
HMM-based singing voice synthesis |
| Keyword(2) |
HMM-based speech synthesis |
| Keyword(3) |
style control |
| Keyword(4) |
multiple-regression HSMM |
| Keyword(5) |
pitch adaptive training |
| Keyword(6) |
vibrato modeling |
| Keyword(7) |
|
| Keyword(8) |
|
| 1st Author's Name |
Takashi Nose |
| 1st Author's Affiliation |
Tokyo Institute of Technology (Tokyo Inst. of Tech.) |
| 2nd Author's Name |
Misa Kanemoto |
| 2nd Author's Affiliation |
Tokyo Institute of Technology (Tokyo Inst. of Tech.) |
| 3rd Author's Name |
Tomoki Koriyama |
| 3rd Author's Affiliation |
Tokyo Institute of Technology (Tokyo Inst. of Tech.) |
| 4th Author's Name |
Takao Kobayashi |
| 4th Author's Affiliation |
Tokyo Institute of Technology (Tokyo Inst. of Tech.) |
| 5th Author's Name |
|
| 5th Author's Affiliation |
() |
| 6th Author's Name |
|
| 6th Author's Affiliation |
() |
| 7th Author's Name |
|
| 7th Author's Affiliation |
() |
| 8th Author's Name |
|
| 8th Author's Affiliation |
() |
| 9th Author's Name |
|
| 9th Author's Affiliation |
() |
| 10th Author's Name |
|
| 10th Author's Affiliation |
() |
| 11th Author's Name |
|
| 11th Author's Affiliation |
() |
| 12th Author's Name |
|
| 12th Author's Affiliation |
() |
| 13th Author's Name |
|
| 13th Author's Affiliation |
() |
| 14th Author's Name |
|
| 14th Author's Affiliation |
() |
| 15th Author's Name |
|
| 15th Author's Affiliation |
() |
| 16th Author's Name |
|
| 16th Author's Affiliation |
() |
| 17th Author's Name |
|
| 17th Author's Affiliation |
() |
| 18th Author's Name |
|
| 18th Author's Affiliation |
() |
| 19th Author's Name |
|
| 19th Author's Affiliation |
() |
| 20th Author's Name |
|
| 20th Author's Affiliation |
() |
| 21st Author's Name |
|
| 21st Author's Affiliation |
() |
| 22nd Author's Name |
|
| 22nd Author's Affiliation |
() |
| 23rd Author's Name |
|
| 23rd Author's Affiliation |
() |
| 24th Author's Name |
|
| 24th Author's Affiliation |
() |
| 25th Author's Name |
|
| 25th Author's Affiliation |
() |
| 26th Author's Name |
/ / |
| 26th Author's Affiliation |
()
() |
| 27th Author's Name |
/ / |
| 27th Author's Affiliation |
()
() |
| 28th Author's Name |
/ / |
| 28th Author's Affiliation |
()
() |
| 29th Author's Name |
/ / |
| 29th Author's Affiliation |
()
() |
| 30th Author's Name |
/ / |
| 30th Author's Affiliation |
()
() |
| 31st Author's Name |
/ / |
| 31st Author's Affiliation |
()
() |
| 32nd Author's Name |
/ / |
| 32nd Author's Affiliation |
()
() |
| 33rd Author's Name |
/ / |
| 33rd Author's Affiliation |
()
() |
| 34th Author's Name |
/ / |
| 34th Author's Affiliation |
()
() |
| 35th Author's Name |
/ / |
| 35th Author's Affiliation |
()
() |
| 36th Author's Name |
/ / |
| 36th Author's Affiliation |
()
() |
| Speaker |
Author-1 |
| Date Time |
2013-01-31 14:45:00 |
| Presentation Time |
30 minutes |
| Registration for |
SP |
| Paper # |
SP2012-111 |
| Volume (vol) |
vol.112 |
| Number (no) |
no.422 |
| Page |
pp.79-84 |
| #Pages |
6 |
| Date of Issue |
2013-01-23 (SP) |