Speech Prosody in Speech Synthesis: Modeling and Generation of Prosody for High Quality and Flexible Speech Synthesis

Speech Prosody in Speech Synthesis: Modeling and Generation of Prosody for High Quality and Flexible Speech Synthesis

  • 定價:7199

分期價:(除不盡餘數於第一期收取) 分期說明

3期0利率每期23996期0利率每期1199
  • 運送方式:
  • 臺灣與離島
  • 海外
  • 可配送點:台灣、蘭嶼、綠島、澎湖、金門、馬祖
  • 可配送點:台灣、蘭嶼、綠島、澎湖、金門、馬祖
載入中...
  • 分享
 

內容簡介

The volume addresses issues concerning prosody generation in speech synthesis, including prosody modeling, how we can convey para- and non-linguistic information in speech synthesis, and prosody control in speech synthesis (including prosody conversions). A high level of quality has already been achieved in speech synthesis by using selection-based methods with segments of human speech. Although the method enables synthetic speech with various voice qualities and speaking styles, it requires large speech corpora with targeted quality and style.

Accordingly, speech conversion techniques are now of growing interest among researchers. HMM/GMM-based methods are widely used, but entail several major problems when viewed from the prosody perspective; prosodic features cover a wider time span than segmental features and their frame-by-frame processing is not always appropriate. The book offers a good overview of state-of-the-art studies on prosody in speech synthesis.

 

作者簡介

Professor Keikichi Hirose received the B. E. degree in electrical engineering in 1972, and the M. E. and Ph. D. degrees in electronic engineering respectively in 1974 and 1977 from the University of Tokyo. From 1977, he is a faculty member at the University of Tokyo, and was a Professor of the Department of Electronic Engineering from 1994. Currently he is professor at the Department of Information and Communication Engineering, Graduate School of Information Science and Technology, University of Tokyo. From March 1987 to January 1988, he was Visiting Scientist at the Research Laboratory of Electronics, Massachusetts Institute of Technology, Cambridge, U.S.A. He has been engaged in a wide range of research on spoken language processing, including analysis, synthesis, recognition, dialogue systems, and computer-assisted language learning. From 2000 to 2004, he was Principal Investigator of the national project "Realization of advanced spoken language information processing utilizing prosodic features," supported by the Japanese Government. He served as Chair of Speech Committee, Institute of Electronics, Information and Communication Engineers (IEICE)/Acoustical Society of Japan (ASJ) from 2003 to 2005. He is Chair of Speech Prosody Special Interest Group (SPro-SIG), ISCA, from October 2010. He has been on the editorial board of Speech Communication journal since 2004 and on the editorial board of ETRI Journal since 2009. He is a Fellow of Institute of Information and Communication Engineering and a member of a number of academic societies, including IEEE, International Speech Communication Association (Board member), Acoustical Society of America, Acoustical Society of Japan, Information Processing Society of Japan, Japanese Society for Artificial Intelligence, and Research Institute of Signal Processing Japan (Board member).

Jianhua Tao received the M.S. degree from Nanjing University in 1996 and the Ph.D. in Computer Science from Tsinghua University in 2001. He is currently the professor at National Laboratory of Pattern Recognition (NLPR) of Chinese Academy of Sciences where he chairs the human computer speech interaction group. He developed quite several earliest versions of Speech systems, multimodal interaction system in China, and published more than 90 papers in IEEE Trans. on ASLP, ICASSP, Interspeech, ICME, ICPR, ICCV, ICIP, etc. He has been the main researcher and contributor of several national scientific projects supported by National Natural Science Foundation of China (NSFC), National High-Tech Program and International Cooperation Projects (863). Currently, He is one of the editorial board members of "International Journal on Computational Linguistics and Chinese Language Processing", "Journal on Multimodal User Interfaces (JMUI)", "International Journal of Synthetic Emotions (IJSE)", and the Steering Committee Member for the IEEE Transactions on Affective Computing. He was elected as vice-chair of ISCA Special Interesting Group of Chinese Spoken Language Processing from 2006, the executive committee member of HUMAINE association from 2007, the board member of COCOSDA from 2007, and is also the Council member of Chinese Speech Information Processing Society and the Acoustical Society of China.

 

詳細資料

  • ISBN:9783662452578
  • 規格:精裝 / 213頁 / 23.4 x 15.5 x 1.5 cm / 普通級
  • 出版地:美國

最近瀏覽商品

 

相關活動

  • 【語言學習】職場高效成長術:越級打怪技能get、職場能力level up!電子書6折起
 

購物說明

外文館商品版本:商品之書封,為出版社提供之樣本。實際出貨商品,以出版社所提供之現有版本為主。關於外文書裝訂、版本上的差異,請參考【外文書的小知識】。

調貨時間:無庫存之商品,在您完成訂單程序之後,將以空運的方式為您下單調貨。原則上約14~20個工作天可以取書(若有將延遲另行告知)。為了縮短等待的時間,建議您將外文書與其它商品分開下單,以獲得最快的取貨速度,但若是海外專案進口的外文商品,調貨時間約1~2個月。 

若您具有法人身份為常態性且大量購書者,或有特殊作業需求,建議您可洽詢「企業採購」。 

退換貨說明 

會員所購買的商品均享有到貨十天的猶豫期(含例假日)。退回之商品必須於猶豫期內寄回。 

辦理退換貨時,商品必須是全新狀態與完整包裝(請注意保持商品本體、配件、贈品、保證書、原廠包裝及所有附隨文件或資料的完整性,切勿缺漏任何配件或損毀原廠外盒)。退回商品無法回復原狀者,恐將影響退貨權益或需負擔部分費用。 

訂購本商品前請務必詳閱商品退換貨原則 

  • 國際書展
  • 爸媽英文分級班
  • 2024