Opening book details…
Can I read Neural Text-to-Speech Synthesis on EtoBox?
Neural Text-to-Speech Synthesis by Xu Tan is a nonfiction available to read on EtoBox.
What is Neural Text-to-Speech Synthesis about?
Text-to-speech (TTS) synthesis is an Artificial Intelligence (AI) technique that renders a preferably naturally sounding speech given an arbitrary text. It is a key technological component in many important applications, including virtual assistants, AI-generated audiobooks, speech-to-speech translation, AI news reporters, audible driving guidance, and digital humans. In the past decade, we have observed significant progress made in TTS. These new developments are mainly attributed to Deep Learning techniques and are usually referred to as neural TTS. Many neural TTS systems have achieved human quality for the tasks they are designed for. This book first introduces the history of TTS technologies and overviews neural TTS, and provides preliminary knowledge on language and speech processing, neural networks and Deep Learning, and deep generative models. It then introduces neural TTS from the perspective of key components (text analyses, acoustic models, vocoders, and end-to-end models) and advanced topics (expressive and controllable, robust, model-efficient, and data-efficient TTS). It also points some future research directions and collects some resources related to TTS. Although
Who reads Neural Text-to-Speech Synthesis?
It is typically read by self-directed learners exploring a subject in depth.
Common subject areas: history, science, philosophy, social sciences.
- Author
- Xu Tan
- Publisher
- Springer Nature Singapore Pte Ltd Fka Springer Science + Business Media Singapore Pte Ltd
- Published
- 2023
- Language
- EN
- ISBN
- 9789819908271
- Category
- nonfiction
- Subjects
- Science, Computer Science, Mathematics
More by Xu Tan
Similar books
- An Introduction to Text-to-Speech Synthesis (Text, Speech and Language Technology, 3) — Thierry Dutoit (1997)
- Trainable Text-to-speech Synthesis for European Portuguese — Maria João Almeida de Sá Barros Weiss (2011)
- Transfer Learning from Speaker Verif i cation to Multispeaker Text-To-Speech Synthesis — Ye Jia, Yu Zhang, Ron J. Weiss, Quan Wang, Jonathan Shen,Fei Ren ,Zhifeng Chen, Patrick Nguyen, Ruoming Pang, Ignacio Lopez Moreno, Yonghui Wu (2017)
- Improvements in Speech Synthesis : COST 258: the Naturalness of Synthetic Speech — Keller E., Bailly G., Monaghan A., Terken J., Huckvale M. (2002)
- From Text to Speech: The MI Talk system (Cambridge Studies in Speech Science and Communication) — Jonathan Allen, M. Sharon Hunnicutt, Dennis Klatt, with Robert C. Armstrong and David Pisoni (1987)
- Audiobook Synthesis with Long-form Neural Text-to-speech — Weicheng Zhang, Cheng-Chieh Yeh, Will Beckman, Tuomo Raitio, Ramya Rasipuram, Ladan Golipour, David Winarsky (2023)