Invention Grant
- Patent Title: Artificial intelligence-based text-to-speech system and method
-
Application No.: US16446833Application Date: 2019-06-20
-
Publication No.: US11244669B2Publication Date: 2022-02-08
- Inventor: Martin Reber , Vijeta Avijeet
- Applicant: Telepathy Labs, Inc.
- Applicant Address: US FL Tampa
- Assignee: Telepathy Labs, Inc.
- Current Assignee: Telepathy Labs, Inc.
- Current Assignee Address: US FL Tampa
- Agency: Holland & Knight LLP
- Agent Michael T. Abramson
- Main IPC: G10L25/30
- IPC: G10L25/30 ; G10L13/08 ; G06K9/62 ; G06N5/02 ; G06N3/02 ; G10L19/00 ; G10L13/04 ; G06N3/04 ; G06N3/08

Abstract:
A technique improves training and speech quality of a text-to-speech (TTS) system having an artificial intelligence, such as a neural network. The TTS system is organized as a front-end subsystem and a back-end subsystem. The front-end subsystem is configured to provide analysis and conversion of text into input vectors, each having at least a base frequency, f0, a phenome duration, and a phoneme sequence that is processed by a signal generation unit of the back-end subsystem. The signal generation unit includes the neural network interacting with a pre-existing knowledgebase of phenomes to generate audible speech from the input vectors. The technique applies an error signal from the neural network to correct imperfections of the pre-existing knowledgebase of phenomes to generate audible speech signals. A back-end training system is configured to train the signal generation unit by applying psychoacoustic principles to improve quality of the generated audible speech signals.
Public/Granted literature
- US20190304434A1 ARTIFICIAL INTELLIGENCE-BASED TEXT-TO-SPEECH SYSTEM AND METHOD Public/Granted day:2019-10-03
Information query