Device text to speech
Abstract:
Systems and processes for generating speech from text are provided. An example method of generating speech from text includes, at an electronic device having at least one processor and memory, obtaining text; generating a plurality of segments of a spectrogram using a first neural network, each spectrogram segment of the plurality of spectrogram segments representing a portion of the text; generating, based on the plurality of spectrogram segments, a plurality of speech segments using a second neural network; and providing the plurality of speech segments as a speech output.
Public/Granted literature
Information query
Patent Agency Ranking
0/0