CNIPA.AI
검색으로 돌아가기
기록

TEXT-TO-SPEECH TRANSDUCER

발명심사 중
6조회수
20청구항 · 3 독립항
§ Ⅰ

개요

발명자

Vladimir Bataev; Subhankar Ghosh; Vitaly Lavrukhin; Boris Ginsburg

IPC 분류

G10L 13/27G10L 13/47G10L 13/8

CPC 분류

G10L13/27G10L13/47G10L13/8

Disclosed are apparatuses, systems, and techniques that use a text-to-speech (TTS) transducer to perform TTS operations. The techniques include generating an initial input for a second model using an output of a first model. The techniques include generating, using the second model and the initial input, a first set of audio codes. The techniques include iteratively generating subsequent sets of audio codes using, at each iteration, the second model and a respective subsequent input for the second model. The respective subsequent input can reflect at least one previous set of audio codes generated by the second model.

원문 (중국어)

Disclosed are apparatuses, systems, and techniques that use a text-to-speech (TTS) transducer to perform TTS operations. The techniques include generating an initial input for a second model using an output of a first model. The techniques include generating, using the second model and the initial input, a first set of audio codes. The techniques include iteratively generating subsequent sets of audio codes using, at each iteration, the second model and a respective subsequent input for the second model. The respective subsequent input can reflect at least one previous set of audio codes generated by the second model.