検索に戻る
案件記録

TEXT-TO-SPEECH TRANSDUCER

発明審査中
1閲覧数
20請求項 · 3 独立
§ Ⅰ

案件概要

発明者

Vladimir Bataev; Subhankar Ghosh; Vitaly Lavrukhin; Boris Ginsburg

IPC分類

G10L 13/27G10L 13/47G10L 13/8

CPC分類

G10L13/27G10L13/47G10L13/8

Disclosed are apparatuses, systems, and techniques that use a text-to-speech (TTS) transducer to perform TTS operations. The techniques include generating an initial input for a second model using an output of a first model. The techniques include generating, using the second model and the initial input, a first set of audio codes. The techniques include iteratively generating subsequent sets of audio codes using, at each iteration, the second model and a respective subsequent input for the second model. The respective subsequent input can reflect at least one previous set of audio codes generated by the second model.

原文(中国語)

Disclosed are apparatuses, systems, and techniques that use a text-to-speech (TTS) transducer to perform TTS operations. The techniques include generating an initial input for a second model using an output of a first model. The techniques include generating, using the second model and the initial input, a first set of audio codes. The techniques include iteratively generating subsequent sets of audio codes using, at each iteration, the second model and a respective subsequent input for the second model. The respective subsequent input can reflect at least one previous set of audio codes generated by the second model.

外部リソース