検索に戻る
案件記録

METHOD FOR GENERATING AUDIO BASED ON LARGE MODEL, ELECTRONIC DEVICE, AND STORAGE MEDIUM

発明審査中
20請求項 · 3 独立
§ Ⅰ

案件概要

発明者

Liqiang ZHANG; Zhe PENG; Xuexin XU; Tao SUN; Lei JIA

IPC分類

G10L 13/8G6N 3/455G10L 19/

CPC分類

G10L13/8G6N3/455G10L19/

The present application provides a method for generating audio based on large model, an electronic device, and a storage medium, which relates to a technical field of artificial intelligence such as an audio synthesis and a large model. A specific implementation includes: obtaining a character that is generated in real time during a process of generating a text using a large model; obtaining an audio feature of each audio unit of the character sequentially by using a pre-trained audio generation model based on the character; the audio feature of the audio unit is a discretized audio feature, and the character includes audio features of a plurality of different audio units; synthesizing a corresponding audio by using a pre-trained vocoder based on the audio feature of each audio unit.

原文(中国語)

The present application provides a method for generating audio based on large model, an electronic device, and a storage medium, which relates to a technical field of artificial intelligence such as an audio synthesis and a large model. A specific implementation includes: obtaining a character that is generated in real time during a process of generating a text using a large model; obtaining an audio feature of each audio unit of the character sequentially by using a pre-trained audio generation model based on the character; the audio feature of the audio unit is a discretized audio feature, and the character includes audio features of a plurality of different audio units; synthesizing a corresponding audio by using a pre-trained vocoder based on the audio feature of each audio unit.

外部リソース