METHOD FOR GENERATING AUDIO BASED ON LARGE MODEL, ELECTRONIC DEVICE, AND STORAGE MEDIUM
案件概要
出願人
BEIJING BAIDU NETCOM SCIENCE TECHNOLOGY CO., LTD.
発明者
Liqiang ZHANG; Zhe PENG; Xuexin XU; Tao SUN; Lei JIA
IPC分類
CPC分類
The present application provides a method for generating audio based on large model, an electronic device, and a storage medium, which relates to a technical field of artificial intelligence such as an audio synthesis and a large model. A specific implementation includes: obtaining a character that is generated in real time during a process of generating a text using a large model; obtaining an audio feature of each audio unit of the character sequentially by using a pre-trained audio generation model based on the character; the audio feature of the audio unit is a discretized audio feature, and the character includes audio features of a plurality of different audio units; synthesizing a corresponding audio by using a pre-trained vocoder based on the audio feature of each audio unit.
原文(中国語)
The present application provides a method for generating audio based on large model, an electronic device, and a storage medium, which relates to a technical field of artificial intelligence such as an audio synthesis and a large model. A specific implementation includes: obtaining a character that is generated in real time during a process of generating a text using a large model; obtaining an audio feature of each audio unit of the character sequentially by using a pre-trained audio generation model based on the character; the audio feature of the audio unit is a discretized audio feature, and the character includes audio features of a plurality of different audio units; synthesizing a corresponding audio by using a pre-trained vocoder based on the audio feature of each audio unit.
外部リソース