REAL TIME MUSIC GENERATION FROM DIRECTED INPUT
개요
발명자
Philip Woods; Ali Thabet; Artsiom Sanakoyeu; David Kant
IPC 분류
CPC 분류
Systems and methods to generate audio content are provided. The systems and methods include converting, at a communication device, user input to a text encoding. The systems and methods also include generating, by a first machine learning model associated with the communication device, at least one token representing acoustic information based on the text encoding. A first token of the at least one token may represent at least one audio feature. The systems and methods further include generating at least one audio vector based on the at least one token and the text encoding. The systems and methods further include transforming the at least one audio vector to an audio waveform including at least one segment of audio content associated with the at least one audio feature.
원문 (중국어)
Systems and methods to generate audio content are provided. The systems and methods include converting, at a communication device, user input to a text encoding. The systems and methods also include generating, by a first machine learning model associated with the communication device, at least one token representing acoustic information based on the text encoding. A first token of the at least one token may represent at least one audio feature. The systems and methods further include generating at least one audio vector based on the at least one token and the text encoding. The systems and methods further include transforming the at least one audio vector to an audio waveform including at least one segment of audio content associated with the at least one audio feature.