METHOD AND APPARATUS FOR TASK-DRIVEN SPEECH SEPARATION BY LEVERAGING SPEAKER DISTANCE INFORMATION
案件概要
発明者
Hao Zhang; Meng Yu; Yong Xu; Dong Yu
IPC分類
CPC分類
A method includes receiving a mixture signal comprising at least a first speaker, a second speaker, and background noise, the first speaker having a first distance to a microphone that outputs the mixture signal, the second speaker having a second distance to the microphone; training one or more neural networks to output a target channel and an interference channel by: inputting, into the one or more neural networks, the mixture signal and a task ID associated with one of the first speaker and the second speaker as a target speaker; determining a loss function based on the first distance and the second distance; and updating the neural network based on the loss function.
原文(中国語)
A method includes receiving a mixture signal comprising at least a first speaker, a second speaker, and background noise, the first speaker having a first distance to a microphone that outputs the mixture signal, the second speaker having a second distance to the microphone; training one or more neural networks to output a target channel and an interference channel by: inputting, into the one or more neural networks, the mixture signal and a task ID associated with one of the first speaker and the second speaker as a target speaker; determining a loss function based on the first distance and the second distance; and updating the neural network based on the loss function.
外部リソース