検索に戻る
案件記録

METHOD AND APPARATUS FOR TASK-DRIVEN SPEECH SEPARATION BY LEVERAGING SPEAKER DISTANCE INFORMATION

発明審査中
2閲覧数
20請求項 · 3 独立
§ Ⅰ

案件概要

発明者

Hao Zhang; Meng Yu; Yong Xu; Dong Yu

IPC分類

G10L 21/272G10L 25/30

CPC分類

G10L21/272G10L25/30

A method includes receiving a mixture signal comprising at least a first speaker, a second speaker, and background noise, the first speaker having a first distance to a microphone that outputs the mixture signal, the second speaker having a second distance to the microphone; training one or more neural networks to output a target channel and an interference channel by: inputting, into the one or more neural networks, the mixture signal and a task ID associated with one of the first speaker and the second speaker as a target speaker; determining a loss function based on the first distance and the second distance; and updating the neural network based on the loss function.

原文(中国語)

A method includes receiving a mixture signal comprising at least a first speaker, a second speaker, and background noise, the first speaker having a first distance to a microphone that outputs the mixture signal, the second speaker having a second distance to the microphone; training one or more neural networks to output a target channel and an interference channel by: inputting, into the one or more neural networks, the mixture signal and a task ID associated with one of the first speaker and the second speaker as a target speaker; determining a loss function based on the first distance and the second distance; and updating the neural network based on the loss function.

外部リソース