CNIPA.AI
검색으로 돌아가기
기록

METHOD AND APPARATUS FOR TASK-DRIVEN SPEECH SEPARATION BY LEVERAGING SPEAKER DISTANCE INFORMATION

발명심사 중
20청구항 · 3 독립항
§ Ⅰ

개요

발명자

Hao Zhang; Meng Yu; Yong Xu; Dong Yu

IPC 분류

G10L 21/272G10L 25/30

CPC 분류

G10L21/272G10L25/30

A method includes receiving a mixture signal comprising at least a first speaker, a second speaker, and background noise, the first speaker having a first distance to a microphone that outputs the mixture signal, the second speaker having a second distance to the microphone; training one or more neural networks to output a target channel and an interference channel by: inputting, into the one or more neural networks, the mixture signal and a task ID associated with one of the first speaker and the second speaker as a target speaker; determining a loss function based on the first distance and the second distance; and updating the neural network based on the loss function.

원문 (중국어)

A method includes receiving a mixture signal comprising at least a first speaker, a second speaker, and background noise, the first speaker having a first distance to a microphone that outputs the mixture signal, the second speaker having a second distance to the microphone; training one or more neural networks to output a target channel and an interference channel by: inputting, into the one or more neural networks, the mixture signal and a task ID associated with one of the first speaker and the second speaker as a target speaker; determining a loss function based on the first distance and the second distance; and updating the neural network based on the loss function.