検索に戻る
案件記録

TECHNIQUES FOR ENHANCING SPEECH LANGUAGE MODELS USING DESCRIPTIVE SPEECH-TEXT ALIGNMENT

発明審査中
1閲覧数
20請求項 · 3 独立
§ Ⅰ

案件概要

発明者

Szu-Wei FU; Yu-Chiang WANG; Zhehuai CHEN; He HUANG; Boris GINSBURG

IPC分類

G10L 15/183G10L 15/2G10L 15/22

CPC分類

G10L15/183G10L15/2G10L15/22

The disclosed method for generating a first depth map for responding to audio input includes processing the audio input using a trained encoder to generate a representation of the audio input, where the audio input includes speech; processing the representation of the audio input using a first trained adapter to generate one or more features; and processing the one or more features and text associated with the audio input using a trained language model to generate a response.

原文(中国語)

The disclosed method for generating a first depth map for responding to audio input includes processing the audio input using a trained encoder to generate a representation of the audio input, where the audio input includes speech; processing the representation of the audio input using a first trained adapter to generate one or more features; and processing the one or more features and text associated with the audio input using a trained language model to generate a response.

外部リソース