CNIPA.AI
返回搜索
档案

TECHNIQUES FOR ENHANCING SPEECH LANGUAGE MODELS USING DESCRIPTIVE SPEECH-TEXT ALIGNMENT

发明专利审中
3浏览
20权利要求 · 3 独立
§ Ⅰ

卷宗概要

发明人

Szu-Wei FU; Yu-Chiang WANG; Zhehuai CHEN; He HUANG; Boris GINSBURG

IPC 分类

G10L 15/6G10L 15/183

CPC 分类

G10L15/63G10L15/183

The disclosed method for generating a first depth map for responding to audio input includes processing the audio input using a trained encoder to generate a representation of the audio input, where the audio input includes speech; processing the representation of the audio input using a first trained adapter to generate one or more features; and processing the one or more features and text associated with the audio input using a trained language model to generate a response.

原文(中文)

The disclosed method for generating a first depth map for responding to audio input includes processing the audio input using a trained encoder to generate a representation of the audio input, where the audio input includes speech; processing the representation of the audio input using a first trained adapter to generate one or more features; and processing the one or more features and text associated with the audio input using a trained language model to generate a response.