検索に戻る
案件記録

Method and System for Tagging, Cataloging, and Retrieving Speaker Identities Using Artificial Intelligence on Time-Synchronized Content

発明審査中
3閲覧数
20請求項 · 3 独立
§ Ⅰ

案件概要

発明者

Francisco BONZI; Stella TAVELLA; Loreto PARISI; Luca TORELLI

IPC分類

G10L 17/18G10L 17/6

CPC分類

G10L17/18G10L17/6

In one embodiment, a computer-implemented method includes receiving, at one or more processing devices, an audio file, tagging, using an artificial intelligence engine, one or more portions of the audio file to generate a modified audio file, wherein the tagging is performed based on the one or more portions corresponding to an audio-fingerprint of a voice stored in a database, performing dynamic cluster adaptation on the modified audio file, and causing the modified audio file to be played via a computing device.

原文(中国語)

In one embodiment, a computer-implemented method includes receiving, at one or more processing devices, an audio file, tagging, using an artificial intelligence engine, one or more portions of the audio file to generate a modified audio file, wherein the tagging is performed based on the one or more portions corresponding to an audio-fingerprint of a voice stored in a database, performing dynamic cluster adaptation on the modified audio file, and causing the modified audio file to be played via a computing device.

外部リソース