検索に戻る
案件記録

HIERARCHICAL AUTO EVALUATION OF GENERATIVE AI SYSTEMS

発明審査中
3閲覧数
20請求項 · 2 独立
§ Ⅰ

案件概要

発明者

Jineet Hiren DOSHI; Maya Vered LIVSHITS; Na XU; Yuan ZHOU; Jeyendran BALAKRISHNAN

IPC分類

G6N 3/91

CPC分類

G6N3/91

An auto evaluation system for evaluating large language models (LLMs). The auto evaluation system loads a base auto evaluation class with core functionalities, selects one or more metrics for evaluation, extends the base auto evaluation class to create a child class with additional functionalities tailored to the selected metrics. A judge LLM receives the evaluation prompts from the auto evaluation server for response generation and computes evaluation scores for the test LLM.

原文(中国語)

An auto evaluation system for evaluating large language models (LLMs). The auto evaluation system loads a base auto evaluation class with core functionalities, selects one or more metrics for evaluation, extends the base auto evaluation class to create a child class with additional functionalities tailored to the selected metrics. A judge LLM receives the evaluation prompts from the auto evaluation server for response generation and computes evaluation scores for the test LLM.

外部リソース