CNIPA.AI
검색으로 돌아가기
기록

HIERARCHICAL AUTO EVALUATION OF GENERATIVE AI SYSTEMS

발명심사 중
2조회수
20청구항 · 2 독립항
§ Ⅰ

개요

발명자

Jineet Hiren DOSHI; Maya Vered LIVSHITS; Na XU; Yuan ZHOU; Jeyendran BALAKRISHNAN

IPC 분류

G6N 3/91

CPC 분류

G6N3/91

An auto evaluation system for evaluating large language models (LLMs). The auto evaluation system loads a base auto evaluation class with core functionalities, selects one or more metrics for evaluation, extends the base auto evaluation class to create a child class with additional functionalities tailored to the selected metrics. A judge LLM receives the evaluation prompts from the auto evaluation server for response generation and computes evaluation scores for the test LLM.

원문 (중국어)

An auto evaluation system for evaluating large language models (LLMs). The auto evaluation system loads a base auto evaluation class with core functionalities, selects one or more metrics for evaluation, extends the base auto evaluation class to create a child class with additional functionalities tailored to the selected metrics. A judge LLM receives the evaluation prompts from the auto evaluation server for response generation and computes evaluation scores for the test LLM.