CNIPA.AI
返回搜索
档案

TASK GROUPING FOR REINFORCEMENT LEARNING WITH MULTIPLE TASKS

发明专利审中
6浏览
20权利要求 · 3 独立
§ Ⅰ

卷宗概要

发明人

Geoffrey Hikaru Harrison; Alexander Walter Cann; Ian Charles Colbert; Mehdi Saeedi

IPC 分类

G6F 9/48G6N 3/92G6N 3/98

CPC 分类

G6F9/4887G6N3/92G6N3/98

To generate reinforcement learning (RL) policies for the multiple tasks performable by a system, a computing device is configured to train an RL model for all tasks of a system to produce a general RL model. For each task, the computing device updates the parameters of the general RL model based on the task to produce a task-specific RL model. Based on comparisons of the general RL model to the task-specific RL models, the computing device determines inter-task similarity scores that represent the impact of a task on other tasks, the impact of other tasks on a task, or both. The computing device then groups the tasks of the system together based on the inter-task similarity scores and generates a task-grouped RL policy for each group of tasks.

原文(中文)

To generate reinforcement learning (RL) policies for the multiple tasks performable by a system, a computing device is configured to train an RL model for all tasks of a system to produce a general RL model. For each task, the computing device updates the parameters of the general RL model based on the task to produce a task-specific RL model. Based on comparisons of the general RL model to the task-specific RL models, the computing device determines inter-task similarity scores that represent the impact of a task on other tasks, the impact of other tasks on a task, or both. The computing device then groups the tasks of the system together based on the inter-task similarity scores and generates a task-grouped RL policy for each group of tasks.