検索に戻る
案件記録

TASK GROUPING FOR REINFORCEMENT LEARNING WITH MULTIPLE TASKS

発明審査中
3閲覧数
20請求項 · 3 独立
§ Ⅰ

案件概要

発明者

Geoffrey Hikaru Harrison; Alexander Walter Cann; Ian Charles Colbert; Mehdi Saeedi

IPC分類

G6F 9/48G6N 3/92G6N 3/98

CPC分類

G6F9/4887G6N3/92G6N3/98

To generate reinforcement learning (RL) policies for the multiple tasks performable by a system, a computing device is configured to train an RL model for all tasks of a system to produce a general RL model. For each task, the computing device updates the parameters of the general RL model based on the task to produce a task-specific RL model. Based on comparisons of the general RL model to the task-specific RL models, the computing device determines inter-task similarity scores that represent the impact of a task on other tasks, the impact of other tasks on a task, or both. The computing device then groups the tasks of the system together based on the inter-task similarity scores and generates a task-grouped RL policy for each group of tasks.

原文(中国語)

To generate reinforcement learning (RL) policies for the multiple tasks performable by a system, a computing device is configured to train an RL model for all tasks of a system to produce a general RL model. For each task, the computing device updates the parameters of the general RL model based on the task to produce a task-specific RL model. Based on comparisons of the general RL model to the task-specific RL models, the computing device determines inter-task similarity scores that represent the impact of a task on other tasks, the impact of other tasks on a task, or both. The computing device then groups the tasks of the system together based on the inter-task similarity scores and generates a task-grouped RL policy for each group of tasks.

外部リソース