Reinforcement learning with scheduled auxiliary control

Invention Grant

US11893480B1 Reinforcement learning with scheduled auxiliary control 有权

Please log in to see more content

Patent Title: Reinforcement learning with scheduled auxiliary control
Application No.: US16289531

Application Date: 2019-02-28
Publication No.: US11893480B1

Publication Date: 2024-02-06
Inventor: Martin Riedmiller , Roland Hafner
Applicant: DeepMind Technologies Limited
Applicant Address: GB London
Assignee: DeepMind Technologies Limited
Current Assignee: DeepMind Technologies Limited
Current Assignee Address: GB London
Agency: Fish & Richardson P.C.
Main IPC: G06N3/08
IPC: G06N3/08 ; G06N3/04 ; G06N7/01

Reinforcement learning with scheduled auxiliary control

Abstract:

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for reinforcement learning with scheduled auxiliary tasks. In one aspect, a method includes maintaining data specifying parameter values for a primary policy neural network and one or more auxiliary neural networks; at each of a plurality of selection time steps during a training episode comprising a plurality of time steps: receiving an observation, selecting a current task for the selection time step using a task scheduling policy, processing an input comprising the observation using the policy neural network corresponding to the selected current task to select an action to be performed by the agent in response to the observation, and causing the agent to perform the selected action.

Information query

Espacenet

IPC分类:

G	物理
G06	计算；推算或计数
G06N	基于特定计算模型的计算机系统
G06N3/00	基于生物学模型的计算机系统
G06N3/02	.采用神经网络模型
G06N3/08	..学习方法