Learning device, simulation system, learning method, and storage medium

Invention Grant

US11544556B2 Learning device, simulation system, learning method, and storage medium 有权

Please log in to see more content

Patent Title: Learning device, simulation system, learning method, and storage medium
Application No.: US16553309

Application Date: 2019-08-28
Publication No.: US11544556B2

Publication Date: 2023-01-03
Inventor: Takeru Goto
Applicant: HONDA MOTOR CO., LTD.
Applicant Address: JP Tokyo
Assignee: HONDA MOTOR CO., LTD.
Current Assignee: HONDA MOTOR CO., LTD.
Current Assignee Address: JP Tokyo
Agency: Amin, Turocy & Watson, LLP
Priority: JPJP2018-161908 20180830
Main IPC: G06N3/08
IPC: G06N3/08 ; G05B13/02 ; G05D1/02 ; G05D1/00

Learning device, simulation system, learning method, and storage medium

Abstract:

A learning device includes a plurality of individual learners. Each of the individual learners includes a planner configured to generate information for defining an operation of the operation subject corresponding to itself, and a reward deriver configured to derive a reward obtained by evaluating information to be evaluated including feedback information obtained from a simulator by inputting information based on the information for defining the operation of the operation subject to the simulator. The planner performs reinforcement learning based on the reward derived by the reward deriver, and at least two of the plurality of individual learners are different in the operations of the operation subject in which the reward derived by the reward deriver is maximized.

Public/Granted literature

US20200074302A1 LEARNING DEVICE, SIMULATION SYSTEM, LEARNING METHOD, AND STORAGE MEDIUM Public/Granted day:2020-03-05

Information query

Espacenet

IPC分类:

G	物理
G06	计算；推算或计数
G06N	基于特定计算模型的计算机系统
G06N3/00	基于生物学模型的计算机系统
G06N3/02	.采用神经网络模型
G06N3/08	..学习方法