Invention Grant
- Patent Title: Learning device, simulation system, learning method, and storage medium
-
Application No.: US16553309Application Date: 2019-08-28
-
Publication No.: US11544556B2Publication Date: 2023-01-03
- Inventor: Takeru Goto
- Applicant: HONDA MOTOR CO., LTD.
- Applicant Address: JP Tokyo
- Assignee: HONDA MOTOR CO., LTD.
- Current Assignee: HONDA MOTOR CO., LTD.
- Current Assignee Address: JP Tokyo
- Agency: Amin, Turocy & Watson, LLP
- Priority: JPJP2018-161908 20180830
- Main IPC: G06N3/08
- IPC: G06N3/08 ; G05B13/02 ; G05D1/02 ; G05D1/00

Abstract:
A learning device includes a plurality of individual learners. Each of the individual learners includes a planner configured to generate information for defining an operation of the operation subject corresponding to itself, and a reward deriver configured to derive a reward obtained by evaluating information to be evaluated including feedback information obtained from a simulator by inputting information based on the information for defining the operation of the operation subject to the simulator. The planner performs reinforcement learning based on the reward derived by the reward deriver, and at least two of the plurality of individual learners are different in the operations of the operation subject in which the reward derived by the reward deriver is maximized.
Public/Granted literature
- US20200074302A1 LEARNING DEVICE, SIMULATION SYSTEM, LEARNING METHOD, AND STORAGE MEDIUM Public/Granted day:2020-03-05
Information query