Invention Grant
- Patent Title: Pre-warming scheme to load machine learning models
-
Application No.: US16146295Application Date: 2018-09-28
-
Publication No.: US11562288B2Publication Date: 2023-01-24
- Inventor: Enrico Sartorello , Stefano Stefani , Nikhil Kandoi , Rama Krishna Sandeep Pokkunuri , Kalpesh N. Sutaria , Navneet Sabbineni , Ganesh Kumar Gella , Cheng Ran Li
- Applicant: Amazon Technologies, Inc.
- Applicant Address: US WA Seattle
- Assignee: Amazon Technologies, Inc.
- Current Assignee: Amazon Technologies, Inc.
- Current Assignee Address: US WA Seattle
- Agency: Nicholson De Vos Webster & Elliott LLP
- Main IPC: G06N20/00
- IPC: G06N20/00 ; G06N5/04

Abstract:
Techniques for hosting adding and warming a host are described. In some instances, a method of determining that at least one group of hosts is to be increased by adding an additional host to the group of hosts; sending a request to the group of hosts for a list of machine learning models loaded per host of the group of hosts; receiving, from each host, the list of loaded machine learning models; loading at least a proper subset of list of loaded machine learning models into random access memory of the at least one group; receiving a request to perform an inference; routing the request to the additional host of the group of hosts; performing an inference using the additional host of the group of hosts; and providing a result of the inference to an external entity is described.
Information query