IP Library Granted Patent US 12,650,821
Granted Patent B2
US 12,650,821 · App. 18/583,660 · Granted Jun 9, 2026

Machine learning repository service

Inventors: Vineet Khare (Redmond, WA); Alexander Johannes Smola (Sunnyvale, CA); Craig Wiley (Redmond, WA)
Assignee: Amazon Technologies, Inc.
G06F8/36H04L67/51
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,650,821
App. No.
18/583,660
Granted
Jun 9, 2026
Kind
B2
Abstract

Techniques for providing and servicing listed repository items such as algorithms, data, models, pipelines, and/or notebooks are described. In some examples, web services provider receives a request for a listed repository item from a requester, the request indicating at least a category of the repository item and each listing of a repository item includes an indication of a category that the listed repository item belongs to and a storage location of the listed repository item, determines a suggestion of at least one listed repository item based on the request, and provides the suggestion of the at least one listed repository item to the requester.

Claims (55)

1 . A computer-implemented method comprising:

publishing, by a web services provider, hosted machine learning models to a hosted machine learning repository, at least some of the hosted machine learning models including a name of the hosted machine learning model and an indication of a category to which the hosted machine learning model belongs;

receiving, by the web services provider from a requester via an application programming interface (API) frontend, a query specifying a requested category of machine learning models;

determining, by the web services provider, one or more of the hosted machine learning models that correspond to the requested category;

returning, by the web services provider to the requester via the API frontend, the one or more of the hosted machine learning models that correspond to the requested category;

receiving, by the web services provider from the requester via the API frontend, a request to use a selected machine learning model from among the one or more of the hosted machine learning models that correspond to the requested category;

adding, by the web services provider, the selected machine learning model as part of a pipeline;

allocating, by the web services provider, execution resources to the pipeline;

receiving, by the web services provider from the requester via the API frontend, a request to perform a task using the selected machine learning model; and

performing, by the web services provider in response to the request to perform a task, the task using the selected machine learning model and the execution resources.

2 . The computer-implemented method of claim 1 , the hosted machine learning repository hosted by the web services provider.

3 . The computer-implemented method of claim 1 , further comprising:

obtaining access rights information from a user account associated with the requester; and

determining access rights for the requester based on the access rights information.

4 . The computer-implemented method of claim 1 , the execution resources including compute resources and storage.

5 . The computer-implemented method of claim 1 , the task comprising training the selected machine learning model.

6 . The computer-implemented method of claim 1 , the task comprising performing inference using the selected machine learning model.

7 . The computer-implemented method of claim 1 , at least some of the hosted machine learning models further including an indication of an input format for the hosted machine learning model.

8 . A computer-implemented method comprising:

publishing, by a web services provider, hosted machine learning models to a hosted machine learning repository, at least some of the hosted machine learning models including a name of the hosted machine learning model and an indication of a category to which the hosted machine learning model belongs;

receiving, by the web services provider from a requester via an application programming interface (API) frontend, a query specifying a name of a requested machine learning model;

determining, by the web services provider, one or more of the hosted machine learning models that correspond to the name of the requested machine learning model;

returning, by the web services provider to the requester via the API frontend, the one or more of the hosted machine learning models that correspond to the name of the requested machine learning model;

receiving, by the web services provider from the requester via the API frontend, a request to use a selected machine learning model from among the one or more of the hosted machine learning models that correspond to the name of the requested machine learning model;

adding, by the web services provider, the selected machine learning model as part of a pipeline;

allocating, by the web services provider, execution resources to the pipeline;

receiving, by the web services provider from the requester via the API frontend, a request to perform a task using the selected machine learning model; and

performing, by the web services provider in response to the request to perform a task, the task using the selected machine learning model and the execution resources.

9 . The computer-implemented method of claim 8 , the hosted machine learning repository hosted by the web services provider.

10 . The computer-implemented method of claim 8 , further comprising:

obtaining access rights information from a user account associated with the requester; and

determining access rights for the requester based on the access rights information.

11 . The computer-implemented method of claim 8 , the execution resources including compute resources and storage.

12 . The computer-implemented method of claim 8 , the task comprising training the selected machine learning model.

13 . The computer-implemented method of claim 8 , the task comprising performing inference using the selected machine learning model.

14 . The computer-implemented method of claim 8 , at least some of the hosted machine learning models further including an indication of an input format for the hosted machine learning model.

15 . A system comprising:

a hosted machine learning repository; and

a web services provider including memory storing instructions that, when executed by at least one processor of the web services provider, cause the web services provider to:

publish hosted machine learning models to the hosted machine learning repository, at least some of the hosted machine learning models including a name of the hosted machine learning model and an indication of a category to which the hosted machine learning model belongs;

receive, from a requester via an application programming interface (API) frontend, a query specifying a requested category of machine learning models;

determine one or more of the hosted machine learning models that correspond to the requested category;

return, to the requester via the API frontend, the one or more of the hosted machine learning models that correspond to the requested category;

receive, from the requester via the API frontend, a request to use a selected machine learning model from among the one or more of the hosted machine learning models that correspond to the requested category;

add the selected machine learning model as part of a pipeline;

allocate execution resources to the pipeline;

receive, from the requester via the API frontend, a request to perform a task using the selected machine learning model; and

perform, in response to the request to perform a task, the task using the selected machine learning model and the execution resources.

16 . The system of claim 15 , the hosted machine learning repository hosted by the web services provider.

17 . The system of claim 15 , the memory storing further instructions that, when executed by the at least one processor of the web services provider, further cause the web services provider to:

obtain access rights information from a user account associated with the requester; and

determine access rights for the requester based on the access rights information.

18 . The system of claim 15 , the execution resources including compute resources and storage.

19 . The system of claim 15 , the task comprising training the selected machine learning model.

20 . The system of claim 15 , the task comprising performing inference using the selected machine learning model.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 3, 2026
From: KHARE, VINEET; SMOLA, ALEXANDER JOHANNES; WILEY, CRAIG
To: AMAZON TECHNOLOGIES, INC.
Reel/Frame 073953/0767 →
Continuity (1)
Related Publication 20250265050A1 · Aug 21, 2025
References Cited (27)
US 11501201B2 · Driscoll · 2022 [cited by examiner]
US 12518181B2 · Rao · 2026 [cited by examiner]
US 20030227487A1 · Hugh · 2003 [cited by applicant]
US 20080215456A1 · West et al. · 2008 [cited by applicant]
US 20140358825A1 · Phillipps · 2014 [cited by examiner]
US 20160019471A1 · Shin et al. · 2016 [cited by applicant]
US 20180107660A1 · Wang et al. · 2018 [cited by applicant]
US 20190019106A1 · Driscoll · 2019 [cited by examiner]
US 20190044829A1 · Balzer et al. · 2019 [cited by applicant]
US 20190114293A1 · Li et al. · 2019 [cited by applicant]
US 20250362963A1 · Poothiyot · 2025 [cited by examiner]
CN 104717098A · 2015 [cited by applicant]
CN 105378699A · 2016 [cited by applicant]
CN 107003977A · 2017 [cited by applicant]
KR 1020170023168A · 2017 [cited by applicant]
WO WO2025209924A1 · 2025 [cited by examiner]
International Preliminary Report on Patentability, PCT App. No. PCT/US2019/021645, Sep. 24, 2020, 8 pages. [cited by applicant]
International Search Report and Written Opinion, PCT App. No. PCT/US2019/021645, May 8, 2019, 9 pages. [cited by applicant]
Non-Final Office Action, U.S. Appl. No. 15/919, 178, May 30, 2019, 11 pages. [cited by applicant]
Non-Final Office Action, U.S. Appl. No. 17/572,470, Aug. 17, 2023, 10 pages. [cited by applicant]
Notice of Allowance, CN App. No. 202210535615.9, Jan. 18, 2023, 04 pages (02 pages of English Translation and 2 pages of Original Document). [cited by applicant]
Notice of Allowance, U.S. App. No. 15/919, 178, Oct. 24, 2019, 7 pages. [cited by applicant]
Notice of Allowance, U.S. Appl. No. 16/799,443, Oct. 15, 2021, 11 pages. [cited by applicant]
Notice of Allowance, U.S. Appl. No. 17/572,470, Nov. 14, 2023, 7 pages. [cited by applicant]
Notice on Grant of Patent Right for Invention, CN App. No. 201980016827.2, Mar. 9, 2022, 4 pages (2 pages of English Translation and 2 pages of Original Document). [cited by applicant]
Notice on the First Office Action, CN App. No. 202210535615.9, Oct. 10, 2022, 10 pages (5 pages of English Translation and 5 pages of Original Document). [cited by applicant]
Office Action, EP App. No. 19714902.4, Jan. 28, 2021, 7 pages. [cited by applicant]