IP Library Granted Patent US 11,232,130
Granted Patent B2
US 11,232,130 · App. 17/326,052 · Granted Jan 25, 2022

Push model for intermediate query results

Inventors: Thierry Cruanes (San Mateo, CA); Benoit Dageville (Foster City, CA); Allison Waingold Lee (San Mateo, CA)
Assignee: Snowflake Inc.
G06F16/27G06F9/4881G06F9/5016G06F9/5044G06F9/5083G06F9/5088G06F16/148G06F16/1827G06F16/211G06F16/221G06F16/2365G06F16/2456G06F16/2471G06F16/24532G06F16/24545G06F16/24552G06F16/951G06F16/9535H04L67/1095H04L67/1097H04L67/2842
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,232,130
App. No.
17/326,052
Granted
Jan 25, 2022
Kind
B2
Abstract

A system and method for managing data storage and data access with querying data in a distributed system without buffering the results on intermediate operations in disk storage.

Claims (42)

1. A method comprising:

processing, within a first processor, a set of data with a first parallel execution process of a query plan for a query to generate an intermediate result of the query plan; and

pushing, by the first processor, during execution of the query plan, the intermediate result to a plurality of second processors for processing by a plurality of secondary parallel execution processes within the plurality of second processors that generate a plurality of second results, wherein each of the first processors and the second set of processors is decoupled from a storage platform that stores the set of data prior to the execution of the query plan.

2. The method of claim 1 , further comprising:

storing a final result to disk storage in a storage platform.

3. The method of claim 1 , further comprising:

accessing the query plan for the set of data references by the query.

4. The method of claim 1 , wherein each of the plurality of second processors store at least some of the plurality of intermediate results in a local cache that corresponds to that second processor.

5. The method of claim 1 , wherein the plurality of second results is generated without writing the intermediate result to the storage platform.

6. The method of claim 1 , wherein each of the plurality of secondary parallel execution processes performs a different operation on the intermediate result.

7. The method of claim 1 , further comprising:

delaying operation of at least one of the plurality of secondary parallel execution processes so as to coordinate timing among other secondary parallel execution processes of the plurality of secondary parallel execution processes.

8. The method of claim 1 , wherein the intermediate result comprises a plurality of rows of database data.

9. The method of claim 1 , wherein each of the plurality of secondary parallel execution processes are unique processes.

10. The method of claim 1 , wherein the intermediate result is not materialized.

11. The method of claim 1 , wherein the intermediate result generated by the first parallel execution process is not materialized to a temporary structure.

12. The method of claim 1 , wherein the each of the first processor and second plurality of processors is coupled to disk storage via a communications network.

13. The method of claim 1 , wherein the instructions further cause the computing device to:

receiving the query for information stored in one or more databases.

14. A system comprising:

a first processor programmed to,

process a set of data with a first parallel execution process of a query plan for a query to generate an intermediate result of the query plan, and

push, during execution of the query plan, the intermediate result to a plurality of second processors for concurrent processing by the plurality of secondary parallel execution processes; and

the plurality of second processors programmed to generate the plurality of second results, wherein each of the first processors and the second set of processors is decoupled from a storage platform that stores the set of data prior to the execution of the query plan.

15. The system of claim 14 , wherein the first processor further programmed to:

access the query plan for the set of data references by the query.

16. The system of claim 14 , wherein each of the plurality of second processors store at least some of the plurality of intermediate results in a local cache that corresponds to that second processor.

17. The system of claim 14 , wherein the plurality of second results is generated without writing the intermediate result to the storage platform.

18. The system of claim 14 , wherein each of the plurality of secondary parallel execution processes performs a different operation on the intermediate result.

19. The system of claim 14 , wherein the first processor further programmed to:

delay operation of at least one of the plurality of secondary parallel execution processes so as to coordinate timing among other secondary parallel execution processes of the plurality of secondary parallel execution processes.

20. The system of claim 14 , wherein the intermediate result comprises a plurality of rows of database data.

21. The system of claim 14 , wherein each of the plurality of secondary parallel execution processes are unique processes.

22. The system of claim 14 , wherein the first processor further programmed to:

receive the query for information stored in one or more databases.

23. A non-transitory computer-readable medium storing instructions which, when executed by one or more third processors of a computing device, cause the computing device to:

process, within a first processor, a set of data with a first parallel execution process of a query plan for a query to generate an intermediate result of the query plan; and

push, by the first processor, during execution of the query plan, the intermediate result to a plurality of second processors for processing by a plurality of secondary parallel execution processes within the plurality of second processors that generate a plurality of second results, wherein each of the first processors and the second set of processors is decoupled from a storage platform that stores the set of data prior to the execution of the query plan.

24. The non-transitory computer-readable medium of claim 23 , wherein the instructions further cause the computing device to:

store a final result to a disk storage within a storage platform.

25. The non-transitory computer-readable medium of claim 23 , wherein the instructions further cause the computing device to:

access the query plan for the set of data references by the query.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 20, 2021
From: CRUANES, THIERRY; DAGEVILLE, BENOIT; LEE, ALLISON WAINGOLD
To: SNOWFLAKE COMPUTING INC.
Reel/Frame 056306/0658 →
CHANGE OF NAME Recorded May 20, 2021
From: SNOWFLAKE COMPUTING INC.
To: SNOWFLAKE INC.
Reel/Frame 056322/0305 →
Continuity (8)
Continuation 17194182 · Mar 5, 2021
Continuation 17125524 · Dec 17, 2020
Continuation 16995599 · Aug 17, 2020
Continuation 16911185 · Jun 24, 2020
Continuation 16741676 · Jan 13, 2020
Continuation 14626853 · Feb 19, 2015
Provisional Application 61941986 · Feb 19, 2014
Related Publication 20210271690A1 · Sep 2, 2021