IP Library Granted Patent US 12,393,608
Granted Patent B1
US 12,393,608 · App. 18/673,931 · Granted Aug 19, 2025

Metastore manager for multi-platform data operations

Inventor: Jiyue Ma (Hangzhou, CN)
Assignee: Zoom Communications, Inc.
G06F16/278G06F16/2433
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,393,608
App. No.
18/673,931
Granted
Aug 19, 2025
Kind
B1
Abstract

A metastore manager facilitates access to data/metadata stored on at least one data management platform. A first data processing operation is initiated, in association with a data set maintained in a first data store associated with a first data management platform, to generate a first processed data set. The first processed data set is stored in the first data store and a first metadata set corresponding to the first processed data set is stored in a first metastore associated with the first data store. The first metadata set includes partition metadata indicative of partitioning information associated with the first processed data set. A synchronization operation is implemented to store, in a second metastore associated with a second data management platform, a second metadata set corresponding to a subset of the first metadata set. The metastore manager may facilitate access, by a data processing pipeline, to the processed data set.

Claims (42)

1. A method, comprising:

communicating a request, from a driver node to a metastore manager, for initiating, in association with a data set maintained in a first data store associated with a first data management platform, a first data processing operation, to be performed by the first data management platform, to generate a first processed data set, wherein the first processed data set is stored in the first data store, and wherein a first metadata set corresponding to the first processed data set is stored in a first metastore associated with the first data store, the first metadata set comprising partition metadata indicative of partitioning information associated with the first processed data set, wherein the metastore manager is configured to:

obtain a directed acyclic graph (DAG) corresponding to the first data processing operation; and

configure the first data processing operation in association with the DAG, wherein the DAG includes an unresolved logical plan, a logical plan, and an optimized logical plan;

initiating, by the metastore manager, a synchronization operation to store, in a second metastore associated with a second data management platform, a second metadata set corresponding to a subset of the first metadata set; and

activating, by the metastore manager, a data processing pipeline that is not communicatively coupled with the first data management platform wherein the data processing pipeline includes an artificial intelligence (AI) pipeline, to cause the data processing pipeline to obtain the partition metadata using a refresh component, and wherein the data processing pipeline is configured to:

obtain, based on the partition metadata, the first processed data set;

perform a second data processing operation on the first processed data set to generate a second processed data set, wherein the second processed data set comprises an output of an AI operation; and

output the second processed data set to a unified communications as a service (UCaaS) platform.

2. The method of claim 1 , wherein the data processing pipeline is communicatively coupled, via a platform interface, with the second data management platform.

3. The method of claim 1 , wherein the data processing pipeline is communicatively coupled, via a platform interface, with only the second data management platform.

4. The method of claim 1 , wherein the first data processing operation comprises a structured query language (SQL) query on a distributed database associated with at least one of the first data store or a second data store associated with the second metastore.

5. The method of claim 1 , further comprising configuring the first data processing operation at least in part by allocating a set of computational resources for the first data processing operation.

6. A non-transitory computer-readable medium storing instructions operable to cause one or more processors to perform operations comprising:

communicating a request, from a driver node to a metastore manager, for initiating, in association with a data set maintained in a first data store associated with a first data management platform, a first data processing operation, to be performed by the first data management platform, to generate a first processed data set, wherein the first processed data set is stored in the first data store, and wherein a first metadata set corresponding to the first processed data set is stored in a first metastore associated with the first data store, the first metadata set comprising partition metadata indicative of partitioning information associated with the first processed data set, wherein the metastore manager is configured to:

obtain a directed acyclic graph (DAG) corresponding to the first data processing operation; and

configure the first data processing operation in association with the DAG, wherein the DAG includes an unresolved logical plan, a logical plan, and an optimized logical plan;

initiating, by the metastore manager, a synchronization operation to store, in a second metastore associated with a second data management platform, a second metadata set corresponding to a subset of the first metadata set; and

activating, by the metastore manager, a data processing pipeline that is not communicatively coupled with the first data management platform wherein the data processing pipeline includes an artificial intelligence (AI) pipeline, to cause the data processing pipeline to:

obtain the partition metadata using a refresh component, and wherein the data processing pipeline is configured to:

obtain, based on the partition metadata, the first processed data set;

perform a second data processing operation on the first processed data set to generate a second processed data set, wherein the second processed data set comprises an output of an AI operation; and

output the second processed data set to a unified communications as a service (UCaaS) platform.

7. The non-transitory computer-readable medium of claim 6 , wherein the data processing pipeline is configured to obtain the first processed data set via a platform interface that connects the data processing pipeline with only the second data management platform.

8. The non-transitory computer-readable medium of claim 6 , wherein the first data processing operation comprises a structured query language (SQL) query on a distributed database associated with at least one of the first data store or a second data store associated with the second metastore.

9. The non-transitory computer-readable medium of claim 6 , the operations further comprising allocating a set of computational resources for the first data processing operation.

10. A system, comprising:

one or more memories; and

one or more processors configured to execute instructions stored in the one or more memories to:

communicating a request, from a driver node to a metastore manager, for initiating, in association with a data set maintained in a first data store associated with a first data management platform, a first data processing operation, to be performed by the first data management platform, to generate a first processed data set, wherein the first processed data set is stored in the first data store, and wherein a first metadata set corresponding to the first processed data set is stored in a first metastore associated with the first data store, the first metadata set comprising partition metadata indicative of partitioning information associated with the first processed data set, wherein the metastore manager is configured to:

obtain a directed acyclic graph (DAG) corresponding to the first data processing operation; and

configure the first data processing operation in association with the DAG, wherein the DAG includes an unresolved logical plan, a logical plan, and an optimized logical plan;

initiate, by the metastore manager, a synchronization operation to store, in a second metastore associated with a second data management platform, a second metadata set corresponding to a subset of the first metadata set; and

activate, by the metastore manager, a data processing pipeline that is not communicatively coupled with the first data management platform wherein the data processing pipeline includes an artificial intelligence (AI) pipeline, to cause the data processing pipeline to obtain the partition metadata using a refresh component, and wherein the data processing pipeline is configured to:

obtain, based on the partition metadata, the first processed data set;

perform a second data processing operation on the first processed data set to generate a second processed data set, wherein the second processed data set comprises an output of an AI operation; and

output the second processed data set to a unified communications as a service (UCaaS) platform.

11. The system of claim 10 , further comprising a platform interface that communicatively couples the data processing pipeline with the second data management platform.

12. The system of claim 10 , further comprising a platform interface that communicatively couples the data processing pipeline with only the second data management platform.

13. The system of claim 10 , further comprising a platform interface that communicatively couples the first data management platform with the second data management platform.

14. The system of claim 10 , further comprising a platform interface that communicatively couples the data processing pipeline with a unified communications as a service (UCaaS) platform.

15. The system of claim 10 , wherein the second data management platform comprises a metastore manager configured to manage access, by the data processing pipeline, to the first data store.

Assignments (2)
CHANGE OF NAME Recorded Jan 7, 2025
From: ZOOM VIDEO COMMUNICATIONS, INC.
To: ZOOM COMMUNICATIONS, INC.
Reel/Frame 069839/0593 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 11, 2024
From: MA, JIYUE
To: ZOOM VIDEO COMMUNICATIONS, INC.
Reel/Frame 067693/0705 →
References Cited (16)
US 8200613B1 · Wan et al. · 2012 [cited by applicant]
US 10102269B2 · Marin · 2018 [cited by applicant]
US 10120904B2 · Ranganathan · 2018 [cited by applicant]
US 10541938B1 · Timmerman et al. · 2020 [cited by applicant]
US 11074107B1 · Nandakumar · 2021 [cited by applicant]
US 11107177B1 · Ashley · 2021 [cited by examiner]
US 11153173B1 · Rebeja · 2021 [cited by applicant]
US 11741119B2 · Orun · 2023 [cited by applicant]
US 11941014B1 · Das · 2024 [cited by examiner]
US 20110282835A1 · Cannon · 2011 [cited by examiner]
US 20180196867A1 · Wiesmaier et al. · 2018 [cited by applicant]
US 20200257596A1 · Gokhale · 2020 [cited by examiner]
US 20210124739A1 · Karanasos et al. · 2021 [cited by applicant]
US 20210294814A1 · Murthy · 2021 [cited by applicant]
US 20220138004A1 · Nandakumar · 2022 [cited by applicant]
US 20240427911A1 · Apostolatos · 2024 [cited by examiner]