IP Library Granted Patent US 10,102,258
Granted Patent B2
US 10,102,258 · App. 15/186,514 · Granted Oct 16, 2018

Collaborative dataset consolidation via distributed computer networks

Inventors: Bryon Kristen Jacob (Austin, TX); David Lee Griffith (Austin, TX); Triet Minh Le (Austin, TX); Arthur Albert Keen (Austin, TX); Alexander John Zelenak (Austin, TX); Jon Loyens (Austin, TX); Brett A. Hurt (Austin, TX); Shad William Reynolds (Austin, TX); Joseph Boutros (Austin, TX)
Assignee: data.world, Inc.
G06F17/30545G06F17/30303
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,102,258
App. No.
15/186,514
Granted
Oct 16, 2018
Kind
B2
Abstract

Various embodiments relate generally to data science and data analysis, computer software and systems, and wired and wireless network communications to provide an interface between repositories of disparate datasets and computing machine-based entities that seek access to the datasets, and, more specifically, to a computing and data storage platform that facilitates consolidation of one or more datasets, whereby a collaborative data layer and associated logic facilitate, for example, efficient access to, and implementation of, collaborative datasets. In some examples, a method may include receiving data representing a query into a collaborative dataset consolidation system, identifying datasets relevant to the query, generating one or more queries to access disparate data repositories, and retrieving data representing query results. In some cases, one or more queries are applied (e.g., as a federated query) to atomized datasets stored in one or more atomized data stores, at least two of which may be different.

Claims (15)

1. A method comprising:

receiving a data file including a dataset into a collaborative dataset consolidation system;

formatting the dataset to form a first atomized dataset including atomized data points each including data representing at least two objects and an association between the two objects;

forming a second atomized dataset including the first atomized dataset and one or more other atomized datasets;

receiving data representing a query into the collaborative dataset consolidation system, the query being associated with an identifier;

identifying a subset of the second atomized dataset relevant to the query, wherein portions of the second atomized dataset are disposed in different data repositories;

generating a plurality of sub-queries each of which is configured to access at least one of the different data repositories; and

retrieving data representing query results from the at least one of the different data repositories.

2. The method of claim 1 wherein generating the plurality of sub-queries comprises:

classifying query portions.

3. The method of claim 2 wherein classifying the query portions comprises:

identifying a classification type for a portion of the query.

4. The method of claim 1 wherein the datasets comprise linked data points.

5. The method of claim 4 wherein linked data points comprise triples.

6. The method of claim 5 wherein at least one triple of the triples are formatted to comply with a Resource Description Framework (“RDF”) data model.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 2, 2025
From: DATA.WORLD, INC.
To: SERVICENOW, INC.
Reel/Frame 073004/0844 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 12, 2016
From: JACOB, BRYON KRISTEN; LOYENS, JON; GRIFFITH, DAVID LEE; HURT, BRETT A; LE, TRIET MINH; REYNOLDS, SHAD; KEEN, ARTHUR ALBERT; BOUTROS, JOSEPH; ZELENAK, ALEXANDER JOHN
To: DATA.WORLD, INC.
Reel/Frame 040001/0068 →
Continuity (1)
Related Publication 20170364569A1 · Dec 21, 2017
Cited By (2)
US 12,292,870 US 12,608,366