IP Library Granted Patent US 10,346,429
Granted Patent B2
US 10,346,429 · App. 15/186,520 · Granted Jul 9, 2019

Management of collaborative datasets via distributed computer networks

Inventors: Bryon Kristen Jacob (Austin, TX); David Lee Griffith (Austin, TX); Triet Minh Le (Austin, TX); Jon Loyens (Austin, TX); Brett A. Hurt (Austin, TX); Arthur Albert Keen (Austin, TX)
Assignee: data.world, Inc.
G06F16/273G06F16/178G06F16/2365G06F21/6227G06F2221/2141
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,346,429
App. No.
15/186,520
Granted
Jul 9, 2019
Kind
B2
Abstract

Various embodiments relate generally to data science and data analysis, computer software and systems, and wired and wireless network communications to provide an interface between repositories of disparate datasets and computing machine-based entities that seek access to the datasets, and, more specifically, to a computing and data storage platform that facilitates consolidation of one or more datasets, whereby a collaborative data layer and associated logic facilitate, for example, efficient access to, and implementation of, collaborative datasets. In some examples, a method may include receiving a dataset and dataset attributes and identifying a first version of the dataset. The method may include identifying data that varies from a first version of the dataset, and generating a second version of the dataset to include a first subset and a second subset of atomized data. The method may include storing subsets of atomized data points as an atomized dataset.

Claims (41)

1. A method comprising:

receiving via a network data representing a dataset having a data format into a collaborative dataset consolidation system, the collaborative dataset consolidation system including one or more processors and one or more repositories configured to convert differently-formatted datasets into atomized datasets as data arrangements to facilitate interoperability among converted datasets;

receiving data representing attributes associated with the dataset, the attributes including an account identifier;

identifying a first version of the dataset associated with a first subset of atomized data points, at least one atomized data point being a triple associated with non-protected data;

identifying a subset of data that varies from the first version of the dataset;

accessing the subset of data as a protected data, access to which is authorized as a function of data representing a level of authorization for the account identifier;

converting the subset of data including a non-atomized data point to a second subset of atomized data points having a specific format similar to the first subset;

generating a second version of the dataset to include the first subset of atomized data points and the second subset of atomized data points; and

storing the first subset of atomized data points and the second subset of atomized data points as an atomized dataset in the one or more repositories.

2. The method of claim 1 wherein each of the atomized data point is data representing an addressable data unit.

3. The method of claim 1 wherein storing the first subset of atomized data points and the second subset of atomized data points as the atomized dataset comprises:

storing atomized data points as triples.

4. The method of claim 3 wherein at least one triple of the triples are formatted to comply with a Resource Description Framework (“RDF”) data model.

5. The method of claim 1 wherein generating the second version of the dataset to include the first subset of atomized data points comprises:

generating a data pointer to a memory location at which the first subset of atomized data points is stored.

6. The method of claim 5 wherein storing the first subset of atomized data points comprises:

storing the data pointer as the first subset of atomized data points.

7. The method of claim 1 further comprising:

identifying that the subset of data is associated with a protected dataset;

determining access to the protected dataset is authorized in association with the account identifier; and

forming the second version of the dataset.

8. The method of claim 1 further comprising:

receiving a request in association with another account identifier to access the atomized dataset;

determining access to the subset of data is not authorized; and

denying access to the atomized dataset in association with the another account identifier.

9. The method of claim 1 further comprising managing dataset attributes associated with the atomized dataset.

10. The method of claim 9 wherein managing the dataset attributes comprises:

analyzing atomized datasets associated with the collaborative dataset consolidation system; and

identifying a number of queries associated with the atomized dataset.

11. The method of claim 9 wherein managing the dataset attributes comprises:

analyzing atomized datasets associated with the collaborative dataset consolidation system; and

identifying a subset of other account identifiers that include descriptive data that correlate to the atomized dataset.

12. The method of claim 11 further comprising:

generating a data signal specifying information for at least one of the account identifiers that accessed the descriptive data; and

causing presentation of the information in an activity feed portion of a user interface.

13. The method of claim 9 wherein managing the dataset attributes comprises:

analyzing atomized datasets associated with the collaborative dataset consolidation system; and

identifying a subset of other atomized datasets including similar classification types.

14. The method of claim 13 further comprising:

generating a data signal specifying information for at least one of the other atomized datasets; and

causing presentation of the information in a recommendation portion of a user interface.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 2, 2025
From: DATA.WORLD, INC.
To: SERVICENOW, INC.
Reel/Frame 073004/0844 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 12, 2016
From: JACOB, BRYON KRISTEN; LOYENS, JON; GRIFFITH, DAVID LEE; HURT, BRETT A; LE, TRIET MINH; KEEN, ARTHUR ALBERT
To: DATA.WORLD, INC.
Reel/Frame 040001/0170 →
Continuity (1)
Related Publication 20170364703A1 · Dec 21, 2017
Cited By (3)
US 12,292,870 US 12,455,865 US 12,608,366