IP Library Granted Patent US 10,645,548
Granted Patent B2
US 10,645,548 · App. 15/454,981 · Granted May 5, 2020

Computerized tool implementation of layered data files to discover, form, or analyze dataset interrelations of networked collaborative datasets

Inventors: Shad William Reynolds (Austin, TX); David Lee Griffith (Austin, TX); Bryon Kristen Jacob (Austin, TX)
Assignee: data.world, Inc.
H04W4/38G06F16/254G06F17/18G06F21/604G06F21/6245H04L63/08H04L63/104H04L67/303G06F3/0482H04L67/306H04W4/90
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,645,548
App. No.
15/454,981
Granted
May 5, 2020
Kind
B2
Abstract

A method may relate generally to data science and data analysis, computer software and systems, and wired and wireless network communications to provide an interface between repositories of disparate datasets and computing machine-based entities that seek access to the datasets, and, more specifically, to a computing and data storage platform that facilitates consolidation of one or more datasets. One or more computerized tools may be configured to discover, form, and analyze via one or more layered data files, interrelations among a system of networked collaborative datasets. A method may include transforming of a set of data to an atomized format to form an atomized dataset that includes a derived dataset attribute. The method may also include presenting data representing an annotation at the user interface based on the derived dataset attribute. An annotation may be associated with a layer file.

Claims (63)

1. A method comprising:

receiving data to form a first input via a first user interface as a first user interface element to initiate creation of a first dataset based on a set of data, the first dataset being associated by a first user identifier;

activating a programmatic interface to facilitate the creation of the first dataset responsive to receiving the first input;

causing transformation at a processor of the set of data from a first format to an atomized format to form an atomized dataset including triples, the transformation of the set of data including deriving a dataset attribute based on a subset of data to form a derived dataset attribute for the set of data;

receiving at least one interaction with a second dataset in a subset of datasets via a second user interface associated with a second user identifier, wherein the subset of datasets including a set of data in a community of datasets associated with a community of networked users, each networked user being associated with one or more datasets as a subset of the community of datasets;

causing presentation of a notification including data representing the at least one interaction in an activity feed in the first user interface, the first user interface being configured to include selectable links configured to activate to form a link between the first dataset to other relevant datasets stored as graph-based data in one or more triplestore repositories, the graph-based data including one or more summary characteristics configured to represent other derived dataset attributes associated with a collaborative dataset generated using a collaborative dataset consolidation system;

receiving data representing an input to select the second set in the subset of datasets to generate another interaction;

presenting data representing an annotation at the first user interface based on the derived dataset attribute for the subset of data, the annotation being automatically determined based on inferred data; and

selecting the annotation for implementation automatically as a function of context based on a system of layer files responsive to the another interaction.

2. The method of claim 1 wherein presenting the data representing the annotation comprises:

accepting a second input via the first user interface as a second user interface element to cause linking between the atomized dataset and another dataset based on the annotation; and

presenting data that represents the annotation for a column of data.

3. The method of claim 1 further comprising:

presenting in the first user interface a data view of the subset of data as a column of data associated with an unknown dataset attribute.

4. The method of claim 3 further comprising:

receiving data to annotate a column header to form the annotation to resolve the unknown dataset attribute.

5. The method of claim 3 further comprising:

receiving data to annotate a datatype to form the annotation to resolve the unknown dataset attribute for the column of data.

6. The method of claim 1 further comprising:

presenting a data view in the first user interface of subsets of data from the set of data, each subset representing a column of data; and

presenting a derived column of data that includes data derived from one or more other columns of data.

7. The method of claim 6 further comprising:

receiving data to form a third input via the first user interface as a third user interface element that is configured to receive data signals to add the derived column of data; and

transmitting the data signals to add the derived column of data.

8. The method of claim 6 further comprising:

receiving data to form a fourth input via the first user interface as a fourth user interface element that is configured to receive data signals to substitute the derived column of data; and

transmitting the data signals to substitute the derived column of data to replace the one or more other columns of data from which the derived column were derived.

9. The method of claim 6 further comprising:

receiving data to form a fifth input via the first user interface as a fifth user interface element that is configured to receive data signals to reject the derived column of data; and

transmitting the data signals to reject the derived column.

10. The method of claim 6 wherein presenting the derived column of data comprises:

expanding the column of data into two or more columns of data.

11. The method of claim 6 wherein presenting the derived column of data comprises:

collapsing two or more columns of data into the column of data.

12. The method of claim 1 wherein causing transformation of the set of data to include deriving the dataset attribute comprises:

causing access to a first subset of layer files specifying datatypes to identify an inferred attribute.

13. The method of claim 1 wherein causing transformation of the set of data to include deriving the dataset attribute comprises:

causing access to a second subset of layer files specifying a data classification to identify an inferred attribute.

14. The method of claim 1 wherein causing transformation of the set of data to include deriving the dataset attribute comprises:

causing access to a third subset of layer files specifying contextual data to identify an inferred attribute.

15. The method of claim 1 further comprising accepting another input to link the atomized dataset to another dataset by:

identifying the another data set as a protected dataset;

transmitting authorization to access the protected dataset; and

causing the atomized dataset to form based on the access to the protected dataset.

16. The method of claim 1 wherein causing transformation of the set of data to include deriving the dataset attribute comprises:

causing a collaborative dataset consolidation system to identify contextual data with which to link the atomized data set to other atomized data sets.

17. The method of claim 1 further comprising:

receiving data from the first user interface to modify units for values of data in a column of data; and

receiving data representing a derived column of data that includes the values of data having modified units.

18. The method of claim 1 wherein the atomized dataset include subsets of linked data points stored in a triplestore.

19. The method of claim 18 wherein linked data points comprise triples, at least one triple of the triples are formatted to comply with a Resource Description Framework (“RDF”) data model.

20. An apparatus comprising:

a memory including executable instructions; and

a processor, the executable instructions executed by the processor to:

receive data to form a first input via a first user interface as a first user interface element to initiate creation of a first dataset based on a set of data, the first dataset being associated by a first user identifier;

activate a programmatic interface to facilitate the creation of the first dataset responsive to receiving the first input;

cause transformation of the set of data from a first format to an atomized format to form an atomized dataset, the transformation of the set of data including deriving a dataset attribute based on a subset of data to form a derived dataset attribute for the set of data;

receive at least one interaction with a second dataset in a subset of datasets via a second user interface associated with a second user identifier, wherein the subset of datasets including a subset of data in a community of datasets associated with a community of networked users;

cause presentation of a notification including data representing the at least one interaction in an activity feed in the first user interface, the first user interface being configured to include selectable links configured to activate to form a link between the first dataset to other relevant datasets stored as graph-based data in one or more triplestore repositories, the graph-based data including one or more summary characteristics configured to represent other derived dataset attributes associated with a collaborative dataset generated using a collaborative dataset consolidation system;

receive data representing an input to select the second dataset in the subset of datasets to generate another interaction;

present data representing an annotation at the first user interface based on the derived dataset attribute for the subset of data, the annotation being automatically determined based on inferred data,

wherein the derived dataset attribute is associated with a system of layer files specifying an inferred datatype as the derived dataset attribute; and

select the annotation for implementation automatically as a function of context based on the system of layer files responsive to the another interaction.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 2, 2025
From: DATA.WORLD, INC.
To: SERVICENOW, INC.
Reel/Frame 073004/0844 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 28, 2017
From: REYNOLDS, SHAD WILLIAM; GRIFFITH, DAVID LEE; JACOB, BRYON KRISTEN
To: DATA.WORLD, INC.
Reel/Frame 041772/0274 →
Continuity (2)
Continuation In Part 15186514 · Jun 19, 2016
Related Publication 20180262864A1 · Sep 13, 2018
Cited By (22)
US 12,190,330 US 12,204,564 US 12,216,794 US 12,259,882 US 12,265,896 US 12,277,232 US 12,288,233 US 12,292,870 US 12,299,065 US 12,353,405 US 12,381,915 US 12,412,140 US 12,536,329 US 12,561,296 US 12,591,828 US 12,608,366 US 12,609,938 US 12,641,108 US 12,645,641 US 12,688,324 US 12,694,044 US 12,718,167