IP Library › Granted Patent US 10,339,038
Granted Patent B1
US 10,339,038 · App. 15/718,457 · Granted Jul 2, 2019

Method and system for generating production data pattern driven test data

Inventors: Jagmohan Singh (Coppell, TX); Priya Ranjan (Kendall Park, NJ)
Assignee: JPMORGAN CHASE BANK, N.A.
G06F11/3684G06F11/3664G06F11/3688
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,339,038
App. No.
15/718,457
Granted
Jul 2, 2019
Kind
B1
Abstract

The invention relates to implementing a test data tool that generates test data based on production data patterns. According to an embodiment of the present invention, the test data tool comprises: a processor configured to: receive, via the data input, production data from the one or more production environments, the production data comprises personally identifiable information; identify a plurality of attributes from the production data; for each attribute, identify one or more data patterns; generate one or more rules that define the one or more data patterns for each attribute; generate a configuration file based on the one or more rules; apply the configuration file to generate test data in a manner that obscures personally identifiable information existing in the production data; and transmit the test data to a UAT environment.

Claims (38)

1. A computer implemented system that implements a test data tool that generates test data based on production data patterns, the test data tool comprising:

a data input that interfaces with one or more production environments;

an output interface that transmits test data to one or more user acceptance testing (UAT) environments;

a communication network that receives production data from the one or more production environments and transmits test data to the one or more UAT environments; and

a computer server comprising at least one processor, coupled to the data input, the interactive user interface and the communication network, the processor configured to:

receive, via the data input, production data from the one or more production environments, the production data comprises personally identifiable information;

identify a plurality of attributes from the production data;

for each attribute, identify one or more data patterns;

generate one or more rules that define the one or more data patterns for each attribute;

generate a configuration file based on the one or more rules;

apply the configuration file to generate test data in a manner that obscures personally identifiable information existing in the production data; and

transmit the test data to a UAT environment.

2. The computer implemented system of claim 1 , wherein the one or more data patterns comprise data type including integer, decimal and character strings.

3. The computer implemented system of claim 1 , wherein the one or more data patterns comprise a representative expression.

4. The computer implemented system of claim 1 , wherein a representation of the configuration file is displayed on the interactive user interface to enable a user to update the configuration file.

5. The computer implemented system of claim 1 , wherein the configuration file identifies a number of test data records to be generated.

6. The computer implemented system of claim 1 , wherein at least one of the plurality of attributes comprise a code value with corresponding logic.

7. The computer implemented system of claim 1 , wherein the computer server is implemented in a cloud based architecture.

8. The computer implemented system of claim 1 , wherein the computer server applies machine learning to generate the one or more rules.

9. The computer implemented system of claim 1 , wherein one or more attributes have a relationship where a first identifier of a first attribute matches a second identifier of a second attribute.

10. The computer implemented system of claim 9 , wherein the configuration file preserves the relationship between the first attribute and the second attribute.

11. A computer implemented method that implements a test data tool that generates test data based on production data patterns, the method comprising the steps of:

receiving, via a data input, production data from one or more production environments, the production data comprises personally identifiable information;

identifying, via computer server, a plurality of attributes from the production data;

for each attribute, identifying one or more data patterns;

generating, via a rules engine, one or more rules that define the one or more data patterns for each attribute;

generating a configuration file based on the one or more rules;

applying the configuration file to generate test data in a manner that obscures personally identifiable information existing in the production data; and

transmitting, via a communication network, the test data to a user acceptance testing (UAT) environment.

12. The computer implemented method of claim 11 , wherein the one or more data patterns comprise data type including integer, decimal and character strings.

13. The computer implemented method of claim 11 , wherein the one or more data patterns comprise a representative expression.

14. The computer implemented method of claim 11 , wherein a representation of the configuration file is displayed on the interactive user interface to enable a user to update the configuration file.

15. The computer implemented method of claim 11 , wherein the configuration file identifies a number of test data records to be generated.

16. The computer implemented method of claim 11 , wherein at least one of the plurality of attributes comprise a code value with corresponding logic.

17. The computer implemented method of claim 11 , wherein the computer server is implemented in a cloud based architecture.

18. The computer implemented method of claim 11 , wherein the computer server applies machine learning to generate the one or more rules.

19. The computer implemented method of claim 11 , wherein one or more attributes have a relationship where a first identifier of a first attribute matches a second identifier of a second attribute.

20. The computer implemented method of claim 19 , wherein the configuration file preserves the relationship between the first attribute and the second attribute.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 5, 2017
From: SINGH, JAGMOHAN; RANJAN, PRIYA
To: JPMORGAN CHASE BANK, N.A.
Reel/Frame 043793/0750 →
Cited By (4)
US 12,248,614 US 12,254,110 US 12,561,455 US 12,724,701