IP Library › Granted Patent US 10,114,907
Granted Patent B2
US 10,114,907 · App. 14/940,239 · Granted Oct 30, 2018

Query processing for XML data using big data technology

Inventor: George F. Wang (San Jose, CA)
Assignee: INTERNATIONAL BUSINESS MACHINES CORPORATION
G06F17/30938G06F17/30076G06F17/30194G06F17/30864H04L67/1097
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,114,907
App. No.
14/940,239
Granted
Oct 30, 2018
Kind
B2
Abstract

A computer-implemented method for offloading extensible markup language (XML) data to a distributed file system may include receiving a command to populate a distributed file system with an XML table of a database. The XML table may be queried in response to the command. The source data in the XML table may be offloaded, by a computer processor, to the distributed file system in response to the querying. The offloading may include converting the source data to a string version of the source data and converting the string version of the source data back into XML format.

Claims (36)

1. A system for offloading extensible markup language (XML) data to a distributed file system, comprising:

a memory having computer readable instructions; and

one or more processors for executing the computer readable instructions, the computer readable instructions comprising:

receiving a command to populate a distributed file system with an XML table of a database;

querying the XML table in response to the command;

offloading source data in the XML table to the distributed file system in response to the querying;

receiving an XML path language (XPath) query against the XML table in the database;

processing the X query based on the source data in the distributed file system to generate a result data; and

storing the result data on the distributed file system in response to the XPath query.

2. The system of claim 1 , wherein the database is a DB2 database and the distributed file system is Hadoop Distributed File System (HDFS), and wherein the processing the XPath query based on the source data in the distributed file system to generate the result data comprises use of MapReduce functionality.

3. The system of claim 2 , the computer readable instructions further comprising:

receiving a DB2 user-defined function to retrieve the result data; and

storing the result data to the DB2 database in response to the DB2 user-defined function.

4. The system of claim 2 , the computer readable instructions further comprising:

receiving a request for the result data from the DB2 database; and

storing the result data to the DB2 database responsive to the request.

5. The system of claim 2 , wherein the offloading is performed by a BigInsights server.

6. The system of claim 2 , wherein the offloading is performed by a server for the DB2 database.

7. A computer program product for offloading extensible markup language (XML) data to a distributed file system, the computer program product comprising a computer readable storage medium having program instructions embodied therewith, the program instructions executable by a processor to cause the processor to perform a method comprising:

receiving a command to populate a distributed file system with an XML table of a database;

querying the XML table in response to the command;

offloading source data in the XML table to the distributed file system in response to the querying;

receiving an XML, path language (XPath) query against the XML, table in the database;

processing the XPath query based on the source data in the distributed file system to generate a result data; and

storing the result data on the distributed file system in response to the XPath query.

8. The computer program product of claim 7 , wherein the database is a DB2 database and the distributed file system is Hadoop Distributed File System (HDFS), and wherein the processing the XPath query based on the source data in the distributed file system to generate the result data comprises use of MapReduce functionality.

9. The computer program product of claim 8 , the method further comprising:

receiving a DB2 user-defined function to retrieve the result data; and

storing the result data to the DB2 database in response to the DB2 user-defined function.

10. The computer program product of claim 8 , the method further comprising:

receiving a request for the result data from the DB2 database; and

storing the result data to the DB2 database responsive to the request.

11. The computer program product of claim 8 , wherein the offloading is performed by a BigInsights server.

12. The system of claim 1 , wherein the offloading comprises converting the source data to a plain text version of the source data and converting the plain text version of the source data back into XML format.

13. The computer program product of claim 7 , wherein the offloading comprises converting the source data to a plain text version of the source data and converting the plain text version of the source data back into XML format.

14. The computer program product of claim 8 , wherein the offloading is performed by a server for the DB2 database.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 13, 2015
From: WANG, GEORGE F.
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 037030/0403 →
Continuity (1)
Related Publication 20170140064A1 · May 18, 2017