IP Library › Granted Patent US 10,296,656
Granted Patent B2
US 10,296,656 · App. 14/821,429 · Granted May 21, 2019

Managing database

Inventors: Li Li (Beijing, CN); Liang Liu (Beijing, CN); Junmei Qu (Beijing, CN); Wen Jun Yin (Beijing, CN); Wei Zhuang (Beijing, CN)
Assignee: International Business Machines Corporation
G06F17/30917G06F17/30297G06F17/30315G06F17/30353G06F17/30486G06F17/30548
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,296,656
App. No.
14/821,429
Granted
May 21, 2019
Kind
B2
Abstract

A method for managing a database, each item of data in the database being associated with a timestamp and a data point, the timestamps being used as row keys for rows of a table in the database, the method comprising: obtaining a behavior characteristic of a user based on a previous data access to the database by the user; partitioning columns in the table into column families based on the obtained behavior characteristic and system configuration of the database; and causing data in the database to be stored in respective column families at least in part based on the associated data point.

Claims (16)

1. A computer-implemented method for managing a database, each item of data in the database being associated with a timestamp and a data point, the timestamps being used as row keys for rows of a table in the database, the rows defined on one of two dimensions of the table, the method comprising:

obtaining, at a processor of a computer system, a behavior characteristic of a user based on a previous data access to the database by the user, said behavior characteristic indicating a time span of the previous data access;

partitioning, using the processor, columns in the table into column families based on the obtained behavior characteristic and a system configuration of the database, the system configuration indicating a size of blocks in a file system associated with the database, the columns defined on the other of the two dimensions of the table; and

storing, by the processor, data in the database in respective column families at least in part based on the associated data points.

2. The method according to claim 1 ,

wherein partitioning columns in the table into column families comprises:

generating, by the processor, at least one logical table from the table in the database, the number of rows in the at least one logical table determined based on the time span; and

determining, by the processor, the number of columns included in each of the column families at least in part based on the number of rows in the at least one logical table and the size of the blocks, such that the number of the blocks occupied by data in the column family is minimized.

3. The method according to claim 2 , wherein the behavior characteristic further indicates a data type of the previous data access, and wherein the number of columns included in each of the column families is further determined based on the data type.

4. The method according to claim 2 , wherein data in the table of the database is partitioned into a plurality of regions to store, the method further comprising:

determining, by the processor, the number of the column families included in the at least one logical table based on a size of a region and the size of the blocks, such that the number of the regions occupied by data in the at least one logical table is minimized.

5. The method according to claim 2 , further comprising:

maintaining, by the processor, an index for the table in the database, the index mapping identifications of the data points to the at least one logical table and the column families.

6. The method according to claim 1 , wherein the storing of data in the database in respective column families at least in part based on the associated data points comprises:

causing, by the processor, data for correlated data points to be stored in a same column family.

7. The method according to claim 1 , wherein the database is a Hadoop database (HBase).

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 7, 2015
From: LI, LI; LIU, LIANG; QU, JUNMEI; YIN, WEN JUN; ZHUANG, WEI
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 036281/0316 →
Priority Claims (1)
CN 2014 1 0119830 · Mar 27, 2014 · national
Continuity (2)
Continuation 14670208 · Mar 26, 2015
Related Publication 20150347622A1 · Dec 3, 2015
Cited By (2)
US 12,276,759 US 12,742,860