IP Library Granted Patent US 10,339,106
Granted Patent B2
US 10,339,106 · App. 14/682,988 · Granted Jul 2, 2019

Highly reusable deduplication database after disaster recovery

Inventors: Manoj Kumar Vijayan (Marlboro, NJ); Ganesh Haridas (Pollachi, IN); Deepak Raghunath Attarde (Marlboro, NJ)
Assignee: Commvault Systems, Inc.
G06F16/162G06F11/1451G06F11/1453G06F11/1464G06F11/1469G06F16/113G06F16/178G06F16/1748G06F16/1752G06F16/24556G06F2201/80G06F2201/805
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,339,106
App. No.
14/682,988
Granted
Jul 2, 2019
Kind
B2
Abstract

According to certain aspects, a method can include receiving, in response to an indication that a data storage database is being restored to a second time before a first time such that the data storage database comprises a plurality of first archive file identifiers associated at the second time, a first instruction from a data storage computer, where the first instruction instructs a media agent to stop scheduled secondary storage operations associated with a deduplication database, and where the deduplication database comprises a plurality of second archive file identifiers; determining at least one second archive file identifier in the plurality of second archive file identifiers that does not correlate with any first archive identifier in the plurality of first archive file identifiers; and, for each of the at least one second archive identifier, instructing the deduplication database to prune an entry associated with the respective second archive file identifier.

Claims (28)

1. A networked information management system configured to verify synchronization of deduplication information, the networked information management system comprising:

a data storage database comprising a plurality of first job identifiers, wherein each first job identifier comprises a time that a respective job occurred;

a data storage computer, the data storage computer comprising computer hardware configured to receive an indication that the data storage database is being restored to an earlier version of the data storage database; and

a media agent that executes on one or more computer processors and that is configured to:

receive a first instruction from the data storage computer in response to the data storage computer receiving the indication, wherein the first instruction instructs the media agent to stop scheduled secondary storage operations associated with a deduplication database, wherein the deduplication database comprises a job identifier table which correlates deduplication information with a plurality of second job identifiers, wherein the deduplication database comprises a plurality of reference counters, each reference counter representing a number of links to a respective data block in a secondary storage system; and

for each of the second job identifiers that does not correlate with any first job identifiers in a subset of the first job identifiers, instruct the deduplication database to prune an entry in the job identifier table associated with the second job identifier and to decrement by one each of the reference counters of data blocks that are associated with the second job identifier.

2. The networked information management system of claim 1 , further comprising a second media agent configured to:

receive a second instruction from the data storage computer in response to the data storage computer receiving the indication, wherein the second instruction instructs the second media agent to stop scheduled secondary storage operations associated with a second deduplication database, wherein the second deduplication database comprises a second job identifier table which correlates deduplication information with a plurality of third job identifiers, wherein the second deduplication database comprises a plurality of second reference counters, each second reference counter representing a number of links to a respective data block in the secondary storage system; and

for each of the third job identifiers that does not correlate with any of the first job identifiers in a subset of the plurality of first job identifiers, instruct the second deduplication database to prune an entry in the second job identifier table associated with the third job identifier and to decrement by one each of the second reference counters of data blocks that are associated with the third job identifier.

3. The networked information management system of claim 2 , wherein the data storage computer transmits the first instruction and the second instruction at a same time.

4. The networked information management system of claim 1 , wherein each second job identifier is associated with a data block stored in secondary storage.

5. The networked information management system of claim 4 , wherein each data block associated with the at least one second job identifier is removed from the secondary storage.

6. A computer-implemented method for verifying synchronization of deduplication information, the computer-implemented method comprising:

receiving, in response to an indication that a data storage database is being restored to an earlier version of the data storage database, a first instruction from a data storage computer, wherein individual first job identifiers correspond with a time that a respective job occurred, wherein the first instruction instructs a media agent to stop scheduled secondary storage operations associated with a deduplication database, wherein the deduplication database comprises a job identifier table which correlates deduplication information with a plurality of second job identifiers, wherein the deduplication database comprises a plurality of reference counters, each reference counter representing a number of links to a respective data block in a secondary storage system; and

for each of the second job identifiers that does not correlate with any first job identifiers in a subset of the first job identifiers, instructing the deduplication database to prune an entry in the job identifier table associated with the second job identifier and to decrement by one each of the reference counters of data blocks that are associated with the second job identifier.

7. The computer-implemented method of claim 6 , wherein each second job identifier is associated with a data block stored in secondary storage.

8. The computer-implemented method of claim 7 , wherein each data block associated with the at least one second job identifier is removed from the secondary storage.

9. A networked information management system configured to verify synchronization of deduplication information, the networked information management system comprising:

a storage manager database comprising a plurality of first job identifiers, wherein each first job identifier corresponds to a respective job;

a storage manager, the storage manager comprising computer hardware configured to receive an indication that the storage manager database is being restored to an earlier version of the storage manager database; and

a deduplication database media agent comprising an electronically stored deduplication database and computer hardware configured to:

receive a first instruction from the storage manager in response to the storage manager receiving the indication, wherein the first instruction instructs the deduplication database media agent to stop scheduled secondary storage operations associated with the deduplication database, wherein the deduplication database comprises a job identifier table which correlates deduplication information with a plurality of second job identifiers, the deduplication database comprises a plurality of reference counters, each reference counter representing a number of links to a respective data block in a secondary storage system; and

for each of the second job identifiers that does not correlate with any first job identifiers in a subset of the plurality of first job identifiers, instruct the deduplication database to prune an entry in the job identifier table associated with that second job identifier and to decrement by one each of the reference counters of the data blocks that are associated with that second job identifier.

10. The networked information management system of claim 9 , wherein the deduplication database media agent is further configured to determine at least one second job identifier that does not match any first job identifiers in the plurality of first job identifiers.

11. The networked information management system of claim 9 , further comprising a second deduplication database media agent comprising an electronically stored second deduplication database and computer hardware configured to:

receive a second instruction from the storage manager in response to the storage manager receiving the indication, wherein the second instruction instructs the second deduplication database media agent to stop scheduled secondary storage operations associated with a second deduplication database, wherein the second deduplication database comprises a second job identifier table which correlates deduplication information with a plurality of third job identifiers, the second deduplication database comprises a plurality of second reference counters, each second reference counter representing a number of links to a respective data block in the secondary storage system; and

for each of the third job identifiers that does not correlate with any first job identifiers in a subset of the plurality of first job identifiers, instruct the second deduplication database to prune an entry in the second job identifier table associated with that third job identifier and to decrement by one each of the second reference counters of the data blocks that are associated with that third job identifier.

12. The networked information management system of claim 11 , wherein the storage manager transmits the first instruction and the second instruction at a same time.

Assignments (3)
SUPPLEMENTAL CONFIRMATORY GRANT OF SECURITY INTEREST IN UNITED STATES PATENTS Recorded Apr 16, 2025
From: COMMVAULT SYSTEMS, INC.
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 070864/0344 →
SECURITY INTEREST Recorded Dec 13, 2021
From: COMMVAULT SYSTEMS, INC.
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 058496/0836 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 10, 2015
From: VIJAYAN, MANOJ KUMAR; HARIDAS, GANESH; ATTARDE, DEEPAK RAGHUNATH
To: COMMVAULT SYSTEMS, INC.
Reel/Frame 035410/0596 →
Continuity (1)
Related Publication 20160299818A1 · Oct 13, 2016
Cited By (2)
US 12,321,313 US 12,547,350