SYSTEM AND METHOD FOR GENERATING FILE SYSTEM AND BLOCK-BASED INCREMENTAL BACKUPS USING ENHANCED DEPENDENCIES AND FILE SYSTEM INFORMATION OF DATA BLOCKS
A method for a backup operation includes obtaining, by a backup agent, a backup request for an incremental backup of a file system, and in response to the backup request: selecting a reference backup from a backup storage system, obtaining a first hash value document associated with the reference backup, generating a hash value for an asset associated with the file system, making a first determination that the hash value matches a second hash value specified in the first hash value document, in response to the first determination, populating an incremental backup with a copy of data associated with the asset, initiating a transfer of the incremental backup to the backup storage system, and storing a second hash value document, wherein the second hash value document comprises the hash value and a backup identifier of the incremental backup.
1 . A method for managing a persistent storage system, the method comprising:
obtaining, by a backup agent, a backup request for an incremental backup of a file system; and
in response to the backup request:
selecting a reference backup from a backup storage system;
obtaining a first hash value document associated with the reference backup;
generating a hash value for an asset associated with the file system;
making a first determination that the hash value matches a second hash value specified in the first hash value document;
in response to the first determination, populating an incremental backup with a copy of data associated with the asset;
initiating a transfer of the incremental backup to the backup storage system; and
storing a second hash value document, wherein the second hash value document comprises the hash value and a backup identifier of the incremental backup.
2 . The method of claim 1 , further comprising:
prior to initiating the transfer of the incremental backup to the backup storage system:
generating a third hash value for a second asset associated with the file system;
making a second determination that the third hash value does not match a fourth hash value; and
in response to the second determination, not populating the incremental backup with a second copy of second data associated with the second asset.
3 . The method of claim 2 , wherein the second hash value document further comprises the third hash value.
4 . The method of claim 1 , wherein the first hash value document comprises a timestamp associated with the reference backup, a second backup identifier associated with the reference backup, and a plurality of hash values.
5 . The method of claim 1 , wherein the reference backup is not a most recent backup of the file system.
6 . The method of claim 1 , wherein the asset is a file in the file system.
7 . The method of claim 1 , further comprising:
obtaining a second backup request for an incremental block-based backup; and
in response to the second backup request:
identifying a plurality of data blocks changed since a most recent block-based backup;
performing a data block file analysis on the plurality of data blocks to identify a plurality of modified files;
generating the incremental block-based backup using the plurality of data blocks;
generating a file change document based on the plurality of modified files;
updating the incremental block-based backup based on the file change document; and
initiating a transfer of the incremental block-based backup to the backup storage system.
8 . A system, comprising:
a processor; and
memory comprising instructions which, when executed by the processor, perform a method, the method comprising:
obtaining, by a backup agent, a backup request for an incremental backup of a file system; and
in response to the backup request:
selecting a reference backup from a backup storage system;
obtaining a first hash value document associated with the reference backup;
generating a hash value for an asset associated with the file system;
making a first determination that the hash value matches a second hash value specified in the first hash value document;
in response to the first determination, populating an incremental backup with a copy of data associated with the asset;
initiating a transfer of the incremental backup to the backup storage system; and
storing a second hash value document, wherein the second hash value document specifies the hash value and a backup identifier of the incremental backup.
9 . The system of claim 8 , the method further comprising:
prior to initiating the transfer of the incremental backup to the backup storage system:
generating a third hash value for a second asset associated with the file system;
making a second determination that the third hash value does not match a fourth hash value; and
in response to the second determination, not populating the incremental backup with a second copy of second data associated with the second asset.
10 . The system of claim 9 , wherein the second hash value document further comprises the third hash value.
11 . The system of claim 8 , wherein the first hash value document comprises a time stamp associated with the reference backup, a second backup identifier associated with the reference backup, and a plurality of hash values.
12 . The system of claim 8 , wherein the reference backup is not a most recent backup of the file system.
13 . The system of claim 8 , wherein the asset is a file in the file system.
14 . The system of claim 8 , the method further comprising:
obtaining a second backup request for an incremental block-based backup; and
in response to the second backup request:
identifying a plurality of data blocks changed since a most recent block-based backup;
performing a data block file analysis on the plurality of data blocks to identify a plurality of modified files;
generating the incremental block-based backup using the plurality of data blocks;
generating a file change document based on the plurality of modified files;
updating the incremental block-based backup based on the file change document; and
initiating a transfer of the incremental block-based backup to the backup storage system.
15 . A non-transitory computer readable medium comprising computer readable program code, which when executed by a computer processor enables the computer processor to perform a method for performing a backup operation, the method comprising:
obtaining, by a backup agent, a backup request for an incremental backup of a file system; and
in response to the backup request:
selecting a reference backup from a backup storage system;
obtaining a first hash value document associated with the reference backup;
generating a hash value for an asset associated with the file system;
making a first determination that the hash value matches a second hash value specified in the first hash value document;
in response to the first determination, populating an incremental backup with a copy of data associated with the asset;
initiating a transfer of the incremental backup to the backup storage system; and
storing a second hash value document, wherein the second hash value document comprises the hash value and a backup identifier of the incremental backup.
16 . The non-transitory computer readable medium of claim 15 , the method further comprising:
prior to initiating the transfer of the incremental backup to the backup storage system:
generating a third hash value for a second asset associated with the file system;
making a second determination that the third hash value does not match a fourth hash value; and
in response to the second determination, not populating the incremental backup with a second copy of second data associated with the second asset.
17 . The non-transitory computer readable medium of claim 16 , wherein the second hash value document further comprises the third hash value.
18 . The non-transitory computer readable medium of claim 15 , wherein the first hash value document comprises a timestamp associated with the reference backup, a second backup identifier associated with the reference backup, and a plurality of hash values.
19 . The non-transitory computer readable medium of claim 15 , wherein the reference backup is not a most recent backup of the file system.
20 . The non-transitory computer readable medium of claim 15 , the method further comprising:
obtaining a second backup request for an incremental block-based backup; and
in response to the second backup request:
identifying a plurality of data blocks changed since a most recent block-based backup;
performing a data block file analysis on the plurality of data blocks to identify a plurality of modified files;
generating the incremental block-based backup based on the plurality of data blocks;
generating a file change document based on the plurality of modified files;
updating the incremental block-based backup based on the file change document; and
initiating a transfer of the incremental block-based backup to the backup storage system.