SCALABLE DEDUPLICATION OF STORED DATA

Patent №

US 8,086,799

Granted

2011-12-27

Filed 2008

Owner

NETAPP, INC.

Lab

AI components

1

kr

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

12190511

In a method and apparatus for scalable deduplication, a data set is partitioned into multiple logical partitions, where each partition can be deduplicated independently. Each data block of the data set is assigned to exactly one partition, so that any two or more data blocks that are duplicates of each are always be assigned to the same logical partition. A hash algorithm generates a fingerprint of each data block in the volume, and the fingerprints are subsequently used to detect possible duplicate data blocks as part of deduplication. In addition, the fingerprints are used to ensure that duplicate data blocks are sent to the same logical partition, prior to deduplication. A portion of the fingerprint of each data block is used as a partition identifier to determine the partition to which the data block should be assigned. Once blocks are assigned to partitions, deduplication can be done on partitions independently.

Knowledge representationG06F 3/0641G06F 3/0608G06F 3/061G06F 3/0689G06F 11/1453

AI classification

Knowledge representation0.75
AI hardware0.04
Vision0.03
Evolutionary computation0.00
Natural language0.00
Planning0.00
Machine learning0.00
Speech0.00

Ownership

NETAPP, INC.

assignment · 215520552

Assignors

MONDAL, SHISHIR, KILLAMSETTI, PRAVEEN

On an employer assignment, the assignors are typically the inventors.

From the same owner

© 2026 NYSGPT2525 LLC