Patent
Fixed size extents for variable size deduplication segments
Goutham P. Rao,Vinod Jayaraman +1 more
- 08 Mar 2012
18
TL;DR: In this article, variable size deduplication segments are maintained using fixed size extents in a datastore suitcase, where a minor increase in storage overhead removes the need for inefficient recompaction when a segment is removed from the data-store suitcase.
read more
Abstract: Mechanisms are provided for maintaining variable size deduplication segments using fixed size extents. Variable size segments are identified and maintained in a datastore suitcase. Duplicate segments need not be maintained redundantly but can be managed by updating reference counts associated with the segments in the datastore suitcase. Segments are maintained using fixed size extents. A minor increase in storage overhead removes the need for inefficient recompaction when a segment is removed from the datastore suitcase. Fixed size extents can be reallocated for storage of new segments.
read more
Chat with Paper
AI Agents for this Paper
Find similar papers on Google Scholar, PubMed and Arxiv
Write a critical review of this paper
Analyze citations of this paper to find unaddressed research gaps
Citations
Patent
Tamper-protected hardware and method for using same
Kreft Heinz
- 12 Mar 2012
TL;DR: In this article, the tamper-resistant hardware may be used in a transaction system that provides the off-line transaction protocol, such as trusted bootstrapping by means of secure software entity modules, a new use of hardware providing a Physical Unclonable Function (PUF), and the use of a configuration fingerprint of a FPGA used within a tamper resistant hardware.
196
Patent
Packing deduplicated data into finite-sized containers
Michael Hirsch,Thorsten Krause +1 more
- 19 Jun 2012
TL;DR: In this paper, a similarity score is calculated between files that are similarly of the deduplicated data, and the similarity scores are used for grouping the similarly compared files of the similar files into subsets for destaging each of the subsets from a system to one a finite-sized container.
16
Patent
Method for optimizing WAN traffic with deduplicated storage
Sean Rhea
- 16 Jan 2013
TL;DR: In this article, a local proxy caches, in one or more transmitted data files (TDFs) in a deduplicated manner, chunks of one or multiple streams that have been transmitted to a remote proxy, each of the streams being identified by a stream identifier (ID).
10
Patent
Multiple sub-string searching
Chi-Wai Cheung,Ying-Chau R. Mak +1 more
- 02 Sep 2015
TL;DR: In this paper, a method for searching for multiple sub-strings of an original text is provided, wherein the search query includes a plurality of substrings, and a hash array is allocated.
8
Patent
Method for optimizing WAN traffic
Sean Rhea
- 16 Jan 2013
TL;DR: In this article, a local stream store of a local proxy caches one or more streams of data transmitted over the WAN to a remote proxy, where each stream is stored in a continuous manner and identified by a unique stream identifier (ID).
7
References
•Proceedings Article
Sparse indexing: large scale, inline deduplication using sampling and locality
Mark Lillibridge,Kave Eshghi,Deepavali Bhagwat,Vinay Deolalikar,Greg Trezise,Peter Thomas Camble +5 more
- 24 Feb 2009
TL;DR: Sparse indexing, a technique that uses sampling and exploits the inherent locality within backup streams to solve for large-scale backup the chunk-lookup disk bottleneck problem that inline, chunk-based deduplication schemes face, is presented.
Patent
Content aligned block-based deduplication
Manoj Kumar Vijayan,Deepak Raghunath Attarde,Srikant Viswanathan +2 more
- 30 Dec 2010
TL;DR: In this paper, a content alignment system according to certain embodiments aligns a sliding window at the beginning of a data segment by performing a block alignment function on the data within the sliding window.
267
Patent
Compressed data objects referenced via address references and compression references
Allen Samuels
- 19 Apr 2010
TL;DR: In this paper, the authors present a mapping of a virtual storage to a physical storage, where the mapping includes address references from data included in the virtual disk to one or more compressed data objects in the physical disk.
221
Patent
Use of similarity hash to route data for improved deduplication in a storage server cluster
Michael N. Condict
- 26 Oct 2009
TL;DR: In this article, a technique for routing data for deduplication in a storage server cluster includes computing, for each node in the cluster, a value collectively representative of the data stored on the node, such as a "geometric center" of the node.
215
Patent
System and method for organizing data to facilitate data deduplication
Subramanian Periyagaram,Rahul Khona,Dnyaneshwar Pawar,Sandeep Yadav +3 more
- 03 Oct 2008
TL;DR: In this article, a technique for organizing data to facilitate data deduplication includes dividing a block-based set of data into multiple “chunks”, where the chunk boundaries are independent of the block boundaries (due to the hashing algorithm).
201