SYSTEM AND METHOD FOR MANAGING LARGE FILESYSTEM-BASED CACHES

Embodiments disclosed herein utilize statistical approximations to manage large filesystem-based caches based on imperfect information. When removing entries from a large cache, which may have a million or more entries, the cache manager does not need to find the absolutely oldest entry that has bee...

Full description

Saved in:
Bibliographic Details
Main Authors SCHEEVEL MARK R, FUNG KINUNG
Format Patent
LanguageEnglish
Published 12.01.2012
Subjects
Online AccessGet full text

Cover

Loading…
More Information
Summary:Embodiments disclosed herein utilize statistical approximations to manage large filesystem-based caches based on imperfect information. When removing entries from a large cache, which may have a million or more entries, the cache manager does not need to find the absolutely oldest entry that has been accessed the least recently. Instead, it suffices to find an entry that is older than most. In embodiments disclosed herein, statistical sampling of the cache is performed to produce models of different properties of the cache, including the number of entries, distribution of access times, distribution of entry sizes, etc. The models are then used to guide decisions that involve those properties. The size of the samples can be adjusted to balance the cost of acquiring the samples against the confidence level of the models produced by the samples. To achieve randomness, entries are stored using prefixes of addresses generated via a message-digest function.
Bibliography:Application Number: US201113237236