Title :
Big Wins with Small Application-Aware Caches
Author :
López, Julio C. ; O´Hallaron, David R. ; Tu, Tiankai
Author_Institution :
Carnegie Mellon University
Abstract :
Large datasets, on the order of GB and TB, are increasingly common as abundant computational resources allow practitioners to collect, produce and store data at higher rates. As dataset sizes grow, it becomes more challenging to interactively manipulate and analyze these datasets due to the large amounts of data that need to be moved and processed. Application-independent caches, such as operating system page caches and database buffer caches, are present throughout the memory hierarchy to reduce data access times and alleviate transfer overheads. We claim that an application-aware cache with relatively modest memory requirements can effectively exploit dataset structure and application information to speed access to large datasets. We demonstrate this idea in the context of a system named the tree cache, to reduce query latency to large octree datasets by an order of magnitude.
Keywords :
Computational modeling; Computer science; Data analysis; Data engineering; Databases; Delay; Distributed computing; Earthquake engineering; Operating systems; USA Councils;
Conference_Titel :
Supercomputing, 2004. Proceedings of the ACM/IEEE SC2004 Conference
Print_ISBN :
0-7695-2153-3