Tight Latency Guarantees for Weighted Caching with Delayed Hits
Driven by the massive parallelism of modern multicore architectures and high-bandwidth networks, system throughput has increasingly outpaced physical latency limits. In such high-throughput environments, the ratio Z between retrieval latency and the request inter-arrival time becomes a dominant performance factor. We s...