DEAP Cache learns prefetching, eviction, and admission using machine learning.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Service-induced congestion in memory-constrained LLM serving
New algorithm improves cache management with delayed feedback and decaying costs.
New algorithm reduces costs and latency for large language model inference.
Parrot learns optimal cache replacement policies using imitation learning.
A new algorithm reduces memory usage for deep learning models.
Critical to evaluating the capacity, scalability, and availability of web systems are realistic web traffic generators. Web traffic generation is a classic research problem, no generator accounts for the characteristics of web robots or crawlers that are now the dominant source of traffic to a web server. Administrator…
This study investigates the use of reinforcement learning to guide a general purpose cache manager decisions. Cache managers directly impact the overall performance of computer systems. They govern decisions about which objects should be cached, the duration they should be cached for, and decides on which objects to ev…
Regularizes decision trees to reduce inference time by up to 4x with minimal accuracy loss.