Linear Elastic Caching: A New Approach to Cache Management Optimizes Cloud Costs
Google Research and Google Cloud have introduced linear elastic caching, a method that dynamically adjusts cache size to minimize total cost. By framing page eviction as a ski rental problem and using a lightweight machine learning model, the approach reduced memory use by up to 30% in Spanner production servers with only a 0.5% increase in I/O costs.
Google/DeepMind




