PerfIso: Performance Isolation for Commercial Latency-Sensitive Services

Large commercial latency-sensitive services, such as web search, run on dedicated clusters provisioned for peak load to ensure responsiveness and tolerate data center outages. As a result, the average load is far lower than the peak load used for provisioning, leading to resource under-utilization. The idle resources can be used to run batch jobs, completing useful work and reducing overall data center provisioning costs. However, this is challenging in practice due to the complexity and stringent tail-latency requirements of latency-sensitive services. Left unmanaged, the competition for machine resources can lead to severe response-time degradation and unmet service-level objectives (SLOs). This work describes PerfIso, a performance isolation framework which has been used for nearly three years in Microsoft Bing, a major search engine, to colocate batch jobs with production latency-sensitive services on over 90,000 servers. We discuss the design and implementation of PerfIso, and conduct an experimental evaluation in a production environment. We show that colocating CPU-intensive jobs with latency-sensitive services increases average CPU utilization from 21% to 66% for off-peak load without impacting tail latency.

PerfIso: Performance Isolation for Commercial Latency-Sensitive Services

Graph Chatbot

Chat with Graph Search

Local relaxation of residual stress in high-strength steel welded joints treated by HFMI

From PV to EV: Mapping the Potential for Electric Vehicle Charging with Solar Energy in Europe

Beyond the average consumer: Mapping the potential of demand-side management among patterns of appliance usage

Local relaxation of residual stress in high-strength steel welded joints treated by HFMI

From PV to EV: Mapping the Potential for Electric Vehicle Charging with Solar Energy in Europe

Beyond the average consumer: Mapping the potential of demand-side management among patterns of appliance usage