ALEXKRI.NET
Resume Contact
← Back

Uber Separates Scaling Intent From Execution on Kubernetes Platform

Uber freed over a million CPU cores by letting its orchestrators disagree out in the open. The new ServiceScale CRD lets normal deployments and failover orchestration each write their own scaling intent into Kubernetes, instead of racing each other over one replica count — steady-state provisioning dropped from 2x to 1.3x across 3 million cores and 1.5 million pod launches a day. One-year rollout, zero customer-impacting outages.

Read the source ↗