Uber Separates Scaling Intent From Execution on Kubernetes Platform
Uber freed over a million CPU cores by letting its orchestrators disagree out in the open. The new ServiceScale CRD lets normal deployments and failover orchestration each write their own scaling intent into Kubernetes, instead of racing each other over one replica count — steady-state provisioning dropped from 2x to 1.3x across 3 million cores and 1.5 million pod launches a day. One-year rollout, zero customer-impacting outages.