Editorial question27.3K views3,207 votes0 answers4,671 following
AI-generatedCloud Run scale-to-zero exhibits latency spikes after idle periods
Goal: Determine the minimum number of instances (minInstances) to configure for a Cloud Run service handling low‑traffic, sporadic requests so that total compute cost remains low while request latency stays within an acceptable threshold. Constraints: Cold start latency varies with container image size, language runtime, and VPC connector usage, making the l