Skip to main content
All of the catalog
Scenario

Autoscaling Under Load

Watch KEDA scale go-api on Prometheus RPS: a spike drives it from 1 to several replicas, then cooldown brings it back. The flagship 'autoscaling actually works' demo, verified under traffic from the load generator.

ScalabilityVerifiedk3dkind
Definition on GitHub

What you'll do

  • Declare metric-driven autoscaling for go-api with a KEDA ScaledObject
  • Drive a traffic spike and watch replicas scale up on Prometheus RPS
  • Confirm latency stays within SLO while scaled, then scales back on cooldown

Stages

  1. 1autoscaler

    Declare the ScaledObject and the replicas-vs-RPS dashboard

    go-api-scaledobjectautoscaling-dashboard

Prerequisites

These are installed into the lab cluster for you — listed so you know what the scenario actually depends on.

ingressmonitoring/metricsmonitoring/grafanaautoscaling/kedago-api

The incident field notes

One real Kubernetes failure a week — the symptom, the commands that found it, and the fix. Written from actual lab runs, not from memory.

You'll get the Kubernetes Incident Response Field Guide, plus occasional emails about new scenarios, posts and paid offerings such as courses and workshops. Unsubscribe any time.