Load Balancing and Traffic Management Questions

Distributing requests across capacity: load-balancing algorithms (round-robin, least-connections, consistent hashing), L4 versus L7 balancing, health checks, and traffic shaping. Covers sticky sessions, canary and blue-green routing, rate limiting, and graceful draining. The traffic-distribution layer that keeps a scaled system balanced and available.

MediumTechnical
43 practiced

What observability signals (metrics, logs, traces) would justify automatically removing an instance from load balancer rotation before users see errors? For each signal, describe roughly what threshold or pattern would trigger removal and one way that signal could produce a false positive.

HardSystem Design
41 practiced

Design an automated control loop that shifts a percentage of production traffic to a canary over a fixed window and rolls back automatically on an SLO breach. Cover how you would smooth the weight changes, what guardrails you'd set (minimum observation windows, maximum error thresholds), and how the rollback itself executes quickly and safely.

HardSystem Design
51 practiced

You need to tune health checks across thousands of instances so that rolling deployments don't cause flapping (healthy instances briefly failing checks and being pulled from rotation) while real failures are still caught quickly. Walk through your recommended probe interval, timeout, consecutive-failure threshold, and jitter strategy, and how you would validate and roll out these settings safely.

EasyTechnical
50 practiced

Explain the difference between readiness and liveness health checks (and startup checks, where supported). Design the checks for a database-backed web service (for example, covering DB connectivity and queue backlog), and explain how a misconfigured probe can cause cascading restarts or route traffic into a blackhole.

HardTechnical
45 practiced

Compare connection management for HTTP/2 and gRPC traffic behind a Layer 7 load balancer: long-lived multiplexed connections versus ephemeral short-lived ones. How does connection pooling and multiplexing change throughput and resource usage compared to HTTP/1.1 keep-alive, what per-connection limits would you tune, and how should the load balancer measure load and apply backpressure to avoid head-of-line effects?

Unlock Full Question Bank

Get access to all Load Balancing and Traffic Management interview questions and detailed answers.

Sign in to Continue

Join thousands of developers preparing for their dream job.