Load Balancing and Traffic Management Questions

Distributing requests across capacity: load-balancing algorithms (round-robin, least-connections, consistent hashing), L4 versus L7 balancing, health checks, and traffic shaping. Covers sticky sessions, canary and blue-green routing, rate limiting, and graceful draining. The traffic-distribution layer that keeps a scaled system balanced and available.

HardSystem Design
47 practiced

Design a Layer 7 load balancer that provides session affinity using consistent hashing on a session cookie. It must support health checks and rebalance sessions gracefully when nodes are added or removed. Discuss hash ring maintenance, virtual nodes, and how you would drain and migrate sessions without dropping in-flight traffic.

MediumSystem Design
38 practiced

Design a monitoring and alerting setup for a load balancer tier so you catch unhealthy backends before users see errors. Which metrics would you include and why, what alert thresholds would you propose, and how would you differentiate severity to avoid alert fatigue?

EasyTechnical
42 practiced

What is traffic mirroring (shadowing) at the load balancer level? Explain a safe approach to mirror production traffic to a staging cluster for testing, how you would sample it to limit load, and the risks, such as side effects in downstream systems or data contamination.

HardTechnical
77 practiced

You observe rising p99 latency on your load balancer while backends show stable p95 latency and healthy CPU. Walk through a troubleshooting checklist covering the network, the LB proxies themselves, TLS handshakes, accept/queue backlogs, kernel limits, and client behavior. What instrumentation would you add to pinpoint the root cause?

MediumSystem Design
43 practiced

Design the load balancing layer for a service with no sticky-session requirement and autoscaling backends. Include L4-vs-L7 placement, service discovery integration, TLS termination, health checks, algorithm choice, and how you would validate the design with load testing. Then explain what changes if the service instead needs to be internet-facing across three regions with a 99.99% uptime target: edge proxies, regional failover and routing, and connection draining during rolling deploys.

Unlock Full Question Bank

Get access to all Load Balancing and Traffic Management interview questions and detailed answers.

Sign in to Continue

Join thousands of developers preparing for their dream job.