Scalability Patterns and Techniques Questions

Scaling a system to handle growth in traffic and data: horizontal versus vertical scaling, statelessness, sharding and partitioning strategies, read replicas, and connection pooling. Covers capacity estimation, identifying bottlenecks, and the tradeoffs each scaling axis introduces. The general toolkit for taking a design from thousands to millions of users.

EasyTechnical
34 practiced

Your relational database allows a maximum of 500 concurrent connections. Your application runs on 25 identical JVM instances plus 3 background job workers. Walk through how you'd compute a safe default connection-pool size per instance, the factors you'd weigh, and your recommended pool size. What would you monitor, and what would you do if you saw connection saturation in production?

MediumSystem Design
26 practiced

Design sharding for a real-time pub/sub service that has a very large number of channels. Compare client-side sharding, where clients pick their partition, against server-side partitioning, where the broker assigns it. Discuss rebalancing cost, fairness, latency, and client churn from mobile users with intermittent connectivity, and recommend an approach for a client base that is mostly mobile and frequently offline.

MediumTechnical
26 practiced

Design an operational plan and technical implementation to reshard a live sharded database with minimal downtime. Cover choosing the new shard key or shard count, the data-migration strategy (online migration, dual writes, change-data-capture), routing updates, throttling the migration, validation steps, and your rollback procedure.

EasyTechnical
31 practiced

Explain how read replicas for relational databases improve read throughput. Describe the common replication modes (asynchronous versus semi-synchronous) and the operational pitfall of replication lag. What monitoring and safeguards would you put in place to detect and handle a lagging replica?

EasyTechnical
31 practiced

Explain how a CDN works and when you'd reach for one in a global application. Cover edge caching, cache-control headers, TTL strategy, surrogate keys, origin failover, cache invalidation, and how you'd handle dynamic versus static content (signed URLs, edge logic). What are the cost and operational trade-offs?

Unlock Full Question Bank

Get access to all Scalability Patterns and Techniques interview questions and detailed answers.

Sign in to Continue

Join thousands of developers preparing for their dream job.