Patching and Upgrade Management at Scale Questions
Keeping fleets of systems current: OS and software patching, coordinated fleet updates, version upgrades, and system migrations. Covers planning and rolling out upgrades across many machines, managing change through migrations, and balancing currency against stability. Focuses on maintaining and evolving existing running systems rather than shipping application releases.
You must patch database servers in a replicated PostgreSQL cluster hosting critical transactional data with strict RTO/RPO. Design a rolling patch strategy that minimizes downtime and ensures consistency for schema or binary-level changes. Address backups, schema migrations, leader/follower promotion, and explicit rollback procedures.
You're responsible for patching kernels across a fleet with strict uptime requirements. Evaluate patching strategies: rolling reboots, kernel live patching solutions (kpatch, ksplice, livepatch), use of HA and failover, and boot segmentation. Discuss testing requirements, limitations of live patching (what it cannot change), and fallback plans if a patch introduces regressions.
That is every published Patching and Upgrade Management at Scale question for Systems Administrator so far. Browse the other topics in this category, or practice this one interactively.