Kubernetes Architecture, Operations, and Troubleshooting Questions

How Kubernetes works, how to run it, and how to debug it. Covers control-plane and node components, the scheduler and API server, cluster design, high availability and multi-cluster topologies, and platform-level operations; the workload primitives (pods, deployments, services, controllers), cluster upgrades, and designing Kubernetes as an internal platform; and the operational depth inside a cluster including pod and service networking, ingress and the CNI model, service mesh, persistent volumes and storage classes, resource requests and limits, and systematically diagnosing scheduling, networking, and storage failures. The full architecture-through-day-two-operations span of Kubernetes.

HardTechnical
85 practiced

The Kubernetes API server is experiencing increased request latency. What metrics, logs, and traces would you collect to diagnose whether the bottleneck is etcd, admission controllers, or API server CPU/memory? Provide a prioritized triage checklist and remedial actions for each root cause.

EasyTechnical
39 practiced

Describe the core components of the Kubernetes control plane (API server, etcd, scheduler, controller-manager, cloud-controller-manager). For each component explain its primary responsibility, how it persists or interacts with cluster state, typical failure modes, and what operational metrics you would monitor to detect trouble.

HardTechnical
46 practiced

Explain Pod Disruption Budgets (PDBs). How do PDBs interact with rolling updates, cluster autoscaler evictions, and maintenance operations? Provide scenarios where an incorrect PDB could block upgrades or autoscaling and how to fix those issues.

EasyTechnical
48 practiced

List common cloud and network-backed storage options used with Kubernetes (examples: AWS EBS, AWS EFS, GCE PD, Azure Disk, NFS) and briefly describe trade-offs in terms of performance, durability, multi-node attach, and typical use-cases.

EasyTechnical
85 practiced

A pod remains in Pending and the scheduler does not bind it. Describe commands and checks to determine why the pod is unscheduled: inspect resource requests/limits, node allocatable capacity, taints/tolerations, affinity rules, and namespace quotas. Mention concrete kubectl commands you would run.

Unlock Full Question Bank

Get access to all Kubernetes Architecture, Operations, and Troubleshooting interview questions and detailed answers.

Sign in to Continue

Join thousands of developers preparing for their dream job.