Kubernetes Interview Questions — All Categories
169 questions tagged with Kubernetes as a technology, across every category it appears in
Related categories:Argo CDAWSAzureCI/CDCloud ArchitectureContainersDockerGitOpsHelmJenkinsKubernetesLinuxNetworkingSystem DesignYAML
How would you design a StorageClass for a high-IOPS database workload versus a cheap-capacity logging workload?
IntermediateKubernetes6 min
A production workload needs persistent storage, and its pods may be rescheduled to different nodes — how do you design storage so data survives that?
BeginnerKubernetes6 min
What does volumeBindingMode: WaitForFirstConsumer actually solve, and what breaks in a multi-zone cluster if you don't set it?
AdvancedKubernetes6 min
How would you set up alerting to catch a CrashLoopBackOff-class issue before it reaches production traffic, rather than discovering it via a user-facing outage?
IntermediateKubernetesPrometheus7 min
A pod goes into CrashLoopBackOff immediately after you roll out a ConfigMap change, but only in one namespace. How do you investigate it?
IntermediateKubernetesContainers10 min
How would your investigation differ if a Pod entered ImagePullBackOff instead of CrashLoopBackOff?
BeginnerKubernetes6 min
How do liveness and readiness probes interact with a Pod that's already crash-looping on startup?
IntermediateKubernetes6 min
How would you decide between a Deployment, a StatefulSet, and a DaemonSet for three different real services (a stateless API, a database, a node agent)?
IntermediateKubernetes6 min
What's the difference between a CronJob's concurrencyPolicy: Forbid and Replace, and what production incident does picking the wrong one cause?
IntermediateKubernetes5 min
A CronJob has been silently creating thousands of failed Jobs over several days — how did this happen, and how would you prevent it?
AdvancedKubernetes7 min
A DaemonSet pod is missing from exactly one node while running fine everywhere else — how do you find out why?
IntermediateKubernetes6 min
A Job is supposed to run to completion exactly once, but it created multiple pods — why, and is that actually a bug?
IntermediateKubernetes6 min
How would you design a Job for a task that must never run twice, even if a pod fails partway through (e.g., a billing charge)?
ExpertKubernetes8 min
A Deployment's new pods keep getting scheduled but immediately evicted, while old pods keep running fine — what changed?
AdvancedKubernetes7 min
A rolling update to a Deployment is stuck at 50% — how do you determine whether it's a bad readiness probe, insufficient capacity, or a PodDisruptionBudget blocking it?
AdvancedKubernetes8 min
What's actually different at the API level between kubectl rollout restart and deleting all of a Deployment's pods manually?
BeginnerKubernetes5 min
How would you safely roll out a breaking change to a DaemonSet running a critical node-level agent across a large production cluster?
AdvancedKubernetes7 min
A StatefulSet pod is deleted but isn't recreated with the same identity fast enough — what's actually blocking it?
AdvancedKubernetes7 min
Why does a StatefulSet's rolling update behave completely differently from a Deployment's, and why can that ordering guarantee become a problem mid-incident?
IntermediateKubernetes6 min
A containerized process shows only 40% average CPU usage, well under its configured limit, but application metrics show frequent latency spikes correlating with CPU throttling events. How is this possible?
AdvancedLinuxKubernetes8 min
How does NodeLocal DNSCache in Kubernetes actually reduce DNS-related failures, mechanically?
AdvancedKubernetesDNS7 min
Design a centralized logging platform for an organization running roughly 500 microservices across multiple Kubernetes clusters, where engineers currently can't find logs during incidents.
ExpertKubernetesObservability14 min
Design a deployment orchestration system that lets any of your organization's 200 services safely adopt canary or blue-green deployments, without every team building their own rollout automation from scratch.
ExpertKubernetesPlatform Engineering14 min
Design a self-service CI/CD platform for an engineering org with roughly 100 teams, each owning multiple services, without a central platform team becoming a bottleneck.
ExpertCI/CDKubernetesPlatform Engineering15 min