>DevOps Interview KB

Kubernetes Interview Questions — All Categories

169 questions tagged with Kubernetes as a technology, across every category it appears in

How would you design a StorageClass for a high-IOPS database workload versus a cheap-capacity logging workload?

IntermediateKubernetes6 min

A production workload needs persistent storage, and its pods may be rescheduled to different nodes — how do you design storage so data survives that?

BeginnerKubernetes6 min

What does volumeBindingMode: WaitForFirstConsumer actually solve, and what breaks in a multi-zone cluster if you don't set it?

AdvancedKubernetes6 min

How would you set up alerting to catch a CrashLoopBackOff-class issue before it reaches production traffic, rather than discovering it via a user-facing outage?

IntermediateKubernetesPrometheus7 min

A pod goes into CrashLoopBackOff immediately after you roll out a ConfigMap change, but only in one namespace. How do you investigate it?

IntermediateKubernetesContainers10 min

How would your investigation differ if a Pod entered ImagePullBackOff instead of CrashLoopBackOff?

BeginnerKubernetes6 min

How do liveness and readiness probes interact with a Pod that's already crash-looping on startup?

IntermediateKubernetes6 min

How would you decide between a Deployment, a StatefulSet, and a DaemonSet for three different real services (a stateless API, a database, a node agent)?

IntermediateKubernetes6 min

What's the difference between a CronJob's concurrencyPolicy: Forbid and Replace, and what production incident does picking the wrong one cause?

IntermediateKubernetes5 min

A CronJob has been silently creating thousands of failed Jobs over several days — how did this happen, and how would you prevent it?

AdvancedKubernetes7 min

A DaemonSet pod is missing from exactly one node while running fine everywhere else — how do you find out why?

IntermediateKubernetes6 min

A Job is supposed to run to completion exactly once, but it created multiple pods — why, and is that actually a bug?

IntermediateKubernetes6 min

How would you design a Job for a task that must never run twice, even if a pod fails partway through (e.g., a billing charge)?

ExpertKubernetes8 min

A Deployment's new pods keep getting scheduled but immediately evicted, while old pods keep running fine — what changed?

AdvancedKubernetes7 min

A rolling update to a Deployment is stuck at 50% — how do you determine whether it's a bad readiness probe, insufficient capacity, or a PodDisruptionBudget blocking it?

AdvancedKubernetes8 min

What's actually different at the API level between kubectl rollout restart and deleting all of a Deployment's pods manually?

BeginnerKubernetes5 min

How would you safely roll out a breaking change to a DaemonSet running a critical node-level agent across a large production cluster?

AdvancedKubernetes7 min

A StatefulSet pod is deleted but isn't recreated with the same identity fast enough — what's actually blocking it?

AdvancedKubernetes7 min

Why does a StatefulSet's rolling update behave completely differently from a Deployment's, and why can that ordering guarantee become a problem mid-incident?

IntermediateKubernetes6 min

A containerized process shows only 40% average CPU usage, well under its configured limit, but application metrics show frequent latency spikes correlating with CPU throttling events. How is this possible?

AdvancedLinuxKubernetes8 min

How does NodeLocal DNSCache in Kubernetes actually reduce DNS-related failures, mechanically?

AdvancedKubernetesDNS7 min

Design a centralized logging platform for an organization running roughly 500 microservices across multiple Kubernetes clusters, where engineers currently can't find logs during incidents.

ExpertKubernetesObservability14 min

Design a deployment orchestration system that lets any of your organization's 200 services safely adopt canary or blue-green deployments, without every team building their own rollout automation from scratch.

ExpertKubernetesPlatform Engineering14 min

Design a self-service CI/CD platform for an engineering org with roughly 100 teams, each owning multiple services, without a central platform team becoming a bottleneck.

ExpertCI/CDKubernetesPlatform Engineering15 min