Kubernetes Probes: Liveness, Readiness, and Startup
Use better Kubernetes probes by choosing the right signal, tuning thresholds, and avoiding false restarts, traffic drops, and noisy rollouts.
35 practical articles on Kubernetes, covering setup, troubleshooting, and production decisions.
Use better Kubernetes probes by choosing the right signal, tuning thresholds, and avoiding false restarts, traffic drops, and noisy rollouts.
Safely inspect a live Pod without baking debugging tools into production images.
A concrete rollout path for Kubernetes NetworkPolicy: start with default deny, whitelist DNS and key dependencies, and avoid breaking production traffic.
Combine Deployment rollingUpdate settings with PodDisruptionBudgets to keep availability during upgrades and node maintenance.
Deploy SGLang and vLLM Prefill/Decode roles with LWS DisaggregatedSet, NIXL startup commands, validation steps, and production safeguards.
How CPU/memory requests and limits actually affect scheduling, throttling, OOMKills, and autoscaling.
All posts in reverse chronological order.
Build a repeatable local Kubernetes cluster with Minikube, then verify networking, image loading, and reset workflows.
A concrete introduction to Kubernetes: what problems it solves, what it does not solve, and how desired state and controllers work in real clusters.
Understand Kubernetes architecture from the control plane to worker nodes, including the API server, scheduler, controllers, kubelet, and reconciliation loops.
A step-by-step Kubernetes learning path covering core concepts, workloads, networking, storage, and troubleshooting so you can study in the right order.
Understand Pods, Deployments, Services, and Namespaces through the way Kubernetes actually reconciles and fails, then verify changes with real commands.