Skip to content
EgyKode
11 · Operating itLab 58 / 59
Guided labkubernetesDestructive

Node Drain, Upgrade & Recovery

Take a node out of service without taking the application with it, and find out which workloads were never ready for it.

Time
55 min
Level
Advanced
Objectives
4 objectives
Cost
Free

Where this fits in the platform

This lab adds

  • A node taken out of service with the application still up

Which lets you

Before you start

You will need

  • kind (multi-node)
  • kubectl 1.28+

You do not need these already — the lab environment below provides them.

You will be able to

  • Cordon and drain a node safely
  • Protect availability during voluntary disruption with a PDB
  • Recognise workloads that cannot survive rescheduling

CostFree

— a multi-node kind cluster. See the setup note; a single node cannot demonstrate rescheduling.

Nothing to pay in the browser. Open the terminal runs this against a simulated cloud — the same API calls and the same commands, with no account and no bill. The figure above applies only if you build it in your own.

How to clean up

The scenario#

The cluster needs a Kubernetes upgrade. That means taking each node out of service in turn, and the first one you try teaches you which of your workloads were only ever running by luck.

This evicts running workloads. Use a throwaway cluster.

Hands-on environment

Run this lab in a real terminal, free and in your browser. The environment is temporary and yours alone — break it as much as you like.

Open the terminal

Opens in Killercoda, in a new tab — keep this page open for the steps.

Run it on your own machine

Run this lab on your own machine. One command starts the environment, with everything the lab needs already installed:

You will need:

  • docker
  • kubectl
  • kind
git clone https://github.com/EgyKode/EgyKode-lab.git
cd EgyKode-lab
./egykode start k8s
./egykode shell

You need Docker and Git installed. Everything else runs inside the environment. The first start downloads it and takes a few minutes; later starts are seconds.

Not sure what you already have? Run: npm run doctor — it checks and changes nothing.

Anything you tick here is your own record. EgyKode cannot see inside that terminal, so the success criteria stay self-assessed even when the environment checks your work for you.

Something to keep alive

Step 1 of 6

Setup: more than one node#

Terminal
cat <<'EOF' | kind create cluster --name ops --config=-
kind: Cluster
apiVersion: kind.x-k8s.io/v1alpha4
nodes:
  - role: control-plane
  - role: worker
  - role: worker
EOF
 
kubectl get nodes

A single-node cluster cannot demonstrate any of this — there is nowhere to reschedule to, and every Pod simply goes Pending.

Clean up#

Run this even if you did not finish.

DestructiveThis removes real resources. Check which environment you are in first.

Terminal
kubectl uncordon $(kubectl get nodes -o name)
kubectl delete deployment web --ignore-not-found
kind delete cluster --name ops

Cost of this lab: Free — a multi-node kind cluster. See the setup note; a single node cannot demonstrate rescheduling.

Success criteria

0 of 4

The concept behind it

Ready to try it without help?Do the challenge

Next up

Lab 58 of 59 on the project path

Production Capstone: Build, Deploy & Operate the PlatformEverything, once, with no instructions — then keep it running while it is deliberately broken.240 minAdvancedBillable — destroy resources when you finish

Previous: Terraform Drift & State Recovery