API Pods Restart Across the Rollout
Every api replica is in CrashLoopBackOff and the Deployment has shown 0 of 3 available for ten minutes.
Free simulated incident practice
A workload can exist without being ready to serve traffic. These Kubernetes exercises let you practise moving from an alert to the relevant workload evidence in a simulated terminal. Start a brief, inspect the system, make a justified change and check whether the incident has recovered.
Describe where the workload stops making progress before choosing a fix. A pod waiting to start, a repeatedly exiting process and an application failing readiness represent different investigation paths.
For a reachability problem, compare the application state with the resources that route traffic to it. For a rollout problem, examine the change and its observed effect. Keep your next check tied to a question you can answer.
A resource showing a healthy status is one observation. Check the behaviour the alert described, complete the scenario recovery checks and explain the smallest change that restored it. The aim is a repeatable investigation, not memorising one command.
Open a brief to see the symptoms, then launch the incident. The solution is yours to find.
Every api replica is in CrashLoopBackOff and the Deployment has shown 0 of 3 available for ten minutes.
The migrated webapp has shown 0 of 1 available replicas from the moment it was applied to the new cluster four minutes ago.
Both order-scoring worker pods die within seconds of every start, and the orders.score backlog keeps climbing.
Ten minutes after tonight's release, the checkout Deployment has 0 of 2 replicas available and neither pod has ever started.
Calls to the catalog Service are refused even though both catalog pods are Running and Ready.
The payment pod has been stuck before startup for ten minutes, leaving the payment Deployment with zero ready replicas.
The report Deployment has run zero workers for ten minutes; its only pod is still waiting to be scheduled.
Both gateway pods have been Running for ten minutes and neither has ever become Ready, so no gateway replica is eligible for traffic.
Both pods from the new ledger revision crash on startup, leaving one old pod to carry every ledger write.
The payments deploy job's service account gets Forbidden on every Deployment patch, so the ledger change it carries cannot roll out.
The endpoint sensor on all five fleet-east nodes has been crash-looping.
FixOps runs simulated systems and a terminal designed for its scenarios. It does not provision a real server or cluster. Use it to rehearse investigation and recovery, and use a real environment when you need unrestricted tooling or production experience.
There is no interview pass guarantee or certification. You can start as a guest in the browser or use the Android app. If the concepts feel unfamiliar, work through a learning track before trying another incident.
Free, instant, and it works on your phone. No signup: start as a guest and save your progress later.