Courseiva

CKA · topic practice

Troubleshooting practice questions

The Troubleshooting domain (30%) covers diagnosing broken workloads, services, and nodes on a live cluster. You are tested through scenario tasks: inspecting Pending pods, tracing Service endpoint failures, reading kubelet and control-plane logs, and repairing misconfigured networking, scheduling, or storage so the cluster recovers.

Courseiva uses original exam-style practice questions designed for learning and revision. The goal is to understand the concepts, recognise exam patterns, and improve through explanations — not memorise copied exam dumps.

Editorial oversight:Johnson Ajibi· MSc IT Security, IEEE Senior Member
20 questionsDomain: Troubleshooting

What the exam tests

What to know about Troubleshooting

You must systematically isolate faults across pods, Services, and nodes using kubectl describe, logs, events, and node-level systemctl/journalctl checks, then apply the fix. The single most important thing is reading events and status conditions before changing any configuration.

Diagnosing Pending pods via kubectl describe, events, taints, and resource requests

Tracing Service reachability using endpoints, selectors, kube-proxy, and CoreDNS

Reading kubelet state with systemctl status and journalctl -u kubelet

Resolving CNI and node NotReady conditions shown by kubectl get nodes

Watch out for

Common Troubleshooting exam traps

  • ▸Checking application logs first instead of running kubectl describe pod and kubectl get events to see scheduling or image errors
  • ▸Assuming a Service is broken when its selector does not match pod labels, leaving endpoints empty
  • ▸Ignoring node taints, cordons, or NotReady conditions that silently block scheduling and pod networking

Practice set

Troubleshooting questions

20 questions · select your answer, then reveal the explanation

Question 1mediummultiple choice
Read the full Troubleshooting explanation →

A pod named 'web-frontend' is in CrashLoopBackOff. You run 'kubectl logs web-frontend' and see: 'Error: listen tcp :8080: bind: address already in use'. What is the most likely cause and how should you fix it?

Which TWO of the following are valid methods to troubleshoot a pod that is stuck in 'Pending' state?

A developer reports that a newly deployed Deployment named 'web-app' is not serving traffic. The Deployment has 3 replicas, a Service of type ClusterIP, and an Ingress. Which TWO commands should you run first to diagnose the issue?

A pod is in CrashLoopBackOff state. Which command should you use to see the logs of the previous instance?

A pod is stuck in ContainerCreating. Which condition is most likely if `kubectl describe pod` shows 'Failed to create pod sandbox'?

A worker node is marked NotReady. Which two checks are most relevant to diagnose the node's kubelet health? (Choose two.)

A Deployment's pods are failing with 'CrashLoopBackOff'. The container exits with code 1. Which two approaches will help identify the issue? (Choose two.)

You have a Deployment that is not scaling up beyond 1 replica despite setting replicas: 3. Which of the following could be the cause? (Select all that apply)

A Pod is in ImagePullBackOff state. Which of the following are valid troubleshooting steps? (Select two.)

Drag and drop the steps to set up a PersistentVolumeClaim for a pod into the correct order.

Drag or tap steps into the slots.

Steps
Order
1Step 1
2Step 2
3Step 3
4Step 4
5Step 5

Match each scheduling concept to its description.

Drag a concept onto its matching description — or click a concept then click the description.

Concepts
Matches

Simple label-based constraint

Expressive scheduling rules using labels

Repels Pods unless they tolerate the taint

Allows a Pod to be scheduled on a tainted Node

Determines scheduling precedence based on priority class

Match each logging/monitoring component to its role.

Drag a concept onto its matching description — or click a concept then click the description.

Concepts
Matches

Log aggregator and forwarder

Metrics collection and alerting system

Dashboard and visualization tool

Exposes cluster object metrics

Provides resource usage metrics for autoscaling

Question 13hardmultiple choice
Read the full Troubleshooting explanation →

You want to test network connectivity from pod A to pod B in the same namespace. Which command would you run from within pod A?

Which TWO of the following are valid methods to view the logs of a container that has terminated?

Question 15hardmultiple choice
Read the full Troubleshooting explanation →

You run 'kubectl get pods' and see a pod in 'CrashLoopBackOff'. 'kubectl logs pod' shows no output, but 'kubectl logs --previous pod' shows an error. Why might the current logs be empty?

Question 16mediummultiple choice
Read the full Troubleshooting explanation →

A pod is stuck in Pending state. 'kubectl describe pod' shows '0/2 nodes are available: 1 node(s) had taint that the pod didn't tolerate, 1 node(s) didn't match pod anti-affinity rules'. What should you check?

Question 17easymultiple choice
Read the full Troubleshooting explanation →

Which command shows all events in the cluster sorted by timestamp?

Question 18mediummultiple choice
Read the full Troubleshooting explanation →

You suspect that the kube-scheduler is not running. Which command checks the scheduler's health?

Which TWO of the following are valid reasons for a pod to be in the Pending state? (Choose two)

Question 20easymultiple choice
Read the full Troubleshooting explanation →

You need to check the logs of a container that previously ran but has crashed. Which command would you use?

Free account

Track your progress over time

Create a free account to save your results and see which topics improve across sessions.

Focused Troubleshooting sessions

Start a Troubleshooting only practice session

Every question in these sessions is drawn from the Troubleshooting domain — nothing else.

Related practice questions

Related CKA topic practice pages

Move into related areas when this topic feels solid.

Frequently asked questions

What does the CKA exam test about Troubleshooting?
You must systematically isolate faults across pods, Services, and nodes using kubectl describe, logs, events, and node-level systemctl/journalctl checks, then apply the fix. The single most important thing is reading events and status conditions before changing any configuration.
How should I use these practice questions?
Select your answer before revealing the explanation. Then read why each option is right or wrong — this active recall approach builds retention far faster than re-reading notes.
Can I practise just Troubleshooting questions in a focused session?
Yes — the session launcher on this page draws every question from the Troubleshooting domain. Use a 10-question session first to gauge your baseline, then move to 20 or 30 once the weak spots are clear.
Where can I practise other CKA topics?
Use the topic links above to move to related areas, or go back to the CKA question bank to see all topics.
Are these real exam questions or dumps?
These are original practice questions written to test the same concepts the CKA exam covers. They are not copied from any real exam or dump site.