Practice SK0-005 troubleshooting questions with full explanations on every answer.
Start practicing
troubleshooting — choose a session length
Free · No account required
Click any question to see the full explanation and answer options, or start a focused practice session above.
A RAID 5 array is degraded due to a failed disk. What is the best practice for recovery?
2A technician is troubleshooting a server that fails to boot. The server displays a 'Boot Device Not Found' error. The technician verifies that the hard drives are spinning and appear in the RAID controller's BIOS. What should the technician check next?
3A server administrator is troubleshooting a performance issue on a VMware vSphere host. The host has multiple VMs, and one VM is experiencing high latency on its virtual disk. The administrator checks the datastore and sees it is a RAID 5 array of 4 SAS drives with 50% utilization. Latency for the VM's virtual disk is reported as 200ms. Which action would most likely improve performance?
4A technician is troubleshooting a server that fails to boot after a scheduled power outage. The server is connected to a UPS. Upon pressing the power button, the server powers on briefly (fans spin, lights flash) and then shuts down after 2 seconds. The POST does not complete. The server has a dual power supply with each connected to a separate PDU. Which of the following is the MOST likely cause?
5A server is running out of disk space on the system drive (C:). The server is running Windows Server 2019. The IT manager wants to add more space without downtime. The server has one free SATA port and one available drive bay. Which of the following is the BEST solution?
6You are a server engineer for a financial services firm. The company recently deployed a new HP ProLiant DL380 Gen10 server running Windows Server 2022 with SQL Server 2019. The server has 2 Intel Xeon Gold processors, 128GB RAM, and a Smart Array P408i-p controller managing two RAID 1 arrays: one for OS (two 300GB 10K SAS) and one for data (four 600GB 10K SAS). After one month, the OS array reports a predictive failure on one drive. You replace the drive via hot-swap, and the RAID controller rebuilds. However, the server now experiences random system crashes with Event ID 1001 (BugCheck) and the SQL database occasionally becomes corrupt requiring restore from backup. The server's RAM has been tested with HP's diagnostic tool and passed, and the CPU temperature is normal. The RAID controller log shows no errors during the rebuild but occasional 'Parity errors' logged before the drive replacement. Which of the following is the MOST likely cause of the current instability?
7A server administrator is troubleshooting a physical server that randomly crashes once or twice a week. The administrator has already verified that the power supply is functioning correctly and has checked the event logs for critical errors. According to standard troubleshooting methodology, what should the administrator do next?
8A technician is troubleshooting a server that fails to boot after a power outage. The server displays a "Non-system disk or disk error" message. Which action should the technician take first?
9A server administrator notices that a database server is responding slowly to queries. The CPU utilization is at 30%, memory at 40%, and disk latency is normal. Which of the following should the administrator check NEXT?
10A server in a virtualized environment is experiencing VM performance degradation. The host server has adequate resources. The hypervisor logs show excessive 'ready' time for the CPU. Which of the following is the BEST action to resolve this issue?
11A technician is troubleshooting a server that is experiencing high disk I/O wait times. The disk queue length is consistently above 10. Which of the following is the MOST likely cause?
12Refer to the exhibit. A server is running slowly. Based on the memory statistics, what is the MOST likely issue?
13Refer to the exhibit. A technician is troubleshooting a server that is experiencing data corruption. What is the MOST likely cause?
14A server fails to boot after installing new memory. The POST beep code indicates a memory error. What is the most likely cause?
15A server with dual redundant power supplies shuts down unexpectedly. One power supply has a solid amber LED. What is the most likely cause?
16Refer to the exhibit. What is the most likely cause of this error?
17A technician is troubleshooting a server that fails to boot. The server powers on, fans spin, but no video output and no beep codes. The technician reseats the RAM and GPU, but the issue persists. Which of the following should the technician check NEXT?
18A server administrator is troubleshooting a network connectivity issue on a server that has recently been moved to a different rack. The server can ping its own IP address but cannot ping the default gateway. Which of the following is the MOST likely cause?
19A data center technician is troubleshooting a server that is overheating and shutting down intermittently. The server is a 2U rackmount with six fans at the front and a power supply with an integrated fan at the rear. The technician checks the ambient temperature (72°F) and verifies that the server intake temperature is normal. The server's system logs show 'CPU temperature threshold exceeded' before each shutdown. The technician has replaced the thermal paste on the CPU and reseated the heat sink, but the issue persists. Which of the following should the technician do NEXT?
20A server fails to boot with an 'Operating System Not Found' error. The BIOS detects the hard drive. What is the MOST likely cause?
21A virtualized server running multiple VMs experiences a sudden performance degradation during peak hours. Storage latency spikes and CPU ready time increases significantly. The hypervisor shows high memory overcommitment but no swapping. Which action would most likely resolve the issue without adding hardware?
22After a power outage, a server fails to boot and displays a "Missing operating system" error. The server uses RAID 1 for the OS disk. The administrator verifies both drives are present in the RAID configuration utility but one drive is marked as failed. What should the administrator do FIRST?
23An administrator notices that a Linux server's /var/log/messages file is filled with repeated "eth0: link up" and "eth0: link down" entries every few seconds. The server is connected to a managed switch. Which of the following is the most likely cause?
24A server administrator is troubleshooting an issue where a database server's performance degrades every night at 2 AM. Resource monitor shows high disk I/O and CPU usage during that time. There are no scheduled tasks on the server. Which of the following should the administrator investigate FIRST?
25Following a firmware update on a server's RAID controller, the server fails to boot and reports "No boot device found." The RAID array status shows healthy in the controller BIOS. Which THREE of the following actions should the administrator take to resolve the issue? (Choose three.)
26Refer to the exhibit. A server administrator receives reports that an internal web server is inaccessible. After connecting locally, the administrator runs a command and receives the following output. Which of the following commands would best resolve the issue?
27A database server with 64 GB of RAM and RAID 5 storage is experiencing intermittent performance degradation. System monitoring shows constant 95% memory utilization, high disk queue length, and frequent page faults. The server is running a critical application that cannot be restarted during business hours. Which action will provide the BEST permanent resolution?
28Refer to the exhibit. A technician receives an alert from the monitoring system showing the error in the exhibit. The server is still online, but performance has degraded. Which of the following is the MOST likely cause and appropriate action?
29A server configured with a RAID 5 array and a hot spare experiences intermittent crashes under heavy disk I/O. A technician suspects a failing drive. Which of the following should the technician do FIRST?
30A server uses NIC teaming with LACP (802.3ad) for load balancing and failover. After replacing a failed switch with a new switch of the same model, the server loses network connectivity. The switch ports show no activity. Which of the following is the MOST likely cause?
31After installing additional RAM modules in a server, the BIOS displays only a portion of the installed memory. The modules are identical in speed and size but from different manufacturers. Which of the following is the BEST initial troubleshooting step?
32A virtualization administrator notices that a critical VM on an ESXi host has high CPU Ready time (%RDY) according to performance charts. The VM has 8 vCPUs assigned, and the host has two physical CPUs with 8 cores each. Other VMs on the host are performing normally. Which of the following actions would MOST likely resolve the high CPU Ready time for this VM?
33A server's performance degrades significantly during peak usage hours. Monitoring shows high memory utilization and consistently high disk I/O. Which of the following should the technician check FIRST?
34A server running a critical database application crashes and fails to boot with the error message 'Operating System not found.' The server is equipped with a hardware RAID 5 array consisting of three identical disks. A technician suspects a disk failure. What is the MOST likely cause of this error?
35The Coho Vineyard company operates a two-node Windows Server failover cluster to provide high availability for a critical SQL database. Node A has been running the database workload without issues, while Node B was taken offline two weeks ago due to a memory module failure. The faulty memory was replaced, and Node B was powered on. The administrator verified that Node B boots correctly, the network links are up, and the cluster service is running. However, the database role fails to start on Node B. Checking the cluster logs reveals that Node B cannot join the cluster, and the event 'Cluster service has lost quorum' is recorded. The administrator confirms both nodes can ping each other and access the shared SAN storage, but the cluster disk resource appears as 'Offline' on Node B. What should the administrator do to resolve the issue and bring the database online?
36A database administrator notices that the primary SQL Server instance has been experiencing severe performance degradation over the past hour. The server is a virtual machine with 8 vCPUs and 64 GB of RAM, and it normally operates with around 40% CPU utilization. Currently, Task Manager shows 100% CPU usage, with the sqlservr.exe process consuming nearly all cycles. The database response times have increased tenfold, and some queries are timing out. There are no scheduled maintenance jobs running, and the transaction log backup completed successfully earlier. The DBA checks active sessions and finds one long-running ad-hoc query from a new reporting tool that was deployed this morning. The query is performing a cross join on two large tables with missing WHERE clauses, causing a Cartesian product. Ending the session will roll back the query and free resources.
37An organization has a server with a hardware RAID 5 array consisting of four 2 TB SAS drives. The server hosts a critical database and is configured with a hot spare. During routine monitoring, the storage administrator discovers that one drive has failed and the hot spare has automatically taken over, with the array currently rebuilding. However, the rebuild process repeatedly fails at approximately 30% completion, and the RAID controller logs show I/O errors on the hot spare drive. The failed drive was replaced with an identical model from inventory, and the rebuild was restarted, but it again fails at the same point. All other drives show healthy SMART status, and the server's firmware and RAID controller firmware are up to date. The database is still online and functioning, but the array is running in a degraded state, putting data at risk. The server is located in a remote data center without onsite staff, and a maintenance window is scheduled in three days.
38Users report slow file server performance. The administrator suspects a single process is consuming excessive CPU and wants to analyse its threads and handles. Which tool should be used?
39Refer to the exhibit. The Print Spooler service on a critical Windows Server 2019 fails to start at every boot. The server was recently updated with the latest cumulative patch. Which of the following is the MOST likely cause?
40A server cannot connect to a specific network share. The administrator can successfully ping the server's IP address from the client. Which of the following is MOST likely causing the issue?
41A server running a critical application fails to boot with the error: 'Boot device not found.' The server uses UEFI and a RAID 5 array for the OS. The administrator verifies that all disks are present and powered. Which of the following should the administrator check FIRST?
42After applying a Windows security patch, a server fails to boot and displays 'Bootmgr is missing'. The server is UEFI-based. Which of the following is the MOST efficient way to resolve the issue?
43A server that uses iSCSI storage has suddenly lost connectivity to all LUNs. The network team confirms no changes have been made. Which of the following is the FIRST step in troubleshooting this issue?
44A server in the datacenter fails to power on after a scheduled power maintenance. The facility team confirms that power is being supplied to the rack. The server's front panel LED is not illuminated. Which of the following should the administrator check FIRST?
45A server configured with RAID 5 has two failed drives. The array is offline and critical data is inaccessible. The administrator has replacement drives available. What should the administrator do to restore the data and array with minimal data loss?
46A web application server is experiencing intermittent 502 Bad Gateway errors during peak usage hours. The server's reverse proxy logs show connections to the backend application server being refused. The application server's resource monitor shows CPU utilization at 95% and memory utilization at 40%. Which of the following actions is MOST likely to resolve the issue?
47After a thunderstorm, a server in a remote office is reachable via remote console but has no network connectivity. The network switch port shows a steady link light, but the server's NIC shows no link light. Which of the following is the MOST likely cause?
48A server fails to complete POST and emits a series of beeps: two long, three short. According to the manufacturer's documentation, this beep code indicates a memory error. The server has eight DIMMs installed. Which of the following steps should the technician perform FIRST?
49A server experienced an unexpected power loss. After power is restored, the server boots successfully, but several critical services fail to start automatically. The administrator checks the system logs and finds errors indicating missing LUNs from a SAN. The SAN administrator confirms the SAN is online and the LUNs are assigned to the server's WWNs. The server uses FC HBAs. Which TWO steps should the administrator perform to diagnose the issue?
50Refer to the exhibit. A Linux server started reporting I/O errors to its local disk. The administrator runs `dmesg` and sees the output shown. Which of the following is the MOST likely cause of the errors?
51Refer to the exhibit. A Windows file server at a branch office lost network connectivity after a scheduled reboot. The administrator logs in via the console and runs `ipconfig /all`, which shows the output in the exhibit. The server should have a static IP of 10.0.0.50. What should the administrator do to restore connectivity?
52A company uses a two-node failover cluster for a critical database application. The cluster consists of Node A and Node B, with shared storage connected via SAS. During a routine check, the administrator discovers that Node A has failed and the cluster resources did not fail over to Node B. The cluster service is running on Node B, but the database resource remains offline. Node B can successfully ping Node A's management IP address but not the dedicated cluster heartbeat IP. The shared storage appears in the operating system on Node B, but attempts to bring the disks online fail. The administrator must restore database service as quickly as possible. Which of the following actions should the administrator take FIRST?
53A server administrator notices that a recently installed PCIe network card is not being recognized by the operating system. The server is running Windows Server 2019. The administrator has verified that the card is securely seated and that the slot is enabled in the BIOS. Which of the following should the administrator try FIRST to resolve the issue?
54A database server running on a Linux VM has started experiencing periodic crashes. The VM is hosted on an ESXi hypervisor. The system logs show entries: 'kernel: Out of memory: Kill process 12345 (mysqld) score 700 or sacrifice child'. The VM has 16GB RAM allocated and no memory overcommitment on the host. The DBA reports that the database workload hasn't increased. Which of the following actions should be taken FIRST to diagnose the issue?
55A server technician has just replaced a failed hard drive in a RAID 5 array. After inserting the new drive, the array begins rebuilding, but after 10 minutes, the rebuild fails and the array status shows 'degraded' again. Which of the following is the MOST likely cause?
56A server administrator is troubleshooting a network connectivity issue where a newly installed server cannot communicate with other devices on the same subnet. The server has a static IP address configured. Other devices on the same switch can communicate. Which TWO of the following could be the cause of the issue? (Select TWO)
57A company has a small virtualized environment with two ESXi 7.0 hosts (HostA and HostB) managed by vCenter Server. They use a shared iSCSI storage array for all VMs. The network consists of a single physical switch that connects the hosts, storage, and management traffic using VLANs. Yesterday, the switch failed and was replaced with an identical model. After restoring the configuration from a backup, all VMs on HostA are working normally, but all VMs on HostB show 'network disconnected' in the vSphere console. The VMs on HostB are still running, but they cannot communicate with any other device on the network. HostB itself is reachable via its management IP and can access the storage array. The administrator has verified that the physical cables are correct and the NICs are up on HostB. The VLAN configuration on the new switch was restored for the ports connecting HostB, but the issue persists. Which of the following actions should the administrator perform FIRST to restore connectivity?
58A small business relies on a single Dell PowerEdge T340 tower server running Windows Server 2019 as its primary domain controller and file server. The server is situated in a locked, air-conditioned IT closet and is protected by an online surge protector. During a severe thunderstorm last evening, the building experienced a momentary power loss. This morning, the administrator found the server completely powered off. When the power button is pressed, the front panel diagnostic LEDs briefly illuminate, and all internal cooling fans start spinning for about two to three seconds, but then the server abruptly powers down. There are no audible beep codes, and the monitor never receives a video signal. The administrator has already verified the power cable is firmly seated in both the server and the surge protector, swapped the power cable with a known-good one, and tested the surge protector with another device, which powered on normally. No spare parts are immediately available, and server downtime must be kept to a minimum. What should the administrator do FIRST to diagnose and resolve the problem?
59A medium-size enterprise runs a vSphere 7.0 cluster with two hosts (HostA and HostB) for production VMs. Each host has two 10GbE uplinks (vmnic0 and vmnic1) connected to separate physical switches (SW1 and SW2) for redundancy. A single vSphere standard switch (vSwitch0) uses both uplinks with teaming policy set to 'Route based on originating virtual port ID,' 'Network failure detection: Link status only,' and 'Notify switches: Yes.' Lately, several VMs running on HostA experience intermittent network disconnections lasting 20–30 seconds, while VMs on HostB are unaffected. The administrator checks vCenter events and sees repeated messages: 'Lost uplink redundancy on vSwitch0. vmnic0 is down.' followed seconds later by 'Uplink redundancy restored. vmnic0 is up.' The physical switch SW1's log shows the port for vmnic0 transitions through spanning-tree listening and learning states each time this occurs, taking about 15 seconds. The link is a trunk allowing all necessary VLANs, with no errors or security violations. The network team has already swapped the fiber cable and SFP+ transceiver for vmnic0 without improvement. The VMs affected are those whose virtual ports are pinned to vmnic0 by the load-balancing policy; VMs pinned to vmnic1 never experience disconnections. The administrator has also rebooted HostA and updated the NIC firmware, but the flapping continues. What should the administrator do to permanently resolve the intermittent connectivity?
60A file server running Windows Server 2016 is critical for a department's daily operations. For the past two weeks, it has been crashing with a blue screen every 1–2 days. The IT team collects minidump files and opens the latest one in WinDbg. After running '!analyze -v', they obtain the following output: ``` DRIVER_IRQL_NOT_LESS_OR_EQUAL (d1) An attempt was made to access a pageable (or completely invalid) address at an interrupt request level (IRQL) that is too high. This is usually caused by drivers using improper addresses. Arguments: Arg1: fffff80012345678, memory referenced Arg2: 0000000000000002, IRQL Arg3: 0000000000000000, value 0 = read operation, 1 = write operation Arg4: fffff80abcde1234, address which referenced memory Debugging Details: ... PROCESS_NAME: fileserver.exe SYMBOL_NAME: NetAdapterCx.sys!NetAdapterCxSetLinkState+0x1234 IMAGE_NAME: NetAdapterCx.sys ... ``` The server has 64GB ECC RAM, two Intel Xeon processors, and a Broadcom NetXtreme quad-port 10GbE NIC. The NIC driver was updated from version 7.12.6 to 7.14.8 two weeks ago as part of patch management. The BSODs started occurring immediately after that driver update; prior to the update, the server had been stable for over six months. The driver was obtained from the server manufacturer's support site and installed without error. The administrator wants to restore stability with minimal downtime and risk. Which action should the administrator take FIRST?
The troubleshooting domain covers the key concepts tested in this area of the SK0-005 exam blueprint published by CompTIA. Courseiva provides free domain-focused practice, mock exams, missed-question review, and readiness tracking across all SK0-005 domains — no account required.
The Courseiva SK0-005 question bank contains 60 questions in the troubleshooting domain. Click any question to see the full explanation and answer breakdown.
Start with a 10-question focused session to identify your baseline accuracy in this domain. Read every explanation — even for questions you answer correctly — to understand the reasoning. Once you score consistently above 80%, move to a 20–30 question session to confirm depth before moving to the next domain.
Yes — the session launcher on this page draws questions exclusively from the troubleshooting domain. Choose 10, 20, 30, or 50 questions for a focused session, or click individual questions to review them one by one.
Save your results, see per-domain analytics, and get readiness scores — free, for every certification.
Sign Up FreeFree forever · Every certification included