LFCS Operation of Running Systems Practice Question
A system administrator is troubleshooting a production web server running CentOS 7 that became unresponsive. The server is still pingable, but SSH connections timeout. The admin performs an out-of-band console login. The server appears frozen; typing commands shows no output. The admin is able to trigger a Magic SysRq key sequence (Alt+SysRq+f) to kill the hung processes. After that, the server resumes normal operation. However, the admin wants to understand the root cause. Upon checking 'dmesg', they see repeated messages: 'NMI watchdog: BUG: soft lockup - CPU#0 stuck for 22s!' followed by stack traces from a kernel thread. Which action should the admin take to prevent recurrence while maintaining system stability?
⚠ Common exam trap
It's easy for candidates to think soft lockup errors are false positives or can be safely ignored by increasing thresholds or disabling the watchdog, when in fact they indicate a genuine kernel or hardware issue that requires a proper fix.
Answer choices
Why each option matters
Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.
Correct answer & explanation
✓
Update the server's BIOS/firmware and check for kernel updates.
Soft lockup errors on CentOS 7 often indicate kernel bugs or hardware/firmware issues that cause CPUs to stall for extended periods. Updating the BIOS/firmware can resolve underlying hardware timing problems, while kernel updates may include patches for known soft lockup bugs. This approach addresses the root cause without disabling or weakening the watchdog mechanism, preserving system stability.
Answer analysis
Option-by-option breakdown
For each option: why learners choose it and why it is or isn't the right answer here.
- ✗
Replace the power supply unit to ensure stable power.
Why it's wrong here
The dmesg soft lockup traces identify a kernel thread monopolising CPU#0, so swapping the power supply addresses no fault the logs describe. It is tempting because hardware instability can freeze a server, but PSU replacement would be correct only where logs showed voltage, thermal or memory errors rather than a CPU stuck in kernel code.
- ✗
Increase the soft lockup threshold via sysctl to reduce false positives.
Why it's wrong here
Raising the soft lockup threshold through sysctl only delays when the watchdog reports a stuck CPU; the kernel thread still spins and the server still hangs. It is tempting to silence noisy warnings, but threshold tuning would be correct only when genuine stalls are shorter than the configured detection interval.
- ✗
Add 'nosoftlockup' to the kernel boot parameters.
Why it's wrong here
Adding nosoftlockup merely suppresses the watchdog's detection and reporting of CPU stalls, leaving the underlying kernel-thread hang unaddressed and recurrence silent. It is tempting as a way to stop alarming dmesg output, but it would be appropriate only when soft lockup warnings are confirmed false positives from virtualisation or heavy scheduling latency.
- ✓
Update the server's BIOS/firmware and check for kernel updates.
Why this is correct
Soft lockups in kernel threads typically stem from firmware bugs or kernel defects rather than user processes, so the Magic SysRq kill only masks symptoms. Updating BIOS/firmware and applying kernel updates addresses the underlying CPU or scheduler defect that caused the 22-second stall, preventing recurrence.
Visual reference
Go deeper
Related to this question
About these practice questions
This LFCS question is part of Courseiva's 406-question bank — original exam-style content with full explanations and wrong-answer analysis, never real exam questions or exam dumps. Learn why practice questions differ from exam dumps →
JA
Written by Johnson Ajibi, MSc IT Security
Senior Network & Security Engineer · founder of Courseiva
This LFCS practice question is part of Courseiva's free Linux Foundation certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the LFCS exam.