Skip to content
zarza zarza
Advertisement

How to Debug a Kernel Panic in Production

25/05/2026 8 min Temporada 1 Episodio 10

Listen "How to Debug a Kernel Panic in Production"

Episode Synopsis

When your production server suddenly drops into a kernel panic, the instinct is to reboot and hope it doesn't happen again. Lucas and Luna walk through a real case from a fintech startup where a misconfigured memory module caused intermittent panics across a 200-node cluster. They explain how to capture crash dumps, read the stack trace to identify the offending driver or hardware, and use tools like kdump, crash, and mcelog to pinpoint the root cause. The episode also covers prophylactic measures: setting up netconsole for remote logging, enabling panic-timeouts for automatic reboot, and testing memory with memtest86 during maintenance windows. By the end, listeners will know exactly what to do the next time their server greets them with a blinking cursor on a black screen.

#KernelPanic #LinuxDebugging #ProductionServer #kdump #crashUtility #mcelog #netconsole #memtest86 #Sysadmin #Linux #Technology #ServerEngineering #FexingoBusiness #BusinessPodcast #TechPodcast #LinuxServerAdmin #Bash #ServerReliability

Keep every episode free: buymeacoffee.com/fexingo

ZARZA Studio — Your station on air today: library, music clock, schedule, studio and reports, from the browser.

Meet ZARZA Studio
on air now stations in the catalogue 1,829,025 podcasts countries