Common random BSODs (WHEA_UNCORRECTABLE_ERROR)
Common random BSODs (WHEA_UNCORRECTABLE_ERROR)
A few years back I assembled a PC with a Ryzen 9 5950x, EVGA RTX 3080, 32GB G.Skill RAM, and a 2TB Samsung 980 Pro. I used an ASUS motherboard but faced problems where my USB devices would disconnect and reconnect repeatedly, making it nearly impossible to operate. I returned the board for a refund, only to receive a used one that was clearly old and damaged. I switched to an MSI X570S Carbon Max Wi-Fi board because it was designed for the 5000 series processors, unlike the older models. While waiting, I had to use the same ASUS board they sent back, which still had the same issues—no improvement. Eventually, I got a new board and started using it. It experienced occasional USB disconnections but became more stable over time. Still, I encountered random BSODs during gameplay, always reporting an unrecoverable error. After several months, my system would crash completely, showing only a black screen and no BIOS information. Eventually, after multiple attempts to replace the motherboard, I realized the problem persisted. After replacing the PSU and some cables, the issue continued. I eventually swapped memory with another machine, which worked fine, indicating the problem was likely with the chipset or motherboard. I had to return the board multiple times before getting a replacement that functioned properly. Over time, I've disassembled and reassembled the system repeatedly, always ensuring everything was perfectly connected. Despite my efforts, these BSODs keep occurring. Recently, after running a stress test on 3D Mark Port Royale, the results were alarming—my PC failed at high frequencies, with increasing heat and load. I considered it a test error, but the pattern persisted. I’ve tried adjusting overclocks in Ryzen Master, using the curve optimizer, and even used software like Precision X1 or MSI Afterburner without success. The RAM was already stock, and neither the CPU nor GPU had any overclocking enabled. My only certainty is that something fundamental about my hardware is causing these persistent failures. I’m at a crossroads, unsure of the next steps to resolve this issue.
Visit C:\Windows\Minidump and verify the presence of any minidump files. If found, return to the Windows directory and transfer the entire Minidump folder to the Downloads folder (desktop works if OneDrive isn't syncing). Compress the copied folder and include it in a post. Please adhere strictly to the instructions provided, as Windows discourages file manipulation in that area. Also, look for a file named Memory.dmp in C:\Windows; this will be significantly larger. If uploading it isn't possible, consider using a file host. If you're not receiving dump files: Examine the crash arguments (sub-errors). If the system remains unresponsive on the BSOD screen (you can't access dumps), this step isn't required. However, if it restarts after a few seconds, proceed to the guide and uncheck the automatic restart option. To restart manually, press the power button. To modify the BSOD display to show extra details, edit the registry. If unsure about registry changes, skip this action. Go to HKEY_LOCAL_MACHINE\System\CurrentControlSet\Control\CrashControl, right-click the blank space in the right pane and select New → DWORD value with the name "DisplayParameters". Right-click it, adjust the value data to 1 (hex or decimal doesn't matter). It should appear correctly after saving. Reboot to apply the change. The next time a BSOD occurs, you should see these additional numbers in the top-left corner.
The test results don't make sense—they’re faulty. I’ve seen some CPU temperatures around 127°C on the CPU occasionally. That’s unusual. I’m mostly sure the issue lies with the CPU itself. Have you tried using a different CPU?