Hi everyone,
I'm posting here because I'm using an MSI GeForce RTX 5090 Suprim Liquid and I'm trying to determine whether this points to an isolated GPU issue or whether other MSI RTX 5090 owners have observed similar crashes and WinDbg signatures.
System
- GPU: MSI GeForce RTX 5090 Suprim Liquid
- Motherboard: ASUS ROG Crosshair X870E Hero
- CPU: Ryzen 9 9900X3D
- RAM: 96 GB DDR5 G.Skill 6000 CL30
- PSU: ASUS ROG Strix Platinum 1200W (ATX 3.1)
- Windows 11 24H2
Symptoms
Random black screens resulting in a VIDEO_TDR_FAILURE (0x116) and system reboot.
The crashes occur in different situations:
- Unreal Engine 5
- Autodesk Maya
- Fortnite
- Crimson Desert
Sometimes during normal desktop use, often after previously running GPU-accelerated applications.
The crashes are not exclusively related to heavy GPU load.
No artifacts are ever visible before the crash.
GPU temperatures remain around 55°C under load.
Timeline
- The issue first appeared in early 2026.
- It then disappeared completely for approximately three months without any hardware changes.
- Since mid-May 2026, the crashes have returned and now occur regularly.
- During both periods, the crashes produced the same WinDbg signature.
WinDbg
Every single minidump shows the same result:
- VIDEO_TDR_FAILURE (0x116)
- IMAGE_NAME: nvlddmkm.sys
- Offset: nvlddmkm+0x1958210
- Arg3 (NTSTATUS): 0xC000009A
- FAILURE_BUCKET_ID: 0x116_IMAGE_nvlddmkm.sys
- FAILURE_ID_HASH: {c89bfe8c-ed39-f658-ef27-f2898997fdbd}
Sometimes Windows also logs:
- Event ID 14 (nvlddmkm)
- Event ID 153 (nvlddmkm)
- "GpuRcReset TDR occurred on GPUID:100"
The WinDbg signature is identical across all crashes.
Already tested
- Clean Windows reinstall
- DDU before every driver installation
- Multiple NVIDIA Game Ready drivers
- Multiple NVIDIA Studio drivers
- Different motherboard BIOS versions
- CMOS reset
- EXPO enabled/disabled
- PCIe Auto and Gen4
- HAGS enabled/disabled
- MPO disabled
- G-Sync enabled/disabled
- Single monitor
- All monitors forced to the same refresh rate
- iGPU disabled
- Hyper-V/VBS investigated
- GPU underclock (-200 MHz core)
- Power Limit reduced to 70%
- Prefer Maximum Performance
- No overlays
- MemTest: PASS
- UserDiag: PASS
The crash remains identical.
Hardware checks
- OCCT 3D (multiple modes): PASS
- OCCT VRAM: PASS
- OCCT Power: PASS
- GPU temperatures remain below 55°C during stress tests
- No artifacts
- No WHEA errors
My question:
At this point I'm trying to determine whether this points to an isolated hardware failure or whether multiple RTX 5090 owners are reproducing the same failure signature.
Has anyone using an MSI RTX 5090 (Suprim, Suprim Liquid, Vanguard, Gaming Trio, etc.) observed the same WinDbg signature, especially:
- nvlddmkm+0x1958210
- Arg3 = 0xC000009A
- VIDEO_TDR_FAILURE (0x116)
or similar crashes?
I'm not trying to conclude that this is necessarily a driver issue or a hardware defect. I'm simply trying to determine whether other MSI RTX 5090 owners are reproducing the same failure signature before sending my card for RMA.
If you've analyzed your crash dumps with WinDbg and see similar results, I'd really appreciate comparing outputs, even if your hardware configuration is different.
I'm posting here because I'm using an MSI GeForce RTX 5090 Suprim Liquid and I'm trying to determine whether this points to an isolated GPU issue or whether other MSI RTX 5090 owners have observed similar crashes and WinDbg signatures.
System
- GPU: MSI GeForce RTX 5090 Suprim Liquid
- Motherboard: ASUS ROG Crosshair X870E Hero
- CPU: Ryzen 9 9900X3D
- RAM: 96 GB DDR5 G.Skill 6000 CL30
- PSU: ASUS ROG Strix Platinum 1200W (ATX 3.1)
- Windows 11 24H2
Symptoms
Random black screens resulting in a VIDEO_TDR_FAILURE (0x116) and system reboot.
The crashes occur in different situations:
- Unreal Engine 5
- Autodesk Maya
- Fortnite
- Crimson Desert
Sometimes during normal desktop use, often after previously running GPU-accelerated applications.
The crashes are not exclusively related to heavy GPU load.
No artifacts are ever visible before the crash.
GPU temperatures remain around 55°C under load.
Timeline
- The issue first appeared in early 2026.
- It then disappeared completely for approximately three months without any hardware changes.
- Since mid-May 2026, the crashes have returned and now occur regularly.
- During both periods, the crashes produced the same WinDbg signature.
WinDbg
Every single minidump shows the same result:
- VIDEO_TDR_FAILURE (0x116)
- IMAGE_NAME: nvlddmkm.sys
- Offset: nvlddmkm+0x1958210
- Arg3 (NTSTATUS): 0xC000009A
- FAILURE_BUCKET_ID: 0x116_IMAGE_nvlddmkm.sys
- FAILURE_ID_HASH: {c89bfe8c-ed39-f658-ef27-f2898997fdbd}
Sometimes Windows also logs:
- Event ID 14 (nvlddmkm)
- Event ID 153 (nvlddmkm)
- "GpuRcReset TDR occurred on GPUID:100"
The WinDbg signature is identical across all crashes.
Already tested
- Clean Windows reinstall
- DDU before every driver installation
- Multiple NVIDIA Game Ready drivers
- Multiple NVIDIA Studio drivers
- Different motherboard BIOS versions
- CMOS reset
- EXPO enabled/disabled
- PCIe Auto and Gen4
- HAGS enabled/disabled
- MPO disabled
- G-Sync enabled/disabled
- Single monitor
- All monitors forced to the same refresh rate
- iGPU disabled
- Hyper-V/VBS investigated
- GPU underclock (-200 MHz core)
- Power Limit reduced to 70%
- Prefer Maximum Performance
- No overlays
- MemTest: PASS
- UserDiag: PASS
The crash remains identical.
Hardware checks
- OCCT 3D (multiple modes): PASS
- OCCT VRAM: PASS
- OCCT Power: PASS
- GPU temperatures remain below 55°C during stress tests
- No artifacts
- No WHEA errors
My question:
At this point I'm trying to determine whether this points to an isolated hardware failure or whether multiple RTX 5090 owners are reproducing the same failure signature.
Has anyone using an MSI RTX 5090 (Suprim, Suprim Liquid, Vanguard, Gaming Trio, etc.) observed the same WinDbg signature, especially:
- nvlddmkm+0x1958210
- Arg3 = 0xC000009A
- VIDEO_TDR_FAILURE (0x116)
or similar crashes?
I'm not trying to conclude that this is necessarily a driver issue or a hardware defect. I'm simply trying to determine whether other MSI RTX 5090 owners are reproducing the same failure signature before sending my card for RMA.
If you've analyzed your crash dumps with WinDbg and see similar results, I'd really appreciate comparing outputs, even if your hardware configuration is different.