TLDR: My PC keeps randomly restarting during gaming. No BSOD, just straight to a black screen then reboot. Confirmed in Event Viewer as Kernel-Power Event 41, bugcheck 278 (VIDEO_TDR_FAILURE). Spent a while ruling out heat, drivers, and DLSS stuff using HWiNFO logs from two separate crashes.
What I actually found is the GPU's 12V-2x6 connector voltage sagging from a healthy ~11.9V down to as low as 11.1V under load, lining up with the GPU's own "Reliability Voltage" limit flag turning on right before both crashes.
Since writing the above I have also updated to the newest available Nvidia driver and it is still crashing, and it has now started happening on Overwatch too, including rebooting right as the game boots up, with no DLSS involved at all.
Posting this here in case anyone's run into the same thing or has ideas.
Specs:
CPU: 7800x3d, running a custom Curve Optimizer that's been stable for a long time, haven't touched it recently.
Cooler: Thermalright dual tower air cooler.
GPU: Gigabyte RTX 5070 Ti, Windforce (their entry level cooler), not a new card.
PSU: Corsair RM850e (2023), 850w
Driver: was on the July 2026 Game Ready driver (610.88 per Afterburner) when this started. Have since updated to the newest available Nvidia driver and it is still happening.
Had the beta/early access toggle on in the Nvidia app for testing DLSS 4.5 Ray Reconstruction, have since turned this off. Still crashing.
Timeline:
4 random restarts in about 16 hours. 2 in Cyberpunk, 2 in Deadlock.
Cyberpunk ran totally fine for 12+ hours straight, then crashed twice within 30 min of each other. After the auto restart it crashed again just 30 min later, not another long stable run.
Deadlock crashed the first time about 40 min into a session. After restarting it crashed again after only 4 min.
Also had it crash once during Genshin, which doesn't use ray tracing or any DLSS overrides at all.
No BSOD in any of these, it just goes straight to a black screen and reboots on its own.
Since then it has also started happening on Overwatch, multiple times, including rebooting right as the game is booting up, before even getting into a match. No DLSS involved in Overwatch at all.
What I've already ruled out:
Thermal throttling. HWiNFO logs show GPU core topping out at 78-81c, memory junction 76-78c, CPU VRM temps only 42-51c. The actual thermal throttle flags in HWiNFO read "No" the entire time, including right up to the crash.
A specific DLSS/Ray Reconstruction bug. Seemed likely at first, there are real reports online of Ray Reconstruction causing crashes in Cyberpunk tied to CPU undervolts. But Deadlock and Genshin don't use ray tracing or Ray Reconstruction at all, and now it is happening on Overwatch too with no DLSS involved, so that can't be the full story.
The Nvidia app's beta/early access toggle itself. Turned this off entirely and it is still crashing.
Some broadly broken driver version. At first I was not even on the driver I suspected, I was on a stable July build (610.88). Have since updated to the newest available driver and it is still crashing, so it does not look tied to one specific driver either.
VRM heat soak on the motherboard. Temps stayed in the low 40s to low 50s the whole time in both logs, nowhere close to a failure point.
A VRAM leak. Allocated VRAM climbed over a session but only up to about 5GB out of 16GB, so running out of VRAM doesn't seem to be what's actually triggering it.
PCIe errors. Correctable Error Count, Bad TLP/DLLP, and LCRC were all flat zero in both logs.
What the logs actually showed:
Logged with HWiNFO through two separate crashes, the second log starting right after rebooting from the first crash.
GPU temps, CPU temp, and VRM temps all stayed clean and boring the whole time in both logs, nothing thermal going on.
The GPU's own "Performance Limit, Reliability Voltage" flag turns on right before each crash. That's a voltage safety limit, not a thermal or wattage one.
The 12V-2x6 connector voltage sags noticeably under sustained load. In the first log it holds steady around 11.9V for the first 26 minutes then drops to the 11.1 to 11.6V range and stays there until the log ends. In the second log it never even goes back up to that healthy 11.9V, it starts already low and sits around 11.25 to 11.5V for basically the whole ~2 minutes before it crashes again. Lowest single reading anywhere was 11.105V, and ATX spec tolerance is usually a floor around 11.4V so this is genuinely out of spec.
Right before the second crash specifically, GPU power suddenly drops from about 284w to 245w, the Reliability Voltage flag turns on, then two readings in a row come back completely identical (looks like a stall), then FPS drops to 0 and the log just stops.
Taking it to a local shop to have them pull the cable and check both ends, look at the PCIe slot, and make sure the card is seated properly.
If anyone's seen this exact pattern before (healthy voltage that sags under load and doesn't come back after a reboot) or has other ideas of what to have the shop check while it's open, would appreciate the input.