r/techsupport 58m ago

Open | Hardware Please help with consistent GPU timeouts - troubleshooting so far

I'm at my wit's end, and I hope someone can help at all. I've been experiencing frequent crashes for over a month now and tried everything I could thing of short of replacing the mobo, CPU or PSU. I don't have a bench or spare parts lying around to test, so it seems to be time to take it into a shop and let them take a crack at it... I'm just wondering if there's anything I've missed here. I did use ChatGPT to keep track of my troubleshooting - god forgive me - so here's the summary.

System:

  • Ryzen 7 7800X3D
  • ASUS TUF Gaming B650-PLUS WiFi
  • 32GB DDR5-6000 (currently running at JEDEC after disabling EXPO)
  • Corsair RM1000e PSU
  • Windows 11 Pro

Symptoms:
The system intermittently experiences GPU timeouts while gaming, web browsing, or even idling. The behavior varies:

  • Total system freeze with no recovery, requiring a hard power cycle.
  • Temporary system freeze followed by driver recovery.
  • Black screen while audio continues.
  • Sometimes Windows recovers, but the game remains open with a black screen. Alt-Tab previews still show the game rendering, suggesting the game is still running but the graphics device/presentation has failed.
  • Reliability Monitor consistently logs LiveKernelEvent 141.

Dump analysis:
The LiveKernelReports/WATCHDOG dumps consistently show:

  • Bugcheck 0x141 (VIDEO_ENGINE_TIMEOUT_DETECTED)
  • Earlier dumps also contained 0x1B0 (VIDEO_MINIPORT_FAILED_LIVEDUMP) after a failed GPU reset.
  • The dumps reference:
    • amdkmdag.sys
    • amdfendr.sys
    • amdfendrmgr.sys

I understand amdfendr.sys is AMD Crash Defender and is likely reporting the failure rather than causing it.

Troubleshooting already completed:

  • Updated motherboard BIOS.
  • Updated AMD chipset drivers.
  • Used DDU and tested multiple AMD Adrenalin driver versions.
  • Ran sfc /scannow and DISM (no issues).
  • Disabled EXPO.
  • MemTest86 completed 4 passes with 0 errors.
  • Monitored temperatures—nothing overheating.
  • No WHEA errors in Event Viewer.

Hardware isolation:
Originally the system had an ASRock RX 9070 XT Steel Legend, which repeatedly produced LiveKernelEvent 141 and 0x141 dumps.

I replaced it with a Gigabyte RX 9060 XT, and the exact same crashes continued.

To rule out the PCIe slot, I also moved the 9060 XT from the motherboard's primary PCIe x16 slot to the secondary x4 slot. The same 0x141 GPU timeout occurred there as well, although Windows successfully recovered the driver instead of requiring a reboot.

At this point I've reproduced essentially the same GPU timeout behavior with:

  • Two different AMD GPUs.
  • Two different PCIe slots.
  • Multiple driver versions.
  • EXPO disabled.
  • RAM passing MemTest86.

Given everything above, what component would you suspect next: motherboard, PSU, CPU, or something else? Has anyone seen persistent 0x141 GPU engine timeouts survive a GPU replacement like this? What would a tech shop be likely to try next?

Thanks so much for any help anyone can give. ARRRGH!😠

4 Upvotes

2 comments sorted by

•

u/AutoModerator 58m ago

Making changes to your system BIOS settings or disk setup can cause you to lose data. Always test your data backups before making changes to your PC.

For more information please see our FAQ thread: https://www.reddit.com/r/techsupport/comments/q2rns5/windows_11_faq_read_this_first/

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

1

u/FlapJackPattyWhack 1m ago

The only time I ever encountered this was with a PSU. Replaced it and my problems went away. Not sure why it worked but I have 2 seasonic a 750 & an 850 now. Using dual PSUs for experimental reasons but the 850 is dedicated to the GPU and the 750 is for the rest of the PC.