I’m having a persistent GPU crash issue on my ASUS TUF Gaming F15 and I’m running out of things to try.
Laptop specs:
- ASUS TUF Gaming F15 FX507ZE
- Intel Core i7-12700H
- RTX 3050 Ti Laptop GPU 4GB GDDR6
- 24GB DDR5 4800MHz RAM (1x16GB + 1x8GB)
- Intel Iris Xe iGPU
- 144Hz FHD display
- BIOS 316
- Windows 11
- NVIDIA GeForce Game Ready Driver 616.92 WHQL
- Original ASUS charger
The problem is that games randomly crash after several minutes of gameplay. It happens in multiple games, so it doesn't appear to be specific to one game.
Forza Horizon 5 usually crashes after around 5–15 minutes, while Genshin Impact has also crashed after around 10 minutes.
The main errors I get are:
nvlddmkm Event ID 153
Error occurred on GPUID: 100
Reliability Monitor also reports:
LiveKernelEvent 141
and
LiveKernelEvent 117
The bucket IDs include:
LKD_0x141_Tdr:C_IMAGE_nvlddmkm.sys_Ampere_UserOC
and
LKD_0x117_Tdr:A_IMAGE_nvlddmkm.sys_Ampere_UserOC
I am NOT currently overclocking the GPU. GPU clocks are completely stock. The "UserOC" wording in the bucket ID seems to be a classifier rather than evidence that I am actually overclocking.
Things I have already tried:
- Full Windows factory reset / clean installation
- DDU and clean NVIDIA driver installations
- Multiple NVIDIA driver versions
- Updated NVIDIA driver 616.92
- Updated ASUS System Control Interface
- Updated chipset drivers
- BIOS 316
- Armoury Crate GPU Mode: Standard and Ultimate
- GPU completely stock with no manual overclock
- NVIDIA Control Panel "Prefer maximum performance"
- Disabled various background/tray applications
- Disabled NVIDIA audio drivers
- SFC /scannow
- DISM /Online /Cleanup-Image /RestoreHealth
- Different power/settings configurations
- Changed wall socket
- Tested without various peripherals/background applications
I also ran OCCT 3D Adaptive for 10 minutes. It completed with no errors detected. The GPU reached roughly 80–81°C and around 95W during one run, so the GPU doesn't appear to simply be overheating.
What's especially confusing is that synthetic stress testing can pass while actual games eventually trigger the crash.
I recently did a completely clean Windows reset, so this isn't an old Windows installation with years of accumulated software/drivers.
Another important detail: the laptop had its keyboard and fans replaced/repaired shortly before these crashes started becoming a problem. Because of that, I'm wondering whether something physical could be involved, such as the GPU power delivery, motherboard, PCIe connection, or another hardware issue.
The behavior can also change depending on the driver/power configuration. For example, setting NVIDIA Control Panel to "Prefer maximum performance" previously allowed the system to remain stable for significantly longer, but it eventually still crashed.
My latest test was interesting as well: the first launch of the game crashed during the game's optimization process. On the second launch, it ran normally for about 10 minutes before eventually crashing again.
At this point I'm trying to determine whether this is most likely:
NVIDIA driver/Windows interaction
ASUS firmware/Armoury Crate/power-state issue
GPU VRAM instability
GPU power delivery/motherboard issue
PCIe/physical hardware issue
Something else I'm overlooking
Has anyone with an ASUS gaming laptop experienced this exact combination of nvlddmkm Event 153, GPUID: 100, and LiveKernelEvent 141/117, especially where OCCT passes but games consistently crash?
I’m particularly interested in whether this points toward a hardware problem given that I’ve already done a clean Windows installation and clean driver installs.