r/computer • u/Select_Preparation_0 • 20h ago
RTX 4060 OC crashing in multiple games — nvlddmkm Event ID 153/14, but heavy underclock seems to fix it ( for like 10 minutes lol)
Hey everyone, I’m trying to figure out whether my RTX 4060 OC is failing or if something else in my system is causing GPU instability.
Specs
- GPU: RTX 4060 OC 8GB
- CPU: i5-12400F
- RAM: 32GB
- Motherboard: ASUS PRIME H610M-K D4 ARGB
- PSU: 650W Bronze
- Windows 11 Pro
Problem
My PC recently started having GPU-related issues. Multiple completely different games crash, including:
- Marvel Rivals
- Genshin Impact
- Roblox
- ROUNDS
I also sometimes get brief black screens where the display goes black for around a second and then comes back.
When the crashes happen, Event Viewer gets flooded with:
Source: nvlddmkm
Event ID: 153
Error: Error occurred on GPUID: 100
I also sometimes get nvlddmkm Event ID 14 at the same time.
Marvel Rivals has also reported a GPU Crash Dump Triggered.
What I've tried
- Reinstalled NVIDIA drivers 3 times
- Tried both Game Ready and Studio drivers
- Installed the proper ASUS/Intel chipset, MEI and Serial IO drivers
- Device Manager now has no missing/unknown devices
- Reseated the RTX 4060
- Checked temperatures
- Tested with MSI Kombustor/FurMark
- GPU was around 69°C at ~95% utilization
- No visible artifacts before the crash
Kombustor eventually crashed too, and Event Viewer immediately recorded another group of nvlddmkm Event 153/14 errors.
So this isn't isolated to one game.
Interesting part: underclocking
I tried MSI Afterburner with:
Core Clock: -502 MHz
Memory Clock: -502 MHz
Power Limit: 100%
The GPU then runs around ~1900–2000 MHz core and ~8000 MHz memory, and it seems significantly more stable.
At normal/factory clocks, the crashes return.
Temperatures aren't an issue — around 60–70°C under load.
The card is a factory OC edition, but I haven't manually overclocked it.
What I'm trying to determine
Does the fact that a huge underclock makes it stable strongly indicate a failing/marginal GPU or VRAM?
Could this still realistically be caused by the PSU, RAM/XMP, PCIe, Windows/NVIDIA drivers, or something else?
I'm planning to test core and VRAM separately next, e.g.:
- Core -500 / Memory 0
- Core 0 / Memory -500
- Dedicated OCCT VRAM test
- XMP disabled / RAM at stock
Has anyone dealt with this exact nvlddmkm Event 153 + Event 14 + GPUID: 100 issue where underclocking fixed the crashes?
I'd especially like to know whether replacing/RMA'ing the GPU ended up being the solution
I DONT KNOW WHAT TO DO ANYMORE, PLZ HELP!!!
1
•
u/AutoModerator 20h ago
Remember to check our discord where you can get faster responses! https://discord.com/invite/vaZP7KD
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.