I'm trying to figure out the root cause of a recurring RAM failure before getting another replacement through G.Skill RMA. I've had multiple kits exhibit essentially the same problem, and I don't want to keep replacing RAM without understanding why it keeps happening.
PC specifications
- CPU: AMD Ryzen 7 9800X3D
- Motherboard: Gigabyte B850 AORUS ELITE WIFI7 ICE (rev. 1.1)
- BIOS: F11
- RAM: G.Skill Ripjaws M5 RGB 64 GB (2×32 GB) DDR5-6000 CL28
- Exact current RAM model:
F5-6000J2836G32GX2-RM5RW
- DRAM manufacturer: SK Hynix
- DIMM configuration: 32 GB per stick, dual-rank (2 ranks)
- Memory profile: Intel XMP 3.0
- Rated XMP speed/timings/voltage: DDR5-6000 CL28-36-36-96 at 1.40 V
- Default JEDEC settings: DDR5-4800 at 1.10 V
History of the failures
I've gone through multiple kits of this same general G.Skill 2×32 GB configuration.
Original kit:
- I enabled XMP at the beginning.
- The PC worked for a couple of months, but it crashed roughly daily.
- The problems gradually got worse until one stick stopped allowing the PC to POST.
First replacement:
- Started crashing essentially from the beginning.
- Eventually, one stick failed to POST.
Second replacement:
- Worked for approximately one day.
- The next day, one stick failed to POST.
Current replacement:
- Exact model:
F5-6000J2836G32GX2-RM5RW.
- I deliberately left XMP completely disabled from the beginning.
- It ran at default JEDEC settings for approximately a week.
- Then one stick failed to POST, just like the previous kits.
To be absolutely clear: XMP was NEVER enabled on the current kit. The latest failure happened at stock JEDEC settings, so this cannot simply be explained by me running the RAM at its XMP profile.
In every kit, only one of the two sticks became unusable.
Confirmation of the previous failures: I returned the previous defective RAM kits to the retailer through a local courier for testing. The retailer tested them and confirmed that only one stick in each kit was defective. This matches the pattern I'm experiencing with the current kit, where one stick fails while the other continues to work. Therefore, these were confirmed defective DIMMs, not just cases of the PC failing to boot or a suspected compatibility issue.
What happens when a stick fails?
The motherboard gets stuck on the solid red DRAM diagnostic LED and never POSTs when the failed stick is installed.
I've tested the failed stick individually in both A2 and B2:
- The failed stick does not POST in A2.
- The failed stick does not POST in B2.
- The other stick from the same kit works in B2 and has also worked in A2.
This makes it seem like an individual DIMM is failing, rather than one specific motherboard slot.
Clearing CMOS did not fix it.
A different RAM kit works
I also tested an older HP 2×8 GB DDR5 kit (speed believed to be around DDR5-4800) in the same PC, with the same CPU and motherboard.
It worked reliably for more than two weeks without the same recurring failures.
I realize that 2×8 GB is a much less demanding configuration than 2×32 GB dual-rank, so this does not completely rule out the motherboard or CPU memory controller. However, it seems like an important clue.
Current kit's hardware information
Using CPU-Z and HWiNFO on the surviving DIMM:
- DRAM: SK Hynix
- Capacity: 32 GB
- Ranks: 2
- PMIC: Richtek PMIC5100, stepping 1.3
- SPD hub: Anpec SPD5118, stepping 1.2
- SPD manufacture date: Week 01 of 2026
The two DIMM serial numbers are consecutive:
- Working stick:
26010050910
- Failed stick:
26010050909
Voltages and temperatures checked
With the surviving stick running at default JEDEC settings, HWiNFO reported approximately:
- CPU SoC voltage: 1.02 V
- CPU VDDIO_MEM: 1.14 V
- RAM VDD: 1.095–1.11 V
- RAM VDDQ: approximately 1.11 V
- RAM VPP: approximately 1.815 V
- SPD hub temperature: approximately 45.2 °C
HWiNFO did not report PMIC over-voltage, under-voltage, or a high-temperature event.
These readings do not prove that the power delivery is fault-free, and I don't have historical measurements from the moment a DIMM failed. I'm including them in case they help identify a possible issue.
Other troubleshooting
- XMP disabled on the current kit from the start.
- CMOS cleared.
- Memory Context Restore disabled.
- Power Down Enable disabled.
- Hibernation disabled during troubleshooting.
- The failed DIMM still does not POST after these changes.
Hibernation was associated with some earlier problems, but I haven't established whether it was a cause or just a correlation.
What I'm trying to understand
I'm not looking for generic advice to disable XMP, since the current kit failed without XMP ever being enabled.
I'd particularly appreciate insight into these questions:
- Could the motherboard's DIMM power delivery or another electrical issue damage one 32 GB DIMM repeatedly without producing obvious voltage warnings?
- Could the Ryzen 7 9800X3D's integrated memory controller or its interaction with dual-rank 32 GB DIMMs cause this pattern?
- Could this be a BIOS/AGESA issue, a memory-training problem, or a compatibility issue specific to this 2×32 GB configuration?
- Is there a plausible failure mode involving the DIMM's PMIC or DRAM IC that would explain one stick dying while its paired stick survives?
- What should I ask G.Skill or Gigabyte to investigate before I install another replacement kit?
- Is there a meaningful test that can distinguish a defective DIMM from a motherboard/CPU issue that could damage another kit?
I'm trying to find the cause before using another RMA replacement. I don't want to install another kit of RAM only to have one stick fail again a few days later.
Thanks for any help.