System Configuration
- CPU: AMD Ryzen 5 9600X (OEM)
- Motherboard: ASRock B650M PG Lightning WiFi
- BIOS: version 4.44 (latest)
- RAM: G.Skill Trident Z5 RGB, 32GB (2x16GB), DDR5-6000, XMP/EXPO enabled
- Cooler: Deepcool AS500
- GPU: RTX 5070 (KFA2 ROCK OC)
- PSU: XPG Core Reactor II VE 750W, 80+ Gold
- Storage: ADATA SX8200PNP 1TB (OS drive), WD Blue SN5000 1TB — both checked via CrystalDiskInfo, health at 97-99%, no issues
- OS: Windows 11 Pro
Timeline
September — PC built. Ran completely stable for about 6 months with CPB enabled and stock AMD PBO settings.
First few months — CPU was hitting 95°C under load (stock PBO limits, unrestricted). Manual PPT/TDC/EDC limits (88000/75000/120000) were set later, bringing temps down to ~65-66°C — but this was already after the instability started.
April — with no hardware, driver, or setting changes, the system started hard-hanging with no BSOD: at idle, in the browser, and during Windows boot (spinning dots screen). Hanging during boot is the most frequent and most reliably reproducible scenario.
Key observation: with Core Performance Boost (CPB) disabled, the system boots and runs 100% stable (locked at 3.9 GHz base clock). With CPB enabled, it hangs on boot in almost every configuration tried.
Until June — stable on Positive Curve Optimizer +10. This was the first setting that let the system boot at all with CPB enabled after a long string of failed Negative CO attempts.
Then Positive CO +15 — stable for about 3 months, the only extended period of stability with boost enabled.
June — bumped to Positive CO +20 after the first repeat failure on +15. Worked for a while.
Now — +20 no longer boots. Reverting to +15 or +10 doesn't help either — none of the previously working CO values let the system boot with CPB enabled anymore.
What's been tried and did NOT fully fix it
CPU power/boost settings
- Curve Optimizer, All Core, Negative 10 — hangs
- Curve Optimizer, All Core, Negative 30 — hangs
- CPU Boost Clock Override Negative 100 — hangs on boot
- CPU Boost Clock Override Negative 400 — hangs on boot
- CPU Boost Clock Override Negative 450 — boots occasionally, then crashes with various BSODs (BAD_POOL_CALLER 0xc2, KERNEL_SECURITY_CHECK_FAILURE 0x139)
- PBO limits, Manual, multiple configs from stock down to heavily restricted (PPT 65000/TDC 55000/EDC 90000) with Curve Optimizer and Boost Override left on Auto — hangs on boot
- Fixed clock via CPU Ratio (non-CPB mode): tested at 4800 MHz @ 1.25V → 1.2V → 1.15V → 1.1V — stable at every step, final numbers in OCCT: 69°C, 90W, 100% load, 4806 MHz. This is the only fully working configuration, but with no dynamic boost.
Global C-States / power saving
- Global C-State Control → Disabled — didn't fix hangs under load/in browser
- Power Supply Idle Control: Typical Current Idle vs Low Current Idle — no noticeable effect
Memory
- XMP/EXPO disabled, tested at stock 4800 MHz — no change
- Full multi-pass MemTest86 was not run (only XMP disable as an indirect check)
BIOS
- Full reset to Optimized Defaults — no change
- BIOS updates were tried, but none of the previously working Positive CO values (+10/+15/+20) let the system boot anymore afterward
Windows / software
- Windows Power Plan: High Performance → Balanced — no change
- sfc /scannow — found and repaired corrupted system files
- Clean GPU driver reinstall via DDU — didn't fully resolve the issue
- Memory Integrity (HVCI) — disabling it removes the strict kernel memory integrity check (which triggers BSOD), but doesn't fix the root cause
- AMD SVM Mode → Disabled — tried, no improvement, and it breaks game anti-cheats that require virtualization (Riot Vanguard, BattlEye, EAC) — dropped this option
What temporarily WORKED
- Positive Curve Optimizer +10 — the first value that let the system boot at all with CPB enabled, after many failed attempts with Negative CO
- Positive Curve Optimizer +15 — stable for ~3 months (until June). OCCT numbers: 70°C, 88W, 100% load, 4800 MHz (dynamic via CPB, not fixed)
- Positive CO +20 (from June) — worked temporarily, then stopped booting
- Right now, none of +10/+15/+20 work.
Diagnostics that ruled out other causes
- HWiNFO64: motherboard sensors normal at fixed clock
- CrystalDiskInfo: both drives healthy (97-99%), rules out SSD degradation as root cause
- Event Viewer / WHEA-Logger: empty at the time of hangs — since most hangs happen before Windows finishes loading, the OS never gets a chance to log anything
- SrtTrail.txt / Startup Repair Log: all checks (disk, registry, bootloader, metadata) passed with code 0x0 — rules out filesystem/bootloader corruption
- CPU/GPU stress tests separately via OCCT: both pass cleanly under full sustained load for 5-20 minutes — meaning under steady load the CPU and GPU are stable, so the problem seems to be specifically in transition moments (frequency/boost changes), not sustained load
Additional context
- Originally suspected CPU degradation from running at 95°C under unrestricted stock AMD PBO settings for the first few months
- One notable pattern: the Positive CO value needed for stability keeps climbing over time (+10 → +15 → +20 → now not even +20 is enough), which looks more like progressive degradation than a one-time tuning issue
Main question
What should I actually do here, and how can I realistically fix this? The Positive CO value needed keeps climbing (+10 → +15 → +20, and now even +20 isn't enough) — but I don't want to just keep pushing it to +25/+30 and beyond, because sooner or later I'll hit a ceiling where there's no more room to compensate, and the underlying problem will still be there. I'd like to understand the actual root cause instead of just kicking the can down the road. Is RMA'ing the CPU worth it given this history (6 months of normal operation, then gradually worsening boost instability that needs more and more voltage compensation over time)? Or could this be something other than the CPU that I haven't checked yet? Any ideas or similar experiences would be appreciated.