Beelink GTR9 Pro (Ryzen AI Max+ 395, 128GB)
Node Kits

Beelink GTR9 Pro (Ryzen AI Max+ 395, 128GB)

4/5

$4,349

The best-connected 128GB Strix Halo box. Same Ryzen AI Max+ 395 silicon and 128GB LPDDR5X-8000 as the rest, but the GTR9 Pro adds dual 10GbE and dual USB4 — the IO you want for clustering boxes or pulling models off a fast NAS. A vapor-chamber cooler keeps sustained ~120W inference quiet. Watch for driver-dependent 10GbE instability under heavy GPU load.

Affiliate links — We earn a commission on qualifying purchases at no cost to you.

Specifications

APUAMD Ryzen AI Max+ 395 (16C/32T, Zen 5)
GPURadeon 8060S (40 CU, RDNA 3.5)
NPU50 TOPS (XDNA 2)
Unified Memory128GB LPDDR5X-8000
Memory Bandwidth~256–273 GB/s
Storage2TB NVMe (dual M.2 2280, up to 16TB)
NetworkingDual 10GbE (Intel E610), Wi-Fi 7
I/O2× USB4 (40Gbps), HDMI 2.1, DP 2.1

Pros

  • Dual 10GbE — rare at this size; fast model/dataset transfer to a NAS or cluster
  • Dual USB4 + vapor-chamber cooling keep sustained inference quiet — 125–128W at 39–41 dBA on gpt-oss 120B (ServeTheHome)
  • 128GB unified runs Q4 70B-class models at a price no consumer dGPU touches

Cons

  • Reported 10GbE NIC instability/BSOD under heavy GPU load — driver-dependent
  • ~256–273 GB/s bandwidth ceiling — 70B is usable but not GPU-fast
  • 128GB soldered, non-upgradable — buy max config or nothing

Related Articles

Guide14 min read

192GB vs 128GB Unified Memory: What the Extra 64GB Actually Buys You

IFA 2026 filled the feeds with 192GB Gorgon Halo mini PCs. Capacity went up 50%; bandwidth went up about 7%. That asymmetry decides the whole purchase — here is the model-by-model list of what only 192GB runs, and why most buyers should still buy 128GB today.

Guide14 min read

Can You Fine-Tune an LLM on a 128GB Mini PC? (And Which Box to Buy in 2026)

LoRA and QLoRA on 20–30B models are practical on 128GB of unified memory; full fine-tunes stop near 12B; dense 70B training doesn't happen on any box in this class. Fine-tuning is the one local-AI workload where the software stack — not memory bandwidth — decides which machine you buy. Here's the capability table, the CUDA tax, and the cloud break-even.

Guide13 min read

Strix Halo Memory Bandwidth: Why 256 GB/s Isn't 256 GB/s (2026)

AMD's Ryzen AI Max+ 395 is rated 256 GB/s. Real sustained bandwidth is around 215 GB/s — about 84% of spec. Here's where the missing 40 GB/s goes, why 32MB of Infinity Cache doesn't rescue it, and how to turn GB/s into a tokens-per-second estimate before you buy.

Guide14 min read

Best Mini PC for a Local Coding Agent in 2026 — Why Prefill, Not Tokens/Sec, Decides Your Box

A coding agent re-sends your whole repo context every single turn, which makes it a prefill-bound workload. That flips the buying logic: ~1,700 tok/s prompt processing on a GB10 box vs ~340 tok/s on Strix Halo, while generation is a near-tie. Here's the per-budget verdict, the memory math, and the boxes to skip — with the caveat that the 2026 DRAM spike has narrowed the price gap to about $1,250.

Disclosure: Some links on this page are affiliate links. We may earn a commission if you make a purchase — at no extra cost to you. This helps support our independent reviews.