Beelink GTR9 Pro (Ryzen AI Max+ 395, 128GB)
Node Kits

Beelink GTR9 Pro (Ryzen AI Max+ 395, 128GB)

4/5

$1,899 – $1,999

The best-connected 128GB Strix Halo box. Same Ryzen AI Max+ 395 silicon and 128GB LPDDR5X-8000 as the rest, but the GTR9 Pro adds dual 10GbE and dual USB4 — the IO you want for clustering boxes or pulling models off a fast NAS. A vapor-chamber cooler keeps sustained ~120W inference quiet. Watch for driver-dependent 10GbE instability under heavy GPU load.

Affiliate links — We earn a commission on qualifying purchases at no cost to you.

Specifications

APUAMD Ryzen AI Max+ 395 (16C/32T, Zen 5)
GPURadeon 8060S (40 CU, RDNA 3.5)
NPU50 TOPS (XDNA 2)
Unified Memory128GB LPDDR5X-8000
Memory Bandwidth~256–273 GB/s
Storage2TB NVMe (dual M.2 2280, up to 16TB)
NetworkingDual 10GbE (Intel E610), Wi-Fi 7
I/O2× USB4 (40Gbps), HDMI 2.1, DP 2.1

Pros

  • Dual 10GbE — rare at this size; fast model/dataset transfer to a NAS or cluster
  • Dual USB4 + vapor-chamber cooling keep sustained ~120W inference quiet (~36–41 dBA)
  • 128GB unified runs Q4 70B-class models at a price no consumer dGPU touches

Cons

  • Reported 10GbE NIC instability/BSOD under heavy GPU load — driver-dependent
  • ~256–273 GB/s bandwidth ceiling — 70B is usable but not GPU-fast
  • 128GB soldered, non-upgradable — buy max config or nothing

Related Articles

Guide13 min read

Beelink GTR9 Pro Review (2026): The 128GB Strix Halo Box With Dual 10GbE — Real Local-LLM Benchmarks & Who Should Buy

The Beelink GTR9 Pro is a ~$1,899–$1,999 128GB Ryzen AI Max+ 395 mini PC that runs GPT-OSS 120B at ~31 tok/s (~120W) but a dense 70B at only ~5 tok/s. Its real hook is dual 10GbE — plus a real caveat: reported 10GbE instability under heavy GPU load. Benchmarks, thermals, and how it stacks up against the EVO-X2, Framework Desktop, and DGX Spark.

Tutorial14 min read

How to Unlock the Full 128GB as VRAM on a Ryzen AI Max+ 395 (Strix Halo): BIOS + Linux GTT Guide

You bought a 128GB Strix Halo box and the GPU only sees ~16GB. Here's the fix: Windows caps GPU-allocatable memory at 96GB via the BIOS UMA frame buffer, but Linux with the amdttm GTT kernel params reaches ~110–120GB — and the winning move is counterintuitive.

Guide13 min read

How Fast Is Strix Halo, Really? Real Tokens-Per-Second for Local LLMs (Dense vs MoE)

On the Ryzen AI Max+ 395, a dense 70B model crawls at ~5 tok/s — but a 30B MoE model hits 70–100 tok/s on the same box. Here's the full tokens-per-second table by model size, why memory bandwidth caps it, and why MoE changes the whole buying decision.

Guide13 min read

GMKtec EVO-X2 Review (2026): The 128GB Ryzen AI Max+ 395 Mini PC That Runs 70B Models — Real Benchmarks & Who Should Buy

The GMKtec EVO-X2 is a ~$2,000 128GB Strix Halo box that holds a 4-bit 70B model and runs it at 5–10 tok/s (7B at 50–80). Real per-model benchmarks, thermals, the ROCm reality, and whether it beats the Framework Desktop, Beelink GTR9 Pro, and $3,999 DGX Spark.

Disclosure: Some links on this page are affiliate links. We may earn a commission if you make a purchase — at no extra cost to you. This helps support our independent reviews.