whichaipc

AI PC · mini-pc

Beelink GTR9 Pro (Ryzen AI Max+ 395)

The premium Strix Halo build - dual 10GbE and a vapour chamber - but you pay a lot more for the same 128GB.

Memory
128 GB
Bandwidth
256 GB/s
AI compute
50 TOPS
£ / GB
£31

What it runs

With 128 GB of memory you can load, at 4-bit, a model up to roughly

~252B params

Verdict

The best-built Strix Halo box - dual 10GbE, quiet, premium - but priced well above the same chip elsewhere on the Western store.

BEST FOR

  • + a premium build with fast networking
  • + AI server clustering

NOT FOR

  • - value hunters
  • - anyone who needs CUDA

The Beelink GTR9 Pro is the same Ryzen AI Max+ 395 and 128GB you’ll find in cheaper boxes, dressed in a much nicer suit. Dual 10-gigabit networking, a full vapour chamber, a built-in power supply and a metal chassis that could pass for a Mac. It’s the premium take on Strix Halo - and on Beelink’s Western store, it’s priced like one.

The take

The build is the best of the Strix Halo bunch, and the pitch is clustering. Two 10GbE ports and dual USB4 mean you can lash several together into a little AI cluster, the cooling is a proper vapour chamber, and there’s a 230W supply built in so there’s no brick on the desk. All lovely. The problem is price. On Beelink’s own UK and US store the 128GB model sits north of four thousand, more than double what the near-identical GMKtec asks, and the same box goes for roughly 1,800 dollars in China. So the hardware’s great, but the value depends enormously on where and how you buy it. Go in with your eyes open.

What it’ll actually run

It runs exactly what any 128GB Strix Halo box runs, because it’s the same silicon. Hand most of the 128GB to the Radeon 8060S and a 70B at Q4 fits with context to spare, MoE models like gpt-oss-120b sit comfortably, and the everyday 8B to 32B range is quick. Beelink quotes a Qwen 32B at Q8 around 6 tokens a second in LM Studio, which is about right for a dense model at that precision. MoE models do far better.

The speed ceiling is bandwidth, same story as its siblings: near 256 GB/s on paper, lower in the real world, so dense large models answer at a measured pace. Worth knowing: owners found the 128GB memory can run in a throttled mode until you unlock it in the BIOS, so check your configuration. It’s Windows 11 out of the box, runs Linux well - owners report Ollama and the Lemonade server working nicely on Ubuntu - and it’s quiet, though reviewers note it isn’t silent when you push it to full power.

Who should buy it

Buy it if you want the nicest-built Strix Halo machine and you’ll actually use the fast networking. For clustering boxes or moving big models around a home lab, the dual 10GbE earns its keep. If you can get it near the China price, it’s a really nice bit of kit.

Think twice at the Western price. You’re paying a heavy premium over the GMKtec for the same chip and memory, so unless the build and the 10GbE matter to you specifically, the cheaper box loads the same models just as well. And if you need CUDA, no AMD machine is your answer. Great hardware, just buy it at the right price.

Settings people actually run

The configs owners land on, pulled from the community. A sensible starting point, not gospel - tune to your own kit.

Biggest model that fits

gpt-oss-120b MoE (~60GB); BIOS memory split set high

Same 128GB pool as any Strix Halo box; hand most of it to the GPU.

Dense large model

Qwen 32B Q8 / Llama 70B Q4

Beelink quotes Qwen 32B Q8 around 6 tokens/s; MoE models are much quicker.

Unlock full bandwidth

Check the memory mode in the BIOS

Owners found 128GB can sit in a throttled mode until unlocked; verify your configuration.

Clustering

Dual 10GbE + dual USB4 for AI server clusters

The reason to pick this over a cheaper box: fast networking for lashing several together.

What owners report

Real first-hand experience gathered from owners and the community.

  • An owner on Ubuntu 24.04 ran a 96GB VRAM / 32GB system split, with idle around 15W and AI-load draw up to about 180W at the wall; Ollama and the Lemonade AMD server both worked well.

    Beelink store owner review

  • The GTR9 Pro's 128GB memory was found to run in a throttled 1:4 mode reporting about 72 GB/s until unlocked, so BIOS configuration matters for bandwidth.

    Beelink community forum

  • Reviewers rate the build and features highly but note it isn't silent at full 140-160W performance; the same chip sells for far less as the GMKtec EVO-X2.

    ServeTheHome

Fact-checked 19 Jul 20263 claims verified against primary sources.
4 claim(s) we couldn't fully verify
  • · memory_bandwidth_gbs 256 - Beelink doesn't publish a bandwidth figure; 256 GB/s is the Strix Halo theoretical, real-world is lower, and owners found a throttled mode reporting about 72 GB/s until unlocked in the BIOS.
  • · Price around USD 4,349 / GBP 3,999 - The figure shown on Beelink's Western store for 128GB + 2TB; the same unit sells for roughly 1,800 dollars in China, so pricing is highly channel-dependent. The GBP figure is an estimate.
  • · power_w 140 - 140W is Beelink's full-performance figure; owners measured up to about 180W at the wall under AI load. Built-in 230W supply.
  • · affiliate ASIN B0GQXDCKN1 - A live Amazon US listing for the GTR9 Pro 128GB; not cross-checked against every regional catalogue.

Hands-on reviews we drew on

We don't just copy the spec sheet. These are the teardowns and hands-on reviews behind this page - worth watching in their own right.

Common questions

Why is the Beelink GTR9 Pro so much more expensive than the GMKtec EVO-X2?+

It's the same chip and the same 128GB, so it's not about capability. Beelink's Western store prices the 128GB GTR9 Pro north of four thousand, more than double the near-identical GMKtec, while the same box goes for roughly 1,800 dollars in China. You're paying for the build and the dual 10GbE, plus a hefty channel premium.

What can the GTR9 Pro run?+

Exactly what any 128GB Strix Halo box runs. Hand most of the 128GB to the Radeon 8060S and a 70B at Q4 fits with context to spare, MoE models like gpt-oss-120b sit comfortably, and the everyday 8B to 32B range is quick.

Is it good for AI clustering?+

That's its party trick. Dual 10GbE ports and dual USB4 let you lash several units into a small AI cluster, which is the main reason to pick this over a cheaper box with the same silicon.

Does it run silently?+

It's quiet, with a proper vapour chamber and dual fans, but reviewers note it isn't silent when you push it to full 140W. For always-on ticking over it's fine; under sustained load you'll hear it.