Skip to content
AIPCs

Best mini PCs for running AI models locally (2026)

By AIPCs Editor Updated Researched, not tested

We earn a commission when you buy through retailer links on this page, at no cost to you. It never decides what we recommend. As an Amazon Associate I earn from qualifying purchases. How we make money

Quick answer

For most people who want to run large models locally, the best mini PC is a 128 GB AMD Ryzen AI Max+ 395 machine such as the GMKtec EVO-X2 ($3,499.99 at GMKtec in September 2026). It holds gpt-oss 120B and should write it at about 40 to 55 tokens per second. Apple's Mac Studio with a 128 GB M5 Max is the fastest option for about $1,900 more, NVIDIA's DGX Spark is the pick if you need CUDA, and a 96 GB Ryzen AI 9 HX 370 mini PC is the budget way in.

Our picks at a glance
Pick Memory Bandwidth Price Store
Best overall GMKtec EVO-X2 128 GB at 256 GB/s for $3,499.99, the lowest price for 128 GB we found 128 GB unified 256 GB/s $3,499.99at GMKtec, Sep 25, 2026 Check GMKtec
Fastest Apple Mac Studio (M5 Max, 128 GB) 614 GB/s with the same 128 GB; the quickest way to run big models 128 GB unified 614 GB/s $5,399list price Check Apple
Best for AI developers NVIDIA DGX Spark NVIDIA's CUDA software stack in a mini PC 128 GB unified 273 GB/s $4,699list price Check Amazon
Best for expansion Minisforum MS-S1 Max PCIe slot, 80 Gbps USB4 v2 and dual 10 Gigabit Ethernet 128 GB unified 256 GB/s $3,799at Minisforum, Sep 25, 2026 Check Minisforum
Best small and quiet Apple Mac mini (M5 Pro, 64 GB) 307 GB/s in a silent box, but 64 GB at most 64 GB unified 307 GB/s $3,199list price Check Apple
Best budget Minisforum AI X1 Pro-370 (96 GB) 96 GB of upgradeable memory for about $2,000; slower, but holds big models 96 GB 89.6 GB/s $1,983at Minisforum, Sep 25, 2026 Check Minisforum

A mini PC for local AI lives or dies by two numbers. Memory size decides which models fit, and memory bandwidth decides how fast they write, because every word means reading the model’s active weights once. Our study of 23 chips shows speed tracking bandwidth almost perfectly. So instead of ranking by processor names, we ranked these machines on what they can actually run, and what that costs in September 2026.

We have not tested these machines ourselves yet; they are marked Researched. Specifications come from the makers, prices from the makers’ own stores on the date shown, and the speed estimates from the method we validated on our own hardware.

What each pick can run

Estimated writing speed, tokens per second
Machine gpt-oss 20B12.11 GB fileQwen3.6 35B-A3B22.13 GB fileLlama 3.3 70B42.52 GB filegpt-oss 120B63.39 GB file
GMKtec EVO-X2 128 GB, 256 GB/s 57-7456-734-541-53
NVIDIA DGX Spark 128 GB, 273 GB/s 61-7960-774-543-56
Apple Mac Studio (M5 Max, 128 GB) 128 GB, 614 GB/s 96-12494-1226-868-88
Minisforum MS-S1 Max 128 GB, 256 GB/s 57-7456-734-541-53
Apple Mac mini (M5 Pro, 64 GB) 64 GB, 307 GB/s 68-8767-864-5No fit
Minisforum AI X1 Pro-370 (96 GB) 96 GB, 89.6 GB/s 16-2015-20111-14
Estimated tokens per second with an 8K-token conversation, from memory bandwidth and each model's file, using the method in our bandwidth study. These machines are not tested by us; treat figures as estimates. Try any combination in the calculator.

About 10 tokens per second reads comfortably. Two things stand out. Mixture-of-experts models (gpt-oss and the “A3B” Qwen models) run fast on every machine that can hold them. Large dense models such as Llama 3.3 70B are where bandwidth really shows: only the Mac Studio runs one at a reasonable pace. (“No fit” means the model does not fit in that machine’s memory.)

The picks

GMKtec EVO-X2

128 GB / 1 TB (also sold with 64 GB). AMD Ryzen AI Max+ 395 (16 cores / 32 threads, up to 5.1 GHz)

Memory
128 GB unified
Bandwidth
256 GB/s
Price
$3,499.99

A Ryzen AI Max+ 395 mini PC with 128 GB of fast unified memory, and the lowest-priced 128 GB machine among those we compared in September 2026. It holds gpt-oss 120B and 70B-class models.

With 128 GB of unified memory at 256 GB/s, the EVO-X2 holds every popular open model up to gpt-oss 120B, and at GMKtec’s September 2026 prices it is the least expensive 128 GB machine we found: $3,499.99 with 1 TB, or $3,649.99 with 2 TB. The same chip and memory cost $150 to $700 more from other brands. If you only need models up to about 35B, the 64 GB version is $2,199.99.

What we like

  • 128 GB of 256 GB/s unified memory for less than rivals
  • Two M.2 slots
  • Also sold with 64 GB for $1,300 less

What we don't

  • Memory is soldered; choose the size you need up front
  • Only 2.5 Gigabit Ethernet, where some rivals have 10 Gigabit

Where to buy

$3,499.99 at GMKtec for 128 GB / 1 TB on Sep 25, 2026; list price $3,999.99 (GMKtec store regular price for 128 GB / 1 TB; checked Sep 25, 2026). Prices change often; the buttons show today's price at each store.

Apple Mac Studio (M5 Max, 128 GB)

M5 Max 18-core CPU, 40-core GPU, 128 GB, 1 TB. Apple M5 Max, 18-core CPU

Memory
128 GB unified
Bandwidth
614 GB/s
Price
$5,399

The fastest way to run 70B to 120B-class models without a server. The 40-core M5 Max has 614 GB/s of memory bandwidth, more than twice a Ryzen AI Max+ 395 or DGX Spark, with the same 128 GB of memory.

The M5 Max’s 614 GB/s is more than twice the bandwidth of the 128 GB mini PCs, which matters most for large dense models: we estimate 6 to 8 tokens per second for Llama 3.3 70B against 4 to 5 on a Ryzen AI Max+ 395, and published results for the M5 Max on our reference model beat our estimate. At $5,399 as configured on Apple’s store it is the most expensive pick, and it is technically a small desktop rather than a mini PC.

What we like

  • 614 GB/s, more than twice the 128 GB mini PCs
  • Quiet under load

What we don't

  • The most expensive 128 GB option here
  • macOS only, no CUDA

Where to buy

list price $5,399 (Apple store as configured, $3,799 for 64 GB / 1 TB plus $1,600 for 128 GB; checked Sep 25, 2026). Prices change often; the buttons show today's price at each store.

Best for AI developers Researched, not tested

NVIDIA DGX Spark

128 GB / 4 TB. 20-core Arm (10 Cortex-X925 + 10 Cortex-A725)

Memory
128 GB unified
Bandwidth
273 GB/s
Price
$4,699

NVIDIA's own desktop AI computer, built on the GB10 chip with 128 GB of unified memory. It writes at about the same speed as a Ryzen AI Max+ 395 machine, but runs NVIDIA's CUDA software stack, which most AI development tools support first.

The DGX Spark writes at about the same speed as the Ryzen AI Max+ 395 machines, so don’t buy it for chat speed. Buy it if you develop AI software: it runs NVIDIA’s CUDA stack, reads long prompts several times faster in published llama.cpp results, and can be linked to a second unit over its 200 Gbps ConnectX-7 port. ASUS, Dell, HP, Lenovo, MSI, Acer and Gigabyte sell machines with the same GB10 chip, often with smaller drives, but prices vary widely, so compare.

What we like

  • Full NVIDIA CUDA software stack
  • 200 Gbps ConnectX-7 to link two units for bigger models
  • Small for its power

What we don't

  • Writes no faster than cheaper Ryzen AI Max+ 395 machines
  • Linux only (DGX OS), not a general-purpose Windows PC
  • Often sells above list price

Where to buy

list price $4,699 (NVIDIA list price since February 23, 2026 (previously $3,999); checked Feb 23, 2026). Prices change often; the buttons show today's price at each store.

Best for expansion Researched, not tested

Minisforum MS-S1 Max

128 GB / 2 TB. AMD Ryzen AI Max+ 395 (16 cores / 32 threads, up to 5.1 GHz)

Memory
128 GB unified
Bandwidth
256 GB/s
Price
$3,799

A Ryzen AI Max+ 395 workstation in mini PC form, with 128 GB of unified memory plus things most rivals lack, including a full-length PCIe slot, 80 Gbps USB4 v2 ports, dual 10 Gigabit Ethernet and a built-in power supply.

Same chip and memory as the EVO-X2, with a full-length PCIe slot, 80 Gbps USB4 v2 ports and two 10 Gigabit Ethernet ports. That makes it the better base for adding a network card, fast external storage, or linking machines together. It is bigger and heavier, and $149 more than the EVO-X2 with 2 TB.

What we like

  • PCIe slot and 80 Gbps USB4 v2 ports for expansion
  • Dual 10 Gigabit Ethernet
  • Internal power supply, no power brick

What we don't

  • Larger and heavier than most mini PCs (2.8 kg)
  • Costs $149 more than the EVO-X2 with 2 TB

Where to buy

$3,799 at Minisforum for 128 GB / 2 TB on Sep 25, 2026; list price $4,749 (Minisforum store regular price for 128 GB / 2 TB; checked Sep 25, 2026). Prices change often; the buttons show today's price at each store.

Best small and quiet Researched, not tested

Apple Mac mini (M5 Pro, 64 GB)

M5 Pro 18-core CPU, 20-core GPU, 64 GB, 1 TB. Apple M5 Pro, 18-core CPU

Memory
64 GB unified
Bandwidth
307 GB/s
Price
$3,199

The small, silent way to run mid-size local models. The M5 Pro's 307 GB/s of unified memory bandwidth is about 20% more than a Ryzen AI Max+ 395, but 64 GB is the maximum, so 120B-class models don't fit.

At 307 GB/s the M5 Pro has about 20% more bandwidth than a Ryzen AI Max+ 395, in a far smaller and quieter box that sips power when idle. The catch is capacity: 64 GB is the maximum, which rules out 120B-class models and leaves Llama 3.3 70B only just fitting. For models up to about 35B it is excellent.

What we like

  • 307 GB/s memory bandwidth in a tiny, near-silent box
  • Very low power draw for always-on use

What we don't

  • 64 GB maximum, fixed at purchase
  • Apple charges a lot for memory upgrades

Where to buy

list price $3,199 (Apple store price for M5 Pro 18-core CPU / 20-core GPU / 64 GB / 1 TB; checked Sep 25, 2026). Prices change often; the buttons show today's price at each store.

Minisforum AI X1 Pro-370 (96 GB)

96 GB DDR5 / 2 TB. AMD Ryzen AI 9 HX 370 (12 cores / 24 threads, up to 5.1 GHz)

Memory
96 GB
Bandwidth
89.6 GB/s
Price
$1,983

A mainstream Copilot+ mini PC that takes up to 128 GB of standard DDR5 memory. Its 89.6 GB/s bandwidth makes it slow for large dense models, but with 96 GB it can hold mixture-of-experts models up to gpt-oss 120B for well under the price of a 128 GB Ryzen AI Max machine.

This is a mainstream Copilot+ mini PC, not an AI workstation, but it takes standard DDR5 memory, and with 96 GB it can hold mixture-of-experts models up to gpt-oss 120B. Its 89.6 GB/s is about a third of a Ryzen AI Max+ 395, so expect roughly 11 to 14 tokens per second on gpt-oss 120B and 15 to 20 on the 35B-A3B models: usable, not fast. Our own desktop processor, with similar memory, wrote gpt-oss 20B at 21.8 tokens per second (test T0001). The OCuLink port lets you add a desktop graphics card later.

What we like

  • Upgradeable DDR5 memory, up to 128 GB
  • OCuLink port to add a desktop graphics card later
  • Copilot+ PC with a 50 TOPS NPU

What we don't

  • 89.6 GB/s is about a third of a Ryzen AI Max+ 395
  • Large dense models run too slowly for chat

Where to buy

$1,983 at Minisforum for 96 GB / 2 TB on Sep 25, 2026; list price $2,479 (Minisforum store regular price for 96 GB / 2 TB; checked Sep 25, 2026). Prices change often; the buttons show today's price at each store.

Also worth a look

  • Other GB10 machines, such as the ASUS Ascent GX10. Same chip as the DGX Spark; worth it when you find one well below the DGX Spark’s price.
  • Framework Desktop. A Ryzen AI Max+ 395 machine sold only by Framework, with 64 or 128 GB. Check Framework’s current price and shipping time.

Why prices are so high right now

The fast memory these machines depend on is in short supply in 2026. NVIDIA raised the DGX Spark’s list price from $3,999 to $4,699 in February, and Beelink’s GTR9 Pro, which launched at $1,985 in 2025, now sells for $4,349. We don’t expect a quick reversal, so we would buy the memory size you need now rather than wait for a small version to get cheaper. Our guide to how much memory local AI needs helps you pick the size.

How we chose

We started from the models people most want to run locally in 2026 and worked out, from each model’s actual file, how much memory and bandwidth they need. We then compared every mini PC we could find with 64 GB or more of fast memory, or 96 GB or more of standard memory, on price per gigabyte and estimated speed, using the prices on each maker’s store on September 25, 2026. None of these machines has been through our own tests yet; when one has, its label changes to Tested and its figures come from our benchmarks.

Questions people ask

How much memory does a mini PC need for local AI?

64 GB runs models up to about 35 billion parameters and the popular mixture-of-experts models such as Qwen3.6 35B-A3B. For 70B dense models and 120B-class models such as gpt-oss 120B you need 96 to 128 GB. Memory in the fastest mini PCs is soldered, so buy the amount you will need.

Is the Ryzen AI Max+ 395 or the DGX Spark faster?

They write at about the same speed, because their memory bandwidth is similar (256 versus 273 GB/s). In published llama.cpp results the DGX Spark's chip reads long prompts several times faster, and it runs NVIDIA's CUDA software. The Ryzen machines cost less and run Windows or Linux as ordinary PCs.

Can a mini PC run a 70B model?

Yes, with 64 GB or more of unified memory, but slowly: we estimate 4 to 5 tokens per second on a 256 GB/s Ryzen AI Max machine and 6 to 8 on a Mac Studio M5 Max. Mixture-of-experts models of similar or larger size, such as gpt-oss 120B, run about ten times faster because they read only a small part of their weights per word.

Why are these mini PCs so expensive in 2026?

A global memory shortage has pushed up the price of the fast memory these machines depend on. NVIDIA raised the DGX Spark's list price from $3,999 to $4,699 in February 2026, and 128 GB Ryzen AI Max+ 395 mini PCs that launched at around $2,000 in 2025 now sell for about $3,500 or more.

Sources

  1. GMKtec EVO-X2 product page (prices), checked Sep 25, 2026
  2. Minisforum MS-S1 Max product page (prices), checked Sep 25, 2026
  3. Apple Mac Studio and Mac mini specifications, checked Sep 25, 2026
  4. NVIDIA DGX Spark specifications, checked Sep 25, 2026
  5. NVIDIA Developer Forums: DGX Spark price change announcement, checked Sep 25, 2026
  6. Wccftech: Beelink GTR9 Pro launched at $1,985, checked Sep 25, 2026
  7. llama.cpp discussion #10879: Vulkan results including NVIDIA GB10, checked Sep 25, 2026
  8. llama.cpp discussion #15021: ROCm results including Radeon 8060S, checked Sep 25, 2026

What changed

  • Sep 25, 2026: First published. Prices checked at each maker's store on September 25, 2026.