Skip to content
AIPCs

The best computer for running AI locally, at every budget (2026)

By AIPCs Editor Updated Researched, not tested

We earn a commission when you buy through retailer links on this page, at no cost to you. It never decides what we recommend. As an Amazon Associate I earn from qualifying purchases. How we make money

Quick answer

The best computer for running AI locally depends on the size of model you want. For about $550, a 16 GB AMD RX 9060 XT graphics card turns a desktop you already own into a fast local AI machine for models up to about 20B. Without a desktop, a $1,499 MacBook Air with 24 GB does the same in a laptop. For 70B and 120B-class models you need 96 to 128 GB of memory: a $1,983 Minisforum AI X1 Pro holds them slowly, a $3,499.99 GMKtec EVO-X2 with 128 GB runs gpt-oss 120B at an estimated 41 to 53 tokens per second, and a $5,399 Mac Studio with M5 Max is fastest.

Our picks at a glance
Pick Memory Bandwidth Price Store
About $550: upgrade your PC XFX Swift Radeon RX 9060 XT 16GB 16 GB graphics card for a desktop you already own; gpt-oss 20B at 54 to 70 tokens/s 16 GB 320 GB/s $549.99at Newegg (sold by XFX), Sep 26, 2026 Check Newegg
About $900: small and silent Apple Mac mini (M6, 16 GB) Runs models up to about 14B at reading speed, and cloud-model agents all day 16 GB unified 153 GB/s $899list price Check Apple
About $1,500: a laptop Apple MacBook Air 13-inch (M5, 24 GB) 24 GB holds gpt-oss 20B, in a fanless laptop 24 GB unified 153 GB/s $1,499list price Check Apple
About $2,000: most memory Minisforum AI X1 Pro-370 (96 GB) 96 GB holds gpt-oss 120B, slowly, for the least money 96 GB 89.6 GB/s $1,983at Minisforum, Sep 25, 2026 Check Minisforum
About $3,500: big models GMKtec EVO-X2 128 GB at 256 GB/s runs gpt-oss 120B at 41 to 53 tokens/s 128 GB unified 256 GB/s $3,499.99at GMKtec, Sep 25, 2026 Check GMKtec
About $5,400: fastest Apple Mac Studio (M5 Max, 128 GB) 614 GB/s across 128 GB, the quickest way to run 120B-class models 128 GB unified 614 GB/s $5,399list price Check Apple

Running AI models on your own computer comes down to two numbers: how much memory the models can use, which decides how big a model fits, and how fast that memory is, which decides how quickly it writes. Our study of 23 chips shows writing speed tracking memory bandwidth almost exactly. So each step up in budget below buys either more memory, faster memory, or both.

Prices are the ones we recorded in September 2026, when a memory shortage had pushed many of them up.

What each budget can run

Estimated writing speed, tokens per second
Machine Qwen3.5 9B5.68 GB filegpt-oss 20B12.11 GB fileQwen3.6 35B-A3B22.13 GB fileLlama 3.3 70B42.52 GB filegpt-oss 120B63.39 GB file
XFX Swift Radeon RX 9060 XT 16GB 16 GB, 320 GB/s 40-5254-7029-39*No fitNo fit
Apple Mac mini (M6, 16 GB) 16 GB, 153 GB/s 17-23No fitNo fitNo fitNo fit
Apple MacBook Air 13-inch (M5, 24 GB) 24 GB, 153 GB/s 17-2334-44No fitNo fitNo fit
Minisforum AI X1 Pro-370 (96 GB) 96 GB, 89.6 GB/s 8-1016-2015-20111-14
GMKtec EVO-X2 128 GB, 256 GB/s 29-3857-7456-734-541-53
Apple Mac Studio (M5 Max, 128 GB) 128 GB, 614 GB/s 49-6396-12494-1226-868-88
Estimated tokens per second with an 8K-token conversation, from memory bandwidth and each model's file, using the method in our bandwidth study. * Part of the model runs from system RAM, which is slower (for a graphics card on its own we assume a PC with 32 GB of DDR5). These machines are not tested by us; treat figures as estimates. Try any combination in the calculator.

About 10 tokens per second reads comfortably. The graphics card row assumes a desktop with 32 GB of system memory; bigger models spill into it and slow down.

The picks

About $550: upgrade your PC Researched, not tested

XFX Swift Radeon RX 9060 XT 16GB

16 GB GDDR6, dual fan (RX-96TSW16BQ). AMD Radeon RX 9060 XT (RDNA 4)

Video memory
16 GB
Bandwidth
320 GB/s
Price
$549.99

The least expensive 16 GB graphics card we found in September 2026. At 320 GB/s it runs models up to about 14B, plus gpt-oss 20B, entirely on the card at well above reading speed, though more slowly than an RTX 5060 Ti.

If you have a desktop with a free graphics slot and a 450 W power supply, this is the most local AI per dollar in 2026. Its 16 GB of video memory holds models up to about 14B and gpt-oss 20B, which it should write at 54 to 70 tokens per second. NVIDIA’s RTX 5060 Ti is faster and simpler to set up but cost about $230 more in September; our graphics card guide compares them.

What we like

  • 16 GB for about $230 less than the cheapest RTX 5060 Ti we found
  • Runs gpt-oss 20B at an estimated 54 to 70 tokens per second
  • 160 W typical board power; AMD recommends a 450 W power supply

What we don't

  • Slower than an RTX 5060 Ti, especially on mixture-of-experts models
  • Some local AI tools need more setup on AMD than on NVIDIA

Where to buy

$549.99 at Newegg (sold by XFX) for 16 GB on Sep 26, 2026; list price $349 (AMD launch price for the 16 GB model, June 2025; checked Sep 26, 2026). Prices change often; the buttons show today's price at each store.

About $900: small and silent Researched, not tested

Apple Mac mini (M6, 16 GB)

M6 12-core CPU, 12-core GPU, 16 GB, 256 GB. Apple M6, 12-core CPU

Memory
16 GB unified
Bandwidth
153 GB/s
Price
$899

Apple's base Mac mini. Small, silent and frugal, it is the default machine for an always-on AI agent that uses cloud models, and it runs small local models; 16 GB rules out anything larger.

The base Mac mini is the cheapest computer we would buy new for local AI. Its 16 GB of unified memory runs models up to about 14B at reading speed, and it is the classic always-on machine for AI agents that use cloud models. It can’t hold gpt-oss 20B; for that, you need 24 GB or more.

What we like

  • Silent and very efficient for 24/7 use
  • Deep macOS integration for agent tools

What we don't

  • 16 GB limits local models to small ones
  • Memory can't be upgraded later

Where to buy

list price $899 (Apple store starting price; checked Sep 25, 2026). Prices change often; the buttons show today's price at each store.

About $1,500: a laptop Researched, not tested

Apple MacBook Air 13-inch (M5, 24 GB)

M5 10-core CPU, 10-core GPU, 24 GB, 512 GB. Apple M5, 10-core CPU

Memory
24 GB unified
Bandwidth
153 GB/s
Price
$1,499

Apple's thin, fanless laptop, configured with enough memory for mid-size local models. 24 GB of unified memory at 153 GB/s runs gpt-oss 20B and models up to about 14B at reading speed, alongside Apple's built-in AI features.

For a laptop, choose the MacBook Air with 24 GB rather than 16: the extra $100 at Apple is what lets it hold gpt-oss 20B, at an estimated 34 to 44 tokens per second. Our AI laptops guide covers Windows alternatives and 128 GB laptops.

What we like

  • 24 GB fits gpt-oss 20B, which the 16 GB model cannot hold
  • Silent, with no fan
  • Only $100 more than the 16 GB model at Apple

What we don't

  • Memory can't be upgraded later
  • No fan, so very long jobs can slow down as it heats up
  • Too little memory for 70B or 120B-class models

Where to buy

list price $1,499 (Apple store price for 24 GB / 512 GB; $1,399 with 16 GB; checked Sep 26, 2026). Prices change often; the buttons show today's price at each store.

About $2,000: most memory Researched, not tested

Minisforum AI X1 Pro-370 (96 GB)

96 GB DDR5 / 2 TB. AMD Ryzen AI 9 HX 370 (12 cores / 24 threads, up to 5.1 GHz)

Memory
96 GB
Bandwidth
89.6 GB/s
Price
$1,983

A mainstream Copilot+ mini PC that takes up to 128 GB of standard DDR5 memory. Its 89.6 GB/s bandwidth makes it slow for large dense models, but with 96 GB it can hold mixture-of-experts models up to gpt-oss 120B for well under the price of a 128 GB Ryzen AI Max machine.

This is the cheapest way we found to hold 120B-class models: 96 GB of upgradeable DDR5 for $1,983 at Minisforum. The memory is slower than the 128 GB machines below, so gpt-oss 120B should write at about 11 to 14 tokens per second, just above reading speed, and 70B dense models are too slow to be practical.

What we like

  • Upgradeable DDR5 memory, up to 128 GB
  • OCuLink port to add a desktop graphics card later
  • Copilot+ PC with a 50 TOPS NPU

What we don't

  • 89.6 GB/s is about a third of a Ryzen AI Max+ 395
  • Large dense models run too slowly for chat

Where to buy

$1,983 at Minisforum for 96 GB / 2 TB on Sep 25, 2026; list price $2,479 (Minisforum store regular price for 96 GB / 2 TB; checked Sep 25, 2026). Prices change often; the buttons show today's price at each store.

About $3,500: big models Researched, not tested

GMKtec EVO-X2

128 GB / 1 TB (also sold with 64 GB). AMD Ryzen AI Max+ 395 (16 cores / 32 threads, up to 5.1 GHz)

Memory
128 GB unified
Bandwidth
256 GB/s
Price
$3,499.99

A Ryzen AI Max+ 395 mini PC with 128 GB of fast unified memory, and the lowest-priced 128 GB machine among those we compared in September 2026. It holds gpt-oss 120B and 70B-class models.

For 70B and 120B-class models at a usable speed, this is the value pick: 128 GB of 256 GB/s unified memory for $3,499.99 at GMKtec, the lowest price we found for 128 GB. It should write gpt-oss 120B at 41 to 53 tokens per second. Our mini PC guide compares it with the other 128 GB machines.

What we like

  • 128 GB of 256 GB/s unified memory for less than rivals
  • Two M.2 slots
  • Also sold with 64 GB for $1,300 less

What we don't

  • Memory is soldered; choose the size you need up front
  • Only 2.5 Gigabit Ethernet, where some rivals have 10 Gigabit

Where to buy

$3,499.99 at GMKtec for 128 GB / 1 TB on Sep 25, 2026; list price $3,999.99 (GMKtec store regular price for 128 GB / 1 TB; checked Sep 25, 2026). Prices change often; the buttons show today's price at each store.

About $5,400: fastest Researched, not tested

Apple Mac Studio (M5 Max, 128 GB)

M5 Max 18-core CPU, 40-core GPU, 128 GB, 1 TB. Apple M5 Max, 18-core CPU

Memory
128 GB unified
Bandwidth
614 GB/s
Price
$5,399

The fastest way to run 70B to 120B-class models without a server. The 40-core M5 Max has 614 GB/s of memory bandwidth, more than twice a Ryzen AI Max+ 395 or DGX Spark, with the same 128 GB of memory.

The fastest way to run big models without a server: the 40-core M5 Max reads its 128 GB at 614 GB/s, more than twice the EVO-X2, for an estimated 68 to 88 tokens per second on gpt-oss 120B. It costs about $1,900 more, and memory can’t be upgraded later.

What we like

  • 614 GB/s, more than twice the 128 GB mini PCs
  • Quiet under load

What we don't

  • The most expensive 128 GB option here
  • macOS only, no CUDA

Where to buy

list price $5,399 (Apple store as configured, $3,799 for 64 GB / 1 TB plus $1,600 for 128 GB; checked Sep 25, 2026). Prices change often; the buttons show today's price at each store.

Between the steps

  • About $3,200: a Mac mini with M5 Pro and 64 GB is fast (an estimated 68 to 87 tokens per second on gpt-oss 20B) and quiet, but 64 GB can’t hold 120B-class models. For about $300 more, the EVO-X2 can.
  • More than $5,400: beyond the Mac Studio, the next steps are multi-card workstations or 256 GB Macs, which are outside what most people need.

Before you buy anything

Check what your current computer can already do in our Can it run? calculator. Our own budget test machine, a $369 ACEMAGIC K1 mini PC, runs small models at 3 to 6 tokens per second (T0003, T0004): slow, but enough to learn what local AI can do before you spend more.

How we chose

We picked one machine per budget on memory size, memory bandwidth and the prices we recorded in September 2026, and estimated speeds with the method in our bandwidth study. These machines are researched, not tested by us; our own test results come from a Ryzen 9 7900X desktop with an RTX 4060 and the ACEMAGIC K1.

Questions people ask

What is the cheapest way to run AI models locally?

Start with the computer you have: our calculator shows what it can run. Most PCs with 16 GB of memory run small models on the processor alone, though slowly; our $369 ACEMAGIC K1 mini PC wrote 3 to 6 tokens per second (tests T0003 and T0004). The cheapest big step up is a 16 GB graphics card in a desktop, about $550 in September 2026.

How much should I spend to run 70B models?

You need about 46 GB of fast memory for a 70B model at 4-bit, so a 64 to 128 GB unified-memory machine, from about $3,200. Even then 70B dense models are slow: we estimate 4 to 8 tokens per second on the 64 and 128 GB machines here. Mixture-of-experts models such as gpt-oss 120B are larger but run about ten times faster on the same hardware.

Mac or PC for local AI?

At the same memory size, Macs with Pro and Max chips have faster memory, so they write faster; the Mac Studio M5 Max is the fastest option here. PCs cost less for the same memory, can be upgraded, and run NVIDIA's CUDA software if you add an NVIDIA card. Both run the popular apps, such as LM Studio and Ollama.

Is it cheaper to use cloud AI instead?

For occasional use, yes: a cloud subscription costs far less than a 128 GB computer. Local AI makes sense when you want privacy, work offline, use AI heavily, or want to experiment with open models, and a machine you already own costs nothing extra to try.

Sources

  1. AMD Radeon RX 9060 XT (16GB) specifications, checked Sep 26, 2026
  2. Apple Mac mini technical specifications, checked Sep 25, 2026
  3. Apple MacBook Air technical specifications, checked Sep 26, 2026
  4. Apple Mac Studio technical specifications, checked Sep 25, 2026
  5. AMD Ryzen AI Max+ 395 specifications, checked Sep 25, 2026

What changed

  • Sep 26, 2026: First published.