Whole systems built (or repurposed) for local AI: turnkey AI supercomputers, NPU-powered mini-PCs, and the Mac Studio. All model picks lead with the newest releases (Aug 2026).
“An AI supercomputer on your desk” — the easiest way to run 100B+ models locally.
| Spec | Value |
|---|---|
| Chip | GB10 Grace Blackwell superchip (20-core Arm CPU: 10× Cortex-X925 + 10× Cortex-A725) |
| AI compute | 1 PFLOPS FP4 (~1,000 AI TOPS), 5th-gen Tensor Cores |
| Memory | 128 GB unified LPDDR5x (256-bit, 273 GB/s) — shared CPU/GPU |
| Storage | 4 TB NVMe SSD |
| Networking | ConnectX-7 Smart NIC (200 GbE) |
| Power | ~200 W (fanless, desktop form factor) |
| Price | $4,699 (was $3,999 at announcement, Mar 2025; rose late 2025) |
| Software | DGX OS (Linux), preloaded NVIDIA AI stack (NIM, CUDA, PyTorch) |
Best for: Running up to 200B-parameter models out of the box, prototyping/agent workloads, and teams that want CUDA without buying a server. Mistral Medium 3 and gpt-oss-120B fit fully in memory.
Recommended AI models (newest first):
Workstation-grade: a desktop that competes with a rack of GPUs.
| Spec | Value |
|---|---|
| Chip | GB300 Grace Blackwell Ultra superchip (72-core Arm Neoverse V2 CPU + Blackwell Ultra GPU) |
| AI compute | 20 PFLOPS FP4 |
| Memory | 748 GB coherent total: 496 GB LPDDR5x (396 GB/s) + 252 GB HBM3e (8 TB/s) |
| Networking | ConnectX-8 (800 GbE) |
| Price | ~$70,000–110,000 (NVIDIA list ~$69.9K; OEM/reseller pricing higher) |
Best for: Teams/researchers running today’s frontier open models (GLM-5.2, Mistral Large 3, MiniMax M2.5, Kimi K3, DeepSeek V4 Flash) fully locally with data privacy, or fine-tuning at workstation scale.
Recommended AI models (newest first):
The first desktop chips officially supporting Microsoft Copilot+ (announced MWC 2026).
| Spec | Value |
|---|---|
| CPU | Zen 5 (Medusa Point), up to 8 cores |
| NPU | XDNA2, up to 60 TOPS |
| iGPU | RDNA (up to 16 CU) |
| Memory | DDR5 (standard desktop DIMMs), 32–128 GB |
| Price (Aug 2026) | OEM desktops ~$500–1,700 (entry pricing from ~$499; loaded 16 GB/512 GB systems ~$1,660) — sold as complete PCs, no retail CPUs |
Best for: Everyday AI in an office/home tower — on-device Copilot+, local transcription, background agents, small models. Not for big LLMs (memory bandwidth is DDR5-class).
Recommended AI models (newest first):
Cheap, quiet, always-on local AI boxes.
| Spec | Value |
|---|---|
| Processors | Intel Core Ultra 200V (Lunar Lake) / 300V (Panther Lake), AMD Ryzen AI 300/400 |
| NPU | 48–60 TOPS |
| Memory | 16–96 GB LPDDR5x/DDR5 |
| Form factor | 0.5–2 L mini-PC |
| Price (Aug 2026) | ~$400–1,000 mainstream (16–32 GB); $1,000–3,000 loaded (64–128 GB Strix Halo-class); flagship AI boxes: Ryzen AI Halo $3,999, DGX Spark $4,699 |
Best for: Home servers, privacy-focused assistants, transcription/voice agents, and light RAG. Many folks run Ollama + Open WebUI on these 24/7.
Recommended AI models (newest first):
See apple-silicon.md for the chips; the Studio is the “big memory” chassis.
| Spec | M4 Max Studio | M5 Ultra Studio (reported) |
|---|---|---|
| CPU / GPU | 16-core / 40-core | 36-core / 84-core (reports) |
| Neural Engine | 16-core | 32-core (reports) |
| Unified memory | up to 128 GB | up to 512 GB (768 GB per latest reports) |
| Memory bandwidth | 546 GB/s | ~1 TB/s class (reports) |
| Storage | up to 8 TB | up to 16 TB |
| Price | from ~$2,000 (base) | from ~$4,000+ (reported) |
Best for: The 512 GB M5 Ultra Studio is the only mainstream machine that can hold the newest frontier open models (Mistral Large 3, GLM-5.2, MiniMax M2.5) entirely in memory — the local-inference extreme.
Recommended AI models (newest first):