local intelligence

can it run

Can the AMD Ryzen AI Halo dev box run LLAMA4 Maverick 400B A17B?

NO

No. 128GB of memory can't hold the 307GB working set this model needs. A smaller quant or a bigger machine is required.

~0.63–0.89 tok/s

weights 245GB + KV cache 0.75GB at 4k context vs ~126GB usable → no. 245.8 GB vs 126 GB usable — does not fit

computed roofline: 256 GB/s × 0.60–0.85 efficiency window / 245.0GB (Q4_K_M) — bands, never points.

explore the full catalog84 machines indexed · live prices · what each one can run

Specs

model file245GB (GGUF Q4_K_M)
minimum memory307GB working set
this machine128GB unified LPDDR5X
paramsMaverick 400B A17B
licensellama4 (gated on HF)
price$3999

Run it

model files: hugging face ↗

easy run: ollama pull llama4:maverick

runs with: ollama · llama.cpp · koboldcpp · lm studio

Alternatives

ASUS ExpertCenter Pro ET900N G3 (GB300) — cheapest machine that runs it ($60000)

Buy

AMD ↗ · Amazon ↗

2026-09-17 · ← full hardware catalog · prices verified at source, may drift