local intelligence

can it run

Can the Jetson Orin Nano Super dev kit run llava 7B?

TIGHT

Yes, but close to the edge. 8GB against a 6GB working set: it runs, with little room for long context windows or other processes.

~3.6–9.4 tok/s

weights 4.7GB + KV cache 2GB at 4k context vs ~6GB usable → offload-partial. 6.7 GB vs 6 GB usable — partial CPU offload, expect large speed loss

computed roofline: 68 GB/s × 0.25–0.65 efficiency window / 4.7GB (Q4_K_M) — bands, never points.

explore the full catalog84 machines indexed · live prices · what each one can run

Specs

model file4.7GB (GGUF Q4_K_M)
minimum memory6GB working set
this machine8GB unified
params7B
licenseapache-2.0
price$499

Run it

model files: hugging face ↗

easy run: ollama pull llava:7b

runs with: ollama · llama.cpp · koboldcpp · lm studio

Alternatives

Tesla P100 (16GB, used) — cheapest machine that runs it ($135)

every model the Jetson Orin Nano Super dev kit can run — full list

Buy

NVIDIA (direct) ↗ · Amazon ↗

2026-09-17 · ← all models on the Jetson Orin Nano Super dev kit · ← full hardware catalog · prices verified at source, may drift