NVIDIA Jetson Orin NX 16GB
16GB Unified memory
/
102 GB/s
/
25W
/
$499 MSRP
/
Released 2023
Overview
NVIDIA’s embedded module family, three generations on this site: Orin Nano Super (8GB, 68 GB/s, $249, 2024 - the “Super” refresh of the 2023 board), Orin NX 16GB (102 GB/s, $499, 2023), and AGX Orin 64GB (205 GB/s, $899, 2022). All run JetPack Linux on ARM with CUDA.
What it does well:
- Power envelope: 10-60W total, passively or lightly cooled - the only CUDA platform here that runs on a battery or a solar rig.
- CUDA on a budget: $249 buys a real NVIDIA stack for 1-8B models, with the full llama.cpp and MNN toolchain.
Where it falls short: 68 GB/s on the Nano Super caps single-stream decode near single-digit tok/s on 7B-class models; 8-64GB means only small dense models or heavily quantized MoE fit.
Run it locally: Jetson AI Lab containers, Ollama, or llama.cpp on JetPack. The modeldex results rank checkpoints by fit and estimated tok/s.
Chip specs
- Nvidia Orin NX 16
- 16GB unified
- 102 GB/s
- 2023
- Quiet
Top models (55 compatible)
Qwen3.8-27B
UD-IQ1_S
Runs comfortably
6.2GB / 8.0GB min
9
ⓘ estimated
Gemma 4 26B A4B
Q4_K_M
Runs tight
13.0GB / 16.0GB min
29
ⓘ estimated
Ornith-1.5-9B
Q4_K_M
Runs comfortably
5.8GB / 6.0GB min
10
ⓘ estimated
Qwen3 30B A3B
Q4_K_M
Runs tight
15.5GB / 15.5GB min
37
ⓘ estimated
Phi-4 14B
Q4_K_M
Runs comfortably
8.5GB / 8.5GB min
7
ⓘ estimated