Hardware / NVIDIA Jetson Orin NX 16GB

NVIDIA Jetson Orin NX 16GB

16GB Unified memory / 102 GB/s / 25W / $499 MSRP / Released 2023
Find compatible models
Spec-sheet schematic of NVIDIA Jetson Orin NX 16GB - Orin NX 16 chip, 16GB unified memory, 102 GB/s
Overview

NVIDIA’s embedded module family, three generations on this site: Orin Nano Super (8GB, 68 GB/s, $249, 2024 - the “Super” refresh of the 2023 board), Orin NX 16GB (102 GB/s, $499, 2023), and AGX Orin 64GB (205 GB/s, $899, 2022). All run JetPack Linux on ARM with CUDA.

What it does well:

  • Power envelope: 10-60W total, passively or lightly cooled - the only CUDA platform here that runs on a battery or a solar rig.
  • CUDA on a budget: $249 buys a real NVIDIA stack for 1-8B models, with the full llama.cpp and MNN toolchain.

Where it falls short: 68 GB/s on the Nano Super caps single-stream decode near single-digit tok/s on 7B-class models; 8-64GB means only small dense models or heavily quantized MoE fit.

Run it locally: Jetson AI Lab containers, Ollama, or llama.cpp on JetPack. The modeldex results rank checkpoints by fit and estimated tok/s.

Chip specs
Chip
Nvidia Orin NX 16
Memory
16GB unified
Bandwidth
102 GB/s
Available since
2023
Noise
Quiet

Top models (55 compatible)

Qwen3.8-27B UD-IQ1_S Runs comfortably
6.2GB / 8.0GB min
9
TOK/S
ⓘ estimated
Gemma 4 26B A4B Q4_K_M Runs tight
13.0GB / 16.0GB min
29
TOK/S
ⓘ estimated
Ornith-1.5-9B Q4_K_M Runs comfortably
5.8GB / 6.0GB min
10
TOK/S
ⓘ estimated
Qwen3 30B A3B Q4_K_M Runs tight
15.5GB / 15.5GB min
37
TOK/S
ⓘ estimated
Phi-4 14B Q4_K_M Runs comfortably
8.5GB / 8.5GB min
7
TOK/S
ⓘ estimated
See all 55 compatible models →