Hardware / NVIDIA Jetson AGX Orin 64GB

NVIDIA Jetson AGX Orin 64GB

64GB Unified memory / 205 GB/s / 60W / $899 MSRP / Released 2022
Find compatible models
Spec-sheet schematic of NVIDIA Jetson AGX Orin 64GB - AGX Orin 64 chip, 64GB unified memory, 205 GB/s
Overview

NVIDIA’s embedded module family, three generations on this site: Orin Nano Super (8GB, 68 GB/s, $249, 2024 - the “Super” refresh of the 2023 board), Orin NX 16GB (102 GB/s, $499, 2023), and AGX Orin 64GB (205 GB/s, $899, 2022). All run JetPack Linux on ARM with CUDA.

What it does well:

  • Power envelope: 10-60W total, passively or lightly cooled - the only CUDA platform here that runs on a battery or a solar rig.
  • CUDA on a budget: $249 buys a real NVIDIA stack for 1-8B models, with the full llama.cpp and MNN toolchain.

Where it falls short: 68 GB/s on the Nano Super caps single-stream decode near single-digit tok/s on 7B-class models; 8-64GB means only small dense models or heavily quantized MoE fit.

Run it locally: Jetson AI Lab containers, Ollama, or llama.cpp on JetPack. The modeldex results rank checkpoints by fit and estimated tok/s.

Chip specs
Chip
Nvidia AGX Orin 64
Memory
64GB unified
Bandwidth
205 GB/s
Available since
2022
Noise
Quiet

Top models (96 compatible)

Qwen3.8-Flash-Next EXL3_2BPW Runs tight
62.7GB / 24.0GB min
54
TOK/S
ⓘ estimated
Qwen3.6 35B A3B Q4_K_M Runs comfortably
20.0GB / 20.0GB min
66
TOK/S
ⓘ estimated
Ornith-1.5-35B-A3B Q4_K_M Runs comfortably
21.7GB / 22.0GB min
62
TOK/S
ⓘ estimated
Pokee-Isaac 28B Q4_K_M Runs comfortably
16.0GB / 18.0GB min
7
TOK/S
ⓘ estimated
Qwen3.8-27B UD-IQ1_S Runs comfortably
6.2GB / 8.0GB min
18
TOK/S
ⓘ estimated
See all 96 compatible models →