Run AI on hardware you own - and follow the models worth running

Stop renting intelligence. Find open models that fit your rig, benchmarked honestly, with a source-cited tracker of who runs what - so no vendor or government order can switch off the model you depend on.

Latest open-weight models

Browse all models

Latest news & guides

All guides
Why run locally?

The local intelligence frontier

You don't need a cloud API to run the best open-weight models. A single high-memory machine or multi-GPU desktop runs Llama, Qwen, GLM, and more privately - no per-token bill, no rate limit, and no vendor switch-off.

Explore the full frontier →
Cloud frontier output price
GPT-5.5
$30 / 1M tokens - OpenAI
Claude Opus 4.8
$25 / 1M tokens - Anthropic
GPT-5.4
$15 / 1M tokens - OpenAI

Latest rigs

All rigs

Top models by quality

Browse all models

Find the right model for your hardware

Already know your rig? Pick it here and see exactly which models you can run locally, with honest speed estimates and cloud-pricing comparisons.

02 - Save your rig (free)

Sign in with GitHub to save your hardware. New here? We'll guide you through picking your rig - Mac, multi-GPU (up to 8x), or custom specs - then show you exactly which models you can run locally.

Sign in with GitHub - free

Who's running what

See all 31 β†’
Databricks runs DeepSeek V4.1 Flash confirmed DeepSeek

Yuchen Jin (Databricks): DeepSeek V4.1 Flash is his pick for easy and moderate tasks - fast and cheap. Databricks reports customers cutting spend dramatically by shifting 20% of internal coding traffic from Claude/GPT to open models.

Databricks runs GLM-5.3 confirmed Z.ai

Yuchen Jin (Databricks): GLM-5.3 is cheap and fast, suited to high volumes of non-complex coding tasks. Part of Databricks engineers' shift to open models as daily drivers.

Databricks runs Kimi K3 confirmed Moonshot AI

Databricks AI systems CTO Yuchen Jin: Kimi K3 is the best open-weight coding model in their experience. Databricks serves it at 239 tok/s, the largest open model they have hosted, and ranks #1 for K3 inference speed on Artificial Analysis.

Databricks runs DeepSeek V4.1 Flash confirmed DeepSeek

Yuchen Jin (Databricks): DeepSeek V4.1 Flash is his pick for easy and moderate tasks - fast and cheap. Databricks reports customers cutting spend dramatically by shifting 20% of internal coding traffic from Claude/GPT to open models.

Databricks runs GLM-5.3 confirmed Z.ai

Yuchen Jin (Databricks): GLM-5.3 is cheap and fast, suited to high volumes of non-complex coding tasks. Part of Databricks engineers' shift to open models as daily drivers.

Statuses: confirmed official source reported credible third-party, not officially confirmed testing evaluating, not deployed.

Curated and source-cited, not a scraper. Built a rig worth sharing? Browse shared builds β†’.

81
Models tracked
37
Tools tracked
70
Hardware configs
31
Adopters tracked