Zimmer
Interactive Hardware & Compatibility Calculator

Which open-source AI model fits your hardware?

Published August 2026Updated for Llama 3.3, Qwen 2.5 & DeepSeek MoEBy Omer Khan, Zimmer (Fihi Labs UG)

Top Recommended Models (13 compatible)

DeepSeek · 8B Best for You

DeepSeek R1 Distill Qwen 8B

DeepSeek's breakthrough reasoning specialist with chain-of-thought problem solving. Incredible STEM and logic on 16 GB.

Uses 6.4 GB of 16 GB~35–55 tok/s
Alibaba Qwen · 7B Best for You

Qwen 2.5 Coder 7B

Alibaba's state-of-the-art compact coding model. Exceptional syntax comprehension, function calling, and fast inference.

Uses 6.1 GB of 16 GB~35–55 tok/s
Alibaba Qwen · 3B Best for You

Qwen 2.5 3B Instruct

High-speed compact coder from Alibaba with strong logic, multilingual fluency, and lightweight RAM usage.

Uses 2.8 GB of 16 GB~35–55 tok/s

Hardware Compatibility Matrix

Explore pre-calculated memory sizing, quantization trade-offs, and inference throughput across devices.

Apple Silicon Macs share unified memory between the CPU and GPU with bandwidth up to 800 GB/s on Max and Ultra chips, making them premier hardware for high-parameter local models.

Unified RAMBest Model FitRecommended QuantContext LimitExpected Speed
8 GB MacGemma 2 2B, Meta Glimmer 3B, Qwen 2.5 3BQ4_K_M8k–16k35–55 tok/s
16 GB Mac (Baseline)DeepSeek R1 Distill 8B, Qwen 2.5 Coder 7B, Gemma 2 9BQ4_K_M / Q5_K_M16k (Default)30–45 tok/s
24 GB MacQwen 2.5 Coder 14B, Kimi k1.5 8B, GLM-4 9BQ4_K_M16k–32k22–35 tok/s
32 GB – 36 GB MacGemma 2 27B, DeepSeek MoE 30B, Qwen 2.5 Coder 14B Q5Q5_K_M / MoE Q432k–64k18–30 tok/s
48 GB – 64 GB MacDeepSeek R1 Distill 32B, Qwen 2.5 32B, Llama 3.3 70B Q4Q4_K_M / Q5_K_M32k–128k12–25 tok/s
96 GB – 128 GB+ MacLlama 3.3 70B Q8_0, DeepSeek V3, Qwen 2.5 72BQ5_K_M / Q8_0128k15–28 tok/s

Why 16k Context is Zimmer AI's Smart Default

While modern model architectures advertise 128k+ tokens, in local inference the key-value (KV) attention cache grows linearly with context length. Zimmer AI defaults to 16k context—giving plenty of room for multi-file coding and conversations while keeping memory consumption under 1.5 GB.

Frequently Asked Questions

Zero-Terminal Setup · Free Forever

Run the right model on the computer you own.

Zimmer AI automatically reads your Mac or Windows hardware specs, sorts every Hugging Face model into compatibility buckets, and lets you download and run them with one click.

Free forever for personal and commercial use · No subscriptions · No API keys needed