Which open-source AI model fits your hardware?
Top Recommended Models (13 compatible)
DeepSeek R1 Distill Qwen 8B
DeepSeek's breakthrough reasoning specialist with chain-of-thought problem solving. Incredible STEM and logic on 16 GB.
Qwen 2.5 Coder 7B
Alibaba's state-of-the-art compact coding model. Exceptional syntax comprehension, function calling, and fast inference.
Qwen 2.5 3B Instruct
High-speed compact coder from Alibaba with strong logic, multilingual fluency, and lightweight RAM usage.
Hardware Compatibility Matrix
Explore pre-calculated memory sizing, quantization trade-offs, and inference throughput across devices.
Apple Silicon Macs share unified memory between the CPU and GPU with bandwidth up to 800 GB/s on Max and Ultra chips, making them premier hardware for high-parameter local models.
| Unified RAM | Best Model Fit | Recommended Quant | Context Limit | Expected Speed |
|---|---|---|---|---|
| 8 GB Mac | Gemma 2 2B, Meta Glimmer 3B, Qwen 2.5 3B | Q4_K_M | 8k–16k | 35–55 tok/s |
| 16 GB Mac (Baseline) | DeepSeek R1 Distill 8B, Qwen 2.5 Coder 7B, Gemma 2 9B | Q4_K_M / Q5_K_M | 16k (Default) | 30–45 tok/s |
| 24 GB Mac | Qwen 2.5 Coder 14B, Kimi k1.5 8B, GLM-4 9B | Q4_K_M | 16k–32k | 22–35 tok/s |
| 32 GB – 36 GB Mac | Gemma 2 27B, DeepSeek MoE 30B, Qwen 2.5 Coder 14B Q5 | Q5_K_M / MoE Q4 | 32k–64k | 18–30 tok/s |
| 48 GB – 64 GB Mac | DeepSeek R1 Distill 32B, Qwen 2.5 32B, Llama 3.3 70B Q4 | Q4_K_M / Q5_K_M | 32k–128k | 12–25 tok/s |
| 96 GB – 128 GB+ Mac | Llama 3.3 70B Q8_0, DeepSeek V3, Qwen 2.5 72B | Q5_K_M / Q8_0 | 128k | 15–28 tok/s |
Why 16k Context is Zimmer AI's Smart Default
While modern model architectures advertise 128k+ tokens, in local inference the key-value (KV) attention cache grows linearly with context length. Zimmer AI defaults to 16k context—giving plenty of room for multi-file coding and conversations while keeping memory consumption under 1.5 GB.
Frequently Asked Questions
Run the right model on the computer you own.
Zimmer AI automatically reads your Mac or Windows hardware specs, sorts every Hugging Face model into compatibility buckets, and lets you download and run them with one click.
Free forever for personal and commercial use · No subscriptions · No API keys needed