General model by Alibaba · 30.53B parameters · Released 2025-04-27
Everything this chip runs, machine by machine: MacBook Pro 14" M5 Pro, MacBook Pro 16" M5 Pro
Qwen3 30B A3B can run but will be tight at 84% of RAM. Consider a lower quantization or closing other apps.
— tok/s
Q4_K_M · 20.2 GB · Benchmark needed
Alternative: Q3_K_M (17.2 GB, ~107 tok/s)
| Quantization | Memory | Speed | Fits? |
|---|---|---|---|
| Q8_0 | 34.6 GB | ~49 tok/s | ✗ |
| Q6_K | 26.9 GB | ~64 tok/s | ✗ |
| Q5_K_M | 23.3 GB | ~75 tok/s | ✗ |
| Q4_K_M Recommended | 20.2 GB | ~88 tok/s | ✓ |
| Q3_K_M | 17.2 GB | ~107 tok/s | ✓ |
| Q2_K | 13.8 GB | ~138 tok/s | ✓ |
GGUF Sources
Affiliate disclosure: the shop icons in this table are Amazon affiliate links. As an Amazon Associate, ModelPiper earns from qualifying purchases. Every fit rating, quantization and tok/s figure below is computed from the hardware specs and is not influenced by that.
Closest matches by use case and parameter count.
ToolPiper downloads, manages, and runs models with one click. Apple Silicon optimized.
Get ToolPiper free