General model by farbodtavakkoli · 9.23B parameters · Released 2026-06-17
Everything this chip runs, machine by machine: MacBook Pro 14" M5 Pro, MacBook Pro 16" M5 Pro
OTel LLM 8B A1B IT fits comfortably in your 24 GB Mac using Q8_0 quantization, using 45% of your RAM.
— tok/s
Q8_0 · 10.8 GB · Benchmark needed
| Quantization | Memory | Speed | Fits? |
|---|---|---|---|
| Q8_0 Recommended | 10.8 GB | ~103 tok/s | ✓ |
| Q6_K | 8.5 GB | ~135 tok/s | ✓ |
| Q5_K_M | 7.4 GB | ~159 tok/s | ✓ |
| Q4_K_M | 6.5 GB | ~187 tok/s | ✓ |
| Q3_K_M | 5.5 GB | ~226 tok/s | ✓ |
| Q2_K | 4.5 GB | ~293 tok/s | ✓ |
Affiliate disclosure: the shop icons in this table are Amazon affiliate links. As an Amazon Associate, ModelPiper earns from qualifying purchases. Every fit rating, quantization and tok/s figure below is computed from the hardware specs and is not influenced by that.
Closest matches by use case and parameter count.
ToolPiper downloads, manages, and runs models with one click. Apple Silicon optimized.
Get ToolPiper free