---
title: "Mac Studio M5 Max: What AI Models Can It Run? | ModelPiper"
description: "Local AI on the M5 Max Mac Studio: 460 to 614 GB/s memory bandwidth and 36 to 128 GB of unified memory. See which LLMs fit, at which quantization, and how fast they run."
canonical: "https://modelpiper.com/mac-studio/m5-max"
---

# Mac Studio M5 Max: What AI Models Can It Run? | ModelPiper

> Local AI on the M5 Max Mac Studio: 460 to 614 GB/s memory bandwidth and 36 to 128 GB of unified memory. See which LLMs fit, at which quantization, and how fast they run.

[← All Mac Studio models](/mac-studio)

# Mac Studio M5 Max

The M5 Max Mac Studio runs local models at 460 to 614 GB/s of memory bandwidth with 36 to 128 GB of unified memory. On Apple Silicon that memory is shared with the GPU, so the whole pool is available for weights: at 128 GB you can hold roughly a 172B dense model at Q4. The Max is where the bus gets wide enough that model size, not bandwidth, becomes the thing you plan around.

By [Ben Racicot](/about#founder), Founder & Lead Engineer— Updated 2026-08-21

## Specifications

ChipApple M5 Max

CPU cores18

GPU cores32 or 40

Unified memory36, 48, 64, or 128 GB

Memory bandwidth460 to 614 GB/s

Neural Engine38 TOPS

Released2026

AvailabilitySold new by Apple

Memory bandwidth is faster than 75% of the Apple Silicon chips shipped in a Mac, against a 1228 GB/s peak.

Its memory ceiling is above 65% of them, against a 512 GB peak.

## Pick your configuration

Every option Apple sells with this chip. The model list below recomputes against the one you pick.

GPU cores

32-core 460 GB/s 40-core 614 GB/s

Unified memory

36 GB

Apple couples memory to the core count on this chip, so the options change with the bin above.

## What each memory option runs

Unified memory is the ceiling and it is soldered, so this is the decision you cannot revisit.

[36 GB](#ram-36gb)[48 GB](#ram-48gb)[64 GB](#ram-64gb)[128 GB](#ram-128gb)

### 36 GB unified memory

31 GB usable for weights · 460 GB/s

6,120 of 6,563 models fit, and 5,827 of them run with headroom rather than as a squeeze.

Largest model at Q4

[Chinese Mixtral 8x7B](/fit/hit-scir-chinese-mixtral-8x7b) · 46.91B

Best all-round pick

[Kimi K3 DSpark](/fit/radixark-kimi-k3-dspark) · Q8\_0 · ~107 tok/s

### 48 GB unified memory

42 GB usable for weights · 614 GB/s

6,273 of 6,563 models fit, and 6,033 of them run with headroom rather than as a squeeze.

Largest model at Q4

[Yi 34Bx2 MoE 60B DPO](/fit/cloudyu-yi-34bx2-moe-60b-dpo) · 60.81B

Best all-round pick

[Kimi K3 DSpark](/fit/radixark-kimi-k3-dspark) · Q8\_0 · ~143 tok/s

### 64 GB unified memory

56 GB usable for weights · 614 GB/s

6,357 of 6,563 models fit, and 6,086 of them run with headroom rather than as a squeeze.

Largest model at Q4

[Hy3 REAP 48e](/fit/sapidlabs-hy3-reap-48e) · 84.02B

Best all-round pick

[Kimi K3 DSpark](/fit/radixark-kimi-k3-dspark) · Q8\_0 · ~143 tok/s

### 128 GB unified memory

112 GB usable for weights · 614 GB/s

6,420 of 6,563 models fit, and 6,334 of them run with headroom rather than as a squeeze.

Largest model at Q4

[Kimi K2 Thinking converted](/fit/lrudl-workshop-kimi-k2-thinking-converted) · 170.27B

Best all-round pick

[Ornith 1.0 35B](/fit/deepreinforce-ai-ornith-1-0-35b) · Q8\_0 · ~118 tok/s

## What a 36 GB M5 Max Mac Studio can run

Every model in the database against this exact configuration, at 460 GB/s. Ratings and speeds are the same numbers the model pages show.

All Categories

Show All

All Sizes

 Best Fit

Showing 6563 of 6563 models

[Kimi K3 DSpark](/fit/radixark-kimi-k3-dspark?chip=m5-max&ram=36)

General · radixark · 2026-07-27

Q8\_0Excellent

3.0 GB8% of RAM~107 tok/sEstimated2.25B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[MiniCPM5 1B](/fit/openbmb-minicpm5-1b?chip=m5-max&ram=36)

General · openbmb · 2026-05-21

Q8\_0Excellent

1.7 GB5% of RAM~223 tok/sEstimated1.08B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[OTel LLM 8B A1B IT](/fit/farbodtavakkoli-otel-llm-8b-a1b-it?chip=m5-max&ram=36)

General · farbodtavakkoli · 2026-06-17

Q8\_0Excellent

10.8 GB30% of RAMBenchmark needed9.23B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2.5 8B A1B](/fit/liquidai-lfm2-5-8b-a1b?chip=m5-max&ram=36)

General · Liquid AI · 2026-05-28

Q8\_0Excellent

9.9 GB28% of RAMBenchmark needed8.47B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[VibeThinker 3B](/fit/weiboai-vibethinker-3b?chip=m5-max&ram=36)

General · weiboai · 2026-06-12

Q8\_0Excellent

3.9 GB11% of RAM~78 tok/sEstimated3.09B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2.5 2.6B](/fit/liquidai-lfm2-5-2-6b?chip=m5-max&ram=36)

General · Liquid AI · 2026-07-28

Q8\_0Excellent

3.5 GB10% of RAM~89 tok/sEstimated2.7B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2.5 230M](/fit/liquidai-lfm2-5-230m?chip=m5-max&ram=36)

General · Liquid AI · 2026-06-24

Q8\_0Excellent

0.8 GB2% of RAM~1,048 tok/sEstimated0.23B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Nanbeige4.2 3B](/fit/nanbeige-nanbeige4-2-3b?chip=m5-max&ram=36)

General · nanbeige · 2026-07-21

Q8\_0Excellent

5.2 GB14% of RAM~58 tok/sEstimated4.17B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3.8\_4B\_Distilled\_GGUF](/fit/ma7ee7-qwen3-8_4b_distilled_gguf?chip=m5-max&ram=36)

General · ma7ee7 · 2026-07-30

Q8\_0Excellent

4.0 GB11% of RAM~77 tok/sEstimated3.13B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[JOSIE 2 2B OSS](/fit/goekdeniz-guelmez-josie-2-2b-oss?chip=m5-max&ram=36)

General · goekdeniz-guelmez · 2026-07-31

Q8\_0Excellent

3.0 GB8% of RAM~106 tok/sEstimated2.27B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[minicpm5 1b hermes toolhv1](/fit/ktruestory-minicpm5-1b-hermes-toolhv1?chip=m5-max&ram=36)

General · ktruestory · 2026-05-28

Q8\_0Excellent

1.7 GB5% of RAM~223 tok/sEstimated1.08B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[FrankenCPM 4x1B A2B](/fit/petrouil-frankencpm-4x1b-a2b?chip=m5-max&ram=36)

General · petrouil · 2026-07-23

Q8\_0Excellent

3.4 GB9% of RAMBenchmark needed2.61B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[granite 4.1 3b](/fit/ibm-granite-granite-4-1-3b?chip=m5-max&ram=36)

General · ibm-granite · 2026-04-06

Q8\_0Excellent

4.3 GB12% of RAM~71 tok/sEstimated3.4B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Agents A1 4B](/fit/internscience-agents-a1-4b?chip=m5-max&ram=36)

General · internscience · 2026-07-13

Q8\_0Excellent

5.6 GB15% of RAM~53 tok/sEstimated4.54B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[granite vision 4.1 4b](/fit/ibm-granite-granite-vision-4-1-4b?chip=m5-max&ram=36)

Multimodal · ibm-granite · 2026-04-16

Q8\_0Excellent

5.0 GB14% of RAM~60 tok/sEstimated4B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2.5 350M](/fit/liquidai-lfm2-5-350m?chip=m5-max&ram=36)

General · Liquid AI · 2026-03-31

Q8\_0Excellent

0.9 GB2% of RAM~688 tok/sEstimated0.35B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3.5 0.8B](/fit/qwen-qwen3-5-0-8b?chip=m5-max&ram=36)

Multimodal · Alibaba · 2026-02-28

Q8\_0Excellent

1.5 GB4% of RAM~277 tok/sEstimated0.87B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3.5 2B](/fit/qwen-qwen3-5-2b?chip=m5-max&ram=36)

Multimodal · Alibaba · 2026-02-28

Q8\_0Excellent

3.0 GB8% of RAM~106 tok/sEstimated2.27B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3.5 0.8B Base](/fit/qwen-qwen3-5-0-8b-base?chip=m5-max&ram=36)

Multimodal · Alibaba · 2026-02-28

Q8\_0Excellent

1.5 GB4% of RAM~277 tok/sEstimated0.87B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3.5 2B Base](/fit/qwen-qwen3-5-2b-base?chip=m5-max&ram=36)

Multimodal · Alibaba · 2026-02-28

Q8\_0Excellent

3.0 GB8% of RAM~106 tok/sEstimated2.27B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Nemotron Labs Diffusion 3B](/fit/nvidia-nemotron-labs-diffusion-3b?chip=m5-max&ram=36)

General · nvidia · 2026-03-02

Q8\_0Excellent

4.8 GB13% of RAM~63 tok/sEstimated3.83B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[pii guard turkish 270m](/fit/cagrigungor-pii-guard-turkish-270m?chip=m5-max&ram=36)

General · cagrigungor · 2026-08-06

Q8\_0Excellent

0.8 GB2% of RAM~892 tok/sEstimated0.27B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3.5 4B](/fit/qwen-qwen3-5-4b?chip=m5-max&ram=36)

Multimodal · Alibaba · 2026-02-27

Q8\_0Excellent

5.7 GB16% of RAM~52 tok/sEstimated4.66B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[gemma 4 E2B it](/fit/google-gemma-4-e2b-it?chip=m5-max&ram=36)

Multimodal · Google · 2026-03-02

Q8\_0Excellent

6.2 GB17% of RAM~47 tok/sEstimated5.12B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Ornith 1.0 9B](/fit/deepreinforce-ai-ornith-1-0-9b?chip=m5-max&ram=36)

General · deepreinforce-ai · 2026-06-21

Q8\_0Excellent

9.7 GB27% of RAM~29 tok/sEstimated8.21B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3.5 4B Base](/fit/qwen-qwen3-5-4b-base?chip=m5-max&ram=36)

Multimodal · Alibaba · 2026-02-27

Q8\_0Excellent

5.7 GB16% of RAM~52 tok/sEstimated4.66B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2.5 VL 1.6B](/fit/liquidai-lfm2-5-vl-1-6b?chip=m5-max&ram=36)

Multimodal · Liquid AI · 2026-01-05

Q8\_0Excellent

2.3 GB6% of RAM~151 tok/sEstimated1.6B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[dspv2\_guide](/fit/lukebailey181pub-dspv2_guide?chip=m5-max&ram=36)

General · lukebailey181pub · 2026-04-21

Q8\_0Excellent

8.2 GB23% of RAM~35 tok/sEstimated6.91B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2.5 1.2B Thinking](/fit/liquidai-lfm2-5-1-2b-thinking?chip=m5-max&ram=36)

General · Liquid AI · 2026-01-20

Q8\_0Excellent

1.8 GB5% of RAM~206 tok/sEstimated1.17B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[AI21 Jamba2 3B](/fit/ai21labs-ai21-jamba2-3b?chip=m5-max&ram=36)

General · ai21labs · 2026-01-06

Q8\_0Excellent

3.9 GB11% of RAMBenchmark needed3.03B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[OneReason 0.8B pretrain competition](/fit/openonerec-onereason-0-8b-pretrain-competition?chip=m5-max&ram=36)

Reasoning · openonerec · 2026-06-09

Q8\_0Excellent

1.4 GB4% of RAM~301 tok/sEstimated0.8B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2.5 1.2B Base](/fit/liquidai-lfm2-5-1-2b-base?chip=m5-max&ram=36)

General · Liquid AI · 2026-01-05

Q8\_0Excellent

1.8 GB5% of RAM~206 tok/sEstimated1.17B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2.5 1.2B JP](/fit/liquidai-lfm2-5-1-2b-jp?chip=m5-max&ram=36)

General · Liquid AI · 2026-01-04

Q8\_0Excellent

1.8 GB5% of RAM~206 tok/sEstimated1.17B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 2.6B Exp](/fit/liquidai-lfm2-2-6b-exp?chip=m5-max&ram=36)

General · Liquid AI · 2025-12-25

Q8\_0Excellent

3.4 GB9% of RAM~94 tok/sEstimated2.57B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 2.6B Transcript](/fit/liquidai-lfm2-2-6b-transcript?chip=m5-max&ram=36)

General · Liquid AI · 2026-01-05

Q8\_0Excellent

3.4 GB9% of RAM~94 tok/sEstimated2.57B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Gemma 3 4B VL it Gemini Pro Heretic Uncensored Thinking](/fit/davidau-gemma-3-4b-vl-it-gemini-pro-heretic-uncensored-thinking?chip=m5-max&ram=36)

Multimodal · davidau · 2026-02-02

Q8\_0Excellent

5.3 GB15% of RAM~56 tok/sEstimated4.3B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[gemma 4 E4B it](/fit/google-gemma-4-e4b-it?chip=m5-max&ram=36)

Multimodal · Google · 2026-03-02

Q8\_0Excellent

9.4 GB26% of RAM~30 tok/sEstimated8B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2.5 1.2B Instruct](/fit/liquidai-lfm2-5-1-2b-instruct?chip=m5-max&ram=36)

Chat · Liquid AI · 2026-01-06

Q8\_0Excellent

1.8 GB5% of RAM~206 tok/sEstimated1.17B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LLaDA2.1 mini](/fit/inclusionai-llada2-1-mini?chip=m5-max&ram=36)

General · inclusionai · 2026-02-09

Q8\_0Excellent

18.6 GB52% of RAMBenchmark needed16.26B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Dolphin3 Cyber 8B](/fit/ravichandranj-dolphin3-cyber-8b-gguf?chip=m5-max&ram=36)

General · ravichandranj · 2026-02-13

Q8\_0Excellent

7.7 GB21% of RAM~37 tok/sEstimated6.43B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 ColBERT 350M](/fit/liquidai-lfm2-colbert-350m?chip=m5-max&ram=36)

General · Liquid AI · 2025-10-28

Q8\_0Excellent

0.9 GB2% of RAM~688 tok/sEstimated0.35B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 8B A1B](/fit/liquidai-lfm2-8b-a1b?chip=m5-max&ram=36)

General · Liquid AI · 2025-10-07

Q8\_0Excellent

9.8 GB27% of RAMBenchmark needed8.34B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[gemma 4 E4B it OBLITERATED](/fit/obliteratus-gemma-4-e4b-it-obliterated?chip=m5-max&ram=36)

General · obliteratus · 2026-04-15

Q8\_0Excellent

9.4 GB26% of RAM~30 tok/sEstimated8B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 VL 3B](/fit/liquidai-lfm2-vl-3b?chip=m5-max&ram=36)

Multimodal · Liquid AI · 2025-10-22

Q8\_0Excellent

3.8 GB11% of RAM~80 tok/sEstimated3B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3.5 9B Claude 4.6 Opus Reasoning Distilled v2](/fit/jackrong-qwen3-5-9b-claude-4-6-opus-reasoning-distilled-v2?chip=m5-max&ram=36)

Reasoning · jackrong · 2026-03-16

Q8\_0Excellent

11.3 GB31% of RAM~25 tok/sEstimated9.65B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Lumma 0.6B Extract](/fit/frontiersmind-lumma-0-6b-extract?chip=m5-max&ram=36)

General · frontiersmind · 2026-08-03

Q8\_0Excellent

1.2 GB3% of RAM~371 tok/sEstimated0.65B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[typhoon2.5 qwen3 4b](/fit/typhoon-ai-typhoon2-5-qwen3-4b?chip=m5-max&ram=36)

General · typhoon-ai · 2025-09-23

Q8\_0Excellent

5.0 GB14% of RAM~60 tok/sEstimated4.02B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 4B Thinking 2507](/fit/qwen-qwen3-4b-thinking-2507?chip=m5-max&ram=36)

General · Alibaba · 2025-08-05

Q8\_0Excellent

5.0 GB14% of RAM~60 tok/sEstimated4.02B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Ilama 3.2 1B](/fit/hmellor-ilama-3-2-1b?chip=m5-max&ram=36)

General · hmellor · 2025-07-22

Q8\_0Excellent

1.9 GB5% of RAM~194 tok/sEstimated1.24B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[granite 4.0 h tiny](/fit/ibm-granite-granite-4-0-h-tiny?chip=m5-max&ram=36)

General · ibm-granite · 2025-09-16

Q8\_0Excellent

8.2 GB23% of RAMBenchmark needed6.94B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Nemotron Labs Diffusion 8B](/fit/nvidia-nemotron-labs-diffusion-8b?chip=m5-max&ram=36)

General · nvidia · 2026-03-18

Q8\_0Excellent

10.0 GB28% of RAM~28 tok/sEstimated8.49B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[granite 4.0 micro](/fit/ibm-granite-granite-4-0-micro?chip=m5-max&ram=36)

General · ibm-granite · 2025-09-16

Q8\_0Excellent

4.3 GB12% of RAM~71 tok/sEstimated3.4B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 VL 450M](/fit/liquidai-lfm2-vl-450m?chip=m5-max&ram=36)

Multimodal · Liquid AI · 2025-08-12

Q8\_0Excellent

1.0 GB3% of RAM~535 tok/sEstimated0.45B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[granite guardian 4.1 8b](/fit/ibm-granite-granite-guardian-4-1-8b?chip=m5-max&ram=36)

General · ibm-granite · 2026-04-16

Q8\_0Excellent

9.8 GB27% of RAM~29 tok/sEstimated8.38B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[OneRec 1.7B](/fit/openonerec-onerec-1-7b?chip=m5-max&ram=36)

General · openonerec · 2025-12-30

Q8\_0Excellent

2.9 GB8% of RAM~113 tok/sEstimated2.13B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[granite 4.0 h micro](/fit/ibm-granite-granite-4-0-h-micro?chip=m5-max&ram=36)

General · ibm-granite · 2025-09-16

Q8\_0Excellent

4.1 GB11% of RAM~76 tok/sEstimated3.19B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Ouro 1.4B](/fit/bytedance-ouro-1-4b?chip=m5-max&ram=36)

General · bytedance · 2025-10-28

Q8\_0Excellent

2.1 GB6% of RAM~168 tok/sEstimated1.43B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 4B DFlash b16](/fit/z-lab-qwen3-4b-dflash-b16?chip=m5-max&ram=36)

General · z-lab · 2026-01-04

Q8\_0Excellent

5.0 GB14% of RAM~60 tok/sEstimated4B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 VL 1.6B](/fit/liquidai-lfm2-vl-1-6b?chip=m5-max&ram=36)

Multimodal · Liquid AI · 2025-08-12

Q8\_0Excellent

2.3 GB6% of RAM~153 tok/sEstimated1.58B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 2.6B](/fit/liquidai-lfm2-2-6b?chip=m5-max&ram=36)

General · Liquid AI · 2025-09-22

Q8\_0Excellent

3.4 GB9% of RAM~94 tok/sEstimated2.57B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 350M PII Extract JP](/fit/liquidai-lfm2-350m-pii-extract-jp?chip=m5-max&ram=36)

General · Liquid AI · 2025-09-30

Q8\_0Excellent

0.9 GB2% of RAM~688 tok/sEstimated0.35B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 1.2B Extract](/fit/liquidai-lfm2-1-2b-extract?chip=m5-max&ram=36)

General · Liquid AI · 2025-08-22

Q8\_0Excellent

1.8 GB5% of RAM~206 tok/sEstimated1.17B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 350M ENJP MT](/fit/liquidai-lfm2-350m-enjp-mt?chip=m5-max&ram=36)

General · Liquid AI · 2025-09-03

Q8\_0Excellent

0.9 GB2% of RAM~688 tok/sEstimated0.35B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 1.2B RAG](/fit/liquidai-lfm2-1-2b-rag?chip=m5-max&ram=36)

General · Liquid AI · 2025-09-03

Q8\_0Excellent

1.8 GB5% of RAM~206 tok/sEstimated1.17B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 350M Math](/fit/liquidai-lfm2-350m-math?chip=m5-max&ram=36)

General · Liquid AI · 2025-08-25

Q8\_0Excellent

0.9 GB2% of RAM~688 tok/sEstimated0.35B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 350M Extract](/fit/liquidai-lfm2-350m-extract?chip=m5-max&ram=36)

General · Liquid AI · 2025-09-03

Q8\_0Excellent

0.9 GB2% of RAM~688 tok/sEstimated0.35B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 1.2B Tool](/fit/liquidai-lfm2-1-2b-tool?chip=m5-max&ram=36)

General · Liquid AI · 2025-09-03

Q8\_0Excellent

1.8 GB5% of RAM~206 tok/sEstimated1.17B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[VAETKI 7B A1B](/fit/nc-ai-consortium-vaetki-7b-a1b?chip=m5-max&ram=36)

General · NCAI · 2025-12-29

Q8\_0Excellent

8.6 GB24% of RAMBenchmark needed7.25B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3.5 9B](/fit/qwen-qwen3-5-9b?chip=m5-max&ram=36)

Multimodal · Alibaba · 2026-02-27

Q8\_0Excellent

11.3 GB31% of RAM~25 tok/sEstimated9.65B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[granite 4.1 8b](/fit/ibm-granite-granite-4-1-8b?chip=m5-max&ram=36)

General · ibm-granite · 2026-04-06

Q8\_0Excellent

10.3 GB29% of RAM~27 tok/sEstimated8.79B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3.5 9B Base](/fit/qwen-qwen3-5-9b-base?chip=m5-max&ram=36)

Multimodal · Alibaba · 2026-02-26

Q8\_0Excellent

11.3 GB31% of RAM~25 tok/sEstimated9.65B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LLaDA2.0 mini](/fit/inclusionai-llada2-0-mini?chip=m5-max&ram=36)

General · inclusionai · 2025-11-25

Q8\_0Excellent

18.6 GB52% of RAMBenchmark needed16.26B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[granite 4.0 tiny preview](/fit/ibm-granite-granite-4-0-tiny-preview?chip=m5-max&ram=36)

General · ibm-granite · 2025-04-30

Q8\_0Excellent

7.9 GB22% of RAMBenchmark needed6.67B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 1.2B](/fit/liquidai-lfm2-1-2b?chip=m5-max&ram=36)

General · Liquid AI · 2025-07-10

Q8\_0Excellent

1.8 GB5% of RAM~206 tok/sEstimated1.17B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Phi 4 mini reasoning](/fit/microsoft-phi-4-mini-reasoning?chip=m5-max&ram=36)

Reasoning · Microsoft · 2025-04-29

Q8\_0Excellent

4.8 GB13% of RAM~63 tok/sEstimated3.84B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Olmo 3 7B Think](/fit/allenai-olmo-3-7b-think?chip=m5-max&ram=36)

General · allenai · 2025-11-18

Q8\_0Excellent

8.6 GB24% of RAM~33 tok/sEstimated7.3B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[PARD Llama 3.2 1B](/fit/amd-pard-llama-3-2-1b?chip=m5-max&ram=36)

General · amd · 2025-05-17

Q8\_0Excellent

2.2 GB6% of RAM~161 tok/sEstimated1.5B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[newsvibe categories multilingual llama 1b](/fit/stefanruseti-newsvibe-categories-multilingual-llama-1b?chip=m5-max&ram=36)

General · stefanruseti · 2025-06-04

Q8\_0Excellent

1.9 GB5% of RAM~194 tok/sEstimated1.24B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwythos 9B Claude Mythos 5 1M](/fit/empero-ai-qwythos-9b-claude-mythos-5-1m?chip=m5-max&ram=36)

General · empero-ai · 2026-06-19

Q8\_0Excellent

11.0 GB31% of RAM~26 tok/sEstimated9.41B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Olmo Hybrid 7B](/fit/allenai-olmo-hybrid-7b?chip=m5-max&ram=36)

General · allenai · 2026-01-28

Q8\_0Excellent

8.8 GB24% of RAM~32 tok/sEstimated7.43B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Llama 3.1 8B Instruct pearl](/fit/pearl-ai-llama-3-1-8b-instruct-pearl?chip=m5-max&ram=36)

Chat · pearl-ai · 2026-02-26

Q8\_0Excellent

9.5 GB26% of RAM~30 tok/sEstimated8.03B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[granite 4.0 350m base](/fit/ibm-granite-granite-4-0-350m-base?chip=m5-max&ram=36)

General · ibm-granite · 2025-10-07

Q8\_0Excellent

0.9 GB2% of RAM~688 tok/sEstimated0.35B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[EXAONE 4.0 1.2B](/fit/lgai-exaone-exaone-4-0-1-2b?chip=m5-max&ram=36)

General · lgai-exaone · 2025-07-11

Q8\_0Excellent

1.9 GB5% of RAM~188 tok/sEstimated1.28B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 350M](/fit/liquidai-lfm2-350m?chip=m5-max&ram=36)

General · Liquid AI · 2025-07-10

Q8\_0Excellent

0.9 GB2% of RAM~688 tok/sEstimated0.35B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Jan nano 128k](/fit/menlo-jan-nano-128k?chip=m5-max&ram=36)

General · menlo · 2025-06-25

Q8\_0Excellent

5.0 GB14% of RAM~60 tok/sEstimated4.02B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LFM2 700M](/fit/liquidai-lfm2-700m?chip=m5-max&ram=36)

General · Liquid AI · 2025-07-10

Q8\_0Excellent

1.3 GB4% of RAM~326 tok/sEstimated0.74B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[JOSIE 2 9B OSS](/fit/goekdeniz-guelmez-josie-2-9b-oss?chip=m5-max&ram=36)

General · goekdeniz-guelmez · 2026-07-31

Q8\_0Excellent

11.3 GB31% of RAM~25 tok/sEstimated9.65B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[VAETKI VL 7B A1B](/fit/nc-ai-consortium-vaetki-vl-7b-a1b?chip=m5-max&ram=36)

Multimodal · NCAI · 2025-12-29

Q8\_0Excellent

9.0 GB25% of RAMBenchmark needed7.58B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[LaSER Qwen3 8B](/fit/alibaba-nlp-laser-qwen3-8b?chip=m5-max&ram=36)

General · alibaba-nlp · 2026-03-31

Q8\_0Excellent

9.6 GB27% of RAM~29 tok/sEstimated8.19B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen2.5 VL 3B Instruct](/fit/qwen-qwen2-5-vl-3b-instruct?chip=m5-max&ram=36)

Multimodal · Alibaba · 2025-01-26

Q8\_0Excellent

4.7 GB13% of RAM~64 tok/sEstimated3.75B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[gemma 4 12b it](/fit/google-gemma-4-12b-it?chip=m5-max&ram=36)

Multimodal · google · 2026-05

Q8\_0Excellent

13.8 GB38% of RAM~20 tok/sEstimated11.96B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[DeepSeek R1 0528 Qwen3 8B](/fit/deepseek-ai-deepseek-r1-0528-qwen3-8b?chip=m5-max&ram=36)

Reasoning · DeepSeek · 2025-05-29

Q8\_0Excellent

9.6 GB27% of RAM~29 tok/sEstimated8.19B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[SmolLM3 3B](/fit/huggingfacetb-smollm3-3b?chip=m5-max&ram=36)

General · huggingfacetb · 2025-07-08

Q8\_0Excellent

3.9 GB11% of RAM~78 tok/sEstimated3.08B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[DeepSeek R1 Distill Qwen 1.5B](/fit/deepseek-ai-deepseek-r1-distill-qwen-1-5b?chip=m5-max&ram=36)

Reasoning · DeepSeek · 2025-01-20

Q8\_0Excellent

2.5 GB7% of RAM~135 tok/sEstimated1.78B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3Guard Gen 4B](/fit/qwen-qwen3guard-gen-4b?chip=m5-max&ram=36)

General · Alibaba · 2025-09-23

Q8\_0Excellent

5.4 GB15% of RAM~55 tok/sEstimated4.41B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[SmolLM3 3B Base](/fit/huggingfacetb-smollm3-3b-base?chip=m5-max&ram=36)

General · huggingfacetb · 2025-06-19

Q8\_0Excellent

3.9 GB11% of RAM~78 tok/sEstimated3.08B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3Guard Gen 0.6B](/fit/qwen-qwen3guard-gen-0-6b?chip=m5-max&ram=36)

General · Alibaba · 2025-09-23

Q8\_0Excellent

1.3 GB4% of RAM~321 tok/sEstimated0.75B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Olmo 3 1025 7B](/fit/allenai-olmo-3-1025-7b?chip=m5-max&ram=36)

General · allenai · 2025-09-12

Q8\_0Excellent

8.6 GB24% of RAM~33 tok/sEstimated7.3B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[TimeOmni 1 7B](/fit/anton-hugging-timeomni-1-7b?chip=m5-max&ram=36)

General · anton-hugging · 2026-02-06

Q8\_0Excellent

9.0 GB25% of RAM~32 tok/sEstimated7.62B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 8B DFlash b16](/fit/z-lab-qwen3-8b-dflash-b16?chip=m5-max&ram=36)

General · z-lab · 2026-01-04

Q8\_0Excellent

9.4 GB26% of RAM~30 tok/sEstimated8B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[XCurOS0.1 8B Instruct](/fit/xcuros-xcuros0-1-8b-instruct?chip=m5-max&ram=36)

Chat · xcuros · 2026-02-28

Q8\_0Excellent

9.0 GB25% of RAM~32 tok/sEstimated7.62B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[plamo 2 1b](/fit/pfnet-plamo-2-1b?chip=m5-max&ram=36)

General · pfnet · 2025-02-05

Q8\_0Excellent

1.9 GB5% of RAM~187 tok/sEstimated1.29B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[academic ds 9B](/fit/bytedance-seed-academic-ds-9b?chip=m5-max&ram=36)

General · bytedance-seed · 2025-04-09

Q8\_0Excellent

11.0 GB30% of RAMBenchmark needed9.37B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Llama 3.2 3B Instruct pythonic](/fit/baseten-llama-3-2-3b-instruct-pythonic?chip=m5-max&ram=36)

Chat · baseten · 2025-09-12

Q8\_0Excellent

4.1 GB11% of RAM~75 tok/sEstimated3.21B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[HRM Text 1B](/fit/sapientinc-hrm-text-1b?chip=m5-max&ram=36)

General · sapientinc · 2026-05-17

Q8\_0Excellent

1.8 GB5% of RAM~204 tok/sEstimated1.18B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[NEXUS Medical](/fit/fableforge-ai-nexus-medical?chip=m5-max&ram=36)

General · fableforge-ai · 2026-07-05

Q8\_0Excellent

2.2 GB6% of RAM~156 tok/sEstimated1.54B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[MicroLlama v2](/fit/viorikaai-org-microllama-v2?chip=m5-max&ram=36)

General · viorikaai-org · 2026-07-05

Q8\_0Excellent

0.6 GB2% of RAM~4,819 tok/sEstimated0.05B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[BananaMind 2 Nano](/fit/bananamind-bananamind-2-nano?chip=m5-max&ram=36)

General · bananamind · 2026-07-17

Q8\_0Excellent

0.5 GB1% of RAM~24,095 tok/sEstimated0.01B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[EXAONE Deep 2.4B](/fit/lgai-exaone-exaone-deep-2-4b?chip=m5-max&ram=36)

General · lgai-exaone · 2025-03-12

Q8\_0Excellent

3.2 GB9% of RAM~100 tok/sEstimated2.41B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Nexus Erebus 50M](/fit/maliosdark-nexus-erebus-50m?chip=m5-max&ram=36)

General · maliosdark · 2026-07-09

Q8\_0Excellent

0.6 GB2% of RAM~4,819 tok/sEstimated0.05B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Jolia](/fit/raidium-jolia?chip=m5-max&ram=36)

General · raidium · 2026-06-15

Q8\_0Excellent

0.5 GB1% of RAM~12,048 tok/sEstimated0.02B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[FastContext 1.0 4B SFT](/fit/shaungves-fastcontext-1-0-4b-sft?chip=m5-max&ram=36)

General · shaungves · 2026-06-19

Q8\_0Excellent

5.0 GB14% of RAM~60 tok/sEstimated4.02B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[ielts 3b](/fit/preparebuddy-ielts-3b?chip=m5-max&ram=36)

General · preparebuddy · 2026-06-02

Q8\_0Excellent

3.9 GB11% of RAM~78 tok/sEstimated3.08B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[embeddinggemma GTAIDE 300m 2605](/fit/taide-embeddinggemma-gtaide-300m-2605?chip=m5-max&ram=36)

Embedding · taide · 2026-06-12

Q8\_0Excellent

0.8 GB2% of RAM~803 tok/sEstimated0.3B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 0.6B](/fit/qwen-qwen3-0-6b?chip=m5-max&ram=36)

General · Alibaba · 2025-04-27

Q8\_0Excellent

1.3 GB4% of RAM~321 tok/sEstimated0.75B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 1.7B](/fit/qwen-qwen3-1-7b?chip=m5-max&ram=36)

General · Alibaba · 2025-04-27

Q8\_0Excellent

2.8 GB8% of RAM~119 tok/sEstimated2.03B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 4B](/fit/qwen-qwen3-4b?chip=m5-max&ram=36)

General · Alibaba · 2025-04-27

Q8\_0Excellent

5.0 GB14% of RAM~60 tok/sEstimated4.02B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 VL 4B Instruct](/fit/qwen-qwen3-vl-4b-instruct?chip=m5-max&ram=36)

Multimodal · Alibaba

Q8\_0Excellent

5.5 GB15% of RAM~54 tok/sEstimated4.44B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[GLM OCR](/fit/zai-org-glm-ocr?chip=m5-max&ram=36)

Multimodal · zai-org

Q8\_0Excellent

2.0 GB6% of RAM~181 tok/sEstimated1.33B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 4B Instruct 2507](/fit/qwen-qwen3-4b-instruct-2507?chip=m5-max&ram=36)

Chat · Alibaba · 2025-08-05

Q8\_0Excellent

5.0 GB14% of RAM~60 tok/sEstimated4.02B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 VL 2B Instruct](/fit/qwen-qwen3-vl-2b-instruct?chip=m5-max&ram=36)

Multimodal · Alibaba

Q8\_0Excellent

2.9 GB8% of RAM~113 tok/sEstimated2.13B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Rio 3.0 Open Mini](/fit/prefeitura-rio-rio-3-0-open-mini?chip=m5-max&ram=36)

General · prefeitura-rio

Q8\_0Excellent

5.0 GB14% of RAM~60 tok/sEstimated4.02B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[surya ocr 2](/fit/datalab-to-surya-ocr-2?chip=m5-max&ram=36)

Multimodal · datalab-to

Q8\_0Excellent

1.3 GB4% of RAM~349 tok/sEstimated0.69B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Phi 3.5 vision instruct](/fit/microsoft-phi-3-5-vision-instruct?chip=m5-max&ram=36)

Multimodal · Microsoft

Q8\_0Excellent

5.1 GB14% of RAM~58 tok/sEstimated4.15B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 1.7B Base](/fit/qwen-qwen3-1-7b-base?chip=m5-max&ram=36)

General · Alibaba · 2025-04-28

Q8\_0Excellent

2.4 GB7% of RAM~140 tok/sEstimated1.72B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[MiniCPM V 4.6](/fit/openbmb-minicpm-v-4-6?chip=m5-max&ram=36)

Multimodal · openbmb

Q8\_0Excellent

2.0 GB5% of RAM~185 tok/sEstimated1.3B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 0.6B Base](/fit/qwen-qwen3-0-6b-base?chip=m5-max&ram=36)

General · Alibaba · 2025-04-28

Q8\_0Excellent

1.2 GB3% of RAM~402 tok/sEstimated0.6B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3 4B Base](/fit/qwen-qwen3-4b-base?chip=m5-max&ram=36)

General · Alibaba · 2025-04-28

Q8\_0Excellent

5.0 GB14% of RAM~60 tok/sEstimated4.02B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Nanonets OCR2 3B](/fit/nanonets-nanonets-ocr2-3b?chip=m5-max&ram=36)

Multimodal · nanonets

Q8\_0Excellent

4.7 GB13% of RAM~64 tok/sEstimated3.75B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[dots.mocr](/fit/rednote-hilab-dots-mocr?chip=m5-max&ram=36)

Multimodal · rednote-hilab

Q8\_0Excellent

3.9 GB11% of RAM~79 tok/sEstimated3.04B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Mage VL](/fit/microsoft-mage-vl?chip=m5-max&ram=36)

Multimodal · Microsoft

Q8\_0Excellent

5.8 GB16% of RAM~51 tok/sEstimated4.74B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Olmo 3 7B Instruct](/fit/allenai-olmo-3-7b-instruct?chip=m5-max&ram=36)

Chat · allenai · 2025-11-19

Q8\_0Excellent

8.6 GB24% of RAM~33 tok/sEstimated7.3B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Kimi VL A3B Instruct](/fit/moonshotai-kimi-vl-a3b-instruct?chip=m5-max&ram=36)

Multimodal · moonshotai

Q8\_0Excellent

18.8 GB52% of RAMBenchmark needed16.41B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[T lite it 2.1](/fit/t-tech-t-lite-it-2-1?chip=m5-max&ram=36)

General · t-tech · 2025-12-22

Q8\_0Excellent

9.6 GB27% of RAM~29 tok/sEstimated8.19B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[typhoon ocr 3b](/fit/typhoon-ai-typhoon-ocr-3b?chip=m5-max&ram=36)

Multimodal · typhoon-ai

Q8\_0Excellent

4.7 GB13% of RAM~64 tok/sEstimated3.75B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[DeepSeek R1 Distill Llama 8B](/fit/deepseek-ai-deepseek-r1-distill-llama-8b?chip=m5-max&ram=36)

Reasoning · DeepSeek · 2025-01-20

Q8\_0Excellent

9.5 GB26% of RAM~30 tok/sEstimated8.03B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[HunyuanOCR](/fit/tencent-hunyuanocr?chip=m5-max&ram=36)

Multimodal · tencent

Q8\_0Excellent

1.7 GB5% of RAMBenchmark needed1.12B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[OTel LLM 4B IT](/fit/farbodtavakkoli-otel-llm-4b-it?chip=m5-max&ram=36)

General · farbodtavakkoli

Q8\_0Excellent

5.3 GB15% of RAM~56 tok/sEstimated4.3B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Olmo 3 7B Instruct SFT](/fit/allenai-olmo-3-7b-instruct-sft?chip=m5-max&ram=36)

Chat · allenai · 2025-11-17

Q8\_0Excellent

8.6 GB24% of RAM~33 tok/sEstimated7.3B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[dots.ocr](/fit/dots-studio-dots-ocr?chip=m5-max&ram=36)

Multimodal · dots-studio

Q8\_0Excellent

3.9 GB11% of RAM~79 tok/sEstimated3.04B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[distil lfm25 shellper](/fit/distil-labs-distil-lfm25-shellper?chip=m5-max&ram=36)

General · distil-labs

Q8\_0Excellent

0.9 GB2% of RAM~688 tok/sEstimated0.35B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[DeepSeek R1 Distill Qwen 7B](/fit/deepseek-ai-deepseek-r1-distill-qwen-7b?chip=m5-max&ram=36)

Reasoning · DeepSeek · 2025-01-20

Q8\_0Excellent

9.0 GB25% of RAM~32 tok/sEstimated7.62B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[OTel LLM 1.2B IT](/fit/farbodtavakkoli-otel-llm-1-2b-it?chip=m5-max&ram=36)

General · farbodtavakkoli

Q8\_0Excellent

1.8 GB5% of RAM~199 tok/sEstimated1.21B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Rax 4.5](/fit/raxcore-dev-rax-4-5?chip=m5-max&ram=36)

Multimodal · raxcore-dev

Q8\_0Excellent

3.0 GB8% of RAM~106 tok/sEstimated2.27B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Kimi VL A3B Thinking](/fit/moonshotai-kimi-vl-a3b-thinking?chip=m5-max&ram=36)

Multimodal · moonshotai

Q8\_0Excellent

18.8 GB52% of RAMBenchmark needed16.41B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[OTel LLM 3B IT](/fit/farbodtavakkoli-otel-llm-3b-it?chip=m5-max&ram=36)

General · farbodtavakkoli

Q8\_0Excellent

5.2 GB15% of RAM~57 tok/sEstimated4.25B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[granite 4.0 3b vision](/fit/ibm-granite-granite-4-0-3b-vision?chip=m5-max&ram=36)

General · ibm-granite

Q8\_0Excellent

5.0 GB14% of RAM~60 tok/sEstimated4B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[Qwen3.5 4B CHATML](/fit/phanviethoang1512-qwen3-5-4b-chatml?chip=m5-max&ram=36)

Multimodal · phanviethoang1512

Q8\_0Excellent

5.7 GB16% of RAM~52 tok/sEstimated4.66B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[ReaderLM v2](/fit/jinaai-readerlm-v2?chip=m5-max&ram=36)

General · jinaai

Q8\_0Excellent

2.2 GB6% of RAM~156 tok/sEstimated1.54B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

[OvisOCR2](/fit/ath-maas-ovisocr2?chip=m5-max&ram=36)

Multimodal · ath-maas

Q8\_0Excellent

1.4 GB4% of RAM~283 tok/sEstimated0.85B params

What Mac for this model? [Run with ToolPiper](/toolpiper)

Show more (6413 remaining)

## 32-core vs 40-core GPU

Apple ties the memory bus to the bin on this chip: 460 GB/s at 32 cores and 614 GB/s at 40. That moves tokens per second. It does not move which models fit, because that is memory, not cores.

| Configuration | Memory bandwidth | Memory options | Models that fit |
| --- | --- | --- | --- |
| 18-core CPU, 32-core GPU | 460 GB/s | 36 GB | Identical |
| 18-core CPU, 40-core GPU | 614 GB/s | 48, 64, 128 GB | Identical |

On [openbuddy zero 56b v21.2 32k](/fit/openbuddy-openbuddy-zero-56b-v21-2-32k) the 32-core generates about 12 tok/s and the 40-core about 16 tok/s, a 33% difference. Both hold the model at the same quantization.

## Measured on the M5 Max

Nobody has submitted a benchmark on the M5 Max yet, so every speed on this page is the formula estimate rather than a measured run. The estimate is bandwidth-driven and calibrated against chips that do have data, which makes it a good guide and not a promise.

ToolPiper contributes a result anonymously when you run the benchmark, and the [leaderboard](/fit/benchmarks) shows every chip that already has one.

## Where to go from here

What each step actually changes for local models, rather than which one is newer.

[One tier up

Mac Studio M5 Ultra

2x the memory bandwidth, up to 512 GB instead of 128 GB

](/mac-studio/m5-ultra)[Used market alternative

Mac Studio M4 Max

11% less memory bandwidth

](/mac-studio/m4-max)[Same chip, other Mac

MacBook Pro 14" M5 Max

The same chip in a different Mac

](/macbook-pro-14/m5-max)[Same chip, other Mac

MacBook Pro 16" M5 Max

The same chip in a different Mac

](/macbook-pro-16/m5-max)

## Why unified memory is the number that matters

On a PC the model has to fit in GPU VRAM, which is a separate pool from system RAM and usually the smaller of the two. Apple Silicon has one pool. The M5 Max's 614 GB/s bus is shared by CPU, GPU, and Neural Engine, so a 128 GB machine can hand almost all of that to a model with no copy across a bus.

The Studio exists for this workload. It carries the widest memory buses and the highest capacities Apple sells, and it runs at full clocks indefinitely. Buy the memory, not the cores: every extra GB raises what you can load, while the core count only moves throughput on models that already fit.

## Common questions

### Can the M5 Max Mac Studio run a 70B model?

Yes, at 128 GB. A 70B model at Q4\_K\_M needs about 46 GB including an 8K context, and 128 GB of unified memory leaves about 112 GB for weights once macOS takes its share. At 36 GB it does not fit at any quantization worth running.

### How much unified memory should I get with the M5 Max Mac Studio?

Memory is the only spec that changes what you can run at all. 36 GB holds about a 47B model at Q4; 128 GB holds about 172B. It is soldered, so this is a one-time decision, and it is the upgrade worth paying for before core count.

### How fast are local LLMs on the M5 Max?

Token generation is bandwidth-bound, so M5 Max throughput scales with its 614 GB/s memory bus. Divide bandwidth by the size of the weights actually read per token to get the ceiling, then expect roughly half of that in practice. A 7B model at Q4 reads about 4 GB per token pass, so the M5 Max lands in the tens of tokens per second and a 70B model lands in the single digits.

### Is the 40-core GPU worth it over the 32-core on the M5 Max?

For throughput, yes: Apple ties bandwidth to the bin here, so the 32-core runs at 460 GB/s and the 40-core at 614 GB/s, about 33% more. For fit, no: both bins hold exactly the same models, because that is set by memory rather than by cores.

### Should I buy the M5 Max Mac Studio now or wait for the next one?

Buy on the memory you need today. Apple raises memory ceilings slowly and bandwidth in steps, and the M5 Max already holds about a 172B model at Q4. If your target model fits in 128 GB, waiting buys throughput rather than capability.

## Run these models on your Mac Studio

ToolPiper downloads, manages, and runs local models on Apple Silicon. Free, and nothing leaves the machine.

[Get ToolPiper](/toolpiper)[Compare against another Mac](/fit?chip=m5-max&ram=36)

[![ModelPiper](modelpiper-logo-sm.png)ModelPiper](/)

Runs on ToolPiper  
Local AI  
Private by default  
Talk to your Mac

#### Products

-   [ToolPiper](/toolpiper)
-   [VisionPiper](/visionpiper)
-   [AudioPiper](/audiopiper)
-   [MediaPiper](/docs/mediapiper)
-   [PiperTest](/pipertest)

#### Mac Hardware

-   [MacBook Air](/macbook-air)
-   [MacBook Pro 14"](/macbook-pro-14)
-   [MacBook Pro 16"](/macbook-pro-16)
-   [Mac mini](/mac-mini)
-   [Mac Studio](/mac-studio)
-   [iMac](/imac)

#### Resources

-   [Model Fit](/fit)
-   [Benchmarks](/fit/benchmarks)
-   [Blog](/blog)
-   [Workflows](/workflow)
-   [Docs](/docs)
-   [MCP Tools](/mcp-tools)

#### Company

-   [About](/about)
-   [Press](/press)
-   [Changelog](/changelog)
-   [Pricing](/pricing)
-   [Download](/download)
-   [Privacy](/privacy)
-   [Terms](/terms)

© 2026 ModelPiper. All rights reserved.
