No — Qwen3-Coder-480B-A35B-Instruct at Q4 needs ≈ 268.61 GB even at 4K context, beyond the MacBook Pro M1 Max 64GB's 64 GB LPDDR5. It would only run with heavy CPU/disk offload. FP16 needs ≈ 968.4 GB — does not fit on this card. Estimates come from memory-bandwidth math; rows tagged measured override estimates. Relative ranking is more reliable than absolute tok/s. Methodology.
Fits = weights + KV(ctx) + 1 GB overhead ≤ 95% of VRAM. Measured anchors are context-agnostic; the fit verdict is recomputed per context. A missing tok/s means the model is far beyond this card (offload-only territory).
VRAM breakdown at 8K context
Quant
Weights
KV cache
Overhead
Total needed
MacBook Pro M1 Max 64GB VRAM
Q4
264.11 GB
7.0 GB
1 GB
272.11 GB
64 GB
FP16
960.4 GB
7.0 GB
1 GB
968.4 GB
64 GB
Other GPUs that run Qwen3-Coder-480B-A35B-Instruct