← AI Hashrate 中文

MacBook Neo A18 Pro 8GB

Apple · 8 GB LPDDR5X · 60 GB/s · 20 W · MSRP $0

METAL

The MacBook Neo A18 Pro 8GB is a Apple accelerator with 8 GB LPDDR5X and about 60 GB/s peak memory bandwidth. AI Hashrate estimates local LLM decode (batch 1) at 8K context: fit counts weights + KV cache + 1 GB runtime overhead — not weights alone. At Q4, about 11 curated models fit fully on this card; at FP16, about 2 fit. Top Q4 speeds: MiniCPM5-1B ≈ 70.0 tok/s (estimated, fits); Qwen3.5-2B ≈ 37.8 tok/s (estimated, fits); Llama-3.2-3B-Instruct ≈ 23.6 tok/s (estimated, fits); Nanbeige4.2-3B ≈ 19.2 tok/s (estimated, fits); Qwen3-4B-Instruct-2507 ≈ 18.9 tok/s (estimated, fits). Relative ranking is more reliable than absolute tok/s. See methodology for the bandwidth formula and measured-anchor policy. Methodology.

🛒 Where to Buy
affiliate link · we earn from qualifying purchases

Table: decode speed estimates (Q4 / FP16). Measured rows override estimates. MSRP is list price, not live retail. Context default 8K.

ModelParams (B)Quanttok/sFits?
MiniCPM5-1B1.08Q470.0 est.Yes
Qwen3.5-2B2.0Q437.8 est.Yes
Llama-3.2-3B-Instruct3.2Q423.6 est.Yes
MiniCPM5-1B1.08FP1619.6 est.Yes
Nanbeige4.2-3B4.0Q419.2 est.Yes
Qwen3-4B-Instruct-25074.0Q418.9 est.Yes
NVIDIA-Nemotron-3-Nano-4B-BF164.0Q418.9 est.Yes
Qwen3-4B-Thinking-25074.0Q418.9 est.Yes
gemma-4-E2B-it5.1Q414.8 est.Yes
Qwen3.5-9B8.95Q411.2 est.No
Ornith-1.0-9B9.0Q411.0 est.No
Qwen3.5-2B2.0FP1610.6 est.Yes
Qwen3-8B8.2Q410.5 est.No
Mistral-7B-Instruct-v0.37.2Q410.5 est.Tight
DeepSeek-R1-0528-Qwen3-8B8.2Q410.5 est.No
Qwen2.5-7B-Instruct7.6Q49.9 est.Tight
Llama-3.1-8B-Instruct8.0Q49.6 est.No
gemma-4-E4B-it8.0Q49.4 est.Tight
gpt-oss-20b21.5Q45.7 est.No
Llama-3.2-3B-Instruct3.2FP165.1 est.No
gemma-4-12b-it11.95Q44.7 est.No
gemma-4-26B-A4B-it25.2Q43.9 est.No
Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF1630.0Q43.5 est.No
Qwen3-30B-A3B-Instruct-250730.5Q43.3 est.No
Qwen3.5-35B-A3B34.7Q42.8 est.No
Ornith-1.0-35B35.0Q42.7 est.No
Qwen3.6-35B-A3B36.0Q42.6 est.No
Nanbeige4.2-3B4.0FP162.6 est.No
Qwen3-4B-Instruct-25074.0FP162.5 est.No
GLM-4.7-Flash31.2Q42.5 est.No
Qwen3-4B-Thinking-25074.0FP162.5 est.No
NVIDIA-Nemotron-3-Nano-4B-BF164.0FP162.4 est.No
Qwen3-14B14.8Q42.2 est.No
Qwen2.5-14B-Instruct14.8Q42.1 est.No
DeepSeek-R1-Distill-Qwen-14B14.8Q42.1 est.No
gemma-4-E2B-it5.1FP161.6 est.No
Devstral-Small-2-24B-Instruct-251224.0Q40.6 est.No
Qwen2.5-7B-Instruct7.6FP160.5 est.No
Mistral-7B-Instruct-v0.37.2FP160.5 est.No
Qwen3.5-27B27.0Q40.5 est.No
Qwen3-8B8.2FP160.4 est.No
Llama-3.1-8B-Instruct8.0FP160.4 est.No
Qwen3.6-27B27.8Q40.4 est.No
gemma-4-E4B-it8.0FP160.4 est.No
DeepSeek-R1-0528-Qwen3-8B8.2FP160.4 est.No
Qwen3.5-9B8.95FP160.3 est.No
gemma-4-31B-it30.7Q40.3 est.No
gemma-3-27b-it27.4Q40.3 est.No
Ornith-1.0-9B9.0FP160.3 est.No
Qwen2.5-32B-Instruct32.8Q40.2 est.No
Qwen2.5-Coder-32B-Instruct32.8Q40.2 est.No
DeepSeek-R1-Distill-Qwen-32B32.8Q40.2 est.No

Fit checks

All 67 model fit checks — see the linked Fits? column in the table above