← AI Hashrate 中文

DGX Spark (GH200 480GB)

NVIDIA · 480 GB LPDDR5X · 500 GB/s · 170 W · MSRP $3,000

CUDA

The DGX Spark (GH200 480GB) is a NVIDIA accelerator with 480 GB LPDDR5X and about 500 GB/s peak memory bandwidth. AI Hashrate estimates local LLM decode (batch 1) at 8K context: fit counts weights + KV cache + 1 GB runtime overhead — not weights alone. At Q4, about 66 curated models fit fully on this card; at FP16, about 55 fit. Top Q4 speeds: MiniCPM5-1B ≈ 294.6 tok/s (estimated, fits); Qwen3.5-2B ≈ 159.1 tok/s (estimated, fits); DeepSeek-R1-0528-Qwen3-8B-layer-mix-bpw-3.8-mlx ≈ 138.3 tok/s (estimated, fits); Qwen3.6-35B-A3B ≈ 106.1 tok/s (estimated, fits); Qwen3.5-35B-A3B ≈ 106.1 tok/s (estimated, fits). Relative ranking is more reliable than absolute tok/s. See methodology for the bandwidth formula and measured-anchor policy. Methodology.

🛒 Where to Buy
affiliate link · we earn from qualifying purchases

Table: decode speed estimates (Q4 / FP16). Measured rows override estimates. MSRP is list price, not live retail. Context default 8K.

ModelParams (B)Quanttok/sFits?
MiniCPM5-1B1.08Q4294.6 est.Yes
Qwen3.5-2B2.0Q4159.1 est.Yes
DeepSeek-R1-0528-Qwen3-8B-layer-mix-bpw-3.8-mlx2.3Q4138.3 est.Yes
Qwen3.6-35B-A3B36.0Q4106.1 est.Yes
Qwen3.5-35B-A3B34.7Q4106.1 est.Yes
GLM-4.7-Flash31.2Q4106.1 est.Yes
Qwen3-30B-A3B-Instruct-250730.5Q4106.1 est.Yes
Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF1630.0Q4106.1 est.Yes
Qwen3-Coder-Next79.7Q4106.1 est.Yes
Ornith-1.0-35B35.0Q4106.1 est.Yes
Qwen3-Next-80B-A3B-Instruct81.3Q4106.1 est.Yes
qwen35b-a3b-fable-sft-abliterated34.7Q4106.1 est.Yes
Llama-3.2-3B-Instruct3.2Q499.4 est.Yes
gpt-oss-20b21.5Q488.4 est.Yes
gemma-4-26B-A4B-it25.2Q483.7 est.Yes
MiniCPM5-1B1.08FP1681.0 est.Yes
Qwen3-4B-Instruct-25074.0Q479.5 est.Yes
NVIDIA-Nemotron-3-Nano-4B-BF164.0Q479.5 est.Yes
Qwen3-4B-Thinking-25074.0Q479.5 est.Yes
gemma-4-E2B-it5.1Q462.4 est.Yes
gpt-oss-120b120.4Q462.4 est.Yes
Mistral-Small-4-119B-2603119.0Q449.0 est.Yes
Mistral-7B-Instruct-v0.37.2Q444.2 est.Yes
Qwen3.5-2B2.0FP1643.8 est.Yes
Qwen2.5-7B-Instruct7.6Q441.9 est.Yes
Llama-3.1-8B-Instruct8.0Q439.8 est.Yes
gemma-4-E4B-it8.0Q439.8 est.Yes
Qwen3-8B8.2Q438.8 est.Yes
DeepSeek-R1-0528-Qwen3-8B8.2Q438.8 est.Yes
DeepSeek-R1-0528-Qwen3-8B-layer-mix-bpw-3.8-mlx2.3FP1638.0 est.Yes
Qwen3.5-9B8.95Q435.6 est.Yes
Ornith-1.0-9B9.0Q435.4 est.Yes
Qwen3.5-122B-A10B122.0Q431.8 est.Yes
MiniMax-M2.7228.7Q431.8 est.Yes
Qwen3.6-35B-A3B36.0FP1629.2 est.Yes
Qwen3.5-35B-A3B34.7FP1629.2 est.Yes
GLM-4.7-Flash31.2FP1629.2 est.Yes
Qwen3-30B-A3B-Instruct-250730.5FP1629.2 est.Yes
Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF1630.0FP1629.2 est.Yes
Qwen3-Coder-Next79.7FP1629.2 est.Yes
Ornith-1.0-35B35.0FP1629.2 est.Yes
Qwen3-Next-80B-A3B-Instruct81.3FP1629.2 est.Yes
qwen35b-a3b-fable-sft-abliterated34.7FP1629.2 est.Yes
Step-3.5-Flash196.8Q428.9 est.Yes
Llama-3.2-3B-Instruct3.2FP1627.3 est.Yes
gemma-4-12b-it11.95Q426.6 est.Yes
NVIDIA-Nemotron-3-Super-120B-A12B-BF16120.0Q426.5 est.Yes
GLM-4.5-Air110.5Q426.5 est.Yes
DeepSeek-V4-Flash158.1Q424.5 est.Yes
gpt-oss-20b21.5FP1624.3 est.Yes
OLMo-2-1124-13B-Instruct13.7Q423.2 est.Yes
gemma-4-26B-A4B-it25.2FP1623.0 est.Yes
Qwen3-4B-Instruct-25074.0FP1621.9 est.Yes
NVIDIA-Nemotron-3-Nano-4B-BF164.0FP1621.9 est.Yes
Qwen3-4B-Thinking-25074.0FP1621.9 est.Yes
Qwen3-14B14.8Q421.5 est.Yes
Qwen2.5-14B-Instruct14.8Q421.5 est.Yes
DeepSeek-R1-Distill-Qwen-14B14.8Q421.5 est.Yes
MiMo-V2.5310.0Q421.2 est.Yes
Kimi-K2.61030.0Q419.9 est.No
Kimi-K2.7-Code1030.0Q419.9 est.No
Llama-4-Scout-17B-16E-Instruct109.0Q418.7 est.Yes
Qwen3.5-397B-A17B397.0Q418.7 est.Yes
gemma-4-E2B-it5.1FP1617.2 est.Yes
gpt-oss-120b120.4FP1617.2 est.Yes
Inkling975.0Q417.2 est.No
Qwen3.5-122B-A10B-Heretic-v2-MLX-mixed-3.8bit19.8Q416.1 est.Yes
Hy3295.0Q415.2 est.Yes
Qwen3-235B-A22B-Instruct-2507235.1Q414.5 est.Yes
MiniMax-M3428.0Q413.8 est.Yes
Mistral-Small-4-119B-2603119.0FP1613.5 est.Yes
Devstral-Small-2-24B-Instruct-251224.0Q413.3 est.Yes
Mistral-7B-Instruct-v0.37.2FP1612.2 est.Yes
Qwen3.5-27B27.0Q411.8 est.Yes
gemma-3-27b-it27.4Q411.6 est.Yes
Qwen2.5-7B-Instruct7.6FP1611.5 est.Yes
Qwen3.6-27B27.8Q411.4 est.Yes
Llama-3.1-8B-Instruct8.0FP1610.9 est.Yes
gemma-4-E4B-it8.0FP1610.9 est.Yes
Qwen3-8B8.2FP1610.7 est.Yes
DeepSeek-R1-0528-Qwen3-8B8.2FP1610.7 est.Yes
gemma-4-31B-it30.7Q410.4 est.Yes
MiMo-V2.5310.0FP169.9 est.No
Qwen3.5-9B8.95FP169.8 est.Yes
Qwen2.5-32B-Instruct32.8Q49.7 est.Yes
Qwen2.5-Coder-32B-Instruct32.8Q49.7 est.Yes
DeepSeek-R1-Distill-Qwen-32B32.8Q49.7 est.Yes
Ornith-1.0-9B9.0FP169.7 est.Yes
Qwen3-Coder-480B-A35B-Instruct480.2Q49.1 est.Yes
Qwen3.5-122B-A10B122.0FP168.8 est.Yes
MiniMax-M2.7228.7FP168.8 est.Yes
DeepSeek-R1684.5Q48.6 est.Yes
Step-3.5-Flash196.8FP168.0 est.Yes
GLM-5.2753.3Q48.0 est.Yes
GLM-5.1753.9Q48.0 est.Yes
Hy3295.0FP167.7 est.No
gemma-4-12b-it11.95FP167.3 est.Yes
NVIDIA-Nemotron-3-Super-120B-A12B-BF16120.0FP167.3 est.Yes
GLM-4.5-Air110.5FP167.3 est.Yes
DeepSeek-V4-Flash158.1FP166.7 est.Yes
DeepSeek-V4-Pro861.6Q46.5 est.Yes
OLMo-2-1124-13B-Instruct13.7FP166.4 est.Yes
Qwen3-14B14.8FP165.9 est.Yes
Qwen2.5-14B-Instruct14.8FP165.9 est.Yes
DeepSeek-R1-Distill-Qwen-14B14.8FP165.9 est.Yes
NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4550.0Q45.8 est.Yes
Qwen3.5-397B-A17B397.0FP165.4 est.No
Llama-4-Scout-17B-16E-Instruct109.0FP165.1 est.Yes
Llama-3.1-70B-Instruct70.6Q44.5 est.Yes
Llama-3.3-70B-Instruct70.6Q44.5 est.Yes
Qwen2.5-72B-Instruct72.7Q44.4 est.Yes
Qwen3.5-122B-A10B-Heretic-v2-MLX-mixed-3.8bit19.8FP164.4 est.Yes
Qwen3-235B-A22B-Instruct-2507235.1FP164.0 est.Yes
Devstral-Small-2-24B-Instruct-251224.0FP163.6 est.Yes
MiniMax-M3428.0FP163.4 est.No
Qwen3.5-27B27.0FP163.2 est.Yes
gemma-3-27b-it27.4FP163.2 est.Yes
Qwen3.6-27B27.8FP163.1 est.Yes
gemma-4-31B-it30.7FP162.9 est.Yes
Qwen2.5-32B-Instruct32.8FP162.7 est.Yes
Qwen2.5-Coder-32B-Instruct32.8FP162.7 est.Yes
DeepSeek-R1-Distill-Qwen-32B32.8FP162.7 est.Yes
Mistral-Medium-3.5-128B128.0Q42.5 est.Yes
Qwen3-Coder-480B-A35B-Instruct480.2FP161.8 est.No
Llama-3.1-70B-Instruct70.6FP161.2 est.Yes
Llama-3.3-70B-Instruct70.6FP161.2 est.Yes
Qwen2.5-72B-Instruct72.7FP161.2 est.Yes
NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4550.0FP160.9 est.No
DeepSeek-R1684.5FP160.8 est.No
AI21-Jamba-Large-1.5398.6Q40.8 est.Yes
Mistral-Medium-3.5-128B128.0FP160.7 est.Yes
GLM-5.2753.3FP160.6 est.No
GLM-5.1753.9FP160.6 est.No
DeepSeek-V4-Pro861.6FP160.4 est.No
AI21-Jamba-Large-1.5398.6FP160.2 est.No

Fit checks

All 69 model fit checks — see the linked Fits? column in the table above