← AI Hashrate 中文

Nanbeige

Nanbeige · 1 models

Nanbeige's compact LLM lab uses Looped Transformer architecture to maximize capacity per parameter. Nanbeige4.2-3B outperforms 9B+ models on agentic and reasoning benchmarks while fitting in under 2.5 GB at Q4.

Models in this family

1 models in the Nanbeige family, grouped by series. Q4 (GB) is weights-only; total VRAM adds KV cache and overhead — the fit check links compute it for a representative retail GPU at 8K context.

Nanbeige 4.2

ModelParams (B)Ctx (K)Q4 (GB)LicenseFit check
Nanbeige4.2-3B4.01282.2apache-2.0on Tesla P4 8GB