MODEL · ALIBABA · 27B DENSE
Qwen 3.6-27B
The April 2026 dense refresh that supersedes Qwen 3.5 27B — claims to beat the prior 397B MoE flagship on coding benchmarks while staying single-GPU deployable at Q4. Superseded in turn by Qwen3.8-27B (August 5, 2026), which holds the same size, licence and 24 GB target but switches to a Gated DeltaNet hybrid that roughly halves the KV cache at long context. This entry is kept as the dated record and as the fallback if you want plain full attention on every layer; new deployments should start from 3.8.
License: Apache 2.0 · Context: 262K native, extendable to ~1M via YaRN · Released: April 22, 2026
The decision in five lines
- The call
- Consider — runnable locally, family reference
- Best for
- Local evaluation and family reference
- Runs on
- 16 hardware picks fit (cheapest: Minisforum UM890 Pro · $463 barebone (no RAM / SSD); $991 with 32 GB / 1 TB — Minisforum US store, Sep 15 2026)
- Watch out
- When you need fastest possible tok/s on 24 GB VRAM — the 3.6-35B-A3B MoE sibling runs ~6× faster for comparable quality because only 3B activate per token.
- Evidence
- Estimated
- 27B dense
- PARAMETERS
- DENSE
- TYPE
- 262K
- CONTEXT
- ~17 GB
- VRAM AT Q4
Where we recommend this
This model isn’t currently in an active planner slot. See the runner notes below if you’re running it anyway.
The call
The April 2026 dense refresh that supersedes Qwen 3.5 27B — claims to beat the prior 397B MoE flagship on coding benchmarks while staying single-GPU deployable at Q4. Superseded in turn by Qwen3.8-27B (August 5, 2026), which holds the same size, licence and 24 GB target but switches to a Gated DeltaNet hybrid that roughly halves the KV cache at long context. This entry is kept as the dated record and as the fallback if you want plain full attention on every layer; new deployments should start from 3.8.
When not to use: When you need fastest possible tok/s on 24 GB VRAM — the 3.6-35B-A3B MoE sibling runs ~6× faster for comparable quality because only 3B activate per token. The dense 27B trades speed for the simpler dense-attention behavior some workflows prefer.
Runner notes
GGUFs available via unsloth and bartowski. Native Ollama tag may take 1–2 weeks to land. Use vLLM or llama.cpp for long-context work; MLX-community builds for Apple Silicon.
Hardware that fits
Every hardware pick whose memory fits this model at the quant we recommend. Sorted cheapest-first — the top row is your best-value fit. Click through for the full buyer’s guide.
- Minisforum UM890 ProGood · 1.3× 32 GB DDR5 (shared) · $463 barebone (no RAM / SSD); $991 with 32 GB / 1 TB — Minisforum US store, Sep 15 2026
- NVIDIA RTX 3090 (used, single)Good · 1.3× 24 GB · $950–$1,200
- AMD Radeon RX 7900 XTXGood · 1.3× 24 GB · $1,300–$1,500 new (no refurb in stock Sep 15 2026)
- MacBook Air M5 24 GBRequires tweak · 1.1× 24 GB unified · $1,499–$1,899
- Mac Mini M4 Pro 24 GBRequires tweak · 1.1× 24 GB unified · $1,699 (M5 Pro successor, 24 GB / 512 GB, pre-order; M4 Pro discontinued Aug 25 2026)
- Dual RTX 3090 (used)Perfect · 2.5× 48 GB · $1,800–$2,500 all-in
- NVIDIA RTX 4090Good · 1.3× 24 GB · $2,200–$2,800
- M5 Pro MacBook Pro 48 GBPerfect · 1.7× 48 GB unified · $2,999–$3,599
- Framework Desktop (Ryzen AI Max+ 395)Perfect · 4.5× 128 GB unified · $3,449 (128 GB config)
- NVIDIA RTX A6000 (48 GB, used)Perfect · 2.5× 48 GB ECC · $3,500–$4,500
- Mac Studio M4 Max 64 GBPerfect · 2.3× 64 GB unified · $3,799 (M5 Max successor, 64 GB / 1 TB, pre-order; M4 Max discontinued Aug 25 2026)
- NVIDIA DGX SparkPerfect · 4.5× 128 GB unified · $4,699
- M5 Max MacBook Pro 64 GBPerfect · 2.3× 64 GB unified · ~$5,199 (est.; June 25 2026 increase)
- Mac Studio M3 Ultra 96 GBPerfect · 3.4× 96 GB unified · $5,499 (M5 Ultra successor, 96 GB / 1 TB, pre-order; M3 Ultra discontinued Aug 25 2026)
- NVIDIA RTX 5090Perfect · 1.7× 32 GB · $6,450–$7,000 (new, in stock)
- Dual RTX 5090Perfect · 3.4× 64 GB (2×32) · $13,300–$14,000 all-in (two cards alone are ~$12,900 at the Sep 15 2026 Newegg floor)
Next step
Find-by-model — see what hardware runs this→