Family: Qwen3.6
Parameters: 27B
Structure: 27B dense
License: Apache License 2.0; allowed. For guidance only; review the model licence before commercial use.
Native context: 262,144 tokens
Extended context: 1,010,000 via YaRN; not a default fit assumption
Quantization: Q4_K_M
Approx. Q4 weights: 17 GB
Default estimate: 23 GB @ 8,192 tokens
Weights / KV / runtime / margin: 17 GB / 2 GB / 1.5 GB / 2.5 GB
CPU/RAM fallback: Not recommended
VRAM: 24 GB minimum / 32 GB recommended
RAM: 48 GB minimum / 64 GB recommended
Current run mode: Partial GPU offload only
Expected experience: Very slow
Full GPU offload: Only when the memory estimate and context fit
Context warning: Repository-scale or very long document context can push a 24GB card beyond a comfortable fit.