COMPARISON · CALCULATED ESTIMATES

Q4 vs Q5 vs Q6 vs Q8

A 32K planning comparison for Qwen3.8-27B on 24GB. Memory estimates are useful; quality claims remain unknown until the same evaluation is run across every artifact.

Estimated memory at 32K context
QuantizationWeight estimatePeak range24GB verdictQuality evidence
Q4_K_M16.7 GB21.4 GB22.3 GBfitsNot yet tested
Q5_K_M19.8 GB24.7 GB25.8 GBdoes not fitNot yet tested
Q6_K22.9 GB28.1 GB29.3 GBdoes not fitNot yet tested
Q8_029.5 GB35.2 GB36.8 GBdoes not fitNot yet tested
BF1655.6 GB63.2 GB66.2 GBdoes not fitNot yet tested
INTERPRETATION

Fit and quality are different questions

Lower-bit weights create more memory headroom, but that does not establish an acceptable quality trade-off. The lab will only publish a comparative quality statement after the evaluation prompts, scoring method, runtime and generation settings are held constant.