also, I think this 12GB Q3_S is the best low-bit quant if you're looking to shave a few GB off of Qwen3.8:
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
huggingface.co
I think the bonsai ternary models are cool but it's worth noting that their '98% benchmark performance' claims are heavily cherry-picked. empirically it is not anywhere close to 98% as good overall as ≥4-bit Qwen 3.8 27B.