We compressed Hy4-preview from 1.5TB to ~200GiB GGUF and it still works well !
Meet MIX-STQ1_0.The trick isn’t just going low, it’s deciding where: calibration data picks each layer’s bit-width, some down to 1.31-bit STQ1_0, some up to 2.06-bit IQ2_XXS. Same budget, lower error.
Accuracy barely moves vs BF16 📊 MCP Atlas 83.7→83.2 📊 SWE-Bench multi 82.9→81.3 📊 MRCR 81.3→81.1 📊 IFBench 73.5→72.5
See the details on HF : AngelSlim/Hy4-preview-GGUF
Weights & low-bit GGUFs 👇 https://huggingface.co/AngelSlim/Hy4-preview-GGUF
#LLM #Quantization #llamacpp #Hy