FIT-GGUF enables size-targeted mixed-precision quantization of MiniCPM5-2B
Original title🚀 FIT-GGUF brings controllable-size mixed-precision quantization to MiniCPM5-2B
AISummary
Developer @Scorp1o_117 used FIT-GGUF to build four MiniCPM5-2B GGUF variants, ranging from about 1.14 GiB to 1.46 GiB, tuned to target file sizes or fidelity tiers.
Instead of fixed presets, FIT-GGUF allocates precision tensor by tensor, with Quality, Balanced, Compact, and Mini options, and its generated files matched predicted sizes.
Builds are evaluated with KL Divergence and Same-top metrics and are available on Hugging Face.
Source: OpenBMB · x.comPublished · added here