Skip to content
Read the original: OpenBMB· Published 34/100AI score34/100

FIT-GGUF enables size-targeted mixed-precision quantization of MiniCPM5-2B

Original title🚀 FIT-GGUF brings controllable-size mixed-precision quantization to MiniCPM5-2B

AISummary

Developer @Scorp1o_117 used FIT-GGUF to build four MiniCPM5-2B GGUF variants, ranging from about 1.14 GiB to 1.46 GiB, tuned to target file sizes or fidelity tiers.

Instead of fixed presets, FIT-GGUF allocates precision tensor by tensor, with Quality, Balanced, Compact, and Mini options, and its generated files matched predicted sizes.

Builds are evaluated with KL Divergence and Same-top metrics and are available on Hugging Face.

Read the original x.com

Source: OpenBMB · x.comPublished · added here