Back to feed
Reddit r/LocalLLaMA·

Here is my llama.cpp NVFP4/MXFP6 GGUF quantizer tool

Signal
72
Hype
25
In three linesOpen-source GGUF quantization tool (MIT license) for creating NVFP4 and MXFP6 models. Uses imatrix and logits KLD to evaluate and blend multiple quantization methods per layer. Shows better performance than ModelOpt on Qwen 27B. Includes detailed reports and reproducible validation.
Read source
Your take?
LlamaOpen sourceToolsBenchmarks

Summary generated by Claude — human-verified