Wondering how much disk space a model will take at Q4_K_M vs Q8? Enter the parameter count and this shows the estimated file size for every common quantization level – no need to check model pages one by one.

Estimates assume ~1.25 bytes per parameter overhead for the tokenizer and metadata. Real sizes vary by model architecture. See the GGUF explained guide for what each quantization level means.

Related