GGUF File Size by Quantisation Level
Work out the download size of a quantised model at every GGUF level, and what quality each one costs you.
A GGUF file is parameters × effective bits per weight ÷ 8. The effective figure is higher than the nominal one because each block of weights carries its own scale — Q4_K_M is 4.83 bits per weight, not 4 — so an 8B model at Q4_K_M is about 4.5 GB rather than 4. This calculator gives the size at all eight common levels with what each costs in quality.
Runs in your browser- Privacy
- Runs entirely in your browser — nothing is uploaded
- Cost
- Free, unlimited, no sign-up
Frequently asked questions
Which quantisation should I download?
Why is Q4_K_M bigger than Q4_0 if both are 4-bit?
How to use the gguf quantisation size calculator
- 1Set the parameters.
- 2Choose the quantisation.
- 3Turn "Show every level" on or off as needed.
- 4The result appears immediately — copy or download it.
Sources & specifications
Embed this tool
Put the working gguf quantisation size calculator on your own site. It runs in your visitors' browsers exactly as it does here — free, no account, nothing uploaded.
Share this tool
Related tools
LLM VRAM CalculatorWork out exactly how much video memory a local model needs, including the KV cache that scales with context length.Base64 to File ConverterDecode Base64, a data URI or hex back into a downloadable file, with the type detected automatically.Paste Image and DownloadPress Ctrl+V to paste a screenshot or copied image, then download it as a real file. Nothing is uploaded.Sample File LibraryDownload a ready-made ladder of 15 sample files: size steps, page counts, aspect ratios, durations or row counts.AI API Cost CalculatorWork out what a prompt costs per call, per day and per month, and compare the same workload across every major model.AI Writing Pattern CheckerMeasure sentence-length variance, vocabulary richness, marker phrases and typographic tells, with per-sentence scoring.
Last updated
More ai & llm tools