What is Quantization?
Quantisation stores a model’s weights at lower numeric precision — 8-bit or 4-bit instead of 16-bit floats — cutting memory use substantially at some cost in accuracy.
2 tools that work with Quantization
Files never leave your browser.
Related terms
Terms that appear alongside Quantization on the same tools.