Skip to content
Convertto

What is Quantization?

Quantisation stores a model’s weights at lower numeric precision — 8-bit or 4-bit instead of 16-bit floats — cutting memory use substantially at some cost in accuracy.

2 tools that work with Quantization

Files never leave your browser.

Related terms

Terms that appear alongside Quantization on the same tools.