Glossary
Quantization
Also: quantization · квантование
Storing model weights at lower precision, say 8 or 4 bits instead of 16. The model takes less memory and runs faster at a small cost in quality.
Example
An open model quantized to 4 bits runs locally on a laptop, where it wouldn't fit at full precision.