TAI BUI
← Glossary
Glossary

What is Quantization?

Reducing the precision of model weights from float32 (4 bytes) to int8 (1 byte) or int4 (0.5 bytes). Trades a small amount of accuracy for 4-8x less memory and faster inference. GPTQ, AWQ, and GGUF are common formats.

What people say

Making the model smaller

Why it's called that