← Glossary

quantization

Storing a model's numbers at lower precision so it fits in less memory and runs faster, at some cost to quality.

ai
Nobody says this one in the transcripts we have yet.