quantization
Storing a model's numbers at lower precision so it fits in less memory and runs faster, at some cost to quality.
ai
Nobody says this one in the transcripts we have yet.
Storing a model's numbers at lower precision so it fits in less memory and runs faster, at some cost to quality.