Back to the Glossary
Training

Quantization

Reducing the numerical precision a model uses (for example from 32-bit to 4-bit) to make it faster and smaller in memory.

More in the same category

Browse the full glossary