GAGraham Andanjeinfata-fanaka.hashnode.dev·Jul 31 · 5 min readModel QuantizationQuantization is a model compression technique that reduces numerical precision of weights and activations from floating-point to lower-bit representations, decreasing model size and computational cost00