GAGraham Andanjeinfata-fanaka.hashnode.dev·1d ago · 5 min readModel QuantizationQuantization is a model compression technique that reduces numerical precision of weights and activations from floating-point to lower-bit representations, decreasing model size and computational cost00