Cache me if you can issue #7
Quantization in Machine Learning 🤖
Quantization stands as an optimization technique aimed at diminishing the computational and memory burdens of inference tasks. It accomplishes this by expressing model weights and activations using lower-precision ...
krietallo.hashnode.dev3 min read