Entity
Quantization
A model compression technique that reduces the precision of weights and activations to lower bit representations.
术语属性
- Purpose:Model optimization
- Benefit:Reduced memory and computation
Entity
A model compression technique that reduces the precision of weights and activations to lower bit representations.