keyword
quantized neural networks
Quantized neural networks are artificial neural networks that employ low-precision numerical representations, such as reduced-bit integers or binary values, rather than standard full-precision floating-point numbers for parameters like weights and activations. By mapping continuous or high-precision values to a discrete set of lower-bit formats, these architectures significantly reduce memory footprint, storage requirements, and memory bandwidth demands. The reduced precision also allows computationally intensive floating-point arithmetic to be replaced by faster, more energy-efficient fixed-point or bitwise calculations during inference and training. Consequently, quantized neural networks enable the efficient execution of deep learning models on resource-constrained platforms, such as mobile and embedded edge devices, while maintaining task performance and accuracy comparable to their full-precision counterparts.
1 item

