Related Experiment Videos
Layer-specific approximate multipliers for energy-precision trade-offs in convolutional neural networks
Ladan Sayadi1, Mohammad Hossein Moaiyeri2, Somayeh Timarchi1
1Faculty of Electrical Engineering, Shahid Beheshti University, Tehran, 1983969411, Iran.
None:
Approximate computing is a promising paradigm for error-resilient applications, such as image and signal processing and neural networks, prioritizing hardware efficiency over precision. This paper presents a novel CNN-specific approximation methodology divided into three sections. The first section identifies the characteristics of suitable approximate multipliers for CNNs, highlighting how variance in weight distribution affects the error tolerance of different layers. Section Two introduces three types of approximate multipliers (AM_5×5, AM_4×4, and AM_3×3), designed using innovative operand truncation techniques. These multipliers enable adjustable accuracy by varying the number of truncated operand bits and are scalable to any multiplier size. Two algorithms are also proposed: one optimizes training with approximate multipliers for improved performance, and another employs a gradual training strategy. Section Three describes two distinct strategies for deploying the methodology within CNN architectures and evaluates its hardware implementation on an ASIC in 28 nm CMOS technology. Comprehensive comparisons using VGG16, VGG10, and AlexNet architectures reveal significant improvements in energy efficiency. The first strategy achieves energy efficiency gains of up to 86%, 95%, and 88% per operation for VGG10, VGG16, and AlexNet, respectively, while the second strategy achieves improvements of 81%, 92%, and 84% for the same networks. This approach effectively balances computational complexity and accuracy while leveraging CNN features to enhance hardware efficiency. Experimental results validate the potential of this methodology to advance CNN designs, optimizing both energy and hardware resources for practical applications.
Related Concept Videos
Convolution Properties II
The width property indicates that if the durations of input signals are T1 and T2, then the width of the output response equals the sum of both durations, irrespective of the shapes of the two functions. For instance, convolving two rectangular pulses with durations of 2 seconds and 1 second results in a function with a width of 3 seconds.
The area property asserts that the area under the...
Linear Approximation in Frequency Domain
In contrast, nonlinear systems do not inherently possess these properties. However, for small deviations around an operating point, a nonlinear system can often be approximated as linear....
Linear Approximation in Time Domain
For a simple pendulum with a mass evenly distributed along its length and the center of mass located at half the pendulum's length,...
Accuracy, limits, and approximation
Accuracy is defined as the closeness of the measured value to the true or actual value. In engineering mechanics, repeated measurements are taken during theoretical or experimental analyses to ensure that the result is precise and accurate.
The accuracy of any solution is based on the...
Improving Translational Accuracy
Improving Translational Accuracy