An FPGA-Based Unified Processing Element for INT8/Binary Quantization and its Scalable Array Design
The widespread deployment of deep neural networks on edge devices faces a severe imbalance between computational demand and available power, while devices frequently switch between low-power standby and highperformance detection modes. Existing general-purpose processors, graphics processing units, and fixed-precision...