Open access
Jul 2026
Design of Resource-Efficient AI Models through Parameter Reduction and Accuracy-Aware Compression
The proposed hybrid pipeline includes structured pruning, INT8 quantization and task-specific knowledge distillation, which is benchmarked against standalone methods and reinforces the idea of upper bound projection based approach for accuracy-oriented, multi-level compression.
Krishna Kumar Tiwari, Komal Tahiliani, Uma Shankar Birthare et al.
· International journal of com... · 0 citations