Skip to content

Author

Asraf Mohamed Moubark

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Open access Jun 2026

Hardware-Aware Optimization of Large Language Models: A System-Level Analysis

A hardware-aware, system-level analysis of key optimization techniques, including pruning, quantization, knowledge distillation, Low-Rank Adaptation (LoRA), and Neural Architecture Search (NAS), shows that quantization consistently achieves the highest inference speedups and memory efficiency.

A. Hamza, Amel Tuama, Asraf Mohamed Moubark · 0 citations