(1)
Structured Pruning of Small Language Models: An Empirical Study on GQA-Aware Attention and MLP Compression With LoRA Recovery. Int. J. Adv. Soft Comput. Appl. 2026, 18 (3), 113–139. https://doi.org/10.15849/ijasca.171.