[1]
BAKKOU, A. et al. trans. 2026. Structured Pruning of Small Language Models: An Empirical Study on GQA-Aware Attention and MLP Compression with LoRA Recovery. International Journal of Advances in Soft Computing and its Applications . 18, 3 (Oct. 2026), 113–139. DOI:https://doi.org/10.15849/ijasca.171.