“Energy-Efficient NLP Through Tiny-Model Distillation, Pruning, and Quantized Inference”. 2026. Kashf Journal of Multidisciplinary Research 3 (05): 1-8. https://doi.org/10.71146/kjmr965.