Energy-Efficient NLP Through Tiny-Model Distillation, Pruning, and Quantized Inference. (2026). Kashf Journal of Multidisciplinary Research, 3(05), 1-8. https://doi.org/10.71146/kjmr965