Jain, Raunit, et al. “Embedding Instability Score for Detecting Backdoor Attacks in NLP Models”. International Journal of Artificial Intelligence and Machine Learning, vol. 6, no. 3, Sept. 2026, pp. 242-50, doi:10.51483/IJAIML.6.3.2026.242-250.