Jain, Raunit, Raaj Shah, Khilan Kanadia, Jinay Jain, Dr. Kranti Ghag, and Dr. Meera Narvekar. “Embedding Instability Score for Detecting Backdoor Attacks in NLP Models”. International Journal of Artificial Intelligence and Machine Learning 6, no. 3 (September 1, 2026): 242–250. Accessed September 14, 2026. https://svedbergopen.com/index.php/ijaiml/article/view/1715.