科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Journal of Intelligent Decision Making and Information Science2026-07-31· Adversarial system

Needle in a Haystack: Decamouflaging Adversarial Examples Using SBERT Embeddings

Sudha Pelluri Sai Reethi Pydi

原始摘要(英文原文)· Original abstract
Natural Language Processing models are vulnerable to adversarial perturbations which can derail the model’s classification ability. Existing works focus on correcting the training data in order to be resilient to these attacks. In this paper, we propose the idea of multi label classification for adversarial attacks. First, we introduce a new anagram based attack into the literature and second, we train our model to learn these representations by adding a novel label to the dataset which is used as a feature for training the model. Our results show that our method succeeds in separating clean samples from adversarial ones without changing the actual data in the process. We conducted our experiments on the internet movie database dataset. Results show that our method makes a model reliable and robust against character-level perturbations.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Needle in a Haystack: Decamouflaging Adversarial Examples Using SBERT Embeddings — 科研速览 Science Skim