Target-Agnostic Adversarial Attacks on Language Understanding Models
Published in arXiv preprint (2106.07047), 2021
Proposes adversarial attacks on language-understanding models that do not assume access to the victim’s label space or target outputs. The attacks transfer across tasks and demonstrate systemic robustness gaps in standard NLU model families.
Recommended citation: Chauhan, J., Bhukar, K., Kaul, M. (2021). "Target-Agnostic Adversarial Attacks on Language Understanding Models." arXiv preprint 2106.07047.
Download Paper
