The Best Defense is Attack: Repairing Semantics in Textual Adversarial Examples
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Heng, Li, Ke |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Defense Against Syntactic Textual Backdoor Attacks with Token Substitution
by: Li, Xinglin, et al.
Published: (2024)
by: Li, Xinglin, et al.
Published: (2024)
Fast Adversarial Training against Textual Adversarial Attacks
by: Yang, Yichen, et al.
Published: (2024)
by: Yang, Yichen, et al.
Published: (2024)
LimeAttack: Local Explainable Method for Textual Hard-Label Adversarial Attack
by: Zhu, Hai, et al.
Published: (2023)
by: Zhu, Hai, et al.
Published: (2023)
Adversarial Attacks and Defense for Conversation Entailment Task
by: Yang, Zhenning, et al.
Published: (2024)
by: Yang, Zhenning, et al.
Published: (2024)
Are ELECTRA's Sentence Embeddings Beyond Repair? The Case of Semantic Textual Similarity
by: Rep, Ivan, et al.
Published: (2024)
by: Rep, Ivan, et al.
Published: (2024)
Attacking Misinformation Detection Using Adversarial Examples Generated by Language Models
by: Przybyła, Piotr, et al.
Published: (2024)
by: Przybyła, Piotr, et al.
Published: (2024)
Robustness of Misinformation Classification Systems to Adversarial Examples Through BeamAttack
by: Fazla, Arnisa, et al.
Published: (2025)
by: Fazla, Arnisa, et al.
Published: (2025)
SEP-Attack: A Simple and Effective Paradigm for Transfer-Based Textual Adversarial Attack
by: Liu, Han, et al.
Published: (2026)
by: Liu, Han, et al.
Published: (2026)
Defensive Dual Masking for Robust Adversarial Defense
by: Yang, Wangli, et al.
Published: (2024)
by: Yang, Wangli, et al.
Published: (2024)
MPAT: Building Robust Deep Neural Networks against Textual Adversarial Attacks
by: Zhang, Fangyuan, et al.
Published: (2024)
by: Zhang, Fangyuan, et al.
Published: (2024)
Mitigating Adversarial Attacks in LLMs through Defensive Suffix Generation
by: Kim, Minkyoung, et al.
Published: (2024)
by: Kim, Minkyoung, et al.
Published: (2024)
Less is More: Understanding Word-level Textual Adversarial Attack via n-gram Frequency Descend
by: Lu, Ning, et al.
Published: (2023)
by: Lu, Ning, et al.
Published: (2023)
Multi-Granularity Tibetan Textual Adversarial Attack Method Based on Masked Language Model
by: Cao, Xi, et al.
Published: (2024)
by: Cao, Xi, et al.
Published: (2024)
Self-Evaluation as a Defense Against Adversarial Attacks on LLMs
by: Brown, Hannah, et al.
Published: (2024)
by: Brown, Hannah, et al.
Published: (2024)
PHISH in MESH: Korean Adversarial Phonetic Substitution and Phonetic-Semantic Feature Integration Defense
by: Kim, Byungjun, et al.
Published: (2025)
by: Kim, Byungjun, et al.
Published: (2025)
Exploring Gradient-Guided Masked Language Model to Detect Textual Adversarial Attacks
by: Zhang, Xiaomei, et al.
Published: (2025)
by: Zhang, Xiaomei, et al.
Published: (2025)
Alert-ME: An Explainability-Driven Defense Against Adversarial Examples in Transformer-Based Text Classification
by: Sabir, Bushra, et al.
Published: (2023)
by: Sabir, Bushra, et al.
Published: (2023)
Multilingual Evaluation of Semantic Textual Relatedness
by: Endait, Sharvi, et al.
Published: (2024)
by: Endait, Sharvi, et al.
Published: (2024)
TF-Attack: Transferable and Fast Adversarial Attacks on Large Language Models
by: Li, Zelin, et al.
Published: (2024)
by: Li, Zelin, et al.
Published: (2024)
Text-CRS: A Generalized Certified Robustness Framework against Textual Adversarial Attacks
by: Zhang, Xinyu, et al.
Published: (2023)
by: Zhang, Xinyu, et al.
Published: (2023)
LSA: Modeling Aspect Sentiment Coherency via Local Sentiment Aggregation
by: Yang, Heng, et al.
Published: (2021)
by: Yang, Heng, et al.
Published: (2021)
BootAug: Boosting Text Augmentation via Hybrid Instance Filtering Framework
by: Yang, Heng, et al.
Published: (2022)
by: Yang, Heng, et al.
Published: (2022)
LLM Lies: Hallucinations are not Bugs, but Features as Adversarial Examples
by: Yao, Jia-Yu, et al.
Published: (2023)
by: Yao, Jia-Yu, et al.
Published: (2023)
Breaking the Reviewer: Assessing the Vulnerability of Large Language Models in Automated Peer Review Under Textual Adversarial Attacks
by: Lin, Tzu-Ling, et al.
Published: (2025)
by: Lin, Tzu-Ling, et al.
Published: (2025)
Pay Attention to the Robustness of Chinese Minority Language Models! Syllable-level Textual Adversarial Attack on Tibetan Script
by: Cao, Xi, et al.
Published: (2024)
by: Cao, Xi, et al.
Published: (2024)
SCOUT: A Defense Against Data Poisoning Attacks in Fine-Tuned Language Models
by: Afane, Mohamed, et al.
Published: (2025)
by: Afane, Mohamed, et al.
Published: (2025)
SSCAE -- Semantic, Syntactic, and Context-aware natural language Adversarial Examples generator
by: Asl, Javad Rafiei, et al.
Published: (2024)
by: Asl, Javad Rafiei, et al.
Published: (2024)
Pcc-tuning: Breaking the Contrastive Learning Ceiling in Semantic Textual Similarity
by: Zhang, Bowen, et al.
Published: (2024)
by: Zhang, Bowen, et al.
Published: (2024)
Unpacking the Resilience of SNLI Contradiction Examples to Attacks
by: Verma, Chetan, et al.
Published: (2024)
by: Verma, Chetan, et al.
Published: (2024)
Semantic-Preserving Adversarial Attacks on LLMs: An Adaptive Greedy Binary Search Approach
by: Zhang, Chong, et al.
Published: (2025)
by: Zhang, Chong, et al.
Published: (2025)
Linguistically Conditioned Semantic Textual Similarity
by: Tu, Jingxuan, et al.
Published: (2024)
by: Tu, Jingxuan, et al.
Published: (2024)
Robust Vision-Language Models via Tensor Decomposition: A Defense Against Adversarial Attacks
by: Patel, Het, et al.
Published: (2025)
by: Patel, Het, et al.
Published: (2025)
UniGuardian: A Unified Defense for Detecting Prompt Injection, Backdoor Attacks and Adversarial Attacks in Large Language Models
by: Lin, Huawei, et al.
Published: (2025)
by: Lin, Huawei, et al.
Published: (2025)
RTD-Guard: A Black-Box Textual Adversarial Detection Framework via Replacement Token Detection
by: Zhu, He, et al.
Published: (2026)
by: Zhu, He, et al.
Published: (2026)
LANE: Lexical Adversarial Negative Examples for Word Sense Disambiguation
by: de Sá, Jader Martins Camboim, et al.
Published: (2025)
by: de Sá, Jader Martins Camboim, et al.
Published: (2025)
Arabic Synonym BERT-based Adversarial Examples for Text Classification
by: Alshahrani, Norah, et al.
Published: (2024)
by: Alshahrani, Norah, et al.
Published: (2024)
Adversarial Attack for Explanation Robustness of Rationalization Models
by: Zhang, Yuankai, et al.
Published: (2024)
by: Zhang, Yuankai, et al.
Published: (2024)
Explainable Semantic Textual Similarity via Dissimilar Span Detection
by: Lozano, Diego Miguel, et al.
Published: (2026)
by: Lozano, Diego Miguel, et al.
Published: (2026)
Token-Level Precise Attack on RAG: Searching for the Best Alternatives to Mislead Generation
by: Li, Zizhong, et al.
Published: (2025)
by: Li, Zizhong, et al.
Published: (2025)
KurdSTS: The Kurdish Semantic Textual Similarity
by: Abdullah, Abdulhady Abas, et al.
Published: (2025)
by: Abdullah, Abdulhady Abas, et al.
Published: (2025)
Similar Items
-
Defense Against Syntactic Textual Backdoor Attacks with Token Substitution
by: Li, Xinglin, et al.
Published: (2024) -
Fast Adversarial Training against Textual Adversarial Attacks
by: Yang, Yichen, et al.
Published: (2024) -
LimeAttack: Local Explainable Method for Textual Hard-Label Adversarial Attack
by: Zhu, Hai, et al.
Published: (2023) -
Adversarial Attacks and Defense for Conversation Entailment Task
by: Yang, Zhenning, et al.
Published: (2024) -
Are ELECTRA's Sentence Embeddings Beyond Repair? The Case of Semantic Textual Similarity
by: Rep, Ivan, et al.
Published: (2024)