Enhancing Adversarial Text Attacks on BERT Models with Projected Gradient Descent
Fuente:
arXiv
Guardado en:
| Autores principales: | Waghela, Hetvi, Sen, Jaydip, Rakshit, Sneha |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Modified Word Saliency-Based Adversarial Attack on Text Classification Models
por: Waghela, Hetvi, et al.
Publicado: (2024)
por: Waghela, Hetvi, et al.
Publicado: (2024)
Saliency Attention and Semantic Similarity-Driven Adversarial Perturbation
por: Waghela, Hetvi, et al.
Publicado: (2024)
por: Waghela, Hetvi, et al.
Publicado: (2024)
Adversarial Robustness through Dynamic Ensemble Learning
por: Waghela, Hetvi, et al.
Publicado: (2024)
por: Waghela, Hetvi, et al.
Publicado: (2024)
Adversarial Text Generation with Dynamic Contextual Perturbation
por: Waghela, Hetvi, et al.
Publicado: (2025)
por: Waghela, Hetvi, et al.
Publicado: (2025)
Robust Image Classification: Defensive Strategies against FGSM and PGD Adversarial Attacks
por: Waghela, Hetvi, et al.
Publicado: (2024)
por: Waghela, Hetvi, et al.
Publicado: (2024)
Privacy in Federated Learning
por: Sen, Jaydip, et al.
Publicado: (2024)
por: Sen, Jaydip, et al.
Publicado: (2024)
Exploring Sectoral Profitability in the Indian Stock Market Using Deep Learning
por: Sen, Jaydip, et al.
Publicado: (2024)
por: Sen, Jaydip, et al.
Publicado: (2024)
Context-Enhanced Contrastive Search for Improved LLM Text Generation
por: Sen, Jaydip, et al.
Publicado: (2025)
por: Sen, Jaydip, et al.
Publicado: (2025)
Semantic Stealth: Adversarial Text Attacks on NLP Using Several Methods
por: Dey, Roopkatha, et al.
Publicado: (2024)
por: Dey, Roopkatha, et al.
Publicado: (2024)
Adversarial Attack on Large Language Models using Exponentiated Gradient Descent
por: Biswas, Sajib, et al.
Publicado: (2025)
por: Biswas, Sajib, et al.
Publicado: (2025)
Confidence-Modulated Speculative Decoding for Large Language Models
por: Sen, Jaydip, et al.
Publicado: (2025)
por: Sen, Jaydip, et al.
Publicado: (2025)
Humanizing Machine-Generated Content: Evading AI-Text Detection through Adversarial Attack
por: Zhou, Ying, et al.
Publicado: (2024)
por: Zhou, Ying, et al.
Publicado: (2024)
Text-CRS: A Generalized Certified Robustness Framework against Textual Adversarial Attacks
por: Zhang, Xinyu, et al.
Publicado: (2023)
por: Zhang, Xinyu, et al.
Publicado: (2023)
Deciphering the Chaos: Enhancing Jailbreak Attacks via Adversarial Prompt Translation
por: Li, Qizhang, et al.
Publicado: (2024)
por: Li, Qizhang, et al.
Publicado: (2024)
Multi-Amateur Contrastive Decoding for Text Generation
por: Sen, Jaydip, et al.
Publicado: (2025)
por: Sen, Jaydip, et al.
Publicado: (2025)
Graded Suspiciousness of Adversarial Texts to Human
por: Tonni, Shakila Mahjabin, et al.
Publicado: (2024)
por: Tonni, Shakila Mahjabin, et al.
Publicado: (2024)
IDT: Dual-Task Adversarial Attacks for Privacy Protection
por: Faustini, Pedro, et al.
Publicado: (2024)
por: Faustini, Pedro, et al.
Publicado: (2024)
Towards More Realistic Extraction Attacks: An Adversarial Perspective
por: More, Yash, et al.
Publicado: (2024)
por: More, Yash, et al.
Publicado: (2024)
MARAGE: Transferable Multi-Model Adversarial Attack for Retrieval-Augmented Generation Data Extraction
por: Hu, Xiao, et al.
Publicado: (2025)
por: Hu, Xiao, et al.
Publicado: (2025)
Adversarial Attacks on Parts of Speech: An Empirical Study in Text-to-Image Generation
por: Shahariar, G M, et al.
Publicado: (2024)
por: Shahariar, G M, et al.
Publicado: (2024)
On Evaluating The Performance of Watermarked Machine-Generated Texts Under Adversarial Attacks
por: Liu, Zesen, et al.
Publicado: (2024)
por: Liu, Zesen, et al.
Publicado: (2024)
Self-Evaluation as a Defense Against Adversarial Attacks on LLMs
por: Brown, Hannah, et al.
Publicado: (2024)
por: Brown, Hannah, et al.
Publicado: (2024)
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs
por: Li, Ran, et al.
Publicado: (2025)
por: Li, Ran, et al.
Publicado: (2025)
Token-Modification Adversarial Attacks for Natural Language Processing: A Survey
por: Roth, Tom, et al.
Publicado: (2021)
por: Roth, Tom, et al.
Publicado: (2021)
MPAT: Building Robust Deep Neural Networks against Textual Adversarial Attacks
por: Zhang, Fangyuan, et al.
Publicado: (2024)
por: Zhang, Fangyuan, et al.
Publicado: (2024)
Data Privacy Preservation on the Internet of Things
por: Sen, Jaydip, et al.
Publicado: (2023)
por: Sen, Jaydip, et al.
Publicado: (2023)
Transferable Embedding Inversion Attack: Uncovering Privacy Risks in Text Embeddings without Model Queries
por: Huang, Yu-Hsiang, et al.
Publicado: (2024)
por: Huang, Yu-Hsiang, et al.
Publicado: (2024)
The Hidden Cost of Modeling P(X): Vulnerability to Membership Inference Attacks in Generative Text Classifiers
por: Makroo, Owais, et al.
Publicado: (2025)
por: Makroo, Owais, et al.
Publicado: (2025)
Bits Leaked per Query: Information-Theoretic Bounds on Adversarial Attacks against LLMs
por: Kaneko, Masahiro, et al.
Publicado: (2025)
por: Kaneko, Masahiro, et al.
Publicado: (2025)
The Resurgence of GCG Adversarial Attacks on Large Language Models
por: Tan, Yuting, et al.
Publicado: (2025)
por: Tan, Yuting, et al.
Publicado: (2025)
ThreatCrawl: A BERT-based Focused Crawler for the Cybersecurity Domain
por: Kuehn, Philipp, et al.
Publicado: (2023)
por: Kuehn, Philipp, et al.
Publicado: (2023)
Membership Inference Attacks and Privacy in Topic Modeling
por: Manzonelli, Nico, et al.
Publicado: (2024)
por: Manzonelli, Nico, et al.
Publicado: (2024)
User Inference Attacks on Large Language Models
por: Kandpal, Nikhil, et al.
Publicado: (2023)
por: Kandpal, Nikhil, et al.
Publicado: (2023)
Adversarial Decoding: Generating Readable Documents for Adversarial Objectives
por: Zhang, Collin, et al.
Publicado: (2024)
por: Zhang, Collin, et al.
Publicado: (2024)
Adversarial Text Purification: A Large Language Model Approach for Defense
por: Moraffah, Raha, et al.
Publicado: (2024)
por: Moraffah, Raha, et al.
Publicado: (2024)
Quantum-Enhanced Adversarial Robustness in Artificial Intelligence
por: Sen, Jaydip
Publicado: (2026)
por: Sen, Jaydip
Publicado: (2026)
Composite Backdoor Attacks Against Large Language Models
por: Huang, Hai, et al.
Publicado: (2023)
por: Huang, Hai, et al.
Publicado: (2023)
Hijacking Large Language Models via Adversarial In-Context Learning
por: Zhou, Xiangyu, et al.
Publicado: (2023)
por: Zhou, Xiangyu, et al.
Publicado: (2023)
Evaluating Adversarial Robustness: A Comparison Of FGSM, Carlini-Wagner Attacks, And The Role of Distillation as Defense Mechanism
por: Sarkar, Trilokesh Ranjan, et al.
Publicado: (2024)
por: Sarkar, Trilokesh Ranjan, et al.
Publicado: (2024)
Advancing Adversarial Suffix Transfer Learning on Aligned Large Language Models
por: Liu, Hongfu, et al.
Publicado: (2024)
por: Liu, Hongfu, et al.
Publicado: (2024)
Ejemplares similares
-
A Modified Word Saliency-Based Adversarial Attack on Text Classification Models
por: Waghela, Hetvi, et al.
Publicado: (2024) -
Saliency Attention and Semantic Similarity-Driven Adversarial Perturbation
por: Waghela, Hetvi, et al.
Publicado: (2024) -
Adversarial Robustness through Dynamic Ensemble Learning
por: Waghela, Hetvi, et al.
Publicado: (2024) -
Adversarial Text Generation with Dynamic Contextual Perturbation
por: Waghela, Hetvi, et al.
Publicado: (2025) -
Robust Image Classification: Defensive Strategies against FGSM and PGD Adversarial Attacks
por: Waghela, Hetvi, et al.
Publicado: (2024)