Advancing NLP Security by Leveraging LLMs as Adversarial Engines
Fuente:
arXiv
Saved in:
| Main Authors: | Srinivasan, Sudarshan, Mahbub, Maria, Sadovnik, Amir |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Redefining "Hallucination" in LLMs: Towards a psychology-informed framework for mitigating misinformation
by: Berberette, Elijah, et al.
Published: (2024)
by: Berberette, Elijah, et al.
Published: (2024)
Leveraging Large Language Models to Extract Information on Substance Use Disorder Severity from Clinical Notes: A Zero-shot Learning Approach
by: Mahbub, Maria, et al.
Published: (2024)
by: Mahbub, Maria, et al.
Published: (2024)
NLP Security and Ethics, in the Wild
by: Lent, Heather, et al.
Published: (2025)
by: Lent, Heather, et al.
Published: (2025)
Faithfulness and the Notion of Adversarial Sensitivity in NLP Explanations
by: Manna, Supriya, et al.
Published: (2024)
by: Manna, Supriya, et al.
Published: (2024)
Hiding-in-Plain-Sight (HiPS) Attack on CLIP for Targetted Object Removal from Images
by: Daw, Arka, et al.
Published: (2024)
by: Daw, Arka, et al.
Published: (2024)
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
by: Jin, Bowen, et al.
Published: (2025)
by: Jin, Bowen, et al.
Published: (2025)
On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs
by: Calderon, Nitay, et al.
Published: (2024)
by: Calderon, Nitay, et al.
Published: (2024)
Ground Truth Generation for Multilingual Historical NLP using LLMs
by: Gladstone, Clovis, et al.
Published: (2025)
by: Gladstone, Clovis, et al.
Published: (2025)
Beyond Weaponization: NLP Security for Medium and Lower-Resourced Languages in Their Own Right
by: Lent, Heather
Published: (2025)
by: Lent, Heather
Published: (2025)
Mitigating Self-Preference by Authorship Obfuscation
by: Mahbub, Taslim, et al.
Published: (2025)
by: Mahbub, Taslim, et al.
Published: (2025)
From Transformers to LLMs: A Systematic Survey of Efficiency Considerations in NLP
by: Ansar, Wazib, et al.
Published: (2024)
by: Ansar, Wazib, et al.
Published: (2024)
How LLMs Fail to Support Fact-Checking
by: Proma, Adiba Mahbub, et al.
Published: (2025)
by: Proma, Adiba Mahbub, et al.
Published: (2025)
Graphusion: Leveraging Large Language Models for Scientific Knowledge Graph Fusion and Construction in NLP Education
by: Yang, Rui, et al.
Published: (2024)
by: Yang, Rui, et al.
Published: (2024)
Empowering Dysarthric Speech: Leveraging Advanced LLMs for Accurate Speech Correction and Multimodal Emotion Analysis
by: Attaluri, Kaushal, et al.
Published: (2024)
by: Attaluri, Kaushal, et al.
Published: (2024)
Machine-Assisted Grading of Nationwide School-Leaving Essay Exams with LLMs and Statistical NLP
by: Karjus, Andres, et al.
Published: (2026)
by: Karjus, Andres, et al.
Published: (2026)
The African Languages Lab: A Collaborative Approach to Advancing Low-Resource African NLP
by: Issaka, Sheriff, et al.
Published: (2025)
by: Issaka, Sheriff, et al.
Published: (2025)
Dynamic Q&A of Clinical Documents with Large Language Models
by: Elgedawy, Ran, et al.
Published: (2024)
by: Elgedawy, Ran, et al.
Published: (2024)
The Pitfalls of Publishing in the Age of LLMs: Strange and Surprising Adventures with a High-Impact NLP Journal
by: Verma, Rakesh M., et al.
Published: (2024)
by: Verma, Rakesh M., et al.
Published: (2024)
An Advanced NLP Framework for Automated Medical Diagnosis with DeBERTa and Dynamic Contextual Positional Gating
by: Khaniki, Mohammad Ali Labbaf, et al.
Published: (2025)
by: Khaniki, Mohammad Ali Labbaf, et al.
Published: (2025)
Comparative Performance of Advanced NLP Models and LLMs in Multilingual Geo-Entity Detection
by: Kopanov, Kalin
Published: (2024)
by: Kopanov, Kalin
Published: (2024)
eSapiens's DEREK Module: Deep Extraction & Reasoning Engine for Knowledge with LLMs
by: Shi, Isaac, et al.
Published: (2025)
by: Shi, Isaac, et al.
Published: (2025)
From Text to Graph: Leveraging Graph Neural Networks for Enhanced Explainability in NLP
by: Yáñez-Romero, Fabio, et al.
Published: (2025)
by: Yáñez-Romero, Fabio, et al.
Published: (2025)
Evaluating Robustness of Generative Search Engine on Adversarial Factual Questions
by: Hu, Xuming, et al.
Published: (2024)
by: Hu, Xuming, et al.
Published: (2024)
Advancing NLP Models with Strategic Text Augmentation: A Comprehensive Study of Augmentation Methods and Curriculum Strategies
by: Kesgin, Himmet Toprak, et al.
Published: (2024)
by: Kesgin, Himmet Toprak, et al.
Published: (2024)
Reasoning Robustness of LLMs to Adversarial Typographical Errors
by: Gan, Esther, et al.
Published: (2024)
by: Gan, Esther, et al.
Published: (2024)
Are You Human? An Adversarial Benchmark to Expose LLMs
by: Gressel, Gilad, et al.
Published: (2024)
by: Gressel, Gilad, et al.
Published: (2024)
Select, Label, Evaluate: Active Testing in NLP
by: Purificato, Antonio, et al.
Published: (2026)
by: Purificato, Antonio, et al.
Published: (2026)
A Multi-Stage Validation Framework for Trustworthy Large-scale Clinical Information Extraction using Large Language Models
by: Mahbub, Maria, et al.
Published: (2026)
by: Mahbub, Maria, et al.
Published: (2026)
Leveraging ChatGPT and Other NLP Methods for Identifying Risk and Protective Behaviors in MSM: Social Media and Dating apps Text Analysis
by: Beikzadeh, Mehrab, et al.
Published: (2026)
by: Beikzadeh, Mehrab, et al.
Published: (2026)
Evaluating Deduplication Techniques for Economic Research Paper Titles with a Focus on Semantic Similarity using NLP and LLMs
by: You, Doohee, et al.
Published: (2024)
by: You, Doohee, et al.
Published: (2024)
Emoti-Attack: Zero-Perturbation Adversarial Attacks on NLP Systems via Emoji Sequences
by: Zhang, Yangshijie
Published: (2025)
by: Zhang, Yangshijie
Published: (2025)
Advancing Prompt Recovery in NLP: A Deep Dive into the Integration of Gemma-2b-it and Phi2 Models
by: Chen, Jianlong, et al.
Published: (2024)
by: Chen, Jianlong, et al.
Published: (2024)
Estimating Causal Effects of Text Interventions Leveraging LLMs
by: Guo, Siyi, et al.
Published: (2024)
by: Guo, Siyi, et al.
Published: (2024)
CorefInst: Leveraging LLMs for Multilingual Coreference Resolution
by: Arslan, Tuğba Pamay, et al.
Published: (2025)
by: Arslan, Tuğba Pamay, et al.
Published: (2025)
Leveraging KV Similarity for Online Structured Pruning in LLMs
by: Lee, Jungmin, et al.
Published: (2025)
by: Lee, Jungmin, et al.
Published: (2025)
Leveraging Natural Language Processing to Unravel the Mystery of Life: A Review of NLP Approaches in Genomics, Transcriptomics, and Proteomics
by: Rannon, Ella, et al.
Published: (2025)
by: Rannon, Ella, et al.
Published: (2025)
Confidence Improves Self-Consistency in LLMs
by: Taubenfeld, Amir, et al.
Published: (2025)
by: Taubenfeld, Amir, et al.
Published: (2025)
Generative Adversarial Reviews: When LLMs Become the Critic
by: Bougie, Nicolas, et al.
Published: (2024)
by: Bougie, Nicolas, et al.
Published: (2024)
Adversarial versification in portuguese as a jailbreak operator in LLMs
by: Queiroz, Joao
Published: (2025)
by: Queiroz, Joao
Published: (2025)
Improving Fairness in LLMs Through Testing-Time Adversaries
by: Gregio, Isabela Pereira, et al.
Published: (2025)
by: Gregio, Isabela Pereira, et al.
Published: (2025)
Similar Items
-
Redefining "Hallucination" in LLMs: Towards a psychology-informed framework for mitigating misinformation
by: Berberette, Elijah, et al.
Published: (2024) -
Leveraging Large Language Models to Extract Information on Substance Use Disorder Severity from Clinical Notes: A Zero-shot Learning Approach
by: Mahbub, Maria, et al.
Published: (2024) -
NLP Security and Ethics, in the Wild
by: Lent, Heather, et al.
Published: (2025) -
Faithfulness and the Notion of Adversarial Sensitivity in NLP Explanations
by: Manna, Supriya, et al.
Published: (2024) -
Hiding-in-Plain-Sight (HiPS) Attack on CLIP for Targetted Object Removal from Images
by: Daw, Arka, et al.
Published: (2024)