Homograph Attacks on Maghreb Sentiment Analyzers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qachfar, Fatima Zahra, Verma, Rakesh M. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLMs can be Dangerous Reasoners: Analyzing-based Jailbreak Attack on Large Language Models
von: Lin, Shi, et al.
Veröffentlicht: (2024)
von: Lin, Shi, et al.
Veröffentlicht: (2024)
A Transfer Attack to Image Watermarks
von: Hu, Yuepeng, et al.
Veröffentlicht: (2024)
von: Hu, Yuepeng, et al.
Veröffentlicht: (2024)
Revisiting the Robustness of Watermarking to Paraphrasing Attacks
von: Rastogi, Saksham, et al.
Veröffentlicht: (2024)
von: Rastogi, Saksham, et al.
Veröffentlicht: (2024)
GCG Attack On A Diffusion LLM
von: Neyroud, Ruben, et al.
Veröffentlicht: (2025)
von: Neyroud, Ruben, et al.
Veröffentlicht: (2025)
Membership Inference Attacks and Privacy in Topic Modeling
von: Manzonelli, Nico, et al.
Veröffentlicht: (2024)
von: Manzonelli, Nico, et al.
Veröffentlicht: (2024)
User Inference Attacks on Large Language Models
von: Kandpal, Nikhil, et al.
Veröffentlicht: (2023)
von: Kandpal, Nikhil, et al.
Veröffentlicht: (2023)
Defending Against Indirect Prompt Injection Attacks With Spotlighting
von: Hines, Keegan, et al.
Veröffentlicht: (2024)
von: Hines, Keegan, et al.
Veröffentlicht: (2024)
Reversible Jump Attack to Textual Classifiers with Modification Reduction
von: Ni, Mingze, et al.
Veröffentlicht: (2024)
von: Ni, Mingze, et al.
Veröffentlicht: (2024)
Humanizing the Machine: Proxy Attacks to Mislead LLM Detectors
von: Wang, Tianchun, et al.
Veröffentlicht: (2024)
von: Wang, Tianchun, et al.
Veröffentlicht: (2024)
IDT: Dual-Task Adversarial Attacks for Privacy Protection
von: Faustini, Pedro, et al.
Veröffentlicht: (2024)
von: Faustini, Pedro, et al.
Veröffentlicht: (2024)
Towards More Realistic Extraction Attacks: An Adversarial Perspective
von: More, Yash, et al.
Veröffentlicht: (2024)
von: More, Yash, et al.
Veröffentlicht: (2024)
Composite Backdoor Attacks Against Large Language Models
von: Huang, Hai, et al.
Veröffentlicht: (2023)
von: Huang, Hai, et al.
Veröffentlicht: (2023)
An Early Categorization of Prompt Injection Attacks on Large Language Models
von: Rossi, Sippo, et al.
Veröffentlicht: (2024)
von: Rossi, Sippo, et al.
Veröffentlicht: (2024)
Self-Evaluation as a Defense Against Adversarial Attacks on LLMs
von: Brown, Hannah, et al.
Veröffentlicht: (2024)
von: Brown, Hannah, et al.
Veröffentlicht: (2024)
Phantom: General Backdoor Attacks on Retrieval Augmented Language Generation
von: Chaudhari, Harsh, et al.
Veröffentlicht: (2024)
von: Chaudhari, Harsh, et al.
Veröffentlicht: (2024)
Blind Baselines Beat Membership Inference Attacks for Foundation Models
von: Das, Debeshee, et al.
Veröffentlicht: (2024)
von: Das, Debeshee, et al.
Veröffentlicht: (2024)
TFL: Targeted Bit-Flip Attack on Large Language Model
von: Guo, Jingkai, et al.
Veröffentlicht: (2026)
von: Guo, Jingkai, et al.
Veröffentlicht: (2026)
Semantic Stealth: Adversarial Text Attacks on NLP Using Several Methods
von: Dey, Roopkatha, et al.
Veröffentlicht: (2024)
von: Dey, Roopkatha, et al.
Veröffentlicht: (2024)
Deciphering the Chaos: Enhancing Jailbreak Attacks via Adversarial Prompt Translation
von: Li, Qizhang, et al.
Veröffentlicht: (2024)
von: Li, Qizhang, et al.
Veröffentlicht: (2024)
AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
von: Zeng, Yifan, et al.
Veröffentlicht: (2024)
von: Zeng, Yifan, et al.
Veröffentlicht: (2024)
Enhancing Adversarial Text Attacks on BERT Models with Projected Gradient Descent
von: Waghela, Hetvi, et al.
Veröffentlicht: (2024)
von: Waghela, Hetvi, et al.
Veröffentlicht: (2024)
Adversarial Attack on Large Language Models using Exponentiated Gradient Descent
von: Biswas, Sajib, et al.
Veröffentlicht: (2025)
von: Biswas, Sajib, et al.
Veröffentlicht: (2025)
SteganoBackdoor: Stealthy and Data-Efficient Backdoor Attacks on Language Models
von: Xue, Eric, et al.
Veröffentlicht: (2025)
von: Xue, Eric, et al.
Veröffentlicht: (2025)
Token-Modification Adversarial Attacks for Natural Language Processing: A Survey
von: Roth, Tom, et al.
Veröffentlicht: (2021)
von: Roth, Tom, et al.
Veröffentlicht: (2021)
Steering Dialogue Dynamics for Robustness against Multi-turn Jailbreaking Attacks
von: Hu, Hanjiang, et al.
Veröffentlicht: (2025)
von: Hu, Hanjiang, et al.
Veröffentlicht: (2025)
Cross-Entropy Attacks to Language Models via Rare Event Simulation
von: Ni, Mingze, et al.
Veröffentlicht: (2025)
von: Ni, Mingze, et al.
Veröffentlicht: (2025)
Confidence Elicitation: A New Attack Vector for Large Language Models
von: Formento, Brian, et al.
Veröffentlicht: (2025)
von: Formento, Brian, et al.
Veröffentlicht: (2025)
Securing Large Language Models (LLMs) from Prompt Injection Attacks
von: Suri, Omar Farooq Khan, et al.
Veröffentlicht: (2025)
von: Suri, Omar Farooq Khan, et al.
Veröffentlicht: (2025)
Graphene: Infrastructure Security Posture Analysis with AI-generated Attack Graphs
von: Jin, Xin, et al.
Veröffentlicht: (2023)
von: Jin, Xin, et al.
Veröffentlicht: (2023)
A Modified Word Saliency-Based Adversarial Attack on Text Classification Models
von: Waghela, Hetvi, et al.
Veröffentlicht: (2024)
von: Waghela, Hetvi, et al.
Veröffentlicht: (2024)
MPAT: Building Robust Deep Neural Networks against Textual Adversarial Attacks
von: Zhang, Fangyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Fangyuan, et al.
Veröffentlicht: (2024)
SoK: Membership Inference Attacks on LLMs are Rushing Nowhere (and How to Fix It)
von: Meeus, Matthieu, et al.
Veröffentlicht: (2024)
von: Meeus, Matthieu, et al.
Veröffentlicht: (2024)
Chain of Attack: a Semantic-Driven Contextual Multi-Turn attacker for LLM
von: Yang, Xikang, et al.
Veröffentlicht: (2024)
von: Yang, Xikang, et al.
Veröffentlicht: (2024)
SOS! Soft Prompt Attack Against Open-Source Large Language Models
von: Yang, Ziqing, et al.
Veröffentlicht: (2024)
von: Yang, Ziqing, et al.
Veröffentlicht: (2024)
Break the Breakout: Reinventing LM Defense Against Jailbreak Attacks with Self-Refinement
von: Kim, Heegyu, et al.
Veröffentlicht: (2024)
von: Kim, Heegyu, et al.
Veröffentlicht: (2024)
SBFA: Single Sneaky Bit Flip Attack to Break Large Language Models
von: Guo, Jingkai, et al.
Veröffentlicht: (2025)
von: Guo, Jingkai, et al.
Veröffentlicht: (2025)
Humanizing Machine-Generated Content: Evading AI-Text Detection through Adversarial Attack
von: Zhou, Ying, et al.
Veröffentlicht: (2024)
von: Zhou, Ying, et al.
Veröffentlicht: (2024)
Cyber-Attack Technique Classification Using Two-Stage Trained Large Language Models
von: You, Weiqiu, et al.
Veröffentlicht: (2024)
von: You, Weiqiu, et al.
Veröffentlicht: (2024)
From Theft to Bomb-Making: The Ripple Effect of Unlearning in Defending Against Jailbreak Attacks
von: Zhang, Zhexin, et al.
Veröffentlicht: (2024)
von: Zhang, Zhexin, et al.
Veröffentlicht: (2024)
Bits Leaked per Query: Information-Theoretic Bounds on Adversarial Attacks against LLMs
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LLMs can be Dangerous Reasoners: Analyzing-based Jailbreak Attack on Large Language Models
von: Lin, Shi, et al.
Veröffentlicht: (2024) -
A Transfer Attack to Image Watermarks
von: Hu, Yuepeng, et al.
Veröffentlicht: (2024) -
Revisiting the Robustness of Watermarking to Paraphrasing Attacks
von: Rastogi, Saksham, et al.
Veröffentlicht: (2024) -
GCG Attack On A Diffusion LLM
von: Neyroud, Ruben, et al.
Veröffentlicht: (2025) -
Membership Inference Attacks and Privacy in Topic Modeling
von: Manzonelli, Nico, et al.
Veröffentlicht: (2024)