Gespeichert in:
| Hauptverfasser: | Alexiou, Michail S., Mertoguno, J. Sukarno |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.09343 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Method for Fast Autonomy Transfer in Reinforcement Learning
von: Sahabandu, Dinuka, et al.
Veröffentlicht: (2024)
von: Sahabandu, Dinuka, et al.
Veröffentlicht: (2024)
Brain Tumor Classifiers Under Attack: Robustness of ResNet Variants Against Transferable FGSM and PGD Attacks
von: Deem, Ryan, et al.
Veröffentlicht: (2026)
von: Deem, Ryan, et al.
Veröffentlicht: (2026)
Adversarial Tuning: Defending Against Jailbreak Attacks for LLMs
von: Liu, Fan, et al.
Veröffentlicht: (2024)
von: Liu, Fan, et al.
Veröffentlicht: (2024)
Adversarial Attacks Against Automated Fact-Checking: A Survey
von: Liu, Fanzhen, et al.
Veröffentlicht: (2025)
von: Liu, Fanzhen, et al.
Veröffentlicht: (2025)
HQA-Attack: Toward High Quality Black-Box Hard-Label Adversarial Attack on Text
von: Liu, Han, et al.
Veröffentlicht: (2024)
von: Liu, Han, et al.
Veröffentlicht: (2024)
Kyrtos: A methodology for automatic deep analysis of graphic charts with curves in technical documents
von: Alexiou, Michail S., et al.
Veröffentlicht: (2026)
von: Alexiou, Michail S., et al.
Veröffentlicht: (2026)
Chain Association-based Attacking and Shielding Natural Language Processing Systems
von: Huang, Jiacheng, et al.
Veröffentlicht: (2024)
von: Huang, Jiacheng, et al.
Veröffentlicht: (2024)
Fast Adversarial Training against Textual Adversarial Attacks
von: Yang, Yichen, et al.
Veröffentlicht: (2024)
von: Yang, Yichen, et al.
Veröffentlicht: (2024)
API-BLEND: A Comprehensive Corpora for Training and Benchmarking API LLMs
von: Basu, Kinjal, et al.
Veröffentlicht: (2024)
von: Basu, Kinjal, et al.
Veröffentlicht: (2024)
A Statistical and Multi-Perspective Revisiting of the Membership Inference Attack in Large Language Models
von: Chen, Bowen, et al.
Veröffentlicht: (2024)
von: Chen, Bowen, et al.
Veröffentlicht: (2024)
Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
von: Wen, Yuchen, et al.
Veröffentlicht: (2024)
von: Wen, Yuchen, et al.
Veröffentlicht: (2024)
STACK: Adversarial Attacks on LLM Safeguard Pipelines
von: McKenzie, Ian R., et al.
Veröffentlicht: (2025)
von: McKenzie, Ian R., et al.
Veröffentlicht: (2025)
Towards Analyzing and Understanding the Limitations of DPO: A Theoretical Perspective
von: Feng, Duanyu, et al.
Veröffentlicht: (2024)
von: Feng, Duanyu, et al.
Veröffentlicht: (2024)
Towards Trustworthy Knowledge Graph Reasoning: An Uncertainty Aware Perspective
von: Ni, Bo, et al.
Veröffentlicht: (2024)
von: Ni, Bo, et al.
Veröffentlicht: (2024)
An Adversarial Perspective on Machine Unlearning for AI Safety
von: Łucki, Jakub, et al.
Veröffentlicht: (2024)
von: Łucki, Jakub, et al.
Veröffentlicht: (2024)
Robust Vision-Language Models via Tensor Decomposition: A Defense Against Adversarial Attacks
von: Patel, Het, et al.
Veröffentlicht: (2025)
von: Patel, Het, et al.
Veröffentlicht: (2025)
Green Shielding: A User-Centric Approach Towards Trustworthy AI
von: Li, Aaron J., et al.
Veröffentlicht: (2026)
von: Li, Aaron J., et al.
Veröffentlicht: (2026)
X-Transfer Attacks: Towards Super Transferable Adversarial Attacks on CLIP
von: Huang, Hanxun, et al.
Veröffentlicht: (2025)
von: Huang, Hanxun, et al.
Veröffentlicht: (2025)
Towards Cross-lingual Values Judgment: A Consensus-Pluralism Perspective
von: Chen, Yukun, et al.
Veröffentlicht: (2026)
von: Chen, Yukun, et al.
Veröffentlicht: (2026)
Perspective Dial: Measuring Perspective of Text and Guiding LLM Outputs
von: Kim, Taejin, et al.
Veröffentlicht: (2025)
von: Kim, Taejin, et al.
Veröffentlicht: (2025)
Combating Adversarial Attacks with Multi-Agent Debate
von: Chern, Steffi, et al.
Veröffentlicht: (2024)
von: Chern, Steffi, et al.
Veröffentlicht: (2024)
Adversarial Attacks and Defense for Conversation Entailment Task
von: Yang, Zhenning, et al.
Veröffentlicht: (2024)
von: Yang, Zhenning, et al.
Veröffentlicht: (2024)
Merging Improves Self-Critique Against Jailbreak Attacks
von: Gallego, Victor
Veröffentlicht: (2024)
von: Gallego, Victor
Veröffentlicht: (2024)
GoEX: Perspectives and Designs Towards a Runtime for Autonomous LLM Applications
von: Patil, Shishir G., et al.
Veröffentlicht: (2024)
von: Patil, Shishir G., et al.
Veröffentlicht: (2024)
SEP-Attack: A Simple and Effective Paradigm for Transfer-Based Textual Adversarial Attack
von: Liu, Han, et al.
Veröffentlicht: (2026)
von: Liu, Han, et al.
Veröffentlicht: (2026)
ShieldLearner: A New Paradigm for Jailbreak Attack Defense in LLMs
von: Ni, Ziyi, et al.
Veröffentlicht: (2025)
von: Ni, Ziyi, et al.
Veröffentlicht: (2025)
A Generative Adversarial Attack for Multilingual Text Classifiers
von: Roth, Tom, et al.
Veröffentlicht: (2024)
von: Roth, Tom, et al.
Veröffentlicht: (2024)
Dual-Modality Multi-Stage Adversarial Safety Training: Robustifying Multimodal Web Agents Against Cross-Modal Attacks
von: Liu, Haoyu, et al.
Veröffentlicht: (2026)
von: Liu, Haoyu, et al.
Veröffentlicht: (2026)
Temperature Matters: Enhancing Watermark Robustness Against Paraphrasing Attacks
von: Idrissi, Badr Youbi, et al.
Veröffentlicht: (2025)
von: Idrissi, Badr Youbi, et al.
Veröffentlicht: (2025)
Perspectives in Play: A Multi-Perspective Approach for More Inclusive NLP Systems
von: Muscato, Benedetta, et al.
Veröffentlicht: (2025)
von: Muscato, Benedetta, et al.
Veröffentlicht: (2025)
Robust Neural Information Retrieval: An Adversarial and Out-of-distribution Perspective
von: Liu, Yu-An, et al.
Veröffentlicht: (2024)
von: Liu, Yu-An, et al.
Veröffentlicht: (2024)
LLM-Generated Negative News Headlines Dataset: Creation and Benchmarking Against Real Journalism
von: Babalola, Olusola, et al.
Veröffentlicht: (2025)
von: Babalola, Olusola, et al.
Veröffentlicht: (2025)
FraudShield: Knowledge Graph Empowered Defense for LLMs against Fraud Attacks
von: Xu, Naen, et al.
Veröffentlicht: (2026)
von: Xu, Naen, et al.
Veröffentlicht: (2026)
API Pack: A Massive Multi-Programming Language Dataset for API Call Generation
von: Guo, Zhen, et al.
Veröffentlicht: (2024)
von: Guo, Zhen, et al.
Veröffentlicht: (2024)
Towards Adaptive, Scalable, and Robust Coordination of LLM Agents: A Dynamic Ad-Hoc Networking Perspective
von: Li, Rui, et al.
Veröffentlicht: (2026)
von: Li, Rui, et al.
Veröffentlicht: (2026)
Towards a More Inclusive AI: Progress and Perspectives in Large Language Model Training for the Sámi Language
von: Paul, Ronny, et al.
Veröffentlicht: (2024)
von: Paul, Ronny, et al.
Veröffentlicht: (2024)
Robustness of Prompting: Enhancing Robustness of Large Language Models Against Prompting Attacks
von: Mu, Lin, et al.
Veröffentlicht: (2025)
von: Mu, Lin, et al.
Veröffentlicht: (2025)
Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language Models
von: Yi, Jingwei, et al.
Veröffentlicht: (2023)
von: Yi, Jingwei, et al.
Veröffentlicht: (2023)
An Audit on the Perspectives and Challenges of Hallucinations in NLP
von: Venkit, Pranav Narayanan, et al.
Veröffentlicht: (2024)
von: Venkit, Pranav Narayanan, et al.
Veröffentlicht: (2024)
Large Knowledge Model: Perspectives and Challenges
von: Chen, Huajun
Veröffentlicht: (2023)
von: Chen, Huajun
Veröffentlicht: (2023)
Ähnliche Einträge
-
A Method for Fast Autonomy Transfer in Reinforcement Learning
von: Sahabandu, Dinuka, et al.
Veröffentlicht: (2024) -
Brain Tumor Classifiers Under Attack: Robustness of ResNet Variants Against Transferable FGSM and PGD Attacks
von: Deem, Ryan, et al.
Veröffentlicht: (2026) -
Adversarial Tuning: Defending Against Jailbreak Attacks for LLMs
von: Liu, Fan, et al.
Veröffentlicht: (2024) -
Adversarial Attacks Against Automated Fact-Checking: A Survey
von: Liu, Fanzhen, et al.
Veröffentlicht: (2025) -
HQA-Attack: Toward High Quality Black-Box Hard-Label Adversarial Attack on Text
von: Liu, Han, et al.
Veröffentlicht: (2024)