Enhance Robustness of Language Models Against Variation Attack through Graph Integration
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiong, Zi, Qing, Lizhi, Kang, Yangyang, Liu, Jiawei, Li, Hongsong, Sun, Changlong, Liu, Xiaozhong, Lu, Wei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Black-Box Opinion Manipulation Attacks to Retrieval-Augmented Generation of Large Language Models
por: Chen, Zhuo, et al.
Publicado: (2024)
por: Chen, Zhuo, et al.
Publicado: (2024)
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks
por: Yin, Ziyi, et al.
Publicado: (2025)
por: Yin, Ziyi, et al.
Publicado: (2025)
Membership Inference Attacks Against Video Large Language Models
por: Song, Wei, et al.
Publicado: (2026)
por: Song, Wei, et al.
Publicado: (2026)
Marlin: Knowledge-Driven Analysis of Provenance Graphs for Efficient and Robust Detection of Cyber Attacks
por: Li, Zhenyuan, et al.
Publicado: (2024)
por: Li, Zhenyuan, et al.
Publicado: (2024)
MirGuard: Towards a Robust Provenance-based Intrusion Detection System Against Graph Manipulation Attacks
por: Sang, Anyuan, et al.
Publicado: (2025)
por: Sang, Anyuan, et al.
Publicado: (2025)
Tit-for-Tat: Safeguarding Large Vision-Language Models Against Jailbreak Attacks via Adversarial Defense
por: Hao, Shuyang, et al.
Publicado: (2025)
por: Hao, Shuyang, et al.
Publicado: (2025)
Real is not True: Backdoor Attacks Against Deepfake Detection
por: Sun, Hong, et al.
Publicado: (2024)
por: Sun, Hong, et al.
Publicado: (2024)
From LLMs to MLLMs to Agents: A Survey of Emerging Paradigms in Jailbreak Attacks and Defenses within LLM Ecosystem
por: Mao, Yanxu, et al.
Publicado: (2025)
por: Mao, Yanxu, et al.
Publicado: (2025)
Cuckoo Attack: Stealthy and Persistent Attacks Against AI-IDE
por: Liu, Xinpeng, et al.
Publicado: (2025)
por: Liu, Xinpeng, et al.
Publicado: (2025)
Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models
por: Gong, Yuyang, et al.
Publicado: (2025)
por: Gong, Yuyang, et al.
Publicado: (2025)
RobustMask: Certified Robustness against Adversarial Neural Ranking Attack via Randomized Masking
por: Liu, Jiawei, et al.
Publicado: (2025)
por: Liu, Jiawei, et al.
Publicado: (2025)
IndirectAD: Practical Data Poisoning Attacks against Recommender Systems for Item Promotion
por: Wang, Zihao, et al.
Publicado: (2025)
por: Wang, Zihao, et al.
Publicado: (2025)
Activation Gradient based Poisoned Sample Detection Against Backdoor Attacks
por: Yuan, Danni, et al.
Publicado: (2023)
por: Yuan, Danni, et al.
Publicado: (2023)
EM-MIAs: Enhancing Membership Inference Attacks in Large Language Models through Ensemble Modeling
por: Song, Zichen, et al.
Publicado: (2024)
por: Song, Zichen, et al.
Publicado: (2024)
PersonaMark: Personalized LLM watermarking for model protection and user attribution
por: Zhang, Yuehan, et al.
Publicado: (2024)
por: Zhang, Yuehan, et al.
Publicado: (2024)
TED-LaST: Towards Robust Backdoor Defense Against Adaptive Attacks
por: Mo, Xiaoxing, et al.
Publicado: (2025)
por: Mo, Xiaoxing, et al.
Publicado: (2025)
Provable Robustness of (Graph) Neural Networks Against Data Poisoning and Backdoor Attacks
por: Gosch, Lukas, et al.
Publicado: (2024)
por: Gosch, Lukas, et al.
Publicado: (2024)
PEFTGuard: Detecting Backdoor Attacks Against Parameter-Efficient Fine-Tuning
por: Sun, Zhen, et al.
Publicado: (2024)
por: Sun, Zhen, et al.
Publicado: (2024)
Backdoor Attacks and Countermeasures in Natural Language Processing Models: A Comprehensive Security Review
por: Cheng, Pengzhou, et al.
Publicado: (2023)
por: Cheng, Pengzhou, et al.
Publicado: (2023)
Membership Inference Attacks Against Vision-Language Models
por: Hu, Yuke, et al.
Publicado: (2025)
por: Hu, Yuke, et al.
Publicado: (2025)
Detection and Defense Against Prominent Attacks on Preconditioned LLM-Integrated Virtual Assistants
por: Chan, Chun Fai, et al.
Publicado: (2024)
por: Chan, Chun Fai, et al.
Publicado: (2024)
CRUcialG: Reconstruct Integrated Attack Scenario Graphs by Cyber Threat Intelligence Reports
por: Cheng, Wenrui, et al.
Publicado: (2024)
por: Cheng, Wenrui, et al.
Publicado: (2024)
Node Injection Attack Based on Label Propagation Against Graph Neural Network
por: Zhu, Peican, et al.
Publicado: (2024)
por: Zhu, Peican, et al.
Publicado: (2024)
DiscourseFlip: An Oblique Discourse-Level Opinion Manipulation Attack against Black-box Retrieval-Augmented Generation
por: Gong, Yuyang, et al.
Publicado: (2026)
por: Gong, Yuyang, et al.
Publicado: (2026)
Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models
por: Chen, Zeyuan, et al.
Publicado: (2026)
por: Chen, Zeyuan, et al.
Publicado: (2026)
Virtual Context: Enhancing Jailbreak Attacks with Special Token Injection
por: Zhou, Yuqi, et al.
Publicado: (2024)
por: Zhou, Yuqi, et al.
Publicado: (2024)
DiffAttack: Evasion Attacks Against Diffusion-Based Adversarial Purification
por: Kang, Mintong, et al.
Publicado: (2023)
por: Kang, Mintong, et al.
Publicado: (2023)
Leakage-abuse Attack Against Substring-SSE with Partially Known Dataset
por: Ba, Xijie, et al.
Publicado: (2025)
por: Ba, Xijie, et al.
Publicado: (2025)
Robustness of Locally Differentially Private Graph Analysis Against Poisoning
por: Imola, Jacob, et al.
Publicado: (2022)
por: Imola, Jacob, et al.
Publicado: (2022)
Rotated Robustness: A Training-Free Defense against Bit-Flip Attacks on Large Language Models
por: Liu, Deng, et al.
Publicado: (2026)
por: Liu, Deng, et al.
Publicado: (2026)
SAGE: Sample-Aware Guarding Engine for Robust Intrusion Detection Against Adversarial Attacks
por: Chen, Jing, et al.
Publicado: (2025)
por: Chen, Jing, et al.
Publicado: (2025)
E-SAGE: Explainability-based Defense Against Backdoor Attacks on Graph Neural Networks
por: Yuan, Dingqiang, et al.
Publicado: (2024)
por: Yuan, Dingqiang, et al.
Publicado: (2024)
Learning-based Privacy-Preserving Graph Publishing Against Sensitive Link Inference Attacks
por: Wu, Yucheng, et al.
Publicado: (2025)
por: Wu, Yucheng, et al.
Publicado: (2025)
Strengthening Polymorphic Prompt Assembling: Dynamic Separator Generation Against Emerging Prompt Injection Attacks
por: Dorzhiev, Nima, et al.
Publicado: (2026)
por: Dorzhiev, Nima, et al.
Publicado: (2026)
Provably Robust Explainable Graph Neural Networks against Graph Perturbation Attacks
por: Li, Jiate, et al.
Publicado: (2025)
por: Li, Jiate, et al.
Publicado: (2025)
Detecting Complex Multi-step Attacks with Explainable Graph Neural Network
por: Liu, Wei, et al.
Publicado: (2024)
por: Liu, Wei, et al.
Publicado: (2024)
LLM-Enhanced Software Patch Localization
por: Yu, Jinhong, et al.
Publicado: (2024)
por: Yu, Jinhong, et al.
Publicado: (2024)
AttackEval: A Systematic Empirical Study of Prompt Injection Attack Effectiveness Against Large Language Models
por: Wang, Jackson
Publicado: (2026)
por: Wang, Jackson
Publicado: (2026)
Defense Against Prompt Injection Attack by Leveraging Attack Techniques
por: Chen, Yulin, et al.
Publicado: (2024)
por: Chen, Yulin, et al.
Publicado: (2024)
UAV Resilience Against Stealthy Attacks
por: Amorim, Arthur, et al.
Publicado: (2025)
por: Amorim, Arthur, et al.
Publicado: (2025)
Ejemplares similares
-
Black-Box Opinion Manipulation Attacks to Retrieval-Augmented Generation of Large Language Models
por: Chen, Zhuo, et al.
Publicado: (2024) -
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks
por: Yin, Ziyi, et al.
Publicado: (2025) -
Membership Inference Attacks Against Video Large Language Models
por: Song, Wei, et al.
Publicado: (2026) -
Marlin: Knowledge-Driven Analysis of Provenance Graphs for Efficient and Robust Detection of Cyber Attacks
por: Li, Zhenyuan, et al.
Publicado: (2024) -
MirGuard: Towards a Robust Provenance-based Intrusion Detection System Against Graph Manipulation Attacks
por: Sang, Anyuan, et al.
Publicado: (2025)