Noise Contrastive Estimation-based Matching Framework for Low-Resource Security Attack Pattern Recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | Nguyen, Tu, Šrndić, Nedim, Neth, Alexander |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Low-Resource Languages Jailbreak GPT-4
di: Yong, Zheng-Xin, et al.
Pubblicazione: (2023)
di: Yong, Zheng-Xin, et al.
Pubblicazione: (2023)
Does Low Rank Adaptation Lead to Lower Robustness against Training-Time Attacks?
di: Liang, Zi, et al.
Pubblicazione: (2025)
di: Liang, Zi, et al.
Pubblicazione: (2025)
Statement-Level Vulnerability Detection: Learning Vulnerability Patterns Through Information Theory and Contrastive Learning
di: Nguyen, Van, et al.
Pubblicazione: (2022)
di: Nguyen, Van, et al.
Pubblicazione: (2022)
Measuring Real-World Prompt Injection Attacks in LLM-based Resume Screening
di: Zhang, Mohan, et al.
Pubblicazione: (2026)
di: Zhang, Mohan, et al.
Pubblicazione: (2026)
MetaDefense: Defending Finetuning-based Jailbreak Attack Before and During Generation
di: Jiang, Weisen, et al.
Pubblicazione: (2025)
di: Jiang, Weisen, et al.
Pubblicazione: (2025)
Less Data, More Security: Advancing Cybersecurity LLMs Specialization via Resource-Efficient Domain-Adaptive Continuous Pre-training with Minimal Tokens
di: Salahuddin, Salahuddin, et al.
Pubblicazione: (2025)
di: Salahuddin, Salahuddin, et al.
Pubblicazione: (2025)
LLMs can be Dangerous Reasoners: Analyzing-based Jailbreak Attack on Large Language Models
di: Lin, Shi, et al.
Pubblicazione: (2024)
di: Lin, Shi, et al.
Pubblicazione: (2024)
Quantifying the Noise of Structural Perturbations on Graph Adversarial Attacks
di: Fang, Junyuan, et al.
Pubblicazione: (2025)
di: Fang, Junyuan, et al.
Pubblicazione: (2025)
Learning-Based Difficulty Calibration for Enhanced Membership Inference Attacks
di: Shi, Haonan, et al.
Pubblicazione: (2024)
di: Shi, Haonan, et al.
Pubblicazione: (2024)
An In-Depth Analysis of Cyber Attacks in Secured Platforms
di: Ozoh, Parick, et al.
Pubblicazione: (2025)
di: Ozoh, Parick, et al.
Pubblicazione: (2025)
Gandalf the Red: Adaptive Security for LLMs
di: Pfister, Niklas, et al.
Pubblicazione: (2025)
di: Pfister, Niklas, et al.
Pubblicazione: (2025)
Formalizing and Benchmarking Prompt Injection Attacks and Defenses
di: Liu, Yupei, et al.
Pubblicazione: (2023)
di: Liu, Yupei, et al.
Pubblicazione: (2023)
CLMIA: Membership Inference Attacks via Unsupervised Contrastive Learning
di: Chen, Depeng, et al.
Pubblicazione: (2024)
di: Chen, Depeng, et al.
Pubblicazione: (2024)
SecureCode: A Production-Grade Multi-Turn Dataset for Training Security-Aware Code Generation Models
di: Thornton, Scott
Pubblicazione: (2025)
di: Thornton, Scott
Pubblicazione: (2025)
Logicbreaks: A Framework for Understanding Subversion of Rule-based Inference
di: Xue, Anton, et al.
Pubblicazione: (2024)
di: Xue, Anton, et al.
Pubblicazione: (2024)
The Resurgence of GCG Adversarial Attacks on Large Language Models
di: Tan, Yuting, et al.
Pubblicazione: (2025)
di: Tan, Yuting, et al.
Pubblicazione: (2025)
Tree of Attacks: Jailbreaking Black-Box LLMs Automatically
di: Mehrotra, Anay, et al.
Pubblicazione: (2023)
di: Mehrotra, Anay, et al.
Pubblicazione: (2023)
Fine-Tuning Language Models with Differential Privacy through Adaptive Noise Allocation
di: Li, Xianzhi, et al.
Pubblicazione: (2024)
di: Li, Xianzhi, et al.
Pubblicazione: (2024)
Jailbreak Foundry: From Papers to Runnable Attacks for Reproducible Benchmarking
di: Fang, Zhicheng, et al.
Pubblicazione: (2026)
di: Fang, Zhicheng, et al.
Pubblicazione: (2026)
Exposing the Systematic Vulnerability of Open-Weight Models to Prefill Attacks
di: Struppek, Lukas, et al.
Pubblicazione: (2026)
di: Struppek, Lukas, et al.
Pubblicazione: (2026)
Best-of-Venom: Attacking RLHF by Injecting Poisoned Preference Data
di: Baumgärtner, Tim, et al.
Pubblicazione: (2024)
di: Baumgärtner, Tim, et al.
Pubblicazione: (2024)
SECA: Semantically Equivalent and Coherent Attacks for Eliciting LLM Hallucinations
di: Liang, Buyun, et al.
Pubblicazione: (2025)
di: Liang, Buyun, et al.
Pubblicazione: (2025)
Bypassing the Safety Training of Open-Source LLMs with Priming Attacks
di: Vega, Jason, et al.
Pubblicazione: (2023)
di: Vega, Jason, et al.
Pubblicazione: (2023)
JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
di: Chu, Junjie, et al.
Pubblicazione: (2024)
di: Chu, Junjie, et al.
Pubblicazione: (2024)
Enhancing Prompt Injection Attacks to LLMs via Poisoning Alignment
di: Shao, Zedian, et al.
Pubblicazione: (2024)
di: Shao, Zedian, et al.
Pubblicazione: (2024)
Exploiting Class Probabilities for Black-box Sentence-level Attacks
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
BadAgent: Inserting and Activating Backdoor Attacks in LLM Agents
di: Wang, Yifei, et al.
Pubblicazione: (2024)
di: Wang, Yifei, et al.
Pubblicazione: (2024)
HSF: Defending against Jailbreak Attacks with Hidden State Filtering
di: Qian, Cheng, et al.
Pubblicazione: (2024)
di: Qian, Cheng, et al.
Pubblicazione: (2024)
REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations
di: Liang, Buyun, et al.
Pubblicazione: (2026)
di: Liang, Buyun, et al.
Pubblicazione: (2026)
AutoBaxBuilder: Bootstrapping Code Security Benchmarking
di: von Arx, Tobias, et al.
Pubblicazione: (2025)
di: von Arx, Tobias, et al.
Pubblicazione: (2025)
SecEncoder: Logs are All You Need in Security
di: Bulut, Muhammed Fatih, et al.
Pubblicazione: (2024)
di: Bulut, Muhammed Fatih, et al.
Pubblicazione: (2024)
Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contrastive Scoring
di: Hua, Peichun, et al.
Pubblicazione: (2025)
di: Hua, Peichun, et al.
Pubblicazione: (2025)
False Data Injection Attack Detection in Edge-based Smart Metering Networks with Federated Learning
di: Uddin, Md Raihan, et al.
Pubblicazione: (2024)
di: Uddin, Md Raihan, et al.
Pubblicazione: (2024)
Towards Understanding the Fragility of Multilingual LLMs against Fine-Tuning Attacks
di: Poppi, Samuele, et al.
Pubblicazione: (2024)
di: Poppi, Samuele, et al.
Pubblicazione: (2024)
Adversarial Attacks on Parts of Speech: An Empirical Study in Text-to-Image Generation
di: Shahariar, G M, et al.
Pubblicazione: (2024)
di: Shahariar, G M, et al.
Pubblicazione: (2024)
Revealing Weaknesses in Text Watermarking Through Self-Information Rewrite Attacks
di: Cheng, Yixin, et al.
Pubblicazione: (2025)
di: Cheng, Yixin, et al.
Pubblicazione: (2025)
On Evaluating The Performance of Watermarked Machine-Generated Texts Under Adversarial Attacks
di: Liu, Zesen, et al.
Pubblicazione: (2024)
di: Liu, Zesen, et al.
Pubblicazione: (2024)
Scalable Defense against In-the-wild Jailbreaking Attacks with Safety Context Retrieval
di: Chen, Taiye, et al.
Pubblicazione: (2025)
di: Chen, Taiye, et al.
Pubblicazione: (2025)
EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage
di: Liao, Zeyi, et al.
Pubblicazione: (2024)
di: Liao, Zeyi, et al.
Pubblicazione: (2024)
PAL: Proxy-Guided Black-Box Attack on Large Language Models
di: Sitawarin, Chawin, et al.
Pubblicazione: (2024)
di: Sitawarin, Chawin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Low-Resource Languages Jailbreak GPT-4
di: Yong, Zheng-Xin, et al.
Pubblicazione: (2023) -
Does Low Rank Adaptation Lead to Lower Robustness against Training-Time Attacks?
di: Liang, Zi, et al.
Pubblicazione: (2025) -
Statement-Level Vulnerability Detection: Learning Vulnerability Patterns Through Information Theory and Contrastive Learning
di: Nguyen, Van, et al.
Pubblicazione: (2022) -
Measuring Real-World Prompt Injection Attacks in LLM-based Resume Screening
di: Zhang, Mohan, et al.
Pubblicazione: (2026) -
MetaDefense: Defending Finetuning-based Jailbreak Attack Before and During Generation
di: Jiang, Weisen, et al.
Pubblicazione: (2025)