How Vulnerable Are Edge LLMs?
Fuente:
arXiv
Salvato in:
| Autori principali: | Ding, Ao, Li, Hongzong, Liang, Zi, Shi, Zhanpeng, Zhuang, Shuxin, Tang, Shiqin, Feng, Rong, Lu, Ping |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multi-Trigger Poisoning Amplifies Backdoor Vulnerabilities in LLMs
di: Sivapiromrat, Sanhanat, et al.
Pubblicazione: (2025)
di: Sivapiromrat, Sanhanat, et al.
Pubblicazione: (2025)
Future Events as Backdoor Triggers: Investigating Temporal Vulnerabilities in LLMs
di: Price, Sara, et al.
Pubblicazione: (2024)
di: Price, Sara, et al.
Pubblicazione: (2024)
Privacy-Preserving Synthetic Review Generation with Diverse Writing Styles Using LLMs
di: Atwal, Tevin, et al.
Pubblicazione: (2025)
di: Atwal, Tevin, et al.
Pubblicazione: (2025)
Can Federated Learning Safeguard Private Data in LLM Training? Vulnerabilities, Attacks, and Defense Evaluation
di: Guo, Wenkai, et al.
Pubblicazione: (2025)
di: Guo, Wenkai, et al.
Pubblicazione: (2025)
Adversary-Free Counterfactual Prediction via Information-Regularized Representations
di: Tang, Shiqin, et al.
Pubblicazione: (2025)
di: Tang, Shiqin, et al.
Pubblicazione: (2025)
SoK: Membership Inference Attacks on LLMs are Rushing Nowhere (and How to Fix It)
di: Meeus, Matthieu, et al.
Pubblicazione: (2024)
di: Meeus, Matthieu, et al.
Pubblicazione: (2024)
How Much Information Can a Vision Token Hold? A Scaling Law for Recognition Limits in VLMs
di: Zhuang, Shuxin, et al.
Pubblicazione: (2026)
di: Zhuang, Shuxin, et al.
Pubblicazione: (2026)
Can LLMs be Fooled? Investigating Vulnerabilities in LLMs
di: Abdali, Sara, et al.
Pubblicazione: (2024)
di: Abdali, Sara, et al.
Pubblicazione: (2024)
Importing Phantoms: Measuring LLM Package Hallucination Vulnerabilities
di: Krishna, Arjun, et al.
Pubblicazione: (2025)
di: Krishna, Arjun, et al.
Pubblicazione: (2025)
Detecting and Understanding Vulnerabilities in Language Models via Mechanistic Interpretability
di: García-Carrasco, Jorge, et al.
Pubblicazione: (2024)
di: García-Carrasco, Jorge, et al.
Pubblicazione: (2024)
Exploring Vulnerabilities and Protections in Large Language Models: A Survey
di: Liu, Frank Weizhen, et al.
Pubblicazione: (2024)
di: Liu, Frank Weizhen, et al.
Pubblicazione: (2024)
A Large-Scale Empirical Analysis of Custom GPTs' Vulnerabilities in the OpenAI Ecosystem
di: Ogundoyin, Sunday Oyinlola, et al.
Pubblicazione: (2025)
di: Ogundoyin, Sunday Oyinlola, et al.
Pubblicazione: (2025)
WET: Overcoming Paraphrasing Vulnerabilities in Embeddings-as-a-Service with Linear Transformation Watermarks
di: Shetty, Anudeex, et al.
Pubblicazione: (2024)
di: Shetty, Anudeex, et al.
Pubblicazione: (2024)
Training Language Model Agents to Find Vulnerabilities with CTF-Dojo
di: Zhuo, Terry Yue, et al.
Pubblicazione: (2025)
di: Zhuo, Terry Yue, et al.
Pubblicazione: (2025)
MTVHunter: Smart Contracts Vulnerability Detection Based on Multi-Teacher Knowledge Translation
di: Sun, Guokai, et al.
Pubblicazione: (2025)
di: Sun, Guokai, et al.
Pubblicazione: (2025)
Permute-and-Flip: An optimally stable and watermarkable decoder for LLMs
di: Zhao, Xuandong, et al.
Pubblicazione: (2024)
di: Zhao, Xuandong, et al.
Pubblicazione: (2024)
The Hidden Cost of Modeling P(X): Vulnerability to Membership Inference Attacks in Generative Text Classifiers
di: Makroo, Owais, et al.
Pubblicazione: (2025)
di: Makroo, Owais, et al.
Pubblicazione: (2025)
LARGO: Latent Adversarial Reflection through Gradient Optimization for Jailbreaking LLMs
di: Li, Ran, et al.
Pubblicazione: (2025)
di: Li, Ran, et al.
Pubblicazione: (2025)
Colluding LoRA: A Compositional Vulnerability in LLM Safety Alignment
di: Ding, Sihao
Pubblicazione: (2026)
di: Ding, Sihao
Pubblicazione: (2026)
UCD: Unlearning in LLMs via Contrastive Decoding
di: Suriyakumar, Vinith M., et al.
Pubblicazione: (2025)
di: Suriyakumar, Vinith M., et al.
Pubblicazione: (2025)
Sockpuppetting: Jailbreaking LLMs by Combining Prefilling with Optimization
di: Dotsinski, Asen, et al.
Pubblicazione: (2026)
di: Dotsinski, Asen, et al.
Pubblicazione: (2026)
Coercing LLMs to do and reveal (almost) anything
di: Geiping, Jonas, et al.
Pubblicazione: (2024)
di: Geiping, Jonas, et al.
Pubblicazione: (2024)
The Landscape of Memorization in LLMs: Mechanisms, Measurement, and Mitigation
di: Xiong, Alexander, et al.
Pubblicazione: (2025)
di: Xiong, Alexander, et al.
Pubblicazione: (2025)
Does Low Rank Adaptation Lead to Lower Robustness against Training-Time Attacks?
di: Liang, Zi, et al.
Pubblicazione: (2025)
di: Liang, Zi, et al.
Pubblicazione: (2025)
The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions
di: Wallace, Eric, et al.
Pubblicazione: (2024)
di: Wallace, Eric, et al.
Pubblicazione: (2024)
Jailbreaking LLMs via Calibration
di: Lu, Yuxuan, et al.
Pubblicazione: (2026)
di: Lu, Yuxuan, et al.
Pubblicazione: (2026)
Bias Amplification in RAG: Poisoning Knowledge Retrieval to Steer LLMs
di: Wang, Linlin, et al.
Pubblicazione: (2025)
di: Wang, Linlin, et al.
Pubblicazione: (2025)
Hidden in Plain Text: Emergence & Mitigation of Steganographic Collusion in LLMs
di: Mathew, Yohan, et al.
Pubblicazione: (2024)
di: Mathew, Yohan, et al.
Pubblicazione: (2024)
Self-Evaluation as a Defense Against Adversarial Attacks on LLMs
di: Brown, Hannah, et al.
Pubblicazione: (2024)
di: Brown, Hannah, et al.
Pubblicazione: (2024)
The Janus Interface: How Fine-Tuning in Large Language Models Amplifies the Privacy Risks
di: Chen, Xiaoyi, et al.
Pubblicazione: (2023)
di: Chen, Xiaoyi, et al.
Pubblicazione: (2023)
Randomized Masked Finetuning: An Efficient Way to Mitigate Memorization of PIIs in LLMs
di: Joshi, Kunj, et al.
Pubblicazione: (2025)
di: Joshi, Kunj, et al.
Pubblicazione: (2025)
Can We Infer Confidential Properties of Training Data from LLMs?
di: Huang, Pengrun, et al.
Pubblicazione: (2025)
di: Huang, Pengrun, et al.
Pubblicazione: (2025)
Securing Large Language Models (LLMs) from Prompt Injection Attacks
di: Suri, Omar Farooq Khan, et al.
Pubblicazione: (2025)
di: Suri, Omar Farooq Khan, et al.
Pubblicazione: (2025)
How Different Tokenization Algorithms Impact LLMs and Transformer Models for Binary Code Analysis
di: Mostafa, Ahmed, et al.
Pubblicazione: (2025)
di: Mostafa, Ahmed, et al.
Pubblicazione: (2025)
Bits Leaked per Query: Information-Theoretic Bounds on Adversarial Attacks against LLMs
di: Kaneko, Masahiro, et al.
Pubblicazione: (2025)
di: Kaneko, Masahiro, et al.
Pubblicazione: (2025)
Implicit Identity Technologies for LLMs: Fingerprinting and Watermarking across Datasets, Models, and Generated Content
di: Liu, Bing, et al.
Pubblicazione: (2026)
di: Liu, Bing, et al.
Pubblicazione: (2026)
Intrinsic Fingerprint of LLMs: Continue Training is NOT All You Need to Steal A Model!
di: Yoon, Do-hyeon, et al.
Pubblicazione: (2025)
di: Yoon, Do-hyeon, et al.
Pubblicazione: (2025)
LLMCloudHunter: Harnessing LLMs for Automated Extraction of Detection Rules from Cloud-Based CTI
di: Schwartz, Yuval, et al.
Pubblicazione: (2024)
di: Schwartz, Yuval, et al.
Pubblicazione: (2024)
On the Learnability of Watermarks for Language Models
di: Gu, Chenchen, et al.
Pubblicazione: (2023)
di: Gu, Chenchen, et al.
Pubblicazione: (2023)
Particle Dynamics for Latent-Variable Energy-Based Models
di: Tang, Shiqin, et al.
Pubblicazione: (2025)
di: Tang, Shiqin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Multi-Trigger Poisoning Amplifies Backdoor Vulnerabilities in LLMs
di: Sivapiromrat, Sanhanat, et al.
Pubblicazione: (2025) -
Future Events as Backdoor Triggers: Investigating Temporal Vulnerabilities in LLMs
di: Price, Sara, et al.
Pubblicazione: (2024) -
Privacy-Preserving Synthetic Review Generation with Diverse Writing Styles Using LLMs
di: Atwal, Tevin, et al.
Pubblicazione: (2025) -
Can Federated Learning Safeguard Private Data in LLM Training? Vulnerabilities, Attacks, and Defense Evaluation
di: Guo, Wenkai, et al.
Pubblicazione: (2025) -
Adversary-Free Counterfactual Prediction via Information-Regularized Representations
di: Tang, Shiqin, et al.
Pubblicazione: (2025)