Universal and Transferable Adversarial Attack on Large Language Models Using Exponentiated Gradient Descent
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Biswas, Sajib, Nishino, Mao, Chacko, Samuel Jacob, Liu, Xiuwen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adversarial Attack on Large Language Models using Exponentiated Gradient Descent
von: Biswas, Sajib, et al.
Veröffentlicht: (2025)
von: Biswas, Sajib, et al.
Veröffentlicht: (2025)
Adversarial Attacks on Large Language Models Using Regularized Relaxation
von: Chacko, Samuel Jacob, et al.
Veröffentlicht: (2024)
von: Chacko, Samuel Jacob, et al.
Veröffentlicht: (2024)
Mechanistic Understandings of Representation Vulnerabilities and Engineering Robust Vision Transformers
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
Attacking Large Language Models with Projected Gradient Descent
von: Geisler, Simon, et al.
Veröffentlicht: (2024)
von: Geisler, Simon, et al.
Veröffentlicht: (2024)
Numerical Instability and Chaos: Quantifying the Unpredictability of Large Language Models
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2026)
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2026)
DeepSeek on a Trip: Inducing Targeted Visual Hallucinations via Representation Vulnerabilities
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
Enhancing Adversarial Text Attacks on BERT Models with Projected Gradient Descent
von: Waghela, Hetvi, et al.
Veröffentlicht: (2024)
von: Waghela, Hetvi, et al.
Veröffentlicht: (2024)
Efficient Extractive Text Summarization for Online News Articles Using Machine Learning
von: Biswas, Sajib, et al.
Veröffentlicht: (2025)
von: Biswas, Sajib, et al.
Veröffentlicht: (2025)
Spatial-ViLT: Enhancing Visual Spatial Reasoning through Multi-Task Learning
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
Exponential Convergence of (Stochastic) Gradient Descent for Separable Logistic Regression
von: Kale, Sacchit, et al.
Veröffentlicht: (2026)
von: Kale, Sacchit, et al.
Veröffentlicht: (2026)
Mirror Descent and Novel Exponentiated Gradient Algorithms Using Trace-Form Entropies and Deformed Logarithms
von: Cichocki, Andrzej, et al.
Veröffentlicht: (2025)
von: Cichocki, Andrzej, et al.
Veröffentlicht: (2025)
Understanding Model Ensemble in Transferable Adversarial Attack
von: Yao, Wei, et al.
Veröffentlicht: (2024)
von: Yao, Wei, et al.
Veröffentlicht: (2024)
Can Gradient Descent Simulate Prompting?
von: Zhang, Eric, et al.
Veröffentlicht: (2025)
von: Zhang, Eric, et al.
Veröffentlicht: (2025)
Using Mechanistic Interpretability to Craft Adversarial Attacks against Large Language Models
von: Winninger, Thomas, et al.
Veröffentlicht: (2025)
von: Winninger, Thomas, et al.
Veröffentlicht: (2025)
Sampling-aware Adversarial Attacks Against Large Language Models
von: Beyer, Tim, et al.
Veröffentlicht: (2025)
von: Beyer, Tim, et al.
Veröffentlicht: (2025)
Fast Adversarial Attacks with Gradient Prediction
von: Ciosek, Kamil, et al.
Veröffentlicht: (2026)
von: Ciosek, Kamil, et al.
Veröffentlicht: (2026)
Transferable Adversarial Attacks on Black-Box Vision-Language Models
von: Hu, Kai, et al.
Veröffentlicht: (2025)
von: Hu, Kai, et al.
Veröffentlicht: (2025)
ReEval: Automatic Hallucination Evaluation for Retrieval-Augmented Large Language Models via Transferable Adversarial Attacks
von: Yu, Xiaodong, et al.
Veröffentlicht: (2023)
von: Yu, Xiaodong, et al.
Veröffentlicht: (2023)
Thermodynamic Natural Gradient Descent
von: Donatella, Kaelan, et al.
Veröffentlicht: (2024)
von: Donatella, Kaelan, et al.
Veröffentlicht: (2024)
Derivatives of Stochastic Gradient Descent in parametric optimization
von: Iutzeler, Franck, et al.
Veröffentlicht: (2024)
von: Iutzeler, Franck, et al.
Veröffentlicht: (2024)
Adversarial Evasion Attack Efficiency against Large Language Models
von: Vitorino, João, et al.
Veröffentlicht: (2024)
von: Vitorino, João, et al.
Veröffentlicht: (2024)
Algebraic Adversarial Attacks on Integrated Gradients
von: Simpson, Lachlan, et al.
Veröffentlicht: (2024)
von: Simpson, Lachlan, et al.
Veröffentlicht: (2024)
REINFORCE Adversarial Attacks on Large Language Models: An Adaptive, Distributional, and Semantic Objective
von: Geisler, Simon, et al.
Veröffentlicht: (2025)
von: Geisler, Simon, et al.
Veröffentlicht: (2025)
Beyond Suffixes: Token Position in GCG Adversarial Attacks on Large Language Models
von: Eddoubi, Hicham, et al.
Veröffentlicht: (2026)
von: Eddoubi, Hicham, et al.
Veröffentlicht: (2026)
DiffGradCAM: A Class Activation Map Using the Full Model Decision to Solve Unaddressed Adversarial Attacks
von: Piland, Jacob, et al.
Veröffentlicht: (2025)
von: Piland, Jacob, et al.
Veröffentlicht: (2025)
Exploring Gradient-Guided Masked Language Model to Detect Textual Adversarial Attacks
von: Zhang, Xiaomei, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaomei, et al.
Veröffentlicht: (2025)
Tight Bounds for Logistic Regression with Large Stepsize Gradient Descent in Low Dimension
von: Crawshaw, Michael, et al.
Veröffentlicht: (2026)
von: Crawshaw, Michael, et al.
Veröffentlicht: (2026)
Bypassing the Exponential Dependency: Looped Transformers Efficiently Learn In-context by Multi-step Gradient Descent
von: Chen, Bo, et al.
Veröffentlicht: (2024)
von: Chen, Bo, et al.
Veröffentlicht: (2024)
Occam Gradient Descent
von: Kausik, B. N.
Veröffentlicht: (2024)
von: Kausik, B. N.
Veröffentlicht: (2024)
Hybrid Coordinate Descent for Efficient Neural Network Learning Using Line Search and Gradient Descent
von: Hsiao, Yen-Che, et al.
Veröffentlicht: (2024)
von: Hsiao, Yen-Che, et al.
Veröffentlicht: (2024)
Advancing Adversarial Suffix Transfer Learning on Aligned Large Language Models
von: Liu, Hongfu, et al.
Veröffentlicht: (2024)
von: Liu, Hongfu, et al.
Veröffentlicht: (2024)
Learning Tree-Based Models with Gradient Descent
von: Marton, Sascha
Veröffentlicht: (2026)
von: Marton, Sascha
Veröffentlicht: (2026)
Bolstering Stochastic Gradient Descent with Model Building
von: Birbil, S. Ilker, et al.
Veröffentlicht: (2021)
von: Birbil, S. Ilker, et al.
Veröffentlicht: (2021)
Benchmarking Transferable Adversarial Attacks
von: Jin, Zhibo, et al.
Veröffentlicht: (2024)
von: Jin, Zhibo, et al.
Veröffentlicht: (2024)
Large Stepsizes Accelerate Gradient Descent for Regularized Logistic Regression
von: Wu, Jingfeng, et al.
Veröffentlicht: (2025)
von: Wu, Jingfeng, et al.
Veröffentlicht: (2025)
Stochastic Normalized Gradient Descent with Momentum for Large-Batch Training
von: Zhao, Shen-Yi, et al.
Veröffentlicht: (2020)
von: Zhao, Shen-Yi, et al.
Veröffentlicht: (2020)
Save It All: Enabling Full Parameter Tuning for Federated Large Language Models via Cycle Block Gradient Descent
von: Wang, Lin, et al.
Veröffentlicht: (2024)
von: Wang, Lin, et al.
Veröffentlicht: (2024)
Auditing Information Disclosure During LLM-Scale Gradient Descent Using Gradient Uniqueness
von: Abdelghafar, Sleem, et al.
Veröffentlicht: (2025)
von: Abdelghafar, Sleem, et al.
Veröffentlicht: (2025)
ConMeZO: Adaptive Descent-Direction Sampling for Gradient-Free Finetuning of Large Language Models
von: Behric, Lejs Deen, et al.
Veröffentlicht: (2025)
von: Behric, Lejs Deen, et al.
Veröffentlicht: (2025)
The Resurgence of GCG Adversarial Attacks on Large Language Models
von: Tan, Yuting, et al.
Veröffentlicht: (2025)
von: Tan, Yuting, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Adversarial Attack on Large Language Models using Exponentiated Gradient Descent
von: Biswas, Sajib, et al.
Veröffentlicht: (2025) -
Adversarial Attacks on Large Language Models Using Regularized Relaxation
von: Chacko, Samuel Jacob, et al.
Veröffentlicht: (2024) -
Mechanistic Understandings of Representation Vulnerabilities and Engineering Robust Vision Transformers
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025) -
Attacking Large Language Models with Projected Gradient Descent
von: Geisler, Simon, et al.
Veröffentlicht: (2024) -
Numerical Instability and Chaos: Quantifying the Unpredictability of Large Language Models
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2026)