Saved in:
| Main Authors: | Haldar, Rajdeep, Xing, Yue, Song, Qifan, Lin, Guang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.06921 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Effect of Ambient-Intrinsic Dimension Gap on Adversarial Vulnerability
by: Haldar, Rajdeep, et al.
Published: (2024)
by: Haldar, Rajdeep, et al.
Published: (2024)
f-GRPO and Beyond: Divergence-Based Reinforcement Learning Algorithms for General LLM Alignment
by: Haldar, Rajdeep, et al.
Published: (2026)
by: Haldar, Rajdeep, et al.
Published: (2026)
LLM Safety Alignment is Divergence Estimation in Disguise
by: Haldar, Rajdeep, et al.
Published: (2025)
by: Haldar, Rajdeep, et al.
Published: (2025)
Better Representations via Adversarial Training in Pre-Training: A Theoretical Perspective
by: Xing, Yue, et al.
Published: (2024)
by: Xing, Yue, et al.
Published: (2024)
Fair Supervised Learning with A Simple Random Sampler of Sensitive Attributes
by: Sohn, Jinwon, et al.
Published: (2023)
by: Sohn, Jinwon, et al.
Published: (2023)
Theoretical Understanding of In-Context Learning in Shallow Transformers with Unstructured Data
by: Xing, Yue, et al.
Published: (2024)
by: Xing, Yue, et al.
Published: (2024)
Task-tailored Pre-processing: Fair Downstream Supervised Learning
by: Sohn, Jinwon, et al.
Published: (2026)
by: Sohn, Jinwon, et al.
Published: (2026)
Parallelly Tempered Generative Adversarial Nets: Toward Stabilized Gradients
by: Sohn, Jinwon, et al.
Published: (2024)
by: Sohn, Jinwon, et al.
Published: (2024)
SQS: Bayesian DNN Compression through Sparse Quantized Sub-distributions
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
Bayesian Federated Learning with Hamiltonian Monte Carlo: Algorithm and Theory
by: Liang, Jiajun, et al.
Published: (2024)
by: Liang, Jiajun, et al.
Published: (2024)
Adversarial Autoencoders in Operator Learning
by: Enyeart, Dustin, et al.
Published: (2024)
by: Enyeart, Dustin, et al.
Published: (2024)
Generalized Discrete Diffusion with Self-Correction
by: Wang, Linxuan, et al.
Published: (2026)
by: Wang, Linxuan, et al.
Published: (2026)
Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training
by: Moya, Christian, et al.
Published: (2026)
by: Moya, Christian, et al.
Published: (2026)
Ensuring Calibration Robustness in Split Conformal Prediction Under Adversarial Attacks
by: Qian, Xunlei, et al.
Published: (2025)
by: Qian, Xunlei, et al.
Published: (2025)
Impact of Positional Encoding: Clean and Adversarial Rademacher Complexity for Transformers under In-Context Regression
by: He, Weiyi, et al.
Published: (2025)
by: He, Weiyi, et al.
Published: (2025)
Quantifying Membership Disclosure Risk for Tabular Synthetic Data Using Kernel Density Estimators
by: Pathak, Rajdeep, et al.
Published: (2026)
by: Pathak, Rajdeep, et al.
Published: (2026)
Deep Generative Spatiotemporal Engression for Probabilistic Forecasting of Epidemics
by: Pathak, Rajdeep, et al.
Published: (2026)
by: Pathak, Rajdeep, et al.
Published: (2026)
Adversarial Reward Auditing for Active Detection and Mitigation of Reward Hacking
by: Beigi, Mohammad, et al.
Published: (2026)
by: Beigi, Mohammad, et al.
Published: (2026)
Vulnerability-Aware Robust Multimodal Adversarial Training
by: Zhang, Junrui, et al.
Published: (2025)
by: Zhang, Junrui, et al.
Published: (2025)
Unveiling and Mitigating Adversarial Vulnerabilities in Iterative Optimizers
by: Sofer, Elad, et al.
Published: (2025)
by: Sofer, Elad, et al.
Published: (2025)
Crafting Imperceptible On-Manifold Adversarial Attacks for Tabular Data
by: He, Zhipeng, et al.
Published: (2025)
by: He, Zhipeng, et al.
Published: (2025)
Enhancing Low-Precision Sampling via Stochastic Gradient Hamiltonian Monte Carlo
by: Wang, Ziyi, et al.
Published: (2023)
by: Wang, Ziyi, et al.
Published: (2023)
Adversarial Purification by Consistency-aware Latent Space Optimization on Data Manifolds
by: Zhang, Shuhai, et al.
Published: (2024)
by: Zhang, Shuhai, et al.
Published: (2024)
How to Enhance Downstream Adversarial Robustness (almost) without Touching the Pre-Trained Foundation Model?
by: Liu, Meiqi, et al.
Published: (2025)
by: Liu, Meiqi, et al.
Published: (2025)
Rectifying Adversarial Examples Using Their Vulnerabilities
by: Morimoto, Fumiya, et al.
Published: (2026)
by: Morimoto, Fumiya, et al.
Published: (2026)
A Geometric Framework for Adversarial Vulnerability in Machine Learning
by: Bell, Brian
Published: (2024)
by: Bell, Brian
Published: (2024)
MEPT: Mixture of Expert Prompt Tuning as a Manifold Mapper
by: Zeng, Runjia, et al.
Published: (2025)
by: Zeng, Runjia, et al.
Published: (2025)
Manifold-Constrained Adversarial Training for Long-Tailed Robustness via Geometric Alignment
by: Xian, Guanmeng, et al.
Published: (2026)
by: Xian, Guanmeng, et al.
Published: (2026)
Golyadkin's Torment: Doppelgängers and Adversarial Vulnerability
by: Kamberov, George I.
Published: (2024)
by: Kamberov, George I.
Published: (2024)
Transcending Adversarial Perturbations: Manifold-Aided Adversarial Examples with Legitimate Semantics
by: Li, Shuai, et al.
Published: (2024)
by: Li, Shuai, et al.
Published: (2024)
Scale-Aware Adversarial Analysis: A Diagnostic for Generative AI in Multiscale Complex Systems
by: Zhao, Mengke, et al.
Published: (2026)
by: Zhao, Mengke, et al.
Published: (2026)
Model-Free Adversarial Purification via Coarse-To-Fine Tensor Network Representation
by: Lin, Guang, et al.
Published: (2025)
by: Lin, Guang, et al.
Published: (2025)
Unveiling the Vulnerability of Graph-LLMs: An Interpretable Multi-Dimensional Adversarial Attack on TAGs
by: Fan, Bowen, et al.
Published: (2025)
by: Fan, Bowen, et al.
Published: (2025)
Score Matching for Truncated Density Estimation on a Manifold
by: Williams, Daniel J., et al.
Published: (2022)
by: Williams, Daniel J., et al.
Published: (2022)
Are Modern Speech Enhancement Systems Vulnerable to Adversarial Attacks?
by: Makarov, Rostislav, et al.
Published: (2025)
by: Makarov, Rostislav, et al.
Published: (2025)
Knowledge Distillation Detection for Open-weights Models
by: Shi, Qin, et al.
Published: (2025)
by: Shi, Qin, et al.
Published: (2025)
LoRA-Mini : Adaptation Matrices Decomposition and Selective Training
by: Singh, Ayush, et al.
Published: (2024)
by: Singh, Ayush, et al.
Published: (2024)
Superposition as Lossy Compression: Measure with Sparse Autoencoders and Connect to Adversarial Vulnerability
by: Bereska, Leonard, et al.
Published: (2025)
by: Bereska, Leonard, et al.
Published: (2025)
Flow Battery Manifold Design with Heterogeneous Inputs Through Generative Adversarial Neural Networks
by: Seng, Eric, et al.
Published: (2025)
by: Seng, Eric, et al.
Published: (2025)
Error Slice Discovery via Manifold Compactness
by: Yu, Han, et al.
Published: (2025)
by: Yu, Han, et al.
Published: (2025)
Similar Items
-
Effect of Ambient-Intrinsic Dimension Gap on Adversarial Vulnerability
by: Haldar, Rajdeep, et al.
Published: (2024) -
f-GRPO and Beyond: Divergence-Based Reinforcement Learning Algorithms for General LLM Alignment
by: Haldar, Rajdeep, et al.
Published: (2026) -
LLM Safety Alignment is Divergence Estimation in Disguise
by: Haldar, Rajdeep, et al.
Published: (2025) -
Better Representations via Adversarial Training in Pre-Training: A Theoretical Perspective
by: Xing, Yue, et al.
Published: (2024) -
Fair Supervised Learning with A Simple Random Sampler of Sensitive Attributes
by: Sohn, Jinwon, et al.
Published: (2023)