Conflict-Aware Adversarial Training
Fuente:
arXiv
Saved in:
| Main Authors: | Xue, Zhiyu, Wang, Haohan, Qin, Yao, Pedarsani, Ramtin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Inverse Reinforcement Learning by Estimating Expertise of Demonstrators
by: Beliaev, Mark, et al.
Published: (2024)
by: Beliaev, Mark, et al.
Published: (2024)
Generalization Properties of Adversarial Training for $\ell_0$-Bounded Adversarial Attacks
by: Delgosha, Payam, et al.
Published: (2024)
by: Delgosha, Payam, et al.
Published: (2024)
Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations
by: Xue, Zhiyu, et al.
Published: (2025)
by: Xue, Zhiyu, et al.
Published: (2025)
Deactivating Refusal Triggers: Understanding and Mitigating Overrefusal in Safety Alignment
by: Xue, Zhiyu, et al.
Published: (2026)
by: Xue, Zhiyu, et al.
Published: (2026)
Robust Adversarial Quantification via Conflict-Aware Evidential Deep Learning
by: Barker, Charmaine, et al.
Published: (2025)
by: Barker, Charmaine, et al.
Published: (2025)
SORA: Free Second-Order Attacks in Fast Adversarial Training
by: Teymourian, Mazdak, et al.
Published: (2026)
by: Teymourian, Mazdak, et al.
Published: (2026)
DAFA: Distance-Aware Fair Adversarial Training
by: Lee, Hyungyu, et al.
Published: (2024)
by: Lee, Hyungyu, et al.
Published: (2024)
No Free Lunch for Defending Against Prefilling Attack by In-Context Learning
by: Xue, Zhiyu, et al.
Published: (2024)
by: Xue, Zhiyu, et al.
Published: (2024)
DistDD: Distributed Data Distillation Aggregation through Gradient Matching
by: Wang, Peiran, et al.
Published: (2024)
by: Wang, Peiran, et al.
Published: (2024)
CARE-RL: Capability-Aware Reinforcement Learning for Mitigating Cross-Domain Conflicts
by: Zhang, Rui, et al.
Published: (2026)
by: Zhang, Rui, et al.
Published: (2026)
SPEX: Scaling Feature Interaction Explanations for LLMs
by: Kang, Justin Singh, et al.
Published: (2025)
by: Kang, Justin Singh, et al.
Published: (2025)
Few-Shot Adversarial Low-Rank Fine-Tuning of Vision-Language Models
by: Ghiasvand, Sajjad, et al.
Published: (2025)
by: Ghiasvand, Sajjad, et al.
Published: (2025)
REALM: Reliable Expertise-Aware Language Model Fine-Tuning from Noisy Annotations
by: Ghiasvand, Sajjad, et al.
Published: (2026)
by: Ghiasvand, Sajjad, et al.
Published: (2026)
Adversarial Training: A Survey
by: Zhao, Mengnan, et al.
Published: (2024)
by: Zhao, Mengnan, et al.
Published: (2024)
Adversarial Training for Process Reward Models
by: Juneja, Gurusha, et al.
Published: (2025)
by: Juneja, Gurusha, et al.
Published: (2025)
LLMPi: Optimizing LLMs for High-Throughput on Raspberry Pi
by: Ardakani, Mahsa, et al.
Published: (2025)
by: Ardakani, Mahsa, et al.
Published: (2025)
Decentralized Low-Rank Fine-Tuning of Large Language Models
by: Ghiasvand, Sajjad, et al.
Published: (2025)
by: Ghiasvand, Sajjad, et al.
Published: (2025)
Leveraging Information Consistency in Frequency and Spatial Domain for Adversarial Attacks
by: Jin, Zhibo, et al.
Published: (2024)
by: Jin, Zhibo, et al.
Published: (2024)
Communication-Efficient and Tensorized Federated Fine-Tuning of Large Language Models
by: Ghiasvand, Sajjad, et al.
Published: (2024)
by: Ghiasvand, Sajjad, et al.
Published: (2024)
Length-Aware Adversarial Training for Variable-Length Trajectories: Digital Twins for Mall Shopper Paths
by: Sun, He, et al.
Published: (2026)
by: Sun, He, et al.
Published: (2026)
Revisiting the Relationship between Adversarial and Clean Training: Why Clean Training Can Make Adversarial Training Better
by: Zhou, MingWei, et al.
Published: (2025)
by: Zhou, MingWei, et al.
Published: (2025)
Benign Overfitting in Adversarial Training for Vision Transformers
by: Zhang, Jiaming, et al.
Published: (2026)
by: Zhang, Jiaming, et al.
Published: (2026)
Matmul or No Matmul in the Era of 1-bit LLMs
by: Malekar, Jinendra, et al.
Published: (2024)
by: Malekar, Jinendra, et al.
Published: (2024)
MemLoss: Enhancing Adversarial Training with Recycling Adversarial Examples
by: Mahdi, Soroush, et al.
Published: (2025)
by: Mahdi, Soroush, et al.
Published: (2025)
On the Duality Between Sharpness-Aware Minimization and Adversarial Training
by: Zhang, Yihao, et al.
Published: (2024)
by: Zhang, Yihao, et al.
Published: (2024)
Ignition Phase : Standard Training for Fast Adversarial Robustness
by: Yu-Hang, Wang, et al.
Published: (2025)
by: Yu-Hang, Wang, et al.
Published: (2025)
Annealing Self-Distillation Rectification Improves Adversarial Training
by: Wu, Yu-Yu, et al.
Published: (2023)
by: Wu, Yu-Yu, et al.
Published: (2023)
CABS: Conflict-Aware and Balanced Sparsification for Enhancing Model Merging
by: Yang, Zongzhen, et al.
Published: (2025)
by: Yang, Zongzhen, et al.
Published: (2025)
Decentralized Adversarial Training over Graphs
by: Cao, Ying, et al.
Published: (2023)
by: Cao, Ying, et al.
Published: (2023)
Crafting Adversarial Inputs for Large Vision-Language Models Using Black-Box Optimization
by: Guan, Jiwei, et al.
Published: (2026)
by: Guan, Jiwei, et al.
Published: (2026)
Fast-Slow Co-advancing Optimizer: Toward Harmonious Adversarial Training of GAN
by: Wang, Lin, et al.
Published: (2025)
by: Wang, Lin, et al.
Published: (2025)
Conflict-Aware Pseudo Labeling via Optimal Transport for Entity Alignment
by: Ding, Qijie, et al.
Published: (2022)
by: Ding, Qijie, et al.
Published: (2022)
SPICE: Submodular Penalized Information-Conflict Selection for Efficient Large Language Model Training
by: Chang, Powei, et al.
Published: (2026)
by: Chang, Powei, et al.
Published: (2026)
TAET: Two-Stage Adversarial Equalization Training on Long-Tailed Distributions
by: YuHang, Wang, et al.
Published: (2025)
by: YuHang, Wang, et al.
Published: (2025)
PAR-AdvGAN: Improving Adversarial Attack Capability with Progressive Auto-Regression AdvGAN
by: Zhang, Jiayu, et al.
Published: (2025)
by: Zhang, Jiayu, et al.
Published: (2025)
CAT Merging: A Training-Free Approach for Resolving Conflicts in Model Merging
by: Sun, Wenju, et al.
Published: (2025)
by: Sun, Wenju, et al.
Published: (2025)
FedCARE: Federated Unlearning with Conflict-Aware Projection and Relearning-Resistant Recovery
by: Li, Yue, et al.
Published: (2026)
by: Li, Yue, et al.
Published: (2026)
Task Arithmetic in Trust Region: A Training-Free Model Merging Approach to Navigate Knowledge Conflicts
by: Sun, Wenju, et al.
Published: (2025)
by: Sun, Wenju, et al.
Published: (2025)
Improving Clean Accuracy via a Tangent-Space Perspective on Adversarial Training
by: Yi, Bongsoo, et al.
Published: (2024)
by: Yi, Bongsoo, et al.
Published: (2024)
NPAT Null-Space Projected Adversarial Training Towards Zero Deterioration
by: Hu, Hanyi, et al.
Published: (2024)
by: Hu, Hanyi, et al.
Published: (2024)
Similar Items
-
Inverse Reinforcement Learning by Estimating Expertise of Demonstrators
by: Beliaev, Mark, et al.
Published: (2024) -
Generalization Properties of Adversarial Training for $\ell_0$-Bounded Adversarial Attacks
by: Delgosha, Payam, et al.
Published: (2024) -
Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations
by: Xue, Zhiyu, et al.
Published: (2025) -
Deactivating Refusal Triggers: Understanding and Mitigating Overrefusal in Safety Alignment
by: Xue, Zhiyu, et al.
Published: (2026) -
Robust Adversarial Quantification via Conflict-Aware Evidential Deep Learning
by: Barker, Charmaine, et al.
Published: (2025)