Secure Tug-of-War (SecTOW): Iterative Defense-Attack Training with Reinforcement Learning for Multimodal Model Security
Fuente:
arXiv
Saved in:
| Main Authors: | Dai, Muzhi, Liu, Shixuan, Zhao, Zhiyuan, Gao, Junyu, Sun, Hao, Li, Xuelong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
System Password Security: Attack and Defense Mechanisms
by: Shi, Chaofang, et al.
Published: (2025)
by: Shi, Chaofang, et al.
Published: (2025)
SecDTD: Dynamic Token Drop for Secure Transformers Inference
by: Cai, Yifei, et al.
Published: (2026)
by: Cai, Yifei, et al.
Published: (2026)
SecPI: Secure Code Generation with Reasoning Models via Security Reasoning Internalization
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
FedSecurity: Benchmarking Attacks and Defenses in Federated Learning and Federated LLMs
by: Han, Shanshan, et al.
Published: (2023)
by: Han, Shanshan, et al.
Published: (2023)
Uncovering Attacks and Defenses in Secure Aggregation for Federated Deep Learning
by: Zhang, Yiwei, et al.
Published: (2024)
by: Zhang, Yiwei, et al.
Published: (2024)
Ethereum Crypto Wallets under Address Poisoning: How Usable and Secure Are They?
by: Guan, Shixuan, et al.
Published: (2025)
by: Guan, Shixuan, et al.
Published: (2025)
Meta SecAlign: A Secure Foundation LLM Against Prompt Injection Attacks
by: Chen, Sizhe, et al.
Published: (2025)
by: Chen, Sizhe, et al.
Published: (2025)
Ditto: Elastic Confidential VMs with Secure and Dynamic CPU Scaling
by: Zhao, Shixuan, et al.
Published: (2024)
by: Zhao, Shixuan, et al.
Published: (2024)
λ-SecAgg: Partial Vector Freezing for Lightweight Secure Aggregation in Federated Learning
by: Zhang, Siqing, et al.
Published: (2023)
by: Zhang, Siqing, et al.
Published: (2023)
Securing Retrieval-Augmented Generation: A Taxonomy of Attacks, Defenses, and Future Directions
by: Xu, Yuming, et al.
Published: (2026)
by: Xu, Yuming, et al.
Published: (2026)
SecIC3: Customizing IC3 for Hardware Security Verification
by: Tan, Qinhan, et al.
Published: (2026)
by: Tan, Qinhan, et al.
Published: (2026)
AntiFLipper: A Secure and Efficient Defense Against Label-Flipping Attacks in Federated Learning
by: Rahman, Aashnan, et al.
Published: (2025)
by: Rahman, Aashnan, et al.
Published: (2025)
SecureLearn -- An Attack-agnostic Defense for Multiclass Machine Learning Against Data Poisoning Attacks
by: Paracha, Anum, et al.
Published: (2025)
by: Paracha, Anum, et al.
Published: (2025)
CellSecInspector: Safeguarding Cellular Networks via Automated Security Analysis on Specifications
by: Xie, Ke, et al.
Published: (2025)
by: Xie, Ke, et al.
Published: (2025)
Enhancing Security in Deep Reinforcement Learning: A Comprehensive Survey on Adversarial Attacks and Defenses
by: Yichao, Wu, et al.
Published: (2025)
by: Yichao, Wu, et al.
Published: (2025)
Reinforcement Learning-Based Approaches for Enhancing Security and Resilience in Smart Control: A Survey on Attack and Defense Methods
by: Zhang, Zheyu
Published: (2024)
by: Zhang, Zheyu
Published: (2024)
Enhancing O-RAN Security: Evasion Attacks and Robust Defenses for Graph Reinforcement Learning-based Connection Management
by: Balakrishnan, Ravikumar, et al.
Published: (2024)
by: Balakrishnan, Ravikumar, et al.
Published: (2024)
SecPLF: Secure Protocols for Loanable Funds against Oracle Manipulation Attacks
by: Arora, Sanidhay, et al.
Published: (2024)
by: Arora, Sanidhay, et al.
Published: (2024)
SmartX Intelligent Sec: A Security Framework Based on Machine Learning and eBPF/XDP
by: Farasat, Talaya, et al.
Published: (2024)
by: Farasat, Talaya, et al.
Published: (2024)
SecurityLingua: Efficient Defense of LLM Jailbreak Attacks via Security-Aware Prompt Compression
by: Li, Yucheng, et al.
Published: (2025)
by: Li, Yucheng, et al.
Published: (2025)
Evolving Security in LLMs: A Study of Jailbreak Attacks and Defenses
by: Shang, Zhengchun, et al.
Published: (2025)
by: Shang, Zhengchun, et al.
Published: (2025)
CredSec: A Blockchain-based Secure Credential Management System for University Adoption
by: Habib, Md. Ahsan, et al.
Published: (2024)
by: Habib, Md. Ahsan, et al.
Published: (2024)
SecGoal: A Benchmark for Extracting Formalizable Security Goals from Protocol Documents
by: Huang, Dawei, et al.
Published: (2026)
by: Huang, Dawei, et al.
Published: (2026)
SecMLOps: A Comprehensive Framework for Integrating Security Throughout the MLOps Lifecycle
by: Zhang, Xinrui, et al.
Published: (2026)
by: Zhang, Xinrui, et al.
Published: (2026)
Securing Large Language Models: Addressing Bias, Misinformation, and Prompt Attacks
by: Peng, Benji, et al.
Published: (2024)
by: Peng, Benji, et al.
Published: (2024)
BMC4TimeSec: Verification Of Timed Security Protocols
by: Zbrzezny, Agnieszka M.
Published: (2026)
by: Zbrzezny, Agnieszka M.
Published: (2026)
Sec5GLoc: Securing 5G Indoor Localization via Adversary-Resilient Deep Learning Architecture
by: Alla, Ildi, et al.
Published: (2025)
by: Alla, Ildi, et al.
Published: (2025)
LLMs Caught in the Crossfire: Malware Requests and Jailbreak Challenges
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
On the (In-)Security of the Shuffling Defense in the Transformer Secure Inference
by: Li, Zhengyi, et al.
Published: (2026)
by: Li, Zhengyi, et al.
Published: (2026)
SecTracer: A Framework for Uncovering the Root Causes of Network Intrusions via Security Provenance
by: Lee, Seunghyeon, et al.
Published: (2025)
by: Lee, Seunghyeon, et al.
Published: (2025)
Covert Attacks on Machine Learning Training in Passively Secure MPC
by: Jagielski, Matthew, et al.
Published: (2025)
by: Jagielski, Matthew, et al.
Published: (2025)
Snail: Secure Single Iteration Localization
by: Choncholas, James, et al.
Published: (2024)
by: Choncholas, James, et al.
Published: (2024)
SecCodePRM: A Process Reward Model for Code Security
by: Yu, Weichen, et al.
Published: (2026)
by: Yu, Weichen, et al.
Published: (2026)
SecONNds: Secure Outsourced Neural Network Inference on ImageNet
by: Balla, Shashank
Published: (2025)
by: Balla, Shashank
Published: (2025)
SecScale: A Scalable and Secure Trusted Execution Environment for Servers
by: Sunny, Ani, et al.
Published: (2024)
by: Sunny, Ani, et al.
Published: (2024)
DaemonSec: Examining the Role of Machine Learning for Daemon Security in Linux Environments
by: Farjad, Sheikh Muhammad
Published: (2025)
by: Farjad, Sheikh Muhammad
Published: (2025)
CellularSpecSec-Bench: A Staged Benchmark for Evidence-Grounded Interpretation and Security Reasoning over 3GPP Specifications
by: Xie, Ke, et al.
Published: (2026)
by: Xie, Ke, et al.
Published: (2026)
Comparative Analysis of AI-Driven Security Approaches in DevSecOps: Challenges, Solutions, and Future Directions
by: Binbeshr, Farid, et al.
Published: (2025)
by: Binbeshr, Farid, et al.
Published: (2025)
Deep Learning Model Security: Threats and Defenses
by: Wang, Tianyang, et al.
Published: (2024)
by: Wang, Tianyang, et al.
Published: (2024)
Smart Grid Security: A Verified Deep Reinforcement Learning Framework to Counter Cyber-Physical Attacks
by: Maiti, Suman, et al.
Published: (2024)
by: Maiti, Suman, et al.
Published: (2024)
Similar Items
-
System Password Security: Attack and Defense Mechanisms
by: Shi, Chaofang, et al.
Published: (2025) -
SecDTD: Dynamic Token Drop for Secure Transformers Inference
by: Cai, Yifei, et al.
Published: (2026) -
SecPI: Secure Code Generation with Reasoning Models via Security Reasoning Internalization
by: Wang, Hao, et al.
Published: (2026) -
FedSecurity: Benchmarking Attacks and Defenses in Federated Learning and Federated LLMs
by: Han, Shanshan, et al.
Published: (2023) -
Uncovering Attacks and Defenses in Secure Aggregation for Federated Deep Learning
by: Zhang, Yiwei, et al.
Published: (2024)