AdvChain: Adversarial Chain-of-Thought Tuning for Robust Safety Alignment of Large Reasoning Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zhu, Zihao, Wu, Xinyu, Hu, Gehan, Lyu, Siwei, Xu, Ke, Wu, Baoyuan
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!

Similar Items