RSafe: Incentivizing proactive reasoning to build robust and adaptive LLM safeguards

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zheng, Jingnan, Ji, Xiangtian, Lu, Yijun, Cui, Chenhang, Zhao, Weixiang, Deng, Gelei, Liang, Zhenkai, Zhang, An, Chua, Tat-Seng
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!