SafeCtrl-RL: Inference-Time Adaptive Behaviour Control for LLM Dialogue via RL-Driven Prompt Optimisation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Orme, Michael, Yu, Yanchao, Tan, Zhiyuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OnRL-RAG: Real-Time Personalized Mental Health Dialogue System
von: Bilal, Ahsan, et al.
Veröffentlicht: (2025)
von: Bilal, Ahsan, et al.
Veröffentlicht: (2025)
AdaCtrl: Towards Adaptive and Controllable Reasoning via Difficulty-Aware Budgeting
von: Huang, Shijue, et al.
Veröffentlicht: (2025)
von: Huang, Shijue, et al.
Veröffentlicht: (2025)
Tailored Conversations beyond LLMs: A RL-Based Dialogue Manager
von: Galland, Lucie, et al.
Veröffentlicht: (2025)
von: Galland, Lucie, et al.
Veröffentlicht: (2025)
Beyond Markovian: Reflective Exploration via Bayes-Adaptive RL for LLM Reasoning
von: Zhang, Shenao, et al.
Veröffentlicht: (2025)
von: Zhang, Shenao, et al.
Veröffentlicht: (2025)
CtrlA: Adaptive Retrieval-Augmented Generation via Inherent Control
von: Liu, Huanshuo, et al.
Veröffentlicht: (2024)
von: Liu, Huanshuo, et al.
Veröffentlicht: (2024)
ReviewRL: Towards Automated Scientific Review with RL
von: Zeng, Sihang, et al.
Veröffentlicht: (2025)
von: Zeng, Sihang, et al.
Veröffentlicht: (2025)
On Designing Effective RL Reward at Training Time for LLM Reasoning
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2024)
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2024)
Webscale-RL: Automated Data Pipeline for Scaling RL Data to Pretraining Levels
von: Cen, Zhepeng, et al.
Veröffentlicht: (2025)
von: Cen, Zhepeng, et al.
Veröffentlicht: (2025)
Evaluating & Reducing Deceptive Dialogue From Language Models with Multi-turn RL
von: Abdulhai, Marwa, et al.
Veröffentlicht: (2025)
von: Abdulhai, Marwa, et al.
Veröffentlicht: (2025)
Hard Prompts Made Interpretable: Sparse Entropy Regularization for Prompt Tuning with RL
von: Choi, Yunseon, et al.
Veröffentlicht: (2024)
von: Choi, Yunseon, et al.
Veröffentlicht: (2024)
Reasoning Core: A Scalable RL Environment for LLM Symbolic Reasoning
von: Lacombe, Valentin, et al.
Veröffentlicht: (2025)
von: Lacombe, Valentin, et al.
Veröffentlicht: (2025)
Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
von: Xie, Tian, et al.
Veröffentlicht: (2025)
von: Xie, Tian, et al.
Veröffentlicht: (2025)
Efficient RL for optimizing conversation level outcomes with an LLM-based tutor
von: Nam, Hyunji, et al.
Veröffentlicht: (2025)
von: Nam, Hyunji, et al.
Veröffentlicht: (2025)
LoopServe: An Adaptive Dual-phase LLM Inference Acceleration System for Multi-Turn Dialogues
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
von: Li, Haoyang, et al.
Veröffentlicht: (2025)
Query-Dependent Prompt Evaluation and Optimization with Offline Inverse RL
von: Sun, Hao, et al.
Veröffentlicht: (2023)
von: Sun, Hao, et al.
Veröffentlicht: (2023)
Propensity Inference: Environmental Contributors to LLM Behaviour
von: Järviniemi, Olli, et al.
Veröffentlicht: (2026)
von: Järviniemi, Olli, et al.
Veröffentlicht: (2026)
Beyond SFT-to-RL: Pre-alignment via Black-Box On-Policy Distillation for Multimodal RL
von: Wang, Sudong, et al.
Veröffentlicht: (2026)
von: Wang, Sudong, et al.
Veröffentlicht: (2026)
DetectRL: Benchmarking LLM-Generated Text Detection in Real-World Scenarios
von: Wu, Junchao, et al.
Veröffentlicht: (2024)
von: Wu, Junchao, et al.
Veröffentlicht: (2024)
Speaking at the Right Level: Literacy-Controlled Counterspeech Generation with RAG-RL
von: Song, Xiaoying, et al.
Veröffentlicht: (2025)
von: Song, Xiaoying, et al.
Veröffentlicht: (2025)
FlowRL: Matching Reward Distributions for LLM Reasoning
von: Zhu, Xuekai, et al.
Veröffentlicht: (2025)
von: Zhu, Xuekai, et al.
Veröffentlicht: (2025)
LLM-Driven Multi-Turn Task-Oriented Dialogue Synthesis for Realistic Reasoning
von: Zhu, Yu, et al.
Veröffentlicht: (2026)
von: Zhu, Yu, et al.
Veröffentlicht: (2026)
Words as Beacons: Guiding RL Agents with High-Level Language Prompts
von: Ruiz-Gonzalez, Unai, et al.
Veröffentlicht: (2024)
von: Ruiz-Gonzalez, Unai, et al.
Veröffentlicht: (2024)
GLIDE-RL: Grounded Language Instruction through DEmonstration in RL
von: Kharyal, Chaitanya, et al.
Veröffentlicht: (2024)
von: Kharyal, Chaitanya, et al.
Veröffentlicht: (2024)
CtrlCoT: Dual-Granularity Chain-of-Thought Compression for Controllable Reasoning
von: Fan, Zhenxuan, et al.
Veröffentlicht: (2026)
von: Fan, Zhenxuan, et al.
Veröffentlicht: (2026)
SafeCtrl: Region-Aware Safety Control for Text-to-Image Diffusion via Detect-Then-Suppress
von: Zhang, Lingyun, et al.
Veröffentlicht: (2026)
von: Zhang, Lingyun, et al.
Veröffentlicht: (2026)
Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL
von: Liu, Che, et al.
Veröffentlicht: (2025)
von: Liu, Che, et al.
Veröffentlicht: (2025)
SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
THOR: Tool-Integrated Hierarchical Optimization via RL for Mathematical Reasoning
von: Chang, Qikai, et al.
Veröffentlicht: (2025)
von: Chang, Qikai, et al.
Veröffentlicht: (2025)
Speak & Spell: LLM-Driven Controllable Phonetic Error Augmentation for Robust Dialogue State Tracking
von: Lee, Jihyun, et al.
Veröffentlicht: (2024)
von: Lee, Jihyun, et al.
Veröffentlicht: (2024)
SafeCtrl: Region-Based Safety Control for Text-to-Image Diffusion via Detect-Then-Suppress
von: Zhang, Lingyun, et al.
Veröffentlicht: (2025)
von: Zhang, Lingyun, et al.
Veröffentlicht: (2025)
Look Inward to Explore Outward: Learning Temperature Policy from LLM Internal States via Hierarchical RL
von: Zhou, Yixiao, et al.
Veröffentlicht: (2026)
von: Zhou, Yixiao, et al.
Veröffentlicht: (2026)
SFT-then-RL Outperforms Mixed-Policy Methods for LLM Reasoning
von: Limozin, Alexis, et al.
Veröffentlicht: (2026)
von: Limozin, Alexis, et al.
Veröffentlicht: (2026)
ContextRL: Enhancing MLLM's Knowledge Discovery Efficiency with Context-Augmented RL
von: Lu, Xingyu, et al.
Veröffentlicht: (2026)
von: Lu, Xingyu, et al.
Veröffentlicht: (2026)
Do Not Let Low-Probability Tokens Over-Dominate in RL for LLMs
von: Yang, Zhihe, et al.
Veröffentlicht: (2025)
von: Yang, Zhihe, et al.
Veröffentlicht: (2025)
Sparse-RL: Breaking the Memory Wall in LLM Reinforcement Learning via Stable Sparse Rollouts
von: Luo, Sijia, et al.
Veröffentlicht: (2026)
von: Luo, Sijia, et al.
Veröffentlicht: (2026)
ReadCtrl: Personalizing text generation with readability-controlled instruction learning
von: Tran, Hieu, et al.
Veröffentlicht: (2024)
von: Tran, Hieu, et al.
Veröffentlicht: (2024)
TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
von: Wei, Zhepei, et al.
Veröffentlicht: (2025)
von: Wei, Zhepei, et al.
Veröffentlicht: (2025)
$Q\sharp$: Provably Optimal Distributional RL for LLM Post-Training
von: Zhou, Jin Peng, et al.
Veröffentlicht: (2025)
von: Zhou, Jin Peng, et al.
Veröffentlicht: (2025)
TSR: Trajectory-Search Rollouts for Multi-Turn RL of LLM Agents
von: Djuhera, Aladin, et al.
Veröffentlicht: (2026)
von: Djuhera, Aladin, et al.
Veröffentlicht: (2026)
Anti-Self-Distillation for Reasoning RL via Pointwise Mutual Information
von: Shen, Guobin, et al.
Veröffentlicht: (2026)
von: Shen, Guobin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
OnRL-RAG: Real-Time Personalized Mental Health Dialogue System
von: Bilal, Ahsan, et al.
Veröffentlicht: (2025) -
AdaCtrl: Towards Adaptive and Controllable Reasoning via Difficulty-Aware Budgeting
von: Huang, Shijue, et al.
Veröffentlicht: (2025) -
Tailored Conversations beyond LLMs: A RL-Based Dialogue Manager
von: Galland, Lucie, et al.
Veröffentlicht: (2025) -
Beyond Markovian: Reflective Exploration via Bayes-Adaptive RL for LLM Reasoning
von: Zhang, Shenao, et al.
Veröffentlicht: (2025) -
CtrlA: Adaptive Retrieval-Augmented Generation via Inherent Control
von: Liu, Huanshuo, et al.
Veröffentlicht: (2024)