HIPO: Instruction Hierarchy via Constrained Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Keru, Luo, Jun, Lin, Sen, Liang, Yingbin, Velasquez, Alvaro, Bastian, Nathaniel, Zou, Shaofeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Fast Safe Online Reinforcement Learning via Policy Finetuning
von: Chen, Keru, et al.
Veröffentlicht: (2024)
von: Chen, Keru, et al.
Veröffentlicht: (2024)
Why Adam Can Beat SGD: Second-Moment Normalization Yields Sharper Tails
von: Jin, Ruinan, et al.
Veröffentlicht: (2026)
von: Jin, Ruinan, et al.
Veröffentlicht: (2026)
Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification
von: Guo, Yiju, et al.
Veröffentlicht: (2026)
von: Guo, Yiju, et al.
Veröffentlicht: (2026)
Instructional Segment Embedding: Improving LLM Safety with Instruction Hierarchy
von: Wu, Tong, et al.
Veröffentlicht: (2024)
von: Wu, Tong, et al.
Veröffentlicht: (2024)
Agentic Transformers Provably Learn to Search via Reinforcement Learning
von: Yang, Tong, et al.
Veröffentlicht: (2026)
von: Yang, Tong, et al.
Veröffentlicht: (2026)
Detector-Evasive LLM Paraphrasing via Constrained Policy Optimization
von: Wang, Mingyi, et al.
Veröffentlicht: (2026)
von: Wang, Mingyi, et al.
Veröffentlicht: (2026)
Model-Free Robust Reinforcement Learning with Sample Complexity Analysis
von: Wang, Yudan, et al.
Veröffentlicht: (2024)
von: Wang, Yudan, et al.
Veröffentlicht: (2024)
Bridging Online and Offline RL: Contextual Bandit Learning for Multi-Turn Code Generation
von: Chen, Ziru, et al.
Veröffentlicht: (2026)
von: Chen, Ziru, et al.
Veröffentlicht: (2026)
Instruction-Tuned, but Not More Verifiable Instruction-Following: A Cross-Task Diagnosis for LoRA Adapters
von: Zou, Junyi
Veröffentlicht: (2026)
von: Zou, Junyi
Veröffentlicht: (2026)
Theory on Mixture-of-Experts in Continual Learning
von: Li, Hongbo, et al.
Veröffentlicht: (2024)
von: Li, Hongbo, et al.
Veröffentlicht: (2024)
Conifer: Improving Complex Constrained Instruction-Following Ability of Large Language Models
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
Near-Optimal Partially Observable Reinforcement Learning with Partial Online State Information
von: Shi, Ming, et al.
Veröffentlicht: (2023)
von: Shi, Ming, et al.
Veröffentlicht: (2023)
IH-Challenge: A Training Dataset to Improve Instruction Hierarchy on Frontier LLMs
von: Guo, Chuan, et al.
Veröffentlicht: (2026)
von: Guo, Chuan, et al.
Veröffentlicht: (2026)
Teaching LLMs for Step-Level Automatic Math Correction via Reinforcement Learning
von: Li, Junsong, et al.
Veröffentlicht: (2025)
von: Li, Junsong, et al.
Veröffentlicht: (2025)
Language Models as Hierarchy Encoders
von: He, Yuan, et al.
Veröffentlicht: (2024)
von: He, Yuan, et al.
Veröffentlicht: (2024)
Replay Failures as Successes: Sample-Efficient Reinforcement Learning for Instruction Following
von: Zhang, Kongcheng, et al.
Veröffentlicht: (2025)
von: Zhang, Kongcheng, et al.
Veröffentlicht: (2025)
SUPERNOVA: Eliciting General Reasoning in LLMs with Reinforcement Learning on Natural Instructions
von: Suvarna, Ashima, et al.
Veröffentlicht: (2026)
von: Suvarna, Ashima, et al.
Veröffentlicht: (2026)
Upcycling Instruction Tuning from Dense to Mixture-of-Experts via Parameter Merging
von: Hui, Tingfeng, et al.
Veröffentlicht: (2024)
von: Hui, Tingfeng, et al.
Veröffentlicht: (2024)
Incentivizing Agentic Reasoning in LLM Judges via Tool-Integrated Reinforcement Learning
von: Xu, Ran, et al.
Veröffentlicht: (2025)
von: Xu, Ran, et al.
Veröffentlicht: (2025)
Sparse-RL: Breaking the Memory Wall in LLM Reinforcement Learning via Stable Sparse Rollouts
von: Luo, Sijia, et al.
Veröffentlicht: (2026)
von: Luo, Sijia, et al.
Veröffentlicht: (2026)
KARL: Mitigating Hallucinations in LLMs via Knowledge-Boundary-Aware Reinforcement Learning
von: Gao, Cheng, et al.
Veröffentlicht: (2026)
von: Gao, Cheng, et al.
Veröffentlicht: (2026)
Interactive Dialogue Agents via Reinforcement Learning on Hindsight Regenerations
von: Hong, Joey, et al.
Veröffentlicht: (2024)
von: Hong, Joey, et al.
Veröffentlicht: (2024)
Teaching Language Models to Critique via Reinforcement Learning
von: Xie, Zhihui, et al.
Veröffentlicht: (2025)
von: Xie, Zhihui, et al.
Veröffentlicht: (2025)
What is it for a Machine Learning Model to Have a Capability?
von: Harding, Jacqueline, et al.
Veröffentlicht: (2024)
von: Harding, Jacqueline, et al.
Veröffentlicht: (2024)
Light-IF: Endowing LLMs with Generalizable Reasoning via Preview and Self-Checking for Complex Instruction Following
von: Wang, Chenyang, et al.
Veröffentlicht: (2025)
von: Wang, Chenyang, et al.
Veröffentlicht: (2025)
Stabilizing MoE Reinforcement Learning by Aligning Training and Inference Routers
von: Ma, Wenhan, et al.
Veröffentlicht: (2025)
von: Ma, Wenhan, et al.
Veröffentlicht: (2025)
Scaling Multimodal Search and Recommendation with Small Language Models via Upside-Down Reinforcement Learning
von: Lin, Yu-Chen, et al.
Veröffentlicht: (2025)
von: Lin, Yu-Chen, et al.
Veröffentlicht: (2025)
Contrastive Reasoning Alignment: Reinforcement Learning from Hidden Representations
von: Luo, Haozheng, et al.
Veröffentlicht: (2026)
von: Luo, Haozheng, et al.
Veröffentlicht: (2026)
Self-Guided Process Reward Optimization with Redefined Step-wise Advantage for Process Reinforcement Learning
von: Fei, Wu, et al.
Veröffentlicht: (2025)
von: Fei, Wu, et al.
Veröffentlicht: (2025)
Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning
von: Shi, Chengshuai, et al.
Veröffentlicht: (2026)
von: Shi, Chengshuai, et al.
Veröffentlicht: (2026)
Many-Tier Instruction Hierarchy in LLM Agents
von: Zhang, Jingyu, et al.
Veröffentlicht: (2026)
von: Zhang, Jingyu, et al.
Veröffentlicht: (2026)
Read and Reap the Rewards: Learning to Play Atari with the Help of Instruction Manuals
von: Wu, Yue, et al.
Veröffentlicht: (2023)
von: Wu, Yue, et al.
Veröffentlicht: (2023)
Natural Language Reinforcement Learning
von: Feng, Xidong, et al.
Veröffentlicht: (2024)
von: Feng, Xidong, et al.
Veröffentlicht: (2024)
AttriLens-Mol: Attribute Guided Reinforcement Learning for Molecular Property Prediction with Large Language Models
von: Lin, Xuan, et al.
Veröffentlicht: (2025)
von: Lin, Xuan, et al.
Veröffentlicht: (2025)
WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
von: Luo, Haipeng, et al.
Veröffentlicht: (2023)
von: Luo, Haipeng, et al.
Veröffentlicht: (2023)
Stabilizing Reinforcement Learning with LLMs: Formulation and Practices
von: Zheng, Chujie, et al.
Veröffentlicht: (2025)
von: Zheng, Chujie, et al.
Veröffentlicht: (2025)
TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
von: Wei, Zhepei, et al.
Veröffentlicht: (2025)
von: Wei, Zhepei, et al.
Veröffentlicht: (2025)
From $P(y|x)$ to $P(y)$: Investigating Reinforcement Learning in Pre-train Space
von: Tan, Yuqiao, et al.
Veröffentlicht: (2026)
von: Tan, Yuqiao, et al.
Veröffentlicht: (2026)
Towards Safe Reinforcement Learning via Constraining Conditional Value-at-Risk
von: Ying, Chengyang, et al.
Veröffentlicht: (2022)
von: Ying, Chengyang, et al.
Veröffentlicht: (2022)
Instructions are all you need: Self-supervised Reinforcement Learning for Instruction Following
von: Ren, Qingyu, et al.
Veröffentlicht: (2025)
von: Ren, Qingyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Towards Fast Safe Online Reinforcement Learning via Policy Finetuning
von: Chen, Keru, et al.
Veröffentlicht: (2024) -
Why Adam Can Beat SGD: Second-Moment Normalization Yields Sharper Tails
von: Jin, Ruinan, et al.
Veröffentlicht: (2026) -
Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification
von: Guo, Yiju, et al.
Veröffentlicht: (2026) -
Instructional Segment Embedding: Improving LLM Safety with Instruction Hierarchy
von: Wu, Tong, et al.
Veröffentlicht: (2024) -
Agentic Transformers Provably Learn to Search via Reinforcement Learning
von: Yang, Tong, et al.
Veröffentlicht: (2026)