Saved in:
| Main Authors: | Prakash, Jatin, Buvanesh, Anirudh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2510.03971 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What LLMs Think When You Don't Tell Them What to Think About?
by: Kwon, Yongchan, et al.
Published: (2026)
by: Kwon, Yongchan, et al.
Published: (2026)
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
by: Park, Young-Jin, et al.
Published: (2025)
by: Park, Young-Jin, et al.
Published: (2025)
Can You Break RLVER? Probing Adversarial Robustness of RL-Trained Empathetic Agents
by: K, Deeraj S, et al.
Published: (2026)
by: K, Deeraj S, et al.
Published: (2026)
Optimisation Is Not What You Need
by: Ibias, Alfredo
Published: (2025)
by: Ibias, Alfredo
Published: (2025)
Discovering What You Can Control: Interventional Boundary Discovery for Reinforcement Learning
by: Liu, Jiaxin, et al.
Published: (2026)
by: Liu, Jiaxin, et al.
Published: (2026)
Why Pool When You Can Flow? Active Learning with GFlowNets
by: Zhang, Renfei, et al.
Published: (2025)
by: Zhang, Renfei, et al.
Published: (2025)
A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals
by: Liu, Grace, et al.
Published: (2024)
by: Liu, Grace, et al.
Published: (2024)
Attention Is Not What You Need
by: Chong, Zhang
Published: (2025)
by: Chong, Zhang
Published: (2025)
What You See is What You Classify: Black Box Attributions
by: Stalder, Steven, et al.
Published: (2022)
by: Stalder, Steven, et al.
Published: (2022)
What You See Is Not Always What You Get: Evaluating GPT's Comprehension of Source Code
by: Wen, Jiawen, et al.
Published: (2024)
by: Wen, Jiawen, et al.
Published: (2024)
More Compute Is What You Need
by: Guo, Zhen
Published: (2024)
by: Guo, Zhen
Published: (2024)
What Would You Ask When You First Saw $a^2+b^2=c^2$? Evaluating LLM on Curiosity-Driven Questioning
by: Javaji, Shashidhar Reddy, et al.
Published: (2024)
by: Javaji, Shashidhar Reddy, et al.
Published: (2024)
Plan Before You Trade: Inference-Time Optimization for RL Trading Agents
by: Go, Eun, et al.
Published: (2026)
by: Go, Eun, et al.
Published: (2026)
Synthetic Data RL: Task Definition Is All You Need
by: Guo, Yiduo, et al.
Published: (2025)
by: Guo, Yiduo, et al.
Published: (2025)
Machine Unlearning Doesn't Do What You Think: Lessons for Generative AI Policy and Research
by: Cooper, A. Feder, et al.
Published: (2024)
by: Cooper, A. Feder, et al.
Published: (2024)
It's Not You, It's Clipping: A Soft Trust-Region via Probability Smoothing for LLM RL
by: Dwyer, Madeleine, et al.
Published: (2025)
by: Dwyer, Madeleine, et al.
Published: (2025)
You Are What You Train: Effects of Data Composition on Training Context-aware Machine Translation Models
by: Mąka, Paweł, et al.
Published: (2025)
by: Mąka, Paweł, et al.
Published: (2025)
You Need Reasoning to Learn Reasoning: The Limitations of Label-Free RL in Weak Base Models
by: Roy, Shuvendu, et al.
Published: (2025)
by: Roy, Shuvendu, et al.
Published: (2025)
Zero-Shot LLMs in Human-in-the-Loop RL: Replacing Human Feedback for Reward Shaping
by: Nazir, Mohammad Saif, et al.
Published: (2025)
by: Nazir, Mohammad Saif, et al.
Published: (2025)
xAI-Drop: Don't Use What You Cannot Explain
by: De Luca, Vincenzo Marco, et al.
Published: (2024)
by: De Luca, Vincenzo Marco, et al.
Published: (2024)
Have You Poisoned My Data? Defending Neural Networks against Data Poisoning
by: De Gaspari, Fabio, et al.
Published: (2024)
by: De Gaspari, Fabio, et al.
Published: (2024)
Do What You Say: Steering Vision-Language-Action Models via Runtime Reasoning-Action Alignment Verification
by: Wu, Yilin, et al.
Published: (2025)
by: Wu, Yilin, et al.
Published: (2025)
Attention is All You Need Until You Need Retention
by: Yaslioglu, M. Murat
Published: (2025)
by: Yaslioglu, M. Murat
Published: (2025)
Towards Improving Reward Design in RL: A Reward Alignment Metric for RL Practitioners
by: Muslimani, Calarina, et al.
Published: (2025)
by: Muslimani, Calarina, et al.
Published: (2025)
Know What You Don't Know: Selective Prediction for Early Exit DNNs
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
by: Bajpai, Divya Jyoti, et al.
Published: (2025)
When Errors Can Be Beneficial: A Categorization of Imperfect Rewards for Policy Gradient
by: Shang, Shuning, et al.
Published: (2026)
by: Shang, Shuning, et al.
Published: (2026)
You are out of context!
by: Cobino, Giancarlo, et al.
Published: (2024)
by: Cobino, Giancarlo, et al.
Published: (2024)
Fake it till You Make it: Reward Modeling as Discriminative Prediction
by: Liu, Runtao, et al.
Published: (2025)
by: Liu, Runtao, et al.
Published: (2025)
Position: Model Collapse Does Not Mean What You Think
by: Schaeffer, Rylan, et al.
Published: (2025)
by: Schaeffer, Rylan, et al.
Published: (2025)
Think When You Need: Self-Adaptive Chain-of-Thought Learning
by: Yang, Junjie, et al.
Published: (2025)
by: Yang, Junjie, et al.
Published: (2025)
What You See is What You GAN: Rendering Every Pixel for High-Fidelity Geometry in 3D GANs
by: Trevithick, Alex, et al.
Published: (2024)
by: Trevithick, Alex, et al.
Published: (2024)
When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL
by: Wang, Youting, et al.
Published: (2026)
by: Wang, Youting, et al.
Published: (2026)
Your Code Agent Can Grow Alongside You with Structured Memory
by: Deng, Yi-Xuan, et al.
Published: (2026)
by: Deng, Yi-Xuan, et al.
Published: (2026)
Can LLMs Help You at Work? A Sandbox for Evaluating LLM Agents in Enterprise Environments
by: Vishwakarma, Harsh, et al.
Published: (2025)
by: Vishwakarma, Harsh, et al.
Published: (2025)
Seek and You Shall Fold
by: Sellam, Nadav Bojan, et al.
Published: (2025)
by: Sellam, Nadav Bojan, et al.
Published: (2025)
Context is All You Need
by: Delanois, Jean Erik, et al.
Published: (2026)
by: Delanois, Jean Erik, et al.
Published: (2026)
When Your AIs Deceive You: Challenges of Partial Observability in Reinforcement Learning from Human Feedback
by: Lang, Leon, et al.
Published: (2024)
by: Lang, Leon, et al.
Published: (2024)
Where You Go is Who You Are: Behavioral Theory-Guided LLMs for Inverse Reinforcement Learning
by: Sun, Yuran, et al.
Published: (2025)
by: Sun, Yuran, et al.
Published: (2025)
Governing What You Cannot Observe: Adaptive Runtime Governance for Autonomous AI Agents
by: Marin, German, et al.
Published: (2026)
by: Marin, German, et al.
Published: (2026)
Do Transformers Have the Ability for Periodicity Generalization?
by: Liu, Huanyu, et al.
Published: (2026)
by: Liu, Huanyu, et al.
Published: (2026)
Similar Items
-
What LLMs Think When You Don't Tell Them What to Think About?
by: Kwon, Yongchan, et al.
Published: (2026) -
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
by: Park, Young-Jin, et al.
Published: (2025) -
Can You Break RLVER? Probing Adversarial Robustness of RL-Trained Empathetic Agents
by: K, Deeraj S, et al.
Published: (2026) -
Optimisation Is Not What You Need
by: Ibias, Alfredo
Published: (2025) -
Discovering What You Can Control: Interventional Boundary Discovery for Reinforcement Learning
by: Liu, Jiaxin, et al.
Published: (2026)