Model-Free RL Agents Demonstrate System 1-Like Intentionality
Fuente:
arXiv
Saved in:
| Main Authors: | Ashton, Hal, Franklin, Matija |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Preferences in AI Alignment
by: Zhi-Xuan, Tan, et al.
Published: (2024)
by: Zhi-Xuan, Tan, et al.
Published: (2024)
Intelligent AI Delegation
by: Tomašev, Nenad, et al.
Published: (2026)
by: Tomašev, Nenad, et al.
Published: (2026)
Virtual Agent Economies
by: Tomasev, Nenad, et al.
Published: (2025)
by: Tomasev, Nenad, et al.
Published: (2025)
Architecting Trust in Artificial Epistemic Agents
by: Marchal, Nahema, et al.
Published: (2026)
by: Marchal, Nahema, et al.
Published: (2026)
AI Governance through Markets
by: Tomei, Philip Moreira, et al.
Published: (2025)
by: Tomei, Philip Moreira, et al.
Published: (2025)
TWISTED-RL: Hierarchical Skilled Agents for Knot-Tying without Human Demonstrations
by: Freund, Guy, et al.
Published: (2026)
by: Freund, Guy, et al.
Published: (2026)
Distributional AGI Safety
by: Tomašev, Nenad, et al.
Published: (2025)
by: Tomašev, Nenad, et al.
Published: (2025)
Demonstration-Free Robotic Control via LLM Agents
by: Tsui, Brian Y., et al.
Published: (2026)
by: Tsui, Brian Y., et al.
Published: (2026)
A Necessary Step toward Faithfulness: Measuring and Improving Consistency in Free-Text Explanations
by: Zhao, Lingjun, et al.
Published: (2025)
by: Zhao, Lingjun, et al.
Published: (2025)
Learning Human-Like RL Agents Through Trajectory Optimization With Action Quantization
by: Guo, Jian-Ting, et al.
Published: (2025)
by: Guo, Jian-Ting, et al.
Published: (2025)
Language Models Exhibit Inconsistent Biases Towards Algorithmic Agents and Human Experts
by: Bo, Jessica Y., et al.
Published: (2026)
by: Bo, Jessica Y., et al.
Published: (2026)
SkyRL-Agent: Efficient RL Training for Multi-turn LLM Agent
by: Cao, Shiyi, et al.
Published: (2025)
by: Cao, Shiyi, et al.
Published: (2025)
Who's the Mole? Modeling and Detecting Intention-Hiding Malicious Agents in LLM-Based Multi-Agent Systems
by: Xie, Yizhe, et al.
Published: (2025)
by: Xie, Yizhe, et al.
Published: (2025)
Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
by: Mai, Xinji, et al.
Published: (2025)
by: Mai, Xinji, et al.
Published: (2025)
ProRL Agent: Rollout-as-a-Service for RL Training of Multi-Turn LLM Agents
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
Agent^2 RL-Bench: Can LLM Agents Engineer Agentic RL Post-Training?
by: Chen, Wanyi, et al.
Published: (2026)
by: Chen, Wanyi, et al.
Published: (2026)
Probabilistic Modeling of Intentions in Socially Intelligent LLM Agents
by: Xia, Feifan, et al.
Published: (2025)
by: Xia, Feifan, et al.
Published: (2025)
The Reasons that Agents Act: Intention and Instrumental Goals
by: Ward, Francis Rhys, et al.
Published: (2024)
by: Ward, Francis Rhys, et al.
Published: (2024)
Intentional Deception as Controllable Capability in LLM Agents
by: Starace, Jason, et al.
Published: (2026)
by: Starace, Jason, et al.
Published: (2026)
Intentionality is a Design Decision: Measuring Functional Intentionality for Accountable AI Systems
by: Chiappetta, Allessia, et al.
Published: (2026)
by: Chiappetta, Allessia, et al.
Published: (2026)
"Give Me an Example Like This": Episodic Active Reinforcement Learning from Demonstrations
by: Hou, Muhan, et al.
Published: (2024)
by: Hou, Muhan, et al.
Published: (2024)
ProCeedRL: Process Critic with Exploratory Demonstration Reinforcement Learning for LLM Agentic Reasoning
by: Gao, Jingyue, et al.
Published: (2026)
by: Gao, Jingyue, et al.
Published: (2026)
Instruction Agent: Enhancing Agent with Expert Demonstration
by: Li, Yinheng, et al.
Published: (2025)
by: Li, Yinheng, et al.
Published: (2025)
Mind the Model, Not the Agent: The Primacy Bias in Model-based RL
by: Qiao, Zhongjian, et al.
Published: (2023)
by: Qiao, Zhongjian, et al.
Published: (2023)
Deep RL Needs Deep Behavior Analysis: Exploring Implicit Planning by Model-Free Agents in Open-Ended Environments
by: Simmons-Edler, Riley, et al.
Published: (2025)
by: Simmons-Edler, Riley, et al.
Published: (2025)
Investigating Memory in Model-Free RL with POPGym Arcade
by: Wang, Zekang, et al.
Published: (2025)
by: Wang, Zekang, et al.
Published: (2025)
Reduced-Order Model-Guided Reinforcement Learning for Demonstration-Free Humanoid Locomotion
by: Liu, Shuai, et al.
Published: (2025)
by: Liu, Shuai, et al.
Published: (2025)
Crossing the Human-Robot Embodiment Gap with Sim-to-Real RL using One Human Demonstration
by: Lum, Tyler Ga Wei, et al.
Published: (2025)
by: Lum, Tyler Ga Wei, et al.
Published: (2025)
Intelligent Switching for Reset-Free RL
by: Patil, Darshan, et al.
Published: (2024)
by: Patil, Darshan, et al.
Published: (2024)
Towards Shutdownable Agents: Generalizing Stochastic Choice in RL Agents and LLMs
by: Cullen, Carissa, et al.
Published: (2026)
by: Cullen, Carissa, et al.
Published: (2026)
Defense Against the Dark Prompts: Mitigating Best-of-N Jailbreaking with Prompt Evaluation
by: Armstrong, Stuart, et al.
Published: (2025)
by: Armstrong, Stuart, et al.
Published: (2025)
AgentV-RL: Scaling Reward Modeling with Agentic Verifier
by: Zhang, Jiazheng, et al.
Published: (2026)
by: Zhang, Jiazheng, et al.
Published: (2026)
MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism
by: Liu, Shulin, et al.
Published: (2025)
by: Liu, Shulin, et al.
Published: (2025)
Position: Stop Acting Like Language Model Agents Are Normal Agents
by: Perrier, Elija, et al.
Published: (2025)
by: Perrier, Elija, et al.
Published: (2025)
Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL
by: Li, Weizhen, et al.
Published: (2025)
by: Li, Weizhen, et al.
Published: (2025)
IntentionESC: An Intention-Centered Framework for Enhancing Emotional Support in Dialogue Systems
by: Zhang, Xinjie, et al.
Published: (2025)
by: Zhang, Xinjie, et al.
Published: (2025)
Intention Knowledge Graph Construction for User Intention Relation Modeling
by: Bai, Jiaxin, et al.
Published: (2024)
by: Bai, Jiaxin, et al.
Published: (2024)
Can RL Improve Generalization of LLM Agents? An Empirical Study
by: Xi, Zhiheng, et al.
Published: (2026)
by: Xi, Zhiheng, et al.
Published: (2026)
Grounding Computer Use Agents on Human Demonstrations
by: Feizi, Aarash, et al.
Published: (2025)
by: Feizi, Aarash, et al.
Published: (2025)
Learning API Functionality from In-Context Demonstrations for Tool-based Agents
by: Patel, Bhrij, et al.
Published: (2025)
by: Patel, Bhrij, et al.
Published: (2025)
Similar Items
-
Beyond Preferences in AI Alignment
by: Zhi-Xuan, Tan, et al.
Published: (2024) -
Intelligent AI Delegation
by: Tomašev, Nenad, et al.
Published: (2026) -
Virtual Agent Economies
by: Tomasev, Nenad, et al.
Published: (2025) -
Architecting Trust in Artificial Epistemic Agents
by: Marchal, Nahema, et al.
Published: (2026) -
AI Governance through Markets
by: Tomei, Philip Moreira, et al.
Published: (2025)