Learning CLI Agents with Structured Action Credit under Selective Observation
Fuente:
arXiv
Saved in:
| Main Authors: | Su, Haoyang, Wen, Ying |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CLI-Gym: Scalable CLI Task Generation via Agentic Environment Inversion
by: Lin, Yusong, et al.
Published: (2026)
by: Lin, Yusong, et al.
Published: (2026)
CLI-RAG: A Retrieval-Augmented Framework for Clinically Structured and Context Aware Text Generation with LLMs
by: Keerthana, Garapati, et al.
Published: (2025)
by: Keerthana, Garapati, et al.
Published: (2025)
Reinforcing Language Agents via Policy Optimization with Action Decomposition
by: Wen, Muning, et al.
Published: (2024)
by: Wen, Muning, et al.
Published: (2024)
LAMMI-Pathology: A Tool-Centric Bottom-Up LVLM-Agent Framework for Molecularly Informed Medical Intelligence in Pathology
by: Su, Haoyang, et al.
Published: (2026)
by: Su, Haoyang, et al.
Published: (2026)
Hindsight Credit Assignment for Long-Horizon LLM Agents
by: Tan, Hui-Ze, et al.
Published: (2026)
by: Tan, Hui-Ze, et al.
Published: (2026)
Evaluating LLM-Based 0-to-1 Software Generation in End-to-End CLI Tool Scenarios
by: Hu, Ruida, et al.
Published: (2026)
by: Hu, Ruida, et al.
Published: (2026)
PMAT: Optimizing Action Generation Order in Multi-Agent Reinforcement Learning
by: Hu, Kun, et al.
Published: (2025)
by: Hu, Kun, et al.
Published: (2025)
EGL-SCA: Structural Credit Assignment for Co-Evolving Instructions and Tools in Graph Reasoning Agents
by: Yuan, Zike, et al.
Published: (2026)
by: Yuan, Zike, et al.
Published: (2026)
Multi-level Advantage Credit Assignment for Cooperative Multi-Agent Reinforcement Learning
by: Zhao, Xutong, et al.
Published: (2025)
by: Zhao, Xutong, et al.
Published: (2025)
Optimus-2: Multimodal Minecraft Agent with Goal-Observation-Action Conditioned Policy
by: Li, Zaijing, et al.
Published: (2025)
by: Li, Zaijing, et al.
Published: (2025)
SKILLC: Learning Autonomous Skill Internalization in LLM Agents via Contrastive Credit Assignment
by: Lin, Hongxiang, et al.
Published: (2026)
by: Lin, Hongxiang, et al.
Published: (2026)
Counterfactual Credit Policy Optimization for Multi-Agent Collaboration
by: Li, Zhongyi, et al.
Published: (2026)
by: Li, Zhongyi, et al.
Published: (2026)
Unifying Temporal and Structural Credit Assignment in LLM-Based Multi-Agent Prompt Optimization
by: Li, Wenwu, et al.
Published: (2026)
by: Li, Wenwu, et al.
Published: (2026)
Meta-learning how to Share Credit among Macro-Actions
by: Hosu, Ionel-Alexandru, et al.
Published: (2025)
by: Hosu, Ionel-Alexandru, et al.
Published: (2025)
Who Deserves the Reward? SHARP: Shapley Credit-based Optimization for Multi-Agent System
by: Li, Yanming, et al.
Published: (2026)
by: Li, Yanming, et al.
Published: (2026)
AGORA: Adapter-Grounded Observation-Action Retention for Inference-Free Prompt Compression in LLM Agents
by: Zhang, Haoran, et al.
Published: (2026)
by: Zhang, Haoran, et al.
Published: (2026)
Challenges in Credit Assignment for Multi-Agent Reinforcement Learning in Open Agent Systems
by: Abadi, Alireza Saleh, et al.
Published: (2025)
by: Abadi, Alireza Saleh, et al.
Published: (2025)
Empowering Large Language Model Agents through Action Learning
by: Zhao, Haiteng, et al.
Published: (2024)
by: Zhao, Haiteng, et al.
Published: (2024)
Think Twice, Act Once: Verifier-Guided Action Selection For Embodied Agents
by: Singhi, Nishad, et al.
Published: (2026)
by: Singhi, Nishad, et al.
Published: (2026)
From History to State: Constant-Context Skill Learning for LLM Agents
by: Xie, Haoyang, et al.
Published: (2026)
by: Xie, Haoyang, et al.
Published: (2026)
CAPF: Guiding Search-Agent Rollouts with Credit-Attenuated Privileged Feedback
by: Chen, Bin, et al.
Published: (2026)
by: Chen, Bin, et al.
Published: (2026)
Structured Debate Improves Corporate Credit Reasoning in Financial AI
by: Lee, Yoonjin, et al.
Published: (2025)
by: Lee, Yoonjin, et al.
Published: (2025)
SCAR: Shapley Credit Assignment for More Efficient RLHF
by: Cao, Meng, et al.
Published: (2025)
by: Cao, Meng, et al.
Published: (2025)
Agent models: Internalizing Chain-of-Action Generation into Reasoning models
by: Zhang, Yuxiang, et al.
Published: (2025)
by: Zhang, Yuxiang, et al.
Published: (2025)
The Emergence of Altruism in Large-Language-Model Agents Society
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
Beyond Syntax: Action Semantics Learning for App Agents
by: Tang, Bohan, et al.
Published: (2025)
by: Tang, Bohan, et al.
Published: (2025)
Exact Is Easier: Credit Assignment for Cooperative LLM Agents
by: Chen, Yanjun, et al.
Published: (2026)
by: Chen, Yanjun, et al.
Published: (2026)
Choosing How to Remember: Adaptive Memory Structures for LLM Agents
by: Lu, Mingfei, et al.
Published: (2026)
by: Lu, Mingfei, et al.
Published: (2026)
Spectral Invariant Learning for Dynamic Graphs under Distribution Shifts
by: Zhang, Zeyang, et al.
Published: (2024)
by: Zhang, Zeyang, et al.
Published: (2024)
CODESTRUCT: Code Agents over Structured Action Spaces
by: Kim, Myeongsoo, et al.
Published: (2026)
by: Kim, Myeongsoo, et al.
Published: (2026)
Economic Rationality under Specialization: Evidence of Decision Bias in AI Agents
by: Wen, ShuiDe
Published: (2025)
by: Wen, ShuiDe
Published: (2025)
Beyond Uniform Credit Assignment: Selective Eligibility Traces for RLVR
by: Mou, Chaoli, et al.
Published: (2026)
by: Mou, Chaoli, et al.
Published: (2026)
A Selective Learning Method for Temporal Graph Continual Learning
by: Liu, Hanmo, et al.
Published: (2025)
by: Liu, Hanmo, et al.
Published: (2025)
HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents
by: Peng, Jiangweizhi, et al.
Published: (2026)
by: Peng, Jiangweizhi, et al.
Published: (2026)
Leveraging Action Relational Structures for Integrated Learning and Planning
by: Wang, Ryan Xiao, et al.
Published: (2025)
by: Wang, Ryan Xiao, et al.
Published: (2025)
Proximity-Based Multi-Turn Optimization: Practical Credit Assignment for LLM Agent Training
by: Fang, Yangyi, et al.
Published: (2026)
by: Fang, Yangyi, et al.
Published: (2026)
Credit-Budgeted ICPC-Style Coding: When Agents Must Pay for Every Decision
by: Zhou, Lingfeng, et al.
Published: (2026)
by: Zhou, Lingfeng, et al.
Published: (2026)
CLIP-Motion: Learning Reward Functions for Robotic Actions Using Consecutive Observations
by: Dang, Xuzhe, et al.
Published: (2023)
by: Dang, Xuzhe, et al.
Published: (2023)
Engagement Process: Rethinking the Temporal Interface of Action and Observation
by: Li, Jialian, et al.
Published: (2026)
by: Li, Jialian, et al.
Published: (2026)
Success in Humanoid Reinforcement Learning under Partial Observation
by: Wang, Wuhao, et al.
Published: (2025)
by: Wang, Wuhao, et al.
Published: (2025)
Similar Items
-
CLI-Gym: Scalable CLI Task Generation via Agentic Environment Inversion
by: Lin, Yusong, et al.
Published: (2026) -
CLI-RAG: A Retrieval-Augmented Framework for Clinically Structured and Context Aware Text Generation with LLMs
by: Keerthana, Garapati, et al.
Published: (2025) -
Reinforcing Language Agents via Policy Optimization with Action Decomposition
by: Wen, Muning, et al.
Published: (2024) -
LAMMI-Pathology: A Tool-Centric Bottom-Up LVLM-Agent Framework for Molecularly Informed Medical Intelligence in Pathology
by: Su, Haoyang, et al.
Published: (2026) -
Hindsight Credit Assignment for Long-Horizon LLM Agents
by: Tan, Hui-Ze, et al.
Published: (2026)