Rethinking Agentic Reinforcement Learning In Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Cui, Fangming, Zhu, Ruixiao, Fang, Cheng, Li, Sunan, Li, Jiahong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Large Language Model-based Multi-Agent Framework for Analog Circuits' Sizing Relationships Extraction
by: Liu, Chengjie, et al.
Published: (2025)
by: Liu, Chengjie, et al.
Published: (2025)
Managing Escalation in Off-the-Shelf Large Language Models
by: Elbaum, Sebastian, et al.
Published: (2025)
by: Elbaum, Sebastian, et al.
Published: (2025)
Benchmark for Planning and Control with Large Language Model Agents: Blocksworld with Model Context Protocol
by: Jobs, Niklas, et al.
Published: (2025)
by: Jobs, Niklas, et al.
Published: (2025)
An Approach to Checking Correctness for Agentic Systems
by: Sheffler, Thomas J
Published: (2025)
by: Sheffler, Thomas J
Published: (2025)
Agentic AI for autonomous anomaly management in complex systems
by: Barenji, Reza Vatankhah, et al.
Published: (2025)
by: Barenji, Reza Vatankhah, et al.
Published: (2025)
Quantum Reinforcement Learning with Transformers for the Capacitated Vehicle Routing Problem
by: Andrés, Eva
Published: (2026)
by: Andrés, Eva
Published: (2026)
Iterative Prompting with Persuasion Skills in Jailbreaking Large Language Models
by: Ke, Shih-Wen, et al.
Published: (2025)
by: Ke, Shih-Wen, et al.
Published: (2025)
SafeDrive: Knowledge- and Data-Driven Risk-Sensitive Decision-Making for Autonomous Vehicles with Large Language Models
by: Zhou, Zhiyuan, et al.
Published: (2024)
by: Zhou, Zhiyuan, et al.
Published: (2024)
Building Trustworthy AI: Transparent AI Systems via Large Language Models, Ontologies, and Logical Reasoning (TranspNet)
by: Machot, Fadi Al, et al.
Published: (2024)
by: Machot, Fadi Al, et al.
Published: (2024)
Governed By Agents: A Survey On The Role Of Agentic AI In Future Computing Environments
by: Murad, Nauman Ali, et al.
Published: (2025)
by: Murad, Nauman Ali, et al.
Published: (2025)
Graph-Enhanced Deep Reinforcement Learning for Multi-Objective Unrelated Parallel Machine Scheduling
by: Soykan, Bulent, et al.
Published: (2026)
by: Soykan, Bulent, et al.
Published: (2026)
Objective Decoupling in Social Reinforcement Learning: Recovering Ground Truth from Sycophantic Majorities
by: Ghasemi, Majid, et al.
Published: (2026)
by: Ghasemi, Majid, et al.
Published: (2026)
Multimodal Large Language Model Driven Scenario Testing for Autonomous Vehicles
by: Lu, Qiujing, et al.
Published: (2024)
by: Lu, Qiujing, et al.
Published: (2024)
Agentic Workflow for Education: Concepts and Applications
by: Jiang, Yuan-Hao, et al.
Published: (2025)
by: Jiang, Yuan-Hao, et al.
Published: (2025)
Artificial Agency and Large Language Models
by: van Lier, Maud, et al.
Published: (2024)
by: van Lier, Maud, et al.
Published: (2024)
Auditing an Automatic Grading Model with deep Reinforcement Learning
by: Condor, Aubrey, et al.
Published: (2024)
by: Condor, Aubrey, et al.
Published: (2024)
Looking Forward: Challenges and Opportunities in Agentic AI Reliability
by: Xing, Liudong, et al.
Published: (2025)
by: Xing, Liudong, et al.
Published: (2025)
Are Large Language Models Reliable Argument Quality Annotators?
by: Mirzakhmedova, Nailia, et al.
Published: (2024)
by: Mirzakhmedova, Nailia, et al.
Published: (2024)
Can Large Language Models Act as Symbolic Reasoners?
by: Sullivan, Rob, et al.
Published: (2024)
by: Sullivan, Rob, et al.
Published: (2024)
Mind the Gap: How Elicitation Protocols Shape the Stated-Revealed Preference Gap in Language Models
by: Mahajan, Pranav, et al.
Published: (2026)
by: Mahajan, Pranav, et al.
Published: (2026)
Coordinating Ride-Pooling with Public Transit using Reward-Guided Conservative Q-Learning: An Offline Training and Online Fine-Tuning Reinforcement Learning Framework
by: Hu, Yulong, et al.
Published: (2025)
by: Hu, Yulong, et al.
Published: (2025)
One Step is Enough: Multi-Agent Reinforcement Learning based on One-Step Policy Optimization for Order Dispatch on Ride-Sharing Platforms
by: Zhao, Zijian, et al.
Published: (2025)
by: Zhao, Zijian, et al.
Published: (2025)
Automatic Qiskit Code Refactoring Using Large Language Models
by: Suárez, José Manuel, et al.
Published: (2025)
by: Suárez, José Manuel, et al.
Published: (2025)
CIRCUITSYNTH: Leveraging Large Language Models for Circuit Topology Synthesis
by: Vijayaraghavan, Prashanth, et al.
Published: (2024)
by: Vijayaraghavan, Prashanth, et al.
Published: (2024)
LSDTs: LLM-Augmented Semantic Digital Twins for Adaptive Knowledge-Intensive Infrastructure Planning
by: Li, Naiyi, et al.
Published: (2025)
by: Li, Naiyi, et al.
Published: (2025)
An Autonomous Network Orchestration Framework Integrating Large Language Models with Continual Reinforcement Learning
by: Shokrnezhad, Masoud, et al.
Published: (2025)
by: Shokrnezhad, Masoud, et al.
Published: (2025)
SpecPylot: Python Specification Generation using Large Language Models
by: Ayon, Ragib Shahariar, et al.
Published: (2026)
by: Ayon, Ragib Shahariar, et al.
Published: (2026)
Predictive Coding and Information Bottleneck for Hallucination Detection in Large Language Models
by: Bhatt, Manish
Published: (2026)
by: Bhatt, Manish
Published: (2026)
Circuit Partitioning Using Large Language Models for Quantum Compilation and Simulations
by: Sinha, Pranav, et al.
Published: (2025)
by: Sinha, Pranav, et al.
Published: (2025)
A Novel Nuanced Conversation Evaluation Framework for Large Language Models in Mental Health
by: Marrapese, Alexander, et al.
Published: (2024)
by: Marrapese, Alexander, et al.
Published: (2024)
LLM4DS: Evaluating Large Language Models for Data Science Code Generation
by: Nascimento, Nathalia, et al.
Published: (2024)
by: Nascimento, Nathalia, et al.
Published: (2024)
A Hybrid Quantum-Classical AI-Based Detection Strategy for Generative Adversarial Network-Based Deepfake Attacks on an Autonomous Vehicle Traffic Sign Classification System
by: Salek, M Sabbir, et al.
Published: (2024)
by: Salek, M Sabbir, et al.
Published: (2024)
KVNAND: Efficient On-Device Large Language Model Inference Using DRAM-Free In-Flash Computing
by: Deng, Lishuo, et al.
Published: (2025)
by: Deng, Lishuo, et al.
Published: (2025)
Towards Intelligent Transportation with Pedestrians and Vehicles In-the-Loop: A Surveillance Video-Assisted Federated Digital Twin Framework
by: Li, Xiaolong, et al.
Published: (2025)
by: Li, Xiaolong, et al.
Published: (2025)
A Collaborative PIM Computing Optimization Framework for Multi-Tenant DNN
by: Li, Bojing, et al.
Published: (2024)
by: Li, Bojing, et al.
Published: (2024)
MI9: An Integrated Runtime Governance Framework for Agentic AI
by: Wang, Charles L., et al.
Published: (2025)
by: Wang, Charles L., et al.
Published: (2025)
Towards a HIPAA Compliant Agentic AI System in Healthcare
by: Neupane, Subash, et al.
Published: (2025)
by: Neupane, Subash, et al.
Published: (2025)
VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications
by: Konenkov, Mikhail, et al.
Published: (2024)
by: Konenkov, Mikhail, et al.
Published: (2024)
Hybrid Action Based Reinforcement Learning for Multi-Objective Compatible Autonomous Driving
by: Jin, Guizhe, et al.
Published: (2025)
by: Jin, Guizhe, et al.
Published: (2025)
Physics-Informed Autonomous LLM Agents for Explainable Power Electronics Modulation Design
by: Liu, Junhua, et al.
Published: (2024)
by: Liu, Junhua, et al.
Published: (2024)
Similar Items
-
A Large Language Model-based Multi-Agent Framework for Analog Circuits' Sizing Relationships Extraction
by: Liu, Chengjie, et al.
Published: (2025) -
Managing Escalation in Off-the-Shelf Large Language Models
by: Elbaum, Sebastian, et al.
Published: (2025) -
Benchmark for Planning and Control with Large Language Model Agents: Blocksworld with Model Context Protocol
by: Jobs, Niklas, et al.
Published: (2025) -
An Approach to Checking Correctness for Agentic Systems
by: Sheffler, Thomas J
Published: (2025) -
Agentic AI for autonomous anomaly management in complex systems
by: Barenji, Reza Vatankhah, et al.
Published: (2025)