RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Zihan, Wang, Kangrui, Wang, Qineng, Zhang, Pingyue, Li, Linjie, Yang, Zhengyuan, Jin, Xing, Yu, Kefan, Nguyen, Minh Nhat, Liu, Licheng, Gottlieb, Eli, Lu, Yiping, Cho, Kyunghyun, Wu, Jiajun, Fei-Fei, Li, Wang, Lijuan, Choi, Yejin, Li, Manling |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RAGEN-2: Reasoning Collapse in Agentic RL
von: Wang, Zihan, et al.
Veröffentlicht: (2026)
von: Wang, Zihan, et al.
Veröffentlicht: (2026)
VAGEN: Reinforcing World Model Reasoning for Multi-Turn VLM Agents
von: Wang, Kangrui, et al.
Veröffentlicht: (2025)
von: Wang, Kangrui, et al.
Veröffentlicht: (2025)
Planning with the Views via Scene Self-Exploration
von: Wang, Kangrui, et al.
Veröffentlicht: (2026)
von: Wang, Kangrui, et al.
Veröffentlicht: (2026)
Theory of Space: Can Foundation Models Construct Spatial Beliefs through Active Exploration?
von: Zhang, Pingyue, et al.
Veröffentlicht: (2026)
von: Zhang, Pingyue, et al.
Veröffentlicht: (2026)
MindCube: Spatial Mental Modeling from Limited Views
von: Wang, Qineng, et al.
Veröffentlicht: (2025)
von: Wang, Qineng, et al.
Veröffentlicht: (2025)
A Simple "Try Again" Can Elicit Multi-Turn LLM Reasoning
von: Liu, Licheng, et al.
Veröffentlicht: (2025)
von: Liu, Licheng, et al.
Veröffentlicht: (2025)
Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making
von: Li, Manling, et al.
Veröffentlicht: (2024)
von: Li, Manling, et al.
Veröffentlicht: (2024)
Beyond Words: Advancing Long-Text Image Generation via Multimodal Autoregressive Models
von: Wang, Alex Jinpeng, et al.
Veröffentlicht: (2025)
von: Wang, Alex Jinpeng, et al.
Veröffentlicht: (2025)
Diagnostic Benchmark and Iterative Inpainting for Layout-Guided Image Generation
von: Cho, Jaemin, et al.
Veröffentlicht: (2023)
von: Cho, Jaemin, et al.
Veröffentlicht: (2023)
Learning from Trials and Errors: Reflective Test-Time Planning for Embodied LLMs
von: Hong, Yining, et al.
Veröffentlicht: (2026)
von: Hong, Yining, et al.
Veröffentlicht: (2026)
ENACT: Evaluating Embodied Cognition with World Modeling of Egocentric Interaction
von: Wang, Qineng, et al.
Veröffentlicht: (2025)
von: Wang, Qineng, et al.
Veröffentlicht: (2025)
ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop
von: Hong, Yining, et al.
Veröffentlicht: (2026)
von: Hong, Yining, et al.
Veröffentlicht: (2026)
Bring Metric Functions into Diffusion Models
von: An, Jie, et al.
Veröffentlicht: (2024)
von: An, Jie, et al.
Veröffentlicht: (2024)
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
von: Hao, Yunzhuo, et al.
Veröffentlicht: (2025)
von: Hao, Yunzhuo, et al.
Veröffentlicht: (2025)
LiVOS: Light Video Object Segmentation with Gated Linear Matching
von: Liu, Qin, et al.
Veröffentlicht: (2024)
von: Liu, Qin, et al.
Veröffentlicht: (2024)
ImageGen-CoT: Enhancing Text-to-Image In-context Learning with Chain-of-Thought Reasoning
von: Liao, Jiaqi, et al.
Veröffentlicht: (2025)
von: Liao, Jiaqi, et al.
Veröffentlicht: (2025)
MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
von: Yu, Weihao, et al.
Veröffentlicht: (2023)
von: Yu, Weihao, et al.
Veröffentlicht: (2023)
Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation
von: Yang, Zhengyuan, et al.
Veröffentlicht: (2023)
von: Yang, Zhengyuan, et al.
Veröffentlicht: (2023)
Entity6K: A Large Open-Domain Evaluation Dataset for Real-World Entity Recognition
von: Qiu, Jielin, et al.
Veröffentlicht: (2024)
von: Qiu, Jielin, et al.
Veröffentlicht: (2024)
Chain-of-Experts: Unlocking the Communication Power of Mixture-of-Experts Models
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
ODESteer: A Unified ODE-Based Steering Framework for LLM Alignment
von: Zhao, Hongjue, et al.
Veröffentlicht: (2026)
von: Zhao, Hongjue, et al.
Veröffentlicht: (2026)
Non-convolutional Graph Neural Networks
von: Wang, Yuanqing, et al.
Veröffentlicht: (2024)
von: Wang, Yuanqing, et al.
Veröffentlicht: (2024)
SENTINEL: A Multi-Level Formal Framework for Safety Evaluation of Foundation Model-based Embodied Agents
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2025)
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2025)
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
von: Yang, Rui, et al.
Veröffentlicht: (2025)
von: Yang, Rui, et al.
Veröffentlicht: (2025)
Point-RFT: Improving Multimodal Reasoning with Visually Grounded Reinforcement Finetuning
von: Ni, Minheng, et al.
Veröffentlicht: (2025)
von: Ni, Minheng, et al.
Veröffentlicht: (2025)
Glance: Accelerating Diffusion Models with 1 Sample
von: Dong, Zhuobai, et al.
Veröffentlicht: (2025)
von: Dong, Zhuobai, et al.
Veröffentlicht: (2025)
V-MAGE: A Game Evaluation Framework for Assessing Vision-Centric Capabilities in Multimodal Large Language Models
von: Zheng, Xiangxi, et al.
Veröffentlicht: (2025)
von: Zheng, Xiangxi, et al.
Veröffentlicht: (2025)
TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering
von: Mao, Dongxing, et al.
Veröffentlicht: (2026)
von: Mao, Dongxing, et al.
Veröffentlicht: (2026)
COSMO: COntrastive Streamlined MultimOdal Model with Interleaved Pre-Training
von: Wang, Alex Jinpeng, et al.
Veröffentlicht: (2024)
von: Wang, Alex Jinpeng, et al.
Veröffentlicht: (2024)
T*: Re-thinking Temporal Search for Long-Form Video Understanding
von: Ye, Jinhui, et al.
Veröffentlicht: (2025)
von: Ye, Jinhui, et al.
Veröffentlicht: (2025)
EdiVal-Agent: An Object-Centric Framework for Automated, Fine-Grained Evaluation of Multi-Turn Editing
von: Chen, Tianyu, et al.
Veröffentlicht: (2025)
von: Chen, Tianyu, et al.
Veröffentlicht: (2025)
Computer-Use Agents as Judges for Generative User Interface
von: Lin, Kevin Qinghong, et al.
Veröffentlicht: (2025)
von: Lin, Kevin Qinghong, et al.
Veröffentlicht: (2025)
Zero-Shot Audio-Visual Editing via Cross-Modal Delta Denoising
von: Lin, Yan-Bo, et al.
Veröffentlicht: (2025)
von: Lin, Yan-Bo, et al.
Veröffentlicht: (2025)
Self‐Reconstruction of Dual‐Morphology Copper‐Iron Selenides for Cost‐Effective Oxygen Evolution Toward Industrial Alkaline Water Splitting
von: Jiajun Wang, et al.
Veröffentlicht: (2025)
von: Jiajun Wang, et al.
Veröffentlicht: (2025)
Gradient-Boosted Pseudo-Weighting: Methods for Population Inference from Nonprobability samples
von: Liu, Kangrui, et al.
Veröffentlicht: (2025)
von: Liu, Kangrui, et al.
Veröffentlicht: (2025)
SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement
von: Wang, Xiyao, et al.
Veröffentlicht: (2025)
von: Wang, Xiyao, et al.
Veröffentlicht: (2025)
Motion Consistency Model: Accelerating Video Diffusion with Disentangled Motion-Appearance Distillation
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
Scaling Inference-Time Search with Vision Value Model for Improved Visual Comprehension
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
DisCo: Disentangled Control for Realistic Human Dance Generation
von: Wang, Tan, et al.
Veröffentlicht: (2023)
von: Wang, Tan, et al.
Veröffentlicht: (2023)
MM-Vet v2: A Challenging Benchmark to Evaluate Large Multimodal Models for Integrated Capabilities
von: Yu, Weihao, et al.
Veröffentlicht: (2024)
von: Yu, Weihao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
RAGEN-2: Reasoning Collapse in Agentic RL
von: Wang, Zihan, et al.
Veröffentlicht: (2026) -
VAGEN: Reinforcing World Model Reasoning for Multi-Turn VLM Agents
von: Wang, Kangrui, et al.
Veröffentlicht: (2025) -
Planning with the Views via Scene Self-Exploration
von: Wang, Kangrui, et al.
Veröffentlicht: (2026) -
Theory of Space: Can Foundation Models Construct Spatial Beliefs through Active Exploration?
von: Zhang, Pingyue, et al.
Veröffentlicht: (2026) -
MindCube: Spatial Mental Modeling from Limited Views
von: Wang, Qineng, et al.
Veröffentlicht: (2025)