StraTA: Incentivizing Agentic Reinforcement Learning with Strategic Trajectory Abstraction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xue, Xiangyuan, Zhou, Yifan, Wang, Zidong, Tang, Shengji, Torr, Philip, Ouyang, Wanli, Bai, Lei, Yin, Zhenfei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CoMAS: Co-Evolving Multi-Agent Systems via Interaction Rewards
von: Xue, Xiangyuan, et al.
Veröffentlicht: (2025)
von: Xue, Xiangyuan, et al.
Veröffentlicht: (2025)
ComfyBench: Benchmarking LLM-based Agents in ComfyUI for Autonomously Designing Collaborative AI Systems
von: Xue, Xiangyuan, et al.
Veröffentlicht: (2024)
von: Xue, Xiangyuan, et al.
Veröffentlicht: (2024)
StraGo: Harnessing Strategic Guidance for Prompt Optimization
von: Wu, Yurong, et al.
Veröffentlicht: (2024)
von: Wu, Yurong, et al.
Veröffentlicht: (2024)
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
VIKI-R: Coordinating Embodied Multi-Agent Cooperation via Reinforcement Learning
von: Kang, Li, et al.
Veröffentlicht: (2025)
von: Kang, Li, et al.
Veröffentlicht: (2025)
Scaling Behaviors of LLM Reinforcement Learning Post-Training: An Empirical Study in Mathematical Reasoning
von: Tan, Zelin, et al.
Veröffentlicht: (2025)
von: Tan, Zelin, et al.
Veröffentlicht: (2025)
CTTS: Collective Test-Time Scaling
von: Song, Zhende, et al.
Veröffentlicht: (2025)
von: Song, Zhende, et al.
Veröffentlicht: (2025)
Native-Resolution Image Synthesis
von: Wang, Zidong, et al.
Veröffentlicht: (2025)
von: Wang, Zidong, et al.
Veröffentlicht: (2025)
ReCrit: Transition-Aware Reinforcement Learning for Scientific Critic Reasoning
von: Xu, Wanghan, et al.
Veröffentlicht: (2026)
von: Xu, Wanghan, et al.
Veröffentlicht: (2026)
Vision-DeepResearch: Incentivizing DeepResearch Capability in Multimodal Large Language Models
von: Huang, Wenxuan, et al.
Veröffentlicht: (2026)
von: Huang, Wenxuan, et al.
Veröffentlicht: (2026)
AI-Driven Automation Can Become the Foundation of Next-Era Science of Science Research
von: Chen, Renqi, et al.
Veröffentlicht: (2025)
von: Chen, Renqi, et al.
Veröffentlicht: (2025)
ReSo: A Reward-driven Self-organizing LLM-based Multi-Agent System for Reasoning Tasks
von: Zhou, Heng, et al.
Veröffentlicht: (2025)
von: Zhou, Heng, et al.
Veröffentlicht: (2025)
Transition Models: Rethinking the Generative Learning Objective
von: Wang, Zidong, et al.
Veröffentlicht: (2025)
von: Wang, Zidong, et al.
Veröffentlicht: (2025)
Small Model as Master Orchestrator: Learning Unified Agent-Tool Orchestration with Parallel Subtask Decomposition
von: Yuan, Wenzhen, et al.
Veröffentlicht: (2026)
von: Yuan, Wenzhen, et al.
Veröffentlicht: (2026)
HiSplat: Hierarchical 3D Gaussian Splatting for Generalizable Sparse-View Reconstruction
von: Tang, Shengji, et al.
Veröffentlicht: (2024)
von: Tang, Shengji, et al.
Veröffentlicht: (2024)
Strategic Exploitation in LLM Agent Markets: A Simulation Framework for E-Commerce Trust
von: Lei, Shijun, et al.
Veröffentlicht: (2026)
von: Lei, Shijun, et al.
Veröffentlicht: (2026)
Ego to World: Collaborative Spatial Reasoning in Embodied Systems via Reinforcement Learning
von: Zhou, Heng, et al.
Veröffentlicht: (2026)
von: Zhou, Heng, et al.
Veröffentlicht: (2026)
Many Heads Are Better Than One: Improved Scientific Idea Generation by A LLM-Based Multi-Agent System
von: Su, Haoyang, et al.
Veröffentlicht: (2024)
von: Su, Haoyang, et al.
Veröffentlicht: (2024)
FiT: Flexible Vision Transformer for Diffusion Model
von: Lu, Zeyu, et al.
Veröffentlicht: (2024)
von: Lu, Zeyu, et al.
Veröffentlicht: (2024)
Diffusion Models Need Visual Priors for Image Generation
von: Yue, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Yue, Xiaoyu, et al.
Veröffentlicht: (2024)
Understand Before You Generate: Self-Guided Training for Autoregressive Image Generation
von: Yue, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Yue, Xiaoyu, et al.
Veröffentlicht: (2025)
A Scalable Multi-LLM Collaboration System with Retrieval-based Selection and Exploration-Exploitation-Driven Enhancement
von: Tang, Shengji, et al.
Veröffentlicht: (2025)
von: Tang, Shengji, et al.
Veröffentlicht: (2025)
Charting Empirical Laws for LLM Fine-Tuning in Scientific Multi-Discipline Learning
von: Wang, Lintao, et al.
Veröffentlicht: (2026)
von: Wang, Lintao, et al.
Veröffentlicht: (2026)
Hierarchical Reinforcement Learning for Temporal Abstraction of Listwise Recommendation
von: Ji, Luo, et al.
Veröffentlicht: (2024)
von: Ji, Luo, et al.
Veröffentlicht: (2024)
EWE: An Agentic Framework for Extreme Weather Analysis
von: Jiang, Zhe, et al.
Veröffentlicht: (2025)
von: Jiang, Zhe, et al.
Veröffentlicht: (2025)
Exploring Representation-Aligned Latent Space for Better Generation
von: Xu, Wanghan, et al.
Veröffentlicht: (2025)
von: Xu, Wanghan, et al.
Veröffentlicht: (2025)
Contextual Bilevel Reinforcement Learning for Incentive Alignment
von: Thoma, Vinzenz, et al.
Veröffentlicht: (2024)
von: Thoma, Vinzenz, et al.
Veröffentlicht: (2024)
ReVSeg: Incentivizing the Reasoning Chain for Video Segmentation with Reinforcement Learning
von: Li, Yifan, et al.
Veröffentlicht: (2025)
von: Li, Yifan, et al.
Veröffentlicht: (2025)
Self-consistent Validation for Machine Learning Electronic Structure
von: Hu, Gengyuan, et al.
Veröffentlicht: (2024)
von: Hu, Gengyuan, et al.
Veröffentlicht: (2024)
Concept and Development of Metal‐Framework Nucleic Acids
von: Li Sun, et al.
Veröffentlicht: (2025)
von: Li Sun, et al.
Veröffentlicht: (2025)
DetToolChain: A New Prompting Paradigm to Unleash Detection Ability of MLLM
von: Wu, Yixuan, et al.
Veröffentlicht: (2024)
von: Wu, Yixuan, et al.
Veröffentlicht: (2024)
Communications-Incentivized Collaborative Reasoning in NetGPT through Agentic Reinforcement Learning
von: Yu, Xiaoxue, et al.
Veröffentlicht: (2026)
von: Yu, Xiaoxue, et al.
Veröffentlicht: (2026)
Incentivizing Agentic Reasoning in LLM Judges via Tool-Integrated Reinforcement Learning
von: Xu, Ran, et al.
Veröffentlicht: (2025)
von: Xu, Ran, et al.
Veröffentlicht: (2025)
Wisdom of the Crowd: Reinforcement Learning from Coevolutionary Collective Feedback
von: Yuan, Wenzhen, et al.
Veröffentlicht: (2025)
von: Yuan, Wenzhen, et al.
Veröffentlicht: (2025)
Beyond Gemini-3-Pro: Revisiting LLM Routing and Aggregation at Scale
von: Tang, Shengji, et al.
Veröffentlicht: (2026)
von: Tang, Shengji, et al.
Veröffentlicht: (2026)
StraTyper: Automated Semantic Type Discovery and Multi-Type Annotation for Dataset Collections
von: Koutras, Christos, et al.
Veröffentlicht: (2026)
von: Koutras, Christos, et al.
Veröffentlicht: (2026)
FengWu-4DVar: Coupling the Data-driven Weather Forecasting Model with 4D Variational Assimilation
von: Xiao, Yi, et al.
Veröffentlicht: (2023)
von: Xiao, Yi, et al.
Veröffentlicht: (2023)
Revealing Decurve Flows for Generalized Graph Propagation
von: Lin, Chen, et al.
Veröffentlicht: (2024)
von: Lin, Chen, et al.
Veröffentlicht: (2024)
Learning to Lead: Incentivizing Strategic Agents in the Dark
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
von: Wu, Yuchen, et al.
Veröffentlicht: (2025)
Counterbalancing Learning and Strategic Incentives in Allocation Markets
von: Ashlagi, Itai, et al.
Veröffentlicht: (2021)
von: Ashlagi, Itai, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
CoMAS: Co-Evolving Multi-Agent Systems via Interaction Rewards
von: Xue, Xiangyuan, et al.
Veröffentlicht: (2025) -
ComfyBench: Benchmarking LLM-based Agents in ComfyUI for Autonomously Designing Collaborative AI Systems
von: Xue, Xiangyuan, et al.
Veröffentlicht: (2024) -
StraGo: Harnessing Strategic Guidance for Prompt Optimization
von: Wu, Yurong, et al.
Veröffentlicht: (2024) -
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey
von: Zhang, Guibin, et al.
Veröffentlicht: (2025) -
VIKI-R: Coordinating Embodied Multi-Agent Cooperation via Reinforcement Learning
von: Kang, Li, et al.
Veröffentlicht: (2025)