FlowSteer: Towards Agents Designing Agentic Workflows via Reinforced Progressive Canvas Editing
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Mingda, Liu, Wenjin, Shen, Tiesunlong, Lin, Qika, Mao, Rui, Cambria, Erik, Tang, Xiaoying, Luo, Haoran |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
by: Tang, Wenjie, et al.
Published: (2026)
by: Tang, Wenjie, et al.
Published: (2026)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
by: Wang, Yuchen, et al.
Published: (2026)
by: Wang, Yuchen, et al.
Published: (2026)
Analysing Factorizations of Action-Value Networks for Cooperative Multi-Agent Reinforcement Learning
by: Castellini, Jacopo, et al.
Published: (2019)
by: Castellini, Jacopo, et al.
Published: (2019)
ME-IGM: Individual-Global-Max in Maximum Entropy Multi-Agent Reinforcement Learning
by: Chen, Wen-Tse, et al.
Published: (2024)
by: Chen, Wen-Tse, et al.
Published: (2024)
ChromaFlow: A Negative Ablation Study of Orchestration Overhead in Tool-Augmented Agent Evaluation
by: Mittal, Tarun
Published: (2026)
by: Mittal, Tarun
Published: (2026)
Robust and Diverse Multi-Agent Learning via Rational Policy Gradient
by: Lauffer, Niklas, et al.
Published: (2025)
by: Lauffer, Niklas, et al.
Published: (2025)
Latent Cache Flow: Model-to-Model Communication Without Text
by: Rossi, Maximillian, et al.
Published: (2026)
by: Rossi, Maximillian, et al.
Published: (2026)
Towards General Negotiation Strategies with End-to-End Reinforcement Learning
by: Renting, Bram M., et al.
Published: (2024)
by: Renting, Bram M., et al.
Published: (2024)
Advancing Multimodal Agent Reasoning with Long-Term Neuro-Symbolic Memory
by: Jiang, Rongjie, et al.
Published: (2026)
by: Jiang, Rongjie, et al.
Published: (2026)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
by: Costa, Rimom
Published: (2025)
by: Costa, Rimom
Published: (2025)
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
by: Hong, Yoosung
Published: (2026)
by: Hong, Yoosung
Published: (2026)
Dynamic Dual-Granularity Skill Bank for Agentic RL
by: Tu, Songjun, et al.
Published: (2026)
by: Tu, Songjun, et al.
Published: (2026)
Dynamic Attentional Context Scoping: Agent-Triggered Focus Sessions for Isolated Per-Agent Steering in Multi-Agent LLM Orchestration
by: Patel, Nickson
Published: (2026)
by: Patel, Nickson
Published: (2026)
N-Agent Ad Hoc Teamwork
by: Wang, Caroline, et al.
Published: (2024)
by: Wang, Caroline, et al.
Published: (2024)
The Stochastic Gap: A Markovian Framework for Pre-Deployment Reliability and Oversight-Cost Auditing in Agentic Artificial Intelligence
by: Pal, Biplab, et al.
Published: (2026)
by: Pal, Biplab, et al.
Published: (2026)
Procedural Game Level Design with Deep Reinforcement Learning
by: Özkan, Miraç Buğra
Published: (2025)
by: Özkan, Miraç Buğra
Published: (2025)
Learning To Help: Training Models to Assist Legacy Devices
by: Wu, Yu, et al.
Published: (2024)
by: Wu, Yu, et al.
Published: (2024)
StatePlane: A Cognitive State Plane for Long-Horizon AI Systems Under Bounded Context
by: Annapureddy, Sasank, et al.
Published: (2026)
by: Annapureddy, Sasank, et al.
Published: (2026)
Do We Always Need Query-Level Workflows? Rethinking Agentic Workflow Generation for Multi-Agent Systems
by: Wang, Zixu, et al.
Published: (2026)
by: Wang, Zixu, et al.
Published: (2026)
When Agents Disagree: The Selection Bottleneck in Multi-Agent LLM Pipelines
by: Maryanskyy, Artem
Published: (2026)
by: Maryanskyy, Artem
Published: (2026)
Decentralized Aerial Manipulation of a Cable-Suspended Load using Multi-Agent Reinforcement Learning
by: Zeng, Jack, et al.
Published: (2025)
by: Zeng, Jack, et al.
Published: (2025)
When Actions Disappear: Adversarial Action Removal in Self-Play Reinforcement Learning
by: Kujur, Arahan
Published: (2026)
by: Kujur, Arahan
Published: (2026)
Extending NGU to Multi-Agent RL: A Preliminary Study
by: Hernandez, Juan, et al.
Published: (2025)
by: Hernandez, Juan, et al.
Published: (2025)
Umwelt Engineering: Designing the Cognitive Worlds of Linguistic Agents
by: Jehu-Appiah, Rodney
Published: (2026)
by: Jehu-Appiah, Rodney
Published: (2026)
A Principle of Targeted Intervention for Multi-Agent Reinforcement Learning
by: Liu, Anjie, et al.
Published: (2025)
by: Liu, Anjie, et al.
Published: (2025)
MACS: Multi-Agent Reinforcement Learning for Optimization of Crystal Structures
by: Zamaraeva, Elena, et al.
Published: (2025)
by: Zamaraeva, Elena, et al.
Published: (2025)
WorkflowGen:an adaptive workflow generation mechanism driven by trajectory experience
by: Wei, Ruocan, et al.
Published: (2026)
by: Wei, Ruocan, et al.
Published: (2026)
QTypeMix: Enhancing Multi-Agent Cooperative Strategies through Heterogeneous and Homogeneous Value Decomposition
by: Fu, Songchen, et al.
Published: (2024)
by: Fu, Songchen, et al.
Published: (2024)
PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments
by: Schipper, Olivier, et al.
Published: (2025)
by: Schipper, Olivier, et al.
Published: (2025)
A Framework for Assessing AI Agent Decisions and Outcomes in AutoML Pipelines
by: Du, Gaoyuan, et al.
Published: (2026)
by: Du, Gaoyuan, et al.
Published: (2026)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
by: Wu, Shuai, et al.
Published: (2026)
by: Wu, Shuai, et al.
Published: (2026)
Towards Resource-Efficient Multimodal Intelligence: Learned Routing among Specialized Expert Models
by: Saini, Mayank, et al.
Published: (2025)
by: Saini, Mayank, et al.
Published: (2025)
Risk-Sensitive Multi-Agent Reinforcement Learning in Network Aggregative Markov Games
by: Ghaemi, Hafez, et al.
Published: (2024)
by: Ghaemi, Hafez, et al.
Published: (2024)
Centrally Coordinated Multi-Agent Reinforcement Learning for Power Grid Topology Control
by: de Mol, Barbera, et al.
Published: (2025)
by: de Mol, Barbera, et al.
Published: (2025)
Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design
by: Pepe, Alberto, et al.
Published: (2026)
by: Pepe, Alberto, et al.
Published: (2026)
Bimanual Robot Manipulation via Multi-Agent In-Context Learning
by: Palma, Alessio, et al.
Published: (2026)
by: Palma, Alessio, et al.
Published: (2026)
Differentiable Model Predictive Safety for Heterogeneous Mobility at Urban Intersections
by: Song, Wenzhe, et al.
Published: (2026)
by: Song, Wenzhe, et al.
Published: (2026)
PilotBench: A Benchmark for General Aviation Agents with Safety Constraints
by: Wu, Yalun, et al.
Published: (2026)
by: Wu, Yalun, et al.
Published: (2026)
Safe and Policy-Compliant Multi-Agent Orchestration for Enterprise AI
by: Pasupuleti, Vinil, et al.
Published: (2026)
by: Pasupuleti, Vinil, et al.
Published: (2026)
Policy Search, Retrieval, and Composition via Task Similarity in Collaborative Agentic Systems
by: Nath, Saptarshi, et al.
Published: (2025)
by: Nath, Saptarshi, et al.
Published: (2025)
Similar Items
-
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
by: Tang, Wenjie, et al.
Published: (2026) -
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
by: Wang, Yuchen, et al.
Published: (2026) -
Analysing Factorizations of Action-Value Networks for Cooperative Multi-Agent Reinforcement Learning
by: Castellini, Jacopo, et al.
Published: (2019) -
ME-IGM: Individual-Global-Max in Maximum Entropy Multi-Agent Reinforcement Learning
by: Chen, Wen-Tse, et al.
Published: (2024) -
ChromaFlow: A Negative Ablation Study of Orchestration Overhead in Tool-Augmented Agent Evaluation
by: Mittal, Tarun
Published: (2026)