Workflow-R1: Group Sub-sequence Policy Optimization for Multi-turn Workflow Construction
Fuente:
arXiv
Saved in:
| Main Authors: | Kong, Mingze, Qu, Zikun, Zhou, Zhongquan, Liang, Pengyu, Li, Xiang, Shang, Zhiwei, Hong, Zhi, Huang, Kaiyu, Wang, Zhiyong, Dai, Zhongxiang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
T-POP: Test-Time Personalization with Online Preference Feedback
by: Qu, Zikun, et al.
Published: (2025)
by: Qu, Zikun, et al.
Published: (2025)
Meta-Prompt Optimization for LLM-Based Sequential Decision Making
by: Kong, Mingze, et al.
Published: (2025)
by: Kong, Mingze, et al.
Published: (2025)
MASPOB: Bandit-Based Prompt Optimization for Multi-Agent Systems with Graph Neural Networks
by: Hong, Zhi, et al.
Published: (2026)
by: Hong, Zhi, et al.
Published: (2026)
Linear and Neural Dueling Bandits with Delayed Feedback
by: Wang, Xiangyi, et al.
Published: (2026)
by: Wang, Xiangyi, et al.
Published: (2026)
UniScale: Adaptive Unified Inference Scaling via Online Joint Optimization of Model Routing and Test-Time Scaling
by: Huang, Kaiyu, et al.
Published: (2026)
by: Huang, Kaiyu, et al.
Published: (2026)
ALSO: Adversarial Online Strategy Optimization for Social Agents
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
FedPOB: Sample-Efficient Federated Prompt Optimization via Bandits
by: Lu, Pingchen, et al.
Published: (2025)
by: Lu, Pingchen, et al.
Published: (2025)
EduAgentQG: A Multi-Agent Workflow Framework for Personalized Question Generation
by: Jia, Rui, et al.
Published: (2025)
by: Jia, Rui, et al.
Published: (2025)
Online Clustering of Dueling Bandits
by: Wang, Zhiyong, et al.
Published: (2025)
by: Wang, Zhiyong, et al.
Published: (2025)
EpiBench: Benchmarking Multi-turn Research Workflows for Multimodal Agents
by: Dong, Xuan, et al.
Published: (2026)
by: Dong, Xuan, et al.
Published: (2026)
EduDial: Constructing a Large-scale Multi-turn Teacher-Student Dialogue Corpus
by: Wei, Shouang, et al.
Published: (2025)
by: Wei, Shouang, et al.
Published: (2025)
FlowCompile: An Optimizing Compiler for Structured LLM Workflows
by: Li, Junyan, et al.
Published: (2026)
by: Li, Junyan, et al.
Published: (2026)
Large Language Models for Constructing and Optimizing Machine Learning Workflows: A Survey
by: Gu, Yang, et al.
Published: (2024)
by: Gu, Yang, et al.
Published: (2024)
When Does Multi-Agent RL Improve LLM Workflows? Workflow, Scale, and Policy-Sharing Tradeoffs
by: Zeng, Yifan, et al.
Published: (2026)
by: Zeng, Yifan, et al.
Published: (2026)
Optimization Problem Solving Can Transition to Evolutionary Agentic Workflows
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
Finch: Benchmarking Finance & Accounting across Spreadsheet-Centric Enterprise Workflows
by: Dong, Haoyu, et al.
Published: (2025)
by: Dong, Haoyu, et al.
Published: (2025)
Workflow Optimization for Parallel Split Learning
by: Tirana, Joana, et al.
Published: (2024)
by: Tirana, Joana, et al.
Published: (2024)
On the extreme order statistics for stationary Gaussian sequences subject to random missing observations
by: Fang, Yuan, et al.
Published: (2023)
by: Fang, Yuan, et al.
Published: (2023)
Translating Workflow Nets into the Partially Ordered Workflow Language
by: Kourani, Humam, et al.
Published: (2025)
by: Kourani, Humam, et al.
Published: (2025)
Towards Enforcing Company Policy Adherence in Agentic Workflows
by: Zwerdling, Naama, et al.
Published: (2025)
by: Zwerdling, Naama, et al.
Published: (2025)
AFlow: Automating Agentic Workflow Generation
by: Zhang, Jiayi, et al.
Published: (2024)
by: Zhang, Jiayi, et al.
Published: (2024)
Agentic Workflow for Education: Concepts and Applications
by: Jiang, Yuan-Hao, et al.
Published: (2025)
by: Jiang, Yuan-Hao, et al.
Published: (2025)
Towards Hierarchical Multi-Agent Workflows for Zero-Shot Prompt Optimization
by: Liu, Yuchi, et al.
Published: (2024)
by: Liu, Yuchi, et al.
Published: (2024)
Writing Workflows
by: Lockridge, Tim, et al.
Published: (2024)
by: Lockridge, Tim, et al.
Published: (2024)
Batch Query Processing and Optimization for Agentic Workflows
by: Shen, Junyi, et al.
Published: (2025)
by: Shen, Junyi, et al.
Published: (2025)
Optimizing Agentic Workflows using Meta-tools
by: Abuzakuk, Sami, et al.
Published: (2026)
by: Abuzakuk, Sami, et al.
Published: (2026)
Automating Nanoindentation: Optimizing Workflows for Precision and Accuracy
by: Chawla, Vivek, et al.
Published: (2025)
by: Chawla, Vivek, et al.
Published: (2025)
GPTArticleExtractor: An Automated Workflow for Magnetic Material Database Construction
by: Zhang, Yibo, et al.
Published: (2024)
by: Zhang, Yibo, et al.
Published: (2024)
Statistical Independence Aware Caching for LLM Workflows
by: Dai, Yihan, et al.
Published: (2025)
by: Dai, Yihan, et al.
Published: (2025)
Opus: A Workflow Intention Framework for Complex Workflow Generation
by: Kingston, Phillip, et al.
Published: (2025)
by: Kingston, Phillip, et al.
Published: (2025)
From Cyber Threat to Data Shield: Constructing Provably Secure File Erasure with Repurposed Ransomware Cryptography
by: Shang, Jiahui, et al.
Published: (2025)
by: Shang, Jiahui, et al.
Published: (2025)
Group Sequence Policy Optimization
by: Zheng, Chujie, et al.
Published: (2025)
by: Zheng, Chujie, et al.
Published: (2025)
Workflows Community Summit 2024: Future Trends and Challenges in Scientific Workflows
by: da Silva, Rafael Ferreira, et al.
Published: (2024)
by: da Silva, Rafael Ferreira, et al.
Published: (2024)
Near-Miss: Latent Policy Failure Detection in Agentic Workflows
by: Rabinovich, Ella, et al.
Published: (2026)
by: Rabinovich, Ella, et al.
Published: (2026)
Guardrails as Infrastructure: Policy-First Control for Tool-Orchestrated Workflows
by: Sigdel, Akshey, et al.
Published: (2026)
by: Sigdel, Akshey, et al.
Published: (2026)
DenoiseFlow: Uncertainty-Aware Denoising for Reliable LLM Agentic Workflows
by: Yan, Yandong, et al.
Published: (2026)
by: Yan, Yandong, et al.
Published: (2026)
ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows
by: Sun, Qiushi, et al.
Published: (2025)
by: Sun, Qiushi, et al.
Published: (2025)
WorkflowLLM: Enhancing Workflow Orchestration Capability of Large Language Models
by: Fan, Shengda, et al.
Published: (2024)
by: Fan, Shengda, et al.
Published: (2024)
WorkTeam: Constructing Workflows from Natural Language with Multi-Agents
by: Liu, Hanchao, et al.
Published: (2025)
by: Liu, Hanchao, et al.
Published: (2025)
Compass: Optimizing Compound AI Workflows for Dynamic Adaptation
by: Gravara, Milos, et al.
Published: (2026)
by: Gravara, Milos, et al.
Published: (2026)
Similar Items
-
T-POP: Test-Time Personalization with Online Preference Feedback
by: Qu, Zikun, et al.
Published: (2025) -
Meta-Prompt Optimization for LLM-Based Sequential Decision Making
by: Kong, Mingze, et al.
Published: (2025) -
MASPOB: Bandit-Based Prompt Optimization for Multi-Agent Systems with Graph Neural Networks
by: Hong, Zhi, et al.
Published: (2026) -
Linear and Neural Dueling Bandits with Delayed Feedback
by: Wang, Xiangyi, et al.
Published: (2026) -
UniScale: Adaptive Unified Inference Scaling via Online Joint Optimization of Model Routing and Test-Time Scaling
by: Huang, Kaiyu, et al.
Published: (2026)