Learning to Hand Off: Provably Convergent Workflow Learning under Interface Constraints
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Jiayu, Zhang, Enpei, Zhou, Dawei, Chen, Elynn, Yan, Yujun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
One-Step Bellman Alignment Enables Provably Efficient Transfer in Online RL
by: Chen, Elynn, et al.
Published: (2026)
by: Chen, Elynn, et al.
Published: (2026)
Seeing Through the Brain: New Insights from Decoding Visual Stimuli with fMRI
by: Huang, Zheng, et al.
Published: (2025)
by: Huang, Zheng, et al.
Published: (2025)
ACT-Tensor: Tensor Completion Framework for Financial Dataset Imputation
by: Mo, Junyi, et al.
Published: (2025)
by: Mo, Junyi, et al.
Published: (2025)
The Path Not Taken: RLVR Provably Learns Off the Principals
by: Zhu, Hanqing, et al.
Published: (2025)
by: Zhu, Hanqing, et al.
Published: (2025)
Tensor-Fused Multi-View Graph Contrastive Learning
by: Wu, Yujia, et al.
Published: (2024)
by: Wu, Yujia, et al.
Published: (2024)
Learning to Reason under Off-Policy Guidance
by: Yan, Jianhao, et al.
Published: (2025)
by: Yan, Jianhao, et al.
Published: (2025)
Provably Learning Diffusion Models under the Manifold Hypothesis: Collapse and Refine
by: Huang, Wei, et al.
Published: (2026)
by: Huang, Wei, et al.
Published: (2026)
Judging with Many Minds: Do More Perspectives Mean Less Prejudice? On Bias Amplifications and Resistance in Multi-Agent Based LLM-as-Judge
by: Ma, Chiyu, et al.
Published: (2025)
by: Ma, Chiyu, et al.
Published: (2025)
Behaviour Policy Optimization: Provably Lower Variance Return Estimates for Off-Policy Reinforcement Learning
by: Goodall, Alexander W., et al.
Published: (2025)
by: Goodall, Alexander W., et al.
Published: (2025)
Provable Zero-Shot Generalization in Offline Reinforcement Learning
by: Wang, Zhiyong, et al.
Published: (2025)
by: Wang, Zhiyong, et al.
Published: (2025)
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
by: Xiao, Minheng, et al.
Published: (2024)
by: Xiao, Minheng, et al.
Published: (2024)
Tensor-view Topological Graph Neural Network
by: Wen, Tao, et al.
Published: (2024)
by: Wen, Tao, et al.
Published: (2024)
Distributed In-Context Learning under Non-IID Among Clients
by: Liang, Siqi, et al.
Published: (2024)
by: Liang, Siqi, et al.
Published: (2024)
Ferret: An Efficient Online Continual Learning Framework under Varying Memory Constraints
by: Zhou, Yuhao, et al.
Published: (2025)
by: Zhou, Yuhao, et al.
Published: (2025)
Provable Distributional Value Iteration under Partial Observability
by: Preuett III, Larry, et al.
Published: (2025)
by: Preuett III, Larry, et al.
Published: (2025)
Accelerating Convergence of Score-Based Diffusion Models, Provably
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Theoretical Analysis of Meta Reinforcement Learning: Generalization Bounds and Convergence Guarantees
by: Wang, Cangqing, et al.
Published: (2024)
by: Wang, Cangqing, et al.
Published: (2024)
RUKA: Rethinking the Design of Humanoid Hands with Learning
by: Zorin, Anya, et al.
Published: (2025)
by: Zorin, Anya, et al.
Published: (2025)
Off-Policy Evaluation and Learning for the Future under Non-Stationarity
by: Shimizu, Tatsuhiro, et al.
Published: (2025)
by: Shimizu, Tatsuhiro, et al.
Published: (2025)
ODRL: A Benchmark for Off-Dynamics Reinforcement Learning
by: Lyu, Jiafei, et al.
Published: (2024)
by: Lyu, Jiafei, et al.
Published: (2024)
Triadic-OCD: Asynchronous Online Change Detection with Provable Robustness, Optimality, and Convergence
by: Huang, Yancheng, et al.
Published: (2024)
by: Huang, Yancheng, et al.
Published: (2024)
Transfer Faster, Price Smarter: Minimax Dynamic Pricing under Cross-Market Preference Shift
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
Improving Open-world Continual Learning under the Constraints of Scarce Labeled Data
by: Li, Yujie, et al.
Published: (2025)
by: Li, Yujie, et al.
Published: (2025)
Provably Efficient Exploration in Inverse Constrained Reinforcement Learning
by: Yue, Bo, et al.
Published: (2024)
by: Yue, Bo, et al.
Published: (2024)
Deciphering Raw Data in Neuro-Symbolic Learning with Provable Guarantees
by: Tao, Lue, et al.
Published: (2023)
by: Tao, Lue, et al.
Published: (2023)
Learning to Solve Combinatorial Optimization under Positive Linear Constraints via Non-Autoregressive Neural Networks
by: Wang, Runzhong, et al.
Published: (2024)
by: Wang, Runzhong, et al.
Published: (2024)
How to Provably Improve Return Conditioned Supervised Learning?
by: Liu, Zhishuai, et al.
Published: (2025)
by: Liu, Zhishuai, et al.
Published: (2025)
Anytime Probabilistically Constrained Provably Convergent Online Belief Space Planning
by: Zhitnikov, Andrey, et al.
Published: (2024)
by: Zhitnikov, Andrey, et al.
Published: (2024)
TEAFormers: TEnsor-Augmented Transformers for Multi-Dimensional Time Series Forecasting
by: Kong, Linghang, et al.
Published: (2024)
by: Kong, Linghang, et al.
Published: (2024)
AudioMotionBench: Evaluating Auditory Motion Perception in Audio LLMs
by: Sun, Zhe, et al.
Published: (2025)
by: Sun, Zhe, et al.
Published: (2025)
Provably Efficient Action-Manipulation Attack Against Continuous Reinforcement Learning
by: Luo, Zhi, et al.
Published: (2024)
by: Luo, Zhi, et al.
Published: (2024)
Off-Policy Evaluation and Learning for Survival Outcomes under Censoring
by: Kubota, Kohsuke, et al.
Published: (2026)
by: Kubota, Kohsuke, et al.
Published: (2026)
Provable Benefits of Task-Specific Prompts for In-context Learning
by: Chang, Xiangyu, et al.
Published: (2025)
by: Chang, Xiangyu, et al.
Published: (2025)
Low-Rank Plus Sparse Matrix Transfer Learning under Growing Representations and Ambient Dimensions
by: Chai, Jinhang, et al.
Published: (2026)
by: Chai, Jinhang, et al.
Published: (2026)
Off-policy Reinforcement Learning with Model-based Exploration Augmentation
by: Wang, Likun, et al.
Published: (2025)
by: Wang, Likun, et al.
Published: (2025)
Incentivizing Inclusive Contributions in Model Sharing Markets
by: Zhang, Enpei, et al.
Published: (2025)
by: Zhang, Enpei, et al.
Published: (2025)
Cost-Aware Dynamic Cloud Workflow Scheduling using Self-Attention and Evolutionary Reinforcement Learning
by: Shen, Ya, et al.
Published: (2024)
by: Shen, Ya, et al.
Published: (2024)
Helpful Agent Meets Deceptive Judge: Understanding Vulnerabilities in Agentic Workflows
by: Ming, Yifei, et al.
Published: (2025)
by: Ming, Yifei, et al.
Published: (2025)
Learning Flexible Job Shop Scheduling under Limited Buffers and Material Kitting Constraints
by: Zhang, Shishun, et al.
Published: (2026)
by: Zhang, Shishun, et al.
Published: (2026)
On Time, Within Budget: Constraint-Driven Online Resource Allocation for Agentic Workflows
by: Wang, Xinglin, et al.
Published: (2026)
by: Wang, Xinglin, et al.
Published: (2026)
Similar Items
-
One-Step Bellman Alignment Enables Provably Efficient Transfer in Online RL
by: Chen, Elynn, et al.
Published: (2026) -
Seeing Through the Brain: New Insights from Decoding Visual Stimuli with fMRI
by: Huang, Zheng, et al.
Published: (2025) -
ACT-Tensor: Tensor Completion Framework for Financial Dataset Imputation
by: Mo, Junyi, et al.
Published: (2025) -
The Path Not Taken: RLVR Provably Learns Off the Principals
by: Zhu, Hanqing, et al.
Published: (2025) -
Tensor-Fused Multi-View Graph Contrastive Learning
by: Wu, Yujia, et al.
Published: (2024)