Interaction as Intelligence Part II: Asynchronous Human-Agent Rollout for Long-Horizon Task Training
Fuente:
arXiv
Saved in:
| Main Authors: | Fu, Dayuan, Wu, Yunze, Cai, Xiaojie, Ye, Lyumanshan, Xia, Shijie, Huang, Zhen, Si, Weiye, Xu, Tianze, Sun, Jie, Li, Keyu, Jiang, Mohan, Wang, Junfei, Hua, Qishuo, Lu, Pengrui, Xiao, Yang, Liu, Pengfei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
InnovatorBench: Evaluating Agents' Ability to Conduct Innovative LLM Research
by: Wu, Yunze, et al.
Published: (2025)
by: Wu, Yunze, et al.
Published: (2025)
Context Engineering 2.0: The Context of Context Engineering
by: Hua, Qishuo, et al.
Published: (2025)
by: Hua, Qishuo, et al.
Published: (2025)
AgencyBench: Benchmarking the Frontiers of Autonomous Agents in 1M-Token Real-World Contexts
by: Li, Keyu, et al.
Published: (2026)
by: Li, Keyu, et al.
Published: (2026)
DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments
by: Zheng, Yuxiang, et al.
Published: (2025)
by: Zheng, Yuxiang, et al.
Published: (2025)
daVinci-Agency: Unlocking Long-Horizon Agency Data-Efficiently
by: Jiang, Mohan, et al.
Published: (2026)
by: Jiang, Mohan, et al.
Published: (2026)
ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry
by: Xu, Tianze, et al.
Published: (2025)
by: Xu, Tianze, et al.
Published: (2025)
DatasetResearch: Benchmarking Agent Systems for Demand-Driven Dataset Discovery
by: Li, Keyu, et al.
Published: (2025)
by: Li, Keyu, et al.
Published: (2025)
Interaction as Intelligence: Deep Research With Human-AI Partnership
by: Ye, Lyumanshan, et al.
Published: (2025)
by: Ye, Lyumanshan, et al.
Published: (2025)
daVinci-Dev: Agent-native Mid-training for Software Engineering
by: Zeng, Ji, et al.
Published: (2026)
by: Zeng, Ji, et al.
Published: (2026)
DiffusionRollout: Uncertainty-Aware Rollout Planning in Long-Horizon PDE Solving
by: Yoo, Seungwoo, et al.
Published: (2026)
by: Yoo, Seungwoo, et al.
Published: (2026)
Harmonizing Real-Time Constraints and Long-Horizon Reasoning: An Asynchronous Agentic Framework for Dynamic Scheduling
by: Cao, Shijie, et al.
Published: (2026)
by: Cao, Shijie, et al.
Published: (2026)
Intelligent Optimization of Mine Environmental Damage Assessment and Repair Strategies Based on Deep Learning
by: Cheng, Qishuo
Published: (2024)
by: Cheng, Qishuo
Published: (2024)
Rubrics to Tokens: Bridging Response-level Rubrics and Token-level Rewards in Instruction Following Tasks
by: Xu, Tianze, et al.
Published: (2026)
by: Xu, Tianze, et al.
Published: (2026)
OlympicArena Medal Ranks: Who Is the Most Intelligent AI So Far?
by: Huang, Zhen, et al.
Published: (2024)
by: Huang, Zhen, et al.
Published: (2024)
Data Darwinism Part I: Unlocking the Value of Scientific Data for Pre-training
by: Qin, Yiwei, et al.
Published: (2026)
by: Qin, Yiwei, et al.
Published: (2026)
Unleashing Efficient Asynchronous RL Post-Training via Staleness-Constrained Rollout Coordination
by: Li, Haoyang, et al.
Published: (2026)
by: Li, Haoyang, et al.
Published: (2026)
Generalizable Dense Reward for Long-Horizon Robotic Tasks
by: Yong, Silong, et al.
Published: (2026)
by: Yong, Silong, et al.
Published: (2026)
SGNO: Spectral Generator Neural Operators for Stable Long Horizon PDE Rollouts
by: Li, Jiayi, et al.
Published: (2026)
by: Li, Jiayi, et al.
Published: (2026)
DyDiff: Long-Horizon Rollout via Dynamics Diffusion for Offline Reinforcement Learning
by: Zhao, Hanye, et al.
Published: (2024)
by: Zhao, Hanye, et al.
Published: (2024)
On Training Large Language Models for Long-Horizon Tasks: An Empirical Study of Horizon Length
by: Kim, Sunghwan, et al.
Published: (2026)
by: Kim, Sunghwan, et al.
Published: (2026)
LIMI: Less is More for Agency
by: Xiao, Yang, et al.
Published: (2025)
by: Xiao, Yang, et al.
Published: (2025)
Horizon Imagination: Efficient On-Policy Rollout in Diffusion World Models
by: Cohen, Lior, et al.
Published: (2026)
by: Cohen, Lior, et al.
Published: (2026)
Coherent Rollout Oracles for Finite-Horizon Sequential Decision Problems
by: Shukla, Nishant
Published: (2026)
by: Shukla, Nishant
Published: (2026)
Asynchronous Heavy-Tailed Optimization
by: Sun, Junfei, et al.
Published: (2026)
by: Sun, Junfei, et al.
Published: (2026)
daVinci-LLM:Towards the Science of Pretraining
by: Qin, Yiwei, et al.
Published: (2026)
by: Qin, Yiwei, et al.
Published: (2026)
A Goal Without a Plan Is Just a Wish: Efficient and Effective Global Planner Training for Long-Horizon Agent Tasks
by: Si, Shuzheng, et al.
Published: (2025)
by: Si, Shuzheng, et al.
Published: (2025)
RollPacker: Mitigating Long-Tail Rollouts for Fast, Synchronous RL Post-Training
by: Gao, Wei, et al.
Published: (2025)
by: Gao, Wei, et al.
Published: (2025)
AlphaGo Moment for Model Architecture Discovery
by: Liu, Yixiu, et al.
Published: (2025)
by: Liu, Yixiu, et al.
Published: (2025)
Policy Library CBF: Finite-Horizon Safety at Runtime via Parallel Rollouts
by: Kim, Taekyung, et al.
Published: (2026)
by: Kim, Taekyung, et al.
Published: (2026)
LHAW: Controllable Underspecification for Long-Horizon Tasks
by: Pu, George, et al.
Published: (2026)
by: Pu, George, et al.
Published: (2026)
SR-Scientist: Scientific Equation Discovery With Agentic AI
by: Xia, Shijie, et al.
Published: (2025)
by: Xia, Shijie, et al.
Published: (2025)
ProjDevBench: Benchmarking AI Coding Agents on End-to-End Project Development
by: Lu, Pengrui, et al.
Published: (2026)
by: Lu, Pengrui, et al.
Published: (2026)
D$^2$HScore: Reasoning-Aware Hallucination Detection via Semantic Breadth and Depth Analysis in LLMs
by: Ding, Yue, et al.
Published: (2025)
by: Ding, Yue, et al.
Published: (2025)
Decentralized Network Topology Design for Task Offloading in Mobile Edge Computing
by: Ma, Ke, et al.
Published: (2024)
by: Ma, Ke, et al.
Published: (2024)
Optimal Sample Splitting for Observational Studies
by: Yin, Qishuo, et al.
Published: (2026)
by: Yin, Qishuo, et al.
Published: (2026)
Knowledge-Guided Attention-Inspired Learning for Task Offloading in Vehicle Edge Computing
by: Ma, Ke, et al.
Published: (2025)
by: Ma, Ke, et al.
Published: (2025)
Beyond Ten Turns: Unlocking Long-Horizon Agentic Search with Large-Scale Asynchronous RL
by: Gao, Jiaxuan, et al.
Published: (2025)
by: Gao, Jiaxuan, et al.
Published: (2025)
DETACH: Cross-domain Learning for Long-Horizon Tasks via Mixture of Disentangled Experts
by: Shen, Yutong, et al.
Published: (2025)
by: Shen, Yutong, et al.
Published: (2025)
Shopping Companion: Benchmarking and Training LLM Agents for Long-Horizon Preference-Grounded E-Commerce Tasks
by: Yu, Zijian, et al.
Published: (2026)
by: Yu, Zijian, et al.
Published: (2026)
Asynchronous Majority Dynamics on Binomial Random Graphs
by: Mohan, Divyarthi, et al.
Published: (2023)
by: Mohan, Divyarthi, et al.
Published: (2023)
Similar Items
-
InnovatorBench: Evaluating Agents' Ability to Conduct Innovative LLM Research
by: Wu, Yunze, et al.
Published: (2025) -
Context Engineering 2.0: The Context of Context Engineering
by: Hua, Qishuo, et al.
Published: (2025) -
AgencyBench: Benchmarking the Frontiers of Autonomous Agents in 1M-Token Real-World Contexts
by: Li, Keyu, et al.
Published: (2026) -
DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments
by: Zheng, Yuxiang, et al.
Published: (2025) -
daVinci-Agency: Unlocking Long-Horizon Agency Data-Efficiently
by: Jiang, Mohan, et al.
Published: (2026)