ScaleEnv: Scaling Environment Synthesis from Scratch for Generalist Interactive Tool-Use Agent Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tu, Dunwei, Hao, Hongyan, Yang, Hansi, Chen, Yihao, Zhang, Yi-Kai, Xia, Zhikang, Yang, Yu, Sun, Yueqing, Liu, Xingchen, Shen, Furao, Gu, Qi, Su, Hui, Cai, Xunliang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL
von: Xu, Minrui, et al.
Veröffentlicht: (2026)
von: Xu, Minrui, et al.
Veröffentlicht: (2026)
TopoCurate:Modeling Interaction Topology for Tool-Use Agent Training
von: Yang, Jinluan, et al.
Veröffentlicht: (2026)
von: Yang, Jinluan, et al.
Veröffentlicht: (2026)
EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis
von: Song, Xiaoshuai, et al.
Veröffentlicht: (2026)
von: Song, Xiaoshuai, et al.
Veröffentlicht: (2026)
daVinci-Env: Open SWE Environment Synthesis at Scale
von: Fu, Dayuan, et al.
Veröffentlicht: (2026)
von: Fu, Dayuan, et al.
Veröffentlicht: (2026)
$V_{0.5}$: Generalist Value Model as a Prior for Sparse RL Rollouts
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2026)
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2026)
$V_0$: A Generalist Value Model for Any Policy at State Zero
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2026)
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2026)
Large-Scale Diverse Synthesis for Mid-Training
von: Zhang, Xuemiao, et al.
Veröffentlicht: (2025)
von: Zhang, Xuemiao, et al.
Veröffentlicht: (2025)
Beyond Point-wise Neural Collapse: A Topology-Aware Hierarchical Classifier for Class-Incremental Learning
von: Yi, Huiyu, et al.
Veröffentlicht: (2026)
von: Yi, Huiyu, et al.
Veröffentlicht: (2026)
Multiple Queries with Multiple Keys: A Precise Prompt Matching Paradigm for Prompt-based Continual Learning
von: Tu, Dunwei, et al.
Veröffentlicht: (2025)
von: Tu, Dunwei, et al.
Veröffentlicht: (2025)
Embedding Space Allocation with Angle-Norm Joint Classifiers for Few-Shot Class-Incremental Learning
von: Tu, Dunwei, et al.
Veröffentlicht: (2024)
von: Tu, Dunwei, et al.
Veröffentlicht: (2024)
VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications
von: He, Wei, et al.
Veröffentlicht: (2025)
von: He, Wei, et al.
Veröffentlicht: (2025)
ShuttleEnv: An Interactive Data-Driven RL Environment for Badminton Strategy Modeling
von: Li, Ang, et al.
Veröffentlicht: (2026)
von: Li, Ang, et al.
Veröffentlicht: (2026)
Prioritization Method for Crowdsourced Test Report by Integrating Text and Image Information
von: Huijie Tu, et al.
Veröffentlicht: (2025)
von: Huijie Tu, et al.
Veröffentlicht: (2025)
World-Env: Leveraging World Model as a Virtual Environment for VLA Post-Training
von: Xiao, Junjin, et al.
Veröffentlicht: (2025)
von: Xiao, Junjin, et al.
Veröffentlicht: (2025)
Tool Zero: Training Tool-Augmented LLMs via Pure RL from Scratch
von: Zeng, Yirong, et al.
Veröffentlicht: (2025)
von: Zeng, Yirong, et al.
Veröffentlicht: (2025)
EnvGen: Generating and Adapting Environments via LLMs for Training Embodied Agents
von: Zala, Abhay, et al.
Veröffentlicht: (2024)
von: Zala, Abhay, et al.
Veröffentlicht: (2024)
Scaling Sparse and Dense Retrieval in Decoder-Only LLMs
von: Zeng, Hansi, et al.
Veröffentlicht: (2025)
von: Zeng, Hansi, et al.
Veröffentlicht: (2025)
AutoEnv: Automated Environments for Measuring Cross-Environment Agent Learning
von: Zhang, Jiayi, et al.
Veröffentlicht: (2025)
von: Zhang, Jiayi, et al.
Veröffentlicht: (2025)
ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents
von: Lu, Yuxing, et al.
Veröffentlicht: (2026)
von: Lu, Yuxing, et al.
Veröffentlicht: (2026)
Inference-Time Scaling for Generalist Reward Modeling
von: Liu, Zijun, et al.
Veröffentlicht: (2025)
von: Liu, Zijun, et al.
Veröffentlicht: (2025)
ResearchEnvBench: Benchmarking Agents on Environment Synthesis for Research Code Execution
von: Wang, Yubang, et al.
Veröffentlicht: (2026)
von: Wang, Yubang, et al.
Veröffentlicht: (2026)
Unraveling the Mystery of Scaling Laws: Part I
von: Su, Hui, et al.
Veröffentlicht: (2024)
von: Su, Hui, et al.
Veröffentlicht: (2024)
Community Detection of Directed Network for Software Ecosystems Based on a Two‐Step Information Dissemination Model
von: Huijie Tu, et al.
Veröffentlicht: (2025)
von: Huijie Tu, et al.
Veröffentlicht: (2025)
EnvSimBench: A Benchmark for Evaluating and Improving LLM-Based Environment Simulation
von: Liu, Yi, et al.
Veröffentlicht: (2026)
von: Liu, Yi, et al.
Veröffentlicht: (2026)
Scaling and Transferability of Annealing Strategies in Large Language Model Training
von: Wang, Siqi, et al.
Veröffentlicht: (2025)
von: Wang, Siqi, et al.
Veröffentlicht: (2025)
Scaling Agents for Computer Use
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2025)
von: Gonzalez-Pumariega, Gonzalo, et al.
Veröffentlicht: (2025)
EnvBridge: Bridging Diverse Environments with Cross-Environment Knowledge Transfer for Embodied AI
von: Kagaya, Tomoyuki, et al.
Veröffentlicht: (2024)
von: Kagaya, Tomoyuki, et al.
Veröffentlicht: (2024)
DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use
von: Chen, Aili, et al.
Veröffentlicht: (2026)
von: Chen, Aili, et al.
Veröffentlicht: (2026)
ToolMind Technical Report: A Large-Scale, Reasoning-Enhanced Tool-Use Dataset
von: Yang, Chen, et al.
Veröffentlicht: (2025)
von: Yang, Chen, et al.
Veröffentlicht: (2025)
EnvPoser: Environment-aware Realistic Human Motion Estimation from Sparse Observations with Uncertainty Modeling
von: Xia, Songpengcheng, et al.
Veröffentlicht: (2024)
von: Xia, Songpengcheng, et al.
Veröffentlicht: (2024)
EnvBench: A Benchmark for Automated Environment Setup
von: Eliseeva, Aleksandra, et al.
Veröffentlicht: (2025)
von: Eliseeva, Aleksandra, et al.
Veröffentlicht: (2025)
Scaling Generalist Data-Analytic Agents
von: Qiao, Shuofei, et al.
Veröffentlicht: (2025)
von: Qiao, Shuofei, et al.
Veröffentlicht: (2025)
Agent S2: A Compositional Generalist-Specialist Framework for Computer Use Agents
von: Agashe, Saaket, et al.
Veröffentlicht: (2025)
von: Agashe, Saaket, et al.
Veröffentlicht: (2025)
ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch
von: Liu, Zheng, et al.
Veröffentlicht: (2026)
von: Liu, Zheng, et al.
Veröffentlicht: (2026)
ToolRM: Towards Agentic Tool-Use Reward Modeling
von: Li, Renhao, et al.
Veröffentlicht: (2025)
von: Li, Renhao, et al.
Veröffentlicht: (2025)
Communication-Efficient and Privacy-Preserving Decentralized Meta-Learning
von: Yang, Hansi, et al.
Veröffentlicht: (2024)
von: Yang, Hansi, et al.
Veröffentlicht: (2024)
CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents
von: Wang, Bowen, et al.
Veröffentlicht: (2026)
von: Wang, Bowen, et al.
Veröffentlicht: (2026)
AutoTool: Automatic Scaling of Tool-Use Capabilities in RL via Decoupled Entropy Constraints
von: Zeng, Yirong, et al.
Veröffentlicht: (2026)
von: Zeng, Yirong, et al.
Veröffentlicht: (2026)
GenEnv: Difficulty-Aligned Co-Evolution Between LLM Agents and Environment Simulators
von: Guo, Jiacheng, et al.
Veröffentlicht: (2025)
von: Guo, Jiacheng, et al.
Veröffentlicht: (2025)
SortingEnv: An Extendable RL-Environment for an Industrial Sorting Process
von: Maus, Tom, et al.
Veröffentlicht: (2025)
von: Maus, Tom, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL
von: Xu, Minrui, et al.
Veröffentlicht: (2026) -
TopoCurate:Modeling Interaction Topology for Tool-Use Agent Training
von: Yang, Jinluan, et al.
Veröffentlicht: (2026) -
EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis
von: Song, Xiaoshuai, et al.
Veröffentlicht: (2026) -
daVinci-Env: Open SWE Environment Synthesis at Scale
von: Fu, Dayuan, et al.
Veröffentlicht: (2026) -
$V_{0.5}$: Generalist Value Model as a Prior for Sparse RL Rollouts
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2026)