Agent JIT Compilation for Latency-Optimizing Web Agent Planning and Scheduling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Winston, Caleb, Wang, Ron Yifeng, Mirhoseini, Azalia, Kozyrakis, Christos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
That Chip Has Sailed: A Critique of Unfounded Skepticism Around AI for Chip Design
von: Goldie, Anna, et al.
Veröffentlicht: (2024)
von: Goldie, Anna, et al.
Veröffentlicht: (2024)
On the Role of Temperature Sampling in Test-Time Scaling
von: Wu, Yuheng, et al.
Veröffentlicht: (2025)
von: Wu, Yuheng, et al.
Veröffentlicht: (2025)
SPRINT: Enabling Interleaved Planning and Parallelized Execution in Reasoning Models
von: Biju, Emil, et al.
Veröffentlicht: (2025)
von: Biju, Emil, et al.
Veröffentlicht: (2025)
Improving Efficiency of GPU Kernel Optimization Agents using a Domain-Specific Language and Speed-of-Light Guidance
von: Hari, Siva Kumar Sastry, et al.
Veröffentlicht: (2026)
von: Hari, Siva Kumar Sastry, et al.
Veröffentlicht: (2026)
ForTIFAI: Fending Off Recursive Training Induced Failure for AI Model Collapse
von: Shabgahi, Soheil Zibakhsh, et al.
Veröffentlicht: (2025)
von: Shabgahi, Soheil Zibakhsh, et al.
Veröffentlicht: (2025)
Hybrid JIT-CUDA Graph Optimization for Low-Latency Large Language Model Inference
von: Yadav, Divakar Kumar, et al.
Veröffentlicht: (2026)
von: Yadav, Divakar Kumar, et al.
Veröffentlicht: (2026)
Astra: A Multi-Agent System for GPU Kernel Performance Optimization
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
CHESS: Contextual Harnessing for Efficient SQL Synthesis
von: Talaei, Shayan, et al.
Veröffentlicht: (2024)
von: Talaei, Shayan, et al.
Veröffentlicht: (2024)
Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use
von: Goldie, Anna, et al.
Veröffentlicht: (2025)
von: Goldie, Anna, et al.
Veröffentlicht: (2025)
Efficient GNN Training Through Structure-Aware Randomized Mini-Batching
von: Balaji, Vignesh, et al.
Veröffentlicht: (2025)
von: Balaji, Vignesh, et al.
Veröffentlicht: (2025)
Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
von: Brown, Bradley, et al.
Veröffentlicht: (2024)
von: Brown, Bradley, et al.
Veröffentlicht: (2024)
Workspace Optimization: How to Train Your Agent
von: Sarafian, Elad, et al.
Veröffentlicht: (2026)
von: Sarafian, Elad, et al.
Veröffentlicht: (2026)
AI Metropolis: Scaling Large Language Model-based Multi-Agent Simulation with Out-of-order Execution
von: Xie, Zhiqiang, et al.
Veröffentlicht: (2024)
von: Xie, Zhiqiang, et al.
Veröffentlicht: (2024)
KernelBlaster: Continual Cross-Task CUDA Optimization via Memory-Augmented In-Context Reinforcement Learning
von: Dong, Kris Shengjun, et al.
Veröffentlicht: (2026)
von: Dong, Kris Shengjun, et al.
Veröffentlicht: (2026)
Agentic Web: Weaving the Next Web with AI Agents
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
von: Yang, Yingxuan, et al.
Veröffentlicht: (2025)
How Do Large Language Monkeys Get Their Power (Laws)?
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2025)
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2025)
KernelBench: Can LLMs Write Efficient GPU Kernels?
von: Ouyang, Anne, et al.
Veröffentlicht: (2025)
von: Ouyang, Anne, et al.
Veröffentlicht: (2025)
Capacity-Aware Planning and Scheduling in Budget-Constrained Multi-Agent MDPs: A Meta-RL Approach
von: Vora, Manav, et al.
Veröffentlicht: (2024)
von: Vora, Manav, et al.
Veröffentlicht: (2024)
TGPO: Tree-Guided Preference Optimization for Robust Web Agent Reinforcement Learning
von: Chen, Ziyuan, et al.
Veröffentlicht: (2025)
von: Chen, Ziyuan, et al.
Veröffentlicht: (2025)
SynthAgent: Adapting Web Agents with Synthetic Supervision
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2025)
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2025)
MPO: Boosting LLM Agents with Meta Plan Optimization
von: Xiong, Weimin, et al.
Veröffentlicht: (2025)
von: Xiong, Weimin, et al.
Veröffentlicht: (2025)
Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents
von: Zambrano, Alejandra, et al.
Veröffentlicht: (2026)
von: Zambrano, Alejandra, et al.
Veröffentlicht: (2026)
Reasoning-SQL: Reinforcement Learning with SQL Tailored Partial Rewards for Reasoning-Enhanced Text-to-SQL
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2025)
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2025)
Plan Before You Trade: Inference-Time Optimization for RL Trading Agents
von: Go, Eun, et al.
Veröffentlicht: (2026)
von: Go, Eun, et al.
Veröffentlicht: (2026)
SWE-Replay: Efficient Test-Time Scaling for Software Engineering Agents
von: Ding, Yifeng, et al.
Veröffentlicht: (2026)
von: Ding, Yifeng, et al.
Veröffentlicht: (2026)
A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis
von: Gur, Izzeddin, et al.
Veröffentlicht: (2023)
von: Gur, Izzeddin, et al.
Veröffentlicht: (2023)
AgentFold: Long-Horizon Web Agents with Proactive Context Management
von: Ye, Rui, et al.
Veröffentlicht: (2025)
von: Ye, Rui, et al.
Veröffentlicht: (2025)
Dynamic Speculative Agent Planning
von: Guan, Yilin, et al.
Veröffentlicht: (2025)
von: Guan, Yilin, et al.
Veröffentlicht: (2025)
Dynamic Co-Optimization Compiler: Leveraging Multi-Agent Reinforcement Learning for Enhanced DNN Accelerator Performance
von: Fayyazi, Arya, et al.
Veröffentlicht: (2024)
von: Fayyazi, Arya, et al.
Veröffentlicht: (2024)
Cogito, Ergo Ludo: An Agent that Learns to Play by Reasoning and Planning
von: Wang, Sai, et al.
Veröffentlicht: (2025)
von: Wang, Sai, et al.
Veröffentlicht: (2025)
Agent-Oriented Planning in Multi-Agent Systems
von: Li, Ao, et al.
Veröffentlicht: (2024)
von: Li, Ao, et al.
Veröffentlicht: (2024)
Evaluating Long-Context Reasoning in LLM-Based WebAgents
von: Chung, Andy, et al.
Veröffentlicht: (2025)
von: Chung, Andy, et al.
Veröffentlicht: (2025)
Reducing Latency of LLM Search Agent via Speculation-based Algorithm-System Co-Design
von: Huang, Zixiao, et al.
Veröffentlicht: (2025)
von: Huang, Zixiao, et al.
Veröffentlicht: (2025)
Spatial Reasoning and Planning for Deep Embodied Agents
von: Ishida, Shu
Veröffentlicht: (2024)
von: Ishida, Shu
Veröffentlicht: (2024)
Cartridges: Lightweight and general-purpose long context representations via self-study
von: Eyuboglu, Sabri, et al.
Veröffentlicht: (2025)
von: Eyuboglu, Sabri, et al.
Veröffentlicht: (2025)
AgentRewardBench: Evaluating Automatic Evaluations of Web Agent Trajectories
von: Lù, Xing Han, et al.
Veröffentlicht: (2025)
von: Lù, Xing Han, et al.
Veröffentlicht: (2025)
AscendOptimizer: Episodic Agent for Ascend NPU Operator Optimization
von: Wu, Jiehao, et al.
Veröffentlicht: (2026)
von: Wu, Jiehao, et al.
Veröffentlicht: (2026)
How to Train Your LLM Web Agent: A Statistical Diagnosis
von: Vattikonda, Dheeraj, et al.
Veröffentlicht: (2025)
von: Vattikonda, Dheeraj, et al.
Veröffentlicht: (2025)
DRIVE: Modeling Skills at the Reasoning and Interaction Levels for Web Agents under Continual Learning
von: Liu, Xirui, et al.
Veröffentlicht: (2026)
von: Liu, Xirui, et al.
Veröffentlicht: (2026)
End-to-end PDDL Planning with Hardcoded and Dynamic Agents
von: La Malfa, Emanuele, et al.
Veröffentlicht: (2025)
von: La Malfa, Emanuele, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
That Chip Has Sailed: A Critique of Unfounded Skepticism Around AI for Chip Design
von: Goldie, Anna, et al.
Veröffentlicht: (2024) -
On the Role of Temperature Sampling in Test-Time Scaling
von: Wu, Yuheng, et al.
Veröffentlicht: (2025) -
SPRINT: Enabling Interleaved Planning and Parallelized Execution in Reasoning Models
von: Biju, Emil, et al.
Veröffentlicht: (2025) -
Improving Efficiency of GPU Kernel Optimization Agents using a Domain-Specific Language and Speed-of-Light Guidance
von: Hari, Siva Kumar Sastry, et al.
Veröffentlicht: (2026) -
ForTIFAI: Fending Off Recursive Training Induced Failure for AI Model Collapse
von: Shabgahi, Soheil Zibakhsh, et al.
Veröffentlicht: (2025)