Agentic Test-Time Scaling for WebAgents
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Nicholas, Erdogan, Lutfi Eren, John, Chris Joseph, Krishnapillai, Surya, Mahoney, Michael W., Keutzer, Kurt, Gholami, Amir |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Characterizing Prompt Compression Methods for Long Context Inference
di: Jha, Siddharth, et al.
Pubblicazione: (2024)
di: Jha, Siddharth, et al.
Pubblicazione: (2024)
Plan-and-Act: Improving Planning of Agents for Long-Horizon Tasks
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2025)
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2025)
TinyAgent: Function Calling at the Edge
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2024)
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2024)
Efficient and Scalable Estimation of Tool Representations in Vector Space
di: Moon, Suhong, et al.
Pubblicazione: (2024)
di: Moon, Suhong, et al.
Pubblicazione: (2024)
An LLM Compiler for Parallel Function Calling
di: Kim, Sehoon, et al.
Pubblicazione: (2023)
di: Kim, Sehoon, et al.
Pubblicazione: (2023)
LLM2LLM: Boosting LLMs with Novel Iterative Data Enhancement
di: Lee, Nicholas, et al.
Pubblicazione: (2024)
di: Lee, Nicholas, et al.
Pubblicazione: (2024)
SqueezeLLM: Dense-and-Sparse Quantization
di: Kim, Sehoon, et al.
Pubblicazione: (2023)
di: Kim, Sehoon, et al.
Pubblicazione: (2023)
Stochastic Communication Avoidance for Recommendation Systems
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2024)
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2024)
MM-WebAgent: A Hierarchical Multimodal Web Agent for Webpage Generation
di: Li, Yan, et al.
Pubblicazione: (2026)
di: Li, Yan, et al.
Pubblicazione: (2026)
Beyond Next-Token Prediction: A Performance Characterization of Diffusion versus Autoregressive Language Models
di: Kim, Minseo, et al.
Pubblicazione: (2025)
di: Kim, Minseo, et al.
Pubblicazione: (2025)
Multipole Attention for Efficient Long Context Reasoning
di: Hooper, Coleman, et al.
Pubblicazione: (2025)
di: Hooper, Coleman, et al.
Pubblicazione: (2025)
Squeezed Attention: Accelerating Long Context Length LLM Inference
di: Hooper, Coleman, et al.
Pubblicazione: (2024)
di: Hooper, Coleman, et al.
Pubblicazione: (2024)
WebLeaper: Empowering Efficiency and Efficacy in WebAgent via Enabling Info-Rich Seeking
di: Tao, Zhengwei, et al.
Pubblicazione: (2025)
di: Tao, Zhengwei, et al.
Pubblicazione: (2025)
Arbitrage: Efficient Reasoning via Advantage-Aware Speculation
di: Maheswaran, Monishwaran, et al.
Pubblicazione: (2025)
di: Maheswaran, Monishwaran, et al.
Pubblicazione: (2025)
WebAgent-R1: Training Web Agents via End-to-End Multi-Turn Reinforcement Learning
di: Wei, Zhepei, et al.
Pubblicazione: (2025)
di: Wei, Zhepei, et al.
Pubblicazione: (2025)
Speculative Interaction Agents: Building Real-Time Agents with Asynchronous I/O and Speculative Tool Calling
di: Hooper, Coleman, et al.
Pubblicazione: (2026)
di: Hooper, Coleman, et al.
Pubblicazione: (2026)
ETS: Efficient Tree Search for Inference-Time Scaling
di: Hooper, Coleman, et al.
Pubblicazione: (2025)
di: Hooper, Coleman, et al.
Pubblicazione: (2025)
A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis
di: Gur, Izzeddin, et al.
Pubblicazione: (2023)
di: Gur, Izzeddin, et al.
Pubblicazione: (2023)
MITRA: A Large-Scale Parallel Corpus and Multilingual Pretrained Language Model for Machine Translation and Semantic Retrieval for Pāli, Sanskrit, Buddhist Chinese, and Tibetan
di: Nehrdich, Sebastian, et al.
Pubblicazione: (2026)
di: Nehrdich, Sebastian, et al.
Pubblicazione: (2026)
Towards Foundation Models for Scientific Machine Learning: Characterizing Scaling and Transfer Behavior
di: Subramanian, Shashank, et al.
Pubblicazione: (2023)
di: Subramanian, Shashank, et al.
Pubblicazione: (2023)
SPEED: Speculative Pipelined Execution for Efficient Decoding
di: Hooper, Coleman, et al.
Pubblicazione: (2023)
di: Hooper, Coleman, et al.
Pubblicazione: (2023)
AI and Memory Wall
di: Gholami, Amir, et al.
Pubblicazione: (2024)
di: Gholami, Amir, et al.
Pubblicazione: (2024)
LoSA: Locality Aware Sparse Attention for Block-Wise Diffusion Language Models
di: Xi, Haocheng, et al.
Pubblicazione: (2026)
di: Xi, Haocheng, et al.
Pubblicazione: (2026)
SciML Agents: Write the Solver, Not the Solution
di: Gaonkar, Saarth, et al.
Pubblicazione: (2025)
di: Gaonkar, Saarth, et al.
Pubblicazione: (2025)
CDLM: Consistency Diffusion Language Models For Faster Sampling
di: Kim, Minseo, et al.
Pubblicazione: (2025)
di: Kim, Minseo, et al.
Pubblicazione: (2025)
Residual Context Diffusion Language Models
di: Hu, Yuezhou, et al.
Pubblicazione: (2026)
di: Hu, Yuezhou, et al.
Pubblicazione: (2026)
$\texttt{SPECS}$: Faster Test-Time Scaling through Speculative Drafts
di: Cemri, Mert, et al.
Pubblicazione: (2025)
di: Cemri, Mert, et al.
Pubblicazione: (2025)
BrowseConf: Confidence-Guided Test-Time Scaling for Web Agents
di: Ou, Litu, et al.
Pubblicazione: (2025)
di: Ou, Litu, et al.
Pubblicazione: (2025)
KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
di: Hooper, Coleman, et al.
Pubblicazione: (2024)
di: Hooper, Coleman, et al.
Pubblicazione: (2024)
One Model is All You Need: ByT5-Sanskrit, a Unified Model for Sanskrit NLP Tasks
di: Nehrdich, Sebastian, et al.
Pubblicazione: (2024)
di: Nehrdich, Sebastian, et al.
Pubblicazione: (2024)
Simple and Effective Input Reformulations for Translation
di: Yu, Brian, et al.
Pubblicazione: (2023)
di: Yu, Brian, et al.
Pubblicazione: (2023)
Reward Under Attack: Analyzing the Robustness and Hackability of Process Reward Models
di: Tiwari, Rishabh, et al.
Pubblicazione: (2026)
di: Tiwari, Rishabh, et al.
Pubblicazione: (2026)
Evaluating Long-Context Reasoning in LLM-Based WebAgents
di: Chung, Andy, et al.
Pubblicazione: (2025)
di: Chung, Andy, et al.
Pubblicazione: (2025)
Timely Machine: Awareness of Time Makes Test-Time Scaling Agentic
di: Ma, Yichuan, et al.
Pubblicazione: (2026)
di: Ma, Yichuan, et al.
Pubblicazione: (2026)
LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling
di: Zheng, Tong, et al.
Pubblicazione: (2026)
di: Zheng, Tong, et al.
Pubblicazione: (2026)
On the Role of Feedback in Test-Time Scaling of Agentic AI Workflows
di: Chakraborty, Souradip, et al.
Pubblicazione: (2025)
di: Chakraborty, Souradip, et al.
Pubblicazione: (2025)
Scaling Test-Time Compute for Agentic Coding
di: Kim, Joongwon, et al.
Pubblicazione: (2026)
di: Kim, Joongwon, et al.
Pubblicazione: (2026)
ARTIS: Agentic Risk-Aware Test-Time Scaling via Iterative Simulation
di: Zeng, Xingshan, et al.
Pubblicazione: (2026)
di: Zeng, Xingshan, et al.
Pubblicazione: (2026)
XQuant: Breaking the Memory Wall for LLM Inference with KV Cache Rematerialization
di: Tomar, Aditya, et al.
Pubblicazione: (2025)
di: Tomar, Aditya, et al.
Pubblicazione: (2025)
A Survey of WebAgents: Towards Next-Generation AI Agents for Web Automation with Large Foundation Models
di: Ning, Liangbo, et al.
Pubblicazione: (2025)
di: Ning, Liangbo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Characterizing Prompt Compression Methods for Long Context Inference
di: Jha, Siddharth, et al.
Pubblicazione: (2024) -
Plan-and-Act: Improving Planning of Agents for Long-Horizon Tasks
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2025) -
TinyAgent: Function Calling at the Edge
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2024) -
Efficient and Scalable Estimation of Tool Representations in Vector Space
di: Moon, Suhong, et al.
Pubblicazione: (2024) -
An LLM Compiler for Parallel Function Calling
di: Kim, Sehoon, et al.
Pubblicazione: (2023)