Scaling Test-Time Compute for Agentic Coding
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Joongwon, Yang, Wannan, Niu, Kelvin, Zhang, Hongming, Zhu, Yun, Helenowski, Eryk, Silva, Ruan, Chen, Zhengxing, Iyer, Srinivasan, Zaheer, Manzil, Fried, Daniel, Hajishirzi, Hannaneh, Arora, Sanjeev, Synnaeve, Gabriel, Salakhutdinov, Ruslan, Goyal, Anirudh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ASTRO: Teaching Language Models to Reason by Reflecting and Backtracking In-Context
di: Kim, Joongwon, et al.
Pubblicazione: (2025)
di: Kim, Joongwon, et al.
Pubblicazione: (2025)
Rethinking Thinking Tokens: LLMs as Improvement Operators
di: Madaan, Lovish, et al.
Pubblicazione: (2025)
di: Madaan, Lovish, et al.
Pubblicazione: (2025)
Husky: A Unified, Open-Source Language Agent for Multi-Step Reasoning
di: Kim, Joongwon, et al.
Pubblicazione: (2024)
di: Kim, Joongwon, et al.
Pubblicazione: (2024)
A Systematic Examination of Preference Learning through the Lens of Instruction-Following
di: Kim, Joongwon, et al.
Pubblicazione: (2024)
di: Kim, Joongwon, et al.
Pubblicazione: (2024)
Multi-Agent Computer Use
di: Koh, Jing Yu, et al.
Pubblicazione: (2026)
di: Koh, Jing Yu, et al.
Pubblicazione: (2026)
Critical Batch Size Revisited: A Simple Empirical Approach to Large-Batch Language Model Training
di: Merrill, William, et al.
Pubblicazione: (2025)
di: Merrill, William, et al.
Pubblicazione: (2025)
On the Impossibility of Retrain Equivalence in Machine Unlearning
di: Yu, Jiatong, et al.
Pubblicazione: (2025)
di: Yu, Jiatong, et al.
Pubblicazione: (2025)
Instruct-SkillMix: A Powerful Pipeline for LLM Instruction Tuning
di: Kaur, Simran, et al.
Pubblicazione: (2024)
di: Kaur, Simran, et al.
Pubblicazione: (2024)
Metacognitive Reuse: Turning Recurring LLM Reasoning Into Concise Behaviors
di: Didolkar, Aniket, et al.
Pubblicazione: (2025)
di: Didolkar, Aniket, et al.
Pubblicazione: (2025)
Machine Reading Comprehension using Case-based Reasoning
di: Thai, Dung, et al.
Pubblicazione: (2023)
di: Thai, Dung, et al.
Pubblicazione: (2023)
Contrastive Difference Predictive Coding
di: Zheng, Chongyi, et al.
Pubblicazione: (2023)
di: Zheng, Chongyi, et al.
Pubblicazione: (2023)
Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks
di: Jang, Lawrence Keunho, et al.
Pubblicazione: (2026)
di: Jang, Lawrence Keunho, et al.
Pubblicazione: (2026)
Tree Search for Language Model Agents
di: Koh, Jing Yu, et al.
Pubblicazione: (2024)
di: Koh, Jing Yu, et al.
Pubblicazione: (2024)
Can Models Learn Skill Composition from Examples?
di: Zhao, Haoyu, et al.
Pubblicazione: (2024)
di: Zhao, Haoyu, et al.
Pubblicazione: (2024)
Differentially Private Model Merging
di: Yin, Qichuan, et al.
Pubblicazione: (2026)
di: Yin, Qichuan, et al.
Pubblicazione: (2026)
APT: Adaptive Pruning and Tuning Pretrained Language Models for Efficient Training and Inference
di: Zhao, Bowen, et al.
Pubblicazione: (2024)
di: Zhao, Bowen, et al.
Pubblicazione: (2024)
Efficient Distributed Optimization under Heavy-Tailed Noise
di: Lee, Su Hyeong, et al.
Pubblicazione: (2025)
di: Lee, Su Hyeong, et al.
Pubblicazione: (2025)
A Statistical Framework for Data-dependent Retrieval-Augmented Models
di: Basu, Soumya, et al.
Pubblicazione: (2024)
di: Basu, Soumya, et al.
Pubblicazione: (2024)
New records for Queensland in Lindernia All. (Linderniaceae)
di: Wannan, B S
Pubblicazione: (2013)
di: Wannan, B S
Pubblicazione: (2013)
Keeping LLMs Aligned After Fine-tuning: The Crucial Role of Prompt Templates
di: Lyu, Kaifeng, et al.
Pubblicazione: (2024)
di: Lyu, Kaifeng, et al.
Pubblicazione: (2024)
Generalizing from SIMPLE to HARD Visual Reasoning: Can We Mitigate Modality Imbalance in VLMs?
di: Park, Simon, et al.
Pubblicazione: (2025)
di: Park, Simon, et al.
Pubblicazione: (2025)
Federation over Text: Insight Sharing for Multi-Agent Reasoning
di: Yao, Dixi, et al.
Pubblicazione: (2026)
di: Yao, Dixi, et al.
Pubblicazione: (2026)
Escaping the Cognitive Well: Efficient Competition Math with Off-the-Shelf Models
di: Dang, Xingyu, et al.
Pubblicazione: (2026)
di: Dang, Xingyu, et al.
Pubblicazione: (2026)
ScienceMeter: Tracking Scientific Knowledge Updates in Language Models
di: Wang, Yike, et al.
Pubblicazione: (2025)
di: Wang, Yike, et al.
Pubblicazione: (2025)
HREF: Human Response-Guided Evaluation of Instruction Following in Language Models
di: Lyu, Xinxi, et al.
Pubblicazione: (2024)
di: Lyu, Xinxi, et al.
Pubblicazione: (2024)
BTR: Binary Token Representations for Efficient Retrieval Augmented Language Models
di: Cao, Qingqing, et al.
Pubblicazione: (2023)
di: Cao, Qingqing, et al.
Pubblicazione: (2023)
Efficient Adaptive Federated Optimization
di: Lee, Su Hyeong, et al.
Pubblicazione: (2024)
di: Lee, Su Hyeong, et al.
Pubblicazione: (2024)
Dissecting Adversarial Robustness of Multimodal LM Agents
di: Wu, Chen Henry, et al.
Pubblicazione: (2024)
di: Wu, Chen Henry, et al.
Pubblicazione: (2024)
Social networks and the firm
di: Sanjeev Goyal
Pubblicazione: (2016)
di: Sanjeev Goyal
Pubblicazione: (2016)
Social networks and the firm
di: Sanjeev Goyal
Pubblicazione: (2016)
di: Sanjeev Goyal
Pubblicazione: (2016)
EvalTree: Profiling Language Model Weaknesses via Hierarchical Capability Trees
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2025)
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2025)
Cite Pretrain: Retrieval-Free Knowledge Attribution for Large Language Models
di: Huang, Yukun, et al.
Pubblicazione: (2025)
di: Huang, Yukun, et al.
Pubblicazione: (2025)
Meta-Reinforcement Learning with Self-Reflection for Agentic Search
di: Xiao, Teng, et al.
Pubblicazione: (2026)
di: Xiao, Teng, et al.
Pubblicazione: (2026)
Data Engineering for Scaling Language Models to 128K Context
di: Fu, Yao, et al.
Pubblicazione: (2024)
di: Fu, Yao, et al.
Pubblicazione: (2024)
Unlearning via Sparse Representations
di: Shah, Vedant, et al.
Pubblicazione: (2023)
di: Shah, Vedant, et al.
Pubblicazione: (2023)
Infini-gram: Scaling Unbounded n-gram Language Models to a Trillion Tokens
di: Liu, Jiacheng, et al.
Pubblicazione: (2024)
di: Liu, Jiacheng, et al.
Pubblicazione: (2024)
TurnWise: The Gap between Single- and Multi-turn Language Model Capabilities
di: Graf, Victoria, et al.
Pubblicazione: (2026)
di: Graf, Victoria, et al.
Pubblicazione: (2026)
Answer, Assemble, Ace: Understanding How LMs Answer Multiple Choice Questions
di: Wiegreffe, Sarah, et al.
Pubblicazione: (2024)
di: Wiegreffe, Sarah, et al.
Pubblicazione: (2024)
Train for Truth, Keep the Skills: Binary Retrieval-Augmented Reward Mitigates Hallucinations
di: Chen, Tong, et al.
Pubblicazione: (2025)
di: Chen, Tong, et al.
Pubblicazione: (2025)
MentorCollab: Selective Large-to-Small Inference-Time Guidance for Efficient Reasoning
di: Wang, Haojin, et al.
Pubblicazione: (2026)
di: Wang, Haojin, et al.
Pubblicazione: (2026)
Documenti analoghi
-
ASTRO: Teaching Language Models to Reason by Reflecting and Backtracking In-Context
di: Kim, Joongwon, et al.
Pubblicazione: (2025) -
Rethinking Thinking Tokens: LLMs as Improvement Operators
di: Madaan, Lovish, et al.
Pubblicazione: (2025) -
Husky: A Unified, Open-Source Language Agent for Multi-Step Reasoning
di: Kim, Joongwon, et al.
Pubblicazione: (2024) -
A Systematic Examination of Preference Learning through the Lens of Instruction-Following
di: Kim, Joongwon, et al.
Pubblicazione: (2024) -
Multi-Agent Computer Use
di: Koh, Jing Yu, et al.
Pubblicazione: (2026)