When to Ponder: Adaptive Compute Allocation for Code Generation via Test-Time Training
Fuente:
arXiv
Guardado en:
| Autor principal: | Sim, Gihyeon |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
FLOP-Efficient Training: Early Stopping Based on Test-Time Compute Awareness
por: Amer, Hossam, et al.
Publicado: (2026)
por: Amer, Hossam, et al.
Publicado: (2026)
Zero-Overhead Introspection for Adaptive Test-Time Compute
por: Manvi, Rohin, et al.
Publicado: (2025)
por: Manvi, Rohin, et al.
Publicado: (2025)
PaT: Planning-after-Trial for Efficient Test-Time Code Generation
por: Yoon, Youngsik, et al.
Publicado: (2026)
por: Yoon, Youngsik, et al.
Publicado: (2026)
CodeACT: Code Adaptive Compute-efficient Tuning Framework for Code LLMs
por: Lv, Weijie, et al.
Publicado: (2024)
por: Lv, Weijie, et al.
Publicado: (2024)
Scaling Test-Time Compute for Agentic Coding
por: Kim, Joongwon, et al.
Publicado: (2026)
por: Kim, Joongwon, et al.
Publicado: (2026)
Dynamic Cheatsheet: Test-Time Learning with Adaptive Memory
por: Suzgun, Mirac, et al.
Publicado: (2025)
por: Suzgun, Mirac, et al.
Publicado: (2025)
Learning How Hard to Think: Input-Adaptive Allocation of LM Computation
por: Damani, Mehul, et al.
Publicado: (2024)
por: Damani, Mehul, et al.
Publicado: (2024)
Test-Time Scaling Makes Overtraining Compute-Optimal
por: Roberts, Nicholas, et al.
Publicado: (2026)
por: Roberts, Nicholas, et al.
Publicado: (2026)
BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute
por: Ding, Dujian, et al.
Publicado: (2025)
por: Ding, Dujian, et al.
Publicado: (2025)
In-Place Test-Time Training
por: Feng, Guhao, et al.
Publicado: (2026)
por: Feng, Guhao, et al.
Publicado: (2026)
CoSPlay: Cooperative Self-Play at Test-Time with Self-Generated Code and Unit Test
por: Hu, Zhangyi, et al.
Publicado: (2026)
por: Hu, Zhangyi, et al.
Publicado: (2026)
Test-Time Training on Nearest Neighbors for Large Language Models
por: Hardt, Moritz, et al.
Publicado: (2023)
por: Hardt, Moritz, et al.
Publicado: (2023)
Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision
por: Xi, Zhiheng, et al.
Publicado: (2024)
por: Xi, Zhiheng, et al.
Publicado: (2024)
FEval-TTC: Fair Evaluation Protocol for Test-Time Compute
por: Rumiantsev, Pavel, et al.
Publicado: (2025)
por: Rumiantsev, Pavel, et al.
Publicado: (2025)
Scaling Test-Time Compute Without Verification or RL is Suboptimal
por: Setlur, Amrith, et al.
Publicado: (2025)
por: Setlur, Amrith, et al.
Publicado: (2025)
Log-Augmented Generation: Scaling Test-Time Reasoning with Reusable Computation
por: Chen, Peter Baile, et al.
Publicado: (2025)
por: Chen, Peter Baile, et al.
Publicado: (2025)
Breaking Symmetry When Training Transformers
por: Zuo, Chunsheng, et al.
Publicado: (2024)
por: Zuo, Chunsheng, et al.
Publicado: (2024)
Optimizing Test-Time Compute via Meta Reinforcement Fine-Tuning
por: Qu, Yuxiao, et al.
Publicado: (2025)
por: Qu, Yuxiao, et al.
Publicado: (2025)
Primal Generation, Dual Judgment: Self-Training from Test-Time Scaling
por: Jiao, Yizhu, et al.
Publicado: (2026)
por: Jiao, Yizhu, et al.
Publicado: (2026)
Test-Time Scaling with Reflective Generative Model
por: Wang, Zixiao, et al.
Publicado: (2025)
por: Wang, Zixiao, et al.
Publicado: (2025)
Self-Trained Verification for Training- and Test-Time Self-Improvement
por: Wu, Chen Henry, et al.
Publicado: (2026)
por: Wu, Chen Henry, et al.
Publicado: (2026)
Adaptive Test-Time Reasoning via Reward-Guided Dual-Phase Search
por: Cui, Yingqian, et al.
Publicado: (2025)
por: Cui, Yingqian, et al.
Publicado: (2025)
TABED: Test-Time Adaptive Ensemble Drafting for Robust Speculative Decoding in LVLMs
por: Lee, Minjae, et al.
Publicado: (2026)
por: Lee, Minjae, et al.
Publicado: (2026)
PonderLM-3: Adaptive Token-Wise Pondering with Differentiable Masking
por: Li, He, et al.
Publicado: (2026)
por: Li, He, et al.
Publicado: (2026)
Test-Time Training Done Right
por: Zhang, Tianyuan, et al.
Publicado: (2025)
por: Zhang, Tianyuan, et al.
Publicado: (2025)
e3: Learning to Explore Enables Extrapolation of Test-Time Compute for LLMs
por: Setlur, Amrith, et al.
Publicado: (2025)
por: Setlur, Amrith, et al.
Publicado: (2025)
Rethinking the Unsolvable: When In-Context Search Meets Test-Time Scaling
por: Xia, Fanzeng, et al.
Publicado: (2025)
por: Xia, Fanzeng, et al.
Publicado: (2025)
Let's (not) just put things in Context: Test-Time Training for Long-Context LLMs
por: Bansal, Rachit, et al.
Publicado: (2025)
por: Bansal, Rachit, et al.
Publicado: (2025)
ATLAS: Adaptive Test-Time Latent Steering with External Verifiers for Enhancing LLMs Reasoning
por: Nguyen, Tuc, et al.
Publicado: (2026)
por: Nguyen, Tuc, et al.
Publicado: (2026)
Personalised Distillation: Empowering Open-Sourced LLMs with Adaptive Learning for Code Generation
por: Chen, Hailin, et al.
Publicado: (2023)
por: Chen, Hailin, et al.
Publicado: (2023)
Latency and Token-Aware Test-Time Compute
por: Huang, Jenny Y., et al.
Publicado: (2025)
por: Huang, Jenny Y., et al.
Publicado: (2025)
MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention
por: MiniMax, et al.
Publicado: (2025)
por: MiniMax, et al.
Publicado: (2025)
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
por: Geiping, Jonas, et al.
Publicado: (2025)
por: Geiping, Jonas, et al.
Publicado: (2025)
VerifierQ: Enhancing LLM Test Time Compute with Q-Learning-based Verifiers
por: Qi, Jianing, et al.
Publicado: (2024)
por: Qi, Jianing, et al.
Publicado: (2024)
Test-Time Detoxification without Training or Learning Anything
por: Saglam, Baturay, et al.
Publicado: (2026)
por: Saglam, Baturay, et al.
Publicado: (2026)
UT-ACA: Uncertainty-Triggered Adaptive Context Allocation for Long-Context Inference
por: Zhou, Lang, et al.
Publicado: (2026)
por: Zhou, Lang, et al.
Publicado: (2026)
When To Solve, When To Verify: Compute-Optimal Problem Solving and Generative Verification for LLM Reasoning
por: Singhi, Nishad, et al.
Publicado: (2025)
por: Singhi, Nishad, et al.
Publicado: (2025)
HyperAdaLoRA: Accelerating LoRA Rank Allocation During Training via Hypernetworks without Sacrificing Performance
por: Zhang, Hao, et al.
Publicado: (2025)
por: Zhang, Hao, et al.
Publicado: (2025)
Plan*RAG: Efficient Test-Time Planning for Retrieval Augmented Generation
por: Verma, Prakhar, et al.
Publicado: (2024)
por: Verma, Prakhar, et al.
Publicado: (2024)
SR-TTT: Surprisal-Aware Residual Test-Time Training
por: P, Swamynathan V
Publicado: (2026)
por: P, Swamynathan V
Publicado: (2026)
Ejemplares similares
-
FLOP-Efficient Training: Early Stopping Based on Test-Time Compute Awareness
por: Amer, Hossam, et al.
Publicado: (2026) -
Zero-Overhead Introspection for Adaptive Test-Time Compute
por: Manvi, Rohin, et al.
Publicado: (2025) -
PaT: Planning-after-Trial for Efficient Test-Time Code Generation
por: Yoon, Youngsik, et al.
Publicado: (2026) -
CodeACT: Code Adaptive Compute-efficient Tuning Framework for Code LLMs
por: Lv, Weijie, et al.
Publicado: (2024) -
Scaling Test-Time Compute for Agentic Coding
por: Kim, Joongwon, et al.
Publicado: (2026)