Timely Machine: Awareness of Time Makes Test-Time Scaling Agentic
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Yichuan, Li, Linyang, chen, Yongkang, Li, Peiji, Li, Xiaozhe, Guo, Qipeng, Lin, Dahua, Chen, Kai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mixing Expert Knowledge: Bring Human Thoughts Back To the Game of Go
von: Ma, Yichuan, et al.
Veröffentlicht: (2026)
von: Ma, Yichuan, et al.
Veröffentlicht: (2026)
What and When to Distill: Selective Hindsight Distillation for Multi-Turn Agents
von: Li, Xiaozhe, et al.
Veröffentlicht: (2026)
von: Li, Xiaozhe, et al.
Veröffentlicht: (2026)
Beyond Mode Collapse: Distribution Matching for Diverse Reasoning
von: Li, Xiaozhe, et al.
Veröffentlicht: (2026)
von: Li, Xiaozhe, et al.
Veröffentlicht: (2026)
UnitCoder: Scalable Iterative Code Synthesis with Unit Test Guidance
von: Ma, Yichuan, et al.
Veröffentlicht: (2025)
von: Ma, Yichuan, et al.
Veröffentlicht: (2025)
InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling
von: Li, Peiji, et al.
Veröffentlicht: (2025)
von: Li, Peiji, et al.
Veröffentlicht: (2025)
Agentic Test-Time Scaling for WebAgents
von: Lee, Nicholas, et al.
Veröffentlicht: (2026)
von: Lee, Nicholas, et al.
Veröffentlicht: (2026)
FastMCTS: A Simple Sampling Strategy for Data Synthesis
von: Li, Peiji, et al.
Veröffentlicht: (2025)
von: Li, Peiji, et al.
Veröffentlicht: (2025)
Scaling Up, Speeding Up: A Benchmark of Speculative Decoding for Efficient LLM Test-Time Scaling
von: Sun, Shengyin, et al.
Veröffentlicht: (2025)
von: Sun, Shengyin, et al.
Veröffentlicht: (2025)
Implicit Reward as the Bridge: A Unified View of SFT and DPO Connections
von: Wang, Bo, et al.
Veröffentlicht: (2025)
von: Wang, Bo, et al.
Veröffentlicht: (2025)
Scaling Test-Time Compute for Agentic Coding
von: Kim, Joongwon, et al.
Veröffentlicht: (2026)
von: Kim, Joongwon, et al.
Veröffentlicht: (2026)
EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems
von: He, Yufei, et al.
Veröffentlicht: (2025)
von: He, Yufei, et al.
Veröffentlicht: (2025)
Towards Thinking-Optimal Scaling of Test-Time Compute for LLM Reasoning
von: Yang, Wenkai, et al.
Veröffentlicht: (2025)
von: Yang, Wenkai, et al.
Veröffentlicht: (2025)
TL-GRPO: Turn-Level RL for Reasoning-Guided Iterative Optimization
von: Li, Peiji, et al.
Veröffentlicht: (2026)
von: Li, Peiji, et al.
Veröffentlicht: (2026)
Time is Not a Label: Continuous Phase Rotation for Temporal Knowledge Graphs and Agentic Memory
von: Li, Weixian Waylon, et al.
Veröffentlicht: (2026)
von: Li, Weixian Waylon, et al.
Veröffentlicht: (2026)
CTTS: Collective Test-Time Scaling
von: Song, Zhende, et al.
Veröffentlicht: (2025)
von: Song, Zhende, et al.
Veröffentlicht: (2025)
Inverse Scaling in Test-Time Compute
von: Gema, Aryo Pradipta, et al.
Veröffentlicht: (2025)
von: Gema, Aryo Pradipta, et al.
Veröffentlicht: (2025)
Balanced Data Sampling for Language Model Training with Clustering
von: Shao, Yunfan, et al.
Veröffentlicht: (2024)
von: Shao, Yunfan, et al.
Veröffentlicht: (2024)
TUMIX: Multi-Agent Test-Time Scaling with Tool-Use Mixture
von: Chen, Yongchao, et al.
Veröffentlicht: (2025)
von: Chen, Yongchao, et al.
Veröffentlicht: (2025)
Benchmark Test-Time Scaling of General LLM Agents
von: Li, Xiaochuan, et al.
Veröffentlicht: (2026)
von: Li, Xiaochuan, et al.
Veröffentlicht: (2026)
Parallel Test-Time Scaling for Latent Reasoning Models
von: You, Runyang, et al.
Veröffentlicht: (2025)
von: You, Runyang, et al.
Veröffentlicht: (2025)
Turn Waste into Worth: Rectifying Top-$k$ Router of MoE
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2024)
Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence
von: Ghasemabadi, Amirhosein, et al.
Veröffentlicht: (2025)
von: Ghasemabadi, Amirhosein, et al.
Veröffentlicht: (2025)
Scaling over Scaling: Exploring Test-Time Scaling Plateau in Large Reasoning Models
von: Wang, Jian, et al.
Veröffentlicht: (2025)
von: Wang, Jian, et al.
Veröffentlicht: (2025)
Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning
von: Bi, Zhenni, et al.
Veröffentlicht: (2024)
von: Bi, Zhenni, et al.
Veröffentlicht: (2024)
A Dual-Directional Context-Aware Test-Time Learning for Text Classification
von: Xu, Dong, et al.
Veröffentlicht: (2025)
von: Xu, Dong, et al.
Veröffentlicht: (2025)
METAL: A Multi-Agent Framework for Chart Generation with Test-Time Scaling
von: Li, Bingxuan, et al.
Veröffentlicht: (2025)
von: Li, Bingxuan, et al.
Veröffentlicht: (2025)
Provable Scaling Laws for the Test-Time Compute of Large Language Models
von: Chen, Yanxi, et al.
Veröffentlicht: (2024)
von: Chen, Yanxi, et al.
Veröffentlicht: (2024)
Generative AI Act II: Test Time Scaling Drives Cognition Engineering
von: Xia, Shijie, et al.
Veröffentlicht: (2025)
von: Xia, Shijie, et al.
Veröffentlicht: (2025)
BrowseConf: Confidence-Guided Test-Time Scaling for Web Agents
von: Ou, Litu, et al.
Veröffentlicht: (2025)
von: Ou, Litu, et al.
Veröffentlicht: (2025)
TimeSage-MT: A Multi-Turn Benchmark for Evaluating Agentic Time Series Reasoning
von: Kong, Yaxuan, et al.
Veröffentlicht: (2026)
von: Kong, Yaxuan, et al.
Veröffentlicht: (2026)
Adaptive Rectification Sampling for Test-Time Compute Scaling
von: Tan, Zhendong, et al.
Veröffentlicht: (2025)
von: Tan, Zhendong, et al.
Veröffentlicht: (2025)
EconProver: Towards More Economical Test-Time Scaling for Automated Theorem Proving
von: Li, Mukai, et al.
Veröffentlicht: (2025)
von: Li, Mukai, et al.
Veröffentlicht: (2025)
MTPChat: A Multimodal Time-Aware Persona Dataset for Conversational Agents
von: Yang, Wanqi, et al.
Veröffentlicht: (2025)
von: Yang, Wanqi, et al.
Veröffentlicht: (2025)
MatryoshkaThinking: Recursive Test-Time Scaling Enables Efficient Reasoning
von: Chen, Hongwei, et al.
Veröffentlicht: (2025)
von: Chen, Hongwei, et al.
Veröffentlicht: (2025)
Atom of Thoughts for Markov LLM Test-Time Scaling
von: Teng, Fengwei, et al.
Veröffentlicht: (2025)
von: Teng, Fengwei, et al.
Veröffentlicht: (2025)
Entropy-Gated Branching for Efficient Test-Time Reasoning
von: Li, Xianzhi, et al.
Veröffentlicht: (2025)
von: Li, Xianzhi, et al.
Veröffentlicht: (2025)
Rethinking the Unsolvable: When In-Context Search Meets Test-Time Scaling
von: Xia, Fanzeng, et al.
Veröffentlicht: (2025)
von: Xia, Fanzeng, et al.
Veröffentlicht: (2025)
SETS: Leveraging Self-Verification and Self-Correction for Improved Test-Time Scaling
von: Chen, Jiefeng, et al.
Veröffentlicht: (2025)
von: Chen, Jiefeng, et al.
Veröffentlicht: (2025)
AdaFuse: Adaptive Ensemble Decoding with Test-Time Scaling for LLMs
von: Cui, Chengming, et al.
Veröffentlicht: (2026)
von: Cui, Chengming, et al.
Veröffentlicht: (2026)
A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Mixing Expert Knowledge: Bring Human Thoughts Back To the Game of Go
von: Ma, Yichuan, et al.
Veröffentlicht: (2026) -
What and When to Distill: Selective Hindsight Distillation for Multi-Turn Agents
von: Li, Xiaozhe, et al.
Veröffentlicht: (2026) -
Beyond Mode Collapse: Distribution Matching for Diverse Reasoning
von: Li, Xiaozhe, et al.
Veröffentlicht: (2026) -
UnitCoder: Scalable Iterative Code Synthesis with Unit Test Guidance
von: Ma, Yichuan, et al.
Veröffentlicht: (2025) -
InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling
von: Li, Peiji, et al.
Veröffentlicht: (2025)