Rethinking Optimal Verification Granularity for Compute-Efficient Test-Time Scaling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Hao Mark, Lu, Guanxi, Okoshi, Yasuyuki, Mo, Zhiwen, Motomura, Masato, Fan, Hongxiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Context Memorization for Efficient Long Context Generation
von: Okoshi, Yasuyuki, et al.
Veröffentlicht: (2026)
von: Okoshi, Yasuyuki, et al.
Veröffentlicht: (2026)
AQPIM: Breaking the PIM Capacity Wall for LLMs with In-Memory Activation Quantization
von: Matsushima, Kosuke, et al.
Veröffentlicht: (2026)
von: Matsushima, Kosuke, et al.
Veröffentlicht: (2026)
The Strong Lottery Ticket Hypothesis for Multi-Head Attention Mechanisms
von: Otsuka, Hikari, et al.
Veröffentlicht: (2025)
von: Otsuka, Hikari, et al.
Veröffentlicht: (2025)
FastTTS: Accelerating Test-Time Scaling for Edge LLM Reasoning
von: Chen, Hao Mark, et al.
Veröffentlicht: (2025)
von: Chen, Hao Mark, et al.
Veröffentlicht: (2025)
Binary Quadratic Quantization: Beyond First-Order Quantization for Real-Valued Matrix Compression
von: Kuroki, Kyo, et al.
Veröffentlicht: (2025)
von: Kuroki, Kyo, et al.
Veröffentlicht: (2025)
AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Size
von: Lu, Guanxi, et al.
Veröffentlicht: (2025)
von: Lu, Guanxi, et al.
Veröffentlicht: (2025)
Partially Frozen Random Networks Contain Compact Strong Lottery Tickets
von: Otsuka, Hikari, et al.
Veröffentlicht: (2024)
von: Otsuka, Hikari, et al.
Veröffentlicht: (2024)
FW-Merging: Scaling Model Merging with Frank-Wolfe Optimization
von: Chen, Hao Mark, et al.
Veröffentlicht: (2025)
von: Chen, Hao Mark, et al.
Veröffentlicht: (2025)
Rethinking Fine-Tuning when Scaling Test-Time Compute: Limiting Confidence Improves Mathematical Reasoning
von: Chen, Feng, et al.
Veröffentlicht: (2025)
von: Chen, Feng, et al.
Veröffentlicht: (2025)
Every Rollout Counts: Optimal Resource Allocation for Efficient Test-Time Scaling
von: Wang, Xinglin, et al.
Veröffentlicht: (2025)
von: Wang, Xinglin, et al.
Veröffentlicht: (2025)
Enhancing Trustworthiness with Mixed Precision: Benchmarks, Opportunities, and Challenges
von: Lu, Guanxi, et al.
Veröffentlicht: (2025)
von: Lu, Guanxi, et al.
Veröffentlicht: (2025)
Rethinking the Unsolvable: When In-Context Search Meets Test-Time Scaling
von: Xia, Fanzeng, et al.
Veröffentlicht: (2025)
von: Xia, Fanzeng, et al.
Veröffentlicht: (2025)
Iterative Deepening Sampling as Efficient Test-Time Scaling
von: Chen, Weizhe, et al.
Veröffentlicht: (2025)
von: Chen, Weizhe, et al.
Veröffentlicht: (2025)
Detecting and Mitigating the Correct-Answer Extinction Window in Test-Time Reinforcement Learning with Majority Voting
von: Lin, Hongxiang, et al.
Veröffentlicht: (2026)
von: Lin, Hongxiang, et al.
Veröffentlicht: (2026)
Labels Matter More Than Models: Rethinking the Unsupervised Paradigm in Time Series Anomaly Detection
von: Zhong, Zhijie, et al.
Veröffentlicht: (2025)
von: Zhong, Zhijie, et al.
Veröffentlicht: (2025)
Advancing AI-assisted Hardware Design with Hierarchical Decentralized Training and Personalized Inference-Time Optimization
von: Chen, Hao Mark, et al.
Veröffentlicht: (2025)
von: Chen, Hao Mark, et al.
Veröffentlicht: (2025)
Beyond the Frontier: Stochastic Backtracking for Efficient Test-Time Scaling
von: Tran, Dao, et al.
Veröffentlicht: (2026)
von: Tran, Dao, et al.
Veröffentlicht: (2026)
GLIMPSE: Holistic Cross-Modal Explainability for Large Vision-Language Models
von: Shen, Guanxi
Veröffentlicht: (2025)
von: Shen, Guanxi
Veröffentlicht: (2025)
Code Generation by Differential Test Time Scaling
von: He, Yifeng, et al.
Veröffentlicht: (2026)
von: He, Yifeng, et al.
Veröffentlicht: (2026)
Multi-Agent Verification: Scaling Test-Time Compute with Multiple Verifiers
von: Lifshitz, Shalev, et al.
Veröffentlicht: (2025)
von: Lifshitz, Shalev, et al.
Veröffentlicht: (2025)
Sample, Scrutinize and Scale: Effective Inference-Time Search by Scaling Verification
von: Zhao, Eric, et al.
Veröffentlicht: (2025)
von: Zhao, Eric, et al.
Veröffentlicht: (2025)
Provable Scaling Laws for the Test-Time Compute of Large Language Models
von: Chen, Yanxi, et al.
Veröffentlicht: (2024)
von: Chen, Yanxi, et al.
Veröffentlicht: (2024)
IsoCompute Playbook: Optimally Scaling Sampling Compute for LLM RL
von: Cheng, Zhoujun, et al.
Veröffentlicht: (2026)
von: Cheng, Zhoujun, et al.
Veröffentlicht: (2026)
Scaling Test-Time Compute for Agentic Coding
von: Kim, Joongwon, et al.
Veröffentlicht: (2026)
von: Kim, Joongwon, et al.
Veröffentlicht: (2026)
Log-Augmented Generation: Scaling Test-Time Reasoning with Reusable Computation
von: Chen, Peter Baile, et al.
Veröffentlicht: (2025)
von: Chen, Peter Baile, et al.
Veröffentlicht: (2025)
Self-Trained Verification for Training- and Test-Time Self-Improvement
von: Wu, Chen Henry, et al.
Veröffentlicht: (2026)
von: Wu, Chen Henry, et al.
Veröffentlicht: (2026)
Towards Theoretical Understanding of Transformer Test-Time Computing: Investigation on In-Context Linear Regression
von: Chen, Xingwu, et al.
Veröffentlicht: (2025)
von: Chen, Xingwu, et al.
Veröffentlicht: (2025)
Budget-aware Test-time Scaling via Discriminative Verification
von: Montgomery, Kyle, et al.
Veröffentlicht: (2025)
von: Montgomery, Kyle, et al.
Veröffentlicht: (2025)
PETS: A Principled Framework Towards Optimal Trajectory Allocation for Efficient Test-Time Self-Consistency
von: Liu, Zhangyi, et al.
Veröffentlicht: (2026)
von: Liu, Zhangyi, et al.
Veröffentlicht: (2026)
Improving LLM Reasoning through Scaling Inference Computation with Collaborative Verification
von: Liang, Zhenwen, et al.
Veröffentlicht: (2024)
von: Liang, Zhenwen, et al.
Veröffentlicht: (2024)
Rethinking the Role of Prompting Strategies in LLM Test-Time Scaling: A Perspective of Probability Theory
von: Liu, Yexiang, et al.
Veröffentlicht: (2025)
von: Liu, Yexiang, et al.
Veröffentlicht: (2025)
Surprisal-Guided Selection: Compute-Optimal Test-Time Strategies for Execution-Grounded Code Generation
von: Barnes, Jarrod
Veröffentlicht: (2026)
von: Barnes, Jarrod
Veröffentlicht: (2026)
Efficient Test-Time Scaling via Self-Calibration
von: Huang, Chengsong, et al.
Veröffentlicht: (2025)
von: Huang, Chengsong, et al.
Veröffentlicht: (2025)
Beyond Memorization: Extending Reasoning Depth with Recurrence, Memory and Test-Time Compute Scaling
von: Rodkin, Ivan, et al.
Veröffentlicht: (2025)
von: Rodkin, Ivan, et al.
Veröffentlicht: (2025)
When More Thinking Hurts: Overthinking in LLM Test-Time Compute Scaling
von: Zhou, Shu, et al.
Veröffentlicht: (2026)
von: Zhou, Shu, et al.
Veröffentlicht: (2026)
Scales++: Compute Efficient Evaluation Subset Selection with Cognitive Scales Embeddings
von: Bean, Andrew M., et al.
Veröffentlicht: (2025)
von: Bean, Andrew M., et al.
Veröffentlicht: (2025)
CoScale-RL: Efficient Post-Training by Co-Scaling Data and Computation
von: Chen, Yutong, et al.
Veröffentlicht: (2026)
von: Chen, Yutong, et al.
Veröffentlicht: (2026)
Strategic Scaling of Test-Time Compute: A Bandit Learning Approach
von: Zuo, Bowen, et al.
Veröffentlicht: (2025)
von: Zuo, Bowen, et al.
Veröffentlicht: (2025)
S*: Test Time Scaling for Code Generation
von: Li, Dacheng, et al.
Veröffentlicht: (2025)
von: Li, Dacheng, et al.
Veröffentlicht: (2025)
MG-TSD: Multi-Granularity Time Series Diffusion Models with Guided Learning Process
von: Fan, Xinyao, et al.
Veröffentlicht: (2024)
von: Fan, Xinyao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Context Memorization for Efficient Long Context Generation
von: Okoshi, Yasuyuki, et al.
Veröffentlicht: (2026) -
AQPIM: Breaking the PIM Capacity Wall for LLMs with In-Memory Activation Quantization
von: Matsushima, Kosuke, et al.
Veröffentlicht: (2026) -
The Strong Lottery Ticket Hypothesis for Multi-Head Attention Mechanisms
von: Otsuka, Hikari, et al.
Veröffentlicht: (2025) -
FastTTS: Accelerating Test-Time Scaling for Edge LLM Reasoning
von: Chen, Hao Mark, et al.
Veröffentlicht: (2025) -
Binary Quadratic Quantization: Beyond First-Order Quantization for Real-Valued Matrix Compression
von: Kuroki, Kyo, et al.
Veröffentlicht: (2025)