Thought calibration: Efficient and confident test-time scaling
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Menghua, Zhou, Cai, Bates, Stephen, Jaakkola, Tommi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Online Reasoning Calibration: Test-Time Training Enables Generalizable Conformal LLM Reasoning
by: Zhou, Cai, et al.
Published: (2026)
by: Zhou, Cai, et al.
Published: (2026)
Identifying biological perturbation targets through causal differential networks
by: Wu, Menghua, et al.
Published: (2024)
by: Wu, Menghua, et al.
Published: (2024)
Learning Diffusion Models with Flexible Representation Guidance
by: Wang, Chenyu, et al.
Published: (2025)
by: Wang, Chenyu, et al.
Published: (2025)
Continuously Tempered Diffusion Samplers
by: Erives, Ezra, et al.
Published: (2025)
by: Erives, Ezra, et al.
Published: (2025)
Harmonic Self-Conditioned Flow Matching for Multi-Ligand Docking and Binding Site Design
by: Stärk, Hannes, et al.
Published: (2023)
by: Stärk, Hannes, et al.
Published: (2023)
Extending confidence calibration to generalised measures of variation
by: Thompson, Andrew, et al.
Published: (2026)
by: Thompson, Andrew, et al.
Published: (2026)
Rethinking Diffusion Models with Symmetries through Canonicalization with Applications to Molecular Graph Generation
by: Zhou, Cai, et al.
Published: (2026)
by: Zhou, Cai, et al.
Published: (2026)
Sample, estimate, aggregate: A recipe for causal discovery foundation models
by: Wu, Menghua, et al.
Published: (2024)
by: Wu, Menghua, et al.
Published: (2024)
GLASS Flows: Transition Sampling for Alignment of Flow and Diffusion Models
by: Holderrieth, Peter, et al.
Published: (2025)
by: Holderrieth, Peter, et al.
Published: (2025)
Diamond Maps: Efficient Reward Alignment via Stochastic Flow Maps
by: Holderrieth, Peter, et al.
Published: (2026)
by: Holderrieth, Peter, et al.
Published: (2026)
DisCo-Diff: Enhancing Continuous Diffusion Models with Discrete Latents
by: Xu, Yilun, et al.
Published: (2024)
by: Xu, Yilun, et al.
Published: (2024)
Learning to refine domain knowledge for biological network inference
by: Li, Peiwen, et al.
Published: (2024)
by: Li, Peiwen, et al.
Published: (2024)
An Information Criterion for Controlled Disentanglement of Multimodal Data
by: Wang, Chenyu, et al.
Published: (2024)
by: Wang, Chenyu, et al.
Published: (2024)
Generator Matching: Generative modeling with arbitrary Markov processes
by: Holderrieth, Peter, et al.
Published: (2024)
by: Holderrieth, Peter, et al.
Published: (2024)
Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts
by: Ahmed, Ammar, et al.
Published: (2025)
by: Ahmed, Ammar, et al.
Published: (2025)
Isotonic Survival Regression: Calibrated Survival Distributions from Deep Cox Models
by: Jain, Anchit, et al.
Published: (2026)
by: Jain, Anchit, et al.
Published: (2026)
Fine-Tuning Discrete Diffusion Models via Reward Optimization with Applications to DNA and Protein Design
by: Wang, Chenyu, et al.
Published: (2024)
by: Wang, Chenyu, et al.
Published: (2024)
s1: Simple test-time scaling
by: Muennighoff, Niklas, et al.
Published: (2025)
by: Muennighoff, Niklas, et al.
Published: (2025)
A calibration test for evaluating set-based epistemic uncertainty representations
by: Jürgens, Mira, et al.
Published: (2025)
by: Jürgens, Mira, et al.
Published: (2025)
Deep Ensembles for Epistemic Uncertainty: A Frequentist Perspective
by: Jain, Anchit, et al.
Published: (2025)
by: Jain, Anchit, et al.
Published: (2025)
Next Semantic Scale Prediction via Hierarchical Diffusion Language Models
by: Zhou, Cai, et al.
Published: (2025)
by: Zhou, Cai, et al.
Published: (2025)
PROflow: An iterative refinement model for PROTAC-induced structure prediction
by: Qiang, Bo, et al.
Published: (2024)
by: Qiang, Bo, et al.
Published: (2024)
Value-Guided Search for Efficient Chain-of-Thought Reasoning
by: Wang, Kaiwen, et al.
Published: (2025)
by: Wang, Kaiwen, et al.
Published: (2025)
Re-FORC: Adaptive Reward Prediction for Efficient Chain-of-Thought Reasoning
by: Zabounidis, Renos, et al.
Published: (2025)
by: Zabounidis, Renos, et al.
Published: (2025)
Scaling Graph Chain-of-Thought Reasoning: A Multi-Agent Framework with Efficient LLM Serving
by: Huan, Chengying, et al.
Published: (2025)
by: Huan, Chengying, et al.
Published: (2025)
Think While You Generate: Discrete Diffusion with Planned Denoising
by: Liu, Sulin, et al.
Published: (2024)
by: Liu, Sulin, et al.
Published: (2024)
Everything of Thoughts: Defying the Law of Penrose Triangle for Thought Generation
by: Ding, Ruomeng, et al.
Published: (2023)
by: Ding, Ruomeng, et al.
Published: (2023)
Curriculum Learning for Efficient Chain-of-Thought Distillation via Structure-Aware Masking and GRPO
by: Yu, Bowen, et al.
Published: (2026)
by: Yu, Bowen, et al.
Published: (2026)
Nonparametric Distribution Regression Re-calibration
by: Jung, Ádám, et al.
Published: (2026)
by: Jung, Ádám, et al.
Published: (2026)
Keeping Medical AI Healthy and Trustworthy: A Review of Detection and Correction Methods for System Degradation
by: Guan, Hao, et al.
Published: (2025)
by: Guan, Hao, et al.
Published: (2025)
Forming Auxiliary High-confident Instance-level Loss to Promote Learning from Label Proportions
by: Ma, Tianhao, et al.
Published: (2024)
by: Ma, Tianhao, et al.
Published: (2024)
Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation
by: Wu, Yecheng, et al.
Published: (2026)
by: Wu, Yecheng, et al.
Published: (2026)
Policy-Guided Search on Tree-of-Thoughts for Efficient Problem Solving with Bounded Language Model Queries
by: Pendurkar, Sumedh, et al.
Published: (2026)
by: Pendurkar, Sumedh, et al.
Published: (2026)
Measuring multi-calibration
by: Guy, Ido, et al.
Published: (2025)
by: Guy, Ido, et al.
Published: (2025)
Continuous Thought Machines
by: Darlow, Luke, et al.
Published: (2025)
by: Darlow, Luke, et al.
Published: (2025)
Efficient 4D fMRI ASD Classification using Spatial-Temporal-Omics-based Learning Framework
by: Weng, Ziqiao, et al.
Published: (2025)
by: Weng, Ziqiao, et al.
Published: (2025)
Prediction-Powered Inference with Imputed Covariates and Nonuniform Sampling
by: Kluger, Dan M., et al.
Published: (2025)
by: Kluger, Dan M., et al.
Published: (2025)
Beyond Rewards in Reinforcement Learning for Cyber Defence
by: Bates, Elizabeth, et al.
Published: (2026)
by: Bates, Elizabeth, et al.
Published: (2026)
Stepwise Penalization for Length-Efficient Chain-of-Thought Reasoning
by: Li, Xintong, et al.
Published: (2026)
by: Li, Xintong, et al.
Published: (2026)
Test-time scaling of diffusions with flow maps
by: Sabour, Amirmojtaba, et al.
Published: (2025)
by: Sabour, Amirmojtaba, et al.
Published: (2025)
Similar Items
-
Online Reasoning Calibration: Test-Time Training Enables Generalizable Conformal LLM Reasoning
by: Zhou, Cai, et al.
Published: (2026) -
Identifying biological perturbation targets through causal differential networks
by: Wu, Menghua, et al.
Published: (2024) -
Learning Diffusion Models with Flexible Representation Guidance
by: Wang, Chenyu, et al.
Published: (2025) -
Continuously Tempered Diffusion Samplers
by: Erives, Ezra, et al.
Published: (2025) -
Harmonic Self-Conditioned Flow Matching for Multi-Ligand Docking and Binding Site Design
by: Stärk, Hannes, et al.
Published: (2023)