Trace Length is a Simple Uncertainty Signal in Reasoning Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Devic, Siddartha, Peale, Charlotte, Bradley, Arwen, Williamson, Sinead, Nakkiran, Preetum, Gollakota, Aravind |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Classifier-Free Guidance is a Predictor-Corrector
von: Bradley, Arwen, et al.
Veröffentlicht: (2024)
von: Bradley, Arwen, et al.
Veröffentlicht: (2024)
Flexible Routing via Uncertainty Decomposition
von: Peale, Charlotte, et al.
Veröffentlicht: (2026)
von: Peale, Charlotte, et al.
Veröffentlicht: (2026)
Step-by-Step Diffusion: An Elementary Tutorial
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2024)
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2024)
When is Multicalibration Post-Processing Necessary?
von: Hansen, Dutch, et al.
Veröffentlicht: (2024)
von: Hansen, Dutch, et al.
Veröffentlicht: (2024)
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
Vanishing Gradients in Reinforcement Finetuning of Language Models
von: Razin, Noam, et al.
Veröffentlicht: (2023)
von: Razin, Noam, et al.
Veröffentlicht: (2023)
Mechanisms of Projective Composition of Diffusion Models
von: Bradley, Arwen, et al.
Veröffentlicht: (2025)
von: Bradley, Arwen, et al.
Veröffentlicht: (2025)
Provable Uncertainty Decomposition via Higher-Order Calibration
von: Ahdritz, Gustaf, et al.
Veröffentlicht: (2024)
von: Ahdritz, Gustaf, et al.
Veröffentlicht: (2024)
Composition and Control with Distilled Energy Diffusion Models and Sequential Monte Carlo
von: Thornton, James, et al.
Veröffentlicht: (2025)
von: Thornton, James, et al.
Veröffentlicht: (2025)
When does a predictor know its own loss?
von: Gollakota, Aravind, et al.
Veröffentlicht: (2025)
von: Gollakota, Aravind, et al.
Veröffentlicht: (2025)
Revisiting Uncertainty Quantification Evaluation in Language Models: Spurious Interactions with Response Length Bias Results
von: Santilli, Andrea, et al.
Veröffentlicht: (2025)
von: Santilli, Andrea, et al.
Veröffentlicht: (2025)
Annotations Mitigate Post-Training Mode Collapse
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2026)
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2026)
Benign, Tempered, or Catastrophic: A Taxonomy of Overfitting
von: Mallinar, Neil, et al.
Veröffentlicht: (2022)
von: Mallinar, Neil, et al.
Veröffentlicht: (2022)
Tracing Uncertainty in Language Model "Reasoning"
von: Grünefeld, Nils, et al.
Veröffentlicht: (2026)
von: Grünefeld, Nils, et al.
Veröffentlicht: (2026)
Tracing the Traces: Latent Temporal Signals for Efficient and Accurate Reasoning
von: Vilas, Martina G., et al.
Veröffentlicht: (2025)
von: Vilas, Martina G., et al.
Veröffentlicht: (2025)
What do your logits know? (The answer may surprise you!)
von: Fedzechkina, Masha, et al.
Veröffentlicht: (2026)
von: Fedzechkina, Masha, et al.
Veröffentlicht: (2026)
Uncertainty-Aware Measurement of Scenario Suite Representativeness for Autonomous Systems
von: Chakherlou, Robab Aghazadeh, et al.
Veröffentlicht: (2025)
von: Chakherlou, Robab Aghazadeh, et al.
Veröffentlicht: (2025)
To Infinity and Beyond: Tool-Use Unlocks Length Generalization in State Space Models
von: Malach, Eran, et al.
Veröffentlicht: (2025)
von: Malach, Eran, et al.
Veröffentlicht: (2025)
Contradiction Detection in RAG Systems: Evaluating LLMs as Context Validators for Improved Information Consistency
von: Gokul, Vignesh, et al.
Veröffentlicht: (2025)
von: Gokul, Vignesh, et al.
Veröffentlicht: (2025)
Reasoning's Razor: Reasoning Improves Accuracy but Can Hurt Recall at Critical Operating Points in Safety and Hallucination Detection
von: Chegini, Atoosa, et al.
Veröffentlicht: (2025)
von: Chegini, Atoosa, et al.
Veröffentlicht: (2025)
From Passive Metric to Active Signal: The Evolving Role of Uncertainty Quantification in Large Language Models
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
Towards Agents That Know When They Don't Know: Uncertainty as a Control Signal for Structured Reasoning
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
Steering into New Embedding Spaces: Analyzing Cross-Lingual Alignment Induced by Model Interventions in Multilingual Language Models
von: Sundar, Anirudh, et al.
Veröffentlicht: (2025)
von: Sundar, Anirudh, et al.
Veröffentlicht: (2025)
ShorterBetter: Guiding Reasoning Models to Find Optimal Inference Length for Efficient Reasoning
von: Yi, Jingyang, et al.
Veröffentlicht: (2025)
von: Yi, Jingyang, et al.
Veröffentlicht: (2025)
The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models
von: Tan, Xue Wen, et al.
Veröffentlicht: (2025)
von: Tan, Xue Wen, et al.
Veröffentlicht: (2025)
Optimizing Length Compression in Large Reasoning Models
von: Cheng, Zhengxiang, et al.
Veröffentlicht: (2025)
von: Cheng, Zhengxiang, et al.
Veröffentlicht: (2025)
Adaptive Overclocking: Dynamic Control of Thinking Path Length via Real-Time Reasoning Signals
von: Jiang, Shuhao, et al.
Veröffentlicht: (2025)
von: Jiang, Shuhao, et al.
Veröffentlicht: (2025)
Boule or Baguette? A Study on Task Topology, Length Generalization, and the Benefit of Reasoning Traces
von: Tong, William L., et al.
Veröffentlicht: (2026)
von: Tong, William L., et al.
Veröffentlicht: (2026)
A Simple Generative Model of Logical Reasoning and Statistical Learning
von: Kido, Hiroyuki
Veröffentlicht: (2023)
von: Kido, Hiroyuki
Veröffentlicht: (2023)
The Impact of Reasoning Step Length on Large Language Models
von: Jin, Mingyu, et al.
Veröffentlicht: (2024)
von: Jin, Mingyu, et al.
Veröffentlicht: (2024)
A Theory for Length Generalization in Learning to Reason
von: Xiao, Changnan, et al.
Veröffentlicht: (2024)
von: Xiao, Changnan, et al.
Veröffentlicht: (2024)
ExpertLens: Activation steering features are highly interpretable
von: Fedzechkina, Masha, et al.
Veröffentlicht: (2025)
von: Fedzechkina, Masha, et al.
Veröffentlicht: (2025)
Listening with Language Models: Using LLMs to Collect and Interpret Classroom Feedback
von: Maram, Sai Siddartha, et al.
Veröffentlicht: (2025)
von: Maram, Sai Siddartha, et al.
Veröffentlicht: (2025)
Leash: Adaptive Length Penalty and Reward Shaping for Efficient Large Reasoning Model
von: Li, Yanhao, et al.
Veröffentlicht: (2025)
von: Li, Yanhao, et al.
Veröffentlicht: (2025)
Reasoning Traces Shape Outputs but Models Won't Say So
von: Hao, Yijie, et al.
Veröffentlicht: (2026)
von: Hao, Yijie, et al.
Veröffentlicht: (2026)
ReasoningShield: Safety Detection over Reasoning Traces of Large Reasoning Models
von: Li, Changyi, et al.
Veröffentlicht: (2025)
von: Li, Changyi, et al.
Veröffentlicht: (2025)
Learning Lifted STRIPS Models from Action Traces Alone: A Simple, General, and Scalable Solution
von: Gösgens, Jonas, et al.
Veröffentlicht: (2024)
von: Gösgens, Jonas, et al.
Veröffentlicht: (2024)
FOL-Traces: Verified First-Order Logic Reasoning Traces at Scale
von: Lee, Isabelle, et al.
Veröffentlicht: (2025)
von: Lee, Isabelle, et al.
Veröffentlicht: (2025)
Filtered Reasoning Score: Evaluating Reasoning Quality on a Model's Most-Confident Traces
von: Pathak, Manas, et al.
Veröffentlicht: (2026)
von: Pathak, Manas, et al.
Veröffentlicht: (2026)
AV-Dialog: Spoken Dialogue Models with Audio-Visual Input
von: Chen, Tuochao, et al.
Veröffentlicht: (2025)
von: Chen, Tuochao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Classifier-Free Guidance is a Predictor-Corrector
von: Bradley, Arwen, et al.
Veröffentlicht: (2024) -
Flexible Routing via Uncertainty Decomposition
von: Peale, Charlotte, et al.
Veröffentlicht: (2026) -
Step-by-Step Diffusion: An Elementary Tutorial
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2024) -
When is Multicalibration Post-Processing Necessary?
von: Hansen, Dutch, et al.
Veröffentlicht: (2024) -
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)