Guardado en:
| Autores principales: | Petullo, James, George, Sonny, Cashman, Dylan, Xue, Nianwen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2605.08070 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CA-SQL: Complexity-Aware Inference Time Reasoning for Text-to-SQL via Exploration and Compute Budget Allocation
por: Petullo, James, et al.
Publicado: (2026)
por: Petullo, James, et al.
Publicado: (2026)
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction
por: George, Sonny, et al.
Publicado: (2024)
por: George, Sonny, et al.
Publicado: (2024)
Confidence Improves Self-Consistency in LLMs
por: Taubenfeld, Amir, et al.
Publicado: (2025)
por: Taubenfeld, Amir, et al.
Publicado: (2025)
Latent Self-Consistency for Reliable Majority-Set Selection in Short- and Long-Answer Reasoning
por: Oh, Jungsuk, et al.
Publicado: (2025)
por: Oh, Jungsuk, et al.
Publicado: (2025)
The Chain Holds, the Answer Folds: Trace-Answer Dissociation in Reasoning Models Under Adversarial Pressure
por: Li, Yubo, et al.
Publicado: (2026)
por: Li, Yubo, et al.
Publicado: (2026)
How Do Answer Tokens Read Reasoning Traces? Self-Reading Patterns in Thinking LLMs for Quantitative Reasoning
por: Chen, Haoyang, et al.
Publicado: (2026)
por: Chen, Haoyang, et al.
Publicado: (2026)
MPN: Leveraging Multilingual Patch Neuron for Cross-lingual Model Editing
por: Si, Nianwen, et al.
Publicado: (2024)
por: Si, Nianwen, et al.
Publicado: (2024)
Reasoning Model Unlearning: Forgetting Traces, Not Just Answers, While Preserving Reasoning Skills
por: Wang, Changsheng, et al.
Publicado: (2025)
por: Wang, Changsheng, et al.
Publicado: (2025)
Maximizing Confidence Alone Improves Reasoning
por: Prabhudesai, Mihir, et al.
Publicado: (2025)
por: Prabhudesai, Mihir, et al.
Publicado: (2025)
Tiny-QMoE
por: Cashman, Jack, et al.
Publicado: (2025)
por: Cashman, Jack, et al.
Publicado: (2025)
Filtered Reasoning Score: Evaluating Reasoning Quality on a Model's Most-Confident Traces
por: Pathak, Manas, et al.
Publicado: (2026)
por: Pathak, Manas, et al.
Publicado: (2026)
Reasoning or Fluency? Dissecting Probabilistic Confidence in Best-of-N Selection
por: Kim, Hojin, et al.
Publicado: (2026)
por: Kim, Hojin, et al.
Publicado: (2026)
Revisiting Self-Consistency from Dynamic Distributional Alignment Perspective on Answer Aggregation
por: Li, Yiwei, et al.
Publicado: (2025)
por: Li, Yiwei, et al.
Publicado: (2025)
Self-Training Meets Consistency: Improving LLMs' Reasoning with Consistency-Driven Rationale Evaluation
por: Lee, Jaehyeok, et al.
Publicado: (2024)
por: Lee, Jaehyeok, et al.
Publicado: (2024)
Improving Score Reliability of Multiple Choice Benchmarks with Consistency Evaluation and Altered Answer Choices
por: Cavalin, Paulo, et al.
Publicado: (2025)
por: Cavalin, Paulo, et al.
Publicado: (2025)
Predicting Winning Captions for Weekly New Yorker Comics
por: Cao, Stanley, et al.
Publicado: (2024)
por: Cao, Stanley, et al.
Publicado: (2024)
A Generalised Approach for Encoding and Reasoning with Qualitative Theories in Answer Set Programming
por: Baryannis, George, et al.
Publicado: (2020)
por: Baryannis, George, et al.
Publicado: (2020)
Sandwich Reasoning: An Answer-Reasoning-Answer Approach for Low-Latency Query Correction
por: Zhang, Chen, et al.
Publicado: (2026)
por: Zhang, Chen, et al.
Publicado: (2026)
Reflective Confidence: Correcting Reasoning Flaws via Online Self-Correction
por: Zeng, Qinglin, et al.
Publicado: (2025)
por: Zeng, Qinglin, et al.
Publicado: (2025)
PiCSAR: Probabilistic Confidence Selection And Ranking for Reasoning Chains
por: Leang, Joshua Ong Jun, et al.
Publicado: (2025)
por: Leang, Joshua Ong Jun, et al.
Publicado: (2025)
Beyond the Last Answer: Your Reasoning Trace Uncovers More than You Think
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2025)
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2025)
The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models
por: Tan, Xue Wen, et al.
Publicado: (2025)
por: Tan, Xue Wen, et al.
Publicado: (2025)
Does Inference Scaling Improve Reasoning Faithfulness? A Multi-Model Analysis of Self-Consistency Tradeoffs
por: Mehta, Deep
Publicado: (2026)
por: Mehta, Deep
Publicado: (2026)
Diagnosing Harmful Continuation in Answer-Correct Long-CoT Training Traces
por: He, Chen, et al.
Publicado: (2026)
por: He, Chen, et al.
Publicado: (2026)
Roundtable Policy: Confidence-Weighted-Consensus Aggregation Improves Multi-Agent-System Reasoning
por: Yao, Yu, et al.
Publicado: (2025)
por: Yao, Yu, et al.
Publicado: (2025)
FANS -- Formal Answer Selection for Natural Language Math Reasoning Using Lean4
por: Yao, Jiarui, et al.
Publicado: (2025)
por: Yao, Jiarui, et al.
Publicado: (2025)
Self-Consistency Boosts Calibration for Math Reasoning
por: Wang, Ante, et al.
Publicado: (2024)
por: Wang, Ante, et al.
Publicado: (2024)
Debiased Multimodal Personality Understanding through Dual Causal Intervention
por: Zhu, Yangfu, et al.
Publicado: (2026)
por: Zhu, Yangfu, et al.
Publicado: (2026)
ReasonOps: Operator Segmentation for LLM Reasoning Traces
por: Lee, Daniel, et al.
Publicado: (2026)
por: Lee, Daniel, et al.
Publicado: (2026)
Self-Consistency of the Internal Reward Models Improves Self-Rewarding Language Models
por: Zhou, Xin, et al.
Publicado: (2025)
por: Zhou, Xin, et al.
Publicado: (2025)
Plantain: Plan-Answer Interleaved Reasoning
por: Liang, Anthony, et al.
Publicado: (2025)
por: Liang, Anthony, et al.
Publicado: (2025)
Optimal Bayesian Stopping for Efficient Inference of Consistent LLM Answers
por: Huang, Jingkai, et al.
Publicado: (2026)
por: Huang, Jingkai, et al.
Publicado: (2026)
Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards
por: Kim, Seungwook, et al.
Publicado: (2026)
por: Kim, Seungwook, et al.
Publicado: (2026)
Reasoning about Study Regulations in Answer Set Programming
por: Hahn, Susana, et al.
Publicado: (2024)
por: Hahn, Susana, et al.
Publicado: (2024)
S2Vec: Self-Supervised Geospatial Embeddings for the Built Environment
por: Choudhury, Shushman, et al.
Publicado: (2025)
por: Choudhury, Shushman, et al.
Publicado: (2025)
CASK: Core-Aware Selective KV Compression for Reasoning Traces
por: Kim, Buseong, et al.
Publicado: (2026)
por: Kim, Buseong, et al.
Publicado: (2026)
Beyond Self-Consistency: Ensemble Reasoning Boosts Consistency and Accuracy of LLMs in Cancer Staging
por: Chang, Chia-Hsuan, et al.
Publicado: (2024)
por: Chang, Chia-Hsuan, et al.
Publicado: (2024)
Do Cognitively Interpretable Reasoning Traces Improve LLM Performance?
por: Bhambri, Siddhant, et al.
Publicado: (2025)
por: Bhambri, Siddhant, et al.
Publicado: (2025)
Incentivizing LLMs to Self-Verify Their Answers
por: Zhang, Fuxiang, et al.
Publicado: (2025)
por: Zhang, Fuxiang, et al.
Publicado: (2025)
Reasoning Aware Self-Consistency: Leveraging Reasoning Paths for Efficient LLM Sampling
por: Wan, Guangya, et al.
Publicado: (2024)
por: Wan, Guangya, et al.
Publicado: (2024)
Ejemplares similares
-
CA-SQL: Complexity-Aware Inference Time Reasoning for Text-to-SQL via Exploration and Compute Budget Allocation
por: Petullo, James, et al.
Publicado: (2026) -
Probing the Capacity of Language Model Agents to Operationalize Disparate Experiential Context Despite Distraction
por: George, Sonny, et al.
Publicado: (2024) -
Confidence Improves Self-Consistency in LLMs
por: Taubenfeld, Amir, et al.
Publicado: (2025) -
Latent Self-Consistency for Reliable Majority-Set Selection in Short- and Long-Answer Reasoning
por: Oh, Jungsuk, et al.
Publicado: (2025) -
The Chain Holds, the Answer Folds: Trace-Answer Dissociation in Reasoning Models Under Adversarial Pressure
por: Li, Yubo, et al.
Publicado: (2026)