Reasoning or Fluency? Dissecting Probabilistic Confidence in Best-of-N Selection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Hojin, Kim, Jaehyung |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated Generation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2026)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2026)
EMCEE: Improving Multilingual Capability of LLMs via Bridging Knowledge and Reasoning with Extracted Synthetic Multilingual Context
von: Koo, Hamin, et al.
Veröffentlicht: (2025)
von: Koo, Hamin, et al.
Veröffentlicht: (2025)
Learning to Correct for QA Reasoning with Black-box LLMs
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
Revisiting the UID Hypothesis in LLM Reasoning Traces
von: Gwak, Minju, et al.
Veröffentlicht: (2025)
von: Gwak, Minju, et al.
Veröffentlicht: (2025)
Structural Reasoning Improves Molecular Understanding of LLM
von: Jang, Yunhui, et al.
Veröffentlicht: (2024)
von: Jang, Yunhui, et al.
Veröffentlicht: (2024)
PiCSAR: Probabilistic Confidence Selection And Ranking for Reasoning Chains
von: Leang, Joshua Ong Jun, et al.
Veröffentlicht: (2025)
von: Leang, Joshua Ong Jun, et al.
Veröffentlicht: (2025)
Revisiting the Uniform Information Density Hypothesis in LLM Reasoning
von: Gwak, Minju, et al.
Veröffentlicht: (2025)
von: Gwak, Minju, et al.
Veröffentlicht: (2025)
InterPol: De-anonymizing LM Arena via Interpolated Preference Learning
von: Cho, Minsung, et al.
Veröffentlicht: (2026)
von: Cho, Minsung, et al.
Veröffentlicht: (2026)
Reasoning Models Better Express Their Confidence
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
Align to Misalign: Automatic LLM Jailbreak with Meta-Optimized LLM Judges
von: Koo, Hamin, et al.
Veröffentlicht: (2025)
von: Koo, Hamin, et al.
Veröffentlicht: (2025)
Enhancing Instruction Following of LLMs via Activation Steering with Dynamic Rejection
von: Kang, Minjae, et al.
Veröffentlicht: (2026)
von: Kang, Minjae, et al.
Veröffentlicht: (2026)
TiTok: Transfer Token-level Knowledge via Contrastive Excess to Transplant LoRA
von: Jung, Chanjoo, et al.
Veröffentlicht: (2025)
von: Jung, Chanjoo, et al.
Veröffentlicht: (2025)
SelectLLM: Can LLMs Select Important Instructions to Annotate?
von: Parkar, Ritik Sachin, et al.
Veröffentlicht: (2024)
von: Parkar, Ritik Sachin, et al.
Veröffentlicht: (2024)
Robot-R1: Reinforcement Learning for Enhanced Embodied Reasoning in Robotics
von: Kim, Dongyoung, et al.
Veröffentlicht: (2025)
von: Kim, Dongyoung, et al.
Veröffentlicht: (2025)
Gap-K%: Measuring Top-1 Prediction Gap for Detecting Pretraining Data
von: Kwak, Minseo, et al.
Veröffentlicht: (2026)
von: Kwak, Minseo, et al.
Veröffentlicht: (2026)
Few-shot Personalization of LLMs with Mis-aligned Responses
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
RoboAlign: Learning Test-Time Reasoning for Language-Action Alignment in Vision-Language-Action Models
von: Kim, Dongyoung, et al.
Veröffentlicht: (2026)
von: Kim, Dongyoung, et al.
Veröffentlicht: (2026)
Optimized Feature Generation for Tabular Data via LLMs with Decision Tree Reasoning
von: Nam, Jaehyun, et al.
Veröffentlicht: (2024)
von: Nam, Jaehyun, et al.
Veröffentlicht: (2024)
PPMI: Privacy-Preserving LLM Interaction with Socratic Chain-of-Thought Reasoning and Homomorphically Encrypted Vector Databases
von: Bae, Yubeen, et al.
Veröffentlicht: (2025)
von: Bae, Yubeen, et al.
Veröffentlicht: (2025)
The Confidence Shortcut: A Reasoning Failure Mode of Masked Diffusion Models
von: Kim, Dueun, et al.
Veröffentlicht: (2026)
von: Kim, Dueun, et al.
Veröffentlicht: (2026)
Revisit What You See: Revealing Visual Semantics in Vision Tokens to Guide LVLM Decoding
von: Cho, Beomsik, et al.
Veröffentlicht: (2025)
von: Cho, Beomsik, et al.
Veröffentlicht: (2025)
Training-free LLM Verification via Recycling Few-shot Examples
von: Lee, Dongseok, et al.
Veröffentlicht: (2025)
von: Lee, Dongseok, et al.
Veröffentlicht: (2025)
Debiasing Online Preference Learning via Preference Feature Preservation
von: Kim, Dongyoung, et al.
Veröffentlicht: (2025)
von: Kim, Dongyoung, et al.
Veröffentlicht: (2025)
Learning from the Undesirable: Robust Adaptation of Language Models without Forgetting
von: Nam, Yunhun, et al.
Veröffentlicht: (2025)
von: Nam, Yunhun, et al.
Veröffentlicht: (2025)
NeuralSVCD for Efficient Swept Volume Collision Detection
von: Son, Dongwon, et al.
Veröffentlicht: (2025)
von: Son, Dongwon, et al.
Veröffentlicht: (2025)
VLM2Rec: Resolving Modality Collapse in Vision-Language Model Embedders for Multimodal Sequential Recommendation
von: Kim, Junyoung, et al.
Veröffentlicht: (2026)
von: Kim, Junyoung, et al.
Veröffentlicht: (2026)
INDIBATOR: Diverse and Fact-Grounded Individuality for Multi-Agent Debate in Molecular Discovery
von: Jang, Yunhui, et al.
Veröffentlicht: (2026)
von: Jang, Yunhui, et al.
Veröffentlicht: (2026)
Spread Preference Annotation: Direct Preference Judgment for Efficient LLM Alignment
von: Kim, Dongyoung, et al.
Veröffentlicht: (2024)
von: Kim, Dongyoung, et al.
Veröffentlicht: (2024)
A Neuro-Symbolic Approach for Probabilistic Reasoning on Graph Data
von: Pojer, Raffaele, et al.
Veröffentlicht: (2025)
von: Pojer, Raffaele, et al.
Veröffentlicht: (2025)
Confidence Optimization for Probabilistic Encoding
von: Xia, Pengjiu, et al.
Veröffentlicht: (2025)
von: Xia, Pengjiu, et al.
Veröffentlicht: (2025)
Model Already Knows the Best Noise: Bayesian Active Noise Selection via Attention in Video Diffusion Model
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2025)
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2025)
Fair Best Arm Identification with Fixed Confidence
von: Russo, Alessio, et al.
Veröffentlicht: (2024)
von: Russo, Alessio, et al.
Veröffentlicht: (2024)
An intuitive multi-frequency feature representation for SO(3)-equivariant networks
von: Son, Dongwon, et al.
Veröffentlicht: (2024)
von: Son, Dongwon, et al.
Veröffentlicht: (2024)
DEF-oriCORN: efficient 3D scene understanding for robust language-directed manipulation without demonstrations
von: Son, Dongwon, et al.
Veröffentlicht: (2024)
von: Son, Dongwon, et al.
Veröffentlicht: (2024)
Learning Generative Selection for Best-of-N
von: Toshniwal, Shubham, et al.
Veröffentlicht: (2026)
von: Toshniwal, Shubham, et al.
Veröffentlicht: (2026)
Personalized LLM Decoding via Contrasting Personal Preference
von: Bu, Hyungjune, et al.
Veröffentlicht: (2025)
von: Bu, Hyungjune, et al.
Veröffentlicht: (2025)
Asking Is Not Enough: Protocol Sensitivity in LLM Confidence Calibration
von: Kim, Hankyeol, et al.
Veröffentlicht: (2026)
von: Kim, Hankyeol, et al.
Veröffentlicht: (2026)
Efficient LLM Collaboration via Planning
von: Lee, Byeongchan, et al.
Veröffentlicht: (2025)
von: Lee, Byeongchan, et al.
Veröffentlicht: (2025)
No Reader Left Behind: Multi-Agent Summaries Everyone Can Understand
von: Jung, Jimin, et al.
Veröffentlicht: (2026)
von: Jung, Jimin, et al.
Veröffentlicht: (2026)
Verbalized Confidence Triggers Self-Verification: Emergent Behavior Without Explicit Reasoning Supervision
von: Jang, Chaeyun, et al.
Veröffentlicht: (2025)
von: Jang, Chaeyun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated Generation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2026) -
EMCEE: Improving Multilingual Capability of LLMs via Bridging Knowledge and Reasoning with Extracted Synthetic Multilingual Context
von: Koo, Hamin, et al.
Veröffentlicht: (2025) -
Learning to Correct for QA Reasoning with Black-box LLMs
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024) -
Revisiting the UID Hypothesis in LLM Reasoning Traces
von: Gwak, Minju, et al.
Veröffentlicht: (2025) -
Structural Reasoning Improves Molecular Understanding of LLM
von: Jang, Yunhui, et al.
Veröffentlicht: (2024)