Measuring Black-Box Confidence via Reasoning Trajectories: Geometry, Coverage, and Verbalization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Martell, Marc Boubnovski, Stoisser, Josefa Lia, Märtens, Kaspar, Yu, Jialin, Kitchen, Robert, Torr, Philip, Ferkinghoff-Borg, Jesper |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MechPert: Mechanistic Consensus as an Inductive Bias for Unseen Perturbation Prediction
von: Martell, Marc Boubnovski, et al.
Veröffentlicht: (2026)
von: Martell, Marc Boubnovski, et al.
Veröffentlicht: (2026)
Ambig-DS: A Benchmark for Task-Framing Ambiguity in Data-Science Agents
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2026)
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2026)
Towards Agents That Know When They Don't Know: Uncertainty as a Control Signal for Structured Reasoning
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
Sparks of Tabular Reasoning via Text2SQL Reinforcement Learning
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
Towards Label-Free Biological Reasoning Synthetic Dataset Creation via Uncertainty Filtering
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
SynthPert: Enhancing LLM Biological Reasoning via Synthetic Reasoning Traces for Cellular Perturbation Prediction
von: Phillips, Lawrence, et al.
Veröffentlicht: (2025)
von: Phillips, Lawrence, et al.
Veröffentlicht: (2025)
STRuCT-LLM: Unifying Tabular and Graph Reasoning with Reinforcement Learning for Semantic Parsing
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
Query, Don't Train: Privacy-Preserving Tabular Prediction from EHR Data via SQL Queries
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
Batched Energy-Entropy acquisition for Bayesian Optimization
von: Teufel, Felix, et al.
Veröffentlicht: (2024)
von: Teufel, Felix, et al.
Veröffentlicht: (2024)
Zero-shot protein stability prediction by inverse folding models: a free energy interpretation
von: Frellsen, Jes, et al.
Veröffentlicht: (2025)
von: Frellsen, Jes, et al.
Veröffentlicht: (2025)
Disentangling shared and private latent factors in multimodal Variational Autoencoders
von: Märtens, Kaspar, et al.
Veröffentlicht: (2024)
von: Märtens, Kaspar, et al.
Veröffentlicht: (2024)
On Verbalized Confidence Scores for LLMs
von: Yang, Daniel, et al.
Veröffentlicht: (2024)
von: Yang, Daniel, et al.
Veröffentlicht: (2024)
Are LLM Decisions Faithful to Verbal Confidence?
von: Wang, Jiawei, et al.
Veröffentlicht: (2026)
von: Wang, Jiawei, et al.
Veröffentlicht: (2026)
Direct Confidence Alignment: Aligning Verbalized Confidence with Internal Confidence In Large Language Models
von: Zhang, Glenn, et al.
Veröffentlicht: (2025)
von: Zhang, Glenn, et al.
Veröffentlicht: (2025)
Verbalized Confidence Triggers Self-Verification: Emergent Behavior Without Explicit Reasoning Supervision
von: Jang, Chaeyun, et al.
Veröffentlicht: (2025)
von: Jang, Chaeyun, et al.
Veröffentlicht: (2025)
ADVICE: Answer-Dependent Verbalized Confidence Estimation
von: Seo, Ki Jung, et al.
Veröffentlicht: (2025)
von: Seo, Ki Jung, et al.
Veröffentlicht: (2025)
How do LLMs Compute Verbal Confidence
von: Kumaran, Dharshan, et al.
Veröffentlicht: (2026)
von: Kumaran, Dharshan, et al.
Veröffentlicht: (2026)
On the Robustness of Verbal Confidence of LLMs in Adversarial Attacks
von: Obadinma, Stephen, et al.
Veröffentlicht: (2025)
von: Obadinma, Stephen, et al.
Veröffentlicht: (2025)
Calibrating Verbalized Confidence with Self-Generated Distractors
von: Wang, Victor, et al.
Veröffentlicht: (2025)
von: Wang, Victor, et al.
Veröffentlicht: (2025)
From Black-Box Confidence to Measurable Trust in Clinical AI: A Framework for Evidence, Supervision, and Staged Autonomy
von: Zabolotnii, Serhii, et al.
Veröffentlicht: (2026)
von: Zabolotnii, Serhii, et al.
Veröffentlicht: (2026)
LiBOG: Lifelong Learning for Black-Box Optimizer Generation
von: Pei, Jiyuan, et al.
Veröffentlicht: (2025)
von: Pei, Jiyuan, et al.
Veröffentlicht: (2025)
A Leakage Bound for Confidence Sets after Black-Box Selection
von: Banerjee, Sayantan
Veröffentlicht: (2026)
von: Banerjee, Sayantan
Veröffentlicht: (2026)
PCS: Perceived Confidence Scoring of Black Box LLMs with Metamorphic Relations
von: Salimian, Sina, et al.
Veröffentlicht: (2025)
von: Salimian, Sina, et al.
Veröffentlicht: (2025)
Large Language Model Confidence Estimation via Black-Box Access
von: Pedapati, Tejaswini, et al.
Veröffentlicht: (2024)
von: Pedapati, Tejaswini, et al.
Veröffentlicht: (2024)
Time of day of cardiac surgery and postoperative outcomes: a reply
von: Gareth Kitchen
Veröffentlicht: (2026)
von: Gareth Kitchen
Veröffentlicht: (2026)
Breaking the Black-Box: Confidence-Guided Model Inversion Attack for Distribution Shift
von: Liu, Xinhao, et al.
Veröffentlicht: (2024)
von: Liu, Xinhao, et al.
Veröffentlicht: (2024)
Realizing an Atomtronic AQUID in a Rotating-Box Potential
von: Görg, Kaspar, et al.
Veröffentlicht: (2025)
von: Görg, Kaspar, et al.
Veröffentlicht: (2025)
Modeling Gene Expression Distributional Shifts for Unseen Genetic Perturbations
von: Ramakrishnan, Kalyan, et al.
Veröffentlicht: (2025)
von: Ramakrishnan, Kalyan, et al.
Veröffentlicht: (2025)
"Moralized" Multi-Step Jailbreak Prompts: Black-Box Testing of Guardrails in Large Language Models for Verbal Attacks
von: Wang, Libo
Veröffentlicht: (2024)
von: Wang, Libo
Veröffentlicht: (2024)
Are Large Language Models More Honest in Their Probabilistic or Verbalized Confidence?
von: Ni, Shiyu, et al.
Veröffentlicht: (2024)
von: Ni, Shiyu, et al.
Veröffentlicht: (2024)
Influential Training Data Retrieval for Explaining Verbalized Confidence of LLMs
von: Xia, Yuxi, et al.
Veröffentlicht: (2026)
von: Xia, Yuxi, et al.
Veröffentlicht: (2026)
Verbal and Numeric Eyewitness Confidence Differentially Affect Decision‐Making
von: Pia Pennekamp
Veröffentlicht: (2025)
von: Pia Pennekamp
Veröffentlicht: (2025)
Certifiable Black-Box Attacks with Randomized Adversarial Examples: Breaking Defenses with Provable Confidence
von: Hong, Hanbin, et al.
Veröffentlicht: (2023)
von: Hong, Hanbin, et al.
Veröffentlicht: (2023)
SLIM: Stealthy Low-Coverage Black-Box Watermarking via Latent-Space Confusion Zones
von: Wu, Hengyu, et al.
Veröffentlicht: (2026)
von: Wu, Hengyu, et al.
Veröffentlicht: (2026)
PUH Theorem 214 — Black Hole Information Preservation in Planck Cores via Polariton Bound States, and the Triple Isolation of the Matter and Antimatter Universes Across the Rebound Throat: A Resolution of the Hawking Information Paradox Within the Photonic Universe Hypothesis
von: Martell, Brian
Veröffentlicht: (2026)
von: Martell, Brian
Veröffentlicht: (2026)
Wired for Overconfidence: A Mechanistic Perspective on Inflated Verbalized Confidence in LLMs
von: Zhao, Tianyi, et al.
Veröffentlicht: (2026)
von: Zhao, Tianyi, et al.
Veröffentlicht: (2026)
ConfTuner: Training Large Language Models to Express Their Confidence Verbally
von: Li, Yibo, et al.
Veröffentlicht: (2025)
von: Li, Yibo, et al.
Veröffentlicht: (2025)
ORCE: Order-Aware Alignment of Verbalized Confidence in Large Language Models
von: Li, Chen, et al.
Veröffentlicht: (2026)
von: Li, Chen, et al.
Veröffentlicht: (2026)
Conformal Sets in Multiple-Choice Question Answering under Black-Box Settings with Provable Coverage Guarantees
von: Yang, Guang, et al.
Veröffentlicht: (2025)
von: Yang, Guang, et al.
Veröffentlicht: (2025)
Frozen/Thawed Samples Can Replace Fresh Samples for Assignment of ISI to Secondary Thromboplastin Standards for Multiple Reagent/Instrument Combinations: Data to Support Possible Revision of WHO Guidelines
von: Matthew Kitchen, et al.
Veröffentlicht: (2024)
von: Matthew Kitchen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MechPert: Mechanistic Consensus as an Inductive Bias for Unseen Perturbation Prediction
von: Martell, Marc Boubnovski, et al.
Veröffentlicht: (2026) -
Ambig-DS: A Benchmark for Task-Framing Ambiguity in Data-Science Agents
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2026) -
Towards Agents That Know When They Don't Know: Uncertainty as a Control Signal for Structured Reasoning
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025) -
Sparks of Tabular Reasoning via Text2SQL Reinforcement Learning
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025) -
Towards Label-Free Biological Reasoning Synthetic Dataset Creation via Uncertainty Filtering
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)