GrACE: A Generative Approach to Better Confidence Elicitation and Efficient Test-Time Scaling in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Zhaohan, Liu, Ziquan, Patras, Ioannis |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Get Confused Cautiously: Textual Sequence Memorization Erasure with Selective Entropy Maximization
von: Zhang, Zhaohan, et al.
Veröffentlicht: (2024)
von: Zhang, Zhaohan, et al.
Veröffentlicht: (2024)
Confidence Should Be Calibrated More Than One Turn Deep
von: Zhang, Zhaohan, et al.
Veröffentlicht: (2026)
von: Zhang, Zhaohan, et al.
Veröffentlicht: (2026)
Calibrating the Confidence of Large Language Models by Eliciting Fidelity
von: Zhang, Mozhi, et al.
Veröffentlicht: (2024)
von: Zhang, Mozhi, et al.
Veröffentlicht: (2024)
Breaking Language Barriers or Reinforcing Bias? A Study of Gender and Racial Disparities in Multilingual Contrastive Vision Language Models
von: Sahili, Zahraa Al, et al.
Veröffentlicht: (2025)
von: Sahili, Zahraa Al, et al.
Veröffentlicht: (2025)
Confidence Elicitation: A New Attack Vector for Large Language Models
von: Formento, Brian, et al.
Veröffentlicht: (2025)
von: Formento, Brian, et al.
Veröffentlicht: (2025)
First Finish Search: Efficient Test-Time Scaling in Large Language Models
von: Agarwal, Aradhye, et al.
Veröffentlicht: (2025)
von: Agarwal, Aradhye, et al.
Veröffentlicht: (2025)
Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence
von: Ghasemabadi, Amirhosein, et al.
Veröffentlicht: (2025)
von: Ghasemabadi, Amirhosein, et al.
Veröffentlicht: (2025)
Confidence in Large Language Model Evaluation: A Bayesian Approach to Limited-Sample Challenges
von: Xiao, Xiao, et al.
Veröffentlicht: (2025)
von: Xiao, Xiao, et al.
Veröffentlicht: (2025)
Knowledge Boundary Discovery for Large Language Models
von: Wang, Ziquan, et al.
Veröffentlicht: (2026)
von: Wang, Ziquan, et al.
Veröffentlicht: (2026)
Faster and Better LLMs via Latency-Aware Test-Time Scaling
von: Wang, Zili, et al.
Veröffentlicht: (2025)
von: Wang, Zili, et al.
Veröffentlicht: (2025)
The Art of Scaling Test-Time Compute for Large Language Models
von: Agarwal, Aradhye, et al.
Veröffentlicht: (2025)
von: Agarwal, Aradhye, et al.
Veröffentlicht: (2025)
Efficiently Seeking Flat Minima for Better Generalization in Fine-Tuning Large Language Models and Beyond
von: Deng, Jiaxin, et al.
Veröffentlicht: (2025)
von: Deng, Jiaxin, et al.
Veröffentlicht: (2025)
Direct Confidence Alignment: Aligning Verbalized Confidence with Internal Confidence In Large Language Models
von: Zhang, Glenn, et al.
Veröffentlicht: (2025)
von: Zhang, Glenn, et al.
Veröffentlicht: (2025)
Confidence-Calibrated Small-Large Language Model Collaboration for Cost-Efficient Reasoning
von: Zhang, Chuang, et al.
Veröffentlicht: (2026)
von: Zhang, Chuang, et al.
Veröffentlicht: (2026)
AMR-Evol: Adaptive Modular Response Evolution Elicits Better Knowledge Distillation for Large Language Models in Code Generation
von: Luo, Ziyang, et al.
Veröffentlicht: (2024)
von: Luo, Ziyang, et al.
Veröffentlicht: (2024)
BrowseConf: Confidence-Guided Test-Time Scaling for Web Agents
von: Ou, Litu, et al.
Veröffentlicht: (2025)
von: Ou, Litu, et al.
Veröffentlicht: (2025)
Chain-of-Dictionary Prompting Elicits Translation in Large Language Models
von: Lu, Hongyuan, et al.
Veröffentlicht: (2023)
von: Lu, Hongyuan, et al.
Veröffentlicht: (2023)
BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence
von: Wu, Sean, et al.
Veröffentlicht: (2026)
von: Wu, Sean, et al.
Veröffentlicht: (2026)
Step-level Verifier-guided Hybrid Test-Time Scaling for Large Language Models
von: Chang, Kaiyan, et al.
Veröffentlicht: (2025)
von: Chang, Kaiyan, et al.
Veröffentlicht: (2025)
Enhancing Test-Time Scaling of Large Language Models with Hierarchical Retrieval-Augmented MCTS
von: Dou, Alex ZH, et al.
Veröffentlicht: (2025)
von: Dou, Alex ZH, et al.
Veröffentlicht: (2025)
Confidence over Time: Confidence Calibration with Temporal Logic for Large Language Model Reasoning
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
Cognition-of-Thought Elicits Social-Aligned Reasoning in Large Language Models
von: Zhang, Xuanming, et al.
Veröffentlicht: (2025)
von: Zhang, Xuanming, et al.
Veröffentlicht: (2025)
EAGLE-3: Scaling up Inference Acceleration of Large Language Models via Training-Time Test
von: Li, Yuhui, et al.
Veröffentlicht: (2025)
von: Li, Yuhui, et al.
Veröffentlicht: (2025)
AutoElicit: Using Large Language Models for Expert Prior Elicitation in Predictive Modelling
von: Capstick, Alexander, et al.
Veröffentlicht: (2024)
von: Capstick, Alexander, et al.
Veröffentlicht: (2024)
Test-Time Scaling with Reflective Generative Model
von: Wang, Zixiao, et al.
Veröffentlicht: (2025)
von: Wang, Zixiao, et al.
Veröffentlicht: (2025)
Reasoning Models Better Express Their Confidence
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
Eliciting Personality Traits in Large Language Models
von: Hilliard, Airlie, et al.
Veröffentlicht: (2024)
von: Hilliard, Airlie, et al.
Veröffentlicht: (2024)
SteerConf: Steering LLMs for Confidence Elicitation
von: Zhou, Ziang, et al.
Veröffentlicht: (2025)
von: Zhou, Ziang, et al.
Veröffentlicht: (2025)
MiGrATe: Mixed-Policy GRPO for Adaptation at Test-Time
von: Phan, Peter, et al.
Veröffentlicht: (2025)
von: Phan, Peter, et al.
Veröffentlicht: (2025)
A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2025)
Eliciting Informative Text Evaluations with Large Language Models
von: Lu, Yuxuan, et al.
Veröffentlicht: (2024)
von: Lu, Yuxuan, et al.
Veröffentlicht: (2024)
Zero-to-Strong Generalization: Eliciting Strong Capabilities of Large Language Models Iteratively without Gold Labels
von: Liu, Chaoqun, et al.
Veröffentlicht: (2024)
von: Liu, Chaoqun, et al.
Veröffentlicht: (2024)
Self-Demos: Eliciting Out-of-Demonstration Generalizability in Large Language Models
von: He, Wei, et al.
Veröffentlicht: (2024)
von: He, Wei, et al.
Veröffentlicht: (2024)
ReProbe: Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models
von: Ni, Jingwei, et al.
Veröffentlicht: (2025)
von: Ni, Jingwei, et al.
Veröffentlicht: (2025)
Advancing Block Diffusion Language Models for Test-Time Scaling
von: Lu, Yi, et al.
Veröffentlicht: (2026)
von: Lu, Yi, et al.
Veröffentlicht: (2026)
Executable Code Actions Elicit Better LLM Agents
von: Wang, Xingyao, et al.
Veröffentlicht: (2024)
von: Wang, Xingyao, et al.
Veröffentlicht: (2024)
Uncertainty Quantification and Confidence Calibration in Large Language Models: A Survey
von: Liu, Xiaoou, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoou, et al.
Veröffentlicht: (2025)
Hypothesis-only Biases in Large Language Model-Elicited Natural Language Inference
von: Proebsting, Grace, et al.
Veröffentlicht: (2024)
von: Proebsting, Grace, et al.
Veröffentlicht: (2024)
m1: Unleash the Potential of Test-Time Scaling for Medical Reasoning with Large Language Models
von: Huang, Xiaoke, et al.
Veröffentlicht: (2025)
von: Huang, Xiaoke, et al.
Veröffentlicht: (2025)
Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions
von: Masoud, Reem I., et al.
Veröffentlicht: (2023)
von: Masoud, Reem I., et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Get Confused Cautiously: Textual Sequence Memorization Erasure with Selective Entropy Maximization
von: Zhang, Zhaohan, et al.
Veröffentlicht: (2024) -
Confidence Should Be Calibrated More Than One Turn Deep
von: Zhang, Zhaohan, et al.
Veröffentlicht: (2026) -
Calibrating the Confidence of Large Language Models by Eliciting Fidelity
von: Zhang, Mozhi, et al.
Veröffentlicht: (2024) -
Breaking Language Barriers or Reinforcing Bias? A Study of Gender and Racial Disparities in Multilingual Contrastive Vision Language Models
von: Sahili, Zahraa Al, et al.
Veröffentlicht: (2025) -
Confidence Elicitation: A New Attack Vector for Large Language Models
von: Formento, Brian, et al.
Veröffentlicht: (2025)