Low-Confidence Gold: Refining Low-Confidence Samples for Efficient Instruction Tuning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cai, Hongyi, Li, Jie, Rahman, Mohammad Mahdinur, Dong, Wenzhen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MCQA-Eval: Efficient Confidence Evaluation in NLG with Gold-Standard Correctness Labels
von: Liu, Xiaoou, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoou, et al.
Veröffentlicht: (2025)
CoRefine: Confidence-Guided Self-Refinement for Adaptive Test-Time Compute
von: Jin, Chen, et al.
Veröffentlicht: (2026)
von: Jin, Chen, et al.
Veröffentlicht: (2026)
Confidence-guided Refinement Reasoning for Zero-shot Question Answering
von: Jang, Youwon, et al.
Veröffentlicht: (2025)
von: Jang, Youwon, et al.
Veröffentlicht: (2025)
CFPFormer: Feature-pyramid like Transformer Decoder for Segmentation and Detection
von: Cai, Hongyi, et al.
Veröffentlicht: (2024)
von: Cai, Hongyi, et al.
Veröffentlicht: (2024)
Confidence is Not Competence
von: Sanyal, Debdeep, et al.
Veröffentlicht: (2025)
von: Sanyal, Debdeep, et al.
Veröffentlicht: (2025)
ConCISE: Confidence-guided Compression in Step-by-step Efficient Reasoning
von: Qiao, Ziqing, et al.
Veröffentlicht: (2025)
von: Qiao, Ziqing, et al.
Veröffentlicht: (2025)
Agentic Confidence Calibration
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
Verbal Confidence Saturation in 3-9B Open-Weight Instruction-Tuned LLMs: A Pre-Registered Psychometric Validity Screen
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence
von: Ghasemabadi, Amirhosein, et al.
Veröffentlicht: (2025)
von: Ghasemabadi, Amirhosein, et al.
Veröffentlicht: (2025)
ConMax: Confidence-Maximizing Compression for Efficient Chain-of-Thought Reasoning
von: Hu, Minda, et al.
Veröffentlicht: (2026)
von: Hu, Minda, et al.
Veröffentlicht: (2026)
Label-Confidence-Aware Uncertainty Estimation in Natural Language Generation
von: Lin, Qinhong, et al.
Veröffentlicht: (2024)
von: Lin, Qinhong, et al.
Veröffentlicht: (2024)
Combining Confidence Elicitation and Sample-based Methods for Uncertainty Quantification in Misinformation Mitigation
von: Rivera, Mauricio, et al.
Veröffentlicht: (2024)
von: Rivera, Mauricio, et al.
Veröffentlicht: (2024)
A Survey of Confidence Estimation and Calibration in Large Language Models
von: Geng, Jiahui, et al.
Veröffentlicht: (2023)
von: Geng, Jiahui, et al.
Veröffentlicht: (2023)
Confident RAG: Enhancing the Performance of LLMs for Mathematics Question Answering through Multi-Embedding and Confidence Scoring
von: Chen, Shiting, et al.
Veröffentlicht: (2025)
von: Chen, Shiting, et al.
Veröffentlicht: (2025)
Fact-Level Confidence Calibration and Self-Correction
von: Yuan, Yige, et al.
Veröffentlicht: (2024)
von: Yuan, Yige, et al.
Veröffentlicht: (2024)
Entropy-Adaptive Fine-Tuning: Resolving Confident Conflicts to Mitigate Forgetting
von: Diao, Muxi, et al.
Veröffentlicht: (2026)
von: Diao, Muxi, et al.
Veröffentlicht: (2026)
Reasoning Models Better Express Their Confidence
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
Confidence Improves Self-Consistency in LLMs
von: Taubenfeld, Amir, et al.
Veröffentlicht: (2025)
von: Taubenfeld, Amir, et al.
Veröffentlicht: (2025)
Confidence-Calibrated Small-Large Language Model Collaboration for Cost-Efficient Reasoning
von: Zhang, Chuang, et al.
Veröffentlicht: (2026)
von: Zhang, Chuang, et al.
Veröffentlicht: (2026)
Multi-Perspective Consistency Enhances Confidence Estimation in Large Language Models
von: Wang, Pei, et al.
Veröffentlicht: (2024)
von: Wang, Pei, et al.
Veröffentlicht: (2024)
Confidence Under the Hood: An Investigation into the Confidence-Probability Alignment in Large Language Models
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
Tuning LLMs with Contrastive Alignment Instructions for Machine Translation in Unseen, Low-resource Languages
von: Mao, Zhuoyuan, et al.
Veröffentlicht: (2024)
von: Mao, Zhuoyuan, et al.
Veröffentlicht: (2024)
Mind the Generation Process: Fine-Grained Confidence Estimation During LLM Generation
von: Han, Jinyi, et al.
Veröffentlicht: (2025)
von: Han, Jinyi, et al.
Veröffentlicht: (2025)
SUGAR: Leveraging Contextual Confidence for Smarter Retrieval
von: Zubkova, Hanna, et al.
Veröffentlicht: (2025)
von: Zubkova, Hanna, et al.
Veröffentlicht: (2025)
Calibrating Verbalized Confidence with Self-Generated Distractors
von: Wang, Victor, et al.
Veröffentlicht: (2025)
von: Wang, Victor, et al.
Veröffentlicht: (2025)
From Confidence to Collapse in LLM Factual Robustness
von: Fastowski, Alina, et al.
Veröffentlicht: (2025)
von: Fastowski, Alina, et al.
Veröffentlicht: (2025)
CAMEL: Confidence-Gated Reflection for Reward Modeling
von: Zhu, Zirui, et al.
Veröffentlicht: (2026)
von: Zhu, Zirui, et al.
Veröffentlicht: (2026)
Beyond Confidence: The Rhythms of Reasoning in Generative Models
von: Liu, Deyuan, et al.
Veröffentlicht: (2026)
von: Liu, Deyuan, et al.
Veröffentlicht: (2026)
ConfSpec: Efficient Step-Level Speculative Reasoning via Confidence-Gated Verification
von: Liu, Siran, et al.
Veröffentlicht: (2026)
von: Liu, Siran, et al.
Veröffentlicht: (2026)
Low-Resource Fine-Tuning for Multi-Task Structured Information Extraction with a Billion-Parameter Instruction-Tuned Model
von: Chih, Yu Cheng, et al.
Veröffentlicht: (2025)
von: Chih, Yu Cheng, et al.
Veröffentlicht: (2025)
LinguaLIFT: An Effective Two-stage Instruction Tuning Framework for Low-Resource Language Reasoning
von: Zhang, Hongbin, et al.
Veröffentlicht: (2024)
von: Zhang, Hongbin, et al.
Veröffentlicht: (2024)
Konkani LLM: Multi-Script Instruction Tuning and Evaluation for a Low-Resource Indian Language
von: Fernandes, Reuben Chagas, et al.
Veröffentlicht: (2026)
von: Fernandes, Reuben Chagas, et al.
Veröffentlicht: (2026)
System-2 Mathematical Reasoning via Enriched Instruction Tuning
von: Cai, Huanqia, et al.
Veröffentlicht: (2024)
von: Cai, Huanqia, et al.
Veröffentlicht: (2024)
PiCSAR: Probabilistic Confidence Selection And Ranking for Reasoning Chains
von: Leang, Joshua Ong Jun, et al.
Veröffentlicht: (2025)
von: Leang, Joshua Ong Jun, et al.
Veröffentlicht: (2025)
ConfTuner: Training Large Language Models to Express Their Confidence Verbally
von: Li, Yibo, et al.
Veröffentlicht: (2025)
von: Li, Yibo, et al.
Veröffentlicht: (2025)
PACR: Progressively Ascending Confidence Reward for LLM Reasoning
von: Yoon, Eunseop, et al.
Veröffentlicht: (2025)
von: Yoon, Eunseop, et al.
Veröffentlicht: (2025)
Knowledge Graph Error Detection with Contrastive Confidence Adaption
von: Liu, Xiangyu, et al.
Veröffentlicht: (2023)
von: Liu, Xiangyu, et al.
Veröffentlicht: (2023)
Closing the Confidence-Faithfulness Gap in Large Language Models
von: Miao, Miranda Muqing, et al.
Veröffentlicht: (2026)
von: Miao, Miranda Muqing, et al.
Veröffentlicht: (2026)
When Quantization Affects Confidence of Large Language Models?
von: Proskurina, Irina, et al.
Veröffentlicht: (2024)
von: Proskurina, Irina, et al.
Veröffentlicht: (2024)
Confidence Estimation for LLM-Based Dialogue State Tracking
von: Sun, Yi-Jyun, et al.
Veröffentlicht: (2024)
von: Sun, Yi-Jyun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MCQA-Eval: Efficient Confidence Evaluation in NLG with Gold-Standard Correctness Labels
von: Liu, Xiaoou, et al.
Veröffentlicht: (2025) -
CoRefine: Confidence-Guided Self-Refinement for Adaptive Test-Time Compute
von: Jin, Chen, et al.
Veröffentlicht: (2026) -
Confidence-guided Refinement Reasoning for Zero-shot Question Answering
von: Jang, Youwon, et al.
Veröffentlicht: (2025) -
CFPFormer: Feature-pyramid like Transformer Decoder for Segmentation and Detection
von: Cai, Hongyi, et al.
Veröffentlicht: (2024) -
Confidence is Not Competence
von: Sanyal, Debdeep, et al.
Veröffentlicht: (2025)