The Effects of Data Augmentation on Confidence Estimation for LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Rui, Zhu, Renyu, Lin, Minmin, Wu, Runze, Lv, Tangjie, Fan, Changjie, Wang, Haobo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Dataset for the Validation of Truth Inference Algorithms Suitable for Online Deployment
by: Wang, Fei, et al.
Published: (2024)
by: Wang, Fei, et al.
Published: (2024)
Digital Player: Evaluating Large Language Models based Human-like Agent in Games
by: Wang, Jiawei, et al.
Published: (2025)
by: Wang, Jiawei, et al.
Published: (2025)
Prioritized Trajectory Replay: A Replay Memory for Data-driven Reinforcement Learning
by: Liu, Jinyi, et al.
Published: (2023)
by: Liu, Jinyi, et al.
Published: (2023)
Reinforcement Learning From Imperfect Corrective Actions And Proxy Rewards
by: Jiang, Zhaohui, et al.
Published: (2024)
by: Jiang, Zhaohui, et al.
Published: (2024)
Rank Aggregation in Crowdsourcing for Listwise Annotations
by: Luo, Wenshui, et al.
Published: (2024)
by: Luo, Wenshui, et al.
Published: (2024)
vMFER: Von Mises-Fisher Experience Resampling Based on Uncertainty of Gradient Directions for Policy Improvement
by: Zhu, Yiwen, et al.
Published: (2024)
by: Zhu, Yiwen, et al.
Published: (2024)
Bayesian Design Principles for Offline-to-Online Reinforcement Learning
by: Hu, Hao, et al.
Published: (2024)
by: Hu, Hao, et al.
Published: (2024)
Fast-DataShapley: Neural Modeling for Training Data Valuation
by: Sun, Haifeng, et al.
Published: (2025)
by: Sun, Haifeng, et al.
Published: (2025)
CrowdAgent: Multi-Agent Managed Multi-Source Annotation System
by: Qin, Maosheng, et al.
Published: (2025)
by: Qin, Maosheng, et al.
Published: (2025)
Prompt Candidates, then Distill: A Teacher-Student Framework for LLM-driven Data Annotation
by: Xia, Mingxuan, et al.
Published: (2025)
by: Xia, Mingxuan, et al.
Published: (2025)
Filtering with Confidence: When Data Augmentation Meets Conformal Prediction
by: Wu, Zixuan, et al.
Published: (2025)
by: Wu, Zixuan, et al.
Published: (2025)
Data Contamination Calibration for Black-box LLMs
by: Ye, Wentao, et al.
Published: (2024)
by: Ye, Wentao, et al.
Published: (2024)
Factual Confidence of LLMs: on Reliability and Robustness of Current Estimators
by: Mahaut, Matéo, et al.
Published: (2024)
by: Mahaut, Matéo, et al.
Published: (2024)
On the Saturation Effect of Kernel Ridge Regression
by: Li, Yicheng, et al.
Published: (2024)
by: Li, Yicheng, et al.
Published: (2024)
A New Baseline Assumption of Integated Gradients Based on Shaply value
by: Liu, Shuyang, et al.
Published: (2023)
by: Liu, Shuyang, et al.
Published: (2023)
Enhancing Confidence Estimation in Telco LLMs via Twin-Pass CoT-Ensembling
by: Saenko, Anton, et al.
Published: (2026)
by: Saenko, Anton, et al.
Published: (2026)
Dynamic Causal Structure Discovery and Causal Effect Estimation
by: Wang, Jianian, et al.
Published: (2025)
by: Wang, Jianian, et al.
Published: (2025)
ConfClip: Confidence-Weighted and Clipped Reward for Reinforcement Learning in LLMs
by: Zhang, Bonan, et al.
Published: (2025)
by: Zhang, Bonan, et al.
Published: (2025)
Enhance GNNs with Reliable Confidence Estimation via Adversarial Calibration Learning
by: Wang, Yilong, et al.
Published: (2025)
by: Wang, Yilong, et al.
Published: (2025)
Predicting Effects, Missing Distributions: Evaluating LLMs as Human Behavior Simulators in Operations Management
by: Zhang, Runze, et al.
Published: (2025)
by: Zhang, Runze, et al.
Published: (2025)
Unsupervised Confidence Calibration for Reasoning LLMs from a Single Generation
by: Zollo, Thomas, et al.
Published: (2026)
by: Zollo, Thomas, et al.
Published: (2026)
An Analysis of Causal Effect Estimation using Outcome Invariant Data Augmentation
by: Akbar, Uzair, et al.
Published: (2025)
by: Akbar, Uzair, et al.
Published: (2025)
On the Saturation Effects of Spectral Algorithms in Large Dimensions
by: Lu, Weihao, et al.
Published: (2025)
by: Lu, Weihao, et al.
Published: (2025)
Revisiting Confidence Estimation: Towards Reliable Failure Prediction
by: Zhu, Fei, et al.
Published: (2024)
by: Zhu, Fei, et al.
Published: (2024)
Disentangle Estimation of Causal Effects from Cross-Silo Data
by: Liu, Yuxuan, et al.
Published: (2024)
by: Liu, Yuxuan, et al.
Published: (2024)
CER: Confidence Enhanced Reasoning in LLMs
by: Razghandi, Ali, et al.
Published: (2025)
by: Razghandi, Ali, et al.
Published: (2025)
Feature Augmentations for High-Dimensional Learning
by: Zhu, Xiaonan, et al.
Published: (2025)
by: Zhu, Xiaonan, et al.
Published: (2025)
Personalized Programming Guidance based on Deep Programming Learning Style Capturing
by: Liu, Yingfan, et al.
Published: (2024)
by: Liu, Yingfan, et al.
Published: (2024)
Optimal Confidence Band for Kernel Gradient Flow Estimator
by: Cheng, Yuqian, et al.
Published: (2026)
by: Cheng, Yuqian, et al.
Published: (2026)
DataInf: Efficiently Estimating Data Influence in LoRA-tuned LLMs and Diffusion Models
by: Kwon, Yongchan, et al.
Published: (2023)
by: Kwon, Yongchan, et al.
Published: (2023)
Multicalibration for Confidence Scoring in LLMs
by: Detommaso, Gianluca, et al.
Published: (2024)
by: Detommaso, Gianluca, et al.
Published: (2024)
Unleashing the Potential of Acquisition Functions in High-Dimensional Bayesian Optimization
by: Zhao, Jiayu, et al.
Published: (2023)
by: Zhao, Jiayu, et al.
Published: (2023)
Instance-Dependent Continuous-Time Reinforcement Learning via Maximum Likelihood Estimation
by: Zhao, Runze, et al.
Published: (2025)
by: Zhao, Runze, et al.
Published: (2025)
SaySelf: Teaching LLMs to Express Confidence with Self-Reflective Rationales
by: Xu, Tianyang, et al.
Published: (2024)
by: Xu, Tianyang, et al.
Published: (2024)
Towards Cross-Table Masked Pretraining for Web Data Mining
by: Ye, Chao, et al.
Published: (2023)
by: Ye, Chao, et al.
Published: (2023)
Reward-Augmented Data Enhances Direct Preference Alignment of LLMs
by: Zhang, Shenao, et al.
Published: (2024)
by: Zhang, Shenao, et al.
Published: (2024)
Confidence Estimation via Sequential Likelihood Mixing
by: Kirschner, Johannes, et al.
Published: (2025)
by: Kirschner, Johannes, et al.
Published: (2025)
Uncertainty Estimation via Hyperspherical Confidence Mapping
by: Choi, Eunseo, et al.
Published: (2026)
by: Choi, Eunseo, et al.
Published: (2026)
Principled Confidence Estimation for Deep Computed Tomography
by: Gätzner, Matteo, et al.
Published: (2026)
by: Gätzner, Matteo, et al.
Published: (2026)
Stronger-MAS: Multi-Agent Reinforcement Learning for Collaborative LLMs
by: Zhao, Yujie, et al.
Published: (2025)
by: Zhao, Yujie, et al.
Published: (2025)
Similar Items
-
A Dataset for the Validation of Truth Inference Algorithms Suitable for Online Deployment
by: Wang, Fei, et al.
Published: (2024) -
Digital Player: Evaluating Large Language Models based Human-like Agent in Games
by: Wang, Jiawei, et al.
Published: (2025) -
Prioritized Trajectory Replay: A Replay Memory for Data-driven Reinforcement Learning
by: Liu, Jinyi, et al.
Published: (2023) -
Reinforcement Learning From Imperfect Corrective Actions And Proxy Rewards
by: Jiang, Zhaohui, et al.
Published: (2024) -
Rank Aggregation in Crowdsourcing for Listwise Annotations
by: Luo, Wenshui, et al.
Published: (2024)