Efficient Epistemic Uncertainty Estimation for Large Language Models via Knowledge Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Park, Seonghyeon, Yeom, Jewon, Sok, Jaewon, Park, Jeongjae, Kim, Heejun, Kim, Taesup |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Garbage Attention in Large Language Models: BOS Sink Heads and Sink-aware Pruning
by: Sok, Jaewon, et al.
Published: (2026)
by: Sok, Jaewon, et al.
Published: (2026)
From Noise to Diversity: Random Embedding Injection in LLM Reasoning
by: Kim, Heejun, et al.
Published: (2026)
by: Kim, Heejun, et al.
Published: (2026)
Hallucination as Commitment Failure: Larger LLMs Misfire Despite Knowing the Answer
by: Yeom, Jewon, et al.
Published: (2026)
by: Yeom, Jewon, et al.
Published: (2026)
EpiCaR: Knowing What You Don't Know Matters for Better Reasoning in LLMs
by: Yeom, Jewon, et al.
Published: (2026)
by: Yeom, Jewon, et al.
Published: (2026)
Stable On-Policy Distillation through Adaptive Target Reformulation
by: Jang, Ijun, et al.
Published: (2026)
by: Jang, Ijun, et al.
Published: (2026)
Beyond Case Law: Evaluating Structure-Aware Retrieval and Safety in Statute-Centric Legal QA
by: Chae, Kyubyung, et al.
Published: (2026)
by: Chae, Kyubyung, et al.
Published: (2026)
PRISP: Privacy-Safe Few-Shot Personalization via Lightweight Adaptation
by: Park, Junho, et al.
Published: (2026)
by: Park, Junho, et al.
Published: (2026)
Towards Robust Real-World Multivariate Time Series Forecasting: A Unified Framework for Dependency, Asynchrony, and Missingness
by: Jang, Jinkwan, et al.
Published: (2025)
by: Jang, Jinkwan, et al.
Published: (2025)
Overcoming Data Inequality across Domains with Semi-Supervised Domain Generalization
by: Park, Jinha, et al.
Published: (2024)
by: Park, Jinha, et al.
Published: (2024)
Random Conditioning with Distillation for Data-Efficient Diffusion Model Compression
by: Kim, Dohyun, et al.
Published: (2025)
by: Kim, Dohyun, et al.
Published: (2025)
Energy-Efficient Wireless LLM Inference via Uncertainty and Importance-Aware Speculative Decoding
by: Park, Jihoon, et al.
Published: (2025)
by: Park, Jihoon, et al.
Published: (2025)
X-PEFT: eXtremely Parameter-Efficient Fine-Tuning for Extreme Multi-Profile Scenarios
by: Kwak, Namju, et al.
Published: (2024)
by: Kwak, Namju, et al.
Published: (2024)
PiCa: Parameter-Efficient Fine-Tuning with Column Space Projection
by: Hwang, Junseo, et al.
Published: (2025)
by: Hwang, Junseo, et al.
Published: (2025)
Efficient Epistemic Uncertainty Estimation in Regression Ensemble Models Using Pairwise-Distance Estimators
by: Berry, Lucas, et al.
Published: (2023)
by: Berry, Lucas, et al.
Published: (2023)
CASK: Core-Aware Selective KV Compression for Reasoning Traces
by: Kim, Buseong, et al.
Published: (2026)
by: Kim, Buseong, et al.
Published: (2026)
DoMIX: An Efficient Framework for Exploiting Domain Knowledge in Fine-Tuning
by: Kim, Dohoon, et al.
Published: (2025)
by: Kim, Dohoon, et al.
Published: (2025)
Latent Bayesian Optimization via Autoregressive Normalizing Flows
by: Lee, Seunghun, et al.
Published: (2025)
by: Lee, Seunghun, et al.
Published: (2025)
Epistemic Diversity and Knowledge Collapse in Large Language Models
by: Wright, Dustin, et al.
Published: (2025)
by: Wright, Dustin, et al.
Published: (2025)
Distillation of Large Language Models via Concrete Score Matching
by: Kim, Yeongmin, et al.
Published: (2025)
by: Kim, Yeongmin, et al.
Published: (2025)
Inversion-based Latent Bayesian Optimization
by: Chu, Jaewon, et al.
Published: (2024)
by: Chu, Jaewon, et al.
Published: (2024)
KaSA: Knowledge-Aware Singular-Value Adaptation of Large Language Models
by: Wang, Fan, et al.
Published: (2024)
by: Wang, Fan, et al.
Published: (2024)
The Role of Teacher Calibration in Knowledge Distillation
by: Kim, Suyoung, et al.
Published: (2025)
by: Kim, Suyoung, et al.
Published: (2025)
ESI: Epistemic Uncertainty Quantification via Semantic-preserving Intervention for Large Language Models
by: Li, Mingda, et al.
Published: (2025)
by: Li, Mingda, et al.
Published: (2025)
Lookahead Unmasking Elicits Accurate Decoding in Diffusion Language Models
by: Lee, Sanghyun, et al.
Published: (2025)
by: Lee, Sanghyun, et al.
Published: (2025)
Data Descriptions from Large Language Models with Influence Estimation
by: Kim, Chaeri, et al.
Published: (2025)
by: Kim, Chaeri, et al.
Published: (2025)
Contrastive Residual Energy Test-time Adaptation
by: Han, Yewon, et al.
Published: (2025)
by: Han, Yewon, et al.
Published: (2025)
MAGNET: Autonomous Expert Model Generation via Decentralized Autoresearch and BitNet Training
by: Kim, Yongwan, et al.
Published: (2026)
by: Kim, Yongwan, et al.
Published: (2026)
Robust Domain Generalization under Divergent Marginal and Conditional Distributions
by: Yeom, Jewon, et al.
Published: (2026)
by: Yeom, Jewon, et al.
Published: (2026)
Action-Sufficient Goal Representations
by: Hyeon, Jinu, et al.
Published: (2026)
by: Hyeon, Jinu, et al.
Published: (2026)
Diffusion-Based Offline RL for Improved Decision-Making in Augmented ARC Task
by: Kim, Yunho, et al.
Published: (2024)
by: Kim, Yunho, et al.
Published: (2024)
CAdam: Context-Adaptive Moment Estimation for 3D Gaussian Densification in Generative Distillation
by: Chung, SeungJeh, et al.
Published: (2026)
by: Chung, SeungJeh, et al.
Published: (2026)
Attention-space Contrastive Guidance for Efficient Hallucination Mitigation in LVLMs
by: Jo, Yujin, et al.
Published: (2026)
by: Jo, Yujin, et al.
Published: (2026)
Quantifying Epistemic Uncertainty in Diffusion Models
by: Gupta, Aditi, et al.
Published: (2026)
by: Gupta, Aditi, et al.
Published: (2026)
PromptKD: Distilling Student-Friendly Knowledge for Generative Language Models via Prompt Tuning
by: Kim, Gyeongman, et al.
Published: (2024)
by: Kim, Gyeongman, et al.
Published: (2024)
Learning to Act Robustly with View-Invariant Latent Actions
by: Jeong, Youngjoon, et al.
Published: (2026)
by: Jeong, Youngjoon, et al.
Published: (2026)
Knowledge Distillation from Language-Oriented to Emergent Communication for Multi-Agent Remote Control
by: Kim, Yongjun, et al.
Published: (2024)
by: Kim, Yongjun, et al.
Published: (2024)
LLM-NEO: Parameter Efficient Knowledge Distillation for Large Language Models
by: Yang, Runming, et al.
Published: (2024)
by: Yang, Runming, et al.
Published: (2024)
Text-Aware Image Restoration with Diffusion Models
by: Min, Jaewon, et al.
Published: (2025)
by: Min, Jaewon, et al.
Published: (2025)
Acceleration of Grokking in Learning Arithmetic Operations via Kolmogorov-Arnold Representation
by: Park, Yeachan, et al.
Published: (2024)
by: Park, Yeachan, et al.
Published: (2024)
Dist2ill: Distributional Distillation for One-Pass Uncertainty Estimation in Large Language Models
by: Zhao, Yicong, et al.
Published: (2025)
by: Zhao, Yicong, et al.
Published: (2025)
Similar Items
-
Garbage Attention in Large Language Models: BOS Sink Heads and Sink-aware Pruning
by: Sok, Jaewon, et al.
Published: (2026) -
From Noise to Diversity: Random Embedding Injection in LLM Reasoning
by: Kim, Heejun, et al.
Published: (2026) -
Hallucination as Commitment Failure: Larger LLMs Misfire Despite Knowing the Answer
by: Yeom, Jewon, et al.
Published: (2026) -
EpiCaR: Knowing What You Don't Know Matters for Better Reasoning in LLMs
by: Yeom, Jewon, et al.
Published: (2026) -
Stable On-Policy Distillation through Adaptive Target Reformulation
by: Jang, Ijun, et al.
Published: (2026)