Dist2ill: Distributional Distillation for One-Pass Uncertainty Estimation in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Yicong, Tsang, King Yeung, Vejendla, Harshil, Shi, Haizhou, Li, Zhuohang, Hua, Zhigang, Xu, Qi, Zhang, Tunyu, Wang, Yi, Han, Ligong, Malin, Bradley A., Wang, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
H1B-KV: Hybrid One-Bit Caches for Memory-Efficient Large Language Model Inference
by: Vejendla, Harshil
Published: (2025)
by: Vejendla, Harshil
Published: (2025)
Teaching by Failure: Counter-Example-Driven Curricula for Transformer Self-Improvement
by: Vejendla, Harshil
Published: (2025)
by: Vejendla, Harshil
Published: (2025)
SliceMoE: Routing Embedding Slices Instead of Tokens for Fine-Grained and Balanced Transformer Scaling
by: Vejendla, Harshil
Published: (2025)
by: Vejendla, Harshil
Published: (2025)
RewriteNets: End-to-End Trainable String-Rewriting for Generative Sequence Modeling
by: Vejendla, Harshil
Published: (2026)
by: Vejendla, Harshil
Published: (2026)
LATTA: Langevin-Anchored Test-Time Adaptation for Enhanced Robustness and Stability
by: Vejendla, Harshil
Published: (2025)
by: Vejendla, Harshil
Published: (2025)
Wave-PDE Nets: Trainable Wave-Equation Layers as an Alternative to Attention
by: Vejendla, Harshil
Published: (2025)
by: Vejendla, Harshil
Published: (2025)
Learning to Predict Chaos: Curriculum-Driven Training for Robust Forecasting of Chaotic Dynamics
by: Vejendla, Harshil
Published: (2025)
by: Vejendla, Harshil
Published: (2025)
Drift-Adapter: A Practical Approach to Near Zero-Downtime Embedding Model Upgrades in Vector Databases
by: Vejendla, Harshil
Published: (2025)
by: Vejendla, Harshil
Published: (2025)
TokUR: Token-Level Uncertainty Estimation for Large Language Model Reasoning
by: Zhang, Tunyu, et al.
Published: (2025)
by: Zhang, Tunyu, et al.
Published: (2025)
DistDD: Distributed Data Distillation Aggregation through Gradient Matching
by: Wang, Peiran, et al.
Published: (2024)
by: Wang, Peiran, et al.
Published: (2024)
Few-Step Diffusion Language Models via Trajectory Self-Distillation
by: Zhang, Tunyu, et al.
Published: (2026)
by: Zhang, Tunyu, et al.
Published: (2026)
Training-Free Bayesianization for Low-Rank Adapters of Large Language Models
by: Shi, Haizhou, et al.
Published: (2024)
by: Shi, Haizhou, et al.
Published: (2024)
BLoB: Bayesian Low-Rank Adaptation by Backpropagation for Large Language Models
by: Wang, Yibin, et al.
Published: (2024)
by: Wang, Yibin, et al.
Published: (2024)
From Passive Metric to Active Signal: The Evolving Role of Uncertainty Quantification in Large Language Models
by: Zhang, Jiaxin, et al.
Published: (2026)
by: Zhang, Jiaxin, et al.
Published: (2026)
Multimodal Needle in a Haystack: Benchmarking Long-Context Capability of Multimodal Large Language Models
by: Wang, Hengyi, et al.
Published: (2024)
by: Wang, Hengyi, et al.
Published: (2024)
SAC3: Reliable Hallucination Detection in Black-Box Language Models via Semantic-aware Cross-check Consistency
by: Zhang, Jiaxin, et al.
Published: (2023)
by: Zhang, Jiaxin, et al.
Published: (2023)
Multi-Modal Dataset Distillation in the Wild
by: Dang, Zhuohang, et al.
Published: (2025)
by: Dang, Zhuohang, et al.
Published: (2025)
Improved Random-Binning Exponent for Distributed Hypothesis Testing
by: Kochman, Yuval, et al.
Published: (2023)
by: Kochman, Yuval, et al.
Published: (2023)
Estimating the Black-box LLM Uncertainty with Distribution-Aligned Adversarial Distillation
by: Cui, Huizi, et al.
Published: (2026)
by: Cui, Huizi, et al.
Published: (2026)
DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
by: Wang, Taiyi, et al.
Published: (2024)
by: Wang, Taiyi, et al.
Published: (2024)
GeoDistNet: An Open-Source Tool for Synthetic Distribution Network Generation
by: Wang, Yunqi, et al.
Published: (2026)
by: Wang, Yunqi, et al.
Published: (2026)
EDUE: Expert Disagreement-Guided One-Pass Uncertainty Estimation for Medical Image Segmentation
by: Abutalip, Kudaibergen, et al.
Published: (2024)
by: Abutalip, Kudaibergen, et al.
Published: (2024)
Infinite-LLM: Efficient LLM Service for Long Context with DistAttention and Distributed KVCache
by: Lin, Bin, et al.
Published: (2024)
by: Lin, Bin, et al.
Published: (2024)
Exploring User-level Gradient Inversion with a Diffusion Prior
by: Li, Zhuohang, et al.
Published: (2024)
by: Li, Zhuohang, et al.
Published: (2024)
Analyzing Inference Privacy Risks Through Gradients in Machine Learning
by: Li, Zhuohang, et al.
Published: (2024)
by: Li, Zhuohang, et al.
Published: (2024)
DistFlow: A Fully Distributed RL Framework for Scalable and Efficient LLM Post-Training
by: Wang, Zhixin, et al.
Published: (2025)
by: Wang, Zhixin, et al.
Published: (2025)
Why Settle for One? Text-to-ImageSet Generation and Evaluation
by: Jia, Chengyou, et al.
Published: (2025)
by: Jia, Chengyou, et al.
Published: (2025)
Improved Turbo Message Passing for Compressive Robust Principal Component Analysis: Algorithm Design and Asymptotic Analysis
by: He, Zhuohang, et al.
Published: (2024)
by: He, Zhuohang, et al.
Published: (2024)
DistDF: Time-Series Forecasting Needs Joint-Distribution Wasserstein Alignment
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
DCR-Consistency: Divide-Conquer-Reasoning for Consistency Evaluation and Improvement of Large Language Models
by: Cui, Wendi, et al.
Published: (2024)
by: Cui, Wendi, et al.
Published: (2024)
StarDist: A Code Generator for Distributed Graph Algorithms
by: Nandy, Barenya Kumar, et al.
Published: (2025)
by: Nandy, Barenya Kumar, et al.
Published: (2025)
DistShap: Scalable GNN Explanations with Distributed Shapley Values
by: Akkas, Selahattin, et al.
Published: (2025)
by: Akkas, Selahattin, et al.
Published: (2025)
Heavy-Tailed Linear Bandits: Huber Regression with One-Pass Update
by: Wang, Jing, et al.
Published: (2025)
by: Wang, Jing, et al.
Published: (2025)
DistZO2: High-Throughput and Memory-Efficient Zeroth-Order Fine-tuning LLMs with Distributed Parallel Computing
by: Wang, Liangyu, et al.
Published: (2025)
by: Wang, Liangyu, et al.
Published: (2025)
Do You Know What You Are Talking About? Characterizing Query-Knowledge Relevance For Reliable Retrieval Augmented Generation
by: Li, Zhuohang, et al.
Published: (2024)
by: Li, Zhuohang, et al.
Published: (2024)
OptDist: Learning Optimal Distribution for Customer Lifetime Value Prediction
by: Weng, Yunpeng, et al.
Published: (2024)
by: Weng, Yunpeng, et al.
Published: (2024)
DistJoin: A Decoupled Join Cardinality Estimator based on Adaptive Neural Predicate Modulation
by: Zhang, Kaixin, et al.
Published: (2025)
by: Zhang, Kaixin, et al.
Published: (2025)
Investigating Sparsity in Recurrent Neural Networks
by: Darji, Harshil
Published: (2024)
by: Darji, Harshil
Published: (2024)
Pseudo-D: Informing Multi-View Uncertainty Estimation with Calibrated Neural Training Dynamics
by: Gu, Ang Nan, et al.
Published: (2025)
by: Gu, Ang Nan, et al.
Published: (2025)
Haste makes waste: Early nutrition prescription for critically ill patients
by: Siying Chen, et al.
Published: (2024)
by: Siying Chen, et al.
Published: (2024)
Similar Items
-
H1B-KV: Hybrid One-Bit Caches for Memory-Efficient Large Language Model Inference
by: Vejendla, Harshil
Published: (2025) -
Teaching by Failure: Counter-Example-Driven Curricula for Transformer Self-Improvement
by: Vejendla, Harshil
Published: (2025) -
SliceMoE: Routing Embedding Slices Instead of Tokens for Fine-Grained and Balanced Transformer Scaling
by: Vejendla, Harshil
Published: (2025) -
RewriteNets: End-to-End Trainable String-Rewriting for Generative Sequence Modeling
by: Vejendla, Harshil
Published: (2026) -
LATTA: Langevin-Anchored Test-Time Adaptation for Enhanced Robustness and Stability
by: Vejendla, Harshil
Published: (2025)