Guardado en:
| Autores principales: | Phillips, Edward, Wu, Sean, Gustafsson, Fredrik K., Gao, Boyan, Clifton, David A. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.04577 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence
por: Wu, Sean, et al.
Publicado: (2026)
por: Wu, Sean, et al.
Publicado: (2026)
Entropy Alone is Insufficient for Safe Selective Prediction in LLMs
por: Phillips, Edward, et al.
Publicado: (2026)
por: Phillips, Edward, et al.
Publicado: (2026)
Geometric Uncertainty for Detecting and Correcting Hallucinations in LLMs
por: Phillips, Edward, et al.
Publicado: (2025)
por: Phillips, Edward, et al.
Publicado: (2025)
Uncertainty Distillation: Teaching Language Models to Express Semantic Confidence
por: Hager, Sophia, et al.
Publicado: (2025)
por: Hager, Sophia, et al.
Publicado: (2025)
SpikeLLM: Scaling up Spiking Neural Network to Large Language Models via Saliency-based Spiking
por: Xing, Xingrun, et al.
Publicado: (2024)
por: Xing, Xingrun, et al.
Publicado: (2024)
Cognition Chain for Explainable Psychological Stress Detection on Social Media
por: Wang, Xin, et al.
Publicado: (2024)
por: Wang, Xin, et al.
Publicado: (2024)
Trust Region On-Policy Distillation
por: Xing, Xingrun, et al.
Publicado: (2026)
por: Xing, Xingrun, et al.
Publicado: (2026)
Semantic Self-Consistency: Enhancing Language Model Reasoning via Semantic Weighting
por: Knappe, Tim, et al.
Publicado: (2024)
por: Knappe, Tim, et al.
Publicado: (2024)
Self-Data Distillation for Recovering Quality in Pruned Large Language Models
por: Thangarasa, Vithursan, et al.
Publicado: (2024)
por: Thangarasa, Vithursan, et al.
Publicado: (2024)
Benchmarking Pathology Foundation Models for Breast Cancer Survival Prediction
por: Gustafsson, Fredrik K., et al.
Publicado: (2026)
por: Gustafsson, Fredrik K., et al.
Publicado: (2026)
Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models
por: Zhao, Siyan, et al.
Publicado: (2026)
por: Zhao, Siyan, et al.
Publicado: (2026)
Uncertainty in Semantic Language Modeling with PIXELS
por: Radu, Stefania, et al.
Publicado: (2025)
por: Radu, Stefania, et al.
Publicado: (2025)
From Large to Tiny: Distilling and Refining Mathematical Expertise for Math Word Problems with Weakly Supervision
por: Lin, Qingwen, et al.
Publicado: (2024)
por: Lin, Qingwen, et al.
Publicado: (2024)
MASSV: Multimodal Adaptation and Self-Data Distillation for Speculative Decoding of Vision-Language Models
por: Ganesan, Mugilan, et al.
Publicado: (2025)
por: Ganesan, Mugilan, et al.
Publicado: (2025)
Multi-Granularity Semantic Revision for Large Language Model Distillation
por: Liu, Xiaoyu, et al.
Publicado: (2024)
por: Liu, Xiaoyu, et al.
Publicado: (2024)
Self-Distillation for Model Stacking Unlocks Cross-Lingual NLU in 200+ Languages
por: Schmidt, Fabian David, et al.
Publicado: (2024)
por: Schmidt, Fabian David, et al.
Publicado: (2024)
FE-Adapter: Adapting Image-based Emotion Classifiers to Videos
por: Gowda, Shreyank N, et al.
Publicado: (2024)
por: Gowda, Shreyank N, et al.
Publicado: (2024)
Can Language Models Be Tricked by Language Illusions? Easier with Syntax, Harder with Semantics
por: Zhang, Yuhan, et al.
Publicado: (2023)
por: Zhang, Yuhan, et al.
Publicado: (2023)
PLPP: Prompt Learning with Perplexity Is Self-Distillation for Vision-Language Models
por: Liu, Biao, et al.
Publicado: (2024)
por: Liu, Biao, et al.
Publicado: (2024)
HumorGen: Cognitive Synergy for Humor Generation in Large Language Models via Persona-Based Distillation
por: Ajayi, Edward, et al.
Publicado: (2026)
por: Ajayi, Edward, et al.
Publicado: (2026)
Self-Updatable Large Language Models by Integrating Context into Model Parameters
por: Wang, Yu, et al.
Publicado: (2024)
por: Wang, Yu, et al.
Publicado: (2024)
D2LLM: Decomposed and Distilled Large Language Models for Semantic Search
por: Liao, Zihan, et al.
Publicado: (2024)
por: Liao, Zihan, et al.
Publicado: (2024)
Language Confusion Gate: Language-Aware Decoding Through Model Self-Distillation
por: Zhang, Collin, et al.
Publicado: (2025)
por: Zhang, Collin, et al.
Publicado: (2025)
On-Policy Context Distillation for Language Models
por: Ye, Tianzhu, et al.
Publicado: (2026)
por: Ye, Tianzhu, et al.
Publicado: (2026)
Are Large Language Models Good Statisticians?
por: Zhu, Yizhang, et al.
Publicado: (2024)
por: Zhu, Yizhang, et al.
Publicado: (2024)
Self-Distillation Bridges Distribution Gap in Language Model Fine-Tuning
por: Yang, Zhaorui, et al.
Publicado: (2024)
por: Yang, Zhaorui, et al.
Publicado: (2024)
Mind's Mirror: Distilling Self-Evaluation Capability and Comprehensive Thinking from Large Language Models
por: Liu, Weize, et al.
Publicado: (2023)
por: Liu, Weize, et al.
Publicado: (2023)
Self-Distilled Trajectory-Aware Boltzmann Modeling: Bridging the Training-Inference Discrepancy in Diffusion Language Models
por: Chen, Kecheng, et al.
Publicado: (2026)
por: Chen, Kecheng, et al.
Publicado: (2026)
Estimating the Black-box LLM Uncertainty with Distribution-Aligned Adversarial Distillation
por: Cui, Huizi, et al.
Publicado: (2026)
por: Cui, Huizi, et al.
Publicado: (2026)
ROSD: Reflective On-Policy Self-Distillation for Language Model Reasoning across Domains
por: Zhao, Ziqi, et al.
Publicado: (2026)
por: Zhao, Ziqi, et al.
Publicado: (2026)
No Reliable Evidence of Self-Reported Sentience in Small Large Language Models
por: Kaiser, Caspar, et al.
Publicado: (2026)
por: Kaiser, Caspar, et al.
Publicado: (2026)
Memorization Dynamics in Knowledge Distillation for Language Models
por: Borkar, Jaydeep, et al.
Publicado: (2026)
por: Borkar, Jaydeep, et al.
Publicado: (2026)
OPSDL: On-Policy Self-Distillation for Long-Context Language Models
por: Zhang, Xinsen, et al.
Publicado: (2026)
por: Zhang, Xinsen, et al.
Publicado: (2026)
Self-Refining Language Model Anonymizers via Adversarial Distillation
por: Kim, Kyuyoung, et al.
Publicado: (2025)
por: Kim, Kyuyoung, et al.
Publicado: (2025)
Semantic Density: Uncertainty Quantification for Large Language Models through Confidence Measurement in Semantic Space
por: Qiu, Xin, et al.
Publicado: (2024)
por: Qiu, Xin, et al.
Publicado: (2024)
BattleAgentBench: A Benchmark for Evaluating Cooperation and Competition Capabilities of Language Models in Multi-Agent Systems
por: Wang, Wei, et al.
Publicado: (2024)
por: Wang, Wei, et al.
Publicado: (2024)
Retrieval-Augmented and Knowledge-Grounded Language Models for Faithful Clinical Medicine
por: Liu, Fenglin, et al.
Publicado: (2022)
por: Liu, Fenglin, et al.
Publicado: (2022)
Advantage-Guided Distillation for Preference Alignment in Small Language Models
por: Gao, Shiping, et al.
Publicado: (2025)
por: Gao, Shiping, et al.
Publicado: (2025)
Quantification of Large Language Model Distillation
por: Lee, Sunbowen, et al.
Publicado: (2025)
por: Lee, Sunbowen, et al.
Publicado: (2025)
Evaluating Deep Regression Models for WSI-Based Gene-Expression Prediction
por: Gustafsson, Fredrik K., et al.
Publicado: (2024)
por: Gustafsson, Fredrik K., et al.
Publicado: (2024)
Ejemplares similares
-
BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence
por: Wu, Sean, et al.
Publicado: (2026) -
Entropy Alone is Insufficient for Safe Selective Prediction in LLMs
por: Phillips, Edward, et al.
Publicado: (2026) -
Geometric Uncertainty for Detecting and Correcting Hallucinations in LLMs
por: Phillips, Edward, et al.
Publicado: (2025) -
Uncertainty Distillation: Teaching Language Models to Express Semantic Confidence
por: Hager, Sophia, et al.
Publicado: (2025) -
SpikeLLM: Scaling up Spiking Neural Network to Large Language Models via Saliency-based Spiking
por: Xing, Xingrun, et al.
Publicado: (2024)