Enregistré dans:
| Auteurs principaux: | Vengertsev, Dmitry, Sherman, Elena |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2401.11365 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Confidence-aware Self-Semantic Distillation on Knowledge Graph Embedding
par: Liu, Yichen, et autres
Publié: (2022)
par: Liu, Yichen, et autres
Publié: (2022)
Uncertainty Distillation: Teaching Language Models to Express Semantic Confidence
par: Hager, Sophia, et autres
Publié: (2025)
par: Hager, Sophia, et autres
Publié: (2025)
Confidence Estimation for Error Detection in Text-to-SQL Systems
par: Somov, Oleg, et autres
Publié: (2025)
par: Somov, Oleg, et autres
Publié: (2025)
Knowledge Distillation with Training Wheels
par: Liu, Guanlin, et autres
Publié: (2025)
par: Liu, Guanlin, et autres
Publié: (2025)
Sinkhorn Distance Minimization for Knowledge Distillation
par: Cui, Xiao, et autres
Publié: (2024)
par: Cui, Xiao, et autres
Publié: (2024)
Detecting Conceptual Abstraction in LLMs
par: Regneri, Michaela, et autres
Publié: (2024)
par: Regneri, Michaela, et autres
Publié: (2024)
LoCa: Logit Calibration for Knowledge Distillation
par: Yang, Runming, et autres
Publié: (2024)
par: Yang, Runming, et autres
Publié: (2024)
Do LLMs Really Forget? Evaluating Unlearning with Knowledge Correlation and Confidence Awareness
par: Wei, Rongzhe, et autres
Publié: (2025)
par: Wei, Rongzhe, et autres
Publié: (2025)
Knowledge Distillation and Dataset Distillation of Large Language Models: Emerging Trends, Challenges, and Future Directions
par: Fang, Luyang, et autres
Publié: (2025)
par: Fang, Luyang, et autres
Publié: (2025)
Efficient Knowledge Injection in LLMs via Self-Distillation
par: Kujanpää, Kalle, et autres
Publié: (2024)
par: Kujanpää, Kalle, et autres
Publié: (2024)
In Good GRACEs: Principled Teacher Selection for Knowledge Distillation
par: Panigrahi, Abhishek, et autres
Publié: (2025)
par: Panigrahi, Abhishek, et autres
Publié: (2025)
Brewing Knowledge in Context: Distillation Perspectives on In-Context Learning
par: Li, Chengye, et autres
Publié: (2025)
par: Li, Chengye, et autres
Publié: (2025)
DistillSpec: Improving Speculative Decoding via Knowledge Distillation
par: Zhou, Yongchao, et autres
Publié: (2023)
par: Zhou, Yongchao, et autres
Publié: (2023)
A Survey on Symbolic Knowledge Distillation of Large Language Models
par: Acharya, Kamal, et autres
Publié: (2024)
par: Acharya, Kamal, et autres
Publié: (2024)
Relational Knowledge Distillation Using Fine-tuned Function Vectors
par: Kang, Andrea, et autres
Publié: (2026)
par: Kang, Andrea, et autres
Publié: (2026)
Tailoring Instructions to Student's Learning Levels Boosts Knowledge Distillation
par: Ren, Yuxin, et autres
Publié: (2023)
par: Ren, Yuxin, et autres
Publié: (2023)
Reinforcement Learning-based Knowledge Distillation with LLM-as-a-Judge
par: Shen, Yiyang, et autres
Publié: (2026)
par: Shen, Yiyang, et autres
Publié: (2026)
Language Model Knowledge Distillation for Efficient Question Answering in Spanish
par: Bazaga, Adrián, et autres
Publié: (2023)
par: Bazaga, Adrián, et autres
Publié: (2023)
Short Data, Long Context: Distilling Positional Knowledge in Transformers
par: Huber, Patrick, et autres
Publié: (2026)
par: Huber, Patrick, et autres
Publié: (2026)
X-Token: Projection-Guided Cross-Tokenizer Knowledge Distillation
par: Sreenivas, Sharath Turuvekere, et autres
Publié: (2026)
par: Sreenivas, Sharath Turuvekere, et autres
Publié: (2026)
Show Your Work with Confidence: Confidence Bands for Tuning Curves
par: Lourie, Nicholas, et autres
Publié: (2023)
par: Lourie, Nicholas, et autres
Publié: (2023)
Feature Alignment and Representation Transfer in Knowledge Distillation for Large Language Models
par: Yang, Junjie, et autres
Publié: (2025)
par: Yang, Junjie, et autres
Publié: (2025)
Knowledge Distillation from Large Language Models for Household Energy Modeling
par: Takrouri, Mohannad, et autres
Publié: (2025)
par: Takrouri, Mohannad, et autres
Publié: (2025)
Smoothing Out Hallucinations: Mitigating LLM Hallucination with Smoothed Knowledge Distillation
par: Nguyen, Hieu, et autres
Publié: (2025)
par: Nguyen, Hieu, et autres
Publié: (2025)
Implicit Word Reordering with Knowledge Distillation for Cross-Lingual Dependency Parsing
par: Li, Zhuoran, et autres
Publié: (2025)
par: Li, Zhuoran, et autres
Publié: (2025)
UniMaia: Steering Chess Policies with Language for Human-like Play
par: Siu, Sherman, et autres
Publié: (2026)
par: Siu, Sherman, et autres
Publié: (2026)
HarmAug: Effective Data Augmentation for Knowledge Distillation of Safety Guard Models
par: Lee, Seanie, et autres
Publié: (2024)
par: Lee, Seanie, et autres
Publié: (2024)
MLKD-BERT: Multi-level Knowledge Distillation for Pre-trained Language Models
par: Zhang, Ying, et autres
Publié: (2024)
par: Zhang, Ying, et autres
Publié: (2024)
Delta Knowledge Distillation for Large Language Models
par: Cao, Yihan, et autres
Publié: (2025)
par: Cao, Yihan, et autres
Publié: (2025)
Multi-modal Anchor Gated Transformer with Knowledge Distillation for Emotion Recognition in Conversation
par: Li, Jie, et autres
Publié: (2025)
par: Li, Jie, et autres
Publié: (2025)
Confidence over Time: Confidence Calibration with Temporal Logic for Large Language Model Reasoning
par: Mao, Zhenjiang, et autres
Publié: (2026)
par: Mao, Zhenjiang, et autres
Publié: (2026)
Capturing Sparks of Abstraction for the ARC Challenge
par: Andrews, Martin
Publié: (2024)
par: Andrews, Martin
Publié: (2024)
CURaTE: Continual Unlearning in Real Time with Ensured Preservation of LLM Knowledge
par: Bae, Seyun, et autres
Publié: (2026)
par: Bae, Seyun, et autres
Publié: (2026)
Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training
par: Rahmati, Elnaz, et autres
Publié: (2026)
par: Rahmati, Elnaz, et autres
Publié: (2026)
Multicalibration for Confidence Scoring in LLMs
par: Detommaso, Gianluca, et autres
Publié: (2024)
par: Detommaso, Gianluca, et autres
Publié: (2024)
Knowledge Distillation in Automated Annotation: Supervised Text Classification with LLM-Generated Training Labels
par: Pangakis, Nicholas, et autres
Publié: (2024)
par: Pangakis, Nicholas, et autres
Publié: (2024)
Exploring the Limits of Model Compression in LLMs: A Knowledge Distillation Study on QA Tasks
par: Datta, Joyeeta, et autres
Publié: (2025)
par: Datta, Joyeeta, et autres
Publié: (2025)
Compact Language Models via Pruning and Knowledge Distillation
par: Muralidharan, Saurav, et autres
Publié: (2024)
par: Muralidharan, Saurav, et autres
Publié: (2024)
Improving Neural Topic Models with Wasserstein Knowledge Distillation
par: Adhya, Suman, et autres
Publié: (2023)
par: Adhya, Suman, et autres
Publié: (2023)
Analogical Reasoning Inside Large Language Models: Concept Vectors and the Limits of Abstraction
par: Opiełka, Gustaw, et autres
Publié: (2025)
par: Opiełka, Gustaw, et autres
Publié: (2025)
Documents similaires
-
Confidence-aware Self-Semantic Distillation on Knowledge Graph Embedding
par: Liu, Yichen, et autres
Publié: (2022) -
Uncertainty Distillation: Teaching Language Models to Express Semantic Confidence
par: Hager, Sophia, et autres
Publié: (2025) -
Confidence Estimation for Error Detection in Text-to-SQL Systems
par: Somov, Oleg, et autres
Publié: (2025) -
Knowledge Distillation with Training Wheels
par: Liu, Guanlin, et autres
Publié: (2025) -
Sinkhorn Distance Minimization for Knowledge Distillation
par: Cui, Xiao, et autres
Publié: (2024)