Knowledge Distillation for Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | La Torre, Alejandro Paredes, Flores, Barbara, Rodriguez, Diego |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PPLqa: An Unsupervised Information-Theoretic Quality Metric for Comparing Generative Large Language Models
von: Friedland, Gerald, et al.
Veröffentlicht: (2024)
von: Friedland, Gerald, et al.
Veröffentlicht: (2024)
ValueDCG: Measuring Comprehensive Human Value Understanding Ability of Language Models
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2023)
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2023)
The QCET Taxonomy of Standard Quality Criterion Names and Definitions for the Evaluation of NLP Systems
von: Belz, Anya, et al.
Veröffentlicht: (2025)
von: Belz, Anya, et al.
Veröffentlicht: (2025)
Semantic Delta: An Interpretable Signal Differentiating Human and LLMs Dialogue
von: Scantamburlo, Riccardo, et al.
Veröffentlicht: (2026)
von: Scantamburlo, Riccardo, et al.
Veröffentlicht: (2026)
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
von: Wen, Yuqiao, et al.
Veröffentlicht: (2025)
von: Wen, Yuqiao, et al.
Veröffentlicht: (2025)
Bielik-Minitron-7B: Compressing Large Language Models via Structured Pruning and Knowledge Distillation for the Polish Language
von: Kinas, Remigiusz, et al.
Veröffentlicht: (2026)
von: Kinas, Remigiusz, et al.
Veröffentlicht: (2026)
Quo Vadis ChatGPT? From Large Language Models to Large Knowledge Models
von: Venkatasubramanian, Venkat, et al.
Veröffentlicht: (2024)
von: Venkatasubramanian, Venkat, et al.
Veröffentlicht: (2024)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
von: Wen, Yuqiao, et al.
Veröffentlicht: (2024)
von: Wen, Yuqiao, et al.
Veröffentlicht: (2024)
Enhancing Retrieval-Augmented Generation for Electric Power Industry Customer Support
von: Chan, Hei Yu, et al.
Veröffentlicht: (2025)
von: Chan, Hei Yu, et al.
Veröffentlicht: (2025)
Leveraging Synthetic Data for Question Answering with Multilingual LLMs in the Agricultural Domain
von: Kaur, Rishemjit, et al.
Veröffentlicht: (2025)
von: Kaur, Rishemjit, et al.
Veröffentlicht: (2025)
A Framework for Collaborating a Large Language Model Tool in Brainstorming for Triggering Creative Thoughts
von: Chang, Hung-Fu, et al.
Veröffentlicht: (2024)
von: Chang, Hung-Fu, et al.
Veröffentlicht: (2024)
A Super-Learner with Large Language Models for Medical Emergency Advising
von: Aityan, Sergey K., et al.
Veröffentlicht: (2025)
von: Aityan, Sergey K., et al.
Veröffentlicht: (2025)
The Battle of LLMs: A Comparative Study in Conversational QA Tasks
von: Rangapur, Aryan, et al.
Veröffentlicht: (2024)
von: Rangapur, Aryan, et al.
Veröffentlicht: (2024)
Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models
von: Wang, Yating, et al.
Veröffentlicht: (2026)
von: Wang, Yating, et al.
Veröffentlicht: (2026)
Evaluating Class Membership Relations in Knowledge Graphs using Large Language Models
von: Allen, Bradley P., et al.
Veröffentlicht: (2024)
von: Allen, Bradley P., et al.
Veröffentlicht: (2024)
KNOW: A Real-World Ontology for Knowledge Capture with Large Language Models
von: Bendiken, Arto
Veröffentlicht: (2024)
von: Bendiken, Arto
Veröffentlicht: (2024)
Dialogic Pedagogy for Large Language Models: Aligning Conversational AI with Proven Theories of Learning
von: Beale, Russell
Veröffentlicht: (2025)
von: Beale, Russell
Veröffentlicht: (2025)
Opinion Mining on Offshore Wind Energy for Environmental Engineering
von: Bittencourt, Isabele, et al.
Veröffentlicht: (2024)
von: Bittencourt, Isabele, et al.
Veröffentlicht: (2024)
Text Clustering with Large Language Model Embeddings
von: Petukhova, Alina, et al.
Veröffentlicht: (2024)
von: Petukhova, Alina, et al.
Veröffentlicht: (2024)
Large Language Models Can Better Understand Knowledge Graphs Than We Thought
von: Dai, Xinbang, et al.
Veröffentlicht: (2024)
von: Dai, Xinbang, et al.
Veröffentlicht: (2024)
Prompt-Time Symbolic Knowledge Capture with Large Language Models
von: Çöplü, Tolga, et al.
Veröffentlicht: (2024)
von: Çöplü, Tolga, et al.
Veröffentlicht: (2024)
Context-Aware Clustering using Large Language Models
von: Tipirneni, Sindhu, et al.
Veröffentlicht: (2024)
von: Tipirneni, Sindhu, et al.
Veröffentlicht: (2024)
Turning the TIDE: Cross-Architecture Distillation for Diffusion Large Language Models
von: Zhang, Gongbo, et al.
Veröffentlicht: (2026)
von: Zhang, Gongbo, et al.
Veröffentlicht: (2026)
Distilling Knowledge from Large Language Models: A Concept Bottleneck Model for Hate and Counter Speech Recognition
von: Labadie-Tamayo, Roberto, et al.
Veröffentlicht: (2025)
von: Labadie-Tamayo, Roberto, et al.
Veröffentlicht: (2025)
Revisiting Parameter-Based Knowledge Editing in Large Language Models: Theoretical Limits and Empirical Evidence
von: Ren, Wanying, et al.
Veröffentlicht: (2026)
von: Ren, Wanying, et al.
Veröffentlicht: (2026)
Relations Prediction for Knowledge Graph Completion using Large Language Models
von: Alqaaidi, Sakher Khalil, et al.
Veröffentlicht: (2024)
von: Alqaaidi, Sakher Khalil, et al.
Veröffentlicht: (2024)
Few-Shot Learning for Mental Disorder Detection: A Continuous Multi-Prompt Engineering Approach with Medical Knowledge Injection
von: Liu, Haoxin, et al.
Veröffentlicht: (2024)
von: Liu, Haoxin, et al.
Veröffentlicht: (2024)
Prompt-Time Ontology-Driven Symbolic Knowledge Capture with Large Language Models
von: Çöplü, Tolga, et al.
Veröffentlicht: (2024)
von: Çöplü, Tolga, et al.
Veröffentlicht: (2024)
Extracting Probabilistic Knowledge from Large Language Models for Bayesian Network Parameterization
von: Nafar, Aliakbar, et al.
Veröffentlicht: (2025)
von: Nafar, Aliakbar, et al.
Veröffentlicht: (2025)
Alif: Advancing Urdu Large Language Models via Multilingual Synthetic Data Distillation
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2025)
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models on Historical Health Crisis Knowledge in Resource-Limited Settings: A Hybrid Multi-Metric Study
von: Hasan, Mohammed Rakibul
Veröffentlicht: (2026)
von: Hasan, Mohammed Rakibul
Veröffentlicht: (2026)
APP: Accelerated Path Patching with Task-Specific Pruning
von: Andersen, Frauke, et al.
Veröffentlicht: (2025)
von: Andersen, Frauke, et al.
Veröffentlicht: (2025)
Beyond Prefixes: Graph-as-Memory Cross-Attention for Knowledge Graph Completion with Large Language Models
von: Liu, Ruitong, et al.
Veröffentlicht: (2025)
von: Liu, Ruitong, et al.
Veröffentlicht: (2025)
Interactive-KBQA: Multi-Turn Interactions for Knowledge Base Question Answering with Large Language Models
von: Xiong, Guanming, et al.
Veröffentlicht: (2024)
von: Xiong, Guanming, et al.
Veröffentlicht: (2024)
Pre-trained Language Model with Prompts for Temporal Knowledge Graph Completion
von: Xu, Wenjie, et al.
Veröffentlicht: (2023)
von: Xu, Wenjie, et al.
Veröffentlicht: (2023)
High-Throughput Phenotyping of Clinical Text Using Large Language Models
von: Hier, Daniel B., et al.
Veröffentlicht: (2024)
von: Hier, Daniel B., et al.
Veröffentlicht: (2024)
What should I say? -- Interacting with AI and Natural Language Interfaces
von: Adkins, Mark
Veröffentlicht: (2024)
von: Adkins, Mark
Veröffentlicht: (2024)
Simulating a Bias Mitigation Scenario in Large Language Models
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
Eliciting Problem Specifications via Large Language Models
von: Wray, Robert E., et al.
Veröffentlicht: (2024)
von: Wray, Robert E., et al.
Veröffentlicht: (2024)
Robustness of Large Language Models to Perturbations in Text
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
PPLqa: An Unsupervised Information-Theoretic Quality Metric for Comparing Generative Large Language Models
von: Friedland, Gerald, et al.
Veröffentlicht: (2024) -
ValueDCG: Measuring Comprehensive Human Value Understanding Ability of Language Models
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2023) -
The QCET Taxonomy of Standard Quality Criterion Names and Definitions for the Evaluation of NLP Systems
von: Belz, Anya, et al.
Veröffentlicht: (2025) -
Semantic Delta: An Interpretable Signal Differentiating Human and LLMs Dialogue
von: Scantamburlo, Riccardo, et al.
Veröffentlicht: (2026) -
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
von: Wen, Yuqiao, et al.
Veröffentlicht: (2025)