Teaching the Teacher: The Role of Teacher-Student Smoothness Alignment in Genetic Programming-based Symbolic Distillation
Fuente:
arXiv
Guardado en:
| Autores principales: | Dhar, Soumyadeep, Fong, Kei Sen, Motani, Mehul |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
From Teacher to Student: Tracking Memorization Through Model Distillation
por: Singh, Simardeep
Publicado: (2025)
por: Singh, Simardeep
Publicado: (2025)
Toward Student-Oriented Teacher Network Training For Knowledge Distillation
por: Dong, Chengyu, et al.
Publicado: (2022)
por: Dong, Chengyu, et al.
Publicado: (2022)
Generalizing Teacher Networks for Effective Knowledge Distillation Across Student Architectures
por: Binici, Kuluhan, et al.
Publicado: (2024)
por: Binici, Kuluhan, et al.
Publicado: (2024)
The Role of Teacher Calibration in Knowledge Distillation
por: Kim, Suyoung, et al.
Publicado: (2025)
por: Kim, Suyoung, et al.
Publicado: (2025)
Multi-Teacher Knowledge Distillation via Teacher-Informed Mixture Priors
por: Fang, Luyang, et al.
Publicado: (2026)
por: Fang, Luyang, et al.
Publicado: (2026)
Model Merging via Multi-Teacher Knowledge Distillation
por: Dalili, Seyed Arshan, et al.
Publicado: (2025)
por: Dalili, Seyed Arshan, et al.
Publicado: (2025)
Linear Projections of Teacher Embeddings for Few-Class Distillation
por: Loo, Noel, et al.
Publicado: (2024)
por: Loo, Noel, et al.
Publicado: (2024)
Masking Teacher and Reinforcing Student for Distilling Vision-Language Models
por: Lee, Byung-Kwan, et al.
Publicado: (2025)
por: Lee, Byung-Kwan, et al.
Publicado: (2025)
On Teacher Hacking in Language Model Distillation
por: Tiapkin, Daniil, et al.
Publicado: (2025)
por: Tiapkin, Daniil, et al.
Publicado: (2025)
Teacher-Student Learning on Complexity in Intelligent Routing
por: Pi, Shu-Ting, et al.
Publicado: (2024)
por: Pi, Shu-Ting, et al.
Publicado: (2024)
DUET: Distilled LLM Unlearning from an Efficiently Contextualized Teacher
por: Zhong, Yisheng, et al.
Publicado: (2026)
por: Zhong, Yisheng, et al.
Publicado: (2026)
A Comparison of Recent Algorithms for Symbolic Regression to Genetic Programming
por: Radwan, Yousef A., et al.
Publicado: (2024)
por: Radwan, Yousef A., et al.
Publicado: (2024)
The Strongest Teacher Is Not Always the Best Teacher: Student-Centric Answer Selection
por: Hu, Zhengyu, et al.
Publicado: (2026)
por: Hu, Zhengyu, et al.
Publicado: (2026)
Heuristic Methods are Good Teachers to Distill MLPs for Graph Link Prediction
por: Qin, Zongyue, et al.
Publicado: (2025)
por: Qin, Zongyue, et al.
Publicado: (2025)
Group Relative Knowledge Distillation: Learning from Teacher's Relational Inductive Bias
por: Li, Chao, et al.
Publicado: (2025)
por: Li, Chao, et al.
Publicado: (2025)
Robust Knowledge Distillation Based on Feature Variance Against Backdoored Teacher Model
por: Chen, Jinyin, et al.
Publicado: (2024)
por: Chen, Jinyin, et al.
Publicado: (2024)
When Are Teacher Tokens Reliable? Position-Weighted On-Policy Self-Distillation for Reasoning
por: Liu, Xiaogeng, et al.
Publicado: (2026)
por: Liu, Xiaogeng, et al.
Publicado: (2026)
CLIP-Embed-KD: Computationally Efficient Knowledge Distillation Using Embeddings as Teachers
por: Nair, Lakshmi
Publicado: (2024)
por: Nair, Lakshmi
Publicado: (2024)
Teacher as a Lenient Expert: Teacher-Agnostic Data-Free Knowledge Distillation
por: Shin, Hyunjune, et al.
Publicado: (2024)
por: Shin, Hyunjune, et al.
Publicado: (2024)
Knowledge Distillation in Wide Neural Networks: Risk Bound, Data Efficiency and Imperfect Teacher
por: Ji, Guangda, et al.
Publicado: (2020)
por: Ji, Guangda, et al.
Publicado: (2020)
Efficient and Robust Knowledge Distillation from A Stronger Teacher Based on Correlation Matching
por: Niu, Wenqi, et al.
Publicado: (2024)
por: Niu, Wenqi, et al.
Publicado: (2024)
Do Students Debias Like Teachers? On the Distillability of Bias Mitigation Methods
por: Cheng, Jiali, et al.
Publicado: (2025)
por: Cheng, Jiali, et al.
Publicado: (2025)
Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation
por: Yang, Wenkai, et al.
Publicado: (2026)
por: Yang, Wenkai, et al.
Publicado: (2026)
Tab-PET: Graph-Based Positional Encodings for Tabular Transformers
por: Leng, Yunze, et al.
Publicado: (2025)
por: Leng, Yunze, et al.
Publicado: (2025)
YODA: Teacher-Student Progressive Learning for Language Models
por: Lu, Jianqiao, et al.
Publicado: (2024)
por: Lu, Jianqiao, et al.
Publicado: (2024)
The Collaboration Paradox: Why Generative AI Requires Both Strategic Intelligence and Operational Stability in Supply Chain Management
por: Dhar, Soumyadeep
Publicado: (2025)
por: Dhar, Soumyadeep
Publicado: (2025)
Prompt-responsive Object Retrieval with Memory-augmented Student-Teacher Learning
por: Mosbach, Malte, et al.
Publicado: (2025)
por: Mosbach, Malte, et al.
Publicado: (2025)
Good Teachers Explain: Explanation-Enhanced Knowledge Distillation
por: Parchami-Araghi, Amin, et al.
Publicado: (2024)
por: Parchami-Araghi, Amin, et al.
Publicado: (2024)
FedMTFI: Feature Importance Based Optimized Multi Teacher Knowledge Distillation in Heterogeneous Federated Learning Environment
por: Shadin, Nazmus Shakib, et al.
Publicado: (2026)
por: Shadin, Nazmus Shakib, et al.
Publicado: (2026)
Multi-Surrogate-Teacher Assistance for Representation Alignment in Fingerprint-based Indoor Localization
por: Nguyen, Son Minh, et al.
Publicado: (2024)
por: Nguyen, Son Minh, et al.
Publicado: (2024)
Quantifying Symptom Causality in Clinical Decision Making: An Exploration Using CausaLM
por: Shetty, Mehul, et al.
Publicado: (2025)
por: Shetty, Mehul, et al.
Publicado: (2025)
Prioritize Alignment in Dataset Distillation
por: Li, Zekai, et al.
Publicado: (2024)
por: Li, Zekai, et al.
Publicado: (2024)
Distilling Symbolic Priors for Concept Learning into Neural Networks
por: Marinescu, Ioana, et al.
Publicado: (2024)
por: Marinescu, Ioana, et al.
Publicado: (2024)
Trajectory as the Teacher: Few-Step Discrete Flow Matching via Energy-Navigated Distillation
por: Monsefi, Amin Karimi, et al.
Publicado: (2026)
por: Monsefi, Amin Karimi, et al.
Publicado: (2026)
AVSD: Adaptive-View Self-Distillation by Balancing Consensus and Teacher-Specific Privileged Signals
por: Nguyen, Duy, et al.
Publicado: (2026)
por: Nguyen, Duy, et al.
Publicado: (2026)
PHI-S: Distribution Balancing for Label-Free Multi-Teacher Distillation
por: Ranzinger, Mike, et al.
Publicado: (2024)
por: Ranzinger, Mike, et al.
Publicado: (2024)
Offline Behavior Distillation
por: Lei, Shiye, et al.
Publicado: (2024)
por: Lei, Shiye, et al.
Publicado: (2024)
Data-Efficient Symbolic Regression via Foundation Model Distillation
por: Ying, Wangyang, et al.
Publicado: (2025)
por: Ying, Wangyang, et al.
Publicado: (2025)
Beyond Answers: Transferring Reasoning Capabilities to Smaller LLMs Using Multi-Teacher Knowledge Distillation
por: Tian, Yijun, et al.
Publicado: (2024)
por: Tian, Yijun, et al.
Publicado: (2024)
DataEnvGym: Data Generation Agents in Teacher Environments with Student Feedback
por: Khan, Zaid, et al.
Publicado: (2024)
por: Khan, Zaid, et al.
Publicado: (2024)
Ejemplares similares
-
From Teacher to Student: Tracking Memorization Through Model Distillation
por: Singh, Simardeep
Publicado: (2025) -
Toward Student-Oriented Teacher Network Training For Knowledge Distillation
por: Dong, Chengyu, et al.
Publicado: (2022) -
Generalizing Teacher Networks for Effective Knowledge Distillation Across Student Architectures
por: Binici, Kuluhan, et al.
Publicado: (2024) -
The Role of Teacher Calibration in Knowledge Distillation
por: Kim, Suyoung, et al.
Publicado: (2025) -
Multi-Teacher Knowledge Distillation via Teacher-Informed Mixture Priors
por: Fang, Luyang, et al.
Publicado: (2026)