Guardado en:
| Autor principal: | Park, Sejik |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2403.19941 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ResidualDroppath: Enhancing Feature Reuse over Residual Connections
por: Park, Sejik
Publicado: (2024)
por: Park, Sejik
Publicado: (2024)
Learning More Generalized Experts by Merging Experts in Mixture-of-Experts
por: Park, Sejik
Publicado: (2024)
por: Park, Sejik
Publicado: (2024)
Self-Normalized Resets for Plasticity in Continual Learning
por: Farias, Vivek F., et al.
Publicado: (2024)
por: Farias, Vivek F., et al.
Publicado: (2024)
Retro: Reusing teacher projection head for efficient embedding distillation on Lightweight Models via Self-supervised Learning
por: Nguyen, Khanh-Binh, et al.
Publicado: (2024)
por: Nguyen, Khanh-Binh, et al.
Publicado: (2024)
Reset-free Reinforcement Learning with World Models
por: Yang, Zhao, et al.
Publicado: (2024)
por: Yang, Zhao, et al.
Publicado: (2024)
The Power of Resets in Online Reinforcement Learning
por: Mhammedi, Zakaria, et al.
Publicado: (2024)
por: Mhammedi, Zakaria, et al.
Publicado: (2024)
Single-Reset Divide & Conquer Imitation Learning
por: Chenu, Alexandre, et al.
Publicado: (2024)
por: Chenu, Alexandre, et al.
Publicado: (2024)
Self-supervised network distillation: an effective approach to exploration in sparse reward environments
por: Pecháč, Matej, et al.
Publicado: (2023)
por: Pecháč, Matej, et al.
Publicado: (2023)
SPARE: Self-distillation for PARameter-Efficient Removal
por: Mola, Natnael, et al.
Publicado: (2026)
por: Mola, Natnael, et al.
Publicado: (2026)
Smooth-Distill: A Self-distillation Framework for Multitask Learning with Wearable Sensor Data
por: Vu, Hoang-Dieu, et al.
Publicado: (2025)
por: Vu, Hoang-Dieu, et al.
Publicado: (2025)
Credit Assignment with Resets in Language Model Reasoning
por: Samanta, Ankur, et al.
Publicado: (2026)
por: Samanta, Ankur, et al.
Publicado: (2026)
Intelligent Switching for Reset-Free RL
por: Patil, Darshan, et al.
Publicado: (2024)
por: Patil, Darshan, et al.
Publicado: (2024)
Non-Stationary Learning of Neural Networks with Automatic Soft Parameter Reset
por: Galashov, Alexandre, et al.
Publicado: (2024)
por: Galashov, Alexandre, et al.
Publicado: (2024)
sDREAMER: Self-distilled Mixture-of-Modality-Experts Transformer for Automatic Sleep Staging
por: Chen, Jingyuan, et al.
Publicado: (2025)
por: Chen, Jingyuan, et al.
Publicado: (2025)
Reset & Distill: A Recipe for Overcoming Negative Transfer in Continual Reinforcement Learning
por: Ahn, Hongjoon, et al.
Publicado: (2024)
por: Ahn, Hongjoon, et al.
Publicado: (2024)
Dataset Reset Policy Optimization for RLHF
por: Chang, Jonathan D., et al.
Publicado: (2024)
por: Chang, Jonathan D., et al.
Publicado: (2024)
Towards Lightweight and Stable Zero-shot TTS with Self-distilled Representation Disentanglement
por: Chen, Qianniu, et al.
Publicado: (2025)
por: Chen, Qianniu, et al.
Publicado: (2025)
Education distillation:getting student models to learn in shcools
por: Feng, Ling, et al.
Publicado: (2023)
por: Feng, Ling, et al.
Publicado: (2023)
A Reinforcement Learning based Reset Policy for CDCL SAT Solvers
por: Li, Chunxiao, et al.
Publicado: (2024)
por: Li, Chunxiao, et al.
Publicado: (2024)
Metis-SPECS: Decoupling Multimodal Learning via Self-distilled Preference-based Cold Start
por: Chen, Kun, et al.
Publicado: (2025)
por: Chen, Kun, et al.
Publicado: (2025)
IDLM: Inverse-distilled Diffusion Language Models
por: Li, David, et al.
Publicado: (2026)
por: Li, David, et al.
Publicado: (2026)
Self-Knowledge Distillation for Learning Ambiguity
por: Park, Hancheol, et al.
Publicado: (2024)
por: Park, Hancheol, et al.
Publicado: (2024)
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation
por: Huang, Feizhen, et al.
Publicado: (2025)
por: Huang, Feizhen, et al.
Publicado: (2025)
OneSearch-V2: The Latent Reasoning Enhanced Self-distillation Generative Search Framework
por: Chen, Ben, et al.
Publicado: (2026)
por: Chen, Ben, et al.
Publicado: (2026)
Learning Diverse Policies with Soft Self-Generated Guidance
por: Wang, Guojian, et al.
Publicado: (2024)
por: Wang, Guojian, et al.
Publicado: (2024)
On student-teacher deviations in distillation: does it pay to disobey?
por: Nagarajan, Vaishnavh, et al.
Publicado: (2023)
por: Nagarajan, Vaishnavh, et al.
Publicado: (2023)
Knowledge distillation through geometry-aware representational alignment
por: Bhattarai, Prajjwal, et al.
Publicado: (2025)
por: Bhattarai, Prajjwal, et al.
Publicado: (2025)
Quality-Diversity Actor-Critic: Learning High-Performing and Diverse Behaviors via Value and Successor Features Critics
por: Grillotti, Luca, et al.
Publicado: (2024)
por: Grillotti, Luca, et al.
Publicado: (2024)
VendiRL: A Framework for Self-Supervised Reinforcement Learning of Diversely Diverse Skills
por: Lintunen, Erik M.
Publicado: (2025)
por: Lintunen, Erik M.
Publicado: (2025)
Towards a theory of model distillation
por: Boix-Adsera, Enric
Publicado: (2024)
por: Boix-Adsera, Enric
Publicado: (2024)
VertiBench: Advancing Feature Distribution Diversity in Vertical Federated Learning Benchmarks
por: Wu, Zhaomin, et al.
Publicado: (2023)
por: Wu, Zhaomin, et al.
Publicado: (2023)
DSLR: Diversity Enhancement and Structure Learning for Rehearsal-based Graph Continual Learning
por: Choi, Seungyoon, et al.
Publicado: (2024)
por: Choi, Seungyoon, et al.
Publicado: (2024)
TabKD: Tabular Knowledge Distillation through Interaction Diversity of Learned Feature Bins
por: Pereira, Shovon Niverd, et al.
Publicado: (2026)
por: Pereira, Shovon Niverd, et al.
Publicado: (2026)
Never Reset Again: A Mathematical Framework for Continual Inference in Recurrent Neural Networks
por: Yin, Bojian, et al.
Publicado: (2024)
por: Yin, Bojian, et al.
Publicado: (2024)
SPAM: Spike-Aware Adam with Momentum Reset for Stable LLM Training
por: Huang, Tianjin, et al.
Publicado: (2025)
por: Huang, Tianjin, et al.
Publicado: (2025)
Trust the uncertain teacher: distilling dark knowledge via calibrated uncertainty
por: Kim, Jeonghyun, et al.
Publicado: (2026)
por: Kim, Jeonghyun, et al.
Publicado: (2026)
ADMEDTAGGER: an annotation framework for distillation of expert knowledge for the Polish medical language
por: Górski, Franciszek, et al.
Publicado: (2025)
por: Górski, Franciszek, et al.
Publicado: (2025)
On Pretraining Data Diversity for Self-Supervised Learning
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2024)
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2024)
Tabular Feature Discovery With Reasoning Type Exploration
por: Han, Sungwon, et al.
Publicado: (2025)
por: Han, Sungwon, et al.
Publicado: (2025)
Stochastic Resetting Mitigates Latent Gradient Bias of SGD from Label Noise
por: Bae, Youngkyoung, et al.
Publicado: (2024)
por: Bae, Youngkyoung, et al.
Publicado: (2024)
Ejemplares similares
-
ResidualDroppath: Enhancing Feature Reuse over Residual Connections
por: Park, Sejik
Publicado: (2024) -
Learning More Generalized Experts by Merging Experts in Mixture-of-Experts
por: Park, Sejik
Publicado: (2024) -
Self-Normalized Resets for Plasticity in Continual Learning
por: Farias, Vivek F., et al.
Publicado: (2024) -
Retro: Reusing teacher projection head for efficient embedding distillation on Lightweight Models via Self-supervised Learning
por: Nguyen, Khanh-Binh, et al.
Publicado: (2024) -
Reset-free Reinforcement Learning with World Models
por: Yang, Zhao, et al.
Publicado: (2024)