OISD: On-Policy Internal Self-Distillation of Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Xinyu, Jacob, Darryl Cherian, Zhou, Yang, Wang, Jindong, He, Pan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Distilling Out-of-Distribution Robustness from Vision-Language Foundation Models
by: Zhou, Andy, et al.
Published: (2023)
by: Zhou, Andy, et al.
Published: (2023)
Score Distillation of Flow Matching Models
by: Zhou, Mingyuan, et al.
Published: (2025)
by: Zhou, Mingyuan, et al.
Published: (2025)
LLMPhy: Parameter-Identifiable Physical Reasoning Combining Large Language Models and Physics Engines
by: Cherian, Anoop, et al.
Published: (2024)
by: Cherian, Anoop, et al.
Published: (2024)
Temporal Pair Consistency for Variance-Reduced Flow Matching
by: Maduabuchi, Chika, et al.
Published: (2026)
by: Maduabuchi, Chika, et al.
Published: (2026)
Can Multimodal Large Language Models Truly Perform Multimodal In-Context Learning?
by: Chen, Shuo, et al.
Published: (2023)
by: Chen, Shuo, et al.
Published: (2023)
EPSD: Early Pruning with Self-Distillation for Efficient Model Compression
by: Chen, Dong, et al.
Published: (2024)
by: Chen, Dong, et al.
Published: (2024)
VERA: Explainable Video Anomaly Detection via Verbalized Learning of Vision-Language Models
by: Ye, Muchao, et al.
Published: (2024)
by: Ye, Muchao, et al.
Published: (2024)
Towards Long-window Anchoring in Vision-Language Model Distillation
by: Zhou, Haoyi, et al.
Published: (2025)
by: Zhou, Haoyi, et al.
Published: (2025)
Circuit Tracing in Vision-Language Models: Understanding the Internal Mechanisms of Multimodal Thinking
by: Yang, Jingcheng, et al.
Published: (2026)
by: Yang, Jingcheng, et al.
Published: (2026)
Generative Dataset Distillation Based on Self-knowledge Distillation
by: Li, Longzhen, et al.
Published: (2025)
by: Li, Longzhen, et al.
Published: (2025)
Identify, Isolate, and Purge: Mitigating Hallucinations in LVLMs via Self-Evolving Distillation
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
Small Scale Data-Free Knowledge Distillation
by: Liu, He, et al.
Published: (2024)
by: Liu, He, et al.
Published: (2024)
Masking Teacher and Reinforcing Student for Distilling Vision-Language Models
by: Lee, Byung-Kwan, et al.
Published: (2025)
by: Lee, Byung-Kwan, et al.
Published: (2025)
COPRA: Conditional Parameter Adaptation with Reinforcement Learning for Video Anomaly Detection
by: Jacob, Darryl Cherian, et al.
Published: (2026)
by: Jacob, Darryl Cherian, et al.
Published: (2026)
Score identity Distillation: Exponentially Fast Distillation of Pretrained Diffusion Models for One-Step Generation
by: Zhou, Mingyuan, et al.
Published: (2024)
by: Zhou, Mingyuan, et al.
Published: (2024)
Dataset Distillation for Pre-Trained Self-Supervised Vision Models
by: Cazenavette, George, et al.
Published: (2025)
by: Cazenavette, George, et al.
Published: (2025)
Soft Label Pruning and Quantization for Large-Scale Dataset Distillation
by: Lingao, Xiao, et al.
Published: (2026)
by: Lingao, Xiao, et al.
Published: (2026)
Dataset Distillation with Neural Characteristic Function: A Minmax Perspective
by: Wang, Shaobo, et al.
Published: (2025)
by: Wang, Shaobo, et al.
Published: (2025)
UniGame: Turning a Unified Multimodal Model Into Its Own Adversary
by: Su, Zhaolong, et al.
Published: (2025)
by: Su, Zhaolong, et al.
Published: (2025)
Co-GRPO: Co-Optimized Group Relative Policy Optimization for Masked Diffusion Model
by: Zhou, Renping, et al.
Published: (2025)
by: Zhou, Renping, et al.
Published: (2025)
Dataset Distillation via the Wasserstein Metric
by: Liu, Haoyang, et al.
Published: (2023)
by: Liu, Haoyang, et al.
Published: (2023)
Privacy-Preserving Model Transcription with Differentially Private Synthetic Distillation
by: Liu, Bochao, et al.
Published: (2026)
by: Liu, Bochao, et al.
Published: (2026)
Dual-Model Weight Selection and Self-Knowledge Distillation for Medical Image Classification
by: Tsutsumi, Ayaka, et al.
Published: (2025)
by: Tsutsumi, Ayaka, et al.
Published: (2025)
Self-Supervised Quantization-Aware Knowledge Distillation
by: Zhao, Kaiqi, et al.
Published: (2024)
by: Zhao, Kaiqi, et al.
Published: (2024)
Visual Question Decomposition on Multimodal Large Language Models
by: Zhang, Haowei, et al.
Published: (2024)
by: Zhang, Haowei, et al.
Published: (2024)
Distilled Prompt Learning for Incomplete Multimodal Survival Prediction
by: Xu, Yingxue, et al.
Published: (2025)
by: Xu, Yingxue, et al.
Published: (2025)
Diversity-Driven Generative Dataset Distillation Based on Diffusion Model with Self-Adaptive Memory
by: Li, Mingzhuo, et al.
Published: (2025)
by: Li, Mingzhuo, et al.
Published: (2025)
VIAssist: Adapting Multi-modal Large Language Models for Users with Visual Impairments
by: Yang, Bufang, et al.
Published: (2024)
by: Yang, Bufang, et al.
Published: (2024)
Self-Attentive Spatio-Temporal Calibration for Precise Intermediate Layer Matching in ANN-to-SNN Distillation
by: Hong, Di, et al.
Published: (2025)
by: Hong, Di, et al.
Published: (2025)
How Well Does GPT-4V(ision) Adapt to Distribution Shifts? A Preliminary Investigation
by: Han, Zhongyi, et al.
Published: (2023)
by: Han, Zhongyi, et al.
Published: (2023)
EgoAdapt: Adaptive Multisensory Distillation and Policy Learning for Efficient Egocentric Perception
by: Chowdhury, Sanjoy, et al.
Published: (2025)
by: Chowdhury, Sanjoy, et al.
Published: (2025)
pi-Flow: Policy-Based Few-Step Generation via Imitation Distillation
by: Chen, Hansheng, et al.
Published: (2025)
by: Chen, Hansheng, et al.
Published: (2025)
SDXL-Lightning: Progressive Adversarial Diffusion Distillation
by: Lin, Shanchuan, et al.
Published: (2024)
by: Lin, Shanchuan, et al.
Published: (2024)
Masked Autoencoders Are Effective Tokenizers for Diffusion Models
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
Evaluating Large Vision-and-Language Models on Children's Mathematical Olympiads
by: Cherian, Anoop, et al.
Published: (2024)
by: Cherian, Anoop, et al.
Published: (2024)
Vision-Language Meets the Skeleton: Progressively Distillation with Cross-Modal Knowledge for 3D Action Representation Learning
by: Chen, Yang, et al.
Published: (2024)
by: Chen, Yang, et al.
Published: (2024)
Personalized Federated Learning via Backbone Self-Distillation
by: Wang, Pengju, et al.
Published: (2024)
by: Wang, Pengju, et al.
Published: (2024)
FedAFD: Multimodal Federated Learning via Adversarial Fusion and Distillation
by: Tan, Min, et al.
Published: (2026)
by: Tan, Min, et al.
Published: (2026)
Vision-Language Models Can Self-Improve Reasoning via Reflection
by: Cheng, Kanzhi, et al.
Published: (2024)
by: Cheng, Kanzhi, et al.
Published: (2024)
Orchestrate Latent Expertise: Advancing Online Continual Learning with Multi-Level Supervision and Reverse Self-Distillation
by: Yan, HongWei, et al.
Published: (2024)
by: Yan, HongWei, et al.
Published: (2024)
Similar Items
-
Distilling Out-of-Distribution Robustness from Vision-Language Foundation Models
by: Zhou, Andy, et al.
Published: (2023) -
Score Distillation of Flow Matching Models
by: Zhou, Mingyuan, et al.
Published: (2025) -
LLMPhy: Parameter-Identifiable Physical Reasoning Combining Large Language Models and Physics Engines
by: Cherian, Anoop, et al.
Published: (2024) -
Temporal Pair Consistency for Variance-Reduced Flow Matching
by: Maduabuchi, Chika, et al.
Published: (2026) -
Can Multimodal Large Language Models Truly Perform Multimodal In-Context Learning?
by: Chen, Shuo, et al.
Published: (2023)