Simplifying Knowledge Transfer in Pretrained Models
Fuente:
arXiv
Saved in:
| Main Authors: | Jain, Siddharth, Karthik, Shyamgopal, Gandhi, Vineet |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models
by: Eyring, Luca, et al.
Published: (2025)
by: Eyring, Luca, et al.
Published: (2025)
It's Never Too Late: Noise Optimization for Collapse Recovery in Trained Diffusion Models
by: Harrington, Anne, et al.
Published: (2025)
by: Harrington, Anne, et al.
Published: (2025)
Data-Efficient Contrastive Language-Image Pretraining: Prioritizing Data Quality over Quantity
by: Joshi, Siddharth, et al.
Published: (2024)
by: Joshi, Siddharth, et al.
Published: (2024)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
by: Pach, Mateusz, et al.
Published: (2025)
by: Pach, Mateusz, et al.
Published: (2025)
Pseudo-labelling meets Label Smoothing for Noisy Partial Label Learning
by: Saravanan, Darshana, et al.
Published: (2024)
by: Saravanan, Darshana, et al.
Published: (2024)
Solving Spatial Supersensing Without Spatial Supersensing
by: Udandarao, Vishaal, et al.
Published: (2025)
by: Udandarao, Vishaal, et al.
Published: (2025)
Post-hoc Probabilistic Vision-Language Models
by: Baumann, Anton, et al.
Published: (2024)
by: Baumann, Anton, et al.
Published: (2024)
Concept-Guided Interpretability via Neural Chunking
by: Wu, Shuchen, et al.
Published: (2025)
by: Wu, Shuchen, et al.
Published: (2025)
Fantastic Gains and Where to Find Them: On the Existence and Prospect of General Knowledge Transfer between Any Pretrained Model
by: Roth, Karsten, et al.
Published: (2023)
by: Roth, Karsten, et al.
Published: (2023)
PEEKABOO: Interactive Video Generation via Masked-Diffusion
by: Jain, Yash, et al.
Published: (2023)
by: Jain, Yash, et al.
Published: (2023)
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues
by: Girmaji, Rohit, et al.
Published: (2025)
by: Girmaji, Rohit, et al.
Published: (2025)
Towards Knowledge Guided Pretraining Approaches for Multimodal Foundation Models: Applications in Remote Sensing
by: Ravirathinam, Praveen, et al.
Published: (2024)
by: Ravirathinam, Praveen, et al.
Published: (2024)
Efficient Multi-Source Knowledge Transfer by Model Merging
by: Osial, Marcin, et al.
Published: (2025)
by: Osial, Marcin, et al.
Published: (2025)
Simplified Diffusion Schrödinger Bridge
by: Tang, Zhicong, et al.
Published: (2024)
by: Tang, Zhicong, et al.
Published: (2024)
Attention based End to end network for Offline Writer Identification on Word level data
by: Kumar, Vineet, et al.
Published: (2024)
by: Kumar, Vineet, et al.
Published: (2024)
Simplified priors for Object-Centric Learning
by: Patil, Vihang, et al.
Published: (2024)
by: Patil, Vihang, et al.
Published: (2024)
VELOCITI: Benchmarking Video-Language Compositional Reasoning with Strict Entailment
by: Saravanan, Darshana, et al.
Published: (2024)
by: Saravanan, Darshana, et al.
Published: (2024)
Pretrain Where? Investigating How Pretraining Data Diversity Impacts Geospatial Foundation Model Performance
by: Kaur, Amandeep, et al.
Published: (2026)
by: Kaur, Amandeep, et al.
Published: (2026)
Simplifying CLIP: Unleashing the Power of Large-Scale Models on Consumer-level Computers
by: Liu, Hongbo
Published: (2024)
by: Liu, Hongbo
Published: (2024)
Knowledge Transfer from Vision Foundation Models for Efficient Training of Small Task-specific Models
by: Vemulapalli, Raviteja, et al.
Published: (2023)
by: Vemulapalli, Raviteja, et al.
Published: (2023)
MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation
by: Joshi, Siddharth, et al.
Published: (2025)
by: Joshi, Siddharth, et al.
Published: (2025)
Generalizable Blood Cell Detection via Unified Dataset and Faster R-CNN
by: Sahay, Siddharth
Published: (2025)
by: Sahay, Siddharth
Published: (2025)
ReMem: Mutual Information-Aware Fine-tuning of Pretrained Vision Transformers for Effective Knowledge Distillation
by: Dong, Chengyu, et al.
Published: (2025)
by: Dong, Chengyu, et al.
Published: (2025)
Pretrained Diffusion Models Are Inherently Skipped-Step Samplers
by: Xu, Wenju
Published: (2025)
by: Xu, Wenju
Published: (2025)
Pretrained Image-Text Models are Secretly Video Captioners
by: Zhang, Chunhui, et al.
Published: (2025)
by: Zhang, Chunhui, et al.
Published: (2025)
A Fast and Efficient Modern BERT based Text-Conditioned Diffusion Model for Medical Image Segmentation
by: Dhara, Venkata Siddharth, et al.
Published: (2025)
by: Dhara, Venkata Siddharth, et al.
Published: (2025)
Foundation Model-oriented Robustness: Robust Image Model Evaluation with Pretrained Models
by: Zhang, Peiyan, et al.
Published: (2023)
by: Zhang, Peiyan, et al.
Published: (2023)
Vision-by-Language for Training-Free Compositional Image Retrieval
by: Karthik, Shyamgopal, et al.
Published: (2023)
by: Karthik, Shyamgopal, et al.
Published: (2023)
Disentangled World Models: Learning to Transfer Semantic Knowledge from Distracting Videos for Reinforcement Learning
by: Wang, Qi, et al.
Published: (2025)
by: Wang, Qi, et al.
Published: (2025)
Diff-Instruct: A Universal Approach for Transferring Knowledge From Pre-trained Diffusion Models
by: Luo, Weijian, et al.
Published: (2023)
by: Luo, Weijian, et al.
Published: (2023)
Pretrained Visual Uncertainties
by: Kirchhof, Michael, et al.
Published: (2024)
by: Kirchhof, Michael, et al.
Published: (2024)
SimpliHuMoN: Simplifying Human Motion Prediction
by: Agrawal, Aadya, et al.
Published: (2026)
by: Agrawal, Aadya, et al.
Published: (2026)
Unsupervised Video Summarization via Iterative Training and Simplified GAN
by: Li, Hanqing, et al.
Published: (2023)
by: Li, Hanqing, et al.
Published: (2023)
Reflecting on the State of Rehearsal-free Continual Learning with Pretrained Models
by: Thede, Lukas, et al.
Published: (2024)
by: Thede, Lukas, et al.
Published: (2024)
GD doesn't make the cut: Three ways that non-differentiability affects neural network training
by: Kumar, Siddharth Krishna
Published: (2024)
by: Kumar, Siddharth Krishna
Published: (2024)
LLVD: LSTM-based Explicit Motion Modeling in Latent Space for Blind Video Denoising
by: Rashid, Loay, et al.
Published: (2025)
by: Rashid, Loay, et al.
Published: (2025)
Continual Learning of Achieving Forgetting-free and Positive Knowledge Transfer
by: Wang, Zhi, et al.
Published: (2026)
by: Wang, Zhi, et al.
Published: (2026)
Unified Multimodal Discrete Diffusion
by: Swerdlow, Alexander, et al.
Published: (2025)
by: Swerdlow, Alexander, et al.
Published: (2025)
ReNO: Enhancing One-step Text-to-Image Models through Reward-based Noise Optimization
by: Eyring, Luca, et al.
Published: (2024)
by: Eyring, Luca, et al.
Published: (2024)
Impact of Pretraining Word Co-occurrence on Compositional Generalization in Multimodal Models
by: Qu, Helen, et al.
Published: (2025)
by: Qu, Helen, et al.
Published: (2025)
Similar Items
-
Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models
by: Eyring, Luca, et al.
Published: (2025) -
It's Never Too Late: Noise Optimization for Collapse Recovery in Trained Diffusion Models
by: Harrington, Anne, et al.
Published: (2025) -
Data-Efficient Contrastive Language-Image Pretraining: Prioritizing Data Quality over Quantity
by: Joshi, Siddharth, et al.
Published: (2024) -
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
by: Pach, Mateusz, et al.
Published: (2025) -
Pseudo-labelling meets Label Smoothing for Noisy Partial Label Learning
by: Saravanan, Darshana, et al.
Published: (2024)