KIND: Knowledge Integration and Diversion for Training Decomposable Models
Fuente:
arXiv
Saved in:
| Main Authors: | Xie, Yucheng, Feng, Fu, Shi, Ruixiao, Wang, Jing, Rui, Yong, Geng, Xin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DivControl: Knowledge Diversion for Controllable Image Generation
by: Xie, Yucheng, et al.
Published: (2025)
by: Xie, Yucheng, et al.
Published: (2025)
FINE: Factorizing Knowledge for Initialization of Variable-sized Diffusion Models
by: Xie, Yucheng, et al.
Published: (2024)
by: Xie, Yucheng, et al.
Published: (2024)
FAD: Frequency Adaptation and Diversion for Cross-domain Few-shot Learning
by: Shi, Ruixiao, et al.
Published: (2025)
by: Shi, Ruixiao, et al.
Published: (2025)
Self-Supervised Weight Templates for Scalable Vision Model Initialization
by: Xie, Yucheng, et al.
Published: (2026)
by: Xie, Yucheng, et al.
Published: (2026)
A Creative Agent is Worth a 64-Token Template
by: Shi, Ruixiao, et al.
Published: (2026)
by: Shi, Ruixiao, et al.
Published: (2026)
Distribution-Conditional Generation: From Class Distribution to Creative Generation
by: Feng, Fu, et al.
Published: (2025)
by: Feng, Fu, et al.
Published: (2025)
Redefining <Creative> in Dictionary: Towards an Enhanced Semantic Understanding of Creative Generation
by: Feng, Fu, et al.
Published: (2024)
by: Feng, Fu, et al.
Published: (2024)
Enriching Knowledge Distillation with Intra-Class Contrastive Learning
by: Yuan, Hua, et al.
Published: (2025)
by: Yuan, Hua, et al.
Published: (2025)
Equivariant Image Modeling
by: Dong, Ruixiao, et al.
Published: (2025)
by: Dong, Ruixiao, et al.
Published: (2025)
Error-Decomposed Class-Conditional Fusion for Statistically Guaranteed Hard-Category Robust Perception
by: Luo, Guowei, et al.
Published: (2026)
by: Luo, Guowei, et al.
Published: (2026)
Frequency-Guided Diffusion Model with Perturbation Training for Skeleton-Based Video Anomaly Detection
by: Tan, Xiaofeng, et al.
Published: (2024)
by: Tan, Xiaofeng, et al.
Published: (2024)
IdealGPT: Iteratively Decomposing Vision and Language Reasoning via Large Language Models
by: You, Haoxuan, et al.
Published: (2023)
by: You, Haoxuan, et al.
Published: (2023)
Extracting Multimodal Learngene in CLIP: Unveiling the Multimodal Generalizable Knowledge
by: Chen, Ruiming, et al.
Published: (2025)
by: Chen, Ruiming, et al.
Published: (2025)
Decompose and Leverage Preferences from Expert Models for Improving Trustworthiness of MLLMs
by: Cao, Rui, et al.
Published: (2024)
by: Cao, Rui, et al.
Published: (2024)
Spectral-Structured Diffusion for Single-Image Rain Removal
by: Xing, Yucheng, et al.
Published: (2026)
by: Xing, Yucheng, et al.
Published: (2026)
Extending Large Vision-Language Model for Diverse Interactive Tasks in Autonomous Driving
by: Zhao, Zongchuang, et al.
Published: (2025)
by: Zhao, Zongchuang, et al.
Published: (2025)
DCFormer: Efficient 3D Vision-Language Modeling with Decomposed Convolutions
by: Ates, Gorkem Can, et al.
Published: (2025)
by: Ates, Gorkem Can, et al.
Published: (2025)
The Victim and The Beneficiary: Exploiting a Poisoned Model to Train a Clean Model on Poisoned Data
by: Zhu, Zixuan, et al.
Published: (2024)
by: Zhu, Zixuan, et al.
Published: (2024)
WinTok: A Win-Win Hybrid Tokenizer via Decomposing Visual Understanding and Generation with Transferable Tokens
by: Guo, Yiwei, et al.
Published: (2026)
by: Guo, Yiwei, et al.
Published: (2026)
FDDM: Frequency-Decomposed Diffusion Model for Rectum Cancer Dose Prediction in Radiotherapy
by: Liao, Xin, et al.
Published: (2024)
by: Liao, Xin, et al.
Published: (2024)
Graph Attention Transformer Network for Multi-Label Image Classification
by: Yuan, Jin, et al.
Published: (2022)
by: Yuan, Jin, et al.
Published: (2022)
Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation
by: Han, Su Ho, et al.
Published: (2025)
by: Han, Su Ho, et al.
Published: (2025)
Enhancing Motion in Text-to-Video Generation with Decomposed Encoding and Conditioning
by: Ruan, Penghui, et al.
Published: (2024)
by: Ruan, Penghui, et al.
Published: (2024)
COMUNI: Decomposing Common and Unique Video Signals for Diffusion-based Video Generation
by: Sun, Mingzhen, et al.
Published: (2024)
by: Sun, Mingzhen, et al.
Published: (2024)
DyMO: Training-Free Diffusion Model Alignment with Dynamic Multi-Objective Scheduling
by: Xie, Xin, et al.
Published: (2024)
by: Xie, Xin, et al.
Published: (2024)
DecoFuse: Decomposing and Fusing the "What", "Where", and "How" for Brain-Inspired fMRI-to-Video Decoding
by: Li, Chong, et al.
Published: (2025)
by: Li, Chong, et al.
Published: (2025)
Exploring Diverse In-Context Configurations for Image Captioning
by: Yang, Xu, et al.
Published: (2023)
by: Yang, Xu, et al.
Published: (2023)
EMMA: Your Text-to-Image Diffusion Model Can Secretly Accept Multi-Modal Prompts
by: Han, Yucheng, et al.
Published: (2024)
by: Han, Yucheng, et al.
Published: (2024)
Affective Behaviour Analysis via Integrating Multi-Modal Knowledge
by: Zhang, Wei, et al.
Published: (2024)
by: Zhang, Wei, et al.
Published: (2024)
VipDiff: Towards Coherent and Diverse Video Inpainting via Training-free Denoising Diffusion Models
by: Xie, Chaohao, et al.
Published: (2025)
by: Xie, Chaohao, et al.
Published: (2025)
ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference
by: Lan, Mengcheng, et al.
Published: (2024)
by: Lan, Mengcheng, et al.
Published: (2024)
Exploring Partial Multi-Label Learning via Integrating Semantic Co-occurrence Knowledge
by: Wu, Xin, et al.
Published: (2025)
by: Wu, Xin, et al.
Published: (2025)
Map-Free Trajectory Prediction with Map Distillation and Hierarchical Encoding
by: Liu, Xiaodong, et al.
Published: (2024)
by: Liu, Xiaodong, et al.
Published: (2024)
Leveraging Structure Knowledge and Deep Models for the Detection of Abnormal Handwritten Text
by: Wang, Zi-Rui
Published: (2024)
by: Wang, Zi-Rui
Published: (2024)
Mask Consistency Regularization in Object Removal
by: Yuan, Hua, et al.
Published: (2025)
by: Yuan, Hua, et al.
Published: (2025)
UKnow: A Unified Knowledge Protocol with Multimodal Knowledge Graph Datasets for Reasoning and Vision-Language Pre-Training
by: Gong, Biao, et al.
Published: (2023)
by: Gong, Biao, et al.
Published: (2023)
StdGEN++: A Comprehensive System for Semantic-Decomposed 3D Character Generation
by: He, Yuze, et al.
Published: (2026)
by: He, Yuze, et al.
Published: (2026)
SRA 2: Variational Autoencoder Self-Representation Alignment for Efficient Diffusion Training
by: Wang, Mengmeng, et al.
Published: (2026)
by: Wang, Mengmeng, et al.
Published: (2026)
N-Tree Diffusion for Long-Horizon Wildfire Risk Forecasting
by: Xing, Yucheng, et al.
Published: (2026)
by: Xing, Yucheng, et al.
Published: (2026)
ExpertGen: Training-Free Expert Guidance for Controllable Text-to-Face Generation
by: Shi, Liang, et al.
Published: (2025)
by: Shi, Liang, et al.
Published: (2025)
Similar Items
-
DivControl: Knowledge Diversion for Controllable Image Generation
by: Xie, Yucheng, et al.
Published: (2025) -
FINE: Factorizing Knowledge for Initialization of Variable-sized Diffusion Models
by: Xie, Yucheng, et al.
Published: (2024) -
FAD: Frequency Adaptation and Diversion for Cross-domain Few-shot Learning
by: Shi, Ruixiao, et al.
Published: (2025) -
Self-Supervised Weight Templates for Scalable Vision Model Initialization
by: Xie, Yucheng, et al.
Published: (2026) -
A Creative Agent is Worth a 64-Token Template
by: Shi, Ruixiao, et al.
Published: (2026)