Exploring Learngene via Stage-wise Weight Sharing for Initializing Variable-sized Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xia, Shi-Yu, Zhu, Wenxuan, Yang, Xu, Geng, Xin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Extracting Multimodal Learngene in CLIP: Unveiling the Multimodal Generalizable Knowledge
von: Chen, Ruiming, et al.
Veröffentlicht: (2025)
von: Chen, Ruiming, et al.
Veröffentlicht: (2025)
Safe Vision-Language Models via Unsafe Weights Manipulation
von: D'Incà, Moreno, et al.
Veröffentlicht: (2025)
von: D'Incà, Moreno, et al.
Veröffentlicht: (2025)
LarvSeg: Exploring Image Classification Data For Large Vocabulary Semantic Segmentation via Category-wise Attentive Classifier
von: Yu, Haojun, et al.
Veröffentlicht: (2025)
von: Yu, Haojun, et al.
Veröffentlicht: (2025)
ShareVerse: Multi-Agent Consistent Video Generation for Shared World Modeling
von: Zhu, Jiayi, et al.
Veröffentlicht: (2026)
von: Zhu, Jiayi, et al.
Veröffentlicht: (2026)
FINE: Factorizing Knowledge for Initialization of Variable-sized Diffusion Models
von: Xie, Yucheng, et al.
Veröffentlicht: (2024)
von: Xie, Yucheng, et al.
Veröffentlicht: (2024)
Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization under Distribution Shift
von: Xu, Qinwu
Veröffentlicht: (2026)
von: Xu, Qinwu
Veröffentlicht: (2026)
Channel-wise Vector Quantization
von: Song, Wei, et al.
Veröffentlicht: (2026)
von: Song, Wei, et al.
Veröffentlicht: (2026)
An Effective Weight Initialization Method for Deep Learning: Application to Satellite Image Classification
von: Boulila, Wadii, et al.
Veröffentlicht: (2024)
von: Boulila, Wadii, et al.
Veröffentlicht: (2024)
Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion
von: Zhang, Yinghui, et al.
Veröffentlicht: (2025)
von: Zhang, Yinghui, et al.
Veröffentlicht: (2025)
Enhanced Structured Lasso Pruning with Class-wise Information
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling
von: Wang, Yuran, et al.
Veröffentlicht: (2025)
von: Wang, Yuran, et al.
Veröffentlicht: (2025)
DP-MDM: Detail-Preserving MR Reconstruction via Multiple Diffusion Models
von: Geng, Mengxiao, et al.
Veröffentlicht: (2024)
von: Geng, Mengxiao, et al.
Veröffentlicht: (2024)
Enhancing Multimodal In-Context Learning for Image Classification through Coreset Optimization
von: Chen, Huiyi, et al.
Veröffentlicht: (2025)
von: Chen, Huiyi, et al.
Veröffentlicht: (2025)
ConsNoTrainLoRA: Data-driven Weight Initialization of Low-rank Adapters using Constraints
von: Das, Debasmit, et al.
Veröffentlicht: (2025)
von: Das, Debasmit, et al.
Veröffentlicht: (2025)
Exploring Partial Multi-Label Learning via Integrating Semantic Co-occurrence Knowledge
von: Wu, Xin, et al.
Veröffentlicht: (2025)
von: Wu, Xin, et al.
Veröffentlicht: (2025)
Semantic Communication based on Large Language Model for Underwater Image Transmission
von: Chen, Weilong, et al.
Veröffentlicht: (2024)
von: Chen, Weilong, et al.
Veröffentlicht: (2024)
CLASP: Class-Adaptive Layer Fusion and Dual-Stage Pruning for Multimodal Large Language Models
von: Dang, Yunkai, et al.
Veröffentlicht: (2026)
von: Dang, Yunkai, et al.
Veröffentlicht: (2026)
Compound Expression Recognition via Multi Model Ensemble
von: Yu, Jun, et al.
Veröffentlicht: (2024)
von: Yu, Jun, et al.
Veröffentlicht: (2024)
StageDesigner: Artistic Stage Generation for Scenography via Theater Scripts
von: Gan, Zhaoxing, et al.
Veröffentlicht: (2025)
von: Gan, Zhaoxing, et al.
Veröffentlicht: (2025)
NeurIPS: Neuro-anatomical Inductive Priors for Sphere-based Brain Decoding
von: Yu, Sijin, et al.
Veröffentlicht: (2026)
von: Yu, Sijin, et al.
Veröffentlicht: (2026)
Variable-frame CNNLSTM for Breast Nodule Classification using Ultrasound Videos
von: Cui, Xiangxiang, et al.
Veröffentlicht: (2025)
von: Cui, Xiangxiang, et al.
Veröffentlicht: (2025)
Block-wise LoRA: Revisiting Fine-grained LoRA for Effective Personalization and Stylization in Text-to-Image Generation
von: Li, Likun, et al.
Veröffentlicht: (2024)
von: Li, Likun, et al.
Veröffentlicht: (2024)
BatStyler: Advancing Multi-category Style Generation for Source-free Domain Generalization
von: Xu, Xiusheng, et al.
Veröffentlicht: (2025)
von: Xu, Xiusheng, et al.
Veröffentlicht: (2025)
Not All Pixels Are Equal: Pixel-wise Meta-Learning for Medical Segmentation with Noisy Labels
von: Mu, Chenyu, et al.
Veröffentlicht: (2025)
von: Mu, Chenyu, et al.
Veröffentlicht: (2025)
DRIVE: Dual-Robustness via Information Variability and Entropic Consistency in Source-Free Unsupervised Domain Adaptation
von: Xiao, Ruiqiang, et al.
Veröffentlicht: (2024)
von: Xiao, Ruiqiang, et al.
Veröffentlicht: (2024)
CSTalk: Correlation Supervised Speech-driven 3D Emotional Facial Animation Generation
von: Liang, Xiangyu, et al.
Veröffentlicht: (2024)
von: Liang, Xiangyu, et al.
Veröffentlicht: (2024)
FAD: Frequency Adaptation and Diversion for Cross-domain Few-shot Learning
von: Shi, Ruixiao, et al.
Veröffentlicht: (2025)
von: Shi, Ruixiao, et al.
Veröffentlicht: (2025)
OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation
von: Zhang, Guohui, et al.
Veröffentlicht: (2026)
von: Zhang, Guohui, et al.
Veröffentlicht: (2026)
DST-Net: A Dual-Stream Transformer with Illumination-Independent Feature Guidance and Multi-Scale Spatial Convolution for Low-Light Image Enhancement
von: Shi, Yicui, et al.
Veröffentlicht: (2026)
von: Shi, Yicui, et al.
Veröffentlicht: (2026)
GBSD: Generative Bokeh with Stage Diffusion
von: Deng, Jieren, et al.
Veröffentlicht: (2023)
von: Deng, Jieren, et al.
Veröffentlicht: (2023)
Nexus-Gen: Unified Image Understanding, Generation, and Editing via Prefilled Autoregression in Shared Embedding Space
von: Zhang, Hong, et al.
Veröffentlicht: (2025)
von: Zhang, Hong, et al.
Veröffentlicht: (2025)
Beyond Point-wise Neural Collapse: A Topology-Aware Hierarchical Classifier for Class-Incremental Learning
von: Yi, Huiyu, et al.
Veröffentlicht: (2026)
von: Yi, Huiyu, et al.
Veröffentlicht: (2026)
Robust Embodied Perception in Dynamic Environments via Disentangled Weight Fusion
von: Guo, Juncen, et al.
Veröffentlicht: (2026)
von: Guo, Juncen, et al.
Veröffentlicht: (2026)
Prompt Tuning with Soft Context Sharing for Vision-Language Models
von: Ding, Kun, et al.
Veröffentlicht: (2022)
von: Ding, Kun, et al.
Veröffentlicht: (2022)
Ultrafast-and-Ultralight ConvNet-Based Intelligent Monitoring System for Diagnosing Early-Stage Mpox Anytime and Anywhere
von: Yue, Yubiao, et al.
Veröffentlicht: (2023)
von: Yue, Yubiao, et al.
Veröffentlicht: (2023)
Exploring Facial Expression Recognition through Semi-Supervised Pretraining and Temporal Modeling
von: Yu, Jun, et al.
Veröffentlicht: (2024)
von: Yu, Jun, et al.
Veröffentlicht: (2024)
DiffLoRA: Generating Personalized Low-Rank Adaptation Weights with Diffusion
von: Wu, Yujia, et al.
Veröffentlicht: (2024)
von: Wu, Yujia, et al.
Veröffentlicht: (2024)
Unified Attention Modeling for Efficient Free-Viewing and Visual Search via Shared Representations
von: Mohammed, Fatma Youssef, et al.
Veröffentlicht: (2025)
von: Mohammed, Fatma Youssef, et al.
Veröffentlicht: (2025)
Multi-Stage Generative Upscaler: Reconstructing Football Broadcast Images via Diffusion Models
von: Martini, Luca, et al.
Veröffentlicht: (2025)
von: Martini, Luca, et al.
Veröffentlicht: (2025)
UniCorn: Towards Self-Improving Unified Multimodal Models through Self-Generated Supervision
von: Han, Ruiyan, et al.
Veröffentlicht: (2026)
von: Han, Ruiyan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Extracting Multimodal Learngene in CLIP: Unveiling the Multimodal Generalizable Knowledge
von: Chen, Ruiming, et al.
Veröffentlicht: (2025) -
Safe Vision-Language Models via Unsafe Weights Manipulation
von: D'Incà, Moreno, et al.
Veröffentlicht: (2025) -
LarvSeg: Exploring Image Classification Data For Large Vocabulary Semantic Segmentation via Category-wise Attentive Classifier
von: Yu, Haojun, et al.
Veröffentlicht: (2025) -
ShareVerse: Multi-Agent Consistent Video Generation for Shared World Modeling
von: Zhu, Jiayi, et al.
Veröffentlicht: (2026) -
FINE: Factorizing Knowledge for Initialization of Variable-sized Diffusion Models
von: Xie, Yucheng, et al.
Veröffentlicht: (2024)