Gespeichert in:
| Hauptverfasser: | Shi, Boyu, Zhou, Junbo, Liu, Chang, Yang, Xu, Wang, Qiufeng, Geng, Xin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2605.08209 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring Learngene via Stage-wise Weight Sharing for Initializing Variable-sized Models
von: Xia, Shi-Yu, et al.
Veröffentlicht: (2024)
von: Xia, Shi-Yu, et al.
Veröffentlicht: (2024)
Transferring Core Knowledge via Learngenes
von: Feng, Fu, et al.
Veröffentlicht: (2024)
von: Feng, Fu, et al.
Veröffentlicht: (2024)
Chain-based Distillation for Effective Initialization of Variable-Sized Small Language Models
von: Shi, Boyu, et al.
Veröffentlicht: (2026)
von: Shi, Boyu, et al.
Veröffentlicht: (2026)
GENE-FL: Gene-Driven Parameter-Efficient Dynamic Federated Learning
von: Guo, Shunxin, et al.
Veröffentlicht: (2025)
von: Guo, Shunxin, et al.
Veröffentlicht: (2025)
NASH: Neural Architecture and Accelerator Search for Multiplication-Reduced Hybrid Models
von: Xu, Yang, et al.
Veröffentlicht: (2024)
von: Xu, Yang, et al.
Veröffentlicht: (2024)
Towards Understanding Feature Learning in Parameter Transfer
von: Yuan, Hua, et al.
Veröffentlicht: (2025)
von: Yuan, Hua, et al.
Veröffentlicht: (2025)
When Forgetting Builds Reliability: LLM Unlearning for Reliable Hardware Code Generation
von: Liang, Yiwen, et al.
Veröffentlicht: (2025)
von: Liang, Yiwen, et al.
Veröffentlicht: (2025)
WAVE: Weight Templates for Adaptive Initialization of Variable-sized Models
von: Feng, Fu, et al.
Veröffentlicht: (2024)
von: Feng, Fu, et al.
Veröffentlicht: (2024)
Extracting Multimodal Learngene in CLIP: Unveiling the Multimodal Generalizable Knowledge
von: Chen, Ruiming, et al.
Veröffentlicht: (2025)
von: Chen, Ruiming, et al.
Veröffentlicht: (2025)
From Isolation to Integration: Building an Adaptive Expert Forest for Pre-Trained Model-based Class-Incremental Learning
von: Liu, Ruiqi, et al.
Veröffentlicht: (2026)
von: Liu, Ruiqi, et al.
Veröffentlicht: (2026)
Efficient Deployment of Deep MIMO Detection Using Learngene
von: Zhang, Jinya, et al.
Veröffentlicht: (2025)
von: Zhang, Jinya, et al.
Veröffentlicht: (2025)
DiffCPS: Diffusion Model based Constrained Policy Search for Offline Reinforcement Learning
von: He, Longxiang, et al.
Veröffentlicht: (2023)
von: He, Longxiang, et al.
Veröffentlicht: (2023)
Learning to Search for Vehicle Routing with Multiple Time Windows
von: Xu, Kuan, et al.
Veröffentlicht: (2025)
von: Xu, Kuan, et al.
Veröffentlicht: (2025)
Diagnosing Retrieval Bias Under Multiple In-Context Knowledge Updates in Large Language Models
von: Qiao, Boyu, et al.
Veröffentlicht: (2026)
von: Qiao, Boyu, et al.
Veröffentlicht: (2026)
XPERT: Expert Knowledge Transfer for Effective Training of Language Models
von: Liu, Chang, et al.
Veröffentlicht: (2026)
von: Liu, Chang, et al.
Veröffentlicht: (2026)
Revisiting Randomization in Greedy Model Search
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
Heterogeneous Data Game: Characterizing the Model Competition Across Multiple Data Sources
von: Xu, Renzhe, et al.
Veröffentlicht: (2025)
von: Xu, Renzhe, et al.
Veröffentlicht: (2025)
Can Class-Priors Help Single-Positive Multi-Label Learning?
von: Liu, Biao, et al.
Veröffentlicht: (2023)
von: Liu, Biao, et al.
Veröffentlicht: (2023)
Constraint-based Pre-training: From Structured Constraints to Scalable Model Initialization
von: Feng, Fu, et al.
Veröffentlicht: (2026)
von: Feng, Fu, et al.
Veröffentlicht: (2026)
Adaptive Federated LoRA in Heterogeneous Wireless Networks with Independent Sampling
von: Hou, Yanzhao, et al.
Veröffentlicht: (2025)
von: Hou, Yanzhao, et al.
Veröffentlicht: (2025)
Un-mixing Test-time Adaptation under Heterogeneous Data Streams
von: Su, Zixian, et al.
Veröffentlicht: (2024)
von: Su, Zixian, et al.
Veröffentlicht: (2024)
DeepVision-103K: A Visually Diverse, Broad-Coverage, and Verifiable Mathematical Dataset for Multimodal Reasoning
von: Sun, Haoxiang, et al.
Veröffentlicht: (2026)
von: Sun, Haoxiang, et al.
Veröffentlicht: (2026)
Exploring the Impact of Dataset Statistical Effect Size on Model Performance and Data Sample Size Sufficiency
von: Hatamian, Arya, et al.
Veröffentlicht: (2025)
von: Hatamian, Arya, et al.
Veröffentlicht: (2025)
Saving for the future: Enhancing generalization via partial logic regularization
von: Tan, Zhaorui, et al.
Veröffentlicht: (2025)
von: Tan, Zhaorui, et al.
Veröffentlicht: (2025)
Towards Unified Task Embeddings Across Multiple Models: Bridging the Gap for Prompt-Based Large Language Models and Beyond
von: Wang, Xinyu, et al.
Veröffentlicht: (2024)
von: Wang, Xinyu, et al.
Veröffentlicht: (2024)
Multi-Study R-Learner for Estimating Heterogeneous Treatment Effects Across Studies Using Statistical Machine Learning
von: Shyr, Cathy, et al.
Veröffentlicht: (2023)
von: Shyr, Cathy, et al.
Veröffentlicht: (2023)
ASSEMBLAGE-DEEPHISTORY: A Cross-Build Binary Dataset with Temporal Coverage
von: Liu, Chang, et al.
Veröffentlicht: (2026)
von: Liu, Chang, et al.
Veröffentlicht: (2026)
AutoSynth: Automated Workflow Optimization for High-Quality Synthetic Dataset Generation via Monte Carlo Tree Search
von: Bi, Shuzhen, et al.
Veröffentlicht: (2025)
von: Bi, Shuzhen, et al.
Veröffentlicht: (2025)
Score Neural Operator: A Generative Model for Learning and Generalizing Across Multiple Probability Distributions
von: Liao, Xinyu, et al.
Veröffentlicht: (2024)
von: Liao, Xinyu, et al.
Veröffentlicht: (2024)
BitStack: Any-Size Compression of Large Language Models in Variable Memory Environments
von: Wang, Xinghao, et al.
Veröffentlicht: (2024)
von: Wang, Xinghao, et al.
Veröffentlicht: (2024)
Federated Continual Learning via Knowledge Fusion: A Survey
von: Yang, Xin, et al.
Veröffentlicht: (2023)
von: Yang, Xin, et al.
Veröffentlicht: (2023)
Unveiling Statistical Significance of Online Regression over Multiple Datasets
von: Abu-Shaira, Mohammad, et al.
Veröffentlicht: (2025)
von: Abu-Shaira, Mohammad, et al.
Veröffentlicht: (2025)
Joint Training Across Multiple Activation Sparsity Regimes
von: Wang, Haotian
Veröffentlicht: (2026)
von: Wang, Haotian
Veröffentlicht: (2026)
Towards Size-invariant Salient Object Detection: A Generic Evaluation and Optimization Approach
von: Bao, Shilong, et al.
Veröffentlicht: (2025)
von: Bao, Shilong, et al.
Veröffentlicht: (2025)
BTS: Building Timeseries Dataset: Empowering Large-Scale Building Analytics
von: Prabowo, Arian, et al.
Veröffentlicht: (2024)
von: Prabowo, Arian, et al.
Veröffentlicht: (2024)
Direction Finding with Sparse Arrays Based on Variable Window Size Spatial Smoothing
von: Leite, Wesley S., et al.
Veröffentlicht: (2025)
von: Leite, Wesley S., et al.
Veröffentlicht: (2025)
Multiple Instance Verification
von: Xu, Xin, et al.
Veröffentlicht: (2024)
von: Xu, Xin, et al.
Veröffentlicht: (2024)
A Sensitivity-Driven Expert Allocation Method in LoRA-MoE for Efficient Fine-Tuning
von: Xu, Junzhou, et al.
Veröffentlicht: (2025)
von: Xu, Junzhou, et al.
Veröffentlicht: (2025)
Scaling Law Phenomena Across Regression Paradigms: Multiple and Kernel Approaches
von: Chen, Yifang, et al.
Veröffentlicht: (2025)
von: Chen, Yifang, et al.
Veröffentlicht: (2025)
When to Commit? Towards Variable-Size Self-Contained Blocks for Discrete Diffusion Language Models
von: Wang, Danny, et al.
Veröffentlicht: (2026)
von: Wang, Danny, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Exploring Learngene via Stage-wise Weight Sharing for Initializing Variable-sized Models
von: Xia, Shi-Yu, et al.
Veröffentlicht: (2024) -
Transferring Core Knowledge via Learngenes
von: Feng, Fu, et al.
Veröffentlicht: (2024) -
Chain-based Distillation for Effective Initialization of Variable-Sized Small Language Models
von: Shi, Boyu, et al.
Veröffentlicht: (2026) -
GENE-FL: Gene-Driven Parameter-Efficient Dynamic Federated Learning
von: Guo, Shunxin, et al.
Veröffentlicht: (2025) -
NASH: Neural Architecture and Accelerator Search for Multiplication-Reduced Hybrid Models
von: Xu, Yang, et al.
Veröffentlicht: (2024)