Bohdi: Heterogeneous LLM Fusion with Automatic Data Exploration
Fuente:
arXiv
Salvato in:
| Autori principali: | Gao, Junqi, Guo, Zhichang, Zhang, Dazhi, Li, Dong, Liu, Runze, Li, Pengfei, Tian, Kai, Qi, Biqing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PDAC: Efficient Coreset Selection for Continual Learning via Probability Density Awareness
di: Gao, Junqi, et al.
Pubblicazione: (2025)
di: Gao, Junqi, et al.
Pubblicazione: (2025)
Auto-FlexSwitch: Efficient Dynamic Model Merging via Learnable Task Vector Compression
di: Gao, Junqi, et al.
Pubblicazione: (2026)
di: Gao, Junqi, et al.
Pubblicazione: (2026)
Perturbation Towards Easy Samples Improves Targeted Adversarial Transferability
di: Gao, Junqi, et al.
Pubblicazione: (2024)
di: Gao, Junqi, et al.
Pubblicazione: (2024)
An Efficient Memory Module for Graph Few-Shot Class-Incremental Learning
di: Li, Dong, et al.
Pubblicazione: (2024)
di: Li, Dong, et al.
Pubblicazione: (2024)
Exploring Adversarial Robustness of Deep State Space Models
di: Qi, Biqing, et al.
Pubblicazione: (2024)
di: Qi, Biqing, et al.
Pubblicazione: (2024)
Fast and Slow Gradient Approximation for Binary Neural Network Optimization
di: Chen, Xinquan, et al.
Pubblicazione: (2024)
di: Chen, Xinquan, et al.
Pubblicazione: (2024)
Online DPO: Online Direct Preference Optimization with Fast-Slow Chasing
di: Qi, Biqing, et al.
Pubblicazione: (2024)
di: Qi, Biqing, et al.
Pubblicazione: (2024)
WIST: Web-Grounded Iterative Self-Play Tree for Domain-Targeted Reasoning Improvement
di: Li, Fangyuan, et al.
Pubblicazione: (2026)
di: Li, Fangyuan, et al.
Pubblicazione: (2026)
SMR: State Memory Replay for Long Sequence Modeling
di: Qi, Biqing, et al.
Pubblicazione: (2024)
di: Qi, Biqing, et al.
Pubblicazione: (2024)
Less is More: Efficient Model Merging with Binary Task Switch
di: Qi, Biqing, et al.
Pubblicazione: (2024)
di: Qi, Biqing, et al.
Pubblicazione: (2024)
Enhancing Adversarial Transferability via Information Bottleneck Constraints
di: Qi, Biqing, et al.
Pubblicazione: (2024)
di: Qi, Biqing, et al.
Pubblicazione: (2024)
Interactive Continual Learning: Fast and Slow Thinking
di: Qi, Biqing, et al.
Pubblicazione: (2024)
di: Qi, Biqing, et al.
Pubblicazione: (2024)
MARTI-MARS$^2$: Scaling Multi-Agent Self-Search via Reinforcement Learning for Code Generation
di: Wang, Shijie, et al.
Pubblicazione: (2026)
di: Wang, Shijie, et al.
Pubblicazione: (2026)
ADO: Automatic Data Optimization for Inputs in LLM Prompts
di: Lin, Sam, et al.
Pubblicazione: (2025)
di: Lin, Sam, et al.
Pubblicazione: (2025)
Mitigating Heterogeneous Token Overfitting in LLM Knowledge Editing
di: Liu, Tianci, et al.
Pubblicazione: (2025)
di: Liu, Tianci, et al.
Pubblicazione: (2025)
Graph Counselor: Adaptive Graph Exploration via Multi-Agent Synergy to Enhance LLM Reasoning
di: Gao, Junqi, et al.
Pubblicazione: (2025)
di: Gao, Junqi, et al.
Pubblicazione: (2025)
SFedHIFI: Fire Rate-Based Heterogeneous Information Fusion for Spiking Federated Learning
di: Tao, Ran, et al.
Pubblicazione: (2026)
di: Tao, Ran, et al.
Pubblicazione: (2026)
Contribution Evaluation of Heterogeneous Participants in Federated Learning via Prototypical Representations
di: Guo, Qi, et al.
Pubblicazione: (2024)
di: Guo, Qi, et al.
Pubblicazione: (2024)
Dynamic Bayesian Optimization Framework for Instruction Tuning in Partial Differential Equation Discovery
di: Qu, Junqi, et al.
Pubblicazione: (2025)
di: Qu, Junqi, et al.
Pubblicazione: (2025)
HeLo: Heterogeneous Multi-Modal Fusion with Label Correlation for Emotion Distribution Learning
di: Zheng, Chuhang, et al.
Pubblicazione: (2025)
di: Zheng, Chuhang, et al.
Pubblicazione: (2025)
Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling
di: Liu, Runze, et al.
Pubblicazione: (2025)
di: Liu, Runze, et al.
Pubblicazione: (2025)
Uncertainty-Aware Data-Based Method for Fast and Reliable Shape Optimization
di: Yang, Yunjia, et al.
Pubblicazione: (2026)
di: Yang, Yunjia, et al.
Pubblicazione: (2026)
Positive and Unlabeled Data: Model, Estimation, Inference, and Classification
di: Liu, Siyan, et al.
Pubblicazione: (2024)
di: Liu, Siyan, et al.
Pubblicazione: (2024)
PINNsFailureRegion Localization and Refinement through White-box AdversarialAttack
di: Shi, Shengzhu, et al.
Pubblicazione: (2023)
di: Shi, Shengzhu, et al.
Pubblicazione: (2023)
Semiparametric Learning from Open-Set Label Shift Data
di: Liu, Siyan, et al.
Pubblicazione: (2025)
di: Liu, Siyan, et al.
Pubblicazione: (2025)
AutoHete: An Automatic and Efficient Heterogeneous Training System for LLMs
di: Zeng, Zihao, et al.
Pubblicazione: (2025)
di: Zeng, Zihao, et al.
Pubblicazione: (2025)
Exploration and Anti-Exploration with Distributional Random Network Distillation
di: Yang, Kai, et al.
Pubblicazione: (2024)
di: Yang, Kai, et al.
Pubblicazione: (2024)
FusionFactory: Fusing LLM Capabilities with Multi-LLM Log Data
di: Feng, Tao, et al.
Pubblicazione: (2025)
di: Feng, Tao, et al.
Pubblicazione: (2025)
HEALNet: Multimodal Fusion for Heterogeneous Biomedical Data
di: Hemker, Konstantin, et al.
Pubblicazione: (2023)
di: Hemker, Konstantin, et al.
Pubblicazione: (2023)
X-SQL: Expert Schema Linking and Understanding of Text-to-SQL with Multi-LLMs
di: Peng, Dazhi
Pubblicazione: (2025)
di: Peng, Dazhi
Pubblicazione: (2025)
SDAR: A Synergistic Diffusion-AutoRegression Paradigm for Scalable Sequence Generation
di: Cheng, Shuang, et al.
Pubblicazione: (2025)
di: Cheng, Shuang, et al.
Pubblicazione: (2025)
Fed MobiLLM: Efficient Federated LLM Fine-Tuning over Heterogeneous Mobile Devices via Server Assisted Side-Tuning
di: Yang, Xingke, et al.
Pubblicazione: (2025)
di: Yang, Xingke, et al.
Pubblicazione: (2025)
Spatial Heterogeneity in Climate Risk and Human Flourishing: An Exploration with Generative AI
di: Iacus, Stefano Maria, et al.
Pubblicazione: (2026)
di: Iacus, Stefano Maria, et al.
Pubblicazione: (2026)
Clustering by Mining Density Distributions and Splitting Manifold Structure
di: Xu, Zhichang, et al.
Pubblicazione: (2024)
di: Xu, Zhichang, et al.
Pubblicazione: (2024)
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
di: Jiang, Yuhua, et al.
Pubblicazione: (2025)
di: Jiang, Yuhua, et al.
Pubblicazione: (2025)
SpecMol: A Spectroscopy-Grounded Foundation Model for Multi-Task Molecular Learning
di: Shen, Shuaike, et al.
Pubblicazione: (2025)
di: Shen, Shuaike, et al.
Pubblicazione: (2025)
A Relative Error-Based Evaluation Framework of Heterogeneous Treatment Effect Estimators
di: Guo, Jiayi, et al.
Pubblicazione: (2025)
di: Guo, Jiayi, et al.
Pubblicazione: (2025)
Neyman-Pearson multiclass classification under label noise via empirical likelihood
di: Zhang, Qiong, et al.
Pubblicazione: (2026)
di: Zhang, Qiong, et al.
Pubblicazione: (2026)
TransFusion: Covariate-Shift Robust Transfer Learning for High-Dimensional Regression
di: He, Zelin, et al.
Pubblicazione: (2024)
di: He, Zelin, et al.
Pubblicazione: (2024)
AutoTailor: Automatic and Efficient Adaptive Model Deployment for Diverse Edge Devices
di: Liu, Mengyang, et al.
Pubblicazione: (2025)
di: Liu, Mengyang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
PDAC: Efficient Coreset Selection for Continual Learning via Probability Density Awareness
di: Gao, Junqi, et al.
Pubblicazione: (2025) -
Auto-FlexSwitch: Efficient Dynamic Model Merging via Learnable Task Vector Compression
di: Gao, Junqi, et al.
Pubblicazione: (2026) -
Perturbation Towards Easy Samples Improves Targeted Adversarial Transferability
di: Gao, Junqi, et al.
Pubblicazione: (2024) -
An Efficient Memory Module for Graph Few-Shot Class-Incremental Learning
di: Li, Dong, et al.
Pubblicazione: (2024) -
Exploring Adversarial Robustness of Deep State Space Models
di: Qi, Biqing, et al.
Pubblicazione: (2024)