Dynamic Fisher-weighted Model Merging via Bayesian Optimization
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Sanwoo, Liu, Jiahao, Wang, Qifan, Wang, Jingang, Cai, Xunliang, Wu, Yunfang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Composable Cross-prompt Essay Scoring by Merging Models
di: Lee, Sanwoo, et al.
Pubblicazione: (2025)
di: Lee, Sanwoo, et al.
Pubblicazione: (2025)
Speculative Decoding via Early-exiting for Faster LLM Inference with Thompson Sampling Control Mechanism
di: Liu, Jiahao, et al.
Pubblicazione: (2024)
di: Liu, Jiahao, et al.
Pubblicazione: (2024)
Unleashing Large Language Models' Proficiency in Zero-shot Essay Scoring
di: Lee, Sanwoo, et al.
Pubblicazione: (2024)
di: Lee, Sanwoo, et al.
Pubblicazione: (2024)
Parallel Decoding via Hidden Transfer for Lossless Large Language Model Acceleration
di: Wu, Pengfei, et al.
Pubblicazione: (2024)
di: Wu, Pengfei, et al.
Pubblicazione: (2024)
Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring
di: Cai, Yida, et al.
Pubblicazione: (2025)
di: Cai, Yida, et al.
Pubblicazione: (2025)
FIRP: Faster LLM inference via future intermediate representation prediction
di: Wu, Pengfei, et al.
Pubblicazione: (2024)
di: Wu, Pengfei, et al.
Pubblicazione: (2024)
FPT: Feature Prompt Tuning for Few-shot Readability Assessment
di: Wang, Ziyang, et al.
Pubblicazione: (2024)
di: Wang, Ziyang, et al.
Pubblicazione: (2024)
Trait-Aware Policy Optimization for Autoregressive Multi-Trait Essay Scoring
di: Wang, Zhengyang, et al.
Pubblicazione: (2026)
di: Wang, Zhengyang, et al.
Pubblicazione: (2026)
A Survey of Uncertainty Estimation in LLMs: Theory Meets Practice
di: Huang, Hsiu-Yuan, et al.
Pubblicazione: (2024)
di: Huang, Hsiu-Yuan, et al.
Pubblicazione: (2024)
ReMamba: Equip Mamba with Effective Long-Sequence Modeling
di: Yuan, Danlong, et al.
Pubblicazione: (2024)
di: Yuan, Danlong, et al.
Pubblicazione: (2024)
Graph-Structured Speculative Decoding
di: Gong, Zhuocheng, et al.
Pubblicazione: (2024)
di: Gong, Zhuocheng, et al.
Pubblicazione: (2024)
Libra: Assessing and Improving Reward Model by Learning to Think
di: Zhou, Meng, et al.
Pubblicazione: (2025)
di: Zhou, Meng, et al.
Pubblicazione: (2025)
MiniDisc: Minimal Distillation Schedule for Language Model Compression
di: Zhang, Chen, et al.
Pubblicazione: (2022)
di: Zhang, Chen, et al.
Pubblicazione: (2022)
mCL-NER: Cross-Lingual Named Entity Recognition via Multi-view Contrastive Learning
di: Mo, Ying, et al.
Pubblicazione: (2023)
di: Mo, Ying, et al.
Pubblicazione: (2023)
SEAS: Self-Evolving Adversarial Safety Optimization for Large Language Models
di: Diao, Muxi, et al.
Pubblicazione: (2024)
di: Diao, Muxi, et al.
Pubblicazione: (2024)
Ltri-LLM: Streaming Long Context Inference for LLMs with Training-Free Dynamic Triangular Attention Pattern
di: Tang, Hongyin, et al.
Pubblicazione: (2024)
di: Tang, Hongyin, et al.
Pubblicazione: (2024)
Checkpoint Merging via Bayesian Optimization in LLM Pretraining
di: Liu, Deyuan, et al.
Pubblicazione: (2024)
di: Liu, Deyuan, et al.
Pubblicazione: (2024)
Multi-Objective and Mixed-Reward Reinforcement Learning via Reward-Decorrelated Policy Optimization
di: Bai, Yang, et al.
Pubblicazione: (2026)
di: Bai, Yang, et al.
Pubblicazione: (2026)
C-ICL: Contrastive In-context Learning for Information Extraction
di: Mo, Ying, et al.
Pubblicazione: (2024)
di: Mo, Ying, et al.
Pubblicazione: (2024)
Optimal Brain Iterative Merging: Mitigating Interference in LLM Merging
di: Wang, Zhixiang, et al.
Pubblicazione: (2025)
di: Wang, Zhixiang, et al.
Pubblicazione: (2025)
Earlier Tokens Contribute More: Learning Direct Preference Optimization From Temporal Decay Perspective
di: Shao, Ruichen, et al.
Pubblicazione: (2025)
di: Shao, Ruichen, et al.
Pubblicazione: (2025)
Length Desensitization in Direct Preference Optimization
di: Liu, Wei, et al.
Pubblicazione: (2024)
di: Liu, Wei, et al.
Pubblicazione: (2024)
NeedleInATable: Exploring Long-Context Capability of Large Language Models towards Long-Structured Tables
di: Wang, Lanrui, et al.
Pubblicazione: (2025)
di: Wang, Lanrui, et al.
Pubblicazione: (2025)
EAVE: Efficient Product Attribute Value Extraction via Lightweight Sparse-layer Interaction
di: Yang, Li, et al.
Pubblicazione: (2024)
di: Yang, Li, et al.
Pubblicazione: (2024)
Beyond Demonstrations: Dynamic Vector Construction from Latent Representations
di: Cai, Wang, et al.
Pubblicazione: (2025)
di: Cai, Wang, et al.
Pubblicazione: (2025)
FIRE: Flexible Integration of Data Quality Ratings for Effective Pre-Training
di: Xu, Liangyu, et al.
Pubblicazione: (2025)
di: Xu, Liangyu, et al.
Pubblicazione: (2025)
Unlocking Implicit Experience: Synthesizing Tool-Use Trajectories from Text
di: Xu, Zhihao, et al.
Pubblicazione: (2026)
di: Xu, Zhihao, et al.
Pubblicazione: (2026)
Predictor-Corrector Enhanced Transformers with Exponential Moving Average Coefficient Learning
di: Li, Bei, et al.
Pubblicazione: (2024)
di: Li, Bei, et al.
Pubblicazione: (2024)
LinkQA: Synthesizing Diverse QA from Multiple Seeds Strongly Linked by Knowledge Points
di: Zhang, Xuemiao, et al.
Pubblicazione: (2025)
di: Zhang, Xuemiao, et al.
Pubblicazione: (2025)
Large-Scale Diverse Synthesis for Mid-Training
di: Zhang, Xuemiao, et al.
Pubblicazione: (2025)
di: Zhang, Xuemiao, et al.
Pubblicazione: (2025)
Expanding Reasoning Potential in Foundation Model by Learning Diverse Chains of Thought Patterns
di: Zhang, Xuemiao, et al.
Pubblicazione: (2025)
di: Zhang, Xuemiao, et al.
Pubblicazione: (2025)
Functionality-Oriented LLM Merging on the Fisher--Rao Manifold
di: Wang, Jiayu, et al.
Pubblicazione: (2026)
di: Wang, Jiayu, et al.
Pubblicazione: (2026)
APP: Adaptive Prototypical Pseudo-Labeling for Few-shot OOD Detection
di: Wang, Pei, et al.
Pubblicazione: (2023)
di: Wang, Pei, et al.
Pubblicazione: (2023)
FRAME: Boosting LLMs with A Four-Quadrant Multi-Stage Pretraining Strategy
di: Zhang, Xuemiao, et al.
Pubblicazione: (2025)
di: Zhang, Xuemiao, et al.
Pubblicazione: (2025)
SyncThink: A Training-Free Strategy to Align Inference Termination with Reasoning Saturation
di: Li, Gengyang, et al.
Pubblicazione: (2026)
di: Li, Gengyang, et al.
Pubblicazione: (2026)
Unsupervised Distractor Generation via Large Language Model Distilling and Counterfactual Contrastive Decoding
di: Qu, Fanyi, et al.
Pubblicazione: (2024)
di: Qu, Fanyi, et al.
Pubblicazione: (2024)
Causal Autoregressive Diffusion Language Model
di: Ruan, Junhao, et al.
Pubblicazione: (2026)
di: Ruan, Junhao, et al.
Pubblicazione: (2026)
1bit-Merging: Dynamic Quantized Merging for Large Language Models
di: Liu, Shuqi, et al.
Pubblicazione: (2025)
di: Liu, Shuqi, et al.
Pubblicazione: (2025)
Beyond the Known: Investigating LLMs Performance on Out-of-Domain Intent Detection
di: Wang, Pei, et al.
Pubblicazione: (2024)
di: Wang, Pei, et al.
Pubblicazione: (2024)
A Preliminary Study on the Promises and Challenges of Native Top-$k$ Sparse Attention
di: Xiu, Di, et al.
Pubblicazione: (2025)
di: Xiu, Di, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Composable Cross-prompt Essay Scoring by Merging Models
di: Lee, Sanwoo, et al.
Pubblicazione: (2025) -
Speculative Decoding via Early-exiting for Faster LLM Inference with Thompson Sampling Control Mechanism
di: Liu, Jiahao, et al.
Pubblicazione: (2024) -
Unleashing Large Language Models' Proficiency in Zero-shot Essay Scoring
di: Lee, Sanwoo, et al.
Pubblicazione: (2024) -
Parallel Decoding via Hidden Transfer for Lossless Large Language Model Acceleration
di: Wu, Pengfei, et al.
Pubblicazione: (2024) -
Rank-Then-Score: Enhancing Large Language Models for Automated Essay Scoring
di: Cai, Yida, et al.
Pubblicazione: (2025)