Unconstrained Model Merging for Enhanced LLM Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yiming, He, Baoyi, Zhang, Shengyu, Fu, Yuhao, Zhou, Qi, Sang, Zhijie, Hong, Zijin, Yang, Kejing, Wang, Wenjun, Yuan, Jianbo, Ning, Guanghan, Li, Linyi, Ji, Chunlin, Wu, Fei, Yang, Hongxia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
InfiFusion: A Unified Framework for Enhanced Cross-Model Reasoning via LLM Fusion
von: Yan, Zhaoyi, et al.
Veröffentlicht: (2025)
von: Yan, Zhaoyi, et al.
Veröffentlicht: (2025)
InfiR : Crafting Effective Small Language Models and Multimodal Small Language Models in Reasoning
von: Xie, Congkai, et al.
Veröffentlicht: (2025)
von: Xie, Congkai, et al.
Veröffentlicht: (2025)
InfiAlign: A Scalable and Sample-Efficient Framework for Aligning LLMs to Enhance Reasoning Capabilities
von: Cai, Shuo, et al.
Veröffentlicht: (2025)
von: Cai, Shuo, et al.
Veröffentlicht: (2025)
Model Merging Scaling Laws in Large Language Models
von: Wang, Yuanyi, et al.
Veröffentlicht: (2025)
von: Wang, Yuanyi, et al.
Veröffentlicht: (2025)
InfiR2: A Comprehensive FP8 Training Recipe for Reasoning-Enhanced Language Models
von: Wang, Wenjun, et al.
Veröffentlicht: (2025)
von: Wang, Wenjun, et al.
Veröffentlicht: (2025)
Learning Stackable and Skippable LEGO Bricks for Efficient, Reconfigurable, and Variable-Resolution Diffusion Modeling
von: Zheng, Huangjie, et al.
Veröffentlicht: (2023)
von: Zheng, Huangjie, et al.
Veröffentlicht: (2023)
Benchmarking LLMs' Mathematical Reasoning with Unseen Random Variables Questions
von: Hong, Zijin, et al.
Veröffentlicht: (2025)
von: Hong, Zijin, et al.
Veröffentlicht: (2025)
InfiMed: Low-Resource Medical MLLMs with Advancing Understanding and Reasoning
von: Liu, Zeyu, et al.
Veröffentlicht: (2025)
von: Liu, Zeyu, et al.
Veröffentlicht: (2025)
E-PMQ: Expert-Guided Post-Merge Quantization with Merged-Weight Anchoring
von: Wang, Wenjun, et al.
Veröffentlicht: (2026)
von: Wang, Wenjun, et al.
Veröffentlicht: (2026)
InfiBench: Evaluating the Question-Answering Capabilities of Code Large Language Models
von: Li, Linyi, et al.
Veröffentlicht: (2024)
von: Li, Linyi, et al.
Veröffentlicht: (2024)
InfiMed-ORBIT: Aligning LLMs on Open-Ended Complex Tasks via Rubric-Based Incremental Training
von: Wang, Pengkai, et al.
Veröffentlicht: (2025)
von: Wang, Pengkai, et al.
Veröffentlicht: (2025)
SORREL: Suboptimal-Demonstration-Guided Reinforcement Learning for Learning to Branch
von: Feng, Shengyu, et al.
Veröffentlicht: (2024)
von: Feng, Shengyu, et al.
Veröffentlicht: (2024)
Regularized Langevin Dynamics for Combinatorial Optimization
von: Feng, Shengyu, et al.
Veröffentlicht: (2025)
von: Feng, Shengyu, et al.
Veröffentlicht: (2025)
InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners
von: Liu, Yuhang, et al.
Veröffentlicht: (2025)
von: Liu, Yuhang, et al.
Veröffentlicht: (2025)
MergePipe: A Budget-Aware Parameter Management System for Scalable LLM Merging
von: Wang, Yuanyi, et al.
Veröffentlicht: (2026)
von: Wang, Yuanyi, et al.
Veröffentlicht: (2026)
Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models
von: Liu, Zeyu, et al.
Veröffentlicht: (2025)
von: Liu, Zeyu, et al.
Veröffentlicht: (2025)
A Memory-Augmented LLM-Driven Method for Autonomous Merging of 3D Printing Work Orders
von: Liu, Yuhao, et al.
Veröffentlicht: (2025)
von: Liu, Yuhao, et al.
Veröffentlicht: (2025)
How Likely Do LLMs with CoT Mimic Human Reasoning?
von: Bao, Guangsheng, et al.
Veröffentlicht: (2024)
von: Bao, Guangsheng, et al.
Veröffentlicht: (2024)
InfiGUIAgent: A Multimodal Generalist GUI Agent with Native Reasoning and Reflection
von: Liu, Yuhang, et al.
Veröffentlicht: (2025)
von: Liu, Yuhang, et al.
Veröffentlicht: (2025)
InfiFPO: Implicit Model Fusion via Preference Optimization in Large Language Models
von: Gu, Yanggan, et al.
Veröffentlicht: (2025)
von: Gu, Yanggan, et al.
Veröffentlicht: (2025)
InfiGFusion: Graph-on-Logits Distillation via Efficient Gromov-Wasserstein for Model Fusion
von: Wang, Yuanyi, et al.
Veröffentlicht: (2025)
von: Wang, Yuanyi, et al.
Veröffentlicht: (2025)
DeepReview: Improving LLM-based Paper Review with Human-like Deep Thinking Process
von: Zhu, Minjun, et al.
Veröffentlicht: (2025)
von: Zhu, Minjun, et al.
Veröffentlicht: (2025)
Self-Infilling Code Generation
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
von: Zheng, Lin, et al.
Veröffentlicht: (2023)
InfiMed-Foundation: Pioneering Advanced Multimodal Medical Models with Compute-Efficient Pre-Training and Multi-Stage Fine-Tuning
von: Zhu, Guanghao, et al.
Veröffentlicht: (2025)
von: Zhu, Guanghao, et al.
Veröffentlicht: (2025)
MergeNet: Knowledge Migration across Heterogeneous Models, Tasks, and Modalities
von: Li, Kunxi, et al.
Veröffentlicht: (2024)
von: Li, Kunxi, et al.
Veröffentlicht: (2024)
Interactive Learning for LLM Reasoning
von: Lin, Hehai, et al.
Veröffentlicht: (2025)
von: Lin, Hehai, et al.
Veröffentlicht: (2025)
SPL-LNS: Sampling-Enhanced Large Neighborhood Search for Solving Integer Linear Programs
von: Feng, Shengyu, et al.
Veröffentlicht: (2025)
von: Feng, Shengyu, et al.
Veröffentlicht: (2025)
Unsupervised Diffusion Solver for Combinatorial Optimization via Combinatorial Adjoint Matching
von: Feng, Shengyu, et al.
Veröffentlicht: (2026)
von: Feng, Shengyu, et al.
Veröffentlicht: (2026)
WildActor: Unconstrained Identity-Preserving Video Generation
von: Guo, Qin, et al.
Veröffentlicht: (2026)
von: Guo, Qin, et al.
Veröffentlicht: (2026)
How Can LLM Guide RL? A Value-Based Approach
von: Zhang, Shenao, et al.
Veröffentlicht: (2024)
von: Zhang, Shenao, et al.
Veröffentlicht: (2024)
FeatCal: Feature Calibration for Post-Merging Models
von: Gu, Yanggan, et al.
Veröffentlicht: (2026)
von: Gu, Yanggan, et al.
Veröffentlicht: (2026)
Rethinking Detecting Salient and Camouflaged Objects in Unconstrained Scenes
von: Zhou, Zhangjun, et al.
Veröffentlicht: (2024)
von: Zhou, Zhangjun, et al.
Veröffentlicht: (2024)
Asymmetric Association Between Trade Uncertainty and Environmental Quality: Evidence From Newly Industrialized Economies
von: Baoyi Ji, et al.
Veröffentlicht: (2025)
von: Baoyi Ji, et al.
Veröffentlicht: (2025)
An In-Depth Investigation of Data Collection in LLM App Ecosystems
von: Wu, Yuhao, et al.
Veröffentlicht: (2024)
von: Wu, Yuhao, et al.
Veröffentlicht: (2024)
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion
von: Lv, Zheqi, et al.
Veröffentlicht: (2025)
von: Lv, Zheqi, et al.
Veröffentlicht: (2025)
MemPot: Defending Against Memory Extraction Attack with Optimized Honeypots
von: Wang, Yuhao, et al.
Veröffentlicht: (2026)
von: Wang, Yuhao, et al.
Veröffentlicht: (2026)
Direct Value Optimization: Improving Chain-of-Thought Reasoning in LLMs with Refined Values
von: Zhang, Hongbo, et al.
Veröffentlicht: (2025)
von: Zhang, Hongbo, et al.
Veröffentlicht: (2025)
JuliaSpacePhysics/GeoAACGM.jl: v0.1.5
von: Zijin Zhang
Veröffentlicht: (2026)
von: Zijin Zhang
Veröffentlicht: (2026)
JuliaSpacePhysics/PySPEDAS.jl: v0.1.7
von: Zijin Zhang
Veröffentlicht: (2026)
von: Zijin Zhang
Veröffentlicht: (2026)
JuliaSpacePhysics/PySPEDAS.jl: v0.1.0
von: Zijin Zhang
Veröffentlicht: (2025)
von: Zijin Zhang
Veröffentlicht: (2025)
Ähnliche Einträge
-
InfiFusion: A Unified Framework for Enhanced Cross-Model Reasoning via LLM Fusion
von: Yan, Zhaoyi, et al.
Veröffentlicht: (2025) -
InfiR : Crafting Effective Small Language Models and Multimodal Small Language Models in Reasoning
von: Xie, Congkai, et al.
Veröffentlicht: (2025) -
InfiAlign: A Scalable and Sample-Efficient Framework for Aligning LLMs to Enhance Reasoning Capabilities
von: Cai, Shuo, et al.
Veröffentlicht: (2025) -
Model Merging Scaling Laws in Large Language Models
von: Wang, Yuanyi, et al.
Veröffentlicht: (2025) -
InfiR2: A Comprehensive FP8 Training Recipe for Reasoning-Enhanced Language Models
von: Wang, Wenjun, et al.
Veröffentlicht: (2025)