Middo: Model-Informed Dynamic Data Optimization for Enhanced LLM Fine-Tuning via Closed-Loop Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Zinan, Gao, Xin, Pei, Qizhi, Pan, Zhuoshi, Cai, Mengzhang, Wu, Jiang, He, Conghui, Wu, Lijun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MetaLadder: Ascending Mathematical Solution Quality via Analogical-Problem Reasoning Transfer
by: Lin, Honglin, et al.
Published: (2025)
by: Lin, Honglin, et al.
Published: (2025)
Closing the Data Loop: Using OpenDataArena to Engineer Superior Training Datasets
by: Gao, Xin, et al.
Published: (2025)
by: Gao, Xin, et al.
Published: (2025)
ScaleDiff: Scaling Difficult Problems for Advanced Mathematical Reasoning
by: Pei, Qizhi, et al.
Published: (2025)
by: Pei, Qizhi, et al.
Published: (2025)
IDEAL: Data Equilibrium Adaptation for Multi-Capability Language Model Alignment
by: Ming, Chenlin, et al.
Published: (2025)
by: Ming, Chenlin, et al.
Published: (2025)
MathFusion: Enhancing Mathematical Problem-solving of LLM through Instruction Fusion
by: Pei, Qizhi, et al.
Published: (2025)
by: Pei, Qizhi, et al.
Published: (2025)
REST: Stress Testing Large Reasoning Models by Asking Multiple Problems at Once
by: Pan, Zhuoshi, et al.
Published: (2025)
by: Pan, Zhuoshi, et al.
Published: (2025)
A Strategic Coordination Framework of Small LLMs Matches Large LLMs in Data Synthesis
by: Gao, Xin, et al.
Published: (2025)
by: Gao, Xin, et al.
Published: (2025)
Scaling Code-Assisted Chain-of-Thoughts and Instructions for Model Reasoning
by: Lin, Honglin, et al.
Published: (2025)
by: Lin, Honglin, et al.
Published: (2025)
LEMMA: Learning from Errors for MatheMatical Advancement in LLMs
by: Pan, Zhuoshi, et al.
Published: (2025)
by: Pan, Zhuoshi, et al.
Published: (2025)
OpenDataArena: A Fair and Open Arena for Benchmarking Post-Training Dataset Value
by: Cai, Mengzhang, et al.
Published: (2025)
by: Cai, Mengzhang, et al.
Published: (2025)
CipherBank: Exploring the Boundary of LLM Reasoning Capabilities through Cryptography Challenges
by: Li, Yu, et al.
Published: (2025)
by: Li, Yu, et al.
Published: (2025)
ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch
by: Liu, Zheng, et al.
Published: (2026)
by: Liu, Zheng, et al.
Published: (2026)
Can One Domain Help Others? A Data-Centric Study on Multi-Domain Reasoning via Reinforcement Learning
by: Li, Yu, et al.
Published: (2025)
by: Li, Yu, et al.
Published: (2025)
Unlocking Data Value in Finance: A Study on Distillation and Difficulty-Aware Training
by: Cao, Chuxue, et al.
Published: (2026)
by: Cao, Chuxue, et al.
Published: (2026)
Tracing the Roots: A Multi-Agent Framework for Uncovering Data Lineage in Post-Training LLMs
by: Li, Yu, et al.
Published: (2026)
by: Li, Yu, et al.
Published: (2026)
FABind+: Enhancing Molecular Docking through Improved Pocket Prediction and Pose Generation
by: Gao, Kaiyuan, et al.
Published: (2024)
by: Gao, Kaiyuan, et al.
Published: (2024)
Heterogeneous Adaptive Policy Optimization: Tailoring Optimization to Every Token's Nature
by: Liu, Zheng, et al.
Published: (2025)
by: Liu, Zheng, et al.
Published: (2025)
Scientific Image Synthesis: Benchmarking, Methodologies, and Downstream Utility
by: Lin, Honglin, et al.
Published: (2026)
by: Lin, Honglin, et al.
Published: (2026)
3D-MolT5: Leveraging Discrete Structural Information for Molecule-Text Modeling
by: Pei, Qizhi, et al.
Published: (2024)
by: Pei, Qizhi, et al.
Published: (2024)
Back to Blackwell: Closing the Loop on Intransitivity in Multi-Objective Preference Fine-Tuning
by: Zhang, Jiahao, et al.
Published: (2026)
by: Zhang, Jiahao, et al.
Published: (2026)
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation
by: Zhu, Runchuan, et al.
Published: (2025)
by: Zhu, Runchuan, et al.
Published: (2025)
BioT5+: Towards Generalized Biological Understanding with IUPAC Integration and Multi-task Tuning
by: Pei, Qizhi, et al.
Published: (2024)
by: Pei, Qizhi, et al.
Published: (2024)
Closed-Loop Supervised Fine-Tuning of Tokenized Traffic Models
by: Zhang, Zhejun, et al.
Published: (2024)
by: Zhang, Zhejun, et al.
Published: (2024)
MobiLLM: Enabling LLM Fine-Tuning on the Mobile Device via Server Assisted Side Tuning
by: Li, Liang, et al.
Published: (2025)
by: Li, Liang, et al.
Published: (2025)
A Survey on Split Learning for LLM Fine-Tuning: Models, Systems, and Privacy Optimizations
by: Liu, Zihan, et al.
Published: (2026)
by: Liu, Zihan, et al.
Published: (2026)
RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
PlannerRFT: Reinforcing Diffusion Planners through Closed-Loop and Sample-Efficient Fine-Tuning
by: Li, Hongchen, et al.
Published: (2026)
by: Li, Hongchen, et al.
Published: (2026)
LENSLLM: Unveiling Fine-Tuning Dynamics for LLM Selection
by: Zeng, Xinyue, et al.
Published: (2025)
by: Zeng, Xinyue, et al.
Published: (2025)
Exploiting Pre-trained Models for Drug Target Affinity Prediction with Nearest Neighbors
by: Pei, Qizhi, et al.
Published: (2024)
by: Pei, Qizhi, et al.
Published: (2024)
BioT5: Enriching Cross-modal Integration in Biology with Chemical Knowledge and Natural Language Associations
by: Pei, Qizhi, et al.
Published: (2023)
by: Pei, Qizhi, et al.
Published: (2023)
Gradual Learning: Optimizing Fine-Tuning with Partially Mastered Knowledge in Large Language Models
by: Li, Bozhou, et al.
Published: (2024)
by: Li, Bozhou, et al.
Published: (2024)
Quantum-Enhanced LLM Efficient Fine Tuning
by: Kong, Xiaofei, et al.
Published: (2025)
by: Kong, Xiaofei, et al.
Published: (2025)
Tokenizing 3D Molecule Structure with Quantized Spherical Coordinates
by: Gao, Kaiyuan, et al.
Published: (2024)
by: Gao, Kaiyuan, et al.
Published: (2024)
Closed‐Loop Recyclable and Thermal‐Conductive Fluoropolyimide Composites for High‐Temperature Capacitors
by: Mingyuan Yang, et al.
Published: (2026)
by: Mingyuan Yang, et al.
Published: (2026)
Envision: Benchmarking Unified Understanding & Generation for Causal World Process Insights
by: Tian, Juanxi, et al.
Published: (2025)
by: Tian, Juanxi, et al.
Published: (2025)
Utilize the Flow before Stepping into the Same River Twice: Certainty Represented Knowledge Flow for Refusal-Aware Instruction Tuning
by: Zhu, Runchuan, et al.
Published: (2024)
by: Zhu, Runchuan, et al.
Published: (2024)
Alignment Dynamics in LLM Fine-Tuning
by: Huang, Yuhan, et al.
Published: (2026)
by: Huang, Yuhan, et al.
Published: (2026)
PAE MobiLLM: Privacy-Aware and Efficient LLM Fine-Tuning on the Mobile Device via Additive Side-Tuning
by: Yang, Xingke, et al.
Published: (2025)
by: Yang, Xingke, et al.
Published: (2025)
MMFineReason: Closing the Multimodal Reasoning Gap via Open Data-Centric Methods
by: Lin, Honglin, et al.
Published: (2026)
by: Lin, Honglin, et al.
Published: (2026)
High‐performance self‐healing epoxy by microencapsulated epoxy‐amine chemistry I: Properties of the adopted healant system
by: Mengzhang Zhu, et al.
Published: (2024)
by: Mengzhang Zhu, et al.
Published: (2024)
Similar Items
-
MetaLadder: Ascending Mathematical Solution Quality via Analogical-Problem Reasoning Transfer
by: Lin, Honglin, et al.
Published: (2025) -
Closing the Data Loop: Using OpenDataArena to Engineer Superior Training Datasets
by: Gao, Xin, et al.
Published: (2025) -
ScaleDiff: Scaling Difficult Problems for Advanced Mathematical Reasoning
by: Pei, Qizhi, et al.
Published: (2025) -
IDEAL: Data Equilibrium Adaptation for Multi-Capability Language Model Alignment
by: Ming, Chenlin, et al.
Published: (2025) -
MathFusion: Enhancing Mathematical Problem-solving of LLM through Instruction Fusion
by: Pei, Qizhi, et al.
Published: (2025)