Gespeichert in:
| Hauptverfasser: | Shen, Qianli, Chen, Daoyuan, Huang, Yilun, Ling, Zhenqing, Li, Yaliang, Ding, Bolin, Zhou, Jingren |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2510.26374 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Diversity as a Reward: Fine-Tuning LLMs on a Mixture of Domain-Undetermined Data
von: Ling, Zhenqing, et al.
Veröffentlicht: (2025)
von: Ling, Zhenqing, et al.
Veröffentlicht: (2025)
Data-Juicer Sandbox: A Feedback-Driven Suite for Multimodal Data-Model Co-development
von: Chen, Daoyuan, et al.
Veröffentlicht: (2024)
von: Chen, Daoyuan, et al.
Veröffentlicht: (2024)
Img-Diff: Contrastive Data Synthesis for Multimodal Large Language Models
von: Jiao, Qirui, et al.
Veröffentlicht: (2024)
von: Jiao, Qirui, et al.
Veröffentlicht: (2024)
MindGYM: What Matters in Question Synthesis for Thinking-Centric Fine-Tuning?
von: Xu, Zhe, et al.
Veröffentlicht: (2025)
von: Xu, Zhe, et al.
Veröffentlicht: (2025)
Designing Algorithms Empowered by Language Models: An Analytical Framework, Case Studies, and Insights
von: Chen, Yanxi, et al.
Veröffentlicht: (2024)
von: Chen, Yanxi, et al.
Veröffentlicht: (2024)
Grounded in Reality: Learning and Deploying Proactive LLM from Offline Logs
von: Wei, Fei, et al.
Veröffentlicht: (2025)
von: Wei, Fei, et al.
Veröffentlicht: (2025)
EE-LLM: Large-Scale Training and Inference of Early-Exit Large Language Models with 3D Parallelism
von: Chen, Yanxi, et al.
Veröffentlicht: (2023)
von: Chen, Yanxi, et al.
Veröffentlicht: (2023)
HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks
von: Zhou, Ting, et al.
Veröffentlicht: (2024)
von: Zhou, Ting, et al.
Veröffentlicht: (2024)
From Training-Free to Adaptive: Empirical Insights into MLLMs' Understanding of Detection Information
von: Jiao, Qirui, et al.
Veröffentlicht: (2024)
von: Jiao, Qirui, et al.
Veröffentlicht: (2024)
Provable Scaling Laws for the Test-Time Compute of Large Language Models
von: Chen, Yanxi, et al.
Veröffentlicht: (2024)
von: Chen, Yanxi, et al.
Veröffentlicht: (2024)
EE-Tuning: An Economical yet Scalable Solution for Tuning Early-Exit Large Language Models
von: Pan, Xuchen, et al.
Veröffentlicht: (2024)
von: Pan, Xuchen, et al.
Veröffentlicht: (2024)
Tree-based Models for Vertical Federated Learning: A Survey
von: Qian, Bingchen, et al.
Veröffentlicht: (2025)
von: Qian, Bingchen, et al.
Veröffentlicht: (2025)
BiMix: A Bivariate Data Mixing Law for Language Model Pretraining
von: Ge, Ce, et al.
Veröffentlicht: (2024)
von: Ge, Ce, et al.
Veröffentlicht: (2024)
The Synergy between Data and Multi-Modal Large Language Models: A Survey from Co-Development Perspective
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
von: Qin, Zhen, et al.
Veröffentlicht: (2024)
DetailMaster: Can Your Text-to-Image Model Handle Long Prompts?
von: Jiao, Qirui, et al.
Veröffentlicht: (2025)
von: Jiao, Qirui, et al.
Veröffentlicht: (2025)
On-Policy RL Meets Off-Policy Experts: Harmonizing Supervised Fine-Tuning and Reinforcement Learning via Dynamic Weighting
von: Zhang, Wenhao, et al.
Veröffentlicht: (2025)
von: Zhang, Wenhao, et al.
Veröffentlicht: (2025)
Towards Anthropomorphic Conversational AI Part I: A Practical Framework
von: Wei, Fei, et al.
Veröffentlicht: (2025)
von: Wei, Fei, et al.
Veröffentlicht: (2025)
Data-Juicer 2.0: Cloud-Scale Adaptive Data Processing for and with Foundation Models
von: Chen, Daoyuan, et al.
Veröffentlicht: (2024)
von: Chen, Daoyuan, et al.
Veröffentlicht: (2024)
Federated Fine-tuning of Large Language Models under Heterogeneous Tasks and Client Resources
von: Bai, Jiamu, et al.
Veröffentlicht: (2024)
von: Bai, Jiamu, et al.
Veröffentlicht: (2024)
On the Convergence of Zeroth-Order Federated Tuning for Large Language Models
von: Ling, Zhenqing, et al.
Veröffentlicht: (2024)
von: Ling, Zhenqing, et al.
Veröffentlicht: (2024)
Trinity-RFT: A General-Purpose and Unified Framework for Reinforcement Fine-Tuning of Large Language Models
von: Pan, Xuchen, et al.
Veröffentlicht: (2025)
von: Pan, Xuchen, et al.
Veröffentlicht: (2025)
Very Large-Scale Multi-Agent Simulation in AgentScope
von: Pan, Xuchen, et al.
Veröffentlicht: (2024)
von: Pan, Xuchen, et al.
Veröffentlicht: (2024)
UniDM: A Unified Framework for Data Manipulation with Large Language Models
von: Qian, Yichen, et al.
Veröffentlicht: (2024)
von: Qian, Yichen, et al.
Veröffentlicht: (2024)
Output Scaling: YingLong-Delayed Chain of Thought in a Large Pretrained Time Series Forecasting Model
von: Wang, Xue, et al.
Veröffentlicht: (2025)
von: Wang, Xue, et al.
Veröffentlicht: (2025)
Decouple-Then-Merge: Finetune Diffusion Models as Multi-Task Learning
von: Ma, Qianli, et al.
Veröffentlicht: (2024)
von: Ma, Qianli, et al.
Veröffentlicht: (2024)
BOTS: Batch Bayesian Optimization of Extended Thompson Sampling for Severely Episode-Limited RL Settings
von: Karine, Karine, et al.
Veröffentlicht: (2024)
von: Karine, Karine, et al.
Veröffentlicht: (2024)
Dynamic Demonstration Retrieval and Cognitive Understanding for Emotional Support Conversation
von: Xu, Zhe, et al.
Veröffentlicht: (2024)
von: Xu, Zhe, et al.
Veröffentlicht: (2024)
An Auction-based Marketplace for Model Trading in Federated Learning
von: Cui, Yue, et al.
Veröffentlicht: (2024)
von: Cui, Yue, et al.
Veröffentlicht: (2024)
API-guided Dataset Synthesis to Finetune Large Code Models
von: Li, Zongjie, et al.
Veröffentlicht: (2024)
von: Li, Zongjie, et al.
Veröffentlicht: (2024)
A Bargaining-based Approach for Feature Trading in Vertical Federated Learning
von: Cui, Yue, et al.
Veröffentlicht: (2024)
von: Cui, Yue, et al.
Veröffentlicht: (2024)
ArtAug: Enhancing Text-to-Image Generation through Synthesis-Understanding Interaction
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024)
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024)
AgentScope: A Flexible yet Robust Multi-Agent Platform
von: Gao, Dawei, et al.
Veröffentlicht: (2024)
von: Gao, Dawei, et al.
Veröffentlicht: (2024)
Understanding Byzantine Robustness in Federated Learning with A Black-box Server
von: Zhao, Fangyuan, et al.
Veröffentlicht: (2024)
von: Zhao, Fangyuan, et al.
Veröffentlicht: (2024)
The Stronger the Diffusion Model, the Easier the Backdoor: Data Poisoning to Induce Copyright Breaches Without Adjusting Finetuning Pipeline
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
Branch-and-Browse: Efficient and Controllable Web Exploration with Tree-Structured Reasoning and Action Memory
von: He, Shiqi, et al.
Veröffentlicht: (2025)
von: He, Shiqi, et al.
Veröffentlicht: (2025)
Agent-Oriented Planning in Multi-Agent Systems
von: Li, Ao, et al.
Veröffentlicht: (2024)
von: Li, Ao, et al.
Veröffentlicht: (2024)
Talk to Right Specialists: Iterative Routing in Multi-agent Systems for Question Answering
von: Wu, Feijie, et al.
Veröffentlicht: (2025)
von: Wu, Feijie, et al.
Veröffentlicht: (2025)
Towards Fast Safe Online Reinforcement Learning via Policy Finetuning
von: Chen, Keru, et al.
Veröffentlicht: (2024)
von: Chen, Keru, et al.
Veröffentlicht: (2024)
Holdout-Loss-Based Data Selection for LLM Finetuning via In-Context Learning
von: Zhang, Ling, et al.
Veröffentlicht: (2025)
von: Zhang, Ling, et al.
Veröffentlicht: (2025)
GenSim: A General Social Simulation Platform with Large Language Model based Agents
von: Tang, Jiakai, et al.
Veröffentlicht: (2024)
von: Tang, Jiakai, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Diversity as a Reward: Fine-Tuning LLMs on a Mixture of Domain-Undetermined Data
von: Ling, Zhenqing, et al.
Veröffentlicht: (2025) -
Data-Juicer Sandbox: A Feedback-Driven Suite for Multimodal Data-Model Co-development
von: Chen, Daoyuan, et al.
Veröffentlicht: (2024) -
Img-Diff: Contrastive Data Synthesis for Multimodal Large Language Models
von: Jiao, Qirui, et al.
Veröffentlicht: (2024) -
MindGYM: What Matters in Question Synthesis for Thinking-Centric Fine-Tuning?
von: Xu, Zhe, et al.
Veröffentlicht: (2025) -
Designing Algorithms Empowered by Language Models: An Analytical Framework, Case Studies, and Insights
von: Chen, Yanxi, et al.
Veröffentlicht: (2024)