Efficient Pretraining Data Selection for Language Models via Multi-Actor Collaboration
Fuente:
arXiv
Saved in:
| Main Authors: | Bai, Tianyi, Yang, Ling, Wong, Zhen Hao, Sun, Fupeng, Peng, Jiahui, Zhuang, Xinlin, Zhang, Chi, Wu, Lijun, Qiu, Jiantao, Zhang, Wentao, Yuan, Binhang, He, Conghui |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Meta-rater: A Multi-dimensional Data Selection Method for Pre-training Language Models
by: Zhuang, Xinlin, et al.
Published: (2025)
by: Zhuang, Xinlin, et al.
Published: (2025)
Hallucination at a Glance: Controlled Visual Edits and Fine-Grained Multimodal Learning
by: Bai, Tianyi, et al.
Published: (2025)
by: Bai, Tianyi, et al.
Published: (2025)
Multi-Step Visual Reasoning with Visual Tokens Scaling and Verification
by: Bai, Tianyi, et al.
Published: (2025)
by: Bai, Tianyi, et al.
Published: (2025)
Harnessing Diversity for Important Data Selection in Pretraining Large Language Models
by: Zhang, Chi, et al.
Published: (2024)
by: Zhang, Chi, et al.
Published: (2024)
VADE: Variance-Aware Dynamic Sampling via Online Sample-Level Difficulty Estimation for Multimodal RL
by: Hu, Zengjie, et al.
Published: (2025)
by: Hu, Zengjie, et al.
Published: (2025)
Topic Over Source: The Key to Effective Data Mixing for Language Models Pre-training
by: Peng, Jiahui, et al.
Published: (2025)
by: Peng, Jiahui, et al.
Published: (2025)
Synthesizing Multimodal Geometry Datasets from Scratch and Enabling Visual Alignment via Plotting Code
by: Lin, Haobo, et al.
Published: (2026)
by: Lin, Haobo, et al.
Published: (2026)
A Survey of Multimodal Large Language Model from A Data-centric Perspective
by: Bai, Tianyi, et al.
Published: (2024)
by: Bai, Tianyi, et al.
Published: (2024)
KeyVideoLLM: Towards Large-scale Video Keyframe Selection
by: Liang, Hao, et al.
Published: (2024)
by: Liang, Hao, et al.
Published: (2024)
Dripper: Token-Efficient Main HTML Extraction with a Lightweight LM
by: Liu, Mengjie, et al.
Published: (2025)
by: Liu, Mengjie, et al.
Published: (2025)
V3DB: Audit-on-Demand Zero-Knowledge Proofs for Verifiable Vector Search over Committed Snapshots
by: Qiu, Zipeng, et al.
Published: (2026)
by: Qiu, Zipeng, et al.
Published: (2026)
TAH-QUANT: Effective Activation Quantization in Pipeline Parallelism over Slow Network
by: He, Guangxin, et al.
Published: (2025)
by: He, Guangxin, et al.
Published: (2025)
SynthVLM: Towards High-Quality and Efficient Synthesis of Image-Caption Datasets for Vision-Language Models
by: Liu, Zheng, et al.
Published: (2024)
by: Liu, Zheng, et al.
Published: (2024)
TQA-Bench: Evaluating LLMs for Multi-Table Question Answering with Scalable Context and Symbolic Extension
by: Qiu, Zipeng, et al.
Published: (2024)
by: Qiu, Zipeng, et al.
Published: (2024)
Selective Prompt Anchoring for Code Generation
by: Tian, Yuan, et al.
Published: (2024)
by: Tian, Yuan, et al.
Published: (2024)
Group-Level Data Selection for Efficient Pretraining
by: Yu, Zichun, et al.
Published: (2025)
by: Yu, Zichun, et al.
Published: (2025)
BEATS: Optimizing LLM Mathematical Capabilities with BackVerify and Adaptive Disambiguate based Efficient Tree Search
by: Sun, Linzhuang, et al.
Published: (2024)
by: Sun, Linzhuang, et al.
Published: (2024)
On the simplest simply connected rational homology $7$-spheres that are not $2$-connected
by: Xu, Fupeng
Published: (2026)
by: Xu, Fupeng
Published: (2026)
Free circle actions on certain simply connected $7-$manifolds
by: Xu, Fupeng
Published: (2024)
by: Xu, Fupeng
Published: (2024)
A Global Dataset for Seasonal to Annual Forecasts of GRACE-like Gridded Terrestrial Water Storage (2010-2024)
by: Li, Fupeng
Published: (2025)
by: Li, Fupeng
Published: (2025)
On $7$-manifolds with $b_{2}=2$: diffeomorphism classification and nonconnected moduli spaces of positive Ricci curvature metrics
by: Xu, Fupeng
Published: (2025)
by: Xu, Fupeng
Published: (2025)
Prompt Inversion Attack against Collaborative Inference of Large Language Models
by: Qu, Wenjie, et al.
Published: (2025)
by: Qu, Wenjie, et al.
Published: (2025)
Actor-Critic Pretraining for Proximal Policy Optimization
by: Kernbach, Andreas, et al.
Published: (2026)
by: Kernbach, Andreas, et al.
Published: (2026)
Distributional Soft Actor-Critic with Harmonic Gradient for Safe and Efficient Autonomous Driving in Multi-lane Scenarios
by: Zhang, Feihong, et al.
Published: (2025)
by: Zhang, Feihong, et al.
Published: (2025)
Co-evolution of Nuclear Star Clusters and Massive Black Holes: Extreme Mass-Ratio Inspirals
by: Zhang, Fupeng, et al.
Published: (2025)
by: Zhang, Fupeng, et al.
Published: (2025)
Self-consistent Solutions of Evolving Nuclear Star Clusters with Two-Dimensional Monte-Carlo Dynamical Simulations
by: Zhang, Fupeng, et al.
Published: (2025)
by: Zhang, Fupeng, et al.
Published: (2025)
KCM: KAN-Based Collaboration Models Enhance Pretrained Large Models
by: Dai, Guangyu, et al.
Published: (2025)
by: Dai, Guangyu, et al.
Published: (2025)
Rational Design of Cobalt Phthalocyanine (CoPc)‐Anchored TiO 2 Nanorods for High‐Efficiency Selective Catalytic Oxidation
by: Simeng Zhu, et al.
Published: (2024)
by: Simeng Zhu, et al.
Published: (2024)
Hybrid Differential Reward: Combining Temporal Difference and Action Gradients for Efficient Multi-Agent Reinforcement Learning in Cooperative Driving
by: Han, Ye, et al.
Published: (2025)
by: Han, Ye, et al.
Published: (2025)
ProtLLM: An Interleaved Protein-Language LLM with Protein-as-Word Pre-Training
by: Zhuo, Le, et al.
Published: (2024)
by: Zhuo, Le, et al.
Published: (2024)
Two-Stage Constrained Actor-Critic for Short Video Recommendation
by: Cai, Qingpeng, et al.
Published: (2023)
by: Cai, Qingpeng, et al.
Published: (2023)
CE-LoRA: Computation-Efficient LoRA Fine-Tuning for Language Models
by: Chen, Guanduo, et al.
Published: (2025)
by: Chen, Guanduo, et al.
Published: (2025)
Pretraining in Actor-Critic Reinforcement Learning for Robot Locomotion
by: Fan, Jiale, et al.
Published: (2025)
by: Fan, Jiale, et al.
Published: (2025)
Efficient-Empathy: Towards Efficient and Effective Selection of Empathy Data
by: Sun, Linzhuang, et al.
Published: (2024)
by: Sun, Linzhuang, et al.
Published: (2024)
PASER: Post-Training Data Selection for Efficient Pruned Large Language Model Recovery
by: He, Bowei, et al.
Published: (2025)
by: He, Bowei, et al.
Published: (2025)
Understanding the Training and Generalization of Pretrained Transformer for Sequential Decision Making
by: Wang, Hanzhao, et al.
Published: (2024)
by: Wang, Hanzhao, et al.
Published: (2024)
IKnow: Instruction-Knowledge-Aware Continual Pretraining for Effective Domain Adaptation
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
Parallax: Efficient LLM Inference Service over Decentralized Environment
by: Tong, Chris, et al.
Published: (2025)
by: Tong, Chris, et al.
Published: (2025)
$k$NNProxy: Efficient Training-Free Proxy Alignment for Black-Box Zero-Shot LLM-Generated Text Detection
by: Wong, Kahim, et al.
Published: (2026)
by: Wong, Kahim, et al.
Published: (2026)
Towards Efficient LLM Grounding for Embodied Multi-Agent Collaboration
by: Zhang, Yang, et al.
Published: (2024)
by: Zhang, Yang, et al.
Published: (2024)
Similar Items
-
Meta-rater: A Multi-dimensional Data Selection Method for Pre-training Language Models
by: Zhuang, Xinlin, et al.
Published: (2025) -
Hallucination at a Glance: Controlled Visual Edits and Fine-Grained Multimodal Learning
by: Bai, Tianyi, et al.
Published: (2025) -
Multi-Step Visual Reasoning with Visual Tokens Scaling and Verification
by: Bai, Tianyi, et al.
Published: (2025) -
Harnessing Diversity for Important Data Selection in Pretraining Large Language Models
by: Zhang, Chi, et al.
Published: (2024) -
VADE: Variance-Aware Dynamic Sampling via Online Sample-Level Difficulty Estimation for Multimodal RL
by: Hu, Zengjie, et al.
Published: (2025)