daVinci-LLM:Towards the Science of Pretraining
Fuente:
arXiv
Saved in:
| Main Authors: | Qin, Yiwei, Liu, Yixiu, Mi, Tiantian, Xie, Muhang, Huang, Zhen, Si, Weiye, Lu, Pengrui, Feng, Siyuan, Wu, Xia, Liu, Liming, Luo, Ye, Hou, Jinlong, Guo, Qipeng, Qiao, Yu, Liu, Pengfei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Data Darwinism Part II: DataEvolve -- AI can Autonomously Evolve Pretraining Data Curation
by: Mi, Tiantian, et al.
Published: (2026)
by: Mi, Tiantian, et al.
Published: (2026)
daVinci-Dev: Agent-native Mid-training for Software Engineering
by: Zeng, Ji, et al.
Published: (2026)
by: Zeng, Ji, et al.
Published: (2026)
daVinci-Env: Open SWE Environment Synthesis at Scale
by: Fu, Dayuan, et al.
Published: (2026)
by: Fu, Dayuan, et al.
Published: (2026)
Data Darwinism Part I: Unlocking the Value of Scientific Data for Pre-training
by: Qin, Yiwei, et al.
Published: (2026)
by: Qin, Yiwei, et al.
Published: (2026)
daVinci-Agency: Unlocking Long-Horizon Agency Data-Efficiently
by: Jiang, Mohan, et al.
Published: (2026)
by: Jiang, Mohan, et al.
Published: (2026)
DIVE: Diversified Iterative Self-Improvement
by: Qin, Yiwei, et al.
Published: (2025)
by: Qin, Yiwei, et al.
Published: (2025)
ASI-Evolve: AI Accelerates AI
by: Xu, Weixian, et al.
Published: (2026)
by: Xu, Weixian, et al.
Published: (2026)
O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?
by: Huang, Zhen, et al.
Published: (2024)
by: Huang, Zhen, et al.
Published: (2024)
O1 Replication Journey: A Strategic Progress Report -- Part 1
by: Qin, Yiwei, et al.
Published: (2024)
by: Qin, Yiwei, et al.
Published: (2024)
AlphaGo Moment for Model Architecture Discovery
by: Liu, Yixiu, et al.
Published: (2025)
by: Liu, Yixiu, et al.
Published: (2025)
InnovatorBench: Evaluating Agents' Ability to Conduct Innovative LLM Research
by: Wu, Yunze, et al.
Published: (2025)
by: Wu, Yunze, et al.
Published: (2025)
SAFETY-J: Evaluating Safety with Critique
by: Liu, Yixiu, et al.
Published: (2024)
by: Liu, Yixiu, et al.
Published: (2024)
MathPile: A Billion-Token-Scale Pretraining Corpus for Math
by: Wang, Zengzhi, et al.
Published: (2023)
by: Wang, Zengzhi, et al.
Published: (2023)
ResearcherBench: Evaluating Deep AI Research Systems on the Frontiers of Scientific Inquiry
by: Xu, Tianze, et al.
Published: (2025)
by: Xu, Tianze, et al.
Published: (2025)
Interaction as Intelligence Part II: Asynchronous Human-Agent Rollout for Long-Horizon Task Training
by: Fu, Dayuan, et al.
Published: (2025)
by: Fu, Dayuan, et al.
Published: (2025)
OmniVinci: Enhancing Architecture and Data for Omni-Modal Understanding LLM
by: Ye, Hanrong, et al.
Published: (2025)
by: Ye, Hanrong, et al.
Published: (2025)
In-Context Learning May Not Elicit Trustworthy Reasoning: A-Not-B Errors in Pretrained Language Models
by: Han, Pengrui, et al.
Published: (2024)
by: Han, Pengrui, et al.
Published: (2024)
Paper Copilot: A Self-Evolving and Efficient LLM System for Personalized Academic Assistance
by: Lin, Guanyu, et al.
Published: (2024)
by: Lin, Guanyu, et al.
Published: (2024)
A Training-free LLM Framework with Interaction between Contextually Related Subtasks in Solving Complex Tasks
by: Liu, Hongjia, et al.
Published: (2025)
by: Liu, Hongjia, et al.
Published: (2025)
Mesoporous Anti‐Perovskite CuNi 3 N for Sustainable Formate Electrosynthesis from Complete Electrooxidation of Biomass Glucose
by: Pengfei Liu, et al.
Published: (2026)
by: Pengfei Liu, et al.
Published: (2026)
Mesoporous Anti‐Perovskite CuNi 3 N for Sustainable Formate Electrosynthesis from Complete Electrooxidation of Biomass Glucose
by: Pengfei Liu, et al.
Published: (2026)
by: Pengfei Liu, et al.
Published: (2026)
ProjDevBench: Benchmarking AI Coding Agents on End-to-End Project Development
by: Lu, Pengrui, et al.
Published: (2026)
by: Lu, Pengrui, et al.
Published: (2026)
Generative AI Act II: Test Time Scaling Drives Cognition Engineering
by: Xia, Shijie, et al.
Published: (2025)
by: Xia, Shijie, et al.
Published: (2025)
Taming Request Imbalance: SLO-Aware Scheduling for Disaggregated LLM Inference
by: Wang, Qipeng
Published: (2026)
by: Wang, Qipeng
Published: (2026)
Strong Teacher Not Needed? On Distillation in LLM Pretraining
by: Lu, Taiming, et al.
Published: (2026)
by: Lu, Taiming, et al.
Published: (2026)
DistFlow: A Fully Distributed RL Framework for Scalable and Efficient LLM Post-Training
by: Wang, Zhixin, et al.
Published: (2025)
by: Wang, Zhixin, et al.
Published: (2025)
Accelerated and Guided Zn 2+ Diffusion via Polarized Interface Engineering Toward High Performance Wearable Zinc‐Ion Batteries
by: Yuhang Zhang, et al.
Published: (2024)
by: Yuhang Zhang, et al.
Published: (2024)
Understanding by Reconstruction: Reversing the Software Development Process for LLM Pretraining
by: Zeng, Zhiyuan, et al.
Published: (2026)
by: Zeng, Zhiyuan, et al.
Published: (2026)
Multi-objective Reinforcement Learning with Nonlinear Preferences: Provable Approximation for Maximizing Expected Scalarized Return
by: Peng, Nianli, et al.
Published: (2023)
by: Peng, Nianli, et al.
Published: (2023)
One-Bit Model Aggregation for Differentially Private and Byzantine-Robust Personalized Federated Learning
by: Lan, Muhang, et al.
Published: (2025)
by: Lan, Muhang, et al.
Published: (2025)
Craw4LLM: Efficient Web Crawling for LLM Pretraining
by: Yu, Shi, et al.
Published: (2025)
by: Yu, Shi, et al.
Published: (2025)
Integrative analysis unveils ECM signatures and pathways driving hepatocellular carcinoma progression: A multi‐omics approach and prognostic model development
by: Zhen Liu, et al.
Published: (2024)
by: Zhen Liu, et al.
Published: (2024)
Parameter-Efficient Fine-Tuning Medical Multimodal Large Language Models for Medical Visual Grounding
by: He, Jinlong, et al.
Published: (2024)
by: He, Jinlong, et al.
Published: (2024)
OlympicArena Medal Ranks: Who Is the Most Intelligent AI So Far?
by: Huang, Zhen, et al.
Published: (2024)
by: Huang, Zhen, et al.
Published: (2024)
Hydrothermal synthesis of highly cross‐linked PtZn@Silicalite‐1 structured catalysts for propane dehydrogenation
by: Liming Xia, et al.
Published: (2024)
by: Liming Xia, et al.
Published: (2024)
Lethe: Layer- and Time-Adaptive KV Cache Pruning for Reasoning-Intensive LLM Serving
by: Zeng, Hui, et al.
Published: (2025)
by: Zeng, Hui, et al.
Published: (2025)
On Evaluating LLM Alignment by Evaluating LLMs as Judges
by: Liu, Yixin, et al.
Published: (2025)
by: Liu, Yixin, et al.
Published: (2025)
Safety Index Synthesis with State-dependent Control Space
by: Chen, Rui, et al.
Published: (2023)
by: Chen, Rui, et al.
Published: (2023)
Implicit Safe Set Algorithm for Provably Safe Reinforcement Learning
by: Zhao, Weiye, et al.
Published: (2024)
by: Zhao, Weiye, et al.
Published: (2024)
Physics-Aware Combinatorial Assembly Sequence Planning using Data-free Action Masking
by: Liu, Ruixuan, et al.
Published: (2024)
by: Liu, Ruixuan, et al.
Published: (2024)
Similar Items
-
Data Darwinism Part II: DataEvolve -- AI can Autonomously Evolve Pretraining Data Curation
by: Mi, Tiantian, et al.
Published: (2026) -
daVinci-Dev: Agent-native Mid-training for Software Engineering
by: Zeng, Ji, et al.
Published: (2026) -
daVinci-Env: Open SWE Environment Synthesis at Scale
by: Fu, Dayuan, et al.
Published: (2026) -
Data Darwinism Part I: Unlocking the Value of Scientific Data for Pre-training
by: Qin, Yiwei, et al.
Published: (2026) -
daVinci-Agency: Unlocking Long-Horizon Agency Data-Efficiently
by: Jiang, Mohan, et al.
Published: (2026)