Investigating Data Pruning for Pretraining Biological Foundation Models at Scale
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Yifan, Jiang, Jiyue, Ye, Xichen, Wang, Yiqi, Zhou, Chang, Xu, Yitao, Chen, Jiayang, Hu, He, Zhang, Weizhong, Jin, Cheng, Yuan, Jiao, Li, Yu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimized Gradient Clipping for Noisy Label Learning
by: Ye, Xichen, et al.
Published: (2024)
by: Ye, Xichen, et al.
Published: (2024)
Active Negative Loss: A Robust Framework for Learning with Noisy Labels
by: Ye, Xichen, et al.
Published: (2024)
by: Ye, Xichen, et al.
Published: (2024)
Stable Long-Horizon Neural ODE Reduced-Order Models via Learned Feedback for Biological Growth and Remodeling
by: Laudo, Joel, et al.
Published: (2026)
by: Laudo, Joel, et al.
Published: (2026)
Epidemic Information Extraction for Event-Based Surveillance using Large Language Models
by: Consoli, Sergio, et al.
Published: (2024)
by: Consoli, Sergio, et al.
Published: (2024)
Stock Movement Prediction with Multimodal Stable Fusion via Gated Cross-Attention Mechanism
by: Zong, Chang, et al.
Published: (2024)
by: Zong, Chang, et al.
Published: (2024)
Deep learning algorithms for solving high dimensional nonlinear backward stochastic differential equations
by: Kapllani, Lorenc, et al.
Published: (2020)
by: Kapllani, Lorenc, et al.
Published: (2020)
Revisiting Energy-Based Model for Out-of-Distribution Detection
by: Wu, Yifan, et al.
Published: (2024)
by: Wu, Yifan, et al.
Published: (2024)
How Pruning Reshapes Features: Sparse Autoencoder Analysis of Weight-Pruned Language Models
by: Borobia, Hector, et al.
Published: (2026)
by: Borobia, Hector, et al.
Published: (2026)
ElliottAgents: A Natural Language-Driven Multi-Agent System for Stock Market Analysis and Prediction
by: Chudziak, Jarosław A., et al.
Published: (2025)
by: Chudziak, Jarosław A., et al.
Published: (2025)
Integrating Traditional Technical Analysis with AI: A Multi-Agent LLM-Based Approach to Stock Market Forecasting
by: Wawer, Michał, et al.
Published: (2025)
by: Wawer, Michał, et al.
Published: (2025)
Tool-Genesis: A Task-Driven Tool Creation Benchmark for Self-Evolving Language Agent
by: Xia, Bowei, et al.
Published: (2026)
by: Xia, Bowei, et al.
Published: (2026)
Hyperbox Mixture Regression for Process Performance Prediction in Antibody Production
by: Nik-Khorasani, Ali, et al.
Published: (2024)
by: Nik-Khorasani, Ali, et al.
Published: (2024)
Physics-based deep kernel learning for parameter estimation in high dimensional PDEs
by: Yan, Weihao, et al.
Published: (2025)
by: Yan, Weihao, et al.
Published: (2025)
Extracting Sentence Embeddings from Pretrained Transformer Models
by: Stankevičius, Lukas, et al.
Published: (2024)
by: Stankevičius, Lukas, et al.
Published: (2024)
Spectra: Surprising Effectiveness of Pretraining Ternary Language Models at Scale
by: Kaushal, Ayush, et al.
Published: (2024)
by: Kaushal, Ayush, et al.
Published: (2024)
Generative Pretrained Embedding and Hierarchical Irregular Time Series Representation for Daily Living Activity Recognition
by: Bouchabou, Damien, et al.
Published: (2024)
by: Bouchabou, Damien, et al.
Published: (2024)
Rethinking Visual Intelligence: Insights from Video Pretraining
by: Acuaviva, Pablo, et al.
Published: (2025)
by: Acuaviva, Pablo, et al.
Published: (2025)
Stage-wise Dynamics of Classifier-Free Guidance in Diffusion Models
by: Jin, Cheng, et al.
Published: (2025)
by: Jin, Cheng, et al.
Published: (2025)
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
by: Mitchell, Rupert, et al.
Published: (2025)
by: Mitchell, Rupert, et al.
Published: (2025)
Semantic Convergence: Investigating Shared Representations Across Scaled LLMs
by: Son, Daniel, et al.
Published: (2025)
by: Son, Daniel, et al.
Published: (2025)
Realizing Scaling Laws in Recommender Systems: A Foundation-Expert Paradigm for Hyperscale Model Deployment
by: Li, Dai, et al.
Published: (2025)
by: Li, Dai, et al.
Published: (2025)
Filtered not Mixed: Stochastic Filtering-Based Online Gating for Mixture of Large Language Models
by: Saqur, Raeid, et al.
Published: (2024)
by: Saqur, Raeid, et al.
Published: (2024)
ValueBlindBench: Agreement-Gated Stress Testing of LLM-Judged Investment Rationales Before Returns Are Observable
by: Chang, Sidi, et al.
Published: (2026)
by: Chang, Sidi, et al.
Published: (2026)
A Comparative Study of Feature Selection in Tsetlin Machines
by: Halenka, Vojtech, et al.
Published: (2025)
by: Halenka, Vojtech, et al.
Published: (2025)
DeepPersona: A Generative Engine for Scaling Deep Synthetic Personas
by: Wang, Zhen, et al.
Published: (2025)
by: Wang, Zhen, et al.
Published: (2025)
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
by: Chahine, Makram, et al.
Published: (2024)
by: Chahine, Makram, et al.
Published: (2024)
FedWCM: Unleashing the Potential of Momentum-based Federated Learning in Long-Tailed Scenarios
by: Li, Tianle, et al.
Published: (2025)
by: Li, Tianle, et al.
Published: (2025)
Asset Pricing in Pre-trained Transformer
by: Lai, Shanyan
Published: (2025)
by: Lai, Shanyan
Published: (2025)
Law-Strength Frontiers and a No-Free-Lunch Result for Law-Seeking Reinforcement Learning on Volatility Law Manifolds
by: Zhang, Jian'an
Published: (2025)
by: Zhang, Jian'an
Published: (2025)
Compressive Meta-Learning
by: Montserrat, Daniel Mas, et al.
Published: (2025)
by: Montserrat, Daniel Mas, et al.
Published: (2025)
Two Is Better Than One: Rotations Scale LoRAs
by: Guo, Hongcan, et al.
Published: (2025)
by: Guo, Hongcan, et al.
Published: (2025)
Foundation Models as World Models: A Foundational Study in Text-Based GridWorlds
by: Sasso, Remo, et al.
Published: (2025)
by: Sasso, Remo, et al.
Published: (2025)
Understanding the Limits of Deep Tabular Methods with Temporal Shift
by: Cai, Hao-Run, et al.
Published: (2025)
by: Cai, Hao-Run, et al.
Published: (2025)
Feature-aware Modulation for Learning from Temporal Tabular Data
by: Cai, Hao-Run, et al.
Published: (2025)
by: Cai, Hao-Run, et al.
Published: (2025)
Adaptive Weighting in Knowledge Distillation: An Axiomatic Framework for Multi-Scale Teacher Ensemble Optimization
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
by: Sarkar, Nilesh, et al.
Published: (2026)
by: Sarkar, Nilesh, et al.
Published: (2026)
Evaluating Model-Agnostic Meta-Learning on MetaWorld ML10 Benchmark: Fast Adaptation in Robotic Manipulation Tasks
by: Atamuradov, Sanjar
Published: (2025)
by: Atamuradov, Sanjar
Published: (2025)
Multi-Scale Graph Learning for Anti-Sparse Downscaling
by: Fan, Yingda, et al.
Published: (2025)
by: Fan, Yingda, et al.
Published: (2025)
CulinaryCut-VLAP: A Vision-Language-Action-Physics Framework for Food Cutting via a Force-Aware Material Point Method
by: Koh, Hyunseo, et al.
Published: (2026)
by: Koh, Hyunseo, et al.
Published: (2026)
Data-Driven Preference Sampling for Pareto Front Learning
by: Ye, Rongguang, et al.
Published: (2024)
by: Ye, Rongguang, et al.
Published: (2024)
Similar Items
-
Optimized Gradient Clipping for Noisy Label Learning
by: Ye, Xichen, et al.
Published: (2024) -
Active Negative Loss: A Robust Framework for Learning with Noisy Labels
by: Ye, Xichen, et al.
Published: (2024) -
Stable Long-Horizon Neural ODE Reduced-Order Models via Learned Feedback for Biological Growth and Remodeling
by: Laudo, Joel, et al.
Published: (2026) -
Epidemic Information Extraction for Event-Based Surveillance using Large Language Models
by: Consoli, Sergio, et al.
Published: (2024) -
Stock Movement Prediction with Multimodal Stable Fusion via Gated Cross-Attention Mechanism
by: Zong, Chang, et al.
Published: (2024)