Investigating Data Pruning for Pretraining Biological Foundation Models at Scale
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Yifan, Jiang, Jiyue, Ye, Xichen, Wang, Yiqi, Zhou, Chang, Xu, Yitao, Chen, Jiayang, Hu, He, Zhang, Weizhong, Jin, Cheng, Yuan, Jiao, Li, Yu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Optimized Gradient Clipping for Noisy Label Learning
di: Ye, Xichen, et al.
Pubblicazione: (2024)
di: Ye, Xichen, et al.
Pubblicazione: (2024)
Active Negative Loss: A Robust Framework for Learning with Noisy Labels
di: Ye, Xichen, et al.
Pubblicazione: (2024)
di: Ye, Xichen, et al.
Pubblicazione: (2024)
Stable Long-Horizon Neural ODE Reduced-Order Models via Learned Feedback for Biological Growth and Remodeling
di: Laudo, Joel, et al.
Pubblicazione: (2026)
di: Laudo, Joel, et al.
Pubblicazione: (2026)
Epidemic Information Extraction for Event-Based Surveillance using Large Language Models
di: Consoli, Sergio, et al.
Pubblicazione: (2024)
di: Consoli, Sergio, et al.
Pubblicazione: (2024)
Stock Movement Prediction with Multimodal Stable Fusion via Gated Cross-Attention Mechanism
di: Zong, Chang, et al.
Pubblicazione: (2024)
di: Zong, Chang, et al.
Pubblicazione: (2024)
Deep learning algorithms for solving high dimensional nonlinear backward stochastic differential equations
di: Kapllani, Lorenc, et al.
Pubblicazione: (2020)
di: Kapllani, Lorenc, et al.
Pubblicazione: (2020)
Revisiting Energy-Based Model for Out-of-Distribution Detection
di: Wu, Yifan, et al.
Pubblicazione: (2024)
di: Wu, Yifan, et al.
Pubblicazione: (2024)
How Pruning Reshapes Features: Sparse Autoencoder Analysis of Weight-Pruned Language Models
di: Borobia, Hector, et al.
Pubblicazione: (2026)
di: Borobia, Hector, et al.
Pubblicazione: (2026)
ElliottAgents: A Natural Language-Driven Multi-Agent System for Stock Market Analysis and Prediction
di: Chudziak, Jarosław A., et al.
Pubblicazione: (2025)
di: Chudziak, Jarosław A., et al.
Pubblicazione: (2025)
Integrating Traditional Technical Analysis with AI: A Multi-Agent LLM-Based Approach to Stock Market Forecasting
di: Wawer, Michał, et al.
Pubblicazione: (2025)
di: Wawer, Michał, et al.
Pubblicazione: (2025)
Tool-Genesis: A Task-Driven Tool Creation Benchmark for Self-Evolving Language Agent
di: Xia, Bowei, et al.
Pubblicazione: (2026)
di: Xia, Bowei, et al.
Pubblicazione: (2026)
Hyperbox Mixture Regression for Process Performance Prediction in Antibody Production
di: Nik-Khorasani, Ali, et al.
Pubblicazione: (2024)
di: Nik-Khorasani, Ali, et al.
Pubblicazione: (2024)
Physics-based deep kernel learning for parameter estimation in high dimensional PDEs
di: Yan, Weihao, et al.
Pubblicazione: (2025)
di: Yan, Weihao, et al.
Pubblicazione: (2025)
Extracting Sentence Embeddings from Pretrained Transformer Models
di: Stankevičius, Lukas, et al.
Pubblicazione: (2024)
di: Stankevičius, Lukas, et al.
Pubblicazione: (2024)
Spectra: Surprising Effectiveness of Pretraining Ternary Language Models at Scale
di: Kaushal, Ayush, et al.
Pubblicazione: (2024)
di: Kaushal, Ayush, et al.
Pubblicazione: (2024)
Generative Pretrained Embedding and Hierarchical Irregular Time Series Representation for Daily Living Activity Recognition
di: Bouchabou, Damien, et al.
Pubblicazione: (2024)
di: Bouchabou, Damien, et al.
Pubblicazione: (2024)
Rethinking Visual Intelligence: Insights from Video Pretraining
di: Acuaviva, Pablo, et al.
Pubblicazione: (2025)
di: Acuaviva, Pablo, et al.
Pubblicazione: (2025)
Stage-wise Dynamics of Classifier-Free Guidance in Diffusion Models
di: Jin, Cheng, et al.
Pubblicazione: (2025)
di: Jin, Cheng, et al.
Pubblicazione: (2025)
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
di: Mitchell, Rupert, et al.
Pubblicazione: (2025)
di: Mitchell, Rupert, et al.
Pubblicazione: (2025)
Semantic Convergence: Investigating Shared Representations Across Scaled LLMs
di: Son, Daniel, et al.
Pubblicazione: (2025)
di: Son, Daniel, et al.
Pubblicazione: (2025)
Realizing Scaling Laws in Recommender Systems: A Foundation-Expert Paradigm for Hyperscale Model Deployment
di: Li, Dai, et al.
Pubblicazione: (2025)
di: Li, Dai, et al.
Pubblicazione: (2025)
Filtered not Mixed: Stochastic Filtering-Based Online Gating for Mixture of Large Language Models
di: Saqur, Raeid, et al.
Pubblicazione: (2024)
di: Saqur, Raeid, et al.
Pubblicazione: (2024)
ValueBlindBench: Agreement-Gated Stress Testing of LLM-Judged Investment Rationales Before Returns Are Observable
di: Chang, Sidi, et al.
Pubblicazione: (2026)
di: Chang, Sidi, et al.
Pubblicazione: (2026)
A Comparative Study of Feature Selection in Tsetlin Machines
di: Halenka, Vojtech, et al.
Pubblicazione: (2025)
di: Halenka, Vojtech, et al.
Pubblicazione: (2025)
DeepPersona: A Generative Engine for Scaling Deep Synthetic Personas
di: Wang, Zhen, et al.
Pubblicazione: (2025)
di: Wang, Zhen, et al.
Pubblicazione: (2025)
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
di: Chahine, Makram, et al.
Pubblicazione: (2024)
di: Chahine, Makram, et al.
Pubblicazione: (2024)
FedWCM: Unleashing the Potential of Momentum-based Federated Learning in Long-Tailed Scenarios
di: Li, Tianle, et al.
Pubblicazione: (2025)
di: Li, Tianle, et al.
Pubblicazione: (2025)
Asset Pricing in Pre-trained Transformer
di: Lai, Shanyan
Pubblicazione: (2025)
di: Lai, Shanyan
Pubblicazione: (2025)
Law-Strength Frontiers and a No-Free-Lunch Result for Law-Seeking Reinforcement Learning on Volatility Law Manifolds
di: Zhang, Jian'an
Pubblicazione: (2025)
di: Zhang, Jian'an
Pubblicazione: (2025)
Compressive Meta-Learning
di: Montserrat, Daniel Mas, et al.
Pubblicazione: (2025)
di: Montserrat, Daniel Mas, et al.
Pubblicazione: (2025)
Two Is Better Than One: Rotations Scale LoRAs
di: Guo, Hongcan, et al.
Pubblicazione: (2025)
di: Guo, Hongcan, et al.
Pubblicazione: (2025)
Foundation Models as World Models: A Foundational Study in Text-Based GridWorlds
di: Sasso, Remo, et al.
Pubblicazione: (2025)
di: Sasso, Remo, et al.
Pubblicazione: (2025)
Understanding the Limits of Deep Tabular Methods with Temporal Shift
di: Cai, Hao-Run, et al.
Pubblicazione: (2025)
di: Cai, Hao-Run, et al.
Pubblicazione: (2025)
Feature-aware Modulation for Learning from Temporal Tabular Data
di: Cai, Hao-Run, et al.
Pubblicazione: (2025)
di: Cai, Hao-Run, et al.
Pubblicazione: (2025)
Adaptive Weighting in Knowledge Distillation: An Axiomatic Framework for Multi-Scale Teacher Ensemble Optimization
di: Flouro, Aaron R., et al.
Pubblicazione: (2026)
di: Flouro, Aaron R., et al.
Pubblicazione: (2026)
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
di: Sarkar, Nilesh, et al.
Pubblicazione: (2026)
di: Sarkar, Nilesh, et al.
Pubblicazione: (2026)
Evaluating Model-Agnostic Meta-Learning on MetaWorld ML10 Benchmark: Fast Adaptation in Robotic Manipulation Tasks
di: Atamuradov, Sanjar
Pubblicazione: (2025)
di: Atamuradov, Sanjar
Pubblicazione: (2025)
Multi-Scale Graph Learning for Anti-Sparse Downscaling
di: Fan, Yingda, et al.
Pubblicazione: (2025)
di: Fan, Yingda, et al.
Pubblicazione: (2025)
CulinaryCut-VLAP: A Vision-Language-Action-Physics Framework for Food Cutting via a Force-Aware Material Point Method
di: Koh, Hyunseo, et al.
Pubblicazione: (2026)
di: Koh, Hyunseo, et al.
Pubblicazione: (2026)
Data-Driven Preference Sampling for Pareto Front Learning
di: Ye, Rongguang, et al.
Pubblicazione: (2024)
di: Ye, Rongguang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Optimized Gradient Clipping for Noisy Label Learning
di: Ye, Xichen, et al.
Pubblicazione: (2024) -
Active Negative Loss: A Robust Framework for Learning with Noisy Labels
di: Ye, Xichen, et al.
Pubblicazione: (2024) -
Stable Long-Horizon Neural ODE Reduced-Order Models via Learned Feedback for Biological Growth and Remodeling
di: Laudo, Joel, et al.
Pubblicazione: (2026) -
Epidemic Information Extraction for Event-Based Surveillance using Large Language Models
di: Consoli, Sergio, et al.
Pubblicazione: (2024) -
Stock Movement Prediction with Multimodal Stable Fusion via Gated Cross-Attention Mechanism
di: Zong, Chang, et al.
Pubblicazione: (2024)