Rethinking Conventional Wisdom in Machine Learning: From Generalization to Scaling
Fuente:
arXiv
Saved in:
| Main Author: | Xiao, Lechao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scaling Collapse Reveals Universal Dynamics in Compute-Optimally Trained Neural Networks
by: Qiu, Shikai, et al.
Published: (2025)
by: Qiu, Shikai, et al.
Published: (2025)
4+3 Phases of Compute-Optimal Neural Scaling Laws
by: Paquette, Elliot, et al.
Published: (2024)
by: Paquette, Elliot, et al.
Published: (2024)
Magic Words or Methodical Work? Challenging Conventional Wisdom in LLM-Based Political Text Annotation
by: McLaren, Lorcan, et al.
Published: (2026)
by: McLaren, Lorcan, et al.
Published: (2026)
Binding Affinity Prediction: From Conventional to Machine Learning-Based Approaches
by: Liu, Xuefeng, et al.
Published: (2024)
by: Liu, Xuefeng, et al.
Published: (2024)
Rethinking Explainable Machine Learning as Applied Statistics
by: Bordt, Sebastian, et al.
Published: (2024)
by: Bordt, Sebastian, et al.
Published: (2024)
Interactive Machine Learning: From Theory to Scale
by: Zhu, Yinglun
Published: (2025)
by: Zhu, Yinglun
Published: (2025)
Text Serialization and Their Relationship with the Conventional Paradigms of Tabular Machine Learning
by: Ono, Kyoka, et al.
Published: (2024)
by: Ono, Kyoka, et al.
Published: (2024)
The Wisdom of the Crowd: High-Fidelity Classification of Cyber-Attacks and Faults in Power Systems Using Ensemble and Machine Learning
by: Abukhousa, Emad, et al.
Published: (2025)
by: Abukhousa, Emad, et al.
Published: (2025)
Rethinking Robustness in Machine Learning: A Posterior Agreement Approach
by: Carvalho, João Borges S., et al.
Published: (2025)
by: Carvalho, João Borges S., et al.
Published: (2025)
Rethinking and Recomputing the Value of Machine Learning Models
by: Sayin, Burcu, et al.
Published: (2022)
by: Sayin, Burcu, et al.
Published: (2022)
Rethinking Federated Learning Over the Air: The Blessing of Scaling Up
by: Zhu, Jiaqi, et al.
Published: (2025)
by: Zhu, Jiaqi, et al.
Published: (2025)
Wisdom and Delusion of LLM Ensembles for Code Generation and Repair
by: Vallecillos-Ruiz, Fernando, et al.
Published: (2025)
by: Vallecillos-Ruiz, Fernando, et al.
Published: (2025)
Rethinking the initialization of Momentum in Federated Learning with Heterogeneous Data
by: Xiao, Chenguang, et al.
Published: (2024)
by: Xiao, Chenguang, et al.
Published: (2024)
From Scaling to Structured Expressivity: Rethinking Transformers for CTR Prediction
by: Yan, Bencheng, et al.
Published: (2025)
by: Yan, Bencheng, et al.
Published: (2025)
Position: Why We Must Rethink Empirical Research in Machine Learning
by: Herrmann, Moritz, et al.
Published: (2024)
by: Herrmann, Moritz, et al.
Published: (2024)
Scaling Exponents Across Parameterizations and Optimizers
by: Everett, Katie, et al.
Published: (2024)
by: Everett, Katie, et al.
Published: (2024)
Rethinking the Personalized Relaxed Initialization in the Federated Learning: Consistency and Generalization
by: Shen, Li, et al.
Published: (2026)
by: Shen, Li, et al.
Published: (2026)
Rethinking Generative Image Pretraining: How Far Are We From Scaling Up Next-Pixel Prediction?
by: Yan, Xinchen, et al.
Published: (2025)
by: Yan, Xinchen, et al.
Published: (2025)
From Absolute to Relative: Rethinking Reward Shaping in Group-Based Reinforcement Learning
by: Niu, Wenzhe, et al.
Published: (2026)
by: Niu, Wenzhe, et al.
Published: (2026)
From Machine Learning to Machine Unlearning: Complying with GDPR's Right to be Forgotten while Maintaining Business Value of Predictive Models
by: Yang, Yuncong, et al.
Published: (2024)
by: Yang, Yuncong, et al.
Published: (2024)
Rethinking Time Series Domain Generalization via Structure-Stratified Calibration
by: Li, Jinyang, et al.
Published: (2026)
by: Li, Jinyang, et al.
Published: (2026)
ClinicRealm: Re-evaluating Large Language Models with Conventional Machine Learning for Non-Generative Clinical Prediction Tasks
by: Zhu, Yinghao, et al.
Published: (2024)
by: Zhu, Yinghao, et al.
Published: (2024)
Machine Learning for Synthetic Data Generation: A Review
by: Lu, Yingzhou, et al.
Published: (2023)
by: Lu, Yingzhou, et al.
Published: (2023)
Scaling Context Requires Rethinking Attention
by: Gelada, Carles, et al.
Published: (2025)
by: Gelada, Carles, et al.
Published: (2025)
SynTSBench: Rethinking Temporal Pattern Learning in Deep Learning Models for Time Series
by: Tan, Qitai, et al.
Published: (2025)
by: Tan, Qitai, et al.
Published: (2025)
From Reward-Free Representations to Preferences: Rethinking Offline Preference-Based Reinforcement Learning
by: Yang, Jun-Jie, et al.
Published: (2026)
by: Yang, Jun-Jie, et al.
Published: (2026)
Rethinking Momentum Knowledge Distillation in Online Continual Learning
by: Michel, Nicolas, et al.
Published: (2023)
by: Michel, Nicolas, et al.
Published: (2023)
Neuro-Symbolic Traders: Assessing the Wisdom of AI Crowds in Markets
by: Stillman, Namid R., et al.
Published: (2024)
by: Stillman, Namid R., et al.
Published: (2024)
Reward Modeling with Ordinal Feedback: Wisdom of the Crowd
by: Liu, Shang, et al.
Published: (2024)
by: Liu, Shang, et al.
Published: (2024)
The Shape of Wisdom: Decision Trajectories in Language Models
by: Rana, Shailesh
Published: (2026)
by: Rana, Shailesh
Published: (2026)
TokenFormer: Rethinking Transformer Scaling with Tokenized Model Parameters
by: Wang, Haiyang, et al.
Published: (2024)
by: Wang, Haiyang, et al.
Published: (2024)
Rethinking Language Model Scaling under Transferable Hypersphere Optimization
by: Ren, Liliang, et al.
Published: (2026)
by: Ren, Liliang, et al.
Published: (2026)
Rethinking Machine Unlearning for Large Language Models
by: Liu, Sijia, et al.
Published: (2024)
by: Liu, Sijia, et al.
Published: (2024)
Scaling Laws for the Value of Individual Data Points in Machine Learning
by: Covert, Ian, et al.
Published: (2024)
by: Covert, Ian, et al.
Published: (2024)
Rethinking Machine Unlearning: Models Designed to Forget via Key Deletion
by: Laguna, Sonia, et al.
Published: (2026)
by: Laguna, Sonia, et al.
Published: (2026)
MaLV-OS: Rethinking the Operating System Architecture for Machine Learning in Virtualized Clouds
by: Bitchebe, Stella, et al.
Published: (2025)
by: Bitchebe, Stella, et al.
Published: (2025)
Kinetics: Rethinking Test-Time Scaling Laws
by: Sadhukhan, Ranajoy, et al.
Published: (2025)
by: Sadhukhan, Ranajoy, et al.
Published: (2025)
Wisdom of Committee: Distilling from Foundation Model to Specialized Application Model
by: Liu, Zichang, et al.
Published: (2024)
by: Liu, Zichang, et al.
Published: (2024)
Bias as a Virtue: Rethinking Generalization under Distribution Shifts
by: Chen, Ruixuan, et al.
Published: (2025)
by: Chen, Ruixuan, et al.
Published: (2025)
The Wisdom of Many Queries: Complexity-Diversity Principle for Dense Retriever Training
by: Feng, Xincan, et al.
Published: (2026)
by: Feng, Xincan, et al.
Published: (2026)
Similar Items
-
Scaling Collapse Reveals Universal Dynamics in Compute-Optimally Trained Neural Networks
by: Qiu, Shikai, et al.
Published: (2025) -
4+3 Phases of Compute-Optimal Neural Scaling Laws
by: Paquette, Elliot, et al.
Published: (2024) -
Magic Words or Methodical Work? Challenging Conventional Wisdom in LLM-Based Political Text Annotation
by: McLaren, Lorcan, et al.
Published: (2026) -
Binding Affinity Prediction: From Conventional to Machine Learning-Based Approaches
by: Liu, Xuefeng, et al.
Published: (2024) -
Rethinking Explainable Machine Learning as Applied Statistics
by: Bordt, Sebastian, et al.
Published: (2024)