Deep Learning is Not So Mysterious or Different
Fuente:
arXiv
Saved in:
| Main Author: | Wilson, Andrew Gordon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Response to Promises and Pitfalls of Deep Kernel Learning
by: Wilson, Andrew Gordon, et al.
Published: (2025)
by: Wilson, Andrew Gordon, et al.
Published: (2025)
Scalable and Flexible Causal Discovery with an Efficient Test for Adjacency
by: Amin, Alan Nawzad, et al.
Published: (2024)
by: Amin, Alan Nawzad, et al.
Published: (2024)
The No Free Lunch Theorem, Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learning
by: Goldblum, Micah, et al.
Published: (2023)
by: Goldblum, Micah, et al.
Published: (2023)
SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity
by: McFadden, Shae, et al.
Published: (2026)
by: McFadden, Shae, et al.
Published: (2026)
In-Context Clustering with Large Language Models
by: Wang, Ying, et al.
Published: (2025)
by: Wang, Ying, et al.
Published: (2025)
Controllable Prompt Tuning For Balancing Group Distributional Robustness
by: Phan, Hoang, et al.
Published: (2024)
by: Phan, Hoang, et al.
Published: (2024)
Why Masking Diffusion Works: Condition on the Jump Schedule for Improved Discrete Diffusion
by: Amin, Alan N., et al.
Published: (2025)
by: Amin, Alan N., et al.
Published: (2025)
Diverse Dictionary Learning
by: Zheng, Yujia, et al.
Published: (2026)
by: Zheng, Yujia, et al.
Published: (2026)
Simplifying Deep Temporal Difference Learning
by: Gallici, Matteo, et al.
Published: (2024)
by: Gallici, Matteo, et al.
Published: (2024)
Robust Testing for Deep Learning using Human Label Noise
by: Lim, Gordon, et al.
Published: (2024)
by: Lim, Gordon, et al.
Published: (2024)
Large Language Models Are Zero-Shot Time Series Forecasters
by: Gruver, Nate, et al.
Published: (2023)
by: Gruver, Nate, et al.
Published: (2023)
Detecting Semantic Backdoors in a Mystery Shopping Scenario
by: Berta, Arpad, et al.
Published: (2026)
by: Berta, Arpad, et al.
Published: (2026)
SoK: Behind the Accuracy of Complex Human Activity Recognition Using Deep Learning
by: Nguyen, Duc-Anh, et al.
Published: (2024)
by: Nguyen, Duc-Anh, et al.
Published: (2024)
"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents
by: Dispoto, Giovanni, et al.
Published: (2025)
by: Dispoto, Giovanni, et al.
Published: (2025)
Reveal the Mystery of DPO: The Connection between DPO and RL Algorithms
by: Su, Xuerui, et al.
Published: (2025)
by: Su, Xuerui, et al.
Published: (2025)
Training Flexible Models of Genetic Variant Effects from Functional Annotations using Accelerated Linear Algebra
by: Amin, Alan N., et al.
Published: (2025)
by: Amin, Alan N., et al.
Published: (2025)
A Study of Bayesian Neural Network Surrogates for Bayesian Optimization
by: Li, Yucen Lily, et al.
Published: (2023)
by: Li, Yucen Lily, et al.
Published: (2023)
The Lie Derivative for Measuring Learned Equivariance
by: Gruver, Nate, et al.
Published: (2022)
by: Gruver, Nate, et al.
Published: (2022)
Unraveling the Mystery of Scaling Laws: Part I
by: Su, Hui, et al.
Published: (2024)
by: Su, Hui, et al.
Published: (2024)
Hyperparameter Transfer Enables Consistent Gains of Matrix-Preconditioned Optimizers Across Scales
by: Qiu, Shikai, et al.
Published: (2025)
by: Qiu, Shikai, et al.
Published: (2025)
Small Batch Size Training for Language Models: When Vanilla SGD Works, and Why Gradient Accumulation Is Wasteful
by: Marek, Martin, et al.
Published: (2025)
by: Marek, Martin, et al.
Published: (2025)
Scaling Collapse Reveals Universal Dynamics in Compute-Optimally Trained Neural Networks
by: Qiu, Shikai, et al.
Published: (2025)
by: Qiu, Shikai, et al.
Published: (2025)
Compute Better Spent: Replacing Dense Layers with Structured Matrices
by: Qiu, Shikai, et al.
Published: (2024)
by: Qiu, Shikai, et al.
Published: (2024)
The Mystery of the Pathological Path-star Task for Language Models
by: Frydenlund, Arvid
Published: (2024)
by: Frydenlund, Arvid
Published: (2024)
SoK: Data Minimization in Machine Learning
by: Staab, Robin, et al.
Published: (2025)
by: Staab, Robin, et al.
Published: (2025)
SoK: What Makes Private Learning Unfair?
by: Yao, Kai, et al.
Published: (2025)
by: Yao, Kai, et al.
Published: (2025)
A Role of Environmental Complexity on Representation Learning in Deep Reinforcement Learning Agents
by: Liu, Andrew, et al.
Published: (2024)
by: Liu, Andrew, et al.
Published: (2024)
Customizing the Inductive Biases of Softmax Attention using Structured Matrices
by: Kuang, Yilun, et al.
Published: (2025)
by: Kuang, Yilun, et al.
Published: (2025)
Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay
by: Marek, Martin, et al.
Published: (2026)
by: Marek, Martin, et al.
Published: (2026)
Unlocking Tokens as Data Points for Generalization Bounds on Larger Language Models
by: Lotfi, Sanae, et al.
Published: (2024)
by: Lotfi, Sanae, et al.
Published: (2024)
Wind Power Prediction across Different Locations using Deep Domain Adaptive Learning
by: Sajol, Md Saiful Islam, et al.
Published: (2024)
by: Sajol, Md Saiful Islam, et al.
Published: (2024)
Deep Reinforcement Learning and The Tale of Two Temporal Difference Errors
by: Rojas, Juan Sebastian, et al.
Published: (2026)
by: Rojas, Juan Sebastian, et al.
Published: (2026)
SoD$^2$: Statically Optimizing Dynamic Deep Neural Network
by: Niu, Wei, et al.
Published: (2024)
by: Niu, Wei, et al.
Published: (2024)
Transferring Knowledge from Large Foundation Models to Small Downstream Models
by: Qiu, Shikai, et al.
Published: (2024)
by: Qiu, Shikai, et al.
Published: (2024)
From Entropy to Epiplexity: Rethinking Information for Computationally Bounded Intelligence
by: Finzi, Marc, et al.
Published: (2026)
by: Finzi, Marc, et al.
Published: (2026)
SoK: Dataset Copyright Auditing in Machine Learning Systems
by: Du, Linkang, et al.
Published: (2024)
by: Du, Linkang, et al.
Published: (2024)
How Do Graph Signals Affect Recommendation: Unveiling the Mystery of Low and High-Frequency Graph Signals
by: Liu, Feng, et al.
Published: (2025)
by: Liu, Feng, et al.
Published: (2025)
Stream of Search (SoS): Learning to Search in Language
by: Gandhi, Kanishk, et al.
Published: (2024)
by: Gandhi, Kanishk, et al.
Published: (2024)
SoK: Machine Learning for Misinformation Detection
by: Xiao, Madelyne, et al.
Published: (2023)
by: Xiao, Madelyne, et al.
Published: (2023)
ODICE: Revealing the Mystery of Distribution Correction Estimation via Orthogonal-gradient Update
by: Mao, Liyuan, et al.
Published: (2024)
by: Mao, Liyuan, et al.
Published: (2024)
Similar Items
-
Response to Promises and Pitfalls of Deep Kernel Learning
by: Wilson, Andrew Gordon, et al.
Published: (2025) -
Scalable and Flexible Causal Discovery with an Efficient Test for Adjacency
by: Amin, Alan Nawzad, et al.
Published: (2024) -
The No Free Lunch Theorem, Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learning
by: Goldblum, Micah, et al.
Published: (2023) -
SoK: The Pitfalls of Deep Reinforcement Learning for Cybersecurity
by: McFadden, Shae, et al.
Published: (2026) -
In-Context Clustering with Large Language Models
by: Wang, Ying, et al.
Published: (2025)