Deep Minds and Shallow Probes
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Su Hyeong, Kondor, Risi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks
by: Lee, Su Hyeong, et al.
Published: (2025)
by: Lee, Su Hyeong, et al.
Published: (2025)
Sign Rank Limitations for Inner Product Graph Decoders
by: Lee, Su Hyeong, et al.
Published: (2024)
by: Lee, Su Hyeong, et al.
Published: (2024)
Toward Super-polynomial Quantum Speedup of Equivariant Quantum Algorithms with SU($d$) Symmetry
by: Zheng, Han, et al.
Published: (2022)
by: Zheng, Han, et al.
Published: (2022)
Unifying O(3) Equivariant Neural Networks Design with Tensor-Network Formalism
by: Li, Zimu, et al.
Published: (2022)
by: Li, Zimu, et al.
Published: (2022)
Simple Baselines are Competitive with Code Evolution
by: Gideoni, Yonatan, et al.
Published: (2026)
by: Gideoni, Yonatan, et al.
Published: (2026)
P-Tensors: a General Formalism for Constructing Higher Order Message Passing Networks
by: Hands, Andrew, et al.
Published: (2023)
by: Hands, Andrew, et al.
Published: (2023)
CARLE: A Hybrid Deep-Shallow Learning Framework for Robust and Explainable RUL Estimation of Rolling Element Bearings
by: Razzaq, Waleed, et al.
Published: (2025)
by: Razzaq, Waleed, et al.
Published: (2025)
Mind the GAP! The Challenges of Scale in Pixel-based Deep Reinforcement Learning
by: Sokar, Ghada, et al.
Published: (2025)
by: Sokar, Ghada, et al.
Published: (2025)
MindCraft: How Concept Trees Take Shape In Deep Models
by: Tian, Bowei, et al.
Published: (2025)
by: Tian, Bowei, et al.
Published: (2025)
Continuous Thought Machines
by: Darlow, Luke, et al.
Published: (2025)
by: Darlow, Luke, et al.
Published: (2025)
Soft Contamination Means Benchmarks Test Shallow Generalization
by: Spiesberger, Ari, et al.
Published: (2026)
by: Spiesberger, Ari, et al.
Published: (2026)
Learning to Solve Multiresolution Matrix Factorization by Manifold Optimization and Evolutionary Metaheuristics
by: Hy, Truong Son, et al.
Published: (2024)
by: Hy, Truong Son, et al.
Published: (2024)
Improving World Models using Deep Supervision with Linear Probes
by: Zahorodnii, Andrii
Published: (2025)
by: Zahorodnii, Andrii
Published: (2025)
Meta-Learning and Meta-Reinforcement Learning -- Tracing the Path towards DeepMind's Adaptive Agent
by: Hoppmann, Björn, et al.
Published: (2026)
by: Hoppmann, Björn, et al.
Published: (2026)
Quotient Geometry, Effective Curvature, and Implicit Bias in Simple Shallow Neural Networks
by: Dong, Hang-Cheng, et al.
Published: (2026)
by: Dong, Hang-Cheng, et al.
Published: (2026)
The Spectral Bias of Shallow Neural Network Learning is Shaped by the Choice of Non-linearity
by: Sahs, Justin, et al.
Published: (2025)
by: Sahs, Justin, et al.
Published: (2025)
Geometric and Dynamic Scaling in Deep Transformers
by: Su, Haoran, et al.
Published: (2026)
by: Su, Haoran, et al.
Published: (2026)
Deep Edge Filter: Return of the Human-Crafted Layer in Deep Learning
by: Lee, Dongkwan, et al.
Published: (2025)
by: Lee, Dongkwan, et al.
Published: (2025)
SafeDPO: A Simple Approach to Direct Preference Optimization with Enhanced Safety
by: Kim, Geon-Hyeong, et al.
Published: (2025)
by: Kim, Geon-Hyeong, et al.
Published: (2025)
Modeling Others' Minds as Code
by: Jha, Kunal, et al.
Published: (2025)
by: Jha, Kunal, et al.
Published: (2025)
Towards Initialization-dependent and Non-vacuous Generalization Bounds for Overparameterized Shallow Neural Networks
by: Lei, Yunwen, et al.
Published: (2026)
by: Lei, Yunwen, et al.
Published: (2026)
Deep Support Vectors
by: Lee, Junhoo, et al.
Published: (2024)
by: Lee, Junhoo, et al.
Published: (2024)
From Shallow Bayesian Neural Networks to Gaussian Processes: General Convergence, Identifiability and Scalable Inference
by: de Araújo, Gracielle Antunes, et al.
Published: (2026)
by: de Araújo, Gracielle Antunes, et al.
Published: (2026)
Zero-Direction Probing: A Linear-Algebraic Framework for Deep Analysis of Large-Language-Model Drift
by: Pandey, Amit
Published: (2025)
by: Pandey, Amit
Published: (2025)
Finite-Time Analysis of Gradient Descent for Shallow Transformers
by: Arda, Enes, et al.
Published: (2026)
by: Arda, Enes, et al.
Published: (2026)
Large Language and Reasoning Models are Shallow Disjunctive Reasoners
by: Khalid, Irtaza, et al.
Published: (2025)
by: Khalid, Irtaza, et al.
Published: (2025)
Rethinking industrial artificial intelligence: a unified foundation framework
by: Lee, Jay, et al.
Published: (2025)
by: Lee, Jay, et al.
Published: (2025)
Machine Learning Approaches for Diagnostics and Prognostics of Industrial Systems Using Open Source Data from PHM Data Challenges: A Review
by: Su, Hanqi, et al.
Published: (2023)
by: Su, Hanqi, et al.
Published: (2023)
A Unified Industrial Large Knowledge Model Framework in Industry 4.0 and Smart Manufacturing
by: Lee, Jay, et al.
Published: (2023)
by: Lee, Jay, et al.
Published: (2023)
Deep Metric Loss for Multimodal Learning
by: Moon, Sehwan, et al.
Published: (2023)
by: Moon, Sehwan, et al.
Published: (2023)
Taming the Adversary: Stable Minimax Deep Deterministic Policy Gradient via Fractional Objectives
by: Lee, Taeho, et al.
Published: (2026)
by: Lee, Taeho, et al.
Published: (2026)
ProvMind: Provenance-grounded reasoning for materials synthesis
by: Zhang, Yiming, et al.
Published: (2026)
by: Zhang, Yiming, et al.
Published: (2026)
When Does Structure Matter in Continual Learning? Dimensionality Controls When Modularity Shapes Representational Geometry
by: Korte, Kathrin, et al.
Published: (2026)
by: Korte, Kathrin, et al.
Published: (2026)
GPT-4o Lacks Core Features of Theory of Mind
by: Muchovej, John, et al.
Published: (2026)
by: Muchovej, John, et al.
Published: (2026)
Autonomous Navigation of an Ultrasound Probe Towards Standard Scan Planes with Deep Reinforcement Learning
by: Li, Keyu, et al.
Published: (2021)
by: Li, Keyu, et al.
Published: (2021)
Convergence of Shallow ReLU Networks on Weakly Interacting Data
by: Dana, Léo, et al.
Published: (2025)
by: Dana, Léo, et al.
Published: (2025)
HiVAE: Hierarchical Latent Variables for Scalable Theory of Mind
by: Doering, Nigel, et al.
Published: (2026)
by: Doering, Nigel, et al.
Published: (2026)
Deep Feature Embedding for Tabular Data
by: Wu, Yuqian, et al.
Published: (2024)
by: Wu, Yuqian, et al.
Published: (2024)
DMWM: Dual-Mind World Model with Long-Term Imagination
by: Wang, Lingyi, et al.
Published: (2025)
by: Wang, Lingyi, et al.
Published: (2025)
Mind the Model, Not the Agent: The Primacy Bias in Model-based RL
by: Qiao, Zhongjian, et al.
Published: (2023)
by: Qiao, Zhongjian, et al.
Published: (2023)
Similar Items
-
Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks
by: Lee, Su Hyeong, et al.
Published: (2025) -
Sign Rank Limitations for Inner Product Graph Decoders
by: Lee, Su Hyeong, et al.
Published: (2024) -
Toward Super-polynomial Quantum Speedup of Equivariant Quantum Algorithms with SU($d$) Symmetry
by: Zheng, Han, et al.
Published: (2022) -
Unifying O(3) Equivariant Neural Networks Design with Tensor-Network Formalism
by: Li, Zimu, et al.
Published: (2022) -
Simple Baselines are Competitive with Code Evolution
by: Gideoni, Yonatan, et al.
Published: (2026)