Predictable Gradient Manifolds in Deep Learning: Temporal Path-Length and Intrinsic Rank as a Complexity Regime
Fuente:
arXiv
Saved in:
| Main Author: | Calvo, Anherutowa |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Nonlinear Regime Transitions via Semi-Parametric State-Space Models
by: Hiremath, Prakul Sunil
Published: (2026)
by: Hiremath, Prakul Sunil
Published: (2026)
Hidden Failure Modes of Gradient Modification under Adam in Continual Learning, and Adaptive Decoupled Moment Routing as a Repair
by: Hu, Yuelin, et al.
Published: (2026)
by: Hu, Yuelin, et al.
Published: (2026)
StepScorer: Accelerating Reinforcement Learning with Step-wise Scoring and Psychological Regret Modeling
by: Xu, Zhe
Published: (2026)
by: Xu, Zhe
Published: (2026)
Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report
by: Wigmore, Jerrod, et al.
Published: (2024)
by: Wigmore, Jerrod, et al.
Published: (2024)
Recurrent Memory-Augmented Transformers with Chunked Attention for Long-Context Language Modeling
by: Kashyap, Ankit
Published: (2025)
by: Kashyap, Ankit
Published: (2025)
Time-to-Injury Forecasting in Elite Female Football: A DeepHit Survival Approach
by: Catterall, Victoria, et al.
Published: (2026)
by: Catterall, Victoria, et al.
Published: (2026)
Unified Unbiased Variance Estimation for Maximum Mean Discrepancy: Robust Finite-Sample Performance with Imbalanced Data and Exact Acceleration under Null and Alternative Hypotheses
by: Zhong, Shijie, et al.
Published: (2026)
by: Zhong, Shijie, et al.
Published: (2026)
A Prescriptive Framework for Determining Optimal Days for Short-Term Traffic Counts
by: Mukwaya, Arthur, et al.
Published: (2025)
by: Mukwaya, Arthur, et al.
Published: (2025)
Emotion-Inspired Learning Signals (EILS): A Homeostatic Framework for Adaptive Autonomous Agents
by: Tiwari, Dhruv
Published: (2025)
by: Tiwari, Dhruv
Published: (2025)
Optimizing Job Allocation using Reinforcement Learning with Graph Neural Networks
by: Quaedvlieg, Lars C. P. M.
Published: (2025)
by: Quaedvlieg, Lars C. P. M.
Published: (2025)
On the approximation capability of GNNs in node classification/regression tasks
by: D'Inverno, Giuseppe Alessio, et al.
Published: (2021)
by: D'Inverno, Giuseppe Alessio, et al.
Published: (2021)
Primal-Dual Sample Complexity Bounds for Constrained Markov Decision Processes with Multiple Constraints
by: Buckley, Max, et al.
Published: (2025)
by: Buckley, Max, et al.
Published: (2025)
CLGNN: A Contrastive Learning-based GNN Model for Betweenness Centrality Prediction on Temporal Graphs
by: Zhang, Tianming, et al.
Published: (2025)
by: Zhang, Tianming, et al.
Published: (2025)
DELTA: Variational Disentangled Learning for Privacy-Preserving Data Reprogramming
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
The Architecture of Errors: From Universal Impossibility to Patch-Local LLM Reliability
by: Arbuzov, Mikhail L., et al.
Published: (2026)
by: Arbuzov, Mikhail L., et al.
Published: (2026)
TRIM: Achieving Extreme Sparsity with Targeted Row-wise Iterative Metric-driven Pruning
by: Beck, Florentin, et al.
Published: (2025)
by: Beck, Florentin, et al.
Published: (2025)
Causality-Driven Neural Network Repair: Challenges and Opportunities
by: Vares, Fatemeh, et al.
Published: (2025)
by: Vares, Fatemeh, et al.
Published: (2025)
Enhanced and Efficient Reasoning in Large Learning Models
by: Valiant, Leslie G.
Published: (2026)
by: Valiant, Leslie G.
Published: (2026)
UniPROT: Uniform Prototype Selection via Partial Optimal Transport with Submodular Guarantees
by: Chanda, Prateek, et al.
Published: (2026)
by: Chanda, Prateek, et al.
Published: (2026)
Sparse Knowledge Distillation: A Mathematical Framework for Probability-Domain Temperature Scaling and Multi-Stage Compression
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
Recursive Meta-Distillation: An Axiomatic Framework for Iterative Knowledge Refinement
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
A Multidisciplinary Approach to Telegram Data Analysis
by: Varbanov, Velizar, et al.
Published: (2024)
by: Varbanov, Velizar, et al.
Published: (2024)
Low-Rank Adapters Initialization via Gradient Surgery for Continual Learning
by: Pasquali, Joana, et al.
Published: (2026)
by: Pasquali, Joana, et al.
Published: (2026)
4OPS: Structural Difficulty Modeling in Integer Arithmetic Puzzles
by: Zeytuncu, Yunus E.
Published: (2026)
by: Zeytuncu, Yunus E.
Published: (2026)
Graph Coloring for Multi-Task Learning
by: Patapati, Santosh
Published: (2025)
by: Patapati, Santosh
Published: (2025)
Fusing Rewards and Preferences in Reinforcement Learning
by: Khorasani, Sadegh, et al.
Published: (2025)
by: Khorasani, Sadegh, et al.
Published: (2025)
Explaining Bayesian Optimization by Shapley Values Facilitates Human-AI Collaboration
by: Rodemann, Julian, et al.
Published: (2024)
by: Rodemann, Julian, et al.
Published: (2024)
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
by: Yousaf, Iqra
Published: (2024)
by: Yousaf, Iqra
Published: (2024)
LLM Agents for Generating Microservice-based Applications: how complex is your specification?
by: Yellin, Daniel M.
Published: (2025)
by: Yellin, Daniel M.
Published: (2025)
Reinforcement Learning Controlled Adaptive PSO for Task Offloading in IIoT Edge Computing
by: Perera, Minod, et al.
Published: (2025)
by: Perera, Minod, et al.
Published: (2025)
Potential-Based Intrinsic Motivation: Preserving Optimality With Complex, Non-Markovian Shaping Rewards
by: Forbes, Grant C., et al.
Published: (2024)
by: Forbes, Grant C., et al.
Published: (2024)
Front-door Reducibility: Reducing ADMGs to the Standard Front-door Setting via a Graphical Criterion
by: Mao, Jianqiao, et al.
Published: (2025)
by: Mao, Jianqiao, et al.
Published: (2025)
MR-GNF: Multi-Resolution Graph Neural Forecasting on Ellipsoidal Meshes for Efficient Regional Weather Prediction
by: Shchur, Andrii, et al.
Published: (2026)
by: Shchur, Andrii, et al.
Published: (2026)
Decoding Rewards in Competitive Games: Inverse Game Theory with Entropy Regularization
by: Liao, Junyi, et al.
Published: (2026)
by: Liao, Junyi, et al.
Published: (2026)
CoGraM: Context-sensitive granular optimization method with rollback for robust model fusion
by: Lenz, Julius
Published: (2025)
by: Lenz, Julius
Published: (2025)
Loss-Complexity Landscape and Model Structure Functions
by: Kolpakov, Alexander
Published: (2025)
by: Kolpakov, Alexander
Published: (2025)
Normalization Layer Per-Example Gradients are Sufficient to Predict Gradient Noise Scale in Transformers
by: Gray, Gavia, et al.
Published: (2024)
by: Gray, Gavia, et al.
Published: (2024)
AI2T: Building Trustable AI Tutors by Interactively Teaching a Self-Aware Learning Agent
by: Weitekamp, Daniel, et al.
Published: (2024)
by: Weitekamp, Daniel, et al.
Published: (2024)
Alignment Verifiability in Large Language Models: Normative Indistinguishability under Behavioral Evaluation
by: Santos-Grueiro, Igor
Published: (2026)
by: Santos-Grueiro, Igor
Published: (2026)
Diffusion-MPC in Discrete Domains: Feasibility Constraints, Horizon Effects, and Critic Alignment: Case study with Tetris
by: Wang, Haochuan Kevin
Published: (2026)
by: Wang, Haochuan Kevin
Published: (2026)
Similar Items
-
Learning Nonlinear Regime Transitions via Semi-Parametric State-Space Models
by: Hiremath, Prakul Sunil
Published: (2026) -
Hidden Failure Modes of Gradient Modification under Adam in Continual Learning, and Adaptive Decoupled Moment Routing as a Repair
by: Hu, Yuelin, et al.
Published: (2026) -
StepScorer: Accelerating Reinforcement Learning with Step-wise Scoring and Psychological Regret Modeling
by: Xu, Zhe
Published: (2026) -
Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report
by: Wigmore, Jerrod, et al.
Published: (2024) -
Recurrent Memory-Augmented Transformers with Chunked Attention for Long-Context Language Modeling
by: Kashyap, Ankit
Published: (2025)