The Benefits of Temporal Correlations: SGD Learns k-Juntas from Random Walks Efficiently
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cornacchia, Elisabetta, Mikulincer, Dan, Mossel, Elchanan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Low-dimensional Functions are Efficiently Learnable under Randomly Biased Distributions
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2025)
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2025)
A Mathematical Model for Curriculum Learning for Parities
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2023)
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2023)
Bypassing the Noisy Parity Barrier: Learning Higher-Order Markov Random Fields from Dynamics
von: Gaitonde, Jason, et al.
Veröffentlicht: (2024)
von: Gaitonde, Jason, et al.
Veröffentlicht: (2024)
Reconstructing the Geometry of Random Geometric Graphs
von: Huang, Han, et al.
Veröffentlicht: (2024)
von: Huang, Han, et al.
Veröffentlicht: (2024)
Sample-Efficient Linear Regression with Self-Selection Bias
von: Gaitonde, Jason, et al.
Veröffentlicht: (2024)
von: Gaitonde, Jason, et al.
Veröffentlicht: (2024)
Noise Sensitivity and Learning Lower Bounds for Hierarchical Functions
von: Li, Rupert, et al.
Veröffentlicht: (2025)
von: Li, Rupert, et al.
Veröffentlicht: (2025)
Some Theoretical Limitations of t-SNE
von: Li, Rupert, et al.
Veröffentlicht: (2026)
von: Li, Rupert, et al.
Veröffentlicht: (2026)
Online Learning of Neural Networks
von: Daniely, Amit, et al.
Veröffentlicht: (2025)
von: Daniely, Amit, et al.
Veröffentlicht: (2025)
Why ReLU? A Bit-Model Dichotomy for Deep Network Training
von: Doron-Arad, Ilan, et al.
Veröffentlicht: (2026)
von: Doron-Arad, Ilan, et al.
Veröffentlicht: (2026)
Better Models and Algorithms for Learning Ising Models from Dynamics
von: Gaitonde, Jason, et al.
Veröffentlicht: (2025)
von: Gaitonde, Jason, et al.
Veröffentlicht: (2025)
A Theory of Online Learning with Autoregressive Chain-of-Thought Reasoning
von: Doron-Arad, Ilan, et al.
Veröffentlicht: (2026)
von: Doron-Arad, Ilan, et al.
Veröffentlicht: (2026)
Learning with Shallow Neural Networks on Cluster-Structured Features
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2026)
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2026)
A Hierarchical Language Model with Predictable Scaling Laws and Provable Benefits of Reasoning
von: Gaitonde, Jason, et al.
Veröffentlicht: (2026)
von: Gaitonde, Jason, et al.
Veröffentlicht: (2026)
Denoising distances beyond the volumetric barrier
von: Huang, Han, et al.
Veröffentlicht: (2026)
von: Huang, Han, et al.
Veröffentlicht: (2026)
Online Realizable Regression and Applications for ReLU Networks
von: Doron-Arad, Ilan, et al.
Veröffentlicht: (2026)
von: Doron-Arad, Ilan, et al.
Veröffentlicht: (2026)
Monotonicity, Topology, and Convexity of Recurrence in Random Walks
von: Li, Rupert, et al.
Veröffentlicht: (2024)
von: Li, Rupert, et al.
Veröffentlicht: (2024)
Is this correct? Let's check!
von: Ben-Eliezer, Omri, et al.
Veröffentlicht: (2022)
von: Ben-Eliezer, Omri, et al.
Veröffentlicht: (2022)
Random Subwords and Billiard Walks in Affine Weyl Groups
von: Defant, Colin, et al.
Veröffentlicht: (2025)
von: Defant, Colin, et al.
Veröffentlicht: (2025)
Learning and Testing Convex Functions
von: Pinto Jr., Renato Ferreira, et al.
Veröffentlicht: (2025)
von: Pinto Jr., Renato Ferreira, et al.
Veröffentlicht: (2025)
Size and depth of monotone neural networks: interpolation and approximation
von: Mikulincer, Dan, et al.
Veröffentlicht: (2022)
von: Mikulincer, Dan, et al.
Veröffentlicht: (2022)
Learning High-Degree Parities: The Crucial Role of the Initialization
von: Abbe, Emmanuel, et al.
Veröffentlicht: (2024)
von: Abbe, Emmanuel, et al.
Veröffentlicht: (2024)
The Refutability Gap: Challenges in Validating Reasoning by Large Language Models
von: Mossel, Elchanan
Veröffentlicht: (2025)
von: Mossel, Elchanan
Veröffentlicht: (2025)
Comparison Theorems for the Mixing Times of Systematic and Random Scan Dynamics
von: Gaitonde, Jason, et al.
Veröffentlicht: (2024)
von: Gaitonde, Jason, et al.
Veröffentlicht: (2024)
Positive Distribution Shift as a Framework for Understanding Tractable Learning
von: Medvedev, Marko, et al.
Veröffentlicht: (2026)
von: Medvedev, Marko, et al.
Veröffentlicht: (2026)
Learning Juntas under Markov Random Fields
von: Chandrasekaran, Gautam, et al.
Veröffentlicht: (2025)
von: Chandrasekaran, Gautam, et al.
Veröffentlicht: (2025)
RQP-SGD: Differential Private Machine Learning through Noisy SGD and Randomized Quantization
von: Feng, Ce, et al.
Veröffentlicht: (2024)
von: Feng, Ce, et al.
Veröffentlicht: (2024)
On the Statistical Benefits of Temporal Difference Learning
von: Cheikhi, David, et al.
Veröffentlicht: (2023)
von: Cheikhi, David, et al.
Veröffentlicht: (2023)
MindFlayer SGD: Efficient Parallel SGD in the Presence of Heterogeneous and Random Worker Compute Times
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2024)
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2024)
Towards Understanding Transformers in Learning Random Walks
von: Shi, Wei, et al.
Veröffentlicht: (2025)
von: Shi, Wei, et al.
Veröffentlicht: (2025)
Biased Local SGD for Efficient Deep Learning on Heterogeneous Systems
von: Lim, Jihyun, et al.
Veröffentlicht: (2025)
von: Lim, Jihyun, et al.
Veröffentlicht: (2025)
Repelling Random Walks
von: Reid, Isaac, et al.
Veröffentlicht: (2023)
von: Reid, Isaac, et al.
Veröffentlicht: (2023)
Revisiting Random Walks for Learning on Graphs
von: Kim, Jinwoo, et al.
Veröffentlicht: (2024)
von: Kim, Jinwoo, et al.
Veröffentlicht: (2024)
Reconstructing Riemannian Metrics From Random Geometric Graphs
von: Huang, Han, et al.
Veröffentlicht: (2025)
von: Huang, Han, et al.
Veröffentlicht: (2025)
Trustworthy Efficient Communication for Distributed Learning using LQ-SGD Algorithm
von: Li, Hongyang, et al.
Veröffentlicht: (2025)
von: Li, Hongyang, et al.
Veröffentlicht: (2025)
Balancing Utility and Privacy: Dynamically Private SGD with Random Projection
von: Jiang, Zhanhong, et al.
Veröffentlicht: (2025)
von: Jiang, Zhanhong, et al.
Veröffentlicht: (2025)
Correlating Cross-Iteration Noise for DP-SGD using Model Curvature
von: Gu, Xin, et al.
Veröffentlicht: (2025)
von: Gu, Xin, et al.
Veröffentlicht: (2025)
Communication-Efficient Federated Learning by Exploiting Spatio-Temporal Correlations of Gradients
von: Zheng, Shenlong, et al.
Veröffentlicht: (2026)
von: Zheng, Shenlong, et al.
Veröffentlicht: (2026)
Differentially Private Decentralized Learning with Random Walks
von: Cyffers, Edwige, et al.
Veröffentlicht: (2024)
von: Cyffers, Edwige, et al.
Veröffentlicht: (2024)
Learning Long Range Dependencies on Graphs via Random Walks
von: Chen, Dexiong, et al.
Veröffentlicht: (2024)
von: Chen, Dexiong, et al.
Veröffentlicht: (2024)
Convolutional Persistence Transforms
von: Solomon, Elchanan, et al.
Veröffentlicht: (2022)
von: Solomon, Elchanan, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Low-dimensional Functions are Efficiently Learnable under Randomly Biased Distributions
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2025) -
A Mathematical Model for Curriculum Learning for Parities
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2023) -
Bypassing the Noisy Parity Barrier: Learning Higher-Order Markov Random Fields from Dynamics
von: Gaitonde, Jason, et al.
Veröffentlicht: (2024) -
Reconstructing the Geometry of Random Geometric Graphs
von: Huang, Han, et al.
Veröffentlicht: (2024) -
Sample-Efficient Linear Regression with Self-Selection Bias
von: Gaitonde, Jason, et al.
Veröffentlicht: (2024)