Learning High-Degree Parities: The Crucial Role of the Initialization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Abbe, Emmanuel, Cornacchia, Elisabetta, Hązła, Jan, Kougang-Yombi, Donald |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Quantitative Version of More Capable Channel Comparison
von: Kougang-Yombi, Donald, et al.
Veröffentlicht: (2024)
von: Kougang-Yombi, Donald, et al.
Veröffentlicht: (2024)
Weight distribution bounds to relate minimum distance, list decoding, and symmetric channel performance
von: Kougang-Yombi, Donald, et al.
Veröffentlicht: (2026)
von: Kougang-Yombi, Donald, et al.
Veröffentlicht: (2026)
A Mathematical Model for Curriculum Learning for Parities
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2023)
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2023)
Intransitive dice tournament is not quasirandom
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2020)
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2020)
Learning with Shallow Neural Networks on Cluster-Structured Features
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2026)
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2026)
On the Minimal Degree Bias in Generalization on the Unseen for non-Boolean Functions
von: Pushkin, Denys, et al.
Veröffentlicht: (2024)
von: Pushkin, Denys, et al.
Veröffentlicht: (2024)
Generalization on the Unseen, Logic Reasoning and Degree Curriculum
von: Abbe, Emmanuel, et al.
Veröffentlicht: (2023)
von: Abbe, Emmanuel, et al.
Veröffentlicht: (2023)
The Benefits of Temporal Correlations: SGD Learns k-Juntas from Random Walks Efficiently
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2026)
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2026)
On the Crucial Role of Initialization for Matrix Factorization
von: Li, Bingcong, et al.
Veröffentlicht: (2024)
von: Li, Bingcong, et al.
Veröffentlicht: (2024)
Low-dimensional Functions are Efficiently Learnable under Randomly Biased Distributions
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2025)
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2025)
Transformation-Invariant Learning and Theoretical Guarantees for OOD Generalization
von: Montasser, Omar, et al.
Veröffentlicht: (2024)
von: Montasser, Omar, et al.
Veröffentlicht: (2024)
Positive Distribution Shift as a Framework for Understanding Tractable Learning
von: Medvedev, Marko, et al.
Veröffentlicht: (2026)
von: Medvedev, Marko, et al.
Veröffentlicht: (2026)
Goldilocks RL: Tuning Task Difficulty to Escape Sparse Rewards for Reasoning
von: Mahrooghi, Ilia, et al.
Veröffentlicht: (2026)
von: Mahrooghi, Ilia, et al.
Veröffentlicht: (2026)
$k$-server-bench: Automating Potential Discovery for the $k$-Server Conjecture
von: Brilliantov, Kirill, et al.
Veröffentlicht: (2026)
von: Brilliantov, Kirill, et al.
Veröffentlicht: (2026)
The Crucial Role of Problem Formulation in Real-World Reinforcement Learning
von: Schäfer, Georg, et al.
Veröffentlicht: (2025)
von: Schäfer, Georg, et al.
Veröffentlicht: (2025)
The merged-staircase property: a necessary and nearly sufficient condition for SGD learning of sparse functions on two-layer neural networks
von: Abbe, Emmanuel, et al.
Veröffentlicht: (2022)
von: Abbe, Emmanuel, et al.
Veröffentlicht: (2022)
The Crucial Role of Samplers in Online Direct Preference Optimization
von: Shi, Ruizhe, et al.
Veröffentlicht: (2024)
von: Shi, Ruizhe, et al.
Veröffentlicht: (2024)
Inductive Domain Transfer In Misspecified Simulation-Based Inference
von: Senouf, Ortal, et al.
Veröffentlicht: (2025)
von: Senouf, Ortal, et al.
Veröffentlicht: (2025)
How Far Can Transformers Reason? The Globality Barrier and Inductive Scratchpad
von: Abbe, Emmanuel, et al.
Veröffentlicht: (2024)
von: Abbe, Emmanuel, et al.
Veröffentlicht: (2024)
Tokenizer Choice For LLM Training: Negligible or Crucial?
von: Ali, Mehdi, et al.
Veröffentlicht: (2023)
von: Ali, Mehdi, et al.
Veröffentlicht: (2023)
Learning High-Dimensional Parity Functions with Product Networks using Gradient Descent
von: Larue, Guillaume, et al.
Veröffentlicht: (2026)
von: Larue, Guillaume, et al.
Veröffentlicht: (2026)
Chain-of-Sketch: Enabling Global Visual Reasoning
von: Lotfi, Aryo, et al.
Veröffentlicht: (2024)
von: Lotfi, Aryo, et al.
Veröffentlicht: (2024)
Hardness of Learning Fixed Parities with Neural Networks
von: Shoshani, Itamar, et al.
Veröffentlicht: (2025)
von: Shoshani, Itamar, et al.
Veröffentlicht: (2025)
Keeping LLMs Aligned After Fine-tuning: The Crucial Role of Prompt Templates
von: Lyu, Kaifeng, et al.
Veröffentlicht: (2024)
von: Lyu, Kaifeng, et al.
Veröffentlicht: (2024)
Loss Gap Parity for Fairness in Heterogeneous Federated Learning
von: Erraji, Brahim, et al.
Veröffentlicht: (2026)
von: Erraji, Brahim, et al.
Veröffentlicht: (2026)
To Infinity and Beyond: Tool-Use Unlocks Length Generalization in State Space Models
von: Malach, Eran, et al.
Veröffentlicht: (2025)
von: Malach, Eran, et al.
Veröffentlicht: (2025)
Boolformer: Symbolic Regression of Logic Functions with Transformers
von: d'Ascoli, Stéphane, et al.
Veröffentlicht: (2023)
von: d'Ascoli, Stéphane, et al.
Veröffentlicht: (2023)
Learning Optimal Individualized Decision Rules with Conditional Demographic Parity
von: Cui, Wenhai, et al.
Veröffentlicht: (2026)
von: Cui, Wenhai, et al.
Veröffentlicht: (2026)
Differential Adjusted Parity for Learning Fair Representations
von: Sahyouni, Bucher, et al.
Veröffentlicht: (2025)
von: Sahyouni, Bucher, et al.
Veröffentlicht: (2025)
Computationally Efficient Replicable Learning of Parities and Applications
von: Noivirt, Moshe, et al.
Veröffentlicht: (2026)
von: Noivirt, Moshe, et al.
Veröffentlicht: (2026)
When can transformers reason with abstract symbols?
von: Boix-Adsera, Enric, et al.
Veröffentlicht: (2023)
von: Boix-Adsera, Enric, et al.
Veröffentlicht: (2023)
Demographic Parity Tails for Regression
von: Le, Naht Sinh, et al.
Veröffentlicht: (2026)
von: Le, Naht Sinh, et al.
Veröffentlicht: (2026)
Approximating the Number of Relevant Variables in a Parity Implies Proper Learning
von: Bshouty, Nader H., et al.
Veröffentlicht: (2024)
von: Bshouty, Nader H., et al.
Veröffentlicht: (2024)
RL for Reasoning by Adaptively Revealing Rationales
von: Amani, Mohammad Hossein, et al.
Veröffentlicht: (2025)
von: Amani, Mohammad Hossein, et al.
Veröffentlicht: (2025)
SDGym: Low-Code Reinforcement Learning Environments using System Dynamics Models
von: Klu, Emmanuel, et al.
Veröffentlicht: (2023)
von: Klu, Emmanuel, et al.
Veröffentlicht: (2023)
Collaboration Between the City and Machine Learning Community is Crucial to Efficient Autonomous Vehicles Routing
von: Psarou, Anastasia, et al.
Veröffentlicht: (2025)
von: Psarou, Anastasia, et al.
Veröffentlicht: (2025)
Parity, Sensitivity, and Transformers
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2026)
von: Kozachinskiy, Alexander, et al.
Veröffentlicht: (2026)
Autocorrelation Matters: Understanding the Role of Initialization Schemes for State Space Models
von: Liu, Fusheng, et al.
Veröffentlicht: (2024)
von: Liu, Fusheng, et al.
Veröffentlicht: (2024)
A Bioinformatic Approach Validated Utilizing Machine Learning Algorithms to Identify Relevant Biomarkers and Crucial Pathways in Gallbladder Cancer
von: Khatun, Rabea, et al.
Veröffentlicht: (2024)
von: Khatun, Rabea, et al.
Veröffentlicht: (2024)
Unravelling the (In)compatibility of Statistical-Parity and Equalized-Odds
von: Bargh, Mortaza S., et al.
Veröffentlicht: (2026)
von: Bargh, Mortaza S., et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Quantitative Version of More Capable Channel Comparison
von: Kougang-Yombi, Donald, et al.
Veröffentlicht: (2024) -
Weight distribution bounds to relate minimum distance, list decoding, and symmetric channel performance
von: Kougang-Yombi, Donald, et al.
Veröffentlicht: (2026) -
A Mathematical Model for Curriculum Learning for Parities
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2023) -
Intransitive dice tournament is not quasirandom
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2020) -
Learning with Shallow Neural Networks on Cluster-Structured Features
von: Cornacchia, Elisabetta, et al.
Veröffentlicht: (2026)