Neuron Block Dynamics for XOR Classification with Zero-Margin
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Braun, Guillaume, Imaizumi, Masaaki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning a Single Index Model from Anisotropic Data with vanilla Stochastic Gradient Descent
von: Braun, Guillaume, et al.
Veröffentlicht: (2025)
von: Braun, Guillaume, et al.
Veröffentlicht: (2025)
Fast Escape, Slow Convergence: Learning Dynamics of Phase Retrieval under Power-Law Data
von: Braun, Guillaume, et al.
Veröffentlicht: (2025)
von: Braun, Guillaume, et al.
Veröffentlicht: (2025)
Spectral Gradient Descent Mitigates Anisotropy-Driven Misalignment: A Case Study in Phase Retrieval
von: Braun, Guillaume, et al.
Veröffentlicht: (2026)
von: Braun, Guillaume, et al.
Veröffentlicht: (2026)
Zero Generalization Error Theorem for Random Interpolators via Algebraic Geometry
von: Yoshida, Naoki, et al.
Veröffentlicht: (2025)
von: Yoshida, Naoki, et al.
Veröffentlicht: (2025)
Precise Dynamics of Diagonal Linear Networks: A Unifying Analysis by Dynamical Mean-Field Theory
von: Nishiyama, Sota, et al.
Veröffentlicht: (2025)
von: Nishiyama, Sota, et al.
Veröffentlicht: (2025)
High-Dimensional Limit of Stochastic Gradient Flow via Dynamical Mean-Field Theory
von: Nishiyama, Sota, et al.
Veröffentlicht: (2026)
von: Nishiyama, Sota, et al.
Veröffentlicht: (2026)
Optimal Dynamic Regret by Transformers for Non-Stationary Reinforcement Learning
von: Chen, Baiyuan, et al.
Veröffentlicht: (2025)
von: Chen, Baiyuan, et al.
Veröffentlicht: (2025)
Spectrum-Adaptive Generalization Bounds for Trained Deep Transformers
von: Sakai, Mana, et al.
Veröffentlicht: (2026)
von: Sakai, Mana, et al.
Veröffentlicht: (2026)
High-dimensional Contextual Bandit Problem without Sparsity
von: Komiyama, Junpei, et al.
Veröffentlicht: (2023)
von: Komiyama, Junpei, et al.
Veröffentlicht: (2023)
Effect of Random Learning Rate: Theoretical Analysis of SGD Dynamics in Non-Convex Optimization via Stationary Distribution
von: Yoshida, Naoki, et al.
Veröffentlicht: (2024)
von: Yoshida, Naoki, et al.
Veröffentlicht: (2024)
Extended Wasserstein-GAN Approach to Causal Distribution Learning: Density-Free Estimation and Minimax Optimality
von: Tamano, Shu, et al.
Veröffentlicht: (2026)
von: Tamano, Shu, et al.
Veröffentlicht: (2026)
Approximation of Permutation Invariant Polynomials by Transformers: Efficient Construction in Column-Size
von: Takeshita, Naoki, et al.
Veröffentlicht: (2025)
von: Takeshita, Naoki, et al.
Veröffentlicht: (2025)
Benign Overfitting in Time Series Linear Models with Over-Parameterization
von: Nakakita, Shogo, et al.
Veröffentlicht: (2022)
von: Nakakita, Shogo, et al.
Veröffentlicht: (2022)
Finite-Sample Inference for Sparsely Permuted Linear Regression
von: Ota, Hirofumi, et al.
Veröffentlicht: (2026)
von: Ota, Hirofumi, et al.
Veröffentlicht: (2026)
Bayesian Analysis for Over-parameterized Linear Model via Effective Spectra
von: Wakayama, Tomoya, et al.
Veröffentlicht: (2023)
von: Wakayama, Tomoya, et al.
Veröffentlicht: (2023)
Minimax Rates of Estimation for Optimal Transport Map between Infinite-Dimensional Spaces
von: Ponnoprat, Donlapark, et al.
Veröffentlicht: (2025)
von: Ponnoprat, Donlapark, et al.
Veröffentlicht: (2025)
Automatic Domain Adaptation by Transformers in In-Context Learning
von: Hataya, Ryuichiro, et al.
Veröffentlicht: (2024)
von: Hataya, Ryuichiro, et al.
Veröffentlicht: (2024)
Infinite-Width Limit of a Single Attention Layer: Analysis via Tensor Programs
von: Sakai, Mana, et al.
Veröffentlicht: (2025)
von: Sakai, Mana, et al.
Veröffentlicht: (2025)
Effect of Weight Quantization on Learning Models by Typical Case Analysis
von: Kashiwamura, Shuhei, et al.
Veröffentlicht: (2024)
von: Kashiwamura, Shuhei, et al.
Veröffentlicht: (2024)
Precise gradient descent training dynamics for finite-width multi-layer neural networks
von: Han, Qiyang, et al.
Veröffentlicht: (2025)
von: Han, Qiyang, et al.
Veröffentlicht: (2025)
High-dimensional Nonparametric Contextual Bandit Problem
von: Iwazaki, Shogo, et al.
Veröffentlicht: (2025)
von: Iwazaki, Shogo, et al.
Veröffentlicht: (2025)
Anti Mode-Collapse in Mean-Field Transformer via Auxiliary Variables
von: Imaizumi, Masaaki, et al.
Veröffentlicht: (2026)
von: Imaizumi, Masaaki, et al.
Veröffentlicht: (2026)
Minimax Optimal Estimation of Transport-Growth Pairs in Unbalanced Optimal Transport
von: Ponnoprat, Donlapark, et al.
Veröffentlicht: (2026)
von: Ponnoprat, Donlapark, et al.
Veröffentlicht: (2026)
Dichotomy of Feature Learning and Unlearning: Fast-Slow Analysis on Neural Networks with Stochastic Gradient Descent
von: Imai, Shota, et al.
Veröffentlicht: (2026)
von: Imai, Shota, et al.
Veröffentlicht: (2026)
Federated Learning with Relative Fairness
von: Nakakita, Shogo, et al.
Veröffentlicht: (2024)
von: Nakakita, Shogo, et al.
Veröffentlicht: (2024)
Demystifying MaskGIT Sampler and Beyond: Adaptive Order Selection in Masked Diffusion
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2025)
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2025)
Distillation of Discrete Diffusion through Dimensional Correlations
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2024)
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2024)
Training-Induced Escape from Token Clustering in a Mean-Field Formulation of Transformers
von: Isobe, Noboru, et al.
Veröffentlicht: (2026)
von: Isobe, Noboru, et al.
Veröffentlicht: (2026)
Transformers Provably Learn Sparse XOR with Polylogarithmic Parameters
von: Han, Yaomengxi, et al.
Veröffentlicht: (2025)
von: Han, Yaomengxi, et al.
Veröffentlicht: (2025)
CITE: Anytime-Valid Statistical Inference in LLM Self-Consistency
von: Ota, Hirofumi, et al.
Veröffentlicht: (2026)
von: Ota, Hirofumi, et al.
Veröffentlicht: (2026)
Iteration and Stochastic First-order Oracle Complexities of Stochastic Gradient Descent using Constant and Decaying Learning Rates
von: Imaizumi, Kento, et al.
Veröffentlicht: (2024)
von: Imaizumi, Kento, et al.
Veröffentlicht: (2024)
SAN: Inducing Metrizability of GAN with Discriminative Normalized Linear Layer
von: Takida, Yuhta, et al.
Veröffentlicht: (2023)
von: Takida, Yuhta, et al.
Veröffentlicht: (2023)
Prototypical Extreme Multi-label Classification with a Dynamic Margin Loss
von: Dahiya, Kunal, et al.
Veröffentlicht: (2024)
von: Dahiya, Kunal, et al.
Veröffentlicht: (2024)
Unified Binary and Multiclass Margin-Based Classification
von: Wang, Yutong, et al.
Veröffentlicht: (2023)
von: Wang, Yutong, et al.
Veröffentlicht: (2023)
Comparing Classical and Quantum Variational Classifiers on the XOR Problem
von: Seilkhan, Miras, et al.
Veröffentlicht: (2026)
von: Seilkhan, Miras, et al.
Veröffentlicht: (2026)
Both Asymptotic and Non-Asymptotic Convergence of Quasi-Hyperbolic Momentum using Increasing Batch Size
von: Imaizumi, Kento, et al.
Veröffentlicht: (2025)
von: Imaizumi, Kento, et al.
Veröffentlicht: (2025)
Noise-Adaptive Conformal Classification with Marginal Coverage
von: Bortolotti, Teresa, et al.
Veröffentlicht: (2025)
von: Bortolotti, Teresa, et al.
Veröffentlicht: (2025)
VEC-SBM: Optimal Community Detection with Vectorial Edges Covariates
von: Braun, Guillaume, et al.
Veröffentlicht: (2024)
von: Braun, Guillaume, et al.
Veröffentlicht: (2024)
High-Dimensional Single-Index Models: Link Estimation and Marginal Inference
von: Sawaya, Kazuma, et al.
Veröffentlicht: (2024)
von: Sawaya, Kazuma, et al.
Veröffentlicht: (2024)
Lightweight Strategy for XOR PUFs as Security Primitives for Resource-constrained IoT device
von: Li, Gaoxiang, et al.
Veröffentlicht: (2022)
von: Li, Gaoxiang, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Learning a Single Index Model from Anisotropic Data with vanilla Stochastic Gradient Descent
von: Braun, Guillaume, et al.
Veröffentlicht: (2025) -
Fast Escape, Slow Convergence: Learning Dynamics of Phase Retrieval under Power-Law Data
von: Braun, Guillaume, et al.
Veröffentlicht: (2025) -
Spectral Gradient Descent Mitigates Anisotropy-Driven Misalignment: A Case Study in Phase Retrieval
von: Braun, Guillaume, et al.
Veröffentlicht: (2026) -
Zero Generalization Error Theorem for Random Interpolators via Algebraic Geometry
von: Yoshida, Naoki, et al.
Veröffentlicht: (2025) -
Precise Dynamics of Diagonal Linear Networks: A Unifying Analysis by Dynamical Mean-Field Theory
von: Nishiyama, Sota, et al.
Veröffentlicht: (2025)