Saved in:
| Main Authors: | Dramko, Evan, Xiong, Yihuang, Zhu, Yizhi, Hautier, Geoffroy, Reps, Thomas, Jermaine, Christopher, Kyrillidis, Anastasios |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2509.24115 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On The Finetuning of MLIPs Through the Lens of Iterated Maps With BPTT
by: Dramko, Evan, et al.
Published: (2025)
by: Dramko, Evan, et al.
Published: (2025)
Teaching and Learning under Deductive Errors
by: Telle, Jan Arne, et al.
Published: (2026)
by: Telle, Jan Arne, et al.
Published: (2026)
Autoencoded UMAP-Enhanced Clustering for Unsupervised Learning
by: Chavooshi, Malihehsadat, et al.
Published: (2025)
by: Chavooshi, Malihehsadat, et al.
Published: (2025)
Backpropagation Through Time For Networks With Long-Term Dependencies
by: Bird, George, et al.
Published: (2021)
by: Bird, George, et al.
Published: (2021)
A Theoretical Study of (Hyper) Self-Attention through the Lens of Interactions: Representation, Training, Generalization
by: Ustaomeroglu, Muhammed, et al.
Published: (2025)
by: Ustaomeroglu, Muhammed, et al.
Published: (2025)
FSD-CAP: Fractional Subgraph Diffusion with Class-Aware Propagation for Graph Feature Imputation
by: Qiao, Xin, et al.
Published: (2026)
by: Qiao, Xin, et al.
Published: (2026)
Understanding the Nature of Generative AI as Threshold Logic in High-Dimensional Space
by: Levin, Ilya
Published: (2026)
by: Levin, Ilya
Published: (2026)
Different Statistical Perspectives for Understanding Generalisation in Graph Neural Networks
by: Ayday, Nil, et al.
Published: (2026)
by: Ayday, Nil, et al.
Published: (2026)
Thanos: A Block-wise Pruning Algorithm for Efficient Large Language Model Compression
by: Ilin, Ivan, et al.
Published: (2025)
by: Ilin, Ivan, et al.
Published: (2025)
A Special Case of Quadratic Extrapolation Under the Neural Tangent Kernel
by: Kim, Abiel
Published: (2025)
by: Kim, Abiel
Published: (2025)
Greedy feature selection: Classifier-dependent feature selection via greedy methods
by: Camattari, Fabiana, et al.
Published: (2024)
by: Camattari, Fabiana, et al.
Published: (2024)
Near-optimal learning of Banach-valued, high-dimensional functions via deep neural networks
by: Adcock, Ben, et al.
Published: (2022)
by: Adcock, Ben, et al.
Published: (2022)
A polynomial-time algorithm for deciding the Hilbert Nullstellensatz over $\mathbb{Z}_2$. A proof of $\mathbf{P}=\mathbf{NP}$ hypothesis
by: Petrov, Petar P.
Published: (2022)
by: Petrov, Petar P.
Published: (2022)
Error Bounds for Learning with Vector-Valued Random Features
by: Lanthaler, Samuel, et al.
Published: (2023)
by: Lanthaler, Samuel, et al.
Published: (2023)
Beyond Discreteness: Sample Complexity Analysis of Straight-Through Estimator for 1-bit Quantization
by: Jeong, Halyun, et al.
Published: (2025)
by: Jeong, Halyun, et al.
Published: (2025)
Binarized Neural Networks Converge Toward Algorithmic Simplicity: Empirical Support for the Learning-as-Compression Hypothesis
by: Sakabe, Eduardo Y., et al.
Published: (2025)
by: Sakabe, Eduardo Y., et al.
Published: (2025)
Tricks and Plug-ins for Gradient Boosting with Transformers
by: Fang, Biyi, et al.
Published: (2025)
by: Fang, Biyi, et al.
Published: (2025)
Think Thrice Before You Speak: Dual knowledge-enhanced Theory-of-Mind Reasoning for Persuasive Agents
by: Ma, Minghui, et al.
Published: (2026)
by: Ma, Minghui, et al.
Published: (2026)
Neural Network Approximation: A View from Polytope Decomposition
by: Li, ZeYu, et al.
Published: (2026)
by: Li, ZeYu, et al.
Published: (2026)
Approaching I/O-optimality for Approximate Attention
by: Papp, Pál András, et al.
Published: (2026)
by: Papp, Pál András, et al.
Published: (2026)
The Current and Future Perspectives of Zinc Oxide Nanoparticles in the Treatment of Diabetes Mellitus
by: Yousaf, Iqra
Published: (2024)
by: Yousaf, Iqra
Published: (2024)
The two clocks and the innovation window: When and how generative models learn rules
by: Wang, Binxu, et al.
Published: (2026)
by: Wang, Binxu, et al.
Published: (2026)
Learning Decentralized Swarms Using Rotation Equivariant Graph Neural Networks
by: Transue, Taos, et al.
Published: (2025)
by: Transue, Taos, et al.
Published: (2025)
Bounds on the Generalization Error in Active Learning
by: Menden, Vincent, et al.
Published: (2024)
by: Menden, Vincent, et al.
Published: (2024)
Swap Agnostic Learning, or Characterizing Omniprediction via Multicalibration
by: Gopalan, Parikshit, et al.
Published: (2023)
by: Gopalan, Parikshit, et al.
Published: (2023)
The Price of Robustness: Stable Classifiers Need Overparameterization
by: von Berg, Jonas, et al.
Published: (2026)
by: von Berg, Jonas, et al.
Published: (2026)
Constructive interpolation and generalization rates for neural ODEs: a control perspective
by: Álvarez-López, Antonio, et al.
Published: (2026)
by: Álvarez-López, Antonio, et al.
Published: (2026)
Theoretical Analysis of Positional Encodings in Transformer Models: Impact on Expressiveness and Generalization
by: Li, Yin
Published: (2025)
by: Li, Yin
Published: (2025)
Bandwidth of Nondeterministic Finite Automata
by: Cho, Da-Jung, et al.
Published: (2026)
by: Cho, Da-Jung, et al.
Published: (2026)
Greedy Matchings in Bipartite Graphs with Ordered Vertex Sets
by: Simon, Hans U.
Published: (2024)
by: Simon, Hans U.
Published: (2024)
Reduced Jeffries-Matusita distance: A Novel Loss Function to Improve Generalization Performance of Deep Classification Models
by: Lashkari, Mohammad, et al.
Published: (2024)
by: Lashkari, Mohammad, et al.
Published: (2024)
Physics-informed features in supervised machine learning
by: Lampani, Margherita, et al.
Published: (2025)
by: Lampani, Margherita, et al.
Published: (2025)
Graded Transformers
by: Shaska Sr, Tony
Published: (2025)
by: Shaska Sr, Tony
Published: (2025)
Random feature-based double Vovk-Azoury-Warmuth algorithm for online multi-kernel learning
by: Rokhlin, Dmitry B., et al.
Published: (2025)
by: Rokhlin, Dmitry B., et al.
Published: (2025)
A hierarchical Vovk-Azoury-Warmuth forecaster with discounting for online regression in RKHS
by: Rokhlin, Dmitry B.
Published: (2025)
by: Rokhlin, Dmitry B.
Published: (2025)
QGraphLIME - Explaining Quantum Graph Neural Networks
by: Jena, Haribandhu, et al.
Published: (2025)
by: Jena, Haribandhu, et al.
Published: (2025)
Roughness and entropy measures of a soft set
by: Acharjee, Santanu, et al.
Published: (2026)
by: Acharjee, Santanu, et al.
Published: (2026)
A Hybrid Deep Learning and Anomaly Detection Framework for Real-Time Malicious URL Classification
by: Khaled, Berkani, et al.
Published: (2025)
by: Khaled, Berkani, et al.
Published: (2025)
Quantifying Concentration Phenomena of Mean-Field Transformers in the Low-Temperature Regime
by: Alcalde, Albert, et al.
Published: (2026)
by: Alcalde, Albert, et al.
Published: (2026)
InhibiDistilbert: Knowledge Distillation for a ReLU and Addition-based Transformer
by: Zhang, Tony, et al.
Published: (2025)
by: Zhang, Tony, et al.
Published: (2025)
Similar Items
-
On The Finetuning of MLIPs Through the Lens of Iterated Maps With BPTT
by: Dramko, Evan, et al.
Published: (2025) -
Teaching and Learning under Deductive Errors
by: Telle, Jan Arne, et al.
Published: (2026) -
Autoencoded UMAP-Enhanced Clustering for Unsupervised Learning
by: Chavooshi, Malihehsadat, et al.
Published: (2025) -
Backpropagation Through Time For Networks With Long-Term Dependencies
by: Bird, George, et al.
Published: (2021) -
A Theoretical Study of (Hyper) Self-Attention through the Lens of Interactions: Representation, Training, Generalization
by: Ustaomeroglu, Muhammed, et al.
Published: (2025)