Saved in:
| Main Authors: | Bergmeister, Andreas, Lal, Manish Krishan, Jegelka, Stefanie, Sra, Suvrit |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.05878 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-fluctuation phase transitions reveal sampling dynamics in diffusion models
by: Ramachandran, Sai Niranjan, et al.
Published: (2025)
by: Ramachandran, Sai Niranjan, et al.
Published: (2025)
Toward generalizable learning of all (linear) first-order methods via memory augmented Transformers
by: Dutta, Sanchayan, et al.
Published: (2024)
by: Dutta, Sanchayan, et al.
Published: (2024)
Linearly Convergent Algorithms for Nonsmooth Problems with Unknown Smooth Pieces
by: Zhang, Zhe, et al.
Published: (2025)
by: Zhang, Zhe, et al.
Published: (2025)
Trees to Flows and Back: Unifying Decision Trees and Diffusion Models
by: Ramachandran, Sai Niranjan, et al.
Published: (2026)
by: Ramachandran, Sai Niranjan, et al.
Published: (2026)
Transformers Implement Functional Gradient Descent to Learn Non-Linear Functions In Context
by: Cheng, Xiang, et al.
Published: (2023)
by: Cheng, Xiang, et al.
Published: (2023)
Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models
by: Bergmeister, Andreas, et al.
Published: (2026)
by: Bergmeister, Andreas, et al.
Published: (2026)
Learning with Boolean threshold functions
by: Elser, Veit, et al.
Published: (2026)
by: Elser, Veit, et al.
Published: (2026)
Implicit Bias in Matrix Factorization and its Explicit Realization in a New Architecture
by: Hou, Yikun, et al.
Published: (2025)
by: Hou, Yikun, et al.
Published: (2025)
Graph Transformers Dream of Electric Flow
by: Cheng, Xiang, et al.
Published: (2024)
by: Cheng, Xiang, et al.
Published: (2024)
Efficient Sampling on Riemannian Manifolds via Langevin MCMC
by: Cheng, Xiang, et al.
Published: (2024)
by: Cheng, Xiang, et al.
Published: (2024)
How to escape sharp minima with random perturbations
by: Ahn, Kwangjun, et al.
Published: (2023)
by: Ahn, Kwangjun, et al.
Published: (2023)
Riemannian Bilevel Optimization
by: Dutta, Sanchayan, et al.
Published: (2024)
by: Dutta, Sanchayan, et al.
Published: (2024)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part I
by: Tian, Yi, et al.
Published: (2022)
by: Tian, Yi, et al.
Published: (2022)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part II
by: Tian, Yi, et al.
Published: (2026)
by: Tian, Yi, et al.
Published: (2026)
Revisiting Frank-Wolfe for Structured Nonconvex Optimization
by: Maskan, Hoomaan, et al.
Published: (2025)
by: Maskan, Hoomaan, et al.
Published: (2025)
Tight Generalization Bounds for Noiseless Inverse Optimization
by: Fatemi, Pouria, et al.
Published: (2026)
by: Fatemi, Pouria, et al.
Published: (2026)
First-Order Methods for Linearly Constrained Bilevel Optimization
by: Kornowski, Guy, et al.
Published: (2024)
by: Kornowski, Guy, et al.
Published: (2024)
Higher-Order Graphon Neural Networks: Approximation and Cut Distance
by: Herbst, Daniel, et al.
Published: (2025)
by: Herbst, Daniel, et al.
Published: (2025)
Sample Complexity Bounds for Estimating Probability Divergences under Invariances
by: Tahmasebi, Behrooz, et al.
Published: (2023)
by: Tahmasebi, Behrooz, et al.
Published: (2023)
The Exact Sample Complexity Gain from Invariances for Kernel Regression
by: Tahmasebi, Behrooz, et al.
Published: (2023)
by: Tahmasebi, Behrooz, et al.
Published: (2023)
Linear attention is (maybe) all you need (to understand transformer optimization)
by: Ahn, Kwangjun, et al.
Published: (2023)
by: Ahn, Kwangjun, et al.
Published: (2023)
A Poincaré Inequality and Consistency Results for Signal Sampling on Large Graphs
by: Le, Thien, et al.
Published: (2023)
by: Le, Thien, et al.
Published: (2023)
Generalization, Expressivity, and Universality of Graph Neural Networks on Attributed Graphs
by: Rauchwerger, Levi, et al.
Published: (2024)
by: Rauchwerger, Levi, et al.
Published: (2024)
Neural Networks With Dense Weights Are Not Universal Approximators
by: Rauchwerger, Levi, et al.
Published: (2026)
by: Rauchwerger, Levi, et al.
Published: (2026)
Counting Substructures with Higher-Order Graph Neural Networks: Possibility and Impossibility Results
by: Tahmasebi, Behrooz, et al.
Published: (2020)
by: Tahmasebi, Behrooz, et al.
Published: (2020)
The Flow-Limit of Reflect-Reflect-Relax: Existence, Stability, and Discrete-Time Behavior
by: Lal, Manish Krishan
Published: (2025)
by: Lal, Manish Krishan
Published: (2025)
Backpropagation from KL Projections: Differential and Exact I-Projection Correspondences
by: Lal, Manish Krishan
Published: (2025)
by: Lal, Manish Krishan
Published: (2025)
On the hardness of learning under symmetries
by: Kiani, Bobak T., et al.
Published: (2024)
by: Kiani, Bobak T., et al.
Published: (2024)
Computing Brascamp-Lieb Constants through the lens of Thompson Geometry
by: Weber, Melanie, et al.
Published: (2022)
by: Weber, Melanie, et al.
Published: (2022)
Efficient and Scalable Graph Generation through Iterative Local Expansion
by: Bergmeister, Andreas, et al.
Published: (2023)
by: Bergmeister, Andreas, et al.
Published: (2023)
On the Emergence of Position Bias in Transformers
by: Wu, Xinyi, et al.
Published: (2025)
by: Wu, Xinyi, et al.
Published: (2025)
Sequential-Parallel Duality in Prefix Scannable Models
by: Yau, Morris, et al.
Published: (2025)
by: Yau, Morris, et al.
Published: (2025)
A Canonicalization Perspective on Invariant and Equivariant Learning
by: Ma, George, et al.
Published: (2024)
by: Ma, George, et al.
Published: (2024)
A Universal Class of Sharpness-Aware Minimization Algorithms
by: Tahmasebi, Behrooz, et al.
Published: (2024)
by: Tahmasebi, Behrooz, et al.
Published: (2024)
LieAugmenter: Equivariant Learning by Discovering Symmetries with Learnable Augmentations
by: Santos-Escriche, Eduardo, et al.
Published: (2025)
by: Santos-Escriche, Eduardo, et al.
Published: (2025)
Learning with Exact Invariances in Polynomial Time
by: Soleymani, Ashkan, et al.
Published: (2025)
by: Soleymani, Ashkan, et al.
Published: (2025)
Survey on Generalization Theory for Graph Neural Networks
by: Vasileiou, Antonis, et al.
Published: (2025)
by: Vasileiou, Antonis, et al.
Published: (2025)
Near-Optimal Algorithms for Group Distributionally Robust Optimization and Beyond
by: Soma, Tasuku, et al.
Published: (2022)
by: Soma, Tasuku, et al.
Published: (2022)
Global Reinforcement Learning: Beyond Linear and Convex Rewards via Submodular Semi-gradient Methods
by: De Santi, Riccardo, et al.
Published: (2024)
by: De Santi, Riccardo, et al.
Published: (2024)
parallelcbf: A composable safety-filter and auditability framework for tensor-parallel reinforcement learning
by: Lu, Yijun, et al.
Published: (2026)
by: Lu, Yijun, et al.
Published: (2026)
Similar Items
-
Cross-fluctuation phase transitions reveal sampling dynamics in diffusion models
by: Ramachandran, Sai Niranjan, et al.
Published: (2025) -
Toward generalizable learning of all (linear) first-order methods via memory augmented Transformers
by: Dutta, Sanchayan, et al.
Published: (2024) -
Linearly Convergent Algorithms for Nonsmooth Problems with Unknown Smooth Pieces
by: Zhang, Zhe, et al.
Published: (2025) -
Trees to Flows and Back: Unifying Decision Trees and Diffusion Models
by: Ramachandran, Sai Niranjan, et al.
Published: (2026) -
Transformers Implement Functional Gradient Descent to Learn Non-Linear Functions In Context
by: Cheng, Xiang, et al.
Published: (2023)