ESPFormer: Doubly-Stochastic Attention with Expected Sliced Transport Plans
Fuente:
arXiv
Saved in:
| Main Authors: | Shahbazi, Ashkan, Akbari, Elaheh, Salehi, Darian, Liu, Xinran, Naderializadeh, Navid, Kolouri, Soheil |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Expected Sliced Transport Plans
by: Liu, Xinran, et al.
Published: (2024)
by: Liu, Xinran, et al.
Published: (2024)
LUNA: Linear Universal Neural Attention with Generalization Guarantees
by: Shahbazi, Ashkan, et al.
Published: (2025)
by: Shahbazi, Ashkan, et al.
Published: (2025)
Constrained Sliced Wasserstein Embedding
by: NaderiAlizadeh, Navid, et al.
Published: (2025)
by: NaderiAlizadeh, Navid, et al.
Published: (2025)
LOTFormer: Doubly-Stochastic Linear Attention via Low-Rank Optimal Transport
by: Shahbazi, Ashkan, et al.
Published: (2025)
by: Shahbazi, Ashkan, et al.
Published: (2025)
Min Generalized Sliced Gromov Wasserstein: A Scalable Path to Gromov Wasserstein
by: Shahbazi, Ashkan, et al.
Published: (2026)
by: Shahbazi, Ashkan, et al.
Published: (2026)
Efficient Transferable Optimal Transport via Min-Sliced Transport Plans
by: Liu, Xinran, et al.
Published: (2025)
by: Liu, Xinran, et al.
Published: (2025)
Linear Spherical Sliced Optimal Transport: A Fast Metric for Comparing Spherical Data
by: Liu, Xinran, et al.
Published: (2024)
by: Liu, Xinran, et al.
Published: (2024)
Stereographic Spherical Sliced Wasserstein Distances
by: Tran, Huy, et al.
Published: (2024)
by: Tran, Huy, et al.
Published: (2024)
Understanding Learning with Sliced-Wasserstein Requires Rethinking Informative Slices
by: Tran, Huy, et al.
Published: (2024)
by: Tran, Huy, et al.
Published: (2024)
Equivariant vs. Invariant Layers: A Comparison of Backbone and Pooling for Point Cloud Classification
by: Kothapalli, Abihith, et al.
Published: (2023)
by: Kothapalli, Abihith, et al.
Published: (2023)
One Category One Prompt: Dataset Distillation using Diffusion Models
by: Abbasi, Ali, et al.
Published: (2024)
by: Abbasi, Ali, et al.
Published: (2024)
EMPEROR: Efficient Moment-Preserving Representation of Distributions
by: Liu, Xinran, et al.
Published: (2025)
by: Liu, Xinran, et al.
Published: (2025)
OT-MeanFlow3D: Bridging Optimal Transport and Meanflow for Efficient 3D Point Cloud Generation
by: Akbari, Elaheh, et al.
Published: (2025)
by: Akbari, Elaheh, et al.
Published: (2025)
ASAP: Amortized Doubly-Stochastic Attention via Sliced Dual Projection
by: Tran, Huy, et al.
Published: (2026)
by: Tran, Huy, et al.
Published: (2026)
Neural-Augmented Kelvinlet for Real-Time Soft Tissue Deformation Modeling
by: Shahbazi, Ashkan, et al.
Published: (2025)
by: Shahbazi, Ashkan, et al.
Published: (2025)
Partial Gromov-Wasserstein Metric
by: Bai, Yikun, et al.
Published: (2024)
by: Bai, Yikun, et al.
Published: (2024)
ConQuR: Corner Aligned Activation Quantization via Optimized Rotations for LLMs
by: Thrash, Chayne, et al.
Published: (2026)
by: Thrash, Chayne, et al.
Published: (2026)
Physics informed cell representations for variational formulation of multiscale problems
by: Gao, Yuxiang, et al.
Published: (2024)
by: Gao, Yuxiang, et al.
Published: (2024)
Vector-Quantized Soft Label Compression for Dataset Distillation
by: Abbasi, Ali, et al.
Published: (2026)
by: Abbasi, Ali, et al.
Published: (2026)
Linear Optimal Partial Transport Embedding
by: Bai, Yikun, et al.
Published: (2023)
by: Bai, Yikun, et al.
Published: (2023)
Low-Rank Prehab: Preparing Neural Networks for SVD Compression
by: Qin, Haoran, et al.
Published: (2025)
by: Qin, Haoran, et al.
Published: (2025)
Sinkhorn-Drifting Generative Models
by: He, Ping, et al.
Published: (2026)
by: He, Ping, et al.
Published: (2026)
IO-SVD: Input-Output Whitened SVD for Adaptive-Rank LLM Compression
by: Abbasi, Ali, et al.
Published: (2026)
by: Abbasi, Ali, et al.
Published: (2026)
Expected Batch Optimal Transport Plans and Consequences for Flow Matching
by: Boïté, Samuel, et al.
Published: (2026)
by: Boïté, Samuel, et al.
Published: (2026)
Fused Partial Gromov-Wasserstein for Structured Objects
by: Bai, Yikun, et al.
Published: (2025)
by: Bai, Yikun, et al.
Published: (2025)
Zero Sum SVD: Balancing Loss Sensitivity for Low Rank LLM Compression
by: Abbasi, Ali, et al.
Published: (2026)
by: Abbasi, Ali, et al.
Published: (2026)
Linear Partial Gromov-Wasserstein Embedding
by: Bai, Yikun, et al.
Published: (2024)
by: Bai, Yikun, et al.
Published: (2024)
Sliced-Regularized Optimal Transport
by: Nguyen, Khai
Published: (2026)
by: Nguyen, Khai
Published: (2026)
Row-Stochastic Matrices Can Provably Outperform Doubly Stochastic Matrices in Decentralized Learning
by: Liu, Bing, et al.
Published: (2025)
by: Liu, Bing, et al.
Published: (2025)
Statistical Context Detection for Deep Lifelong Reinforcement Learning
by: Dick, Jeffery, et al.
Published: (2024)
by: Dick, Jeffery, et al.
Published: (2024)
The Homogeneity Trap: Spectral Collapse in Doubly-Stochastic Deep Networks
by: Liu, Yizhi
Published: (2026)
by: Liu, Yizhi
Published: (2026)
Demystifying SGD with Doubly Stochastic Gradients
by: Kim, Kyurae, et al.
Published: (2024)
by: Kim, Kyurae, et al.
Published: (2024)
Streaming Sliced Optimal Transport
by: Nguyen, Khai
Published: (2025)
by: Nguyen, Khai
Published: (2025)
Slicing Unbalanced Optimal Transport
by: Bonet, Clément, et al.
Published: (2023)
by: Bonet, Clément, et al.
Published: (2023)
MCNC: Manifold-Constrained Reparameterization for Neural Compression
by: Thrash, Chayne, et al.
Published: (2024)
by: Thrash, Chayne, et al.
Published: (2024)
Differentiable Generalized Sliced Wasserstein Plans
by: Chapel, Laetitia, et al.
Published: (2025)
by: Chapel, Laetitia, et al.
Published: (2025)
Beyond the Laplacian: Doubly Stochastic Matrices for Graph Neural Networks
by: Hu, Zhaobo, et al.
Published: (2026)
by: Hu, Zhaobo, et al.
Published: (2026)
Expected Free Energy-based Planning as Variational Inference
by: de Vries, Bert, et al.
Published: (2025)
by: de Vries, Bert, et al.
Published: (2025)
Doubly Stochastic Mean-Shift Clustering
by: Trigano, Tom, et al.
Published: (2026)
by: Trigano, Tom, et al.
Published: (2026)
Convergence Rates for Distribution Matching with Sliced Optimal Transport
by: Thurin, Gauthier, et al.
Published: (2026)
by: Thurin, Gauthier, et al.
Published: (2026)
Similar Items
-
Expected Sliced Transport Plans
by: Liu, Xinran, et al.
Published: (2024) -
LUNA: Linear Universal Neural Attention with Generalization Guarantees
by: Shahbazi, Ashkan, et al.
Published: (2025) -
Constrained Sliced Wasserstein Embedding
by: NaderiAlizadeh, Navid, et al.
Published: (2025) -
LOTFormer: Doubly-Stochastic Linear Attention via Low-Rank Optimal Transport
by: Shahbazi, Ashkan, et al.
Published: (2025) -
Min Generalized Sliced Gromov Wasserstein: A Scalable Path to Gromov Wasserstein
by: Shahbazi, Ashkan, et al.
Published: (2026)