On Fitting Flow Models with Large Sinkhorn Couplings
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Stephen, Mousavi-Hosseini, Alireza, Klein, Michal, Cuturi, Marco |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Flow Matching with Semidiscrete Couplings
by: Mousavi-Hosseini, Alireza, et al.
Published: (2025)
by: Mousavi-Hosseini, Alireza, et al.
Published: (2025)
Multivariate Conformal Prediction using Optimal Transport
by: Klein, Michal, et al.
Published: (2025)
by: Klein, Michal, et al.
Published: (2025)
Contrasting Multiple Representations with the Multi-Marginal Matching Gap
by: Piran, Zoe, et al.
Published: (2024)
by: Piran, Zoe, et al.
Published: (2024)
The Coupling Within: Flow Matching via Distilled Normalizing Flows
by: Berthelot, David, et al.
Published: (2026)
by: Berthelot, David, et al.
Published: (2026)
GENOT: Entropic (Gromov) Wasserstein Flow Matching with Applications to Single-Cell Genomics
by: Klein, Dominik, et al.
Published: (2023)
by: Klein, Dominik, et al.
Published: (2023)
Nectar: Neural Estimation of Cached-Token Attention via Regression
by: Monteiro, João, et al.
Published: (2026)
by: Monteiro, João, et al.
Published: (2026)
Amortizing Maximum Inner Product Search with Learned Support Functions
by: Olausson, Theo X., et al.
Published: (2026)
by: Olausson, Theo X., et al.
Published: (2026)
Simple ReFlow: Improved Techniques for Fast Flow Models
by: Kim, Beomsu, et al.
Published: (2024)
by: Kim, Beomsu, et al.
Published: (2024)
Post-Training with Policy Gradients: Optimality and the Base Model Barrier
by: Mousavi-Hosseini, Alireza, et al.
Published: (2026)
by: Mousavi-Hosseini, Alireza, et al.
Published: (2026)
Robust Feature Learning for Multi-Index Models in High Dimensions
by: Mousavi-Hosseini, Alireza, et al.
Published: (2024)
by: Mousavi-Hosseini, Alireza, et al.
Published: (2024)
Learning Multi-Index Models with Neural Networks via Mean-Field Langevin Dynamics
by: Mousavi-Hosseini, Alireza, et al.
Published: (2024)
by: Mousavi-Hosseini, Alireza, et al.
Published: (2024)
Neural Sinkhorn Gradient Flow
by: Zhu, Huminhao, et al.
Published: (2024)
by: Zhu, Huminhao, et al.
Published: (2024)
Careful with that Scalpel: Improving Gradient Surgery with an EMA
by: Hsieh, Yu-Guan, et al.
Published: (2024)
by: Hsieh, Yu-Guan, et al.
Published: (2024)
On a Neural Implementation of Brenier's Polar Factorization
by: Vesseron, Nina, et al.
Published: (2024)
by: Vesseron, Nina, et al.
Published: (2024)
Mean-Field Langevin Dynamics for Signed Measures via a Bilevel Approach
by: Wang, Guillaume, et al.
Published: (2024)
by: Wang, Guillaume, et al.
Published: (2024)
Learning Elastic Costs to Shape Monge Displacements
by: Klein, Michal, et al.
Published: (2023)
by: Klein, Michal, et al.
Published: (2023)
Progressive Entropic Optimal Transport Solvers
by: Kassraie, Parnian, et al.
Published: (2024)
by: Kassraie, Parnian, et al.
Published: (2024)
The Data-Quality Illusion: Rethinking Classifier-Based Quality Filtering for LLM Pretraining
by: Saada, Thiziri Nait, et al.
Published: (2025)
by: Saada, Thiziri Nait, et al.
Published: (2025)
From Information to Generative Exponent: Learning Rate Induces Phase Transitions in SGD
by: Tsiolis, Konstantinos Christopher, et al.
Published: (2025)
by: Tsiolis, Konstantinos Christopher, et al.
Published: (2025)
Completed Hyperparameter Transfer across Modules, Width, Depth, Batch and Duration
by: Mlodozeniec, Bruno, et al.
Published: (2025)
by: Mlodozeniec, Bruno, et al.
Published: (2025)
When Do Transformers Outperform Feedforward and Recurrent Networks? A Statistical Perspective
by: Mousavi-Hosseini, Alireza, et al.
Published: (2025)
by: Mousavi-Hosseini, Alireza, et al.
Published: (2025)
Sample and Map from a Single Convex Potential: Generation using Conjugate Moment Measures
by: Vesseron, Nina, et al.
Published: (2025)
by: Vesseron, Nina, et al.
Published: (2025)
The Geometries of Truth Are Orthogonal Across Tasks
by: Azizian, Waiss, et al.
Published: (2025)
by: Azizian, Waiss, et al.
Published: (2025)
Sinkhorn-Drifting Generative Models
by: He, Ping, et al.
Published: (2026)
by: He, Ping, et al.
Published: (2026)
Scaling Categorical Flow Maps
by: Davis, Oscar, et al.
Published: (2026)
by: Davis, Oscar, et al.
Published: (2026)
A Separation in Heavy-Tailed Sampling: Gaussian vs. Stable Oracles for Proximal Samplers
by: He, Ye, et al.
Published: (2024)
by: He, Ye, et al.
Published: (2024)
Locking Pretrained Weights via Deep Low-Rank Residual Distillation
by: Sakamoto, Keitaro, et al.
Published: (2026)
by: Sakamoto, Keitaro, et al.
Published: (2026)
On Sinkhorn's Algorithm and Choice Modeling
by: Qu, Zhaonan, et al.
Published: (2023)
by: Qu, Zhaonan, et al.
Published: (2023)
A Specialized Semismooth Newton Method for Kernel-Based Optimal Transport
by: Lin, Tianyi, et al.
Published: (2023)
by: Lin, Tianyi, et al.
Published: (2023)
Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing
by: Filippova, Anastasiia, et al.
Published: (2026)
by: Filippova, Anastasiia, et al.
Published: (2026)
Federated Sinkhorn
by: Kulcsar, Jeremy, et al.
Published: (2025)
by: Kulcsar, Jeremy, et al.
Published: (2025)
Confident Sinkhorn Allocation for Pseudo-Labeling
by: Nguyen, Vu, et al.
Published: (2022)
by: Nguyen, Vu, et al.
Published: (2022)
Controlling Language and Diffusion Models by Transporting Activations
by: Rodriguez, Pau, et al.
Published: (2024)
by: Rodriguez, Pau, et al.
Published: (2024)
Sinkhorn Distance Minimization for Knowledge Distillation
by: Cui, Xiao, et al.
Published: (2024)
by: Cui, Xiao, et al.
Published: (2024)
Sinkhorn Distributionally Robust Optimization
by: Wang, Jie, et al.
Published: (2021)
by: Wang, Jie, et al.
Published: (2021)
Accelerated Sinkhorn Algorithms for Partial Optimal Transport
by: Truong, Nghia Thu, et al.
Published: (2026)
by: Truong, Nghia Thu, et al.
Published: (2026)
HyperTransport: Amortized Conditioning of T2I Generative Models
by: Maiorca, Valentino, et al.
Published: (2026)
by: Maiorca, Valentino, et al.
Published: (2026)
Prior-Fitted Functional Flow: In-Context Generative Models for Pharmacokinetics
by: Ojeda, César, et al.
Published: (2026)
by: Ojeda, César, et al.
Published: (2026)
Selective Sinkhorn Routing for Improved Sparse Mixture of Experts
by: Nguyen, Duc Anh, et al.
Published: (2025)
by: Nguyen, Duc Anh, et al.
Published: (2025)
Iterative Sampling Methods for Sinkhorn Distributionally Robust Optimization
by: Wang, Jie
Published: (2025)
by: Wang, Jie
Published: (2025)
Similar Items
-
Flow Matching with Semidiscrete Couplings
by: Mousavi-Hosseini, Alireza, et al.
Published: (2025) -
Multivariate Conformal Prediction using Optimal Transport
by: Klein, Michal, et al.
Published: (2025) -
Contrasting Multiple Representations with the Multi-Marginal Matching Gap
by: Piran, Zoe, et al.
Published: (2024) -
The Coupling Within: Flow Matching via Distilled Normalizing Flows
by: Berthelot, David, et al.
Published: (2026) -
GENOT: Entropic (Gromov) Wasserstein Flow Matching with Applications to Single-Cell Genomics
by: Klein, Dominik, et al.
Published: (2023)