Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Chefer, Hila, Esser, Patrick, Lorenz, Dominik, Podell, Dustin, Raja, Vikash, Tong, Vinh, Torralba, Antonio, Rombach, Robin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scaling Rectified Flow Transformers for High-Resolution Image Synthesis
by: Esser, Patrick, et al.
Published: (2024)
by: Esser, Patrick, et al.
Published: (2024)
FlowMo: Variance-Based Flow Guidance for Coherent Motion in Video Generation
by: Shaulov, Ariel, et al.
Published: (2025)
by: Shaulov, Ariel, et al.
Published: (2025)
Fast High-Resolution Image Synthesis with Latent Adversarial Diffusion Distillation
by: Sauer, Axel, et al.
Published: (2024)
by: Sauer, Axel, et al.
Published: (2024)
A Meaningful Perturbation Metric for Evaluating Explainability Methods
by: Cohen, Danielle, et al.
Published: (2025)
by: Cohen, Danielle, et al.
Published: (2025)
MultiModal Action Conditioned Video Generation
by: Li, Yichen, et al.
Published: (2025)
by: Li, Yichen, et al.
Published: (2025)
FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space
by: Labs, Black Forest, et al.
Published: (2025)
by: Labs, Black Forest, et al.
Published: (2025)
Dataset Distillation for Pre-Trained Self-Supervised Vision Models
by: Cazenavette, George, et al.
Published: (2025)
by: Cazenavette, George, et al.
Published: (2025)
aMUSEd: An Open MUSE Reproduction
by: Patil, Suraj, et al.
Published: (2024)
by: Patil, Suraj, et al.
Published: (2024)
Contrast-X: A Multi-Modal Contrast Image Synthesis Benchmark and Universal Modality Flow Matching
by: Chen, Yifan, et al.
Published: (2026)
by: Chen, Yifan, et al.
Published: (2026)
WFM: 3D Wavelet Flow Matching for Ultrafast Multi-Modal MRI Synthesis
by: Tur, Yalcin, et al.
Published: (2026)
by: Tur, Yalcin, et al.
Published: (2026)
MMMS: Multi-Modal Multi-Surface Interactive Segmentation
by: Schön, Robin, et al.
Published: (2025)
by: Schön, Robin, et al.
Published: (2025)
Discriminative Class Tokens for Text-to-Image Diffusion Models
by: Schwartz, Idan, et al.
Published: (2023)
by: Schwartz, Idan, et al.
Published: (2023)
Multi-Modal Self-Supervised Semantic Communication
by: Zhao, Hang, et al.
Published: (2025)
by: Zhao, Hang, et al.
Published: (2025)
VideoJAM: Joint Appearance-Motion Representations for Enhanced Motion Generation in Video Models
by: Chefer, Hila, et al.
Published: (2025)
by: Chefer, Hila, et al.
Published: (2025)
Self-Supervised Partial Cycle-Consistency for Multi-View Matching
by: Taggenbrock, Fedor, et al.
Published: (2025)
by: Taggenbrock, Fedor, et al.
Published: (2025)
Decoupling Multi-Contrast Super-Resolution: Self-Supervised Implicit Re-Representation for Unpaired Cross-Modal Synthesis
by: Wu, Yinzhe, et al.
Published: (2025)
by: Wu, Yinzhe, et al.
Published: (2025)
Towards Flexible, Scalable, and Adaptive Multi-Modal Conditioned Face Synthesis
by: Ren, Jingjing, et al.
Published: (2023)
by: Ren, Jingjing, et al.
Published: (2023)
Multi-Task Multi-Modal Self-Supervised Learning for Facial Expression Recognition
by: Halawa, Marah, et al.
Published: (2024)
by: Halawa, Marah, et al.
Published: (2024)
Revisiting LRP: Positional Attribution as the Missing Ingredient for Transformer Explainability
by: Bakish, Yarden, et al.
Published: (2025)
by: Bakish, Yarden, et al.
Published: (2025)
Mixture Prototype Flow Matching for Open-Set Supervised Anomaly Detection
by: Wang, Fuyun, et al.
Published: (2026)
by: Wang, Fuyun, et al.
Published: (2026)
TeFlow: Enabling Multi-frame Supervision for Self-Supervised Feed-forward Scene Flow Estimation
by: Zhang, Qingwen, et al.
Published: (2026)
by: Zhang, Qingwen, et al.
Published: (2026)
FusionFM: All-in-One Multi-Modal Image Fusion with Flow Matching
by: Zhu, Huayi, et al.
Published: (2025)
by: Zhu, Huayi, et al.
Published: (2025)
Efficient Dataset Distillation for Pre-Trained Self-Supervised Models via Statistical Flow Matching
by: Xia, Qianxin, et al.
Published: (2026)
by: Xia, Qianxin, et al.
Published: (2026)
PoseFM: Relative Camera Pose Estimation Through Flow Matching
by: Kuczkowski, Dominik, et al.
Published: (2026)
by: Kuczkowski, Dominik, et al.
Published: (2026)
Cycle-Consistent Multi-Graph Matching for Self-Supervised Annotation of C.Elegans
by: Karg, Christoph, et al.
Published: (2025)
by: Karg, Christoph, et al.
Published: (2025)
SV3D: Novel Multi-view Synthesis and 3D Generation from a Single Image using Latent Video Diffusion
by: Voleti, Vikram, et al.
Published: (2024)
by: Voleti, Vikram, et al.
Published: (2024)
Self-Supervised Multiview Xray Matching
by: Dabboussi, Mohamad, et al.
Published: (2025)
by: Dabboussi, Mohamad, et al.
Published: (2025)
Semi-Supervised Semantic Segmentation using Redesigned Self-Training for White Blood Cells
by: Luu, Vinh Quoc, et al.
Published: (2024)
by: Luu, Vinh Quoc, et al.
Published: (2024)
MIRROR: Multi-Modal Pathological Self-Supervised Representation Learning via Modality Alignment and Retention
by: Wang, Tianyi, et al.
Published: (2025)
by: Wang, Tianyi, et al.
Published: (2025)
NorMatch: Matching Normalizing Flows with Discriminative Classifiers for Semi-Supervised Learning
by: Deng, Zhongying, et al.
Published: (2022)
by: Deng, Zhongying, et al.
Published: (2022)
Self-Supervised Spatial Correspondence Across Modalities
by: Shrivastava, Ayush, et al.
Published: (2025)
by: Shrivastava, Ayush, et al.
Published: (2025)
Still-Moving: Customized Video Generation without Customized Video Data
by: Chefer, Hila, et al.
Published: (2024)
by: Chefer, Hila, et al.
Published: (2024)
Self-Supervised Multi-Frame Neural Scene Flow
by: Liu, Dongrui, et al.
Published: (2024)
by: Liu, Dongrui, et al.
Published: (2024)
DoGFlow: Self-Supervised LiDAR Scene Flow via Cross-Modal Doppler Guidance
by: Khoche, Ajinkya, et al.
Published: (2025)
by: Khoche, Ajinkya, et al.
Published: (2025)
FlowCast: Trajectory Forecasting for Scalable Zero-Cost Speculative Flow Matching
by: Bajpai, Divya Jyoti, et al.
Published: (2026)
by: Bajpai, Divya Jyoti, et al.
Published: (2026)
Multi-Modal Monocular Endoscopic Depth and Pose Estimation with Edge-Guided Self-Supervision
by: Ju, Xinwei, et al.
Published: (2026)
by: Ju, Xinwei, et al.
Published: (2026)
Beyond Instance-Level Self-Supervision in 3D Multi-Modal Medical Imaging
by: Pan, Tan, et al.
Published: (2026)
by: Pan, Tan, et al.
Published: (2026)
Enhancing Cardiovascular Disease Prediction through Multi-Modal Self-Supervised Learning
by: Girlanda, Francesco, et al.
Published: (2024)
by: Girlanda, Francesco, et al.
Published: (2024)
FlowSDF: Flow Matching for Medical Image Segmentation Using Distance Transforms
by: Bogensperger, Lea, et al.
Published: (2024)
by: Bogensperger, Lea, et al.
Published: (2024)
Information Flow in Self-Supervised Learning
by: Tan, Zhiquan, et al.
Published: (2023)
by: Tan, Zhiquan, et al.
Published: (2023)
Similar Items
-
Scaling Rectified Flow Transformers for High-Resolution Image Synthesis
by: Esser, Patrick, et al.
Published: (2024) -
FlowMo: Variance-Based Flow Guidance for Coherent Motion in Video Generation
by: Shaulov, Ariel, et al.
Published: (2025) -
Fast High-Resolution Image Synthesis with Latent Adversarial Diffusion Distillation
by: Sauer, Axel, et al.
Published: (2024) -
A Meaningful Perturbation Metric for Evaluating Explainability Methods
by: Cohen, Danielle, et al.
Published: (2025) -
MultiModal Action Conditioned Video Generation
by: Li, Yichen, et al.
Published: (2025)