OneFlow: Concurrent Mixed-Modal and Interleaved Generation with Edit Flows
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, John, Havasi, Marton, Berrada, Tariq, Zettlemoyer, Luke, Chen, Ricky T. Q. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Edit Flows: Flow Matching with Edit Operations
by: Havasi, Marton, et al.
Published: (2025)
by: Havasi, Marton, et al.
Published: (2025)
Flowception: Temporally Expansive Flow Matching for Video Generation
by: Ifriqi, Tariq Berrada, et al.
Published: (2025)
by: Ifriqi, Tariq Berrada, et al.
Published: (2025)
Flow Matching with General Discrete Paths: A Kinetic-Optimal Perspective
by: Shaul, Neta, et al.
Published: (2024)
by: Shaul, Neta, et al.
Published: (2024)
Flow Matching on General Geometries
by: Chen, Ricky T. Q., et al.
Published: (2023)
by: Chen, Ricky T. Q., et al.
Published: (2023)
Generator Matching: Generative modeling with arbitrary Markov processes
by: Holderrieth, Peter, et al.
Published: (2024)
by: Holderrieth, Peter, et al.
Published: (2024)
FlowMM: Generating Materials with Riemannian Flow Matching
by: Miller, Benjamin Kurt, et al.
Published: (2024)
by: Miller, Benjamin Kurt, et al.
Published: (2024)
FlowLLM: Flow Matching for Material Generation with Large Language Models as Base Distributions
by: Sriram, Anuroop, et al.
Published: (2024)
by: Sriram, Anuroop, et al.
Published: (2024)
Entropy Rectifying Guidance for Diffusion and Flow Models
by: Ifriqi, Tariq Berrada, et al.
Published: (2025)
by: Ifriqi, Tariq Berrada, et al.
Published: (2025)
GLASS Flows: Transition Sampling for Alignment of Flow and Diffusion Models
by: Holderrieth, Peter, et al.
Published: (2025)
by: Holderrieth, Peter, et al.
Published: (2025)
Distributional GFlowNets with Quantile Flows
by: Zhang, Dinghuai, et al.
Published: (2023)
by: Zhang, Dinghuai, et al.
Published: (2023)
Understanding and Mitigating Tokenization Bias in Language Models
by: Phan, Buu, et al.
Published: (2024)
by: Phan, Buu, et al.
Published: (2024)
Diffusion Generative Flow Samplers: Improving learning signals through partial trajectory optimization
by: Zhang, Dinghuai, et al.
Published: (2023)
by: Zhang, Dinghuai, et al.
Published: (2023)
Discrete Flow Matching
by: Gat, Itai, et al.
Published: (2024)
by: Gat, Itai, et al.
Published: (2024)
On Improved Conditioning Mechanisms and Pre-training Strategies for Diffusion Models
by: Ifriqi, Tariq Berrada, et al.
Published: (2024)
by: Ifriqi, Tariq Berrada, et al.
Published: (2024)
Training-free Linear Image Inverses via Flows
by: Pokle, Ashwini, et al.
Published: (2023)
by: Pokle, Ashwini, et al.
Published: (2023)
Comparing Hallucination Detection Metrics for Multilingual Generation
by: Kang, Haoqiang, et al.
Published: (2024)
by: Kang, Haoqiang, et al.
Published: (2024)
From Flow to One Step: Real-Time Multi-Modal Trajectory Policies via Implicit Maximum Likelihood Estimation-based Distribution Distillation
by: Dong, Ju, et al.
Published: (2026)
by: Dong, Ju, et al.
Published: (2026)
Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
by: Zhou, Chunting, et al.
Published: (2024)
by: Zhou, Chunting, et al.
Published: (2024)
Variational Flow Models: Flowing in Your Style
by: Do, Kien, et al.
Published: (2024)
by: Do, Kien, et al.
Published: (2024)
FusionFM: All-in-One Multi-Modal Image Fusion with Flow Matching
by: Zhu, Huayi, et al.
Published: (2025)
by: Zhu, Huayi, et al.
Published: (2025)
Unlocking Pre-trained Image Backbones for Semantic Image Synthesis
by: Berrada, Tariq, et al.
Published: (2023)
by: Berrada, Tariq, et al.
Published: (2023)
TV2TV: A Unified Framework for Interleaved Language and Video Generation
by: Han, Xiaochuang, et al.
Published: (2025)
by: Han, Xiaochuang, et al.
Published: (2025)
Bespoke Non-Stationary Solvers for Fast Sampling of Diffusion and Flow Models
by: Shaul, Neta, et al.
Published: (2024)
by: Shaul, Neta, et al.
Published: (2024)
LPDP: Inference-Time Reward Control for Variable-Length DNA Generation with Edit Flows
by: Kim, Jeongchan, et al.
Published: (2026)
by: Kim, Jeongchan, et al.
Published: (2026)
ContextFlow++: Generalist-Specialist Flow-based Generative Models with Mixed-Variable Context Encoding
by: Gudovskiy, Denis, et al.
Published: (2024)
by: Gudovskiy, Denis, et al.
Published: (2024)
Interleaved-Modal Chain-of-Thought
by: Gao, Jun, et al.
Published: (2024)
by: Gao, Jun, et al.
Published: (2024)
AlphaFlowTSE: One-Step Generative Target Speaker Extraction via Conditional AlphaFlow
by: Li, Duojia, et al.
Published: (2026)
by: Li, Duojia, et al.
Published: (2026)
Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity
by: Liang, Weixin, et al.
Published: (2025)
by: Liang, Weixin, et al.
Published: (2025)
DiFlowDubber: Discrete Flow Matching for Automated Video Dubbing via Cross-Modal Alignment and Synchronization
by: Nguyen, Ngoc-Son, et al.
Published: (2026)
by: Nguyen, Ngoc-Son, et al.
Published: (2026)
Multimodal RewardBench: Holistic Evaluation of Reward Models for Vision Language Models
by: Yasunaga, Michihiro, et al.
Published: (2025)
by: Yasunaga, Michihiro, et al.
Published: (2025)
Orthus: Autoregressive Interleaved Image-Text Generation with Modality-Specific Heads
by: Kou, Siqi, et al.
Published: (2024)
by: Kou, Siqi, et al.
Published: (2024)
(Mis)Fitting: A Survey of Scaling Laws
by: Li, Margaret, et al.
Published: (2025)
by: Li, Margaret, et al.
Published: (2025)
One-Step Generative Policies with Q-Learning: A Reformulation of MeanFlow
by: Wang, Zeyuan, et al.
Published: (2025)
by: Wang, Zeyuan, et al.
Published: (2025)
VITA-Audio: Fast Interleaved Cross-Modal Token Generation for Efficient Large Speech-Language Model
by: Long, Zuwei, et al.
Published: (2025)
by: Long, Zuwei, et al.
Published: (2025)
Mean-Flow based One-Step Vision-Language-Action
by: Chen, Yang, et al.
Published: (2026)
by: Chen, Yang, et al.
Published: (2026)
Mean Flow Policy with Instantaneous Velocity Constraint for One-step Action Generation
by: Zhan, Guojian, et al.
Published: (2026)
by: Zhan, Guojian, et al.
Published: (2026)
OZSpeech: One-step Zero-shot Speech Synthesis with Learned-Prior-Conditioned Flow Matching
by: Huynh-Nguyen, Hieu-Nghia, et al.
Published: (2025)
by: Huynh-Nguyen, Hieu-Nghia, et al.
Published: (2025)
MoMa: Efficient Early-Fusion Pre-training with Mixture of Modality-Aware Experts
by: Lin, Xi Victoria, et al.
Published: (2024)
by: Lin, Xi Victoria, et al.
Published: (2024)
SNR-Edit: Structure-Aware Noise Rectification for Inversion-Free Flow-Based Editing
by: Jiang, Lifan, et al.
Published: (2026)
by: Jiang, Lifan, et al.
Published: (2026)
Multimodal Crystal Flow: Any-to-Any Modality Generation for Unified Crystal Modeling
by: Seong, Kiyoung, et al.
Published: (2026)
by: Seong, Kiyoung, et al.
Published: (2026)
Similar Items
-
Edit Flows: Flow Matching with Edit Operations
by: Havasi, Marton, et al.
Published: (2025) -
Flowception: Temporally Expansive Flow Matching for Video Generation
by: Ifriqi, Tariq Berrada, et al.
Published: (2025) -
Flow Matching with General Discrete Paths: A Kinetic-Optimal Perspective
by: Shaul, Neta, et al.
Published: (2024) -
Flow Matching on General Geometries
by: Chen, Ricky T. Q., et al.
Published: (2023) -
Generator Matching: Generative modeling with arbitrary Markov processes
by: Holderrieth, Peter, et al.
Published: (2024)