Flexible Language Modeling in Continuous Space with Transformer-based Autoregressive Flows
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Ruixiang, Zhai, Shuangfei, Gu, Jiatao, Zhang, Yizhe, Zheng, Huangjie, Chen, Tianrong, Bautista, Miguel Angel, Susskind, Josh, Jaitly, Navdeep |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Normalizing Flows are Capable Generative Models
by: Zhai, Shuangfei, et al.
Published: (2024)
by: Zhai, Shuangfei, et al.
Published: (2024)
Matryoshka Diffusion Models
by: Gu, Jiatao, et al.
Published: (2023)
by: Gu, Jiatao, et al.
Published: (2023)
Improving GFlowNets for Text-to-Image Diffusion Alignment
by: Zhang, Dinghuai, et al.
Published: (2024)
by: Zhang, Dinghuai, et al.
Published: (2024)
DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation
by: Gu, Jiatao, et al.
Published: (2024)
by: Gu, Jiatao, et al.
Published: (2024)
PLANNER: Generating Diversified Paragraph via Latent Language Diffusion Model
by: Zhang, Yizhe, et al.
Published: (2023)
by: Zhang, Yizhe, et al.
Published: (2023)
How Far Are We from Intelligent Visual Deductive Reasoning?
by: Zhang, Yizhe, et al.
Published: (2024)
by: Zhang, Yizhe, et al.
Published: (2024)
TypeScore: A Text Fidelity Metric for Text-to-Image Generative Models
by: Sampaio, Georgia Gabriela, et al.
Published: (2024)
by: Sampaio, Georgia Gabriela, et al.
Published: (2024)
Kaleido Diffusion: Improving Conditional Diffusion Models with Autoregressive Latent Modeling
by: Gu, Jiatao, et al.
Published: (2024)
by: Gu, Jiatao, et al.
Published: (2024)
Continuously Augmented Discrete Diffusion model for Categorical Generative Modeling
by: Zheng, Huangjie, et al.
Published: (2025)
by: Zheng, Huangjie, et al.
Published: (2025)
STARFlow: Scaling Latent Normalizing Flows for High-resolution Image Synthesis
by: Gu, Jiatao, et al.
Published: (2025)
by: Gu, Jiatao, et al.
Published: (2025)
TADA: Improved Diffusion Sampling with Training-free Augmented Dynamics
by: Chen, Tianrong, et al.
Published: (2025)
by: Chen, Tianrong, et al.
Published: (2025)
DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation
by: Gong, Shansan, et al.
Published: (2025)
by: Gong, Shansan, et al.
Published: (2025)
Target Concrete Score Matching: A Holistic Framework for Discrete Diffusion
by: Zhang, Ruixiang, et al.
Published: (2025)
by: Zhang, Ruixiang, et al.
Published: (2025)
Normalizing Flows with Iterative Denoising
by: Chen, Tianrong, et al.
Published: (2026)
by: Chen, Tianrong, et al.
Published: (2026)
STARFlow2: Bridging Language Models and Normalizing Flows for Unified Multimodal Generation
by: Shen, Ying, et al.
Published: (2026)
by: Shen, Ying, et al.
Published: (2026)
STARFlow-V: End-to-End Video Generative Modeling with Normalizing Flows
by: Gu, Jiatao, et al.
Published: (2025)
by: Gu, Jiatao, et al.
Published: (2025)
Normalizing Trajectory Models
by: Gu, Jiatao, et al.
Published: (2026)
by: Gu, Jiatao, et al.
Published: (2026)
Embarrassingly Simple Self-Distillation Improves Code Generation
by: Zhang, Ruixiang, et al.
Published: (2026)
by: Zhang, Ruixiang, et al.
Published: (2026)
The Coupling Within: Flow Matching via Distilled Normalizing Flows
by: Berthelot, David, et al.
Published: (2026)
by: Berthelot, David, et al.
Published: (2026)
SimpleFold: Folding Proteins is Simpler than You Think
by: Wang, Yuyang, et al.
Published: (2025)
by: Wang, Yuyang, et al.
Published: (2025)
SAGE: Steering Dialog Generation with Future-Aware State-Action Augmentation
by: Zhang, Yizhe, et al.
Published: (2025)
by: Zhang, Yizhe, et al.
Published: (2025)
Generative Modeling with Phase Stochastic Bridges
by: Chen, Tianrong, et al.
Published: (2023)
by: Chen, Tianrong, et al.
Published: (2023)
What Makes the Preferred Thinking Direction for LLMs in Multiple-choice Questions?
by: Zhang, Yizhe, et al.
Published: (2025)
by: Zhang, Yizhe, et al.
Published: (2025)
Many-to-many Image Generation with Auto-regressive Diffusion Models
by: Shen, Ying, et al.
Published: (2024)
by: Shen, Ying, et al.
Published: (2024)
World-consistent Video Diffusion with Explicit 3D Modeling
by: Zhang, Qihang, et al.
Published: (2024)
by: Zhang, Qihang, et al.
Published: (2024)
ChipChat: Low-Latency Cascaded Conversational Agent in MLX
by: Likhomanenko, Tatiana, et al.
Published: (2025)
by: Likhomanenko, Tatiana, et al.
Published: (2025)
Probing the Multi-turn Planning Capabilities of LLMs via 20 Question Games
by: Zhang, Yizhe, et al.
Published: (2023)
by: Zhang, Yizhe, et al.
Published: (2023)
INRFlow: Flow Matching for INRs in Ambient Space
by: Wang, Yuyang, et al.
Published: (2024)
by: Wang, Yuyang, et al.
Published: (2024)
Swallowing the Bitter Pill: Simplified Scalable Conformer Generation
by: Wang, Yuyang, et al.
Published: (2023)
by: Wang, Yuyang, et al.
Published: (2023)
Aggregate-and-Adapt Natural Language Prompts for Downstream Generalization of CLIP
by: Huang, Chen, et al.
Published: (2024)
by: Huang, Chen, et al.
Published: (2024)
Divide-or-Conquer? Which Part Should You Distill Your LLM?
by: Wu, Zhuofeng, et al.
Published: (2024)
by: Wu, Zhuofeng, et al.
Published: (2024)
Scalable Pre-training of Large Autoregressive Image Models
by: El-Nouby, Alaaeldin, et al.
Published: (2024)
by: El-Nouby, Alaaeldin, et al.
Published: (2024)
3D Shape Tokenization via Latent Flow Matching
by: Chang, Jen-Hao Rick, et al.
Published: (2024)
by: Chang, Jen-Hao Rick, et al.
Published: (2024)
KGLens: Towards Efficient and Effective Knowledge Probing of Large Language Models with Knowledge Graphs
by: Zheng, Shangshang, et al.
Published: (2023)
by: Zheng, Shangshang, et al.
Published: (2023)
Rephrasing the Web: A Recipe for Compute and Data-Efficient Language Modeling
by: Maini, Pratyush, et al.
Published: (2024)
by: Maini, Pratyush, et al.
Published: (2024)
dMel: Speech Tokenization made Simple
by: Bai, Richard He, et al.
Published: (2024)
by: Bai, Richard He, et al.
Published: (2024)
CLaRa: Bridging Retrieval and Generation with Continuous Latent Reasoning
by: He, Jie, et al.
Published: (2025)
by: He, Jie, et al.
Published: (2025)
Omni-Router: Sharing Routing Decisions in Sparse Mixture-of-Experts for Speech Recognition
by: Gu, Zijin, et al.
Published: (2025)
by: Gu, Zijin, et al.
Published: (2025)
Eliciting In-context Retrieval and Reasoning for Long-context Large Language Models
by: Qiu, Yifu, et al.
Published: (2025)
by: Qiu, Yifu, et al.
Published: (2025)
Score Distillation of Flow Matching Models
by: Zhou, Mingyuan, et al.
Published: (2025)
by: Zhou, Mingyuan, et al.
Published: (2025)
Similar Items
-
Normalizing Flows are Capable Generative Models
by: Zhai, Shuangfei, et al.
Published: (2024) -
Matryoshka Diffusion Models
by: Gu, Jiatao, et al.
Published: (2023) -
Improving GFlowNets for Text-to-Image Diffusion Alignment
by: Zhang, Dinghuai, et al.
Published: (2024) -
DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation
by: Gu, Jiatao, et al.
Published: (2024) -
PLANNER: Generating Diversified Paragraph via Latent Language Diffusion Model
by: Zhang, Yizhe, et al.
Published: (2023)