Latent Forcing: Reordering the Diffusion Trajectory for Pixel-Space Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Baade, Alan, Chan, Eric Ryan, Sargent, Kyle, Chen, Changan, Johnson, Justin, Adeli, Ehsan, Fei-Fei, Li |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Flow to the Mode: Mode-Seeking Diffusion Autoencoders for State-of-the-Art Image Tokenization
di: Sargent, Kyle, et al.
Pubblicazione: (2025)
di: Sargent, Kyle, et al.
Pubblicazione: (2025)
OccFusion: Rendering Occluded Humans with Generative Diffusion Priors
di: Sun, Adam, et al.
Pubblicazione: (2024)
di: Sun, Adam, et al.
Pubblicazione: (2024)
Discovering Latent Graphs with GFlowNets for Diverse Conditional Image Generation
di: Trang, Bailey, et al.
Pubblicazione: (2025)
di: Trang, Bailey, et al.
Pubblicazione: (2025)
The Language of Motion: Unifying Verbal and Non-verbal Language of 3D Human Motion
di: Chen, Changan, et al.
Pubblicazione: (2024)
di: Chen, Changan, et al.
Pubblicazione: (2024)
GPIC: A Giant Permissive Image Corpus for Visual Generation
di: Chandrasegaran, Keshigeyan, et al.
Pubblicazione: (2026)
di: Chandrasegaran, Keshigeyan, et al.
Pubblicazione: (2026)
BAgger: Backwards Aggregation for Mitigating Drift in Autoregressive Video Diffusion Models
di: Po, Ryan, et al.
Pubblicazione: (2025)
di: Po, Ryan, et al.
Pubblicazione: (2025)
Repurposing 2D Diffusion Models with Gaussian Atlas for 3D Generation
di: Xiang, Tiange, et al.
Pubblicazione: (2025)
di: Xiang, Tiange, et al.
Pubblicazione: (2025)
Anatomically Guided Latent Diffusion for Brain MRI Progression Modeling
di: Wan, Cheng, et al.
Pubblicazione: (2026)
di: Wan, Cheng, et al.
Pubblicazione: (2026)
Cycle Diffusion Model for Counterfactual Image Generation
di: Huang, Fangrui, et al.
Pubblicazione: (2025)
di: Huang, Fangrui, et al.
Pubblicazione: (2025)
HumanScore: Benchmarking Human Motions in Generated Videos
di: Fang, Yusu, et al.
Pubblicazione: (2026)
di: Fang, Yusu, et al.
Pubblicazione: (2026)
Latent Drifting in Diffusion Models for Counterfactual Medical Image Synthesis
di: Yeganeh, Yousef, et al.
Pubblicazione: (2024)
di: Yeganeh, Yousef, et al.
Pubblicazione: (2024)
ZeroNVS: Zero-Shot 360-Degree View Synthesis from a Single Image
di: Sargent, Kyle, et al.
Pubblicazione: (2023)
di: Sargent, Kyle, et al.
Pubblicazione: (2023)
Physics-Informed Latent Diffusion for Multimodal Brain MRI Synthesis
di: Lüpke, Sven, et al.
Pubblicazione: (2024)
di: Lüpke, Sven, et al.
Pubblicazione: (2024)
Wild2Avatar: Rendering Humans Behind Occlusions
di: Xiang, Tiange, et al.
Pubblicazione: (2023)
di: Xiang, Tiange, et al.
Pubblicazione: (2023)
VLIC: Vision-Language Models As Perceptual Judges for Human-Aligned Image Compression
di: Sargent, Kyle, et al.
Pubblicazione: (2025)
di: Sargent, Kyle, et al.
Pubblicazione: (2025)
Language-Guided Trajectory Traversal in Disentangled Stable Diffusion Latent Space for Factorized Medical Image Generation
di: TehraniNasab, Zahra, et al.
Pubblicazione: (2025)
di: TehraniNasab, Zahra, et al.
Pubblicazione: (2025)
PixelDiT: Pixel Diffusion Transformers for Image Generation
di: Yu, Yongsheng, et al.
Pubblicazione: (2025)
di: Yu, Yongsheng, et al.
Pubblicazione: (2025)
SocialGen: Modeling Multi-Human Social Interaction with Language Models
di: Yu, Heng, et al.
Pubblicazione: (2025)
di: Yu, Heng, et al.
Pubblicazione: (2025)
Diffusion MRI Transformer with a Diffusion Space Rotary Positional Embedding (D-RoPE)
di: Kung, Gustavo Chau Loo, et al.
Pubblicazione: (2026)
di: Kung, Gustavo Chau Loo, et al.
Pubblicazione: (2026)
MambaOcc: Visual State Space Model for BEV-based Occupancy Prediction with Local Adaptive Reordering
di: Tian, Yonglin, et al.
Pubblicazione: (2024)
di: Tian, Yonglin, et al.
Pubblicazione: (2024)
PixPerfect: Seamless Latent Diffusion Local Editing with Discriminative Pixel-Space Refinement
di: Zheng, Haitian, et al.
Pubblicazione: (2025)
di: Zheng, Haitian, et al.
Pubblicazione: (2025)
Artist-Created Mesh Generation from Raw Observation
di: He, Yao, et al.
Pubblicazione: (2025)
di: He, Yao, et al.
Pubblicazione: (2025)
Pixel-Space Post-Training of Latent Diffusion Models
di: Zhang, Christina, et al.
Pubblicazione: (2024)
di: Zhang, Christina, et al.
Pubblicazione: (2024)
Enforcing Conditional Independence for Fair Representation Learning and Causal Image Generation
di: Hwa, Jensen, et al.
Pubblicazione: (2024)
di: Hwa, Jensen, et al.
Pubblicazione: (2024)
A Tool Bottleneck Framework for Clinically-Informed and Interpretable Medical Image Understanding
di: Liu, Christina, et al.
Pubblicazione: (2025)
di: Liu, Christina, et al.
Pubblicazione: (2025)
Repurposing 2D Diffusion Models for 3D Shape Completion
di: He, Yao, et al.
Pubblicazione: (2025)
di: He, Yao, et al.
Pubblicazione: (2025)
Transforming Weather Data from Pixel to Latent Space
di: Zhao, Sijie, et al.
Pubblicazione: (2025)
di: Zhao, Sijie, et al.
Pubblicazione: (2025)
Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation
di: Zhang, David Junhao, et al.
Pubblicazione: (2023)
di: Zhang, David Junhao, et al.
Pubblicazione: (2023)
One-step Latent-free Image Generation with Pixel Mean Flows
di: Lu, Yiyang, et al.
Pubblicazione: (2026)
di: Lu, Yiyang, et al.
Pubblicazione: (2026)
Integrating Anatomical Priors into a Causal Diffusion Model
di: Li, Binxu, et al.
Pubblicazione: (2025)
di: Li, Binxu, et al.
Pubblicazione: (2025)
ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body
di: Zhang, Juze, et al.
Pubblicazione: (2025)
di: Zhang, Juze, et al.
Pubblicazione: (2025)
Edify Image: High-Quality Image Generation with Pixel Space Laplacian Diffusion Models
di: NVIDIA, et al.
Pubblicazione: (2024)
di: NVIDIA, et al.
Pubblicazione: (2024)
Towards Fine-Grained Video Question Answering
di: Dai, Wei, et al.
Pubblicazione: (2025)
di: Dai, Wei, et al.
Pubblicazione: (2025)
Latent Space Super-Resolution for Higher-Resolution Image Generation with Diffusion Models
di: Jeong, Jinho, et al.
Pubblicazione: (2025)
di: Jeong, Jinho, et al.
Pubblicazione: (2025)
PixelFlow: Pixel-Space Generative Models with Flow
di: Chen, Shoufa, et al.
Pubblicazione: (2025)
di: Chen, Shoufa, et al.
Pubblicazione: (2025)
Probability Density Geodesics in Image Diffusion Latent Space
di: Yu, Qingtao, et al.
Pubblicazione: (2025)
di: Yu, Qingtao, et al.
Pubblicazione: (2025)
Registers Matter for Pixel-Space Diffusion Transformers
di: Starodubcev, Nikita, et al.
Pubblicazione: (2026)
di: Starodubcev, Nikita, et al.
Pubblicazione: (2026)
Few-Shot Classification of Interactive Activities of Daily Living (InteractADL)
di: Durante, Zane, et al.
Pubblicazione: (2024)
di: Durante, Zane, et al.
Pubblicazione: (2024)
PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion
di: Lu, Yifan, et al.
Pubblicazione: (2026)
di: Lu, Yifan, et al.
Pubblicazione: (2026)
Towards Robust 3D Pose Transfer with Adversarial Learning
di: Chen, Haoyu, et al.
Pubblicazione: (2024)
di: Chen, Haoyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Flow to the Mode: Mode-Seeking Diffusion Autoencoders for State-of-the-Art Image Tokenization
di: Sargent, Kyle, et al.
Pubblicazione: (2025) -
OccFusion: Rendering Occluded Humans with Generative Diffusion Priors
di: Sun, Adam, et al.
Pubblicazione: (2024) -
Discovering Latent Graphs with GFlowNets for Diverse Conditional Image Generation
di: Trang, Bailey, et al.
Pubblicazione: (2025) -
The Language of Motion: Unifying Verbal and Non-verbal Language of 3D Human Motion
di: Chen, Changan, et al.
Pubblicazione: (2024) -
GPIC: A Giant Permissive Image Corpus for Visual Generation
di: Chandrasegaran, Keshigeyan, et al.
Pubblicazione: (2026)