DetailFlow: 1D Coarse-to-Fine Autoregressive Image Generation via Next-Detail Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Yiheng, Qu, Liao, Zhang, Huichao, Wang, Xu, Jiang, Yi, Gao, Yiming, Ye, Hu, Li, Xian, Wang, Shuai, Du, Daniel K., Chen, Fangmin, Yuan, Zehuan, Wu, Xinglong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation
by: Qu, Liao, et al.
Published: (2024)
by: Qu, Liao, et al.
Published: (2024)
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction
by: Tian, Keyu, et al.
Published: (2024)
by: Tian, Keyu, et al.
Published: (2024)
VAR RL Done Right: Tackling Asynchronous Policy Conflicts in Visual Autoregressive Generation
by: Sun, Shikun, et al.
Published: (2026)
by: Sun, Shikun, et al.
Published: (2026)
ARMesh: Autoregressive Mesh Generation via Next-Level-of-Detail Prediction
by: Lei, Jiabao, et al.
Published: (2025)
by: Lei, Jiabao, et al.
Published: (2025)
NextFlow: Unified Sequential Modeling Activates Multimodal Understanding and Generation
by: Zhang, Huichao, et al.
Published: (2026)
by: Zhang, Huichao, et al.
Published: (2026)
PointNSP: Autoregressive 3D Point Cloud Generation with Next-Scale Level-of-Detail Prediction
by: Meng, Ziqiao, et al.
Published: (2025)
by: Meng, Ziqiao, et al.
Published: (2025)
PointNSP: Autoregressive 3D Point Cloud Generation with Next-Scale Level-of-Detail Prediction
by: Meng, Ziqiao, et al.
Published: (2025)
by: Meng, Ziqiao, et al.
Published: (2025)
Missing Fine Details in Images: Last Seen in High Frequencies
by: Medi, Tejaswini, et al.
Published: (2025)
by: Medi, Tejaswini, et al.
Published: (2025)
FlashVideo: Flowing Fidelity to Detail for Efficient High-Resolution Video Generation
by: Zhang, Shilong, et al.
Published: (2025)
by: Zhang, Shilong, et al.
Published: (2025)
DetailCLIP: Injecting Image Details into CLIP's Feature Space
by: Zhang, Zilun, et al.
Published: (2022)
by: Zhang, Zilun, et al.
Published: (2022)
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks
by: Monsefi, Amin Karimi, et al.
Published: (2024)
by: Monsefi, Amin Karimi, et al.
Published: (2024)
IAR2: Improving Autoregressive Visual Generation with Semantic-Detail Associated Token Prediction
by: Yi, Ran, et al.
Published: (2025)
by: Yi, Ran, et al.
Published: (2025)
Benchmarking and Improving Detail Image Caption
by: Dong, Hongyuan, et al.
Published: (2024)
by: Dong, Hongyuan, et al.
Published: (2024)
Detail++: Training-Free Detail Enhancer for Text-to-Image Diffusion Models
by: Chen, Lifeng, et al.
Published: (2025)
by: Chen, Lifeng, et al.
Published: (2025)
Improving Autoregressive Image Generation through Coarse-to-Fine Token Prediction
by: Guo, Ziyao, et al.
Published: (2025)
by: Guo, Ziyao, et al.
Published: (2025)
DetailFusion: A Dual-branch Framework with Detail Enhancement for Composed Image Retrieval
by: Yang, Yuxin, et al.
Published: (2025)
by: Yang, Yuxin, et al.
Published: (2025)
Devil is in the Details: Density Guidance for Detail-Aware Generation with Flow Models
by: Karczewski, Rafał, et al.
Published: (2025)
by: Karczewski, Rafał, et al.
Published: (2025)
Devil is in the Detail: Towards Injecting Fine Details of Image Prompt in Image Generation via Conflict-free Guidance and Stratified Attention
by: Jo, Kyungmin, et al.
Published: (2025)
by: Jo, Kyungmin, et al.
Published: (2025)
DetailSemNet: Elevating Signature Verification through Detail-Semantic Integration
by: Shih, Meng-Cheng, et al.
Published: (2025)
by: Shih, Meng-Cheng, et al.
Published: (2025)
CARP: Visuomotor Policy Learning via Coarse-to-Fine Autoregressive Prediction
by: Gong, Zhefei, et al.
Published: (2024)
by: Gong, Zhefei, et al.
Published: (2024)
Detail-Enhancing Framework for Reference-Based Image Super-Resolution
by: Wang, Zihan, et al.
Published: (2024)
by: Wang, Zihan, et al.
Published: (2024)
Generating Fine Details of Entity Interactions
by: Gu, Xinyi, et al.
Published: (2025)
by: Gu, Xinyi, et al.
Published: (2025)
FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow Matching
by: Wang, Hui, et al.
Published: (2025)
by: Wang, Hui, et al.
Published: (2025)
Preserving Full Degradation Details for Blind Image Super-Resolution
by: Liu, Hongda, et al.
Published: (2024)
by: Liu, Hongda, et al.
Published: (2024)
DEIG: Detail-Enhanced Instance Generation with Fine-Grained Semantic Control
by: Du, Shiyan, et al.
Published: (2026)
by: Du, Shiyan, et al.
Published: (2026)
Enhancing CNNs With Detail Feature Module for High‐Pixel Image Classification
by: Shanshan Qin, et al.
Published: (2025)
by: Shanshan Qin, et al.
Published: (2025)
Emergent Detailed Balance in Human Mobility under Temporal Coarse-Graining
by: Dong, Lei
Published: (2026)
by: Dong, Lei
Published: (2026)
Block and Detail: Scaffolding Sketch-to-Image Generation
by: Sarukkai, Vishnu, et al.
Published: (2024)
by: Sarukkai, Vishnu, et al.
Published: (2024)
Derived Fields Preserve Fine-Scale Detail in Budgeted Neural Simulators
by: Wang, Wenshuo, et al.
Published: (2026)
by: Wang, Wenshuo, et al.
Published: (2026)
Finding Needles in Images: Can Multimodal LLMs Locate Fine Details?
by: Thakkar, Parth, et al.
Published: (2025)
by: Thakkar, Parth, et al.
Published: (2025)
CorrDetail: Visual Detail Enhanced Self-Correction for Face Forgery Detection
by: Zhou, Binjia, et al.
Published: (2025)
by: Zhou, Binjia, et al.
Published: (2025)
AbsGS: Recovering Fine Details for 3D Gaussian Splatting
by: Ye, Zongxin, et al.
Published: (2024)
by: Ye, Zongxin, et al.
Published: (2024)
Success is in the Details: Evaluate and Enhance Details Sensitivity of Code LLMs through Counterfactuals
by: Luo, Xianzhen, et al.
Published: (2025)
by: Luo, Xianzhen, et al.
Published: (2025)
Indoor Scene Reconstruction with Fine-Grained Details Using Hybrid Representation and Normal Prior Enhancement
by: Ye, Sheng, et al.
Published: (2023)
by: Ye, Sheng, et al.
Published: (2023)
RDTE-UNet: A Boundary and Detail Aware UNet for Precise Medical Image Segmentation
by: Qu, Jierui, et al.
Published: (2025)
by: Qu, Jierui, et al.
Published: (2025)
Animator-Centric Skeleton Generation on Objects with Fine-Grained Details
by: Sun, Mingze, et al.
Published: (2026)
by: Sun, Mingze, et al.
Published: (2026)
Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation
by: Sun, Peize, et al.
Published: (2024)
by: Sun, Peize, et al.
Published: (2024)
No Detail Left Behind: Revisiting Self-Retrieval for Fine-Grained Image Captioning
by: Gaur, Manu, et al.
Published: (2024)
by: Gaur, Manu, et al.
Published: (2024)
Beyond Illumination: Fine-Grained Detail Preservation in Extreme Dark Image Restoration
by: Zhang, Tongshun, et al.
Published: (2025)
by: Zhang, Tongshun, et al.
Published: (2025)
Combined Image Filter for Robust Fine Detail Preservation and Efficient Noise Supression
by: Volodymyr I. Ponomaryov
Published: (2001)
by: Volodymyr I. Ponomaryov
Published: (2001)
Similar Items
-
TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation
by: Qu, Liao, et al.
Published: (2024) -
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction
by: Tian, Keyu, et al.
Published: (2024) -
VAR RL Done Right: Tackling Asynchronous Policy Conflicts in Visual Autoregressive Generation
by: Sun, Shikun, et al.
Published: (2026) -
ARMesh: Autoregressive Mesh Generation via Next-Level-of-Detail Prediction
by: Lei, Jiabao, et al.
Published: (2025) -
NextFlow: Unified Sequential Modeling Activates Multimodal Understanding and Generation
by: Zhang, Huichao, et al.
Published: (2026)