Auto-Regressively Generating Multi-View Consistent Images
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, JiaKui, Yang, Yuxiao, Liu, Jialun, Wu, Jinbo, Zhao, Chen, Lu, Yanye |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bridging Degradation Discrimination and Generation for Universal Image Restoration
by: Hu, JiaKui, et al.
Published: (2026)
by: Hu, JiaKui, et al.
Published: (2026)
Universal Image Restoration Pre-training via Degradation Classification
by: Hu, JiaKui, et al.
Published: (2025)
by: Hu, JiaKui, et al.
Published: (2025)
Omni-View: Unlocking How Generation Facilitates Understanding in Unified 3D Model based on Multiview images
by: Hu, JiaKui, et al.
Published: (2025)
by: Hu, JiaKui, et al.
Published: (2025)
Universal Image Restoration Pre-training via Masked Degradation Classification
by: Hu, JiaKui, et al.
Published: (2025)
by: Hu, JiaKui, et al.
Published: (2025)
Enhancing Image Restoration Transformer via Adaptive Translation Equivariance
by: Hu, JiaKui, et al.
Published: (2025)
by: Hu, JiaKui, et al.
Published: (2025)
Geometry-as-context: Modulating Explicit 3D in Scene-consistent Video Generation to Geometry Context
by: Hu, JiaKui, et al.
Published: (2026)
by: Hu, JiaKui, et al.
Published: (2026)
GarmentPainter: Efficient 3D Garment Texture Synthesis with Character-Guided Diffusion Model
by: Wu, Jinbo, et al.
Published: (2026)
by: Wu, Jinbo, et al.
Published: (2026)
HYDRA: Unifying Multi-modal Generation and Understanding via Representation-Harmonized Tokenization
by: Qiu, Xuerui, et al.
Published: (2026)
by: Qiu, Xuerui, et al.
Published: (2026)
A Watermark for Auto-Regressive Image Generation Models
by: Wu, Yihan, et al.
Published: (2025)
by: Wu, Yihan, et al.
Published: (2025)
AutoStudio: Crafting Consistent Subjects in Multi-turn Interactive Image Generation
by: Cheng, Junhao, et al.
Published: (2024)
by: Cheng, Junhao, et al.
Published: (2024)
Personalized Text-to-Image Generation with Auto-Regressive Models
by: Sun, Kaiyue, et al.
Published: (2025)
by: Sun, Kaiyue, et al.
Published: (2025)
TexGaussian: Generating High-quality PBR Material via Octree-based 3D Gaussian Splatting
by: Xiong, Bojun, et al.
Published: (2024)
by: Xiong, Bojun, et al.
Published: (2024)
PlanGen: Towards Unified Layout Planning and Image Generation in Auto-Regressive Vision Language Models
by: He, Runze, et al.
Published: (2025)
by: He, Runze, et al.
Published: (2025)
Toward a Multi-View Brain Network Foundation Model: Cross-View Consistency Learning Across Arbitrary Atlases
by: Xu, Jiaxing, et al.
Published: (2026)
by: Xu, Jiaxing, et al.
Published: (2026)
Narrative Weaver: Towards Controllable Long-Range Visual Consistency with Multi-Modal Conditioning
by: Yao, Zhengjian, et al.
Published: (2026)
by: Yao, Zhengjian, et al.
Published: (2026)
Dragen3D: Multiview Geometry Consistent 3D Gaussian Generation with Drag-Based Control
by: Yan, Jinbo, et al.
Published: (2025)
by: Yan, Jinbo, et al.
Published: (2025)
FlexPainter: Flexible and Multi-View Consistent Texture Generation
by: Yan, Dongyu, et al.
Published: (2025)
by: Yan, Dongyu, et al.
Published: (2025)
TexRO: Generating Delicate Textures of 3D Models by Recursive Optimization
by: Wu, Jinbo, et al.
Published: (2024)
by: Wu, Jinbo, et al.
Published: (2024)
VAR-CLIP: Text-to-Image Generator with Visual Auto-Regressive Modeling
by: Zhang, Qian, et al.
Published: (2024)
by: Zhang, Qian, et al.
Published: (2024)
Phase-Consistent Magnetic Spectral Learning for Multi-View Clustering
by: Lu, Mingdong, et al.
Published: (2026)
by: Lu, Mingdong, et al.
Published: (2026)
Resurrect Mask AutoRegressive Modeling for Efficient and Scalable Image Generation
by: Xin, Yi, et al.
Published: (2025)
by: Xin, Yi, et al.
Published: (2025)
ViewMask-1-to-3: Multi-View Consistent Image Generation via Multimodal Diffusion Models
by: Zhu, Ruishu, et al.
Published: (2025)
by: Zhu, Ruishu, et al.
Published: (2025)
MCGS: Multiview Consistency Enhancement for Sparse-View 3D Gaussian Radiance Fields
by: Xiao, Yuru, et al.
Published: (2024)
by: Xiao, Yuru, et al.
Published: (2024)
GVA: Reconstructing Vivid 3D Gaussian Avatars from Monocular Videos
by: Liu, Xinqi, et al.
Published: (2024)
by: Liu, Xinqi, et al.
Published: (2024)
ConsID-Gen: View-Consistent and Identity-Preserving Image-to-Video Generation
by: Wu, Mingyang, et al.
Published: (2026)
by: Wu, Mingyang, et al.
Published: (2026)
ConsistentDreamer: View-Consistent Meshes Through Balanced Multi-View Gaussian Optimization
by: Şahin, Onat, et al.
Published: (2025)
by: Şahin, Onat, et al.
Published: (2025)
BEV-VAE: Multi-view Image Generation with Spatial Consistency for Autonomous Driving
by: Chen, Zeming, et al.
Published: (2025)
by: Chen, Zeming, et al.
Published: (2025)
AutoScape: Geometry-Consistent Long-Horizon Scene Generation
by: Chen, Jiacheng, et al.
Published: (2025)
by: Chen, Jiacheng, et al.
Published: (2025)
Multi-Level Embedding and Alignment Network with Consistency and Invariance Learning for Cross-View Geo-Localization
by: Chen, Zhongwei, et al.
Published: (2024)
by: Chen, Zhongwei, et al.
Published: (2024)
Beyond Text: Frozen Large Language Models in Visual Signal Comprehension
by: Zhu, Lei, et al.
Published: (2024)
by: Zhu, Lei, et al.
Published: (2024)
Multi-scale Image Super Resolution with a Single Auto-Regressive Model
by: Sanchez, Enrique, et al.
Published: (2025)
by: Sanchez, Enrique, et al.
Published: (2025)
SceneLCM: End-to-End Layout-Guided Interactive Indoor Scene Generation with Latent Consistency Model
by: Lin, Yangkai, et al.
Published: (2025)
by: Lin, Yangkai, et al.
Published: (2025)
Physics-Informed Conditional Diffusion for Motion-Robust Retinal Temporal Laser Speckle Contrast Imaging
by: Chen, Qian, et al.
Published: (2026)
by: Chen, Qian, et al.
Published: (2026)
Sculpt3D: Multi-View Consistent Text-to-3D Generation with Sparse 3D Prior
by: Chen, Cheng, et al.
Published: (2024)
by: Chen, Cheng, et al.
Published: (2024)
CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion
by: Zheng, Wendi, et al.
Published: (2024)
by: Zheng, Wendi, et al.
Published: (2024)
FARMER: Flow AutoRegressive Transformer over Pixels
by: Zheng, Guangting, et al.
Published: (2025)
by: Zheng, Guangting, et al.
Published: (2025)
Training-free Test-time Improvement for Explainable Medical Image Classification
by: He, Hangzhou, et al.
Published: (2025)
by: He, Hangzhou, et al.
Published: (2025)
Scaling the Codebook Size of VQGAN to 100,000 with a Utilization Rate of 99%
by: Zhu, Lei, et al.
Published: (2024)
by: Zhu, Lei, et al.
Published: (2024)
CART: Compositional Auto-Regressive Transformer for Image Generation
by: Roheda, Siddharth, et al.
Published: (2024)
by: Roheda, Siddharth, et al.
Published: (2024)
InfinityStar: Unified Spacetime AutoRegressive Modeling for Visual Generation
by: Liu, Jinlai, et al.
Published: (2025)
by: Liu, Jinlai, et al.
Published: (2025)
Similar Items
-
Bridging Degradation Discrimination and Generation for Universal Image Restoration
by: Hu, JiaKui, et al.
Published: (2026) -
Universal Image Restoration Pre-training via Degradation Classification
by: Hu, JiaKui, et al.
Published: (2025) -
Omni-View: Unlocking How Generation Facilitates Understanding in Unified 3D Model based on Multiview images
by: Hu, JiaKui, et al.
Published: (2025) -
Universal Image Restoration Pre-training via Masked Degradation Classification
by: Hu, JiaKui, et al.
Published: (2025) -
Enhancing Image Restoration Transformer via Adaptive Translation Equivariance
by: Hu, JiaKui, et al.
Published: (2025)