Lumina-mGPT 2.0: Stand-Alone AutoRegressive Image Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xin, Yi, Yan, Juncheng, Qin, Qi, Li, Zhen, Liu, Dongyang, Li, Shicheng, Huang, Victor Shea-Jay, Zhou, Yupeng, Zhang, Renrui, Zhuo, Le, Han, Tiancheng, Sun, Xiaoqing, Luo, Siqi, Wang, Mengmeng, Fu, Bin, Cao, Yuewen, Li, Hongsheng, Zhai, Guangtao, Liu, Xiaohong, Qiao, Yu, Gao, Peng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining
von: Liu, Dongyang, et al.
Veröffentlicht: (2024)
von: Liu, Dongyang, et al.
Veröffentlicht: (2024)
Resurrect Mask AutoRegressive Modeling for Efficient and Scalable Image Generation
von: Xin, Yi, et al.
Veröffentlicht: (2025)
von: Xin, Yi, et al.
Veröffentlicht: (2025)
Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT
von: Liu, Dongyang, et al.
Veröffentlicht: (2025)
von: Liu, Dongyang, et al.
Veröffentlicht: (2025)
Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT
von: Zhuo, Le, et al.
Veröffentlicht: (2024)
von: Zhuo, Le, et al.
Veröffentlicht: (2024)
Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding
von: Xin, Yi, et al.
Veröffentlicht: (2025)
von: Xin, Yi, et al.
Veröffentlicht: (2025)
Lumina-Image 2.0: A Unified and Efficient Image Generative Framework
von: Qin, Qi, et al.
Veröffentlicht: (2025)
von: Qin, Qi, et al.
Veröffentlicht: (2025)
Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers
von: Gao, Peng, et al.
Veröffentlicht: (2024)
von: Gao, Peng, et al.
Veröffentlicht: (2024)
NoiseAR: AutoRegressing Initial Noise Prior for Diffusion Models
von: Li, Zeming, et al.
Veröffentlicht: (2025)
von: Li, Zeming, et al.
Veröffentlicht: (2025)
Nested AutoRegressive Models
von: Wu, Hongyu, et al.
Veröffentlicht: (2025)
von: Wu, Hongyu, et al.
Veröffentlicht: (2025)
PTQ4ARVG: Post-Training Quantization for AutoRegressive Visual Generation Models
von: Liu, Xuewen, et al.
Veröffentlicht: (2026)
von: Liu, Xuewen, et al.
Veröffentlicht: (2026)
Lumina
Veröffentlicht: (2017)
Veröffentlicht: (2017)
A$^2$-Edit: Precise Reference-Guided Image Editing of Arbitrary Objects and Ambiguous Masks
von: Zheng, Huayu, et al.
Veröffentlicht: (2026)
von: Zheng, Huayu, et al.
Veröffentlicht: (2026)
MARS: Mesh AutoRegressive Model for 3D Shape Detailization
von: Gao, Jingnan, et al.
Veröffentlicht: (2025)
von: Gao, Jingnan, et al.
Veröffentlicht: (2025)
AvatarPointillist: AutoRegressive 4D Gaussian Avatarization
von: Liu, Hongyu, et al.
Veröffentlicht: (2026)
von: Liu, Hongyu, et al.
Veröffentlicht: (2026)
Sketch Then Paint: Hierarchical Reinforcement Learning for Diffusion Multi-Modal Large Language Models
von: Luo, Siqi, et al.
Veröffentlicht: (2026)
von: Luo, Siqi, et al.
Veröffentlicht: (2026)
A Behaviour-Aware Federated Forecasting Framework for Distributed Stand-Alone Wind Turbines
von: Li, Bowen, et al.
Veröffentlicht: (2026)
von: Li, Bowen, et al.
Veröffentlicht: (2026)
AutoRegressive Generation with B-rep Holistic Token Sequence Representation
von: Li, Jiahao, et al.
Veröffentlicht: (2026)
von: Li, Jiahao, et al.
Veröffentlicht: (2026)
MVAR: MultiVariate AutoRegressive Air Pollutants Forecasting Model
von: Fan, Xu, et al.
Veröffentlicht: (2025)
von: Fan, Xu, et al.
Veröffentlicht: (2025)
Resolution-Agnostic Neural Compression for High-Fidelity Portrait Video Conferencing via Implicit Radiance Fields
von: Li, Yifei, et al.
Veröffentlicht: (2024)
von: Li, Yifei, et al.
Veröffentlicht: (2024)
Enhancing Test Time Adaptation with Few-shot Guidance
von: Luo, Siqi, et al.
Veröffentlicht: (2024)
von: Luo, Siqi, et al.
Veröffentlicht: (2024)
Inference-Time Scaling for Visual AutoRegressive modeling by Searching Representative Samples
von: Tang, Weidong, et al.
Veröffentlicht: (2026)
von: Tang, Weidong, et al.
Veröffentlicht: (2026)
Implementing Adaptations for Vision AutoRegressive Model
von: Shaikh, Kaif, et al.
Veröffentlicht: (2025)
von: Shaikh, Kaif, et al.
Veröffentlicht: (2025)
Privacy Attacks on Image AutoRegressive Models
von: Kowalczuk, Antoni, et al.
Veröffentlicht: (2025)
von: Kowalczuk, Antoni, et al.
Veröffentlicht: (2025)
STAR: Scale-wise Text-conditioned AutoRegressive image generation
von: Ma, Xiaoxiao, et al.
Veröffentlicht: (2024)
von: Ma, Xiaoxiao, et al.
Veröffentlicht: (2024)
Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
von: Han, Jian, et al.
Veröffentlicht: (2024)
von: Han, Jian, et al.
Veröffentlicht: (2024)
FUMO: Prior-Modulated Diffusion for Single Image Reflection Removal
von: Xu, Telang, et al.
Veröffentlicht: (2026)
von: Xu, Telang, et al.
Veröffentlicht: (2026)
LayerT2V: A Unified Multi-Layer Video Generation Framework
von: Li, Guangzhao, et al.
Veröffentlicht: (2025)
von: Li, Guangzhao, et al.
Veröffentlicht: (2025)
InfinityStar: Unified Spacetime AutoRegressive Modeling for Visual Generation
von: Liu, Jinlai, et al.
Veröffentlicht: (2025)
von: Liu, Jinlai, et al.
Veröffentlicht: (2025)
FARMER: Flow AutoRegressive Transformer over Pixels
von: Zheng, Guangting, et al.
Veröffentlicht: (2025)
von: Zheng, Guangting, et al.
Veröffentlicht: (2025)
HiPART: Hierarchical Pose AutoRegressive Transformer for Occluded 3D Human Pose Estimation
von: Zheng, Hongwei, et al.
Veröffentlicht: (2025)
von: Zheng, Hongwei, et al.
Veröffentlicht: (2025)
SparVAR: Exploring Sparsity in Visual AutoRegressive Modeling for Training-Free Acceleration
von: Li, Zekun, et al.
Veröffentlicht: (2026)
von: Li, Zekun, et al.
Veröffentlicht: (2026)
ARMFlow: AutoRegressive MeanFlow for Online 3D Human Reaction Generation
von: Geng, Zichen, et al.
Veröffentlicht: (2025)
von: Geng, Zichen, et al.
Veröffentlicht: (2025)
Stands Alone, Faces, and Other Poems
von: Russell LeBeau, Patrick
Veröffentlicht: (2025)
von: Russell LeBeau, Patrick
Veröffentlicht: (2025)
Integrated vs. Stand-Alone Systems.
von: Kline, Norman
Veröffentlicht: (1986)
von: Kline, Norman
Veröffentlicht: (1986)
TR-PTS: Task-Relevant Parameter and Token Selection for Efficient Tuning
von: Luo, Siqi, et al.
Veröffentlicht: (2025)
von: Luo, Siqi, et al.
Veröffentlicht: (2025)
Charting the Path Forward: CT Image Quality Assessment -- An In-Depth Review
von: Xun, Siyi, et al.
Veröffentlicht: (2024)
von: Xun, Siyi, et al.
Veröffentlicht: (2024)
Quality Assessment in the Era of Large Models: A Survey
von: Zhang, Zicheng, et al.
Veröffentlicht: (2024)
von: Zhang, Zicheng, et al.
Veröffentlicht: (2024)
FAR-Drive: Frame-AutoRegressive Video Generation in Closed-Loop Autonomous Driving
von: Li, Yaoru, et al.
Veröffentlicht: (2026)
von: Li, Yaoru, et al.
Veröffentlicht: (2026)
STAR: STacked AutoRegressive Scheme for Unified Multimodal Learning
von: Qin, Jie, et al.
Veröffentlicht: (2025)
von: Qin, Jie, et al.
Veröffentlicht: (2025)
Vector AutoRegressive Moving Average Models: A Review
von: Düker, Marie-Christine, et al.
Veröffentlicht: (2024)
von: Düker, Marie-Christine, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining
von: Liu, Dongyang, et al.
Veröffentlicht: (2024) -
Resurrect Mask AutoRegressive Modeling for Efficient and Scalable Image Generation
von: Xin, Yi, et al.
Veröffentlicht: (2025) -
Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT
von: Liu, Dongyang, et al.
Veröffentlicht: (2025) -
Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT
von: Zhuo, Le, et al.
Veröffentlicht: (2024) -
Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding
von: Xin, Yi, et al.
Veröffentlicht: (2025)