X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again
Fuente:
arXiv
Saved in:
| Main Authors: | Geng, Zigang, Wang, Yibing, Ma, Yeyao, Li, Chen, Rao, Yongming, Gu, Shuyang, Zhong, Zhao, Lu, Qinglin, Hu, Han, Zhang, Xiaosong, Linus, Wang, Di, Jiang, Jie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tokenize Image as a Set
by: Geng, Zigang, et al.
Published: (2025)
by: Geng, Zigang, et al.
Published: (2025)
Equivariant Image Modeling
by: Dong, Ruixiao, et al.
Published: (2025)
by: Dong, Ruixiao, et al.
Published: (2025)
FAIL: Flow Matching Adversarial Imitation Learning for Image Generation
by: Ma, Yeyao, et al.
Published: (2026)
by: Ma, Yeyao, et al.
Published: (2026)
SAM-SP: Self-Prompting Makes SAM Great Again
by: Zhou, Chunpeng, et al.
Published: (2024)
by: Zhou, Chunpeng, et al.
Published: (2024)
BlazeBVD: Make Scale-Time Equalization Great Again for Blind Video Deflickering
by: Qiu, Xinmin, et al.
Published: (2024)
by: Qiu, Xinmin, et al.
Published: (2024)
FCN+: Global Receptive Convolution Makes FCN Great Again
by: Ren, Xiaoyu, et al.
Published: (2023)
by: Ren, Xiaoyu, et al.
Published: (2023)
Make Literature-Based Discovery Great Again through Reproducible Pipelines
by: Cestnik, Bojan, et al.
Published: (2025)
by: Cestnik, Bojan, et al.
Published: (2025)
Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models
by: Ma, Tianren, et al.
Published: (2025)
by: Ma, Tianren, et al.
Published: (2025)
Make America Great Again: ¿expresión de un nativismo blanco contemporáneo?
by: Antonio Alejo
Published: (2018)
by: Antonio Alejo
Published: (2018)
Ola: Pushing the Frontiers of Omni-Modal Language Model
by: Liu, Zuyan, et al.
Published: (2025)
by: Liu, Zuyan, et al.
Published: (2025)
Autoregressive Omni-Aware Outpainting for Open-Vocabulary 360-Degree Image Generation
by: Lu, Zhuqiang, et al.
Published: (2023)
by: Lu, Zhuqiang, et al.
Published: (2023)
Make LoRA Great Again: Boosting LoRA with Adaptive Singular Values and Mixture-of-Experts Optimization Alignment
by: Fan, Chenghao, et al.
Published: (2025)
by: Fan, Chenghao, et al.
Published: (2025)
#MakeBeefGreatAgain: A Cross-Platform Analysis of Early #MAHA Discourse
by: Xue, Haoning, et al.
Published: (2026)
by: Xue, Haoning, et al.
Published: (2026)
DRPCA-Net: Make Robust PCA Great Again for Infrared Small Target Detection
by: Xiong, Zihao, et al.
Published: (2025)
by: Xiong, Zihao, et al.
Published: (2025)
Make Graph-based Referring Expression Comprehension Great Again through Expression-guided Dynamic Gating and Regression
by: Ke, Jingcheng, et al.
Published: (2024)
by: Ke, Jingcheng, et al.
Published: (2024)
Make Graph Neural Networks Great Again: A Generic Integration Paradigm of Topology-Free Patterns for Traffic Speed Prediction
by: Zhou, Yicheng, et al.
Published: (2024)
by: Zhou, Yicheng, et al.
Published: (2024)
ASPIRE: Make Spectral Graph Collaborative Filtering Great Again via Adaptive Filter Learning
by: He, Yunhang, et al.
Published: (2026)
by: He, Yunhang, et al.
Published: (2026)
Distillation-Driven Diffusion Model for Multi-Scale MRI Super-Resolution: Make 1.5T MRI Great Again
by: Wang, Zhe, et al.
Published: (2025)
by: Wang, Zhe, et al.
Published: (2025)
Making Images Real Again: A Comprehensive Survey on Deep Image Composition
by: Niu, Li, et al.
Published: (2021)
by: Niu, Li, et al.
Published: (2021)
D2C: Unlocking the Potential of Continuous Autoregressive Image Generation with Discrete Tokens
by: Wang, Panpan, et al.
Published: (2025)
by: Wang, Panpan, et al.
Published: (2025)
Gluon: Making Muon & Scion Great Again! (Bridging Theory and Practice of LMO-based Optimizers for LLMs)
by: Riabinin, Artem, et al.
Published: (2025)
by: Riabinin, Artem, et al.
Published: (2025)
OmniPro: A Comprehensive Benchmark for Omni-Proactive Streaming Video Understanding
by: Zhao, Ruixiang, et al.
Published: (2026)
by: Zhao, Ruixiang, et al.
Published: (2026)
Make Planning Research Rigorous Again!
by: Katz, Michael, et al.
Published: (2025)
by: Katz, Michael, et al.
Published: (2025)
Make Deep Networks Shallow Again
by: Bermeitinger, Bernhard, et al.
Published: (2023)
by: Bermeitinger, Bernhard, et al.
Published: (2023)
CCA: Collaborative Competitive Agents for Image Editing
by: Hang, Tiankai, et al.
Published: (2024)
by: Hang, Tiankai, et al.
Published: (2024)
When One-Parameter Dark Energy Makes Neutrinos Physical Again
by: Yang, Weiqiang, et al.
Published: (2026)
by: Yang, Weiqiang, et al.
Published: (2026)
Drug classification based on X-ray spectroscopy combined with machine learning
by: Li, Yongming, et al.
Published: (2025)
by: Li, Yongming, et al.
Published: (2025)
Making Sigmoid-MSE Great Again: Output Reset Challenges Softmax Cross-Entropy in Neural Network Classification
by: Tyagi, Kanishka, et al.
Published: (2024)
by: Tyagi, Kanishka, et al.
Published: (2024)
Make Shuffling Great Again: A Side-Channel Resistant Fisher-Yates Algorithm for Protecting Neural Networks
by: Puškáč, Leonard, et al.
Published: (2025)
by: Puškáč, Leonard, et al.
Published: (2025)
Optimal Power Flow in Renewable-Integrated Power Systems: A Comprehensive Review
by: Chen, Zigang
Published: (2024)
by: Chen, Zigang
Published: (2024)
Several questions of visual generation in 2024
by: Gu, Shuyang
Published: (2024)
by: Gu, Shuyang
Published: (2024)
Fine-grained Spatiotemporal Grounding on Egocentric Videos
by: Liang, Shuo, et al.
Published: (2025)
by: Liang, Shuo, et al.
Published: (2025)
Discrete Noise Inversion for Next-scale Autoregressive Text-based Image Editing
by: Dao, Quan, et al.
Published: (2025)
by: Dao, Quan, et al.
Published: (2025)
AR-Omni: A Unified Autoregressive Model for Any-to-Any Generation
by: Cheng, Dongjie, et al.
Published: (2026)
by: Cheng, Dongjie, et al.
Published: (2026)
Towards Omni-RAG: Comprehensive Retrieval-Augmented Generation for Large Language Models in Medical Applications
by: Chen, Zhe, et al.
Published: (2025)
by: Chen, Zhe, et al.
Published: (2025)
Rethinking Discrete Tokens: Treating Them as Conditions for Continuous Autoregressive Image Synthesis
by: Zheng, Peng, et al.
Published: (2025)
by: Zheng, Peng, et al.
Published: (2025)
Improved Noise Schedule for Diffusion Training
by: Hang, Tiankai, et al.
Published: (2024)
by: Hang, Tiankai, et al.
Published: (2024)
Understanding vs. Generation: Navigating Optimization Dilemma in Multimodal Models
by: Ye, Sen, et al.
Published: (2026)
by: Ye, Sen, et al.
Published: (2026)
Development of Focused X-ray Luminescence Compute Tomography Imaging
by: Fang, Yile, et al.
Published: (2024)
by: Fang, Yile, et al.
Published: (2024)
OmniVideo-R1: Reinforcing Audio-visual Reasoning with Query Intention and Modality Attention
by: Chen, Zhangquan, et al.
Published: (2026)
by: Chen, Zhangquan, et al.
Published: (2026)
Similar Items
-
Tokenize Image as a Set
by: Geng, Zigang, et al.
Published: (2025) -
Equivariant Image Modeling
by: Dong, Ruixiao, et al.
Published: (2025) -
FAIL: Flow Matching Adversarial Imitation Learning for Image Generation
by: Ma, Yeyao, et al.
Published: (2026) -
SAM-SP: Self-Prompting Makes SAM Great Again
by: Zhou, Chunpeng, et al.
Published: (2024) -
BlazeBVD: Make Scale-Time Equalization Great Again for Blind Video Deflickering
by: Qiu, Xinmin, et al.
Published: (2024)