LLaDA-o: An Effective and Length-Adaptive Omni Diffusion Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | You, Zebin, Zhang, Xiaolu, Zhou, Jun, Li, Chongxuan, Wen, Ji-Rong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning
von: You, Zebin, et al.
Veröffentlicht: (2025)
von: You, Zebin, et al.
Veröffentlicht: (2025)
Effective and Efficient Masked Image Generation Models
von: You, Zebin, et al.
Veröffentlicht: (2025)
von: You, Zebin, et al.
Veröffentlicht: (2025)
LLaDA 1.5: Variance-Reduced Preference Optimization for Large Language Diffusion Models
von: Zhu, Fengqi, et al.
Veröffentlicht: (2025)
von: Zhu, Fengqi, et al.
Veröffentlicht: (2025)
LLaDA-VLA: Vision Language Diffusion Action Models
von: Wen, Yuqing, et al.
Veröffentlicht: (2025)
von: Wen, Yuqing, et al.
Veröffentlicht: (2025)
Are Images Indistinguishable to Humans Also Indistinguishable to Classifiers?
von: You, Zebin, et al.
Veröffentlicht: (2024)
von: You, Zebin, et al.
Veröffentlicht: (2024)
Efficient Token Pruning for LLaDA-V
von: Wan, Zhewen, et al.
Veröffentlicht: (2026)
von: Wan, Zhewen, et al.
Veröffentlicht: (2026)
DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models
von: Lu, Cheng, et al.
Veröffentlicht: (2022)
von: Lu, Cheng, et al.
Veröffentlicht: (2022)
LLaDA-MedV: Exploring Large Language Diffusion Models for Biomedical Image Understanding
von: Dong, Xuanzhao, et al.
Veröffentlicht: (2025)
von: Dong, Xuanzhao, et al.
Veröffentlicht: (2025)
LLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion Large Language Model
von: AI, Inclusion, et al.
Veröffentlicht: (2026)
von: AI, Inclusion, et al.
Veröffentlicht: (2026)
LLaDA2.0: Scaling Up Diffusion Language Models to 100B
von: Bie, Tiwei, et al.
Veröffentlicht: (2025)
von: Bie, Tiwei, et al.
Veröffentlicht: (2025)
The Blessing of Randomness: SDE Beats ODE in General Diffusion-based Image Editing
von: Nie, Shen, et al.
Veröffentlicht: (2023)
von: Nie, Shen, et al.
Veröffentlicht: (2023)
On Memorization in Diffusion Models
von: Gu, Xiangming, et al.
Veröffentlicht: (2023)
von: Gu, Xiangming, et al.
Veröffentlicht: (2023)
BayesDiff: Estimating Pixel-wise Uncertainty in Diffusion via Bayesian Inference
von: Kou, Siqi, et al.
Veröffentlicht: (2023)
von: Kou, Siqi, et al.
Veröffentlicht: (2023)
Improving Long-Text Alignment for Text-to-Image Diffusion Models
von: Liu, Luping, et al.
Veröffentlicht: (2024)
von: Liu, Luping, et al.
Veröffentlicht: (2024)
Large Language Diffusion Models
von: Nie, Shen, et al.
Veröffentlicht: (2025)
von: Nie, Shen, et al.
Veröffentlicht: (2025)
Adaptive Moments are Surprisingly Effective for Plug-and-Play Diffusion Sampling
von: Belardi, Christian, et al.
Veröffentlicht: (2026)
von: Belardi, Christian, et al.
Veröffentlicht: (2026)
Scaling Diffusion Transformers Efficiently via $μ$P
von: Zheng, Chenyu, et al.
Veröffentlicht: (2025)
von: Zheng, Chenyu, et al.
Veröffentlicht: (2025)
BADiff: Bandwidth Adaptive Diffusion Model
von: Zhang, Xi, et al.
Veröffentlicht: (2025)
von: Zhang, Xi, et al.
Veröffentlicht: (2025)
Adaptive Non-uniform Timestep Sampling for Accelerating Diffusion Model Training
von: Kim, Myunsoo, et al.
Veröffentlicht: (2024)
von: Kim, Myunsoo, et al.
Veröffentlicht: (2024)
OmniGAIA: Towards Native Omni-Modal AI Agents
von: Li, Xiaoxi, et al.
Veröffentlicht: (2026)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2026)
LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model
von: Wang, Xiyao, et al.
Veröffentlicht: (2025)
von: Wang, Xiyao, et al.
Veröffentlicht: (2025)
ODE$_t$(ODE$_l$): Shortcutting the Time and the Length in Diffusion and Flow Models for Faster Sampling
von: Gudovskiy, Denis, et al.
Veröffentlicht: (2025)
von: Gudovskiy, Denis, et al.
Veröffentlicht: (2025)
CRM: Single Image to 3D Textured Mesh with Convolutional Reconstruction Model
von: Wang, Zhengyi, et al.
Veröffentlicht: (2024)
von: Wang, Zhengyi, et al.
Veröffentlicht: (2024)
Enhancing Diffusion Model Stability for Image Restoration via Gradient Management
von: Wu, Hongjie, et al.
Veröffentlicht: (2025)
von: Wu, Hongjie, et al.
Veröffentlicht: (2025)
Diffusion Models With Learned Adaptive Noise
von: Sahoo, Subham Sekhar, et al.
Veröffentlicht: (2023)
von: Sahoo, Subham Sekhar, et al.
Veröffentlicht: (2023)
R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning
von: Zhao, Jiaxing, et al.
Veröffentlicht: (2025)
von: Zhao, Jiaxing, et al.
Veröffentlicht: (2025)
LLaVA-Zip: Adaptive Visual Token Compression with Intrinsic Image Information
von: Wang, Ke, et al.
Veröffentlicht: (2024)
von: Wang, Ke, et al.
Veröffentlicht: (2024)
Consistency Model is an Effective Posterior Sample Approximation for Diffusion Inverse Solvers
von: Xu, Tongda, et al.
Veröffentlicht: (2024)
von: Xu, Tongda, et al.
Veröffentlicht: (2024)
Robust PCA Based on Adaptive Weighted Least Squares and Low-Rank Matrix Factorization
von: Li, Kexin, et al.
Veröffentlicht: (2024)
von: Li, Kexin, et al.
Veröffentlicht: (2024)
OmniCache: A Trajectory-Oriented Global Perspective on Training-Free Cache Reuse for Diffusion Transformer Models
von: Chu, Huanpeng, et al.
Veröffentlicht: (2025)
von: Chu, Huanpeng, et al.
Veröffentlicht: (2025)
Adaptive Training Meets Progressive Scaling: Elevating Efficiency in Diffusion Models
von: Li, Wenhao, et al.
Veröffentlicht: (2023)
von: Li, Wenhao, et al.
Veröffentlicht: (2023)
Estimation of Confidence Bounds in Binary Classification using Wilson Score Kernel Density Estimation
von: Iversen, Thorbjørn Mosekjær, et al.
Veröffentlicht: (2026)
von: Iversen, Thorbjørn Mosekjær, et al.
Veröffentlicht: (2026)
Yo'LLaVA: Your Personalized Language and Vision Assistant
von: Nguyen, Thao, et al.
Veröffentlicht: (2024)
von: Nguyen, Thao, et al.
Veröffentlicht: (2024)
VideoDPO: Omni-Preference Alignment for Video Diffusion Generation
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
Variance-Aware Adaptive Weighting for Diffusion Model Training
von: Sun, Nanlong, et al.
Veröffentlicht: (2026)
von: Sun, Nanlong, et al.
Veröffentlicht: (2026)
Owl-1: Omni World Model for Consistent Long Video Generation
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers
von: Zhao, Min, et al.
Veröffentlicht: (2025)
von: Zhao, Min, et al.
Veröffentlicht: (2025)
MALT Diffusion: Memory-Augmented Latent Transformers for Any-Length Video Generation
von: Yu, Sihyun, et al.
Veröffentlicht: (2025)
von: Yu, Sihyun, et al.
Veröffentlicht: (2025)
Spatio-Temporal Branching for Motion Prediction using Motion Increments
von: Wang, Jiexin, et al.
Veröffentlicht: (2023)
von: Wang, Jiexin, et al.
Veröffentlicht: (2023)
Learning Stackable and Skippable LEGO Bricks for Efficient, Reconfigurable, and Variable-Resolution Diffusion Modeling
von: Zheng, Huangjie, et al.
Veröffentlicht: (2023)
von: Zheng, Huangjie, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning
von: You, Zebin, et al.
Veröffentlicht: (2025) -
Effective and Efficient Masked Image Generation Models
von: You, Zebin, et al.
Veröffentlicht: (2025) -
LLaDA 1.5: Variance-Reduced Preference Optimization for Large Language Diffusion Models
von: Zhu, Fengqi, et al.
Veröffentlicht: (2025) -
LLaDA-VLA: Vision Language Diffusion Action Models
von: Wen, Yuqing, et al.
Veröffentlicht: (2025) -
Are Images Indistinguishable to Humans Also Indistinguishable to Classifiers?
von: You, Zebin, et al.
Veröffentlicht: (2024)