A Systematic Post-Train Framework for Video Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Xue, Zeyue, Fu, Siming, Huang, Jie, Lu, Shuai, Li, Haoran, Liu, Yijun, Li, Yuming, He, Xiaoxuan, Chen, Mengzhao, Huang, Haoyang, Duan, Nan, Luo, Ping |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization
by: He, Xiaoxuan, et al.
Published: (2026)
by: He, Xiaoxuan, et al.
Published: (2026)
OmniForcing: Unleashing Real-time Joint Audio-Visual Generation
by: Su, Yaofeng, et al.
Published: (2026)
by: Su, Yaofeng, et al.
Published: (2026)
Astrolabe: Steering Forward-Process Reinforcement Learning for Distilled Autoregressive Video Models
by: Zhang, Songchun, et al.
Published: (2026)
by: Zhang, Songchun, et al.
Published: (2026)
OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation
by: Zhang, Guohui, et al.
Published: (2026)
by: Zhang, Guohui, et al.
Published: (2026)
RAE-AR: Taming Autoregressive Models with Representation Autoencoders
by: Yu, Hu, et al.
Published: (2026)
by: Yu, Hu, et al.
Published: (2026)
Scaling Law for Quantization-Aware Training
by: Chen, Mengzhao, et al.
Published: (2025)
by: Chen, Mengzhao, et al.
Published: (2025)
WorldWeaver: Generating Long-Horizon Video Worlds via Rich Perception
by: Liu, Zhiheng, et al.
Published: (2025)
by: Liu, Zhiheng, et al.
Published: (2025)
DanceGRPO: Unleashing GRPO on Visual Generation
by: Xue, Zeyue, et al.
Published: (2025)
by: Xue, Zeyue, et al.
Published: (2025)
Unleashing Efficient Asynchronous RL Post-Training via Staleness-Constrained Rollout Coordination
by: Li, Haoyang, et al.
Published: (2026)
by: Li, Haoyang, et al.
Published: (2026)
INT v.s. FP: A Comprehensive Study of Fine-Grained Low-bit Quantization Formats
by: Chen, Mengzhao, et al.
Published: (2025)
by: Chen, Mengzhao, et al.
Published: (2025)
A study on the relationship between online gaming behavior and suicidal ideation among college students
by: Shouting Lu, et al.
Published: (2026)
by: Shouting Lu, et al.
Published: (2026)
Thinking with Novel Views: A Systematic Analysis of Generative-Augmented Spatial Intelligence
by: Zhang, Yanbing, et al.
Published: (2026)
by: Zhang, Yanbing, et al.
Published: (2026)
Video-OPD: Efficient Post-Training of Multimodal Large Language Models for Temporal Video Grounding via On-Policy Distillation
by: Li, Jiaze, et al.
Published: (2026)
by: Li, Jiaze, et al.
Published: (2026)
Frame-Level Captions for Long Video Generation with Complex Multi Scenes
by: Zheng, Guangcong, et al.
Published: (2025)
by: Zheng, Guangcong, et al.
Published: (2025)
AEGPO: Adaptive Entropy-Guided Policy Optimization for Diffusion Models
by: Li, Yuming, et al.
Published: (2026)
by: Li, Yuming, et al.
Published: (2026)
LinVideo: A Post-Training Framework towards O(n) Attention in Efficient Video Generation
by: Huang, Yushi, et al.
Published: (2025)
by: Huang, Yushi, et al.
Published: (2025)
SAIL: Self-Amplified Iterative Learning for Diffusion Model Alignment with Minimal Human Feedback
by: He, Xiaoxuan, et al.
Published: (2026)
by: He, Xiaoxuan, et al.
Published: (2026)
Harvest Video Foundation Models via Efficient Post-Pretraining
by: Li, Yizhuo, et al.
Published: (2023)
by: Li, Yizhuo, et al.
Published: (2023)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
by: Zhong, Yunshan, et al.
Published: (2023)
by: Zhong, Yunshan, et al.
Published: (2023)
Relationship Analysis of Image-Text Pair in SNS Posts
by: Nabeoka, Takuto, et al.
Published: (2025)
by: Nabeoka, Takuto, et al.
Published: (2025)
Enhancing Large Language Models in Coding Through Multi-Perspective Self-Consistency
by: Huang, Baizhou, et al.
Published: (2023)
by: Huang, Baizhou, et al.
Published: (2023)
TempFlow-GRPO: When Timing Matters for GRPO in Flow Models
by: He, Xiaoxuan, et al.
Published: (2025)
by: He, Xiaoxuan, et al.
Published: (2025)
Alchemy: Amplifying Theorem-Proving Capability through Symbolic Mutation
by: Wu, Shaonan, et al.
Published: (2024)
by: Wu, Shaonan, et al.
Published: (2024)
Nonsurgical Approach to Treat Post‐Rhinoplasty Vascular Complication With Hyaluronic Acid Injection
by: Shuai Qiang, et al.
Published: (2025)
by: Shuai Qiang, et al.
Published: (2025)
EfficientQAT: Efficient Quantization-Aware Training for Large Language Models
by: Chen, Mengzhao, et al.
Published: (2024)
by: Chen, Mengzhao, et al.
Published: (2024)
HunyuanVideo: A Systematic Framework For Large Video Generative Models
by: Kong, Weijie, et al.
Published: (2024)
by: Kong, Weijie, et al.
Published: (2024)
Sparse Reasoning is Enough: Biological-Inspired Framework for Video Anomaly Detection with Large Pre-trained Models
by: Huang, He, et al.
Published: (2025)
by: Huang, He, et al.
Published: (2025)
Selene: Pioneering Automated Proof in Software Verification
by: Zhang, Lichen, et al.
Published: (2024)
by: Zhang, Lichen, et al.
Published: (2024)
AdaptiveLoad: Towards Efficient Video Diffusion Transformer Training
by: Guo, Yucheng, et al.
Published: (2026)
by: Guo, Yucheng, et al.
Published: (2026)
Post-Fair Federated Learning: Achieving Group and Community Fairness in Federated Learning via Post-processing
by: Duan, Yuying, et al.
Published: (2024)
by: Duan, Yuying, et al.
Published: (2024)
Vivim: a Video Vision Mamba for Medical Video Segmentation
by: Yang, Yijun, et al.
Published: (2024)
by: Yang, Yijun, et al.
Published: (2024)
Widespread occurrence of cyhalofop‐resistant Leptochloa chinensis in eastern China was mainly propelled by target‐site mutations
by: Guoqi Chen, et al.
Published: (2025)
by: Guoqi Chen, et al.
Published: (2025)
Safety-Constrained Reinforcement Learning with Post-Training Reachability Verification for Robot Navigation
by: He, Qisong, et al.
Published: (2026)
by: He, Qisong, et al.
Published: (2026)
RAPHAEL: Text-to-Image Generation via Large Mixture of Diffusion Paths
by: Xue, Zeyue, et al.
Published: (2023)
by: Xue, Zeyue, et al.
Published: (2023)
DiRL: An Efficient Post-Training Framework for Diffusion Language Models
by: Zhu, Ying, et al.
Published: (2025)
by: Zhu, Ying, et al.
Published: (2025)
MS-Diffusion: Multi-subject Zero-shot Image Personalization with Layout Guidance
by: Wang, Xierui, et al.
Published: (2024)
by: Wang, Xierui, et al.
Published: (2024)
A Unified Gradient-based Framework for Task-agnostic Continual Learning-Unlearning
by: Huang, Zhehao, et al.
Published: (2025)
by: Huang, Zhehao, et al.
Published: (2025)
Multi-View Representation is What You Need for Point-Cloud Pre-Training
by: Yan, Siming, et al.
Published: (2023)
by: Yan, Siming, et al.
Published: (2023)
Time-dependent Flows and Their Applications in Parabolic-parabolic Patlak-Keller-Segel Systems Part II: Shear Flows
by: He, Siming
Published: (2024)
by: He, Siming
Published: (2024)
Time-dependent Flows and Their Applications in Parabolic-parabolic Patlak-Keller-Segel Systems Part I: Alternating Flows
by: He, Siming
Published: (2024)
by: He, Siming
Published: (2024)
Similar Items
-
Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization
by: He, Xiaoxuan, et al.
Published: (2026) -
OmniForcing: Unleashing Real-time Joint Audio-Visual Generation
by: Su, Yaofeng, et al.
Published: (2026) -
Astrolabe: Steering Forward-Process Reinforcement Learning for Distilled Autoregressive Video Models
by: Zhang, Songchun, et al.
Published: (2026) -
OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation
by: Zhang, Guohui, et al.
Published: (2026) -
RAE-AR: Taming Autoregressive Models with Representation Autoencoders
by: Yu, Hu, et al.
Published: (2026)