ContentV: Efficient Training of Video Generation Models with Limited Compute
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Wenfeng, Chen, Renjie, Liu, Boyuan, Yan, Shiyue, Feng, Ruoyu, Wei, Jiangchuan, Zhang, Yichen, Zhou, Yimeng, Feng, Chao, Ran, Jiao, Wu, Qi, Liu, Zuotao, Guo, Mingyu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CascadeV: An Implementation of Wurstchen Architecture for Video Generation
by: Lin, Wenfeng, et al.
Published: (2025)
by: Lin, Wenfeng, et al.
Published: (2025)
EchoVideo: Identity-Preserving Human Video Generation by Multimodal Feature Fusion
by: Wei, Jiangchuan, et al.
Published: (2025)
by: Wei, Jiangchuan, et al.
Published: (2025)
Towards Self-Improvement of Diffusion Models via Group Preference Optimization
by: Chen, Renjie, et al.
Published: (2025)
by: Chen, Renjie, et al.
Published: (2025)
LVFace: Progressive Cluster Optimization for Large Vision Models in Face Recognition
by: You, Jinghan, et al.
Published: (2025)
by: You, Jinghan, et al.
Published: (2025)
Semantics Lead the Way: Harmonizing Semantic and Texture Modeling with Asynchronous Latent Diffusion
by: Pan, Yueming, et al.
Published: (2025)
by: Pan, Yueming, et al.
Published: (2025)
Towards User-level QoE: Large-scale Practice in Personalized Optimization of Adaptive Video Streaming
by: Jia, Lianchen, et al.
Published: (2025)
by: Jia, Lianchen, et al.
Published: (2025)
AdaLRS: Loss-Guided Adaptive Learning Rate Search for Efficient Foundation Model Pretraining
by: Dong, Hongyuan, et al.
Published: (2025)
by: Dong, Hongyuan, et al.
Published: (2025)
Scalable Vision Language Model Training via High Quality Data Curation
by: Dong, Hongyuan, et al.
Published: (2025)
by: Dong, Hongyuan, et al.
Published: (2025)
Let Me Show You: Learning by Retrieving from Egocentric Video for Robotic Manipulation
by: Zhu, Yichen, et al.
Published: (2025)
by: Zhu, Yichen, et al.
Published: (2025)
StreamGVE: Training-Free Video Editing via Few-Step Streaming Video Generation
by: Jiao, Guanlong, et al.
Published: (2026)
by: Jiao, Guanlong, et al.
Published: (2026)
Interest Clock: Time Perception in Real-Time Streaming Recommendation System
by: Zhu, Yongchun, et al.
Published: (2024)
by: Zhu, Yongchun, et al.
Published: (2024)
Long-Term Interest Clock: Fine-Grained Time Perception in Streaming Recommendation System
by: Zhu, Yongchun, et al.
Published: (2025)
by: Zhu, Yongchun, et al.
Published: (2025)
Asymmetric Diffusion Recommendation Model
by: Zhu, Yongchun, et al.
Published: (2025)
by: Zhu, Yongchun, et al.
Published: (2025)
BVI-UGC: A Video Quality Database for User-Generated Content Transcoding
by: Qi, Zihao, et al.
Published: (2024)
by: Qi, Zihao, et al.
Published: (2024)
Trinity: Syncretizing Multi-/Long-tail/Long-term Interests All in One
by: Yan, Jing, et al.
Published: (2024)
by: Yan, Jing, et al.
Published: (2024)
UniFormer: Unified and Efficient Transformer for Reasoning Across General and Custom Computing
by: Ran, Zhuoheng, et al.
Published: (2025)
by: Ran, Zhuoheng, et al.
Published: (2025)
Full-reference Video Quality Assessment for User Generated Content Transcoding
by: Qi, Zihao, et al.
Published: (2023)
by: Qi, Zihao, et al.
Published: (2023)
Generative AI for Multimedia Communication: Recent Advances, An Information-Theoretic Framework, and Future Opportunities
by: Jin, Yili, et al.
Published: (2025)
by: Jin, Yili, et al.
Published: (2025)
Beyond Stochastic Exploration: What Makes Training Data Valuable for Agentic Search
by: Hao, Chuzhan, et al.
Published: (2026)
by: Hao, Chuzhan, et al.
Published: (2026)
Fre-Res: Frequency-Residual Video Token Compression for Efficient Video MLLMs
by: Feng, Yigui, et al.
Published: (2026)
by: Feng, Yigui, et al.
Published: (2026)
StarStream: Live Video Analytics over Space Networking
by: Zhang, Miao, et al.
Published: (2025)
by: Zhang, Miao, et al.
Published: (2025)
Generative Flow Networks for Personalized Multimedia Systems: A Case Study on Short Video Feeds
by: Jin, Yili, et al.
Published: (2025)
by: Jin, Yili, et al.
Published: (2025)
Byzantine-Robust Distributed SGD: A Unified Analysis and Tight Error Bounds
by: Ruan, Boyuan, et al.
Published: (2026)
by: Ruan, Boyuan, et al.
Published: (2026)
MambaVSR: Content-Aware Scanning State Space Model for Video Super-Resolution
by: He, Linfeng, et al.
Published: (2025)
by: He, Linfeng, et al.
Published: (2025)
Optimizing Mobile-Friendly Viewport Prediction for Live 360-Degree Video Streaming
by: Zhang, Lei, et al.
Published: (2024)
by: Zhang, Lei, et al.
Published: (2024)
OmniSense: Towards Edge-Assisted Online Analytics for 360-Degree Videos
by: Zhang, Miao, et al.
Published: (2025)
by: Zhang, Miao, et al.
Published: (2025)
ReMA: A Training-Free Plug-and-Play Mixing Augmentation for Video Behavior Recognition
by: Cui, Feng-Qi, et al.
Published: (2026)
by: Cui, Feng-Qi, et al.
Published: (2026)
StreamingAssistant: Efficient Visual Token Pruning for Accelerating Online Video Understanding
by: Jin, Xinqi, et al.
Published: (2025)
by: Jin, Xinqi, et al.
Published: (2025)
AdaF^2M^2: Comprehensive Learning and Responsive Leveraging Features in Recommendation System
by: Zhu, Yongchun, et al.
Published: (2025)
by: Zhu, Yongchun, et al.
Published: (2025)
Interaction Design for Human-AI Choreography Co-creation
by: Liu, Yimeng
Published: (2024)
by: Liu, Yimeng
Published: (2024)
Engineering.ai: A Platform for Teams of AI Engineers in Computational Design
by: Xu, Ran, et al.
Published: (2025)
by: Xu, Ran, et al.
Published: (2025)
Multi-Mode Pneumatic Artificial Muscles Driven by Hybrid Positive-Negative Pressure
by: Feng, Siyuan, et al.
Published: (2026)
by: Feng, Siyuan, et al.
Published: (2026)
Smallest distances between zeros of Gaussian analytic functions
by: Feng, Renjie, et al.
Published: (2026)
by: Feng, Renjie, et al.
Published: (2026)
Large gaps of CUE and GUE
by: Feng, Renjie, et al.
Published: (2018)
by: Feng, Renjie, et al.
Published: (2018)
Poisson approximation of the largest gaps between zeros of a stationary Gaussian process
by: Feng, Renjie, et al.
Published: (2026)
by: Feng, Renjie, et al.
Published: (2026)
AdaOcc: Adaptive-Resolution Occupancy Prediction
by: Chen, Chao, et al.
Published: (2024)
by: Chen, Chao, et al.
Published: (2024)
EACO-RAG: Towards Distributed Tiered LLM Deployment using Edge-Assisted and Collaborative RAG with Adaptive Knowledge Update
by: Li, Jiaxing, et al.
Published: (2024)
by: Li, Jiaxing, et al.
Published: (2024)
Lumos: Efficient Performance Modeling and Estimation for Large-scale LLM Training
by: Liang, Mingyu, et al.
Published: (2025)
by: Liang, Mingyu, et al.
Published: (2025)
EDT: An Efficient Diffusion Transformer Framework Inspired by Human-like Sketching
by: Chen, Xinwang, et al.
Published: (2024)
by: Chen, Xinwang, et al.
Published: (2024)
Reconfigurable Intelligent Surface-Enabled Array Radar for Interference Mitigation
by: Chen, Shengyao, et al.
Published: (2024)
by: Chen, Shengyao, et al.
Published: (2024)
Similar Items
-
CascadeV: An Implementation of Wurstchen Architecture for Video Generation
by: Lin, Wenfeng, et al.
Published: (2025) -
EchoVideo: Identity-Preserving Human Video Generation by Multimodal Feature Fusion
by: Wei, Jiangchuan, et al.
Published: (2025) -
Towards Self-Improvement of Diffusion Models via Group Preference Optimization
by: Chen, Renjie, et al.
Published: (2025) -
LVFace: Progressive Cluster Optimization for Large Vision Models in Face Recognition
by: You, Jinghan, et al.
Published: (2025) -
Semantics Lead the Way: Harmonizing Semantic and Texture Modeling with Asynchronous Latent Diffusion
by: Pan, Yueming, et al.
Published: (2025)