FilmWeaver: Weaving Consistent Multi-Shot Videos with Cache-Guided Autoregressive Diffusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Luo, Xiangyang, Li, Qingyu, Liu, Xiaokun, Qin, Wenyu, Yang, Miao, Wang, Meng, Wan, Pengfei, Zhang, Di, Gai, Kun, Huang, Shao-Lun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond the Golden Data: Resolving the Motion-Vision Quality Dilemma via Timestep Selective Training
von: Luo, Xiangyang, et al.
Veröffentlicht: (2026)
von: Luo, Xiangyang, et al.
Veröffentlicht: (2026)
AEGPO: Adaptive Entropy-Guided Policy Optimization for Diffusion Models
von: Li, Yuming, et al.
Veröffentlicht: (2026)
von: Li, Yuming, et al.
Veröffentlicht: (2026)
HairWeaver: Few-Shot Photorealistic Hair Motion Synthesis with Sim-to-Real Guided Video Diffusion
von: Chang, Di, et al.
Veröffentlicht: (2026)
von: Chang, Di, et al.
Veröffentlicht: (2026)
HumanAesExpert: Advancing a Multi-Modality Foundation Model for Human Image Aesthetic Assessment
von: Liao, Zhichao, et al.
Veröffentlicht: (2025)
von: Liao, Zhichao, et al.
Veröffentlicht: (2025)
ConceptWeaver: Weaving Disentangled Concepts with Flow
von: Chen, Jintao, et al.
Veröffentlicht: (2026)
von: Chen, Jintao, et al.
Veröffentlicht: (2026)
Analytic Score Optimization for Multi Dimension Video Quality Assessment
von: Lin, Boda, et al.
Veröffentlicht: (2026)
von: Lin, Boda, et al.
Veröffentlicht: (2026)
BindWeave: Subject-Consistent Video Generation via Cross-Modal Integration
von: Li, Zhaoyang, et al.
Veröffentlicht: (2025)
von: Li, Zhaoyang, et al.
Veröffentlicht: (2025)
Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models
von: Ji, Yicheng, et al.
Veröffentlicht: (2026)
von: Ji, Yicheng, et al.
Veröffentlicht: (2026)
MultiShotMaster: A Controllable Multi-Shot Video Generation Framework
von: Wang, Qinghe, et al.
Veröffentlicht: (2025)
von: Wang, Qinghe, et al.
Veröffentlicht: (2025)
VulWeaver: Weaving Broken Semantics for Grounded Vulnerability Detection
von: Cao, Yiheng, et al.
Veröffentlicht: (2026)
von: Cao, Yiheng, et al.
Veröffentlicht: (2026)
CanonSwap: High-Fidelity and Consistent Video Face Swapping via Canonical Space Modulation
von: Luo, Xiangyang, et al.
Veröffentlicht: (2025)
von: Luo, Xiangyang, et al.
Veröffentlicht: (2025)
CrossWeaver: Cross-modal Weaving for Arbitrary-Modality Semantic Segmentation
von: Zhang, Zelin, et al.
Veröffentlicht: (2026)
von: Zhang, Zelin, et al.
Veröffentlicht: (2026)
ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning
von: Huang, Yuzhou, et al.
Veröffentlicht: (2025)
von: Huang, Yuzhou, et al.
Veröffentlicht: (2025)
FullDiT2: Efficient In-Context Conditioning for Video Diffusion Transformers
von: He, Xuanhua, et al.
Veröffentlicht: (2025)
von: He, Xuanhua, et al.
Veröffentlicht: (2025)
Scaling Image and Video Generation via Test-Time Evolutionary Search
von: He, Haoran, et al.
Veröffentlicht: (2025)
von: He, Haoran, et al.
Veröffentlicht: (2025)
Boosting Resolution Generalization of Diffusion Transformers with Randomized Positional Encodings
von: Hou, Liang, et al.
Veröffentlicht: (2025)
von: Hou, Liang, et al.
Veröffentlicht: (2025)
PatchVSR: Breaking Video Diffusion Resolution Limits with Patch-wise Video Super-Resolution
von: Du, Shian, et al.
Veröffentlicht: (2025)
von: Du, Shian, et al.
Veröffentlicht: (2025)
MemWeaver: Weaving Hybrid Memories for Traceable Long-Horizon Agentic Reasoning
von: Ye, Juexiang, et al.
Veröffentlicht: (2026)
von: Ye, Juexiang, et al.
Veröffentlicht: (2026)
A Survey of Interactive Generative Video
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
VidEmo: Affective-Tree Reasoning for Emotion-Centric Video Foundation Models
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2025)
VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2026)
von: Yesiltepe, Hidir, et al.
Veröffentlicht: (2026)
ToolWeaver: Weaving Collaborative Semantics for Scalable Tool Use in Large Language Models
von: Fang, Bowen, et al.
Veröffentlicht: (2026)
von: Fang, Bowen, et al.
Veröffentlicht: (2026)
Past- and Future-Informed KV Cache Policy with Salience Estimation in Autoregressive Video Diffusion
von: Chen, Hanmo, et al.
Veröffentlicht: (2026)
von: Chen, Hanmo, et al.
Veröffentlicht: (2026)
Fast Autoregressive Video Diffusion and World Models with Temporal Cache Compression and Sparse Attention
von: Samuel, Dvir, et al.
Veröffentlicht: (2026)
von: Samuel, Dvir, et al.
Veröffentlicht: (2026)
Score Augmentation for Diffusion Models
von: Hou, Liang, et al.
Veröffentlicht: (2025)
von: Hou, Liang, et al.
Veröffentlicht: (2025)
Efficient Video Diffusion Models: Advancements and Challenges
von: Shao, Shitong, et al.
Veröffentlicht: (2026)
von: Shao, Shitong, et al.
Veröffentlicht: (2026)
A$^2$RD: Agentic Autoregressive Diffusion for Long Video Consistency
von: Long, Do Xuan, et al.
Veröffentlicht: (2026)
von: Long, Do Xuan, et al.
Veröffentlicht: (2026)
ScaleWeaver: Weaving Efficient Controllable T2I Generation with Multi-Scale Reference Attention
von: Liu, Keli, et al.
Veröffentlicht: (2025)
von: Liu, Keli, et al.
Veröffentlicht: (2025)
UniVideo: Unified Understanding, Generation, and Editing for Videos
von: Wei, Cong, et al.
Veröffentlicht: (2025)
von: Wei, Cong, et al.
Veröffentlicht: (2025)
Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
von: Yu, Jiwen, et al.
Veröffentlicht: (2025)
One-Shot Heterogeneous Federated Learning with Local Model-Guided Diffusion Models
von: Yang, Mingzhao, et al.
Veröffentlicht: (2023)
von: Yang, Mingzhao, et al.
Veröffentlicht: (2023)
Visual-Aware CoT: Achieving High-Fidelity Visual Consistency in Unified Models
von: Ye, Zixuan, et al.
Veröffentlicht: (2025)
von: Ye, Zixuan, et al.
Veröffentlicht: (2025)
Motion-Aware Caching for Efficient Autoregressive Video Generation
von: Xu, Jing, et al.
Veröffentlicht: (2026)
von: Xu, Jing, et al.
Veröffentlicht: (2026)
Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing
von: Gao, Kaifeng, et al.
Veröffentlicht: (2024)
von: Gao, Kaifeng, et al.
Veröffentlicht: (2024)
VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization
von: Cheng, Junhao, et al.
Veröffentlicht: (2026)
von: Cheng, Junhao, et al.
Veröffentlicht: (2026)
FullDiT: Multi-Task Video Generative Foundation Model with Full Attention
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
von: Ju, Xuan, et al.
Veröffentlicht: (2025)
CamCloneMaster: Enabling Reference-based Camera Control for Video Generation
von: Luo, Yawen, et al.
Veröffentlicht: (2025)
von: Luo, Yawen, et al.
Veröffentlicht: (2025)
DualWeaver: Synergistic Feature Weaving Surrogates for Multivariate Forecasting with Univariate Time Series Foundation Models
von: Li, Jinpeng, et al.
Veröffentlicht: (2026)
von: Li, Jinpeng, et al.
Veröffentlicht: (2026)
AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
von: Wang, Zun, et al.
Veröffentlicht: (2026)
von: Wang, Zun, et al.
Veröffentlicht: (2026)
Calibrated Cache Model for Few-Shot Vision-Language Model Adaptation
von: Ding, Kun, et al.
Veröffentlicht: (2024)
von: Ding, Kun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Beyond the Golden Data: Resolving the Motion-Vision Quality Dilemma via Timestep Selective Training
von: Luo, Xiangyang, et al.
Veröffentlicht: (2026) -
AEGPO: Adaptive Entropy-Guided Policy Optimization for Diffusion Models
von: Li, Yuming, et al.
Veröffentlicht: (2026) -
HairWeaver: Few-Shot Photorealistic Hair Motion Synthesis with Sim-to-Real Guided Video Diffusion
von: Chang, Di, et al.
Veröffentlicht: (2026) -
HumanAesExpert: Advancing a Multi-Modality Foundation Model for Human Image Aesthetic Assessment
von: Liao, Zhichao, et al.
Veröffentlicht: (2025) -
ConceptWeaver: Weaving Disentangled Concepts with Flow
von: Chen, Jintao, et al.
Veröffentlicht: (2026)