Brick-Diffusion: Generating Long Videos with Brick-to-Wall Denoising
Fuente:
arXiv
Saved in:
| Main Authors: | Yuan, Yunlong, Guo, Yuanfan, Wang, Chunwei, Xu, Hang, Zhang, Li |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BrickNet: Graph-Backed Generative Brick Assembly
by: Kulits, Peter, et al.
Published: (2026)
by: Kulits, Peter, et al.
Published: (2026)
Stacking Brick by Brick: Aligned Feature Isolation for Incremental Face Forgery Detection
by: Cheng, Jikang, et al.
Published: (2024)
by: Cheng, Jikang, et al.
Published: (2024)
SlowFocus: Enhancing Fine-grained Temporal Understanding in Video LLM
by: Nie, Ming, et al.
Published: (2026)
by: Nie, Ming, et al.
Published: (2026)
TreeSBA: Tree-Transformer for Self-Supervised Sequential Brick Assembly
by: Guo, Mengqi, et al.
Published: (2024)
by: Guo, Mengqi, et al.
Published: (2024)
Generating Physically Stable and Buildable Brick Structures from Text
by: Pun, Ava, et al.
Published: (2025)
by: Pun, Ava, et al.
Published: (2025)
EasyControl: Transfer ControlNet to Video Diffusion for Controllable Generation and Interpolation
by: Wang, Cong, et al.
Published: (2024)
by: Wang, Cong, et al.
Published: (2024)
KFFocus: Highlighting Keyframes for Enhanced Video Understanding
by: Nie, Ming, et al.
Published: (2025)
by: Nie, Ming, et al.
Published: (2025)
Learning Stackable and Skippable LEGO Bricks for Efficient, Reconfigurable, and Variable-Resolution Diffusion Modeling
by: Zheng, Huangjie, et al.
Published: (2023)
by: Zheng, Huangjie, et al.
Published: (2023)
Lego-Edit: A General Image Editing Framework with Model-Level Bricks and MLLM Builder
by: Jia, Qifei, et al.
Published: (2025)
by: Jia, Qifei, et al.
Published: (2025)
Brick Kiln Dataset for Pakistan's IGP Region Using AI
by: Hamdani, Muhammad Suleman Ali, et al.
Published: (2024)
by: Hamdani, Muhammad Suleman Ali, et al.
Published: (2024)
Eye in the Sky: Detection and Compliance Monitoring of Brick Kilns using Satellite Imagery
by: Mondal, Rishabh, et al.
Published: (2024)
by: Mondal, Rishabh, et al.
Published: (2024)
VTimeCoT: Thinking by Drawing for Video Temporal Grounding and Reasoning
by: Zhang, Jinglei, et al.
Published: (2025)
by: Zhang, Jinglei, et al.
Published: (2025)
Towards Unified Multimodal Interleaved Generation via Group Relative Policy Optimization
by: Nie, Ming, et al.
Published: (2026)
by: Nie, Ming, et al.
Published: (2026)
HiAR: Efficient Autoregressive Long Video Generation via Hierarchical Denoising
by: Zou, Kai, et al.
Published: (2026)
by: Zou, Kai, et al.
Published: (2026)
ILLUME+: Illuminating Unified MLLM with Dual Visual Tokenization and Diffusion Refinement
by: Huang, Runhui, et al.
Published: (2025)
by: Huang, Runhui, et al.
Published: (2025)
GLaVE-Cap: Global-Local Aligned Video Captioning with Vision Expert Integration
by: Xu, Wan, et al.
Published: (2025)
by: Xu, Wan, et al.
Published: (2025)
From Summary to Action: Enhancing Large Language Models for Complex Tasks with Open World APIs
by: Liu, Yulong, et al.
Published: (2024)
by: Liu, Yulong, et al.
Published: (2024)
Scalable Methods for Brick Kiln Detection and Compliance Monitoring from Satellite Imagery: A Deployment Case Study in India
by: Mondal, Rishabh, et al.
Published: (2024)
by: Mondal, Rishabh, et al.
Published: (2024)
Self-Adaptive Reality-Guided Diffusion for Artifact-Free Super-Resolution
by: Zheng, Qingping, et al.
Published: (2024)
by: Zheng, Qingping, et al.
Published: (2024)
Denoise to Track: Harnessing Video Diffusion Priors for Robust Correspondence
by: Yuan, Tianyu, et al.
Published: (2025)
by: Yuan, Tianyu, et al.
Published: (2025)
Detecting Brick Kiln Infrastructure at Scale: Graph, Foundation, and Remote Sensing Models for Satellite Imagery Data
by: Nazir, Usman, et al.
Published: (2026)
by: Nazir, Usman, et al.
Published: (2026)
Video Summarization using Denoising Diffusion Probabilistic Model
by: Shang, Zirui, et al.
Published: (2024)
by: Shang, Zirui, et al.
Published: (2024)
Asynchronous Denoising Diffusion Models for Aligning Text-to-Image Generation
by: Hu, Zijing, et al.
Published: (2025)
by: Hu, Zijing, et al.
Published: (2025)
HumanRefiner: Benchmarking Abnormal Human Generation and Refining with Coarse-to-fine Pose-Reversible Guidance
by: Fang, Guian, et al.
Published: (2024)
by: Fang, Guian, et al.
Published: (2024)
ARLON: Boosting Diffusion Transformers with Autoregressive Models for Long Video Generation
by: Li, Zongyi, et al.
Published: (2024)
by: Li, Zongyi, et al.
Published: (2024)
EgoLCD: Egocentric Video Generation with Long Context Diffusion
by: Zhang, Liuzhou, et al.
Published: (2025)
by: Zhang, Liuzhou, et al.
Published: (2025)
Aligning Generative Denoising with Discriminative Objectives Unleashes Diffusion for Visual Perception
by: Pang, Ziqi, et al.
Published: (2025)
by: Pang, Ziqi, et al.
Published: (2025)
SEDiT: Mask-Free Video Subtitle Erasure via One-step Diffusion Transformer
by: Hui, Zheng, et al.
Published: (2026)
by: Hui, Zheng, et al.
Published: (2026)
Directly Denoising Diffusion Models
by: Zhang, Dan, et al.
Published: (2024)
by: Zhang, Dan, et al.
Published: (2024)
Efficient Autoregressive Video Diffusion with Dummy Head
by: Guo, Hang, et al.
Published: (2026)
by: Guo, Hang, et al.
Published: (2026)
Temporal Residual Guided Diffusion Framework for Event-Driven Video Reconstruction
by: Zhu, Lin, et al.
Published: (2024)
by: Zhu, Lin, et al.
Published: (2024)
DoubleDiffusion: Combining Heat Diffusion with Denoising Diffusion for Texture Generation on 3D Meshes
by: Wang, Xuyang, et al.
Published: (2025)
by: Wang, Xuyang, et al.
Published: (2025)
LVC: A Lightweight Compression Framework for Enhancing VLMs in Long Video Understanding
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
CaRDiff: Video Salient Object Ranking Chain of Thought Reasoning for Saliency Prediction with Diffusion
by: Tang, Yolo Yunlong, et al.
Published: (2024)
by: Tang, Yolo Yunlong, et al.
Published: (2024)
VideoShield: Regulating Diffusion-based Video Generation Models via Watermarking
by: Hu, Runyi, et al.
Published: (2025)
by: Hu, Runyi, et al.
Published: (2025)
Reason2Drive: Towards Interpretable and Chain-based Reasoning for Autonomous Driving
by: Nie, Ming, et al.
Published: (2023)
by: Nie, Ming, et al.
Published: (2023)
Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation
by: Huang, Tianyu, et al.
Published: (2025)
by: Huang, Tianyu, et al.
Published: (2025)
Reversing Skin Cancer Adversarial Examples by Multiscale Diffusive and Denoising Aggregation Mechanism
by: Wang, Yongwei, et al.
Published: (2022)
by: Wang, Yongwei, et al.
Published: (2022)
Helios: Real Real-Time Long Video Generation Model
by: Yuan, Shenghai, et al.
Published: (2026)
by: Yuan, Shenghai, et al.
Published: (2026)
BIVDiff: A Training-Free Framework for General-Purpose Video Synthesis via Bridging Image and Video Diffusion Models
by: Shi, Fengyuan, et al.
Published: (2023)
by: Shi, Fengyuan, et al.
Published: (2023)
Similar Items
-
BrickNet: Graph-Backed Generative Brick Assembly
by: Kulits, Peter, et al.
Published: (2026) -
Stacking Brick by Brick: Aligned Feature Isolation for Incremental Face Forgery Detection
by: Cheng, Jikang, et al.
Published: (2024) -
SlowFocus: Enhancing Fine-grained Temporal Understanding in Video LLM
by: Nie, Ming, et al.
Published: (2026) -
TreeSBA: Tree-Transformer for Self-Supervised Sequential Brick Assembly
by: Guo, Mengqi, et al.
Published: (2024) -
Generating Physically Stable and Buildable Brick Structures from Text
by: Pun, Ava, et al.
Published: (2025)