Enhancing Video Inpainting with Aligned Frame Interval Guidance
Fuente:
arXiv
Salvato in:
| Autori principali: | Xie, Ming, Yu, Junqiu, Dong, Qiaole, Xue, Xiangyang, Fu, Yanwei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Aligned Stable Inpainting: Mitigating Unwanted Object Insertion and Preserving Color Consistency
di: Wang, Yikai, et al.
Pubblicazione: (2026)
di: Wang, Yikai, et al.
Pubblicazione: (2026)
Towards Enhanced Image Inpainting: Mitigating Unwanted Object Insertion and Preserving Color Consistency
di: Wang, Yikai, et al.
Pubblicazione: (2023)
di: Wang, Yikai, et al.
Pubblicazione: (2023)
Online Dense Point Tracking with Streaming Memory
di: Dong, Qiaole, et al.
Pubblicazione: (2025)
di: Dong, Qiaole, et al.
Pubblicazione: (2025)
MemFlow: Optical Flow Estimation and Prediction with Memory
di: Dong, Qiaole, et al.
Pubblicazione: (2024)
di: Dong, Qiaole, et al.
Pubblicazione: (2024)
EgoSound: Benchmarking Sound Understanding in Egocentric Videos
di: Zhu, Bingwen, et al.
Pubblicazione: (2026)
di: Zhu, Bingwen, et al.
Pubblicazione: (2026)
Repositioning the Subject within Image
di: Wang, Yikai, et al.
Pubblicazione: (2024)
di: Wang, Yikai, et al.
Pubblicazione: (2024)
MVInpainter: Learning Multi-View Consistent Inpainting to Bridge 2D and 3D Editing
di: Cao, Chenjie, et al.
Pubblicazione: (2024)
di: Cao, Chenjie, et al.
Pubblicazione: (2024)
Exploiting Optical Flow Guidance for Transformer-Based Video Inpainting
di: Zhang, Kaidong, et al.
Pubblicazione: (2023)
di: Zhang, Kaidong, et al.
Pubblicazione: (2023)
LeftRefill: Filling Right Canvas based on Left Reference through Generalized Text-to-Image Diffusion Model
di: Cao, Chenjie, et al.
Pubblicazione: (2023)
di: Cao, Chenjie, et al.
Pubblicazione: (2023)
AnyRefill: A Unified, Data-Efficient Framework for Left-Prompt-Guided Vision Tasks
di: Xie, Ming, et al.
Pubblicazione: (2025)
di: Xie, Ming, et al.
Pubblicazione: (2025)
Uni3C: Unifying Precisely 3D-Enhanced Camera and Human Motion Controls for Video Generation
di: Cao, Chenjie, et al.
Pubblicazione: (2025)
di: Cao, Chenjie, et al.
Pubblicazione: (2025)
MVGenMaster: Scaling Multi-View Generation from Any Image via 3D Priors Enhanced Diffusion Model
di: Cao, Chenjie, et al.
Pubblicazione: (2024)
di: Cao, Chenjie, et al.
Pubblicazione: (2024)
Polaris: Open-ended Interactive Robotic Manipulation via Syn2Real Visual Grounding and Large Language Models
di: Wang, Tianyu, et al.
Pubblicazione: (2024)
di: Wang, Tianyu, et al.
Pubblicazione: (2024)
Improving Neural Surface Reconstruction with Feature Priors from Multi-View Image
di: Ren, Xinlin, et al.
Pubblicazione: (2024)
di: Ren, Xinlin, et al.
Pubblicazione: (2024)
EDEN: Enhanced Diffusion for High-quality Large-motion Video Frame Interpolation
di: Zhang, Zihao, et al.
Pubblicazione: (2025)
di: Zhang, Zihao, et al.
Pubblicazione: (2025)
DecoFuse: Decomposing and Fusing the "What", "Where", and "How" for Brain-Inspired fMRI-to-Video Decoding
di: Li, Chong, et al.
Pubblicazione: (2025)
di: Li, Chong, et al.
Pubblicazione: (2025)
CorrFill: Enhancing Faithfulness in Reference-based Inpainting with Correspondence Guidance in Diffusion Models
di: Liu, Kuan-Hung, et al.
Pubblicazione: (2025)
di: Liu, Kuan-Hung, et al.
Pubblicazione: (2025)
Beyond 'Templates': Category-Agnostic Object Pose, Size, and Shape Estimation from a Single View
di: Zhang, Jinyu, et al.
Pubblicazione: (2025)
di: Zhang, Jinyu, et al.
Pubblicazione: (2025)
Image-Text-Image Knowledge Transfer for Lifelong Person Re-Identification with Hybrid Clothing States
di: Wang, Qizao, et al.
Pubblicazione: (2024)
di: Wang, Qizao, et al.
Pubblicazione: (2024)
Exploring Fine-Grained Representation and Recomposition for Cloth-Changing Person Re-Identification
di: Wang, Qizao, et al.
Pubblicazione: (2023)
di: Wang, Qizao, et al.
Pubblicazione: (2023)
Deep Learning-based Image and Video Inpainting: A Survey
di: Quan, Weize, et al.
Pubblicazione: (2024)
di: Quan, Weize, et al.
Pubblicazione: (2024)
PPMStereo: Pick-and-Play Memory Construction for Consistent Dynamic Stereo Matching
di: Wang, Yun, et al.
Pubblicazione: (2025)
di: Wang, Yun, et al.
Pubblicazione: (2025)
Raformer: Redundancy-Aware Transformer for Video Wire Inpainting
di: Ji, Zhong, et al.
Pubblicazione: (2024)
di: Ji, Zhong, et al.
Pubblicazione: (2024)
Object-Aware Video Matting with Cross-Frame Guidance
di: Zhang, Huayu, et al.
Pubblicazione: (2025)
di: Zhang, Huayu, et al.
Pubblicazione: (2025)
VIP: Video Inpainting Pipeline for Real World Human Removal
di: Sun, Huiming, et al.
Pubblicazione: (2025)
di: Sun, Huiming, et al.
Pubblicazione: (2025)
RAG-6DPose: Retrieval-Augmented 6D Pose Estimation via Leveraging CAD as Knowledge Base
di: Wang, Kuanning, et al.
Pubblicazione: (2025)
di: Wang, Kuanning, et al.
Pubblicazione: (2025)
Frame Guidance: Training-Free Guidance for Frame-Level Control in Video Diffusion Models
di: Jang, Sangwon, et al.
Pubblicazione: (2025)
di: Jang, Sangwon, et al.
Pubblicazione: (2025)
DiffuEraser: A Diffusion Model for Video Inpainting
di: Li, Xiaowen, et al.
Pubblicazione: (2025)
di: Li, Xiaowen, et al.
Pubblicazione: (2025)
Content and Salient Semantics Collaboration for Cloth-Changing Person Re-Identification
di: Wang, Qizao, et al.
Pubblicazione: (2024)
di: Wang, Qizao, et al.
Pubblicazione: (2024)
SparseGrasp: Robotic Grasping via 3D Semantic Gaussian Splatting from Sparse Multi-View RGB Images
di: Yu, Junqiu, et al.
Pubblicazione: (2024)
di: Yu, Junqiu, et al.
Pubblicazione: (2024)
Flow-Guided Diffusion for Video Inpainting
di: Gu, Bohai, et al.
Pubblicazione: (2023)
di: Gu, Bohai, et al.
Pubblicazione: (2023)
MTV-Inpaint: Multi-Task Long Video Inpainting
di: Yang, Shiyuan, et al.
Pubblicazione: (2025)
di: Yang, Shiyuan, et al.
Pubblicazione: (2025)
You Only Estimate Once: Unified, One-stage, Real-Time Category-level Articulated Object 6D Pose Estimation for Robotic Grasping
di: Huang, Jingshun, et al.
Pubblicazione: (2025)
di: Huang, Jingshun, et al.
Pubblicazione: (2025)
DocCogito: Aligning Layout Cognition and Step-Level Grounded Reasoning for Document Understanding
di: Wu, Yuchuan, et al.
Pubblicazione: (2026)
di: Wu, Yuchuan, et al.
Pubblicazione: (2026)
Distribution Aligned Semantics Adaption for Lifelong Person Re-Identification
di: Wang, Qizao, et al.
Pubblicazione: (2024)
di: Wang, Qizao, et al.
Pubblicazione: (2024)
FFP-300K: Scaling First-Frame Propagation for Generalizable Video Editing
di: Huang, Xijie, et al.
Pubblicazione: (2026)
di: Huang, Xijie, et al.
Pubblicazione: (2026)
TRACE: Object Motion Editing in Videos with First-Frame Trajectory Guidance
di: Phung, Quynh, et al.
Pubblicazione: (2026)
di: Phung, Quynh, et al.
Pubblicazione: (2026)
E-Commerce Inpainting with Mask Guidance in Controlnet for Reducing Overcompletion
di: Li, Guandong
Pubblicazione: (2024)
di: Li, Guandong
Pubblicazione: (2024)
CAP-Net: A Unified Network for 6D Pose and Size Estimation of Categorical Articulated Parts from a Single RGB-D Image
di: Huang, Jingshun, et al.
Pubblicazione: (2025)
di: Huang, Jingshun, et al.
Pubblicazione: (2025)
Shot-Aware Frame Sampling for Video Understanding
di: Zhao, Mengyu, et al.
Pubblicazione: (2026)
di: Zhao, Mengyu, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Aligned Stable Inpainting: Mitigating Unwanted Object Insertion and Preserving Color Consistency
di: Wang, Yikai, et al.
Pubblicazione: (2026) -
Towards Enhanced Image Inpainting: Mitigating Unwanted Object Insertion and Preserving Color Consistency
di: Wang, Yikai, et al.
Pubblicazione: (2023) -
Online Dense Point Tracking with Streaming Memory
di: Dong, Qiaole, et al.
Pubblicazione: (2025) -
MemFlow: Optical Flow Estimation and Prediction with Memory
di: Dong, Qiaole, et al.
Pubblicazione: (2024) -
EgoSound: Benchmarking Sound Understanding in Egocentric Videos
di: Zhu, Bingwen, et al.
Pubblicazione: (2026)