Preserve, Reveal, Expand: Faithful 4D Video Editing with Region-Aware Conditioning
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Zhangchi, Sun, Wenzhang, Yin, Xiangchen, Yuan, Jiahui, Wang, Chunfeng, Li, Hao, Zhan, Kun, Sun, Xiaoyan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DeCo-VAE: Learning Compact Latents for Video Reconstruction via Decoupled Representation
by: Yin, Xiangchen, et al.
Published: (2025)
by: Yin, Xiangchen, et al.
Published: (2025)
MUSE: A Multi-agent Framework for Unconstrained Story Envisioning via Closed-Loop Cognitive Orchestration
by: Sun, Wenzhang, et al.
Published: (2026)
by: Sun, Wenzhang, et al.
Published: (2026)
RiO-DETR: DETR for Real-time Oriented Object Detection
by: Hu, Zhangchi, et al.
Published: (2026)
by: Hu, Zhangchi, et al.
Published: (2026)
DASH: 4D Hash Encoding with Self-Supervised Decomposition for Real-Time Dynamic Scene Rendering
by: Chen, Jie, et al.
Published: (2025)
by: Chen, Jie, et al.
Published: (2025)
DrivingScene: A Multi-Task Online Feed-Forward 3D Gaussian Splatting Method for Dynamic Driving Scenes
by: Hou, Qirui, et al.
Published: (2025)
by: Hou, Qirui, et al.
Published: (2025)
PAGS: Priority-Adaptive Gaussian Splatting for Dynamic Driving Scenes
by: A, Ying, et al.
Published: (2025)
by: A, Ying, et al.
Published: (2025)
UniCP: A Unified Caching and Pruning Framework for Efficient Video Generation
by: Sun, Wenzhang, et al.
Published: (2025)
by: Sun, Wenzhang, et al.
Published: (2025)
Omni-Video 2: Scaling MLLM-Conditioned Diffusion for Unified Video Generation and Editing
by: Yang, Hao, et al.
Published: (2026)
by: Yang, Hao, et al.
Published: (2026)
DirectTryOn: One-Step Virtual Try-On via Straightened Conditional Transport
by: Sun, Xianbing, et al.
Published: (2026)
by: Sun, Xianbing, et al.
Published: (2026)
FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation
by: Wang, Yuanzhi, et al.
Published: (2026)
by: Wang, Yuanzhi, et al.
Published: (2026)
DEFormer: DCT-driven Enhancement Transformer for Low-light Image and Dark Vision
by: Yin, Xiangchen, et al.
Published: (2023)
by: Yin, Xiangchen, et al.
Published: (2023)
MoEE: Mixture of Emotion Experts for Audio-Driven Portrait Animation
by: Liu, Huaize, et al.
Published: (2025)
by: Liu, Huaize, et al.
Published: (2025)
Dome-DETR: DETR with Density-Oriented Feature-Query Manipulation for Efficient Tiny Object Detection
by: Hu, Zhangchi, et al.
Published: (2025)
by: Hu, Zhangchi, et al.
Published: (2025)
Generative Photographic Control for Scene-Consistent Video Cinematic Editing
by: Sun, Huiqiang, et al.
Published: (2025)
by: Sun, Huiqiang, et al.
Published: (2025)
Hi-VAE: Efficient Video Autoencoding with Global and Detailed Motion
by: Liu, Huaize, et al.
Published: (2025)
by: Liu, Huaize, et al.
Published: (2025)
ReLayout: Versatile and Structure-Preserving Design Layout Editing via Relation-Aware Design Reconstruction
by: Lin, Jiawei, et al.
Published: (2026)
by: Lin, Jiawei, et al.
Published: (2026)
Look Before You Fuse: 2D-Guided Cross-Modal Alignment for Robust 3D Detection
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
Decoupling Perception from Reasoning for Hallucination-Resistant Video Understanding
by: Pu, Bowei, et al.
Published: (2025)
by: Pu, Bowei, et al.
Published: (2025)
Event-assisted Low-Light Video Object Segmentation
by: Li, Hebei, et al.
Published: (2024)
by: Li, Hebei, et al.
Published: (2024)
Generative Video Motion Editing with 3D Point Tracks
by: Lee, Yao-Chih, et al.
Published: (2025)
by: Lee, Yao-Chih, et al.
Published: (2025)
Enhancing Low-Cost Video Editing with Lightweight Adaptors and Temporal-Aware Inversion
by: He, Yangfan, et al.
Published: (2025)
by: He, Yangfan, et al.
Published: (2025)
EA-VTR: Event-Aware Video-Text Retrieval
by: Ma, Zongyang, et al.
Published: (2024)
by: Ma, Zongyang, et al.
Published: (2024)
Efficient Spiking Point Mamba for Point Cloud Analysis
by: Wu, Peixi, et al.
Published: (2025)
by: Wu, Peixi, et al.
Published: (2025)
3D-aware Image Generation and Editing with Multi-modal Conditions
by: Li, Bo, et al.
Published: (2024)
by: Li, Bo, et al.
Published: (2024)
Residual-Conditioned Optimal Transport: Towards Structure-Preserving Unpaired and Paired Image Restoration
by: Tang, Xiaole, et al.
Published: (2024)
by: Tang, Xiaole, et al.
Published: (2024)
A Self-supervised Motion Representation for Portrait Video Generation
by: Zhang, Qiyuan, et al.
Published: (2025)
by: Zhang, Qiyuan, et al.
Published: (2025)
3D4D: An Interactive, Editable, 4D World Model via 3D Video Generation
by: He, Yunhong, et al.
Published: (2025)
by: He, Yunhong, et al.
Published: (2025)
HiVid: LLM-Guided Video Saliency For Content-Aware VOD And Live Streaming
by: Chen, Jiahui, et al.
Published: (2026)
by: Chen, Jiahui, et al.
Published: (2026)
DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing
by: Cai, Lingling, et al.
Published: (2025)
by: Cai, Lingling, et al.
Published: (2025)
CAMEO: A Conditional and Quality-Aware Multi-Agent Image Editing Orchestrator
by: Pu, Yuhan, et al.
Published: (2026)
by: Pu, Yuhan, et al.
Published: (2026)
3D Human Pose Estimation Based on 2D-3D Consistency with Synchronized Adversarial Training
by: Deng, Yicheng, et al.
Published: (2021)
by: Deng, Yicheng, et al.
Published: (2021)
In-N-Out: Faithful 3D GAN Inversion with Volumetric Decomposition for Face Editing
by: Xu, Yiran, et al.
Published: (2023)
by: Xu, Yiran, et al.
Published: (2023)
VicaSplat: A Single Run is All You Need for 3D Gaussian Splatting and Camera Estimation from Unposed Video Frames
by: Li, Zhiqi, et al.
Published: (2025)
by: Li, Zhiqi, et al.
Published: (2025)
Physics-Aware 3D Gaussian Editing for Driving Scene Generation
by: Zhou, Feng, et al.
Published: (2026)
by: Zhou, Feng, et al.
Published: (2026)
SDD-4DGS: Static-Dynamic Aware Decoupling in Gaussian Splatting for 4D Scene Reconstruction
by: Sun, Dai, et al.
Published: (2025)
by: Sun, Dai, et al.
Published: (2025)
Towards Source-Aware Object Swapping with Initial Noise Perturbation
by: Zhan, Jiahui, et al.
Published: (2026)
by: Zhan, Jiahui, et al.
Published: (2026)
TimeMachine: Fine-Grained Facial Age Editing with Identity Preservation
by: Mi, Yilin, et al.
Published: (2025)
by: Mi, Yilin, et al.
Published: (2025)
Benchmarking Segmentation Models with Mask-Preserved Attribute Editing
by: Yin, Zijin, et al.
Published: (2024)
by: Yin, Zijin, et al.
Published: (2024)
CamDirector: Towards Long-Term Coherent Video Trajectory Editing
by: Shi, Zhihao, et al.
Published: (2026)
by: Shi, Zhihao, et al.
Published: (2026)
VideoHandles: Editing 3D Object Compositions in Videos Using Video Generative Priors
by: Koo, Juil, et al.
Published: (2025)
by: Koo, Juil, et al.
Published: (2025)
Similar Items
-
DeCo-VAE: Learning Compact Latents for Video Reconstruction via Decoupled Representation
by: Yin, Xiangchen, et al.
Published: (2025) -
MUSE: A Multi-agent Framework for Unconstrained Story Envisioning via Closed-Loop Cognitive Orchestration
by: Sun, Wenzhang, et al.
Published: (2026) -
RiO-DETR: DETR for Real-time Oriented Object Detection
by: Hu, Zhangchi, et al.
Published: (2026) -
DASH: 4D Hash Encoding with Self-Supervised Decomposition for Real-Time Dynamic Scene Rendering
by: Chen, Jie, et al.
Published: (2025) -
DrivingScene: A Multi-Task Online Feed-Forward 3D Gaussian Splatting Method for Dynamic Driving Scenes
by: Hou, Qirui, et al.
Published: (2025)