Visual Prompting for One-shot Controllable Video Editing without Inversion
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Zhengbo, Zhou, Yuxi, Peng, Duo, Lim, Joo-Hwee, Tu, Zhigang, Soh, De Wen, Foo, Lin Geng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Leveraging Text-to-Image Diffusion Models for Unsupervised Visual Object Tracking
by: Zhang, Zhengbo, et al.
Published: (2026)
by: Zhang, Zhengbo, et al.
Published: (2026)
Frequency-Enhanced Diffusion Models: Curriculum-Guided Semantic Alignment for Zero-Shot Skeleton Action Recognition
by: Zhou, Yuxi, et al.
Published: (2026)
by: Zhou, Yuxi, et al.
Published: (2026)
Instance Temperature Knowledge Distillation
by: Zhang, Zhengbo, et al.
Published: (2024)
by: Zhang, Zhengbo, et al.
Published: (2024)
MEDIC: Zero-shot Music Editing with Disentangled Inversion Control
by: Liu, Huadai, et al.
Published: (2024)
by: Liu, Huadai, et al.
Published: (2024)
Controllable Hand Grasp Generation for HOI and Efficient Evaluation Methods
by: Ishant, et al.
Published: (2025)
by: Ishant, et al.
Published: (2025)
Masked Diffusion Vision-Language Models for Temporal Action Localization
by: Wang, Fengshun, et al.
Published: (2026)
by: Wang, Fengshun, et al.
Published: (2026)
OnlineSplatter: Pose-Free Online 3D Reconstruction for Free-Moving Objects
by: Huang, Mark He, et al.
Published: (2025)
by: Huang, Mark He, et al.
Published: (2025)
FADE: A Dataset for Detecting Falling Objects around Buildings in Video
by: Tu, Zhigang, et al.
Published: (2024)
by: Tu, Zhigang, et al.
Published: (2024)
Diffusion Time-step Curriculum for One Image to 3D Generation
by: Yi, Xuanyu, et al.
Published: (2024)
by: Yi, Xuanyu, et al.
Published: (2024)
UAV-OVO: Out-of-Viewpoint Generalization in UAV Action Recognition
by: Xia, Yu, et al.
Published: (2026)
by: Xia, Yu, et al.
Published: (2026)
Avatar Concept Slider: Controllable Editing of Concepts in 3D Human Avatars
by: Foo, Lin Geng, et al.
Published: (2024)
by: Foo, Lin Geng, et al.
Published: (2024)
Informative Sample Selection Model for Skeleton-based Action Recognition with Limited Training Samples
by: Tu, Zhigang, et al.
Published: (2025)
by: Tu, Zhigang, et al.
Published: (2025)
VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification
by: Zhang, Wanyue, et al.
Published: (2025)
by: Zhang, Wanyue, et al.
Published: (2025)
Zero-shot Face Editing via ID-Attribute Decoupled Inversion
by: Hou, Yang, et al.
Published: (2025)
by: Hou, Yang, et al.
Published: (2025)
Bridging the Intent Gap: Knowledge-Enhanced Visual Generation
by: Cheng, Yi, et al.
Published: (2024)
by: Cheng, Yi, et al.
Published: (2024)
MoCha:End-to-End Video Character Replacement without Structural Guidance
by: Xu, Zhengbo, et al.
Published: (2026)
by: Xu, Zhengbo, et al.
Published: (2026)
ProEdit: Inversion-based Editing From Prompts Done Right
by: Ouyang, Zhi, et al.
Published: (2025)
by: Ouyang, Zhi, et al.
Published: (2025)
An Interlibrary Loan System on the World-Wide Web.
by: Foo, Schubert, et al.
Published: (1998)
by: Foo, Schubert, et al.
Published: (1998)
Adaptive Prototype Model for Attribute-based Multi-label Few-shot Action Recognition
by: Xiao, Juefeng, et al.
Published: (2025)
by: Xiao, Juefeng, et al.
Published: (2025)
MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm
by: Guo, Ziyan, et al.
Published: (2025)
by: Guo, Ziyan, et al.
Published: (2025)
FlowAnchor: Stabilizing the Editing Signal for Inversion-Free Video Editing
by: Chen, Ze, et al.
Published: (2026)
by: Chen, Ze, et al.
Published: (2026)
DreamColour: Controllable Video Colour Editing without Training
by: Utintu, Chaitat, et al.
Published: (2024)
by: Utintu, Chaitat, et al.
Published: (2024)
UPAM: Unified Prompt Attack in Text-to-Image Generation Models Against Both Textual Filters and Visual Checkers
by: Peng, Duo, et al.
Published: (2024)
by: Peng, Duo, et al.
Published: (2024)
StyleDiffusion: Prompt-Embedding Inversion for Text-Based Editing
by: Li, Senmao, et al.
Published: (2023)
by: Li, Senmao, et al.
Published: (2023)
Diff-Tracker: Text-to-Image Diffusion Models are Unsupervised Trackers
by: Zhang, Zhengbo, et al.
Published: (2024)
by: Zhang, Zhengbo, et al.
Published: (2024)
Precise Action-to-Video Generation Through Visual Action Prompts
by: Wang, Yuang, et al.
Published: (2025)
by: Wang, Yuang, et al.
Published: (2025)
One-Bit-Aided Modulo Sampling for DOA Estimation
by: Zhang, Qi, et al.
Published: (2023)
by: Zhang, Qi, et al.
Published: (2023)
GenVideo: One-shot Target-image and Shape Aware Video Editing using T2I Diffusion Models
by: Harsha, Sai Sree, et al.
Published: (2024)
by: Harsha, Sai Sree, et al.
Published: (2024)
Prompt-based Visual Alignment for Zero-shot Policy Transfer
by: Gao, Haihan, et al.
Published: (2024)
by: Gao, Haihan, et al.
Published: (2024)
PathCoT: Chain-of-Thought Prompting for Zero-shot Pathology Visual Reasoning
by: Zhou, Junjie, et al.
Published: (2025)
by: Zhou, Junjie, et al.
Published: (2025)
A Mixed Precision Eigensolver Based on the Jacobi Algorithm
by: Zhou, Zhengbo
Published: (2025)
by: Zhou, Zhengbo
Published: (2025)
DisControlFace: Adding Disentangled Control to Diffusion Autoencoder for One-shot Explicit Facial Image Editing
by: Jia, Haozhe, et al.
Published: (2023)
by: Jia, Haozhe, et al.
Published: (2023)
Unveiling the Tapestry: the Interplay of Generalization and Forgetting in Continual Learning
by: Shi, Zenglin, et al.
Published: (2022)
by: Shi, Zenglin, et al.
Published: (2022)
Transient entanglement generation in driven chiral networks beyond the secular approximation
by: Foo, Yan Xi, et al.
Published: (2026)
by: Foo, Yan Xi, et al.
Published: (2026)
AdaptiveDrag: Semantic-Driven Dragging on Diffusion-Based Image Editing
by: Chen, DuoSheng, et al.
Published: (2024)
by: Chen, DuoSheng, et al.
Published: (2024)
Meta-Prompting for Automating Zero-shot Visual Recognition with LLMs
by: Mirza, M. Jehanzeb, et al.
Published: (2024)
by: Mirza, M. Jehanzeb, et al.
Published: (2024)
VENUS: Visual Editing with Noise Inversion Using Scene Graphs
by: Vo, Thanh-Nhan, et al.
Published: (2026)
by: Vo, Thanh-Nhan, et al.
Published: (2026)
Object-aware Inversion and Reassembly for Image Editing
by: Yang, Zhen, et al.
Published: (2023)
by: Yang, Zhen, et al.
Published: (2023)
One-shot Training for Video Object Segmentation
by: Chen, Baiyu, et al.
Published: (2024)
by: Chen, Baiyu, et al.
Published: (2024)
Identifying Hard Noise in Long-Tailed Sample Distribution
by: Yi, Xuanyu, et al.
Published: (2022)
by: Yi, Xuanyu, et al.
Published: (2022)
Similar Items
-
Leveraging Text-to-Image Diffusion Models for Unsupervised Visual Object Tracking
by: Zhang, Zhengbo, et al.
Published: (2026) -
Frequency-Enhanced Diffusion Models: Curriculum-Guided Semantic Alignment for Zero-Shot Skeleton Action Recognition
by: Zhou, Yuxi, et al.
Published: (2026) -
Instance Temperature Knowledge Distillation
by: Zhang, Zhengbo, et al.
Published: (2024) -
MEDIC: Zero-shot Music Editing with Disentangled Inversion Control
by: Liu, Huadai, et al.
Published: (2024) -
Controllable Hand Grasp Generation for HOI and Efficient Evaluation Methods
by: Ishant, et al.
Published: (2025)