Point2Insert: Video Object Insertion via Sparse Point Guidance
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhou, Yu, Yang, Xiaoyan, Zi, Bojia, Zhang, Lihan, Sun, Ruijie, Zheng, Weishi, Huang, Haibin, Zhang, Chi, Li, Xuelong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
TeleStyle: Content-Preserving Style Transfer in Images and Videos
por: Zhang, Shiwen, et al.
Publicado: (2026)
por: Zhang, Shiwen, et al.
Publicado: (2026)
SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion
por: Chen, Xinyu, et al.
Publicado: (2026)
por: Chen, Xinyu, et al.
Publicado: (2026)
Insert Anything: Image Insertion via In-Context Editing in DiT
por: Song, Wensong, et al.
Publicado: (2025)
por: Song, Wensong, et al.
Publicado: (2025)
FreeInsert: Personalized Object Insertion with Geometric and Style Control
por: Zhang, Yuhong, et al.
Publicado: (2025)
por: Zhang, Yuhong, et al.
Publicado: (2025)
Point-to-Point: Sparse Motion Guidance for Controllable Video Editing
por: Song, Yeji, et al.
Publicado: (2025)
por: Song, Yeji, et al.
Publicado: (2025)
DreamInsert: Zero-Shot Image-to-Video Object Insertion from A Single Image
por: Zhao, Qi, et al.
Publicado: (2025)
por: Zhao, Qi, et al.
Publicado: (2025)
OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models
por: Chen, Jinshu, et al.
Publicado: (2025)
por: Chen, Jinshu, et al.
Publicado: (2025)
QwenStyle: Content-Preserving Style Transfer with Qwen-Image-Edit
por: Zhang, Shiwen, et al.
Publicado: (2026)
por: Zhang, Shiwen, et al.
Publicado: (2026)
InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Video Object Insertion
por: Jin, Hoiyeong, et al.
Publicado: (2025)
por: Jin, Hoiyeong, et al.
Publicado: (2025)
GAP: Gaussianize Any Point Clouds with Text Guidance
por: Zhang, Weiqi, et al.
Publicado: (2025)
por: Zhang, Weiqi, et al.
Publicado: (2025)
TempoMaster: Efficient Long Video Generation via Next-Frame-Rate Prediction
por: Ma, Yukuo, et al.
Publicado: (2025)
por: Ma, Yukuo, et al.
Publicado: (2025)
Multi-Point Positional Insertion Tuning for Small Object Detection
por: Goto, Kanoko, et al.
Publicado: (2024)
por: Goto, Kanoko, et al.
Publicado: (2024)
Promoting SAM for Camouflaged Object Detection via Selective Key Point-based Guidance
por: Liang, Guoying, et al.
Publicado: (2025)
por: Liang, Guoying, et al.
Publicado: (2025)
Anything in Any Scene: Photorealistic Video Object Insertion
por: Bai, Chen, et al.
Publicado: (2024)
por: Bai, Chen, et al.
Publicado: (2024)
Conditional Polarization Guidance for Camouflaged Object Detection
por: Zhang, QIfan, et al.
Publicado: (2026)
por: Zhang, QIfan, et al.
Publicado: (2026)
PointCNN++: Performant Convolution on Native Points
por: Li, Lihan, et al.
Publicado: (2025)
por: Li, Lihan, et al.
Publicado: (2025)
UniModel: A Visual-Only Framework for Unified Multimodal Understanding and Generation
por: Zhang, Chi, et al.
Publicado: (2025)
por: Zhang, Chi, et al.
Publicado: (2025)
PointGS: Point Attention-Aware Sparse View Synthesis with Gaussian Splatting
por: Xiang, Lintao, et al.
Publicado: (2025)
por: Xiang, Lintao, et al.
Publicado: (2025)
DISPLAY: Directable Human-Object Interaction Video Generation via Sparse Motion Guidance and Multi-Task Auxiliary
por: Guan, Jiazhi, et al.
Publicado: (2026)
por: Guan, Jiazhi, et al.
Publicado: (2026)
Refaçade: Editing Object with Given Reference Texture
por: Huang, Youze, et al.
Publicado: (2025)
por: Huang, Youze, et al.
Publicado: (2025)
Point-VOS: Pointing Up Video Object Segmentation
por: Zulfikar, Idil Esen, et al.
Publicado: (2024)
por: Zulfikar, Idil Esen, et al.
Publicado: (2024)
DetVPCC: RoI-based Point Cloud Sequence Compression for 3D Object Detection
por: Yan, Mingxuan, et al.
Publicado: (2025)
por: Yan, Mingxuan, et al.
Publicado: (2025)
ObjectGS: Object-aware Scene Reconstruction and Scene Understanding via Gaussian Splatting
por: Zhu, Ruijie, et al.
Publicado: (2025)
por: Zhu, Ruijie, et al.
Publicado: (2025)
TelePhysics: Physics-Grounded Multi-Object Scene Generation from a Single Image with Real-Time Interaction
por: Zhang, Xin, et al.
Publicado: (2026)
por: Zhang, Xin, et al.
Publicado: (2026)
Efficient Spiking Point Mamba for Point Cloud Analysis
por: Wu, Peixi, et al.
Publicado: (2025)
por: Wu, Peixi, et al.
Publicado: (2025)
MiniMax-Remover: Taming Bad Noise Helps Video Object Removal
por: Zi, Bojia, et al.
Publicado: (2025)
por: Zi, Bojia, et al.
Publicado: (2025)
FreeInsert: Disentangled Text-Guided Object Insertion in 3D Gaussian Scene without Spatial Priors
por: Li, Chenxi, et al.
Publicado: (2025)
por: Li, Chenxi, et al.
Publicado: (2025)
PIGEON: VLM-Driven Object Navigation via Points of Interest Selection
por: Peng, Cheng, et al.
Publicado: (2025)
por: Peng, Cheng, et al.
Publicado: (2025)
CtrlVDiff: Controllable Video Generation via Unified Multimodal Video Diffusion
por: Xi, Dianbing, et al.
Publicado: (2025)
por: Xi, Dianbing, et al.
Publicado: (2025)
Voxel or Pillar: Exploring Efficient Point Cloud Representation for 3D Object Detection
por: Huang, Yuhao, et al.
Publicado: (2023)
por: Huang, Yuhao, et al.
Publicado: (2023)
Controllable Video Object Insertion via Multiview Priors
por: Qi, Xia, et al.
Publicado: (2026)
por: Qi, Xia, et al.
Publicado: (2026)
UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models
por: Jiang, Hong, et al.
Publicado: (2026)
por: Jiang, Hong, et al.
Publicado: (2026)
Seeing What Matters: Visual Preference Policy Optimization for Visual Generation
por: Ni, Ziqi, et al.
Publicado: (2025)
por: Ni, Ziqi, et al.
Publicado: (2025)
EffectErase: Joint Video Object Removal and Insertion for High-Quality Effect Erasing
por: Fu, Yang, et al.
Publicado: (2026)
por: Fu, Yang, et al.
Publicado: (2026)
PointMT: Efficient Point Cloud Analysis with Hybrid MLP-Transformer Architecture
por: Zheng, Qiang, et al.
Publicado: (2024)
por: Zheng, Qiang, et al.
Publicado: (2024)
Scene Adaptive Sparse Transformer for Event-based Object Detection
por: Peng, Yansong, et al.
Publicado: (2024)
por: Peng, Yansong, et al.
Publicado: (2024)
Shape-Prior-Based Point Cloud Completion for Single-Stage Fully Sparse 3D Object Detection
por: Wang, Kaizheng, et al.
Publicado: (2026)
por: Wang, Kaizheng, et al.
Publicado: (2026)
3DKeyAD: High-Resolution 3D Point Cloud Anomaly Detection via Keypoint-Guided Point Clustering
por: Wang, Zi, et al.
Publicado: (2025)
por: Wang, Zi, et al.
Publicado: (2025)
PSReg: Prior-guided Sparse Mixture of Experts for Point Cloud Registration
por: Huang, Xiaoshui, et al.
Publicado: (2025)
por: Huang, Xiaoshui, et al.
Publicado: (2025)
PointGauss: Point Cloud-Guided Multi-Object Segmentation for Gaussian Splatting
por: Sun, Wentao, et al.
Publicado: (2025)
por: Sun, Wentao, et al.
Publicado: (2025)
Ejemplares similares
-
TeleStyle: Content-Preserving Style Transfer in Images and Videos
por: Zhang, Shiwen, et al.
Publicado: (2026) -
SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion
por: Chen, Xinyu, et al.
Publicado: (2026) -
Insert Anything: Image Insertion via In-Context Editing in DiT
por: Song, Wensong, et al.
Publicado: (2025) -
FreeInsert: Personalized Object Insertion with Geometric and Style Control
por: Zhang, Yuhong, et al.
Publicado: (2025) -
Point-to-Point: Sparse Motion Guidance for Controllable Video Editing
por: Song, Yeji, et al.
Publicado: (2025)