Point Prompting: Counterfactual Tracking with Video Diffusion Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Shrivastava, Ayush, Mehta, Sanyam, Geng, Daniel, Owens, Andrew |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Self-Supervised Any-Point Tracking by Contrastive Random Walks
por: Shrivastava, Ayush, et al.
Publicado: (2024)
por: Shrivastava, Ayush, et al.
Publicado: (2024)
Self-Supervised Spatial Correspondence Across Modalities
por: Shrivastava, Ayush, et al.
Publicado: (2025)
por: Shrivastava, Ayush, et al.
Publicado: (2025)
Motion Guidance: Diffusion-Based Image Editing with Differentiable Motion Estimators
por: Geng, Daniel, et al.
Publicado: (2024)
por: Geng, Daniel, et al.
Publicado: (2024)
Visual Anagrams: Generating Multi-View Optical Illusions with Diffusion Models
por: Geng, Daniel, et al.
Publicado: (2023)
por: Geng, Daniel, et al.
Publicado: (2023)
Factorized Diffusion: Perceptual Illusions by Noise Decomposition
por: Geng, Daniel, et al.
Publicado: (2024)
por: Geng, Daniel, et al.
Publicado: (2024)
Efficient Continuous Video Flow Model for Video Prediction
por: Shrivastava, Gaurav, et al.
Publicado: (2024)
por: Shrivastava, Gaurav, et al.
Publicado: (2024)
Fine-grained Defocus Blur Control for Generative Image Models
por: Shrivastava, Ayush, et al.
Publicado: (2025)
por: Shrivastava, Ayush, et al.
Publicado: (2025)
Repurposing Video Diffusion Transformers for Robust Point Tracking
por: Son, Soowon, et al.
Publicado: (2025)
por: Son, Soowon, et al.
Publicado: (2025)
InVi: Object Insertion In Videos Using Off-the-Shelf Diffusion Models
por: Saini, Nirat, et al.
Publicado: (2024)
por: Saini, Nirat, et al.
Publicado: (2024)
Generative Animations: A Multi-Model Pipeline for Prompt-Driven Motion Synthesis
por: Khurana, Mannat, et al.
Publicado: (2026)
por: Khurana, Mannat, et al.
Publicado: (2026)
Motion Prompting: Controlling Video Generation with Motion Trajectories
por: Geng, Daniel, et al.
Publicado: (2024)
por: Geng, Daniel, et al.
Publicado: (2024)
Counterfactual World Models via Digital Twin-conditioned Video Diffusion
por: Shen, Yiqing, et al.
Publicado: (2025)
por: Shen, Yiqing, et al.
Publicado: (2025)
Continuous Video Process: Modeling Videos as Continuous Multi-Dimensional Processes for Video Prediction
por: Shrivastava, Gaurav, et al.
Publicado: (2024)
por: Shrivastava, Gaurav, et al.
Publicado: (2024)
InJecteD: Analyzing Trajectories and Drift Dynamics in Denoising Diffusion Probabilistic Models for 2D Point Cloud Generation
por: Jain, Sanyam, et al.
Publicado: (2025)
por: Jain, Sanyam, et al.
Publicado: (2025)
NeRV-Diffusion: Diffuse Implicit Neural Representations for Video Synthesis
por: Ren, Yixuan, et al.
Publicado: (2025)
por: Ren, Yixuan, et al.
Publicado: (2025)
EgoPoints: Advancing Point Tracking for Egocentric Videos
por: Darkhalil, Ahmad, et al.
Publicado: (2024)
por: Darkhalil, Ahmad, et al.
Publicado: (2024)
DeepSeaNet: Improving Underwater Object Detection using EfficientDet
por: Jain, Sanyam
Publicado: (2023)
por: Jain, Sanyam
Publicado: (2023)
Customize-A-Video: One-Shot Motion Customization of Text-to-Video Diffusion Models
por: Ren, Yixuan, et al.
Publicado: (2024)
por: Ren, Yixuan, et al.
Publicado: (2024)
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
por: Jeong, Hyeonho, et al.
Publicado: (2024)
por: Jeong, Hyeonho, et al.
Publicado: (2024)
Counterfactual Stress Testing for Image Classification Models
por: Stammel, Moritz, et al.
Publicado: (2026)
por: Stammel, Moritz, et al.
Publicado: (2026)
Video Decomposition Prior: A Methodology to Decompose Videos into Layers
por: Shrivastava, Gaurav, et al.
Publicado: (2024)
por: Shrivastava, Gaurav, et al.
Publicado: (2024)
Evolutionary Caching to Accelerate Your Off-the-Shelf Diffusion Model
por: Aggarwal, Anirud, et al.
Publicado: (2025)
por: Aggarwal, Anirud, et al.
Publicado: (2025)
Mitigating Hallucinations in Diffusion Models through Adaptive Attention Modulation
por: Oorloff, Trevine, et al.
Publicado: (2025)
por: Oorloff, Trevine, et al.
Publicado: (2025)
What is Point Supervision Worth in Video Instance Segmentation?
por: Huang, Shuaiyi, et al.
Publicado: (2024)
por: Huang, Shuaiyi, et al.
Publicado: (2024)
LD-ViCE: Latent Diffusion Model for Video Counterfactual Explanations
por: Varshney, Payal, et al.
Publicado: (2025)
por: Varshney, Payal, et al.
Publicado: (2025)
Hazedefy: A Lightweight Real-Time Image and Video Dehazing Pipeline for Practical Deployment
por: Bhavsar, Ayush
Publicado: (2025)
por: Bhavsar, Ayush
Publicado: (2025)
ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models
por: Fang, Zixun, et al.
Publicado: (2025)
por: Fang, Zixun, et al.
Publicado: (2025)
CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models
por: Begiristain, León, et al.
Publicado: (2026)
por: Begiristain, León, et al.
Publicado: (2026)
Masked Diffusion Captioning for Visual Feature Learning
por: Feng, Chao, et al.
Publicado: (2025)
por: Feng, Chao, et al.
Publicado: (2025)
Towards Unbiased and Robust Spatio-Temporal Scene Graph Generation and Anticipation
por: Peddi, Rohith, et al.
Publicado: (2024)
por: Peddi, Rohith, et al.
Publicado: (2024)
Causally Steered Diffusion for Automated Video Counterfactual Generation
por: Spyrou, Nikos, et al.
Publicado: (2025)
por: Spyrou, Nikos, et al.
Publicado: (2025)
DiffusionTrack: Diffusion Model For Multi-Object Tracking
por: Luo, Run, et al.
Publicado: (2023)
por: Luo, Run, et al.
Publicado: (2023)
MV-TAP: Tracking Any Point in Multi-View Videos
por: Koo, Jahyeok, et al.
Publicado: (2025)
por: Koo, Jahyeok, et al.
Publicado: (2025)
Generative Video Motion Editing with 3D Point Tracks
por: Lee, Yao-Chih, et al.
Publicado: (2025)
por: Lee, Yao-Chih, et al.
Publicado: (2025)
DiTVR: Zero-Shot Diffusion Transformer for Video Restoration
por: Gao, Sicheng, et al.
Publicado: (2025)
por: Gao, Sicheng, et al.
Publicado: (2025)
Efficient Camera-Controlled Video Generation of Static Scenes via Sparse Diffusion and 3D Rendering
por: Chen, Jieying, et al.
Publicado: (2026)
por: Chen, Jieying, et al.
Publicado: (2026)
Adversarial Attack On Yolov5 For Traffic And Road Sign Detection
por: Jain, Sanyam
Publicado: (2023)
por: Jain, Sanyam
Publicado: (2023)
Prompt-A-Video: Prompt Your Video Diffusion Model via Preference-Aligned LLM
por: Ji, Yatai, et al.
Publicado: (2024)
por: Ji, Yatai, et al.
Publicado: (2024)
From Prompt to Progression: Taming Video Diffusion Models for Seamless Attribute Transition
por: Lo, Ling, et al.
Publicado: (2025)
por: Lo, Ling, et al.
Publicado: (2025)
Frequency-Guided Diffusion Model with Perturbation Training for Skeleton-Based Video Anomaly Detection
por: Tan, Xiaofeng, et al.
Publicado: (2024)
por: Tan, Xiaofeng, et al.
Publicado: (2024)
Ejemplares similares
-
Self-Supervised Any-Point Tracking by Contrastive Random Walks
por: Shrivastava, Ayush, et al.
Publicado: (2024) -
Self-Supervised Spatial Correspondence Across Modalities
por: Shrivastava, Ayush, et al.
Publicado: (2025) -
Motion Guidance: Diffusion-Based Image Editing with Differentiable Motion Estimators
por: Geng, Daniel, et al.
Publicado: (2024) -
Visual Anagrams: Generating Multi-View Optical Illusions with Diffusion Models
por: Geng, Daniel, et al.
Publicado: (2023) -
Factorized Diffusion: Perceptual Illusions by Noise Decomposition
por: Geng, Daniel, et al.
Publicado: (2024)