Studying Image Diffusion Features for Zero-Shot Video Object Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Delatolas, Thanos, Kalogeiton, Vicky, Papadopoulos, Dim P. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HiddenObjects: Scalable Diffusion-Distilled Spatial Priors for Object Placement
by: Schouten, Marco, et al.
Published: (2026)
by: Schouten, Marco, et al.
Published: (2026)
Boosting Unsupervised Video Instance Segmentation with Automatic Quality-Guided Self-Training
by: Lu, Kaixuan, et al.
Published: (2025)
by: Lu, Kaixuan, et al.
Published: (2025)
AutoQ-VIS: Improving Unsupervised Video Instance Segmentation via Automatic Quality Assessment
by: Lu, Kaixuan, et al.
Published: (2025)
by: Lu, Kaixuan, et al.
Published: (2025)
Visual Autoregressive Models Beat Diffusion Models on Inference Time Scaling
by: Riise, Erik, et al.
Published: (2025)
by: Riise, Erik, et al.
Published: (2025)
Towards High-Quality Image Segmentation: Improving Topology Accuracy by Penalizing Neighbor Pixels
by: Valverde, Juan Miguel, et al.
Published: (2026)
by: Valverde, Juan Miguel, et al.
Published: (2026)
POEM: Precise Object-level Editing via MLLM control
by: Schouten, Marco, et al.
Published: (2025)
by: Schouten, Marco, et al.
Published: (2025)
Track Anything Behind Everything: Zero-Shot Amodal Video Object Segmentation
by: Hudson, Finlay G. C., et al.
Published: (2024)
by: Hudson, Finlay G. C., et al.
Published: (2024)
MUSE: Manipulating Unified Framework for Synthesizing Emotions in Images via Test-Time Optimization
by: Xia, Yingjie, et al.
Published: (2025)
by: Xia, Yingjie, et al.
Published: (2025)
Long Story Short: Story-level Video Understanding from 20K Short Films
by: Ghermi, Ridouane, et al.
Published: (2024)
by: Ghermi, Ridouane, et al.
Published: (2024)
AgentRVOS: Reasoning over Object Tracks for Zero-Shot Referring Video Object Segmentation
by: Jin, Woojeong, et al.
Published: (2026)
by: Jin, Woojeong, et al.
Published: (2026)
Motion-Zero: Zero-Shot Moving Object Control Framework for Diffusion-Based Video Generation
by: Chen, Changgu, et al.
Published: (2024)
by: Chen, Changgu, et al.
Published: (2024)
SF20K Competition 2025: Summary and findings
by: Ghermi, Ridouane, et al.
Published: (2026)
by: Ghermi, Ridouane, et al.
Published: (2026)
Diffusion Reinforcement Learning via Centered Reward Distillation
by: Zhu, Yuanzhi, et al.
Published: (2026)
by: Zhu, Yuanzhi, et al.
Published: (2026)
Zero-Shot Video Semantic Segmentation based on Pre-Trained Diffusion Models
by: Wang, Qian, et al.
Published: (2024)
by: Wang, Qian, et al.
Published: (2024)
One-step Diffusion Models with Bregman Density Ratio Matching
by: Zhu, Yuanzhi, et al.
Published: (2025)
by: Zhu, Yuanzhi, et al.
Published: (2025)
Visual Context-Aware Person Fall Detection
by: Nagaj, Aleksander, et al.
Published: (2024)
by: Nagaj, Aleksander, et al.
Published: (2024)
pix2pockets: Shot Suggestions in 8-Ball Pool from a Single Image in the Wild
by: Schiøtt, Jonas Myhre, et al.
Published: (2025)
by: Schiøtt, Jonas Myhre, et al.
Published: (2025)
Di$\mathtt{[M]}$O: Distilling Masked Diffusion Models into One-step Generator
by: Zhu, Yuanzhi, et al.
Published: (2025)
by: Zhu, Yuanzhi, et al.
Published: (2025)
ZS-VCOS: Zero-Shot Video Camouflaged Object Segmentation By Optical Flow and Open Vocabulary Object Detection
by: Guo, Wenqi, et al.
Published: (2025)
by: Guo, Wenqi, et al.
Published: (2025)
Zero-Shot Video Deraining with Video Diffusion Models
by: Varanka, Tuomas, et al.
Published: (2025)
by: Varanka, Tuomas, et al.
Published: (2025)
What about gravity in video generation? Post-Training Newton's Laws with Verifiable Rewards
by: Le, Minh-Quan, et al.
Published: (2025)
by: Le, Minh-Quan, et al.
Published: (2025)
Bridging Text and Image for Artist Style Transfer via Contrastive Learning
by: Liu, Zhi-Song, et al.
Published: (2024)
by: Liu, Zhi-Song, et al.
Published: (2024)
DreamInsert: Zero-Shot Image-to-Video Object Insertion from A Single Image
by: Zhao, Qi, et al.
Published: (2025)
by: Zhao, Qi, et al.
Published: (2025)
Diffuse, Attend, and Segment: Unsupervised Zero-Shot Segmentation using Stable Diffusion
by: Tian, Junjiao, et al.
Published: (2023)
by: Tian, Junjiao, et al.
Published: (2023)
TI2V-Zero: Zero-Shot Image Conditioning for Text-to-Video Diffusion Models
by: Ni, Haomiao, et al.
Published: (2024)
by: Ni, Haomiao, et al.
Published: (2024)
DiffCut: Catalyzing Zero-Shot Semantic Segmentation with Diffusion Features and Recursive Normalized Cut
by: Couairon, Paul, et al.
Published: (2024)
by: Couairon, Paul, et al.
Published: (2024)
Latent Directions: A Simple Pathway to Bias Mitigation in Generative AI
by: Olmos, Carolina Lopez, et al.
Published: (2024)
by: Olmos, Carolina Lopez, et al.
Published: (2024)
Soft-Di[M]O: Improving One-Step Discrete Image Generation with Soft Embeddings
by: Zhu, Yuanzhi, et al.
Published: (2025)
by: Zhu, Yuanzhi, et al.
Published: (2025)
Collaborating Foundation Models for Domain Generalized Semantic Segmentation
by: Benigmim, Yasser, et al.
Published: (2023)
by: Benigmim, Yasser, et al.
Published: (2023)
Weak Cube R-CNN: Weakly Supervised 3D Detection using only 2D Bounding Boxes
by: Hansen, Andreas Lau, et al.
Published: (2025)
by: Hansen, Andreas Lau, et al.
Published: (2025)
Efficient Test-Time Scaling for Small Vision-Language Models
by: Kaya, Mehmet Onurcan, et al.
Published: (2025)
by: Kaya, Mehmet Onurcan, et al.
Published: (2025)
Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models
by: Wang, Wen, et al.
Published: (2023)
by: Wang, Wen, et al.
Published: (2023)
FunnyNet-W: Multimodal Learning of Funny Moments in Videos in the Wild
by: Liu, Zhi-Song, et al.
Published: (2024)
by: Liu, Zhi-Song, et al.
Published: (2024)
Zero-1-to-A: Zero-Shot One Image to Animatable Head Avatars Using Video Diffusion
by: Zhou, Zhenglin, et al.
Published: (2025)
by: Zhou, Zhenglin, et al.
Published: (2025)
E.T. the Exceptional Trajectories: Text-to-camera-trajectory generation with character awareness
by: Courant, Robin, et al.
Published: (2024)
by: Courant, Robin, et al.
Published: (2024)
Make me an Expert: Distilling from Generalist Black-Box Models into Specialized Models for Semantic Segmentation
by: Benigmim, Yasser, et al.
Published: (2025)
by: Benigmim, Yasser, et al.
Published: (2025)
Around the World in 80 Timesteps: A Generative Approach to Global Visual Geolocation
by: Dufour, Nicolas, et al.
Published: (2024)
by: Dufour, Nicolas, et al.
Published: (2024)
Don't drop your samples! Coherence-aware training benefits Conditional diffusion
by: Dufour, Nicolas, et al.
Published: (2024)
by: Dufour, Nicolas, et al.
Published: (2024)
Zero-Shot Video Restoration and Enhancement with Assistance of Video Diffusion Models
by: Cao, Cong, et al.
Published: (2026)
by: Cao, Cong, et al.
Published: (2026)
Conditional Latent Diffusion Models for Zero-Shot Instance Segmentation
by: Ulmer, Maximilian, et al.
Published: (2025)
by: Ulmer, Maximilian, et al.
Published: (2025)
Similar Items
-
HiddenObjects: Scalable Diffusion-Distilled Spatial Priors for Object Placement
by: Schouten, Marco, et al.
Published: (2026) -
Boosting Unsupervised Video Instance Segmentation with Automatic Quality-Guided Self-Training
by: Lu, Kaixuan, et al.
Published: (2025) -
AutoQ-VIS: Improving Unsupervised Video Instance Segmentation via Automatic Quality Assessment
by: Lu, Kaixuan, et al.
Published: (2025) -
Visual Autoregressive Models Beat Diffusion Models on Inference Time Scaling
by: Riise, Erik, et al.
Published: (2025) -
Towards High-Quality Image Segmentation: Improving Topology Accuracy by Penalizing Neighbor Pixels
by: Valverde, Juan Miguel, et al.
Published: (2026)