Where's Waldo: Diffusion Features for Personalized Segmentation and Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Samuel, Dvir, Ben-Ari, Rami, Levy, Matan, Darshan, Nir, Chechik, Gal |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OmnimatteZero: Fast Training-free Omnimatte with Pre-trained Video Diffusion Models
by: Samuel, Dvir, et al.
Published: (2025)
by: Samuel, Dvir, et al.
Published: (2025)
Fast Autoregressive Video Diffusion and World Models with Temporal Cache Compression and Sparse Attention
by: Samuel, Dvir, et al.
Published: (2026)
by: Samuel, Dvir, et al.
Published: (2026)
Task-Specific Adaptation with Restricted Model Access
by: Levy, Matan, et al.
Published: (2025)
by: Levy, Matan, et al.
Published: (2025)
Find your Needle: Small Object Image Retrieval via Multi-Object Attention Optimization
by: Green, Michael, et al.
Published: (2025)
by: Green, Michael, et al.
Published: (2025)
VidVec: Unlocking Video MLLM Embeddings for Video-Text Retrieval
by: Tzachor, Issar, et al.
Published: (2026)
by: Tzachor, Issar, et al.
Published: (2026)
Lightning-Fast Image Inversion and Editing for Text-to-Image Diffusion Models
by: Samuel, Dvir, et al.
Published: (2023)
by: Samuel, Dvir, et al.
Published: (2023)
EffoVPR: Effective Foundation Model Utilization for Visual Place Recognition
by: Tzachor, Issar, et al.
Published: (2024)
by: Tzachor, Issar, et al.
Published: (2024)
Retrieval-Augmented Gaussian Avatars: Improving Expression Generalization
by: Levy, Matan, et al.
Published: (2026)
by: Levy, Matan, et al.
Published: (2026)
Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models
by: Tewel, Yoad, et al.
Published: (2024)
by: Tewel, Yoad, et al.
Published: (2024)
Key-Locked Rank One Editing for Text-to-Image Personalization
by: Tewel, Yoad, et al.
Published: (2023)
by: Tewel, Yoad, et al.
Published: (2023)
Active Learning via Classifier Impact and Greedy Selection for Interactive Image Retrieval
by: Bar, Leah, et al.
Published: (2024)
by: Bar, Leah, et al.
Published: (2024)
Per-Query Visual Concept Learning
by: Malca, Ori, et al.
Published: (2025)
by: Malca, Ori, et al.
Published: (2025)
Single Image Iterative Subject-driven Generation and Editing
by: Shpitzer, Yair, et al.
Published: (2025)
by: Shpitzer, Yair, et al.
Published: (2025)
CarGait: Cross-Attention based Re-ranking for Gait recognition
by: Habib, Gavriel, et al.
Published: (2025)
by: Habib, Gavriel, et al.
Published: (2025)
Bringing Objects to Life: training-free 4D generation from 3D objects through view consistent noise
by: Rahamim, Ohad, et al.
Published: (2024)
by: Rahamim, Ohad, et al.
Published: (2024)
Fast 4D Mesh Generation by Spatio-Temporal Attention Chains
by: Samuel, Dvir, et al.
Published: (2026)
by: Samuel, Dvir, et al.
Published: (2026)
Policy Optimized Text-to-Image Pipeline Design
by: Gadot, Uri, et al.
Published: (2025)
by: Gadot, Uri, et al.
Published: (2025)
Assessing Image Quality Using a Simple Generative Representation
by: Raviv, Simon, et al.
Published: (2024)
by: Raviv, Simon, et al.
Published: (2024)
Abnormal Event Detection In Videos Using Deep Embedding
by: Venkatrayappa, Darshan
Published: (2024)
by: Venkatrayappa, Darshan
Published: (2024)
Make It Count: Text-to-Image Generation with an Accurate Number of Objects
by: Binyamin, Lital, et al.
Published: (2024)
by: Binyamin, Lital, et al.
Published: (2024)
Personalized Federated Segmentation with Shared Feature Aggregation and Boundary-Focused Calibration
by: Tashdeed, Ishmam, et al.
Published: (2025)
by: Tashdeed, Ishmam, et al.
Published: (2025)
Diffusion Features to Bridge Domain Gap for Semantic Segmentation
by: Ji, Yuxiang, et al.
Published: (2024)
by: Ji, Yuxiang, et al.
Published: (2024)
Training-Free Consistent Text-to-Image Generation
by: Tewel, Yoad, et al.
Published: (2024)
by: Tewel, Yoad, et al.
Published: (2024)
Look Where It Matters: High-Resolution Crops Retrieval for Efficient VLMs
by: Shabtay, Nimrod, et al.
Published: (2026)
by: Shabtay, Nimrod, et al.
Published: (2026)
On the Robustness of Diffusion-Based Image Compression to Bit-Flip Errors
by: Vaisman, Amit, et al.
Published: (2026)
by: Vaisman, Amit, et al.
Published: (2026)
pFLFE: Cross-silo Personalized Federated Learning via Feature Enhancement on Medical Image Segmentation
by: Xie, Luyuan, et al.
Published: (2024)
by: Xie, Luyuan, et al.
Published: (2024)
Where It Moves, It Matters: Referring Surgical Instrument Segmentation via Motion
by: Wei, Meng, et al.
Published: (2026)
by: Wei, Meng, et al.
Published: (2026)
Clustering via Self-Supervised Diffusion
by: Uziel, Roy, et al.
Published: (2025)
by: Uziel, Roy, et al.
Published: (2025)
Spanning the Visual Analogy Space with a Weight Basis of LoRAs
by: Manor, Hila, et al.
Published: (2026)
by: Manor, Hila, et al.
Published: (2026)
Where are we with calibration under dataset shift in image classification?
by: Roschewitz, Mélanie, et al.
Published: (2025)
by: Roschewitz, Mélanie, et al.
Published: (2025)
Personalized Feature Translation for Expression Recognition: An Efficient Source-Free Domain Adaptation Method
by: Sharafi, Masoumeh, et al.
Published: (2025)
by: Sharafi, Masoumeh, et al.
Published: (2025)
Text-based Aerial-Ground Person Retrieval
by: Zhou, Xinyu, et al.
Published: (2025)
by: Zhou, Xinyu, et al.
Published: (2025)
TextureSAM: Towards a Texture Aware Foundation Model for Segmentation
by: Cohen, Inbal, et al.
Published: (2025)
by: Cohen, Inbal, et al.
Published: (2025)
Finer-Personalization Rank: Fine-Grained Retrieval Examines Identity Preservation for Personalized Generation
by: Kilrain, Connor, et al.
Published: (2025)
by: Kilrain, Connor, et al.
Published: (2025)
Learner Attentiveness and Engagement Analysis in Online Education Using Computer Vision
by: Gogawale, Sharva, et al.
Published: (2024)
by: Gogawale, Sharva, et al.
Published: (2024)
Novel Extraction of Discriminative Fine-Grained Feature to Improve Retinal Vessel Segmentation
by: Zeng, Shuang, et al.
Published: (2025)
by: Zeng, Shuang, et al.
Published: (2025)
Story2Board: A Training-Free Approach for Expressive Storyboard Generation
by: Dinkevich, David, et al.
Published: (2025)
by: Dinkevich, David, et al.
Published: (2025)
Learning to See Inside Opaque Liquid Containers using Speckle Vibrometry
by: Kichler, Matan, et al.
Published: (2025)
by: Kichler, Matan, et al.
Published: (2025)
RadiomicsRetrieval: A Customizable Framework for Medical Image Retrieval Using Radiomics Features
by: Na, Inye, et al.
Published: (2025)
by: Na, Inye, et al.
Published: (2025)
GEA: Generation-Enhanced Alignment for Text-to-Image Person Retrieval
by: Zou, Hao, et al.
Published: (2025)
by: Zou, Hao, et al.
Published: (2025)
Similar Items
-
OmnimatteZero: Fast Training-free Omnimatte with Pre-trained Video Diffusion Models
by: Samuel, Dvir, et al.
Published: (2025) -
Fast Autoregressive Video Diffusion and World Models with Temporal Cache Compression and Sparse Attention
by: Samuel, Dvir, et al.
Published: (2026) -
Task-Specific Adaptation with Restricted Model Access
by: Levy, Matan, et al.
Published: (2025) -
Find your Needle: Small Object Image Retrieval via Multi-Object Attention Optimization
by: Green, Michael, et al.
Published: (2025) -
VidVec: Unlocking Video MLLM Embeddings for Video-Text Retrieval
by: Tzachor, Issar, et al.
Published: (2026)