SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Cuttano, Claudia, Trivigno, Gabriele, Rosi, Gabriele, Masone, Carlo, Averta, Giuseppe |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SANSA: Unleashing the Hidden Semantics in SAM2 for Few-Shot Segmentation
by: Cuttano, Claudia, et al.
Published: (2025)
by: Cuttano, Claudia, et al.
Published: (2025)
What does CLIP know about peeling a banana?
by: Cuttano, Claudia, et al.
Published: (2024)
by: Cuttano, Claudia, et al.
Published: (2024)
MARCO: Navigating the Unseen Space of Semantic Correspondence
by: Cuttano, Claudia, et al.
Published: (2026)
by: Cuttano, Claudia, et al.
Published: (2026)
The revenge of BiSeNet: Efficient Multi-Task Image Segmentation
by: Rosi, Gabriele, et al.
Published: (2024)
by: Rosi, Gabriele, et al.
Published: (2024)
INSID3: Training-Free In-Context Segmentation with DINOv3
by: Cuttano, Claudia, et al.
Published: (2026)
by: Cuttano, Claudia, et al.
Published: (2026)
PrAda: Few-Shot Visual Adaptation for Text-Prompted Segmentation
by: Rosi, Gabriele, et al.
Published: (2026)
by: Rosi, Gabriele, et al.
Published: (2026)
PEM: Prototype-based Efficient MaskFormer for Image Segmentation
by: Cavagnero, Niccolò, et al.
Published: (2024)
by: Cavagnero, Niccolò, et al.
Published: (2024)
JIST: Joint Image and Sequence Training for Sequential Visual Place Recognition
by: Berton, Gabriele, et al.
Published: (2024)
by: Berton, Gabriele, et al.
Published: (2024)
To Match or Not to Match: Revisiting Image Matching for Reliable Visual Place Recognition
by: Sferrazza, Davide, et al.
Published: (2025)
by: Sferrazza, Davide, et al.
Published: (2025)
The Unreasonable Effectiveness of Pre-Trained Features for Camera Pose Refinement
by: Trivigno, Gabriele, et al.
Published: (2024)
by: Trivigno, Gabriele, et al.
Published: (2024)
EarthMatch: Iterative Coregistration for Fine-grained Localization of Astronaut Photography
by: Berton, Gabriele, et al.
Published: (2024)
by: Berton, Gabriele, et al.
Published: (2024)
Collaborative Visual Place Recognition through Federated Learning
by: Dutto, Mattia, et al.
Published: (2024)
by: Dutto, Mattia, et al.
Published: (2024)
Cross-Domain Transfer Learning with CoRTe: Consistent and Reliable Transfer from Black-Box to Lightweight Segmentation Model
by: Cuttano, Claudia, et al.
Published: (2024)
by: Cuttano, Claudia, et al.
Published: (2024)
MegaLoc: One Retrieval to Place Them All
by: Berton, Gabriele, et al.
Published: (2025)
by: Berton, Gabriele, et al.
Published: (2025)
AstroLoc: Robust Space to Ground Image Localizer
by: Berton, Gabriele, et al.
Published: (2025)
by: Berton, Gabriele, et al.
Published: (2025)
All You Need to Know About Training Image Retrieval Models
by: Berton, Gabriele, et al.
Published: (2025)
by: Berton, Gabriele, et al.
Published: (2025)
Show or Tell? A Benchmark To Evaluate Visual and Textual Prompts in Semantic Segmentation
by: Rosi, Gabriele, et al.
Published: (2025)
by: Rosi, Gabriele, et al.
Published: (2025)
Scale-Free Image Keypoints Using Differentiable Persistent Homology
by: Barbarani, Giovanni, et al.
Published: (2024)
by: Barbarani, Giovanni, et al.
Published: (2024)
AMEGO: Active Memory from long EGOcentric videos
by: Goletto, Gabriele, et al.
Published: (2024)
by: Goletto, Gabriele, et al.
Published: (2024)
EarthLoc: Astronaut Photography Localization by Indexing Earth from Space
by: Berton, Gabriele, et al.
Published: (2024)
by: Berton, Gabriele, et al.
Published: (2024)
Road Obstacle Video Segmentation
by: Rai, Shyam Nandan, et al.
Published: (2025)
by: Rai, Shyam Nandan, et al.
Published: (2025)
MeshVPR: Citywide Visual Place Recognition Using 3D Meshes
by: Berton, Gabriele, et al.
Published: (2024)
by: Berton, Gabriele, et al.
Published: (2024)
FS-SAM2: Adapting Segment Anything Model 2 for Few-Shot Semantic Segmentation via Low-Rank Adaptation
by: Forni, Bernardo, et al.
Published: (2025)
by: Forni, Bernardo, et al.
Published: (2025)
Egocentric zone-aware action recognition across environments
by: Peirone, Simone Alberto, et al.
Published: (2024)
by: Peirone, Simone Alberto, et al.
Published: (2024)
Evaluating SAM2 for Video Semantic Segmentation
by: Ariff, Syed Hesham Syed, et al.
Published: (2025)
by: Ariff, Syed Hesham Syed, et al.
Published: (2025)
VideoSAM: Open-World Video Segmentation
by: Guo, Pinxue, et al.
Published: (2024)
by: Guo, Pinxue, et al.
Published: (2024)
Fast SAM2 with Text-Driven Token Pruning
by: Mandal, Avilasha, et al.
Published: (2025)
by: Mandal, Avilasha, et al.
Published: (2025)
Hier-EgoPack: Hierarchical Egocentric Video Understanding with Diverse Task Perspectives
by: Peirone, Simone Alberto, et al.
Published: (2025)
by: Peirone, Simone Alberto, et al.
Published: (2025)
HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos
by: Peirone, Simone Alberto, et al.
Published: (2025)
by: Peirone, Simone Alberto, et al.
Published: (2025)
Biologically-inspired Semi-supervised Semantic Segmentation for Biomedical Imaging
by: Ciampi, Luca, et al.
Published: (2024)
by: Ciampi, Luca, et al.
Published: (2024)
A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives
by: Peirone, Simone Alberto, et al.
Published: (2024)
by: Peirone, Simone Alberto, et al.
Published: (2024)
Biomedical SAM 2: Segment Anything in Biomedical Images and Videos
by: Yan, Zhiling, et al.
Published: (2024)
by: Yan, Zhiling, et al.
Published: (2024)
MirrorSAM2: Segment Mirror in Videos with Depth Perception
by: Xu, Mingchen, et al.
Published: (2025)
by: Xu, Mingchen, et al.
Published: (2025)
Semi-Supervised Biomedical Image Segmentation via Diffusion Models and Teacher-Student Co-Training
by: Ciampi, Luca, et al.
Published: (2025)
by: Ciampi, Luca, et al.
Published: (2025)
SAM2Long: Enhancing SAM 2 for Long Video Segmentation with a Training-Free Memory Tree
by: Ding, Shuangrui, et al.
Published: (2024)
by: Ding, Shuangrui, et al.
Published: (2024)
MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation
by: Rong, Fu, et al.
Published: (2025)
by: Rong, Fu, et al.
Published: (2025)
PanoSAM2: Lightweight Distortion- and Memory-aware Adaptions of SAM2 for 360 Video Object Segmentation
by: Xiao, Dingwen, et al.
Published: (2026)
by: Xiao, Dingwen, et al.
Published: (2026)
Hi-SAM: Marrying Segment Anything Model for Hierarchical Text Segmentation
by: Ye, Maoyuan, et al.
Published: (2024)
by: Ye, Maoyuan, et al.
Published: (2024)
FORESCENE: FOREcasting human activity via latent SCENE graphs diffusion
by: Alliegro, Antonio, et al.
Published: (2025)
by: Alliegro, Antonio, et al.
Published: (2025)
Memory-Augmented SAM2 for Training-Free Surgical Video Segmentation
by: Yin, Ming, et al.
Published: (2025)
by: Yin, Ming, et al.
Published: (2025)
Similar Items
-
SANSA: Unleashing the Hidden Semantics in SAM2 for Few-Shot Segmentation
by: Cuttano, Claudia, et al.
Published: (2025) -
What does CLIP know about peeling a banana?
by: Cuttano, Claudia, et al.
Published: (2024) -
MARCO: Navigating the Unseen Space of Semantic Correspondence
by: Cuttano, Claudia, et al.
Published: (2026) -
The revenge of BiSeNet: Efficient Multi-Task Image Segmentation
by: Rosi, Gabriele, et al.
Published: (2024) -
INSID3: Training-Free In-Context Segmentation with DINOv3
by: Cuttano, Claudia, et al.
Published: (2026)