DeiSAM: Segment Anything with Deictic Prompting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shindo, Hikaru, Brack, Manuel, Sudhakaran, Gopika, Dhami, Devendra Singh, Schramowski, Patrick, Kersting, Kristian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ART: Adaptive Relation Tuning for Generalized Relation Prediction
von: Sudhakaran, Gopika, et al.
Veröffentlicht: (2025)
von: Sudhakaran, Gopika, et al.
Veröffentlicht: (2025)
Learning Differentiable Logic Programs for Abstract Visual Reasoning
von: Shindo, Hikaru, et al.
Veröffentlicht: (2023)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2023)
V-LoL: A Diagnostic Dataset for Visual Logical Learning
von: Helff, Lukas, et al.
Veröffentlicht: (2023)
von: Helff, Lukas, et al.
Veröffentlicht: (2023)
Core Tokensets for Data-efficient Sequential Training of Transformers
von: Paul, Subarnaduti, et al.
Veröffentlicht: (2024)
von: Paul, Subarnaduti, et al.
Veröffentlicht: (2024)
LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and Models
von: Helff, Lukas, et al.
Veröffentlicht: (2024)
von: Helff, Lukas, et al.
Veröffentlicht: (2024)
Exploiting Cultural Biases via Homoglyphs in Text-to-Image Synthesis
von: Struppek, Lukas, et al.
Veröffentlicht: (2022)
von: Struppek, Lukas, et al.
Veröffentlicht: (2022)
How to Train your Text-to-Image Model: Evaluating Design Choices for Synthetic Training Captions
von: Brack, Manuel, et al.
Veröffentlicht: (2025)
von: Brack, Manuel, et al.
Veröffentlicht: (2025)
Pix2Code: Learning to Compose Neural Visual Concepts as Programs
von: Wüst, Antonia, et al.
Veröffentlicht: (2024)
von: Wüst, Antonia, et al.
Veröffentlicht: (2024)
BlendRL: A Framework for Merging Symbolic and Neural Policy Learning
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
LEDITS++: Limitless Image Editing using Text-to-Image Models
von: Brack, Manuel, et al.
Veröffentlicht: (2023)
von: Brack, Manuel, et al.
Veröffentlicht: (2023)
EXPIL: Explanatory Predicate Invention for Learning in Games
von: Sha, Jingyuan, et al.
Veröffentlicht: (2024)
von: Sha, Jingyuan, et al.
Veröffentlicht: (2024)
CycliST: A Video Language Model Benchmark for Reasoning on Cyclical State Transitions
von: Kohaut, Simon, et al.
Veröffentlicht: (2025)
von: Kohaut, Simon, et al.
Veröffentlicht: (2025)
SAM 3: Segment Anything with Concepts
von: Carion, Nicolas, et al.
Veröffentlicht: (2025)
von: Carion, Nicolas, et al.
Veröffentlicht: (2025)
X-SAM: From Segment Anything to Any Segmentation
von: Wang, Hao, et al.
Veröffentlicht: (2025)
von: Wang, Hao, et al.
Veröffentlicht: (2025)
SGP-SAM: Self-Gated Prompting for Transferring 3D Segment Anything Models to Lesion Segmentation
von: Tang, Zixuan, et al.
Veröffentlicht: (2026)
von: Tang, Zixuan, et al.
Veröffentlicht: (2026)
Compress Any Segment Anything Model (SAM)
von: Fan, Juntong, et al.
Veröffentlicht: (2025)
von: Fan, Juntong, et al.
Veröffentlicht: (2025)
The SAM2-to-SAM3 Gap in the Segment Anything Model Family: Why Prompt-Based Expertise Fails in Concept-Driven Image Segmentation
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
Synthesizing Visual Concepts as Vision-Language Programs
von: Wüst, Antonia, et al.
Veröffentlicht: (2025)
von: Wüst, Antonia, et al.
Veröffentlicht: (2025)
SparseSAM: Structured Sparsification of Activations in Segment Anything Models
von: Tran, Hoai-Chau, et al.
Veröffentlicht: (2026)
von: Tran, Hoai-Chau, et al.
Veröffentlicht: (2026)
MedSAM3: Delving into Segment Anything with Medical Concepts
von: Liu, Anglin, et al.
Veröffentlicht: (2025)
von: Liu, Anglin, et al.
Veröffentlicht: (2025)
CamSAM2: Segment Anything Accurately in Camouflaged Videos
von: Zhou, Yuli, et al.
Veröffentlicht: (2025)
von: Zhou, Yuli, et al.
Veröffentlicht: (2025)
Med-PerSAM: One-Shot Visual Prompt Tuning for Personalized Segment Anything Model in Medical Domain
von: Yoon, Hangyul, et al.
Veröffentlicht: (2024)
von: Yoon, Hangyul, et al.
Veröffentlicht: (2024)
SAM 2: Segment Anything in Images and Videos
von: Ravi, Nikhila, et al.
Veröffentlicht: (2024)
von: Ravi, Nikhila, et al.
Veröffentlicht: (2024)
Does CLIP Know My Face?
von: Hintersdorf, Dominik, et al.
Veröffentlicht: (2022)
von: Hintersdorf, Dominik, et al.
Veröffentlicht: (2022)
Q-SAM2: Accurate Quantization for Segment Anything Model 2
von: Farronato, Nicola, et al.
Veröffentlicht: (2025)
von: Farronato, Nicola, et al.
Veröffentlicht: (2025)
RefSAM: Efficiently Adapting Segmenting Anything Model for Referring Video Object Segmentation
von: Li, Yonglin, et al.
Veröffentlicht: (2023)
von: Li, Yonglin, et al.
Veröffentlicht: (2023)
SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems
von: Shindo, Hikaru, et al.
Veröffentlicht: (2026)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2026)
Region-Guided Attack on the Segment Anything Model (SAM)
von: Liu, Xiaoliang, et al.
Veröffentlicht: (2024)
von: Liu, Xiaoliang, et al.
Veröffentlicht: (2024)
WeakSAM: Segment Anything Meets Weakly-supervised Instance-level Recognition
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
DIAGen: Semantically Diverse Image Augmentation with Generative Models for Few-Shot Learning
von: Lingenberg, Tobias, et al.
Veröffentlicht: (2024)
von: Lingenberg, Tobias, et al.
Veröffentlicht: (2024)
STORM: Segment, Track, and Object Re-Localization from a Single Image
von: Deng, Yu, et al.
Veröffentlicht: (2025)
von: Deng, Yu, et al.
Veröffentlicht: (2025)
SAMAug: Point Prompt Augmentation for Segment Anything Model
von: Dai, Haixing, et al.
Veröffentlicht: (2023)
von: Dai, Haixing, et al.
Veröffentlicht: (2023)
AMA-SAM: Adversarial Multi-Domain Alignment of Segment Anything Model for High-Fidelity Histology Nuclei Segmentation
von: Qian, Jiahe, et al.
Veröffentlicht: (2025)
von: Qian, Jiahe, et al.
Veröffentlicht: (2025)
RobustSAM: Segment Anything Robustly on Degraded Images
von: Chen, Wei-Ting, et al.
Veröffentlicht: (2024)
von: Chen, Wei-Ting, et al.
Veröffentlicht: (2024)
GeoSAM: Fine-tuning SAM with Multi-Modal Prompts for Mobility Infrastructure Segmentation
von: Sultan, Rafi Ibn, et al.
Veröffentlicht: (2023)
von: Sultan, Rafi Ibn, et al.
Veröffentlicht: (2023)
AutoProSAM: Automated Prompting SAM for 3D Multi-Organ Segmentation
von: Li, Chengyin, et al.
Veröffentlicht: (2023)
von: Li, Chengyin, et al.
Veröffentlicht: (2023)
AoP-SAM: Automation of Prompts for Efficient Segmentation
von: Chen, Yi, et al.
Veröffentlicht: (2025)
von: Chen, Yi, et al.
Veröffentlicht: (2025)
ActivationReasoning: Logical Reasoning in Latent Activation Spaces
von: Helff, Lukas, et al.
Veröffentlicht: (2025)
von: Helff, Lukas, et al.
Veröffentlicht: (2025)
Eye image segmentation using visual and concept prompts with Segment Anything Model 3 (SAM3)
von: Niehorster, Diederick C., et al.
Veröffentlicht: (2026)
von: Niehorster, Diederick C., et al.
Veröffentlicht: (2026)
BALD-SAM: Disagreement-based Active Prompting in Interactive Segmentation
von: Chowdhury, Prithwijit, et al.
Veröffentlicht: (2026)
von: Chowdhury, Prithwijit, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ART: Adaptive Relation Tuning for Generalized Relation Prediction
von: Sudhakaran, Gopika, et al.
Veröffentlicht: (2025) -
Learning Differentiable Logic Programs for Abstract Visual Reasoning
von: Shindo, Hikaru, et al.
Veröffentlicht: (2023) -
V-LoL: A Diagnostic Dataset for Visual Logical Learning
von: Helff, Lukas, et al.
Veröffentlicht: (2023) -
Core Tokensets for Data-efficient Sequential Training of Transformers
von: Paul, Subarnaduti, et al.
Veröffentlicht: (2024) -
LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and Models
von: Helff, Lukas, et al.
Veröffentlicht: (2024)