Show or Tell? A Benchmark To Evaluate Visual and Textual Prompts in Semantic Segmentation
Fuente:
arXiv
Guardado en:
| Autores principales: | Rosi, Gabriele, Cermelli, Fabio |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PrAda: Few-Shot Visual Adaptation for Text-Prompted Segmentation
por: Rosi, Gabriele, et al.
Publicado: (2026)
por: Rosi, Gabriele, et al.
Publicado: (2026)
The revenge of BiSeNet: Efficient Multi-Task Image Segmentation
por: Rosi, Gabriele, et al.
Publicado: (2024)
por: Rosi, Gabriele, et al.
Publicado: (2024)
PEM: Prototype-based Efficient MaskFormer for Image Segmentation
por: Cavagnero, Niccolò, et al.
Publicado: (2024)
por: Cavagnero, Niccolò, et al.
Publicado: (2024)
Cross-Domain Transfer Learning with CoRTe: Consistent and Reliable Transfer from Black-Box to Lightweight Segmentation Model
por: Cuttano, Claudia, et al.
Publicado: (2024)
por: Cuttano, Claudia, et al.
Publicado: (2024)
SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation
por: Cuttano, Claudia, et al.
Publicado: (2024)
por: Cuttano, Claudia, et al.
Publicado: (2024)
How Do Optical Flow and Textual Prompts Collaborate to Assist in Audio-Visual Semantic Segmentation?
por: Lee, Yujian, et al.
Publicado: (2026)
por: Lee, Yujian, et al.
Publicado: (2026)
Visual Textualization for Image Prompted Object Detection
por: Wu, Yongjian, et al.
Publicado: (2025)
por: Wu, Yongjian, et al.
Publicado: (2025)
Medal S: Spatio-Textual Prompt Model for Medical Segmentation
por: Shi, Pengcheng, et al.
Publicado: (2025)
por: Shi, Pengcheng, et al.
Publicado: (2025)
Diffusion Is Your Friend in Show, Suggest and Tell
por: Hu, Jia Cheng, et al.
Publicado: (2025)
por: Hu, Jia Cheng, et al.
Publicado: (2025)
Visual and Textual Prompts in VLLMs for Enhancing Emotion Recognition
por: Wang, Zhifeng, et al.
Publicado: (2025)
por: Wang, Zhifeng, et al.
Publicado: (2025)
Show and Tell: Visually Explainable Deep Neural Nets via Spatially-Aware Concept Bottleneck Models
por: Benou, Itay, et al.
Publicado: (2025)
por: Benou, Itay, et al.
Publicado: (2025)
What does CLIP know about peeling a banana?
por: Cuttano, Claudia, et al.
Publicado: (2024)
por: Cuttano, Claudia, et al.
Publicado: (2024)
Show, Don't Tell: Morphing Latent Reasoning into Image Generation
por: Chen, Harold Haodong, et al.
Publicado: (2026)
por: Chen, Harold Haodong, et al.
Publicado: (2026)
Can Textual Semantics Mitigate Sounding Object Segmentation Preference?
por: Wang, Yaoting, et al.
Publicado: (2024)
por: Wang, Yaoting, et al.
Publicado: (2024)
CAT: Coordinating Anatomical-Textual Prompts for Multi-Organ and Tumor Segmentation
por: Huang, Zhongzhen, et al.
Publicado: (2024)
por: Huang, Zhongzhen, et al.
Publicado: (2024)
Probing CLIP's Comprehension of 360-Degree Textual and Visual Semantics
por: Wang, Hai, et al.
Publicado: (2026)
por: Wang, Hai, et al.
Publicado: (2026)
VP Lab: a PEFT-Enabled Visual Prompting Laboratory for Semantic Segmentation
por: Avogaro, Niccolo, et al.
Publicado: (2025)
por: Avogaro, Niccolo, et al.
Publicado: (2025)
Label Anything: Multi-Class Few-Shot Semantic Segmentation with Visual Prompts
por: De Marinis, Pasquale, et al.
Publicado: (2024)
por: De Marinis, Pasquale, et al.
Publicado: (2024)
Visual Transformation Telling
por: Cui, Wanqing, et al.
Publicado: (2023)
por: Cui, Wanqing, et al.
Publicado: (2023)
Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization
por: Zhang, Zhiwang, et al.
Publicado: (2025)
por: Zhang, Zhiwang, et al.
Publicado: (2025)
TCP:Textual-based Class-aware Prompt tuning for Visual-Language Model
por: Yao, Hantao, et al.
Publicado: (2023)
por: Yao, Hantao, et al.
Publicado: (2023)
On the Effectiveness of Textual Prompting with Lightweight Fine-Tuning for SAM3 Remote Sensing Segmentation
por: Blushtein-Livnon, Roni, et al.
Publicado: (2025)
por: Blushtein-Livnon, Roni, et al.
Publicado: (2025)
Tell Me What's Next: Textual Foresight for Generic UI Representations
por: Burns, Andrea, et al.
Publicado: (2024)
por: Burns, Andrea, et al.
Publicado: (2024)
DiffPrompter: Differentiable Implicit Visual Prompts for Semantic-Segmentation in Adverse Conditions
por: Kalwar, Sanket, et al.
Publicado: (2023)
por: Kalwar, Sanket, et al.
Publicado: (2023)
PromptTea: Let Prompts Tell TeaCache the Optimal Threshold
por: Huang, Zishen, et al.
Publicado: (2025)
por: Huang, Zishen, et al.
Publicado: (2025)
OmniEval: A Benchmark for Evaluating Omni-modal Models with Visual, Auditory, and Textual Inputs
por: Zhang, Yiman, et al.
Publicado: (2025)
por: Zhang, Yiman, et al.
Publicado: (2025)
BioVITA: Biological Dataset, Model, and Benchmark for Visual-Textual-Acoustic Alignment
por: Shinoda, Risa, et al.
Publicado: (2026)
por: Shinoda, Risa, et al.
Publicado: (2026)
Textualize Visual Prompt for Image Editing via Diffusion Bridge
por: Xu, Pengcheng, et al.
Publicado: (2025)
por: Xu, Pengcheng, et al.
Publicado: (2025)
VP-Hype: A Hybrid Mamba-Transformer Framework with Visual-Textual Prompting for Hyperspectral Image Classification
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026)
Adaptive Prompt Learning with Negative Textual Semantics and Uncertainty Modeling for Universal Multi-Source Domain Adaptation
por: Yang, Yuxiang, et al.
Publicado: (2024)
por: Yang, Yuxiang, et al.
Publicado: (2024)
Alleviating Textual Reliance in Medical Language-guided Segmentation via Prototype-driven Semantic Approximation
por: Ye, Shuchang, et al.
Publicado: (2025)
por: Ye, Shuchang, et al.
Publicado: (2025)
Show or Tell? Effectively prompting Vision-Language Models for semantic segmentation
por: Avogaro, Niccolo, et al.
Publicado: (2025)
por: Avogaro, Niccolo, et al.
Publicado: (2025)
Contrastive Prompt Clustering for Weakly Supervised Semantic Segmentation
por: Wu, Wangyu, et al.
Publicado: (2025)
por: Wu, Wangyu, et al.
Publicado: (2025)
Semantic-aware SAM for Point-Prompted Instance Segmentation
por: Wei, Zhaoyang, et al.
Publicado: (2023)
por: Wei, Zhaoyang, et al.
Publicado: (2023)
Prompt Categories Cluster for Weakly Supervised Semantic Segmentation
por: Wu, Wangyu, et al.
Publicado: (2024)
por: Wu, Wangyu, et al.
Publicado: (2024)
Advancing Textual Prompt Learning with Anchored Attributes
por: Li, Zheng, et al.
Publicado: (2024)
por: Li, Zheng, et al.
Publicado: (2024)
X-Prompt: Multi-modal Visual Prompt for Video Object Segmentation
por: Guo, Pinxue, et al.
Publicado: (2024)
por: Guo, Pinxue, et al.
Publicado: (2024)
Benchmarking Human and Automated Prompting in the Segment Anything Model
por: Quesada, Jorge, et al.
Publicado: (2024)
por: Quesada, Jorge, et al.
Publicado: (2024)
Prompting classes: Exploring the Power of Prompt Class Learning in Weakly Supervised Semantic Segmentation
por: Murugesan, Balamurali, et al.
Publicado: (2023)
por: Murugesan, Balamurali, et al.
Publicado: (2023)
InvSeg: Test-Time Prompt Inversion for Semantic Segmentation
por: Lin, Jiayi, et al.
Publicado: (2024)
por: Lin, Jiayi, et al.
Publicado: (2024)
Ejemplares similares
-
PrAda: Few-Shot Visual Adaptation for Text-Prompted Segmentation
por: Rosi, Gabriele, et al.
Publicado: (2026) -
The revenge of BiSeNet: Efficient Multi-Task Image Segmentation
por: Rosi, Gabriele, et al.
Publicado: (2024) -
PEM: Prototype-based Efficient MaskFormer for Image Segmentation
por: Cavagnero, Niccolò, et al.
Publicado: (2024) -
Cross-Domain Transfer Learning with CoRTe: Consistent and Reliable Transfer from Black-Box to Lightweight Segmentation Model
por: Cuttano, Claudia, et al.
Publicado: (2024) -
SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation
por: Cuttano, Claudia, et al.
Publicado: (2024)