CPM: Class-conditional Prompting Machine for Audio-visual Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Yuanhong, Wang, Chong, Liu, Yuyuan, Wang, Hu, Carneiro, Gustavo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unraveling Instance Associations: A Closer Look for Audio-Visual Segmentation
von: Chen, Yuanhong, et al.
Veröffentlicht: (2023)
von: Chen, Yuanhong, et al.
Veröffentlicht: (2023)
AuralSAM2: Enabling SAM2 Hear Through Pyramid Audio-Visual Feature Prompting
von: Liu, Yuyuan, et al.
Veröffentlicht: (2025)
von: Liu, Yuyuan, et al.
Veröffentlicht: (2025)
ItTakesTwo: Leveraging Peer Representations for Semi-supervised LiDAR Semantic Segmentation
von: Liu, Yuyuan, et al.
Veröffentlicht: (2024)
von: Liu, Yuyuan, et al.
Veröffentlicht: (2024)
Bridging Generative and Discriminative Noisy-Label Learning via Direction-Agnostic EM Formulation
von: Liu, Fengbei, et al.
Veröffentlicht: (2023)
von: Liu, Fengbei, et al.
Veröffentlicht: (2023)
Translation Consistent Semi-supervised Segmentation for 3D Medical Images
von: Liu, Yuyuan, et al.
Veröffentlicht: (2022)
von: Liu, Yuyuan, et al.
Veröffentlicht: (2022)
Mixture of Gaussian-distributed Prototypes with Generative Modelling for Interpretable and Trustworthy Image Recognition
von: Wang, Chong, et al.
Veröffentlicht: (2023)
von: Wang, Chong, et al.
Veröffentlicht: (2023)
Cross- and Intra-image Prototypical Learning for Multi-label Disease Diagnosis and Interpretation
von: Wang, Chong, et al.
Veröffentlicht: (2024)
von: Wang, Chong, et al.
Veröffentlicht: (2024)
Multi-modal Learning with Missing Modality via Shared-Specific Feature Modelling
von: Wang, Hu, et al.
Veröffentlicht: (2023)
von: Wang, Hu, et al.
Veröffentlicht: (2023)
Leveraging Labelled Data Knowledge: A Cooperative Rectification Learning Network for Semi-supervised 3D Medical Image Segmentation
von: Wang, Yanyan, et al.
Veröffentlicht: (2025)
von: Wang, Yanyan, et al.
Veröffentlicht: (2025)
Meta-Learned Modality-Weighted Knowledge Distillation for Robust Multi-Modal Learning with Missing Data
von: Wang, Hu, et al.
Veröffentlicht: (2024)
von: Wang, Hu, et al.
Veröffentlicht: (2024)
Prompt-and-Transfer: Dynamic Class-aware Enhancement for Few-shot Segmentation
von: Bi, Hanbo, et al.
Veröffentlicht: (2024)
von: Bi, Hanbo, et al.
Veröffentlicht: (2024)
AEON: Adaptive Estimation of Instance-Dependent In-Distribution and Out-of-Distribution Label Noise for Robust Learning
von: Garg, Arpit, et al.
Veröffentlicht: (2025)
von: Garg, Arpit, et al.
Veröffentlicht: (2025)
BRAIxDet: Learning to Detect Malignant Breast Lesion with Incomplete Annotations
von: Chen, Yuanhong, et al.
Veröffentlicht: (2023)
von: Chen, Yuanhong, et al.
Veröffentlicht: (2023)
MiniCPM-V: A GPT-4V Level MLLM on Your Phone
von: Yao, Yuan, et al.
Veröffentlicht: (2024)
von: Yao, Yuan, et al.
Veröffentlicht: (2024)
GroPrompt: Efficient Grounded Prompting and Adaptation for Referring Video Object Segmentation
von: Lin, Ci-Siang, et al.
Veröffentlicht: (2024)
von: Lin, Ci-Siang, et al.
Veröffentlicht: (2024)
Text-guided Visual Prompt DINO for Generic Segmentation
von: Guan, Yuchen, et al.
Veröffentlicht: (2025)
von: Guan, Yuchen, et al.
Veröffentlicht: (2025)
Prompting classes: Exploring the Power of Prompt Class Learning in Weakly Supervised Semantic Segmentation
von: Murugesan, Balamurali, et al.
Veröffentlicht: (2023)
von: Murugesan, Balamurali, et al.
Veröffentlicht: (2023)
Open-Vocabulary SAM: Segment and Recognize Twenty-thousand Classes Interactively
von: Yuan, Haobo, et al.
Veröffentlicht: (2024)
von: Yuan, Haobo, et al.
Veröffentlicht: (2024)
SOUPLE: Enhancing Audio-Visual Localization and Segmentation with Learnable Prompt Contexts
von: Nguyen, Khanh Binh, et al.
Veröffentlicht: (2026)
von: Nguyen, Khanh Binh, et al.
Veröffentlicht: (2026)
PRISM: A Promptable and Robust Interactive Segmentation Model with Visual Prompts
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
Spectral Prompt Tuning:Unveiling Unseen Classes for Zero-Shot Semantic Segmentation
von: Xu, Wenhao, et al.
Veröffentlicht: (2023)
von: Xu, Wenhao, et al.
Veröffentlicht: (2023)
Robust Audio-Visual Segmentation via Audio-Guided Visual Convergent Alignment
von: Liu, Chen, et al.
Veröffentlicht: (2025)
von: Liu, Chen, et al.
Veröffentlicht: (2025)
Progressive Vision-Language Prompt for Multi-Organ Multi-Class Cell Semantic Segmentation with Single Branch
von: Zhang, Qing, et al.
Veröffentlicht: (2024)
von: Zhang, Qing, et al.
Veröffentlicht: (2024)
Temporal Prompting Matters: Rethinking Referring Video Object Segmentation
von: Lin, Ci-Siang, et al.
Veröffentlicht: (2025)
von: Lin, Ci-Siang, et al.
Veröffentlicht: (2025)
Unveiling and Mitigating Bias in Audio Visual Segmentation
von: Sun, Peiwen, et al.
Veröffentlicht: (2024)
von: Sun, Peiwen, et al.
Veröffentlicht: (2024)
Prompting Segmentation with Sound Is Generalizable Audio-Visual Source Localizer
von: Wang, Yaoting, et al.
Veröffentlicht: (2023)
von: Wang, Yaoting, et al.
Veröffentlicht: (2023)
Label Anything: Multi-Class Few-Shot Semantic Segmentation with Visual Prompts
von: De Marinis, Pasquale, et al.
Veröffentlicht: (2024)
von: De Marinis, Pasquale, et al.
Veröffentlicht: (2024)
EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2024)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2024)
Implicit Counterfactual Learning for Audio-Visual Segmentation
von: Zha, Mingfeng, et al.
Veröffentlicht: (2025)
von: Zha, Mingfeng, et al.
Veröffentlicht: (2025)
AV-Unified: A Unified Framework for Audio-visual Scene Understanding
von: Li, Guangyao, et al.
Veröffentlicht: (2026)
von: Li, Guangyao, et al.
Veröffentlicht: (2026)
Class Similarity Transition: Decoupling Class Similarities and Imbalance from Generalized Few-shot Segmentation
von: Wang, Shihong, et al.
Veröffentlicht: (2024)
von: Wang, Shihong, et al.
Veröffentlicht: (2024)
Boosting Audio-visual Zero-shot Learning with Large Language Models
von: Chen, Haoxing, et al.
Veröffentlicht: (2023)
von: Chen, Haoxing, et al.
Veröffentlicht: (2023)
Multi-rater Prompting for Ambiguous Medical Image Segmentation
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
MedCutMix: A Data-Centric Approach to Improve Radiology Vision-Language Pre-training with Disease Awareness
von: Wang, Sinuo, et al.
Veröffentlicht: (2025)
von: Wang, Sinuo, et al.
Veröffentlicht: (2025)
Narrating For You: Prompt-guided Audio-visual Narrating Face Generation Employing Multi-entangled Latent Space
von: Chandra, Aashish, et al.
Veröffentlicht: (2026)
von: Chandra, Aashish, et al.
Veröffentlicht: (2026)
Prompt-Guided Image Editing with Masked Logit Nudging in Visual Autoregressive Models
von: El-Ghoussani, Amir, et al.
Veröffentlicht: (2026)
von: El-Ghoussani, Amir, et al.
Veröffentlicht: (2026)
Detect Anything in Real Time: From Single-Prompt Segmentation to Multi-Class Detection
von: Turkcan, Mehmet Kerem
Veröffentlicht: (2026)
von: Turkcan, Mehmet Kerem
Veröffentlicht: (2026)
RESAnything: Attribute Prompting for Arbitrary Referring Segmentation
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe
von: Yu, Tianyu, et al.
Veröffentlicht: (2025)
von: Yu, Tianyu, et al.
Veröffentlicht: (2025)
Det-SAM2:Technical Report on the Self-Prompting Segmentation Framework Based on Segment Anything Model 2
von: Wang, Zhiting, et al.
Veröffentlicht: (2024)
von: Wang, Zhiting, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Unraveling Instance Associations: A Closer Look for Audio-Visual Segmentation
von: Chen, Yuanhong, et al.
Veröffentlicht: (2023) -
AuralSAM2: Enabling SAM2 Hear Through Pyramid Audio-Visual Feature Prompting
von: Liu, Yuyuan, et al.
Veröffentlicht: (2025) -
ItTakesTwo: Leveraging Peer Representations for Semi-supervised LiDAR Semantic Segmentation
von: Liu, Yuyuan, et al.
Veröffentlicht: (2024) -
Bridging Generative and Discriminative Noisy-Label Learning via Direction-Agnostic EM Formulation
von: Liu, Fengbei, et al.
Veröffentlicht: (2023) -
Translation Consistent Semi-supervised Segmentation for 3D Medical Images
von: Liu, Yuyuan, et al.
Veröffentlicht: (2022)