Masked Multi-Query Slot Attention for Unsupervised Object Discovery
Fuente:
arXiv
Salvato in:
| Autori principali: | Pramanik, Rishav, Villa-Vásquez, José-Fabian, Pedersoli, Marco |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Divided Attention: Unsupervised Multi-Object Discovery with Contextually Separated Slots
di: Lao, Dong, et al.
Pubblicazione: (2023)
di: Lao, Dong, et al.
Pubblicazione: (2023)
Unsupervised Object Discovery: A Comprehensive Survey and Unified Taxonomy
di: Villa-Vásquez, José-Fabian, et al.
Pubblicazione: (2024)
di: Villa-Vásquez, José-Fabian, et al.
Pubblicazione: (2024)
Source-Free Domain Adaptation for YOLO Object Detection
di: Varailhon, Simon, et al.
Pubblicazione: (2024)
di: Varailhon, Simon, et al.
Pubblicazione: (2024)
Adaptive Slot Attention: Object Discovery with Dynamic Slot Number
di: Fan, Ke, et al.
Pubblicazione: (2024)
di: Fan, Ke, et al.
Pubblicazione: (2024)
Visual Modality Prompt for Adapting Vision-Language Object Detectors
di: Medeiros, Heitor R., et al.
Pubblicazione: (2024)
di: Medeiros, Heitor R., et al.
Pubblicazione: (2024)
Object-Centric Learning with Slot Mixture Module
di: Kirilenko, Daniil, et al.
Pubblicazione: (2023)
di: Kirilenko, Daniil, et al.
Pubblicazione: (2023)
SlotLifter: Slot-guided Feature Lifting for Learning Object-centric Radiance Fields
di: Liu, Yu, et al.
Pubblicazione: (2024)
di: Liu, Yu, et al.
Pubblicazione: (2024)
SlotPi: Physics-informed Object-centric Reasoning Models
di: Li, Jian, et al.
Pubblicazione: (2025)
di: Li, Jian, et al.
Pubblicazione: (2025)
Temporally Consistent Object-Centric Learning by Contrasting Slots
di: Manasyan, Anna, et al.
Pubblicazione: (2024)
di: Manasyan, Anna, et al.
Pubblicazione: (2024)
Multi-layer Learnable Attention Mask for Multimodal Tasks
di: Barrios, Wayner, et al.
Pubblicazione: (2024)
di: Barrios, Wayner, et al.
Pubblicazione: (2024)
Attention-based Class-Conditioned Alignment for Multi-Source Domain Adaptation of Object Detectors
di: Belal, Atif, et al.
Pubblicazione: (2024)
di: Belal, Atif, et al.
Pubblicazione: (2024)
Multiple Object Stitching for Unsupervised Representation Learning
di: Shen, Chengchao, et al.
Pubblicazione: (2025)
di: Shen, Chengchao, et al.
Pubblicazione: (2025)
Adversarially Trained Object Detector for Unsupervised Domain Adaptation
di: Fujii, Kazuma, et al.
Pubblicazione: (2021)
di: Fujii, Kazuma, et al.
Pubblicazione: (2021)
Distilling Specialized Orders for Visual Generation
di: Pramanik, Rishav, et al.
Pubblicazione: (2025)
di: Pramanik, Rishav, et al.
Pubblicazione: (2025)
unMORE: Unsupervised Multi-Object Segmentation via Center-Boundary Reasoning
di: Yang, Yafei, et al.
Pubblicazione: (2025)
di: Yang, Yafei, et al.
Pubblicazione: (2025)
Transformers and Slot Encoding for Sample Efficient Physical World Modelling
di: Petri, Francesco, et al.
Pubblicazione: (2024)
di: Petri, Francesco, et al.
Pubblicazione: (2024)
SlotMatch: Distilling Object-Centric Representations for Unsupervised Video Segmentation
di: Grigore, Diana-Nicoleta, et al.
Pubblicazione: (2025)
di: Grigore, Diana-Nicoleta, et al.
Pubblicazione: (2025)
Align Your Query: Representation Alignment for Multimodality Medical Object Detection
di: Seo, Ara, et al.
Pubblicazione: (2025)
di: Seo, Ara, et al.
Pubblicazione: (2025)
UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation
di: Kang, Wonjun, et al.
Pubblicazione: (2025)
di: Kang, Wonjun, et al.
Pubblicazione: (2025)
Improving Zero-shot Generalization of Learned Prompts via Unsupervised Knowledge Distillation
di: Mistretta, Marco, et al.
Pubblicazione: (2024)
di: Mistretta, Marco, et al.
Pubblicazione: (2024)
MAPSeg: Unified Unsupervised Domain Adaptation for Heterogeneous Medical Image Segmentation Based on 3D Masked Autoencoding and Pseudo-Labeling
di: Zhang, Xuzhe, et al.
Pubblicazione: (2023)
di: Zhang, Xuzhe, et al.
Pubblicazione: (2023)
Progressive Confident Masking Attention Network for Audio-Visual Segmentation
di: Wang, Yuxuan, et al.
Pubblicazione: (2024)
di: Wang, Yuxuan, et al.
Pubblicazione: (2024)
Multi-Source Domain Adaptation for Object Detection with Prototype-based Mean-teacher
di: Belal, Atif, et al.
Pubblicazione: (2023)
di: Belal, Atif, et al.
Pubblicazione: (2023)
MoH: Multi-Head Attention as Mixture-of-Head Attention
di: Jin, Peng, et al.
Pubblicazione: (2024)
di: Jin, Peng, et al.
Pubblicazione: (2024)
Prediction Accuracy & Reliability: Classification and Object Localization under Distribution Shift
di: Diet, Fabian, et al.
Pubblicazione: (2024)
di: Diet, Fabian, et al.
Pubblicazione: (2024)
FORLA: Federated Object-centric Representation Learning with Slot Attention
di: Liao, Guiqiu, et al.
Pubblicazione: (2025)
di: Liao, Guiqiu, et al.
Pubblicazione: (2025)
Generalized Neighborhood Attention: Multi-dimensional Sparse Attention at the Speed of Light
di: Hassani, Ali, et al.
Pubblicazione: (2025)
di: Hassani, Ali, et al.
Pubblicazione: (2025)
Out-of-Distribution Detection with Attention Head Masking for Multimodal Document Classification
di: Constantinou, Christos, et al.
Pubblicazione: (2024)
di: Constantinou, Christos, et al.
Pubblicazione: (2024)
MuM: Multi-View Masked Image Modeling for 3D Vision
di: Nordström, David, et al.
Pubblicazione: (2025)
di: Nordström, David, et al.
Pubblicazione: (2025)
MixMask: Revisiting Masking Strategy for Siamese ConvNets
di: Vishniakov, Kirill, et al.
Pubblicazione: (2022)
di: Vishniakov, Kirill, et al.
Pubblicazione: (2022)
CNC: Cross-modal Normality Constraint for Unsupervised Multi-class Anomaly Detection
di: Wang, Xiaolei, et al.
Pubblicazione: (2024)
di: Wang, Xiaolei, et al.
Pubblicazione: (2024)
Mitigating Object Hallucinations in Vision-Language Models through Region-Aware Attention Recalibration
di: Xu, Yuanzhi, et al.
Pubblicazione: (2026)
di: Xu, Yuanzhi, et al.
Pubblicazione: (2026)
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation
di: Yariv, Guy, et al.
Pubblicazione: (2025)
di: Yariv, Guy, et al.
Pubblicazione: (2025)
Multi-modal Masked Siamese Network Improves Chest X-Ray Representation Learning
di: Shurrab, Saeed, et al.
Pubblicazione: (2024)
di: Shurrab, Saeed, et al.
Pubblicazione: (2024)
Gradient-Sign Masking for Task Vector Transport Across Pre-Trained Models
di: Rinaldi, Filippo, et al.
Pubblicazione: (2025)
di: Rinaldi, Filippo, et al.
Pubblicazione: (2025)
Beyond Labels: A Self-Supervised Framework with Masked Autoencoders and Random Cropping for Breast Cancer Subtype Classification
di: Chiocchetti, Annalisa, et al.
Pubblicazione: (2024)
di: Chiocchetti, Annalisa, et al.
Pubblicazione: (2024)
Reproducibility Study of CDUL: CLIP-Driven Unsupervised Learning for Multi-Label Image Classification
di: Shah, Manan, et al.
Pubblicazione: (2024)
di: Shah, Manan, et al.
Pubblicazione: (2024)
Not All Attention Heads Are What You Need: Refining CLIP's Image Representation with Attention Ablation
di: Lin, Feng, et al.
Pubblicazione: (2025)
di: Lin, Feng, et al.
Pubblicazione: (2025)
TrackNetV4: Enhancing Fast Sports Object Tracking with Motion Attention Maps
di: Raj, Arjun, et al.
Pubblicazione: (2024)
di: Raj, Arjun, et al.
Pubblicazione: (2024)
Decoupling Amplitude and Phase Attention in Frequency Domain for RGB-Event based Visual Object Tracking
di: Wang, Shiao, et al.
Pubblicazione: (2026)
di: Wang, Shiao, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Divided Attention: Unsupervised Multi-Object Discovery with Contextually Separated Slots
di: Lao, Dong, et al.
Pubblicazione: (2023) -
Unsupervised Object Discovery: A Comprehensive Survey and Unified Taxonomy
di: Villa-Vásquez, José-Fabian, et al.
Pubblicazione: (2024) -
Source-Free Domain Adaptation for YOLO Object Detection
di: Varailhon, Simon, et al.
Pubblicazione: (2024) -
Adaptive Slot Attention: Object Discovery with Dynamic Slot Number
di: Fan, Ke, et al.
Pubblicazione: (2024) -
Visual Modality Prompt for Adapting Vision-Language Object Detectors
di: Medeiros, Heitor R., et al.
Pubblicazione: (2024)