PEM: Prototype-based Efficient MaskFormer for Image Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cavagnero, Niccolò, Rosi, Gabriele, Cuttano, Claudia, Pistilli, Francesca, Ciccone, Marco, Averta, Giuseppe, Cermelli, Fabio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The revenge of BiSeNet: Efficient Multi-Task Image Segmentation
von: Rosi, Gabriele, et al.
Veröffentlicht: (2024)
von: Rosi, Gabriele, et al.
Veröffentlicht: (2024)
Cross-Domain Transfer Learning with CoRTe: Consistent and Reliable Transfer from Black-Box to Lightweight Segmentation Model
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
What does CLIP know about peeling a banana?
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
Show or Tell? A Benchmark To Evaluate Visual and Textual Prompts in Semantic Segmentation
von: Rosi, Gabriele, et al.
Veröffentlicht: (2025)
von: Rosi, Gabriele, et al.
Veröffentlicht: (2025)
SANSA: Unleashing the Hidden Semantics in SAM2 for Few-Shot Segmentation
von: Cuttano, Claudia, et al.
Veröffentlicht: (2025)
von: Cuttano, Claudia, et al.
Veröffentlicht: (2025)
PrAda: Few-Shot Visual Adaptation for Text-Prompted Segmentation
von: Rosi, Gabriele, et al.
Veröffentlicht: (2026)
von: Rosi, Gabriele, et al.
Veröffentlicht: (2026)
HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
FORESCENE: FOREcasting human activity via latent SCENE graphs diffusion
von: Alliegro, Antonio, et al.
Veröffentlicht: (2025)
von: Alliegro, Antonio, et al.
Veröffentlicht: (2025)
Your ViT is Secretly an Image Segmentation Model
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
PMT: Plain Mask Transformer for Image and Video Segmentation with Frozen Vision Encoders
von: Cavagnero, Niccolò, et al.
Veröffentlicht: (2026)
von: Cavagnero, Niccolò, et al.
Veröffentlicht: (2026)
Transient Fault Tolerant Semantic Segmentation for Autonomous Driving
von: Iurada, Leonardo, et al.
Veröffentlicht: (2024)
von: Iurada, Leonardo, et al.
Veröffentlicht: (2024)
HiERO-StepG @ Ego4D Step Grounding Challenge: hierarchical activity understanding enables zero-shot step grounding
von: Zenotto, Andrea, et al.
Veröffentlicht: (2026)
von: Zenotto, Andrea, et al.
Veröffentlicht: (2026)
A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2024)
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2024)
Towards Data-Efficient Video Pre-training with Frozen Image Foundation Models
von: Orlova, Svetlana, et al.
Veröffentlicht: (2026)
von: Orlova, Svetlana, et al.
Veröffentlicht: (2026)
Hier-EgoPack: Hierarchical Egocentric Video Understanding with Diverse Task Perspectives
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
Learning reusable concepts across different egocentric video understanding tasks
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
A Partial Replication of MaskFormer in TensorFlow on TPUs for the TensorFlow Model Garden
von: Purohit, Vishal, et al.
Veröffentlicht: (2024)
von: Purohit, Vishal, et al.
Veröffentlicht: (2024)
Orion-Lite: Distilling LLM Reasoning into Efficient Vision-Only Driving Models
von: Gu, Jing, et al.
Veröffentlicht: (2026)
von: Gu, Jing, et al.
Veröffentlicht: (2026)
Searching on a Budget: HW-NAS with 10 Latency Probes
von: Capuano, Francesco, et al.
Veröffentlicht: (2025)
von: Capuano, Francesco, et al.
Veröffentlicht: (2025)
Domain Generalization using Action Sequences for Egocentric Action Recognition
von: Nasirimajd, Amirshayan, et al.
Veröffentlicht: (2025)
von: Nasirimajd, Amirshayan, et al.
Veröffentlicht: (2025)
INSID3: Training-Free In-Context Segmentation with DINOv3
von: Cuttano, Claudia, et al.
Veröffentlicht: (2026)
von: Cuttano, Claudia, et al.
Veröffentlicht: (2026)
MARCO: Navigating the Unseen Space of Semantic Correspondence
von: Cuttano, Claudia, et al.
Veröffentlicht: (2026)
von: Cuttano, Claudia, et al.
Veröffentlicht: (2026)
Secondary Vertex Reconstruction with MaskFormers
von: Van Stroud, Samuel, et al.
Veröffentlicht: (2023)
von: Van Stroud, Samuel, et al.
Veröffentlicht: (2023)
AMEGO: Active Memory from long EGOcentric videos
von: Goletto, Gabriele, et al.
Veröffentlicht: (2024)
von: Goletto, Gabriele, et al.
Veröffentlicht: (2024)
Mask4Former: Mask Transformer for 4D Panoptic Segmentation
von: Yilmaz, Kadir, et al.
Veröffentlicht: (2023)
von: Yilmaz, Kadir, et al.
Veröffentlicht: (2023)
PrototypeFormer: Learning to Explore Prototype Relationships for Few-shot Image Classification
von: Su, Meijuan, et al.
Veröffentlicht: (2023)
von: Su, Meijuan, et al.
Veröffentlicht: (2023)
ProtoMask: Segmentation-Guided Prototype Learning
von: Meinert, Steffen, et al.
Veröffentlicht: (2025)
von: Meinert, Steffen, et al.
Veröffentlicht: (2025)
VidEoMT: Your ViT is Secretly Also a Video Segmentation Model
von: Norouzi, Narges, et al.
Veröffentlicht: (2026)
von: Norouzi, Narges, et al.
Veröffentlicht: (2026)
From Prototypes to General Distributions: An Efficient Curriculum for Masked Image Modeling
von: Lin, Jinhong, et al.
Veröffentlicht: (2024)
von: Lin, Jinhong, et al.
Veröffentlicht: (2024)
Egocentric zone-aware action recognition across environments
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2024)
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2024)
SegFormer3D: an Efficient Transformer for 3D Medical Image Segmentation
von: Perera, Shehan, et al.
Veröffentlicht: (2024)
von: Perera, Shehan, et al.
Veröffentlicht: (2024)
Biologically-inspired Semi-supervised Semantic Segmentation for Biomedical Imaging
von: Ciampi, Luca, et al.
Veröffentlicht: (2024)
von: Ciampi, Luca, et al.
Veröffentlicht: (2024)
ShadowMaskFormer: Mask Augmented Patch Embeddings for Shadow Removal
von: Li, Zhuohao, et al.
Veröffentlicht: (2024)
von: Li, Zhuohao, et al.
Veröffentlicht: (2024)
AMBER -- Advanced SegFormer for Multi-Band Image Segmentation: an application to Hyperspectral Imaging
von: Dosi, Andrea, et al.
Veröffentlicht: (2024)
von: Dosi, Andrea, et al.
Veröffentlicht: (2024)
Prior2Former -- Evidential Modeling of Mask Transformers for Assumption-Free Open-World Panoptic Segmentation
von: Schmidt, Sebastian, et al.
Veröffentlicht: (2025)
von: Schmidt, Sebastian, et al.
Veröffentlicht: (2025)
Accuracy Improvement of Cell Image Segmentation Using Feedback Former
von: Mitsuoka, Hinako, et al.
Veröffentlicht: (2024)
von: Mitsuoka, Hinako, et al.
Veröffentlicht: (2024)
Proto-Former: Unified Facial Landmark Detection by Prototype Transformer
von: Hu, Shengkai, et al.
Veröffentlicht: (2025)
von: Hu, Shengkai, et al.
Veröffentlicht: (2025)
Efficient Model Editing with Task-Localized Sparse Fine-tuning
von: Iurada, Leonardo, et al.
Veröffentlicht: (2025)
von: Iurada, Leonardo, et al.
Veröffentlicht: (2025)
Efficient Progressive Image Compression with Variance-aware Masking
von: Presta, Alberto, et al.
Veröffentlicht: (2024)
von: Presta, Alberto, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
The revenge of BiSeNet: Efficient Multi-Task Image Segmentation
von: Rosi, Gabriele, et al.
Veröffentlicht: (2024) -
Cross-Domain Transfer Learning with CoRTe: Consistent and Reliable Transfer from Black-Box to Lightweight Segmentation Model
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024) -
What does CLIP know about peeling a banana?
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024) -
SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024) -
Show or Tell? A Benchmark To Evaluate Visual and Textual Prompts in Semantic Segmentation
von: Rosi, Gabriele, et al.
Veröffentlicht: (2025)