DatUS^2: Data-driven Unsupervised Semantic Segmentation with Pre-trained Self-supervised Vision Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kumar, Sonal, Sur, Arijit, Baruah, Rashmi Dutta |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Kolmogorov-Arnold Attention: Is Learnable Attention Better For Vision Transformers?
von: Maity, Subhajit, et al.
Veröffentlicht: (2025)
von: Maity, Subhajit, et al.
Veröffentlicht: (2025)
Domain-Specific Self-Supervised Pre-training for Agricultural Disease Classification: A Hierarchical Vision Transformer Study
von: Sonavane, Arnav S.
Veröffentlicht: (2026)
von: Sonavane, Arnav S.
Veröffentlicht: (2026)
AttEntropy: On the Generalization Ability of Supervised Semantic Segmentation Transformers to New Objects in New Domains
von: Lis, Krzysztof, et al.
Veröffentlicht: (2022)
von: Lis, Krzysztof, et al.
Veröffentlicht: (2022)
SAM Fewshot Finetuning for Anatomical Segmentation in Medical Images
von: Xie, Weiyi, et al.
Veröffentlicht: (2024)
von: Xie, Weiyi, et al.
Veröffentlicht: (2024)
Low-Cost Tree Crown Dieback Estimation Using Deep Learning-Based Segmentation
von: Allen, M. J., et al.
Veröffentlicht: (2024)
von: Allen, M. J., et al.
Veröffentlicht: (2024)
Pan-Arctic Permafrost Landform and Human-built Infrastructure Feature Detection with Vision Transformers and Location Embeddings
von: Perera, Amal S., et al.
Veröffentlicht: (2025)
von: Perera, Amal S., et al.
Veröffentlicht: (2025)
Manual Labelling Artificially Inflates Deep Learning-Based Segmentation Performance on RGB Images of Closed Canopy: Validation Using TLS
von: Allen, Matthew J., et al.
Veröffentlicht: (2025)
von: Allen, Matthew J., et al.
Veröffentlicht: (2025)
Polyp-SES: Automatic Polyp Segmentation with Self-Enriched Semantic Model
von: Nguyen, Quang Vinh, et al.
Veröffentlicht: (2024)
von: Nguyen, Quang Vinh, et al.
Veröffentlicht: (2024)
TUMLS: Trustful Fully Unsupervised Multi-Level Segmentation for Whole Slide Images of Histology
von: Rehamnia, Walid, et al.
Veröffentlicht: (2025)
von: Rehamnia, Walid, et al.
Veröffentlicht: (2025)
CFFormer: Cross CNN-Transformer Channel Attention and Spatial Feature Fusion for Improved Segmentation of Heterogeneous Medical Images
von: Li, Jiaxuan, et al.
Veröffentlicht: (2025)
von: Li, Jiaxuan, et al.
Veröffentlicht: (2025)
Canonical Space Representation for 4D Panoptic Segmentation of Articulated Objects
von: Gomes, Manuel, et al.
Veröffentlicht: (2025)
von: Gomes, Manuel, et al.
Veröffentlicht: (2025)
U-R-VEDA: Integrating UNET, Residual Links, Edge and Dual Attention, and Vision Transformer for Accurate Semantic Segmentation of CMRs
von: Mukisa, Racheal, et al.
Veröffentlicht: (2025)
von: Mukisa, Racheal, et al.
Veröffentlicht: (2025)
AOI-SSL: Self-Supervised Framework for Efficient Segmentation of Wire-bonded Semiconductors In Optical Inspection
von: Figueira, Joaquín, et al.
Veröffentlicht: (2026)
von: Figueira, Joaquín, et al.
Veröffentlicht: (2026)
Unified Local and Global Attention Interaction Modeling for Vision Transformers
von: Nguyen, Tan, et al.
Veröffentlicht: (2024)
von: Nguyen, Tan, et al.
Veröffentlicht: (2024)
Optimal Blackjack Strategy Recommender: A Comprehensive Study on Computer Vision Integration for Enhanced Gameplay
von: Gupta, Krishnanshu, et al.
Veröffentlicht: (2024)
von: Gupta, Krishnanshu, et al.
Veröffentlicht: (2024)
M3LEO: A Multi-Modal, Multi-Label Earth Observation Dataset Integrating Interferometric SAR and Multispectral Data
von: Allen, Matthew J, et al.
Veröffentlicht: (2024)
von: Allen, Matthew J, et al.
Veröffentlicht: (2024)
HSDA: High-frequency Shuffle Data Augmentation for Bird's-Eye-View Map Segmentation
von: Glisson, Calvin, et al.
Veröffentlicht: (2024)
von: Glisson, Calvin, et al.
Veröffentlicht: (2024)
FLD+: Data-efficient Evaluation Metric for Generative Models
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
A Sensorimotor Vision Transformer
von: Gadzicki, Konrad, et al.
Veröffentlicht: (2025)
von: Gadzicki, Konrad, et al.
Veröffentlicht: (2025)
Grounding Synthetic Data Generation With Vision and Language Models
von: Çağlar, Ümit Mert, et al.
Veröffentlicht: (2026)
von: Çağlar, Ümit Mert, et al.
Veröffentlicht: (2026)
Fusion and Grouping Strategies in Deep Learning for Local Climate Zone Classification of Multimodal Remote Sensing Data
von: Thomas, Ancymol, et al.
Veröffentlicht: (2026)
von: Thomas, Ancymol, et al.
Veröffentlicht: (2026)
UTAL-GNN: Unsupervised Temporal Action Localization using Graph Neural Networks
von: Badatya, Bikash Kumar, et al.
Veröffentlicht: (2025)
von: Badatya, Bikash Kumar, et al.
Veröffentlicht: (2025)
Cytoarchitecture in Words: Weakly Supervised Vision-Language Modeling for Human Brain Microscopy
von: Sutton, Matthew, et al.
Veröffentlicht: (2026)
von: Sutton, Matthew, et al.
Veröffentlicht: (2026)
SelvaMask: Segmenting Trees in Tropical Forests and Beyond
von: Duguay, Simon-Olivier, et al.
Veröffentlicht: (2026)
von: Duguay, Simon-Olivier, et al.
Veröffentlicht: (2026)
Which Backbone to Use: A Resource-efficient Domain Specific Comparison for Computer Vision
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
SpATr: MoCap 3D Human Action Recognition based on Spiral Auto-encoder and Transformer Network
von: Bouzid, Hamza, et al.
Veröffentlicht: (2023)
von: Bouzid, Hamza, et al.
Veröffentlicht: (2023)
Do All Vision Transformers Need Registers? A Cross-Architectural Reassessment
von: Baxevanakis, Spiros, et al.
Veröffentlicht: (2026)
von: Baxevanakis, Spiros, et al.
Veröffentlicht: (2026)
Adaptation of Distinct Semantics for Uncertain Areas in Polyp Segmentation
von: Nguyen, Quang Vinh, et al.
Veröffentlicht: (2024)
von: Nguyen, Quang Vinh, et al.
Veröffentlicht: (2024)
Exploring Visual Embedding Spaces Induced by Vision Transformers for Online Auto Parts Marketplaces
von: Armijo, Cameron, et al.
Veröffentlicht: (2025)
von: Armijo, Cameron, et al.
Veröffentlicht: (2025)
Semi-Supervised Segmentation via Embedding Matching
von: Xie, Weiyi, et al.
Veröffentlicht: (2024)
von: Xie, Weiyi, et al.
Veröffentlicht: (2024)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
Attention-Aware Transformer-Based Aggregation Network for Video Periocular Recognition
von: Carreira, Luiz G F, et al.
Veröffentlicht: (2026)
von: Carreira, Luiz G F, et al.
Veröffentlicht: (2026)
treeX: Unsupervised Tree Instance Segmentation in Dense Forest Point Clouds
von: Burmeister, Josafat-Mattias, et al.
Veröffentlicht: (2025)
von: Burmeister, Josafat-Mattias, et al.
Veröffentlicht: (2025)
DNRSelect: Active Best View Selection for Deferred Neural Rendering
von: Wu, Dongli, et al.
Veröffentlicht: (2025)
von: Wu, Dongli, et al.
Veröffentlicht: (2025)
Pairwise Spatiotemporal Partial Trajectory Matching for Co-movement Analysis
von: Cardei, Maria, et al.
Veröffentlicht: (2024)
von: Cardei, Maria, et al.
Veröffentlicht: (2024)
Event-based Solutions for Human-centered Applications: A Comprehensive Review
von: Adra, Mira, et al.
Veröffentlicht: (2025)
von: Adra, Mira, et al.
Veröffentlicht: (2025)
Prototype Contrastive Consistency Learning for Semi-Supervised Medical Image Segmentation
von: He, Shihuan, et al.
Veröffentlicht: (2025)
von: He, Shihuan, et al.
Veröffentlicht: (2025)
Feature-Augmented Deep Networks for Multiscale Building Segmentation in High-Resolution UAV and Satellite Imagery
von: Maniyar, Chintan B., et al.
Veröffentlicht: (2025)
von: Maniyar, Chintan B., et al.
Veröffentlicht: (2025)
Learning from Semantic Dictionaries: Discriminative Codebook Contrastive Learning for Unified Visual Representation and Generation
von: Estepa, Imanol G., et al.
Veröffentlicht: (2026)
von: Estepa, Imanol G., et al.
Veröffentlicht: (2026)
Parameter-efficient fine-tuning (PEFT) of Vision Foundation Models for Atypical Mitotic Figure Classification
von: Ramchandani, Lavish, et al.
Veröffentlicht: (2025)
von: Ramchandani, Lavish, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Kolmogorov-Arnold Attention: Is Learnable Attention Better For Vision Transformers?
von: Maity, Subhajit, et al.
Veröffentlicht: (2025) -
Domain-Specific Self-Supervised Pre-training for Agricultural Disease Classification: A Hierarchical Vision Transformer Study
von: Sonavane, Arnav S.
Veröffentlicht: (2026) -
AttEntropy: On the Generalization Ability of Supervised Semantic Segmentation Transformers to New Objects in New Domains
von: Lis, Krzysztof, et al.
Veröffentlicht: (2022) -
SAM Fewshot Finetuning for Anatomical Segmentation in Medical Images
von: Xie, Weiyi, et al.
Veröffentlicht: (2024) -
Low-Cost Tree Crown Dieback Estimation Using Deep Learning-Based Segmentation
von: Allen, M. J., et al.
Veröffentlicht: (2024)