DINO-MX: A Modular & Flexible Framework for Self-Supervised Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Gokmen, Mahmut Selman, Bumgardner, Cody |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
High Noise Scheduling is a Must
by: Gokmen, Mahmut S., et al.
Published: (2024)
by: Gokmen, Mahmut S., et al.
Published: (2024)
Curriculum-Driven 3D CT Report Generation via Language-Free Visual Grafting and Zone-Constrained Compression
by: Bumgardner, V. K. Cody, et al.
Published: (2026)
by: Bumgardner, V. K. Cody, et al.
Published: (2026)
DINO-LG: Enhancing Vision Transformers with Label Guidance for Coronary Artery Calcium Detection
by: Gokmen, Mahmut S., et al.
Published: (2024)
by: Gokmen, Mahmut S., et al.
Published: (2024)
Enhancing Low Dose Computed Tomography Images Using Consistency Training Techniques
by: Gokmen, Mahmut S., et al.
Published: (2024)
by: Gokmen, Mahmut S., et al.
Published: (2024)
A Framework for Cross-Domain Generalization in Coronary Artery Calcium Scoring Across Gated and Non-Gated Computed Tomography
by: Gokmen, Mahmut S., et al.
Published: (2026)
by: Gokmen, Mahmut S., et al.
Published: (2026)
Magnification-Aware Distillation (MAD): A Self-Supervised Framework for Unified Representation Learning in Gigapixel Whole-Slide Images
by: Gokmen, Mahmut S., et al.
Published: (2025)
by: Gokmen, Mahmut S., et al.
Published: (2025)
AdvDINO: Domain-Adversarial Self-Supervised Representation Learning for Spatial Proteomics
by: Su, Stella, et al.
Published: (2025)
by: Su, Stella, et al.
Published: (2025)
MammoDINO: Anatomically Aware Self-Supervision for Mammographic Images
by: Zhou, Sicheng, et al.
Published: (2025)
by: Zhou, Sicheng, et al.
Published: (2025)
On Partial Prototype Collapse in the DINO Family of Self-Supervised Methods
by: Govindarajan, Hariprasath, et al.
Published: (2024)
by: Govindarajan, Hariprasath, et al.
Published: (2024)
DINO-YOLO: Self-Supervised Pre-training for Data-Efficient Object Detection in Civil Engineering Applications
by: P, Malaisree, et al.
Published: (2025)
by: P, Malaisree, et al.
Published: (2025)
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation
by: Barsellotti, Luca, et al.
Published: (2024)
by: Barsellotti, Luca, et al.
Published: (2024)
Vision Foundry: A System for Training Foundational Vision AI Models
by: Gokmen, Mahmut S., et al.
Published: (2025)
by: Gokmen, Mahmut S., et al.
Published: (2025)
Efficient License Plate Recognition via Pseudo-Labeled Supervision with Grounding DINO and YOLOv8
by: Vargoorani, Zahra Ebrahimi, et al.
Published: (2025)
by: Vargoorani, Zahra Ebrahimi, et al.
Published: (2025)
OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance
by: Zeng, Haoxi, et al.
Published: (2026)
by: Zeng, Haoxi, et al.
Published: (2026)
Simplifying DINO via Coding Rate Regularization
by: Wu, Ziyang, et al.
Published: (2025)
by: Wu, Ziyang, et al.
Published: (2025)
Surgical-DINO: Adapter Learning of Foundation Models for Depth Estimation in Endoscopic Surgery
by: Cui, Beilei, et al.
Published: (2024)
by: Cui, Beilei, et al.
Published: (2024)
A Mixed Diet Makes DINO An Omnivorous Vision Encoder
by: Kabra, Rishabh, et al.
Published: (2026)
by: Kabra, Rishabh, et al.
Published: (2026)
DIVE: Taming DINO for Subject-Driven Video Editing
by: Huang, Yi, et al.
Published: (2024)
by: Huang, Yi, et al.
Published: (2024)
Swiss DINO: Efficient and Versatile Vision Framework for On-device Personal Object Search
by: Paramonov, Kirill, et al.
Published: (2024)
by: Paramonov, Kirill, et al.
Published: (2024)
Flexible ViG: Learning the Self-Saliency for Flexible Object Recognition
by: Zuo, Lin, et al.
Published: (2024)
by: Zuo, Lin, et al.
Published: (2024)
Learning Generalized and Flexible Trajectory Models from Omni-Semantic Supervision
by: Zhu, Yuanshao, et al.
Published: (2025)
by: Zhu, Yuanshao, et al.
Published: (2025)
CoMAD: A Multiple-Teacher Self-Supervised Distillation Framework
by: Mandalika, Sriram, et al.
Published: (2025)
by: Mandalika, Sriram, et al.
Published: (2025)
Self-Contrastive Weakly Supervised Learning Framework for Prognostic Prediction Using Whole Slide Images
by: Fuster, Saul, et al.
Published: (2024)
by: Fuster, Saul, et al.
Published: (2024)
Self-Supervised Learning for Endoscopic Video Analysis
by: Hirsch, Roy, et al.
Published: (2023)
by: Hirsch, Roy, et al.
Published: (2023)
DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models
by: Pan, Chenbin, et al.
Published: (2025)
by: Pan, Chenbin, et al.
Published: (2025)
Into the Rabbit Hull: From Task-Relevant Concepts in DINO to Minkowski Geometry
by: Fel, Thomas, et al.
Published: (2025)
by: Fel, Thomas, et al.
Published: (2025)
Zero-Shot and Supervised Bird Image Segmentation Using Foundation Models: A Dual-Pipeline Approach with Grounding DINO~1.5, YOLOv11, and SAM~2.1
by: Munagala, Abhinav
Published: (2026)
by: Munagala, Abhinav
Published: (2026)
Data or Language Supervision: What Makes CLIP Better than DINO?
by: Liu, Yiming, et al.
Published: (2025)
by: Liu, Yiming, et al.
Published: (2025)
Automatized Self-Supervised Learning for Skin Lesion Screening
by: Useini, Vullnet, et al.
Published: (2023)
by: Useini, Vullnet, et al.
Published: (2023)
Cross-Task Attack: A Self-Supervision Generative Framework Based on Attention Shift
by: Zeng, Qingyuan, et al.
Published: (2024)
by: Zeng, Qingyuan, et al.
Published: (2024)
A Framework For Image Synthesis Using Supervised Contrastive Learning
by: Liu, Yibin, et al.
Published: (2024)
by: Liu, Yibin, et al.
Published: (2024)
DinoTwins: Combining DINO and Barlow Twins for Robust, Label-Efficient Vision Transformers
by: Podsiadly, Michael, et al.
Published: (2025)
by: Podsiadly, Michael, et al.
Published: (2025)
MVEB: Self-Supervised Learning with Multi-View Entropy Bottleneck
by: Wen, Liangjian, et al.
Published: (2024)
by: Wen, Liangjian, et al.
Published: (2024)
Subspace Clustering on Incomplete Data with Self-Supervised Contrastive Learning
by: Li, Huanran, et al.
Published: (2026)
by: Li, Huanran, et al.
Published: (2026)
BrainDINO: A Brain MRI Foundation Model for Generalizable Clinical Representation Learning
by: Wu, Yizhou, et al.
Published: (2026)
by: Wu, Yizhou, et al.
Published: (2026)
A 3DGS-Diffusion Self-Supervised Framework for Normal Estimation from a Single Image
by: Liang, Yanxing, et al.
Published: (2025)
by: Liang, Yanxing, et al.
Published: (2025)
Masked Scene Modeling: Narrowing the Gap Between Supervised and Self-Supervised Learning in 3D Scene Understanding
by: Hermosilla, Pedro, et al.
Published: (2025)
by: Hermosilla, Pedro, et al.
Published: (2025)
Mask & Match: Learning to Recognize Handwritten Math with Self-Supervised Attention
by: Mitra, Shree, et al.
Published: (2025)
by: Mitra, Shree, et al.
Published: (2025)
Self-Supervised Cross-Modal Learning for Image-to-Point Cloud Registration
by: Wang, Xingmei, et al.
Published: (2025)
by: Wang, Xingmei, et al.
Published: (2025)
Shifting to Machine Supervision: Annotation-Efficient Semi and Self-Supervised Learning for Automatic Medical Image Segmentation and Classification
by: Singh, Pranav, et al.
Published: (2023)
by: Singh, Pranav, et al.
Published: (2023)
Similar Items
-
High Noise Scheduling is a Must
by: Gokmen, Mahmut S., et al.
Published: (2024) -
Curriculum-Driven 3D CT Report Generation via Language-Free Visual Grafting and Zone-Constrained Compression
by: Bumgardner, V. K. Cody, et al.
Published: (2026) -
DINO-LG: Enhancing Vision Transformers with Label Guidance for Coronary Artery Calcium Detection
by: Gokmen, Mahmut S., et al.
Published: (2024) -
Enhancing Low Dose Computed Tomography Images Using Consistency Training Techniques
by: Gokmen, Mahmut S., et al.
Published: (2024) -
A Framework for Cross-Domain Generalization in Coronary Artery Calcium Scoring Across Gated and Non-Gated Computed Tomography
by: Gokmen, Mahmut S., et al.
Published: (2026)