Saved in:
| Main Authors: | Wang, Rui-Feng, Petti, Daniel, Chen, Yue, Li, Changying |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2603.02419 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DINOv3
by: Siméoni, Oriane, et al.
Published: (2025)
by: Siméoni, Oriane, et al.
Published: (2025)
AD-DINOv3: Enhancing DINOv3 for Zero-Shot Anomaly Detection with Anomaly-Aware Calibration
by: Yuan, Jingyi, et al.
Published: (2025)
by: Yuan, Jingyi, et al.
Published: (2025)
DINOv2: Learning Robust Visual Features without Supervision
by: Oquab, Maxime, et al.
Published: (2023)
by: Oquab, Maxime, et al.
Published: (2023)
NeuroSeg Meets DINOv3: Transferring 2D Self-Supervised Visual Priors to 3D Neuron Segmentation via DINOv3 Initialization
by: Cheng, Yik San, et al.
Published: (2026)
by: Cheng, Yik San, et al.
Published: (2026)
INSID3: Training-Free In-Context Segmentation with DINOv3
by: Cuttano, Claudia, et al.
Published: (2026)
by: Cuttano, Claudia, et al.
Published: (2026)
DINOv3-Diffusion Policy: Self-Supervised Large Visual Model for Visuomotor Diffusion Policy Learning
by: Egbe, ThankGod, et al.
Published: (2025)
by: Egbe, ThankGod, et al.
Published: (2025)
DINOv3 Beats Specialized Detectors: A Simple Foundation Model Baseline for Image Forensics
by: Yu, Jieming, et al.
Published: (2026)
by: Yu, Jieming, et al.
Published: (2026)
Optimizing DINOv2 with Registers for Face Anti-Spoofing
by: Feng, Mika, et al.
Published: (2025)
by: Feng, Mika, et al.
Published: (2025)
Visual Bridge: Universal Visual Perception Representations Generating
by: Gao, Yilin, et al.
Published: (2025)
by: Gao, Yilin, et al.
Published: (2025)
DINOv2 Driven Gait Representation Learning for Video-Based Visible-Infrared Person Re-identification
by: Yang, Yujie, et al.
Published: (2025)
by: Yang, Yujie, et al.
Published: (2025)
SINDER: Repairing the Singular Defects of DINOv2
by: Wang, Haoqi, et al.
Published: (2024)
by: Wang, Haoqi, et al.
Published: (2024)
Real-Time Object Detection Meets DINOv3
by: Huang, Shihua, et al.
Published: (2025)
by: Huang, Shihua, et al.
Published: (2025)
NegoCollab: A Common Representation Negotiation Approach for Heterogeneous Collaborative Perception
by: Shao, Congzhang, et al.
Published: (2025)
by: Shao, Congzhang, et al.
Published: (2025)
Revisiting Birds Eye View Perception Models with Frozen Foundation Models: DINOv2 and Metric3Dv2
by: Hayes, Seamie, et al.
Published: (2025)
by: Hayes, Seamie, et al.
Published: (2025)
Rethinking Cross-Generator Image Forgery Detection through DINOv3
by: Huang, Zhenglin, et al.
Published: (2025)
by: Huang, Zhenglin, et al.
Published: (2025)
DINO-MVR: Multi-View Readout of Frozen DINOv3 for Annotation-Efficient Medical Segmentation
by: Jiang, Wei, et al.
Published: (2026)
by: Jiang, Wei, et al.
Published: (2026)
SUGAR: Pre-training 3D Visual Representations for Robotics
by: Chen, Shizhe, et al.
Published: (2024)
by: Chen, Shizhe, et al.
Published: (2024)
Benchmarking DINOv3 for Multi-Task Stroke Analysis on Non-Contrast CT
by: Zhang, Donghao, et al.
Published: (2025)
by: Zhang, Donghao, et al.
Published: (2025)
DINOv3 with Test-Time Training for Medical Image Registration
by: Wang, Shansong, et al.
Published: (2025)
by: Wang, Shansong, et al.
Published: (2025)
DINO-BOLDNet: A DINOv3-Guided Multi-Slice Attention Network for T1-to-BOLD Generation
by: Wang, Jianwei, et al.
Published: (2025)
by: Wang, Jianwei, et al.
Published: (2025)
Simultaneous Tactile-Visual Perception for Learning Multimodal Robot Manipulation
by: Li, Yuyang, et al.
Published: (2025)
by: Li, Yuyang, et al.
Published: (2025)
DinoDental: Benchmarking DINOv3 as a Unified Vision Encoder for Dental Image Analysis
by: Tang, Kun, et al.
Published: (2026)
by: Tang, Kun, et al.
Published: (2026)
Parameter-Efficient Fine-Tuning of DINOv2 for Large-Scale Font Classification
by: Chen, Daniel, et al.
Published: (2026)
by: Chen, Daniel, et al.
Published: (2026)
CHMv2: Improvements in Global Canopy Height Mapping using DINOv3
by: Brandt, John, et al.
Published: (2026)
by: Brandt, John, et al.
Published: (2026)
MedDINOv3: How to adapt vision foundation models for medical image segmentation?
by: Li, Yuheng, et al.
Published: (2025)
by: Li, Yuheng, et al.
Published: (2025)
Spatial Autoregressive Modeling of DINOv3 Embeddings for Unsupervised Anomaly Detection
by: Erdil, Ertunc, et al.
Published: (2026)
by: Erdil, Ertunc, et al.
Published: (2026)
Towards Adversarial Robustness and Uncertainty Quantification in DINOv2-based Few-Shot Anomaly Detection
by: Khan, Akib Mohammed, et al.
Published: (2025)
by: Khan, Akib Mohammed, et al.
Published: (2025)
DINOv3-Guided Cross Fusion Framework for Semantic-aware CT generation from MRI and CBCT
by: Zhou, Xianhao, et al.
Published: (2025)
by: Zhou, Xianhao, et al.
Published: (2025)
Does DINOv3 Set a New Medical Vision Standard? Benchmarking 2D and 3D Classification, Segmentation, and Registration
by: Liu, Che, et al.
Published: (2025)
by: Liu, Che, et al.
Published: (2025)
From Web to Pixels: Bringing Agentic Search into Visual Perception
by: Yang, Bokang, et al.
Published: (2026)
by: Yang, Bokang, et al.
Published: (2026)
From SAM to DINOv2: Towards Distilling Foundation Models to Lightweight Baselines for Generalized Polyp Segmentation
by: Agnihotri, Shivanshu, et al.
Published: (2025)
by: Agnihotri, Shivanshu, et al.
Published: (2025)
Dyn-Adapter: Towards Disentangled Representation for Efficient Visual Recognition
by: Zhang, Yurong, et al.
Published: (2024)
by: Zhang, Yurong, et al.
Published: (2024)
What Is The Best 3D Scene Representation for Robotics? From Geometric to Foundation Models
by: Deng, Tianchen, et al.
Published: (2025)
by: Deng, Tianchen, et al.
Published: (2025)
Learning to Adapt Foundation Model DINOv2 for Capsule Endoscopy Diagnosis
by: Zhang, Bowen, et al.
Published: (2024)
by: Zhang, Bowen, et al.
Published: (2024)
Enhancing Sentiment Analysis through Multimodal Fusion: A BERT-DINOv2 Approach
by: Zhao, Taoxu, et al.
Published: (2025)
by: Zhao, Taoxu, et al.
Published: (2025)
Real-Time Privacy Preservation for Robot Visual Perception
by: Choi, Minkyu, et al.
Published: (2025)
by: Choi, Minkyu, et al.
Published: (2025)
Temporal vs. Spatial: Comparing DINOv3 and V-JEPA2 Feature Representations for Video Action Analysis
by: Kodathala, Sai Varun, et al.
Published: (2025)
by: Kodathala, Sai Varun, et al.
Published: (2025)
Accurate Crop Yield Estimation of Blueberries using Deep Learning and Smart Drones
by: Nguyen, Hieu D., et al.
Published: (2025)
by: Nguyen, Hieu D., et al.
Published: (2025)
Lightweight Distillation of SAM 3 and DINOv3 for Edge-Deployable Individual-Level Livestock Monitoring and Longitudinal Visual Analytics
by: Yang, Haiyu, et al.
Published: (2026)
by: Yang, Haiyu, et al.
Published: (2026)
DINO Soars: DINOv3 for Open-Vocabulary Semantic Segmentation of Remote Sensing Imagery
by: Faulkenberry, Ryan, et al.
Published: (2026)
by: Faulkenberry, Ryan, et al.
Published: (2026)
Similar Items
-
DINOv3
by: Siméoni, Oriane, et al.
Published: (2025) -
AD-DINOv3: Enhancing DINOv3 for Zero-Shot Anomaly Detection with Anomaly-Aware Calibration
by: Yuan, Jingyi, et al.
Published: (2025) -
DINOv2: Learning Robust Visual Features without Supervision
by: Oquab, Maxime, et al.
Published: (2023) -
NeuroSeg Meets DINOv3: Transferring 2D Self-Supervised Visual Priors to 3D Neuron Segmentation via DINOv3 Initialization
by: Cheng, Yik San, et al.
Published: (2026) -
INSID3: Training-Free In-Context Segmentation with DINOv3
by: Cuttano, Claudia, et al.
Published: (2026)