AD-DINO: Attention-Dynamic DINO for Distance-Aware Embodied Reference Understanding
Fuente:
arXiv
Guardado en:
| Autores principales: | Guo, Hao, Fan, Wei, Wei, Baichun, Zhu, Jianfei, Tian, Jin, Yi, Chunzhi, Jiang, Feng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Beyond Object Categories: Multi-Attribute Reference Understanding for Visual Grounding
por: Guo, Hao, et al.
Publicado: (2025)
por: Guo, Hao, et al.
Publicado: (2025)
DINO-Tok: Adapting DINO for Visual Tokenizers
por: Jia, Mingkai, et al.
Publicado: (2025)
por: Jia, Mingkai, et al.
Publicado: (2025)
Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
por: Liu, Shilong, et al.
Publicado: (2023)
por: Liu, Shilong, et al.
Publicado: (2023)
DINO-Foresight: Looking into the Future with DINO
por: Karypidis, Efstathios, et al.
Publicado: (2024)
por: Karypidis, Efstathios, et al.
Publicado: (2024)
DINO-AD: Unsupervised Anomaly Detection with Frozen DINO-V3 Features
por: Huo, Jiayu, et al.
Publicado: (2026)
por: Huo, Jiayu, et al.
Publicado: (2026)
ReferDINO: Referring Video Object Segmentation with Visual Grounding Foundations
por: Liang, Tianming, et al.
Publicado: (2025)
por: Liang, Tianming, et al.
Publicado: (2025)
Evaluating Stenosis Detection with Grounding DINO, YOLO, and DINO-DETR
por: Ansari, Muhammad Musab
Publicado: (2025)
por: Ansari, Muhammad Musab
Publicado: (2025)
SegDINO: An Efficient Design for Medical and Natural Image Segmentation with DINO-V3
por: Yang, Sicheng, et al.
Publicado: (2025)
por: Yang, Sicheng, et al.
Publicado: (2025)
OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion
por: Wang, Hao, et al.
Publicado: (2024)
por: Wang, Hao, et al.
Publicado: (2024)
Deploy DINO with Many-to-Many Association
por: Jiang, Haodong, et al.
Publicado: (2026)
por: Jiang, Haodong, et al.
Publicado: (2026)
PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training
por: Fu, Weifu, et al.
Publicado: (2026)
por: Fu, Weifu, et al.
Publicado: (2026)
DINO-MVR: Multi-View Readout of Frozen DINOv3 for Annotation-Efficient Medical Segmentation
por: Jiang, Wei, et al.
Publicado: (2026)
por: Jiang, Wei, et al.
Publicado: (2026)
ReferDINO-Plus: 2nd Solution for 4th PVUW MeViS Challenge at CVPR 2025
por: Liang, Tianming, et al.
Publicado: (2025)
por: Liang, Tianming, et al.
Publicado: (2025)
DINO-Tracker: Taming DINO for Self-Supervised Point Tracking in a Single Video
por: Tumanyan, Narek, et al.
Publicado: (2024)
por: Tumanyan, Narek, et al.
Publicado: (2024)
DINO-SLAM: DINO-informed RGB-D SLAM for Neural Implicit and Explicit Representations
por: Gong, Ziren, et al.
Publicado: (2025)
por: Gong, Ziren, et al.
Publicado: (2025)
Unlocking Generalization in Polyp Segmentation with DINO Self-Attention "keys"
por: Monteiro, Carla, et al.
Publicado: (2025)
por: Monteiro, Carla, et al.
Publicado: (2025)
DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding
por: Ren, Tianhe, et al.
Publicado: (2024)
por: Ren, Tianhe, et al.
Publicado: (2024)
OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance
por: Zeng, Haoxi, et al.
Publicado: (2026)
por: Zeng, Haoxi, et al.
Publicado: (2026)
DIVE: Taming DINO for Subject-Driven Video Editing
por: Huang, Yi, et al.
Publicado: (2024)
por: Huang, Yi, et al.
Publicado: (2024)
GuiDINO: Rethinking Vision Foundation Model in Medical Image Segmentation
por: Liang, Zhuonan, et al.
Publicado: (2026)
por: Liang, Zhuonan, et al.
Publicado: (2026)
From CLIP to DINO: Visual Encoders Shout in Multi-modal Large Language Models
por: Jiang, Dongsheng, et al.
Publicado: (2023)
por: Jiang, Dongsheng, et al.
Publicado: (2023)
Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Object Detection
por: Lu, Yehao, et al.
Publicado: (2025)
por: Lu, Yehao, et al.
Publicado: (2025)
DACoN: DINO for Anime Paint Bucket Colorization with Any Number of Reference Images
por: Nagata, Kazuma, et al.
Publicado: (2025)
por: Nagata, Kazuma, et al.
Publicado: (2025)
Grounding DINO 1.5: Advance the "Edge" of Open-Set Object Detection
por: Ren, Tianhe, et al.
Publicado: (2024)
por: Ren, Tianhe, et al.
Publicado: (2024)
Simplifying DINO via Coding Rate Regularization
por: Wu, Ziyang, et al.
Publicado: (2025)
por: Wu, Ziyang, et al.
Publicado: (2025)
Empowering DINO Representations for Underwater Instance Segmentation via Aligner and Prompter
por: Chen, Zhiyang, et al.
Publicado: (2025)
por: Chen, Zhiyang, et al.
Publicado: (2025)
PixelDINO: Semi-Supervised Semantic Segmentation for Detecting Permafrost Disturbances
por: Heidler, Konrad, et al.
Publicado: (2024)
por: Heidler, Konrad, et al.
Publicado: (2024)
Cross-DINO: Cross the Deep MLP and Transformer for Small Object Detection
por: Cao, Guiping, et al.
Publicado: (2025)
por: Cao, Guiping, et al.
Publicado: (2025)
Text-guided Visual Prompt DINO for Generic Segmentation
por: Guan, Yuchen, et al.
Publicado: (2025)
por: Guan, Yuchen, et al.
Publicado: (2025)
Few-Shot Adaptation of Grounding DINO for Agricultural Domain
por: Singh, Rajhans, et al.
Publicado: (2025)
por: Singh, Rajhans, et al.
Publicado: (2025)
Tri-path DINO: Feature Complementary Learning for Remote Sensing Multi-Class Change Detection
por: Zheng, Kai, et al.
Publicado: (2026)
por: Zheng, Kai, et al.
Publicado: (2026)
FreqDINO: Frequency-Guided Adaptation for Generalized Boundary-Aware Ultrasound Image Segmentation
por: Zhang, Yixuan, et al.
Publicado: (2025)
por: Zhang, Yixuan, et al.
Publicado: (2025)
MammoDINO: Anatomically Aware Self-Supervision for Mammographic Images
por: Zhou, Sicheng, et al.
Publicado: (2025)
por: Zhou, Sicheng, et al.
Publicado: (2025)
Back to the Features: DINO as a Foundation for Video World Models
por: Baldassarre, Federico, et al.
Publicado: (2025)
por: Baldassarre, Federico, et al.
Publicado: (2025)
Feed-Forward SceneDINO for Unsupervised Semantic Scene Completion
por: Jevtić, Aleksandar, et al.
Publicado: (2025)
por: Jevtić, Aleksandar, et al.
Publicado: (2025)
Multi-task Image Restoration Guided By Robust DINO Features
por: Lin, Xin, et al.
Publicado: (2023)
por: Lin, Xin, et al.
Publicado: (2023)
Unlocking the Potential of Grounding DINO in Videos: Parameter-Efficient Adaptation for Limited-Data Spatial-Temporal Localization
por: Wang, Zanyi, et al.
Publicado: (2026)
por: Wang, Zanyi, et al.
Publicado: (2026)
DINO-BOLDNet: A DINOv3-Guided Multi-Slice Attention Network for T1-to-BOLD Generation
por: Wang, Jianwei, et al.
Publicado: (2025)
por: Wang, Jianwei, et al.
Publicado: (2025)
DINO-SD: Champion Solution for ICRA 2024 RoboDepth Challenge
por: Mao, Yifan, et al.
Publicado: (2024)
por: Mao, Yifan, et al.
Publicado: (2024)
SpectraDINO: Bridging the Spectral Gap in Vision Foundation Models via Lightweight Adapters
por: Nalcakan, Yagiz, et al.
Publicado: (2026)
por: Nalcakan, Yagiz, et al.
Publicado: (2026)
Ejemplares similares
-
Beyond Object Categories: Multi-Attribute Reference Understanding for Visual Grounding
por: Guo, Hao, et al.
Publicado: (2025) -
DINO-Tok: Adapting DINO for Visual Tokenizers
por: Jia, Mingkai, et al.
Publicado: (2025) -
Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
por: Liu, Shilong, et al.
Publicado: (2023) -
DINO-Foresight: Looking into the Future with DINO
por: Karypidis, Efstathios, et al.
Publicado: (2024) -
DINO-AD: Unsupervised Anomaly Detection with Frozen DINO-V3 Features
por: Huo, Jiayu, et al.
Publicado: (2026)