Deploy DINO with Many-to-Many Association
Fuente:
arXiv
Salvato in:
| Autori principali: | Jiang, Haodong, Li, Mingzhe, Wu, Junfeng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Inteval Analysis for two spherical functions arising from robust Perspective-n-Lines problem
di: Zheng, Xiang, et al.
Pubblicazione: (2025)
di: Zheng, Xiang, et al.
Pubblicazione: (2025)
FlexGS: Train Once, Deploy Everywhere with Many-in-One Flexible 3D Gaussian Splatting
di: Liu, Hengyu, et al.
Pubblicazione: (2025)
di: Liu, Hengyu, et al.
Pubblicazione: (2025)
Many-for-Many: Unify the Training of Multiple Video and Image Generation and Manipulation Tasks
di: Li, Ruibin, et al.
Pubblicazione: (2025)
di: Li, Ruibin, et al.
Pubblicazione: (2025)
DINO-Tok: Adapting DINO for Visual Tokenizers
di: Jia, Mingkai, et al.
Pubblicazione: (2025)
di: Jia, Mingkai, et al.
Pubblicazione: (2025)
DINO-Foresight: Looking into the Future with DINO
di: Karypidis, Efstathios, et al.
Pubblicazione: (2024)
di: Karypidis, Efstathios, et al.
Pubblicazione: (2024)
Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
di: Liu, Shilong, et al.
Pubblicazione: (2023)
di: Liu, Shilong, et al.
Pubblicazione: (2023)
DINO-SLAM: DINO-informed RGB-D SLAM for Neural Implicit and Explicit Representations
di: Gong, Ziren, et al.
Pubblicazione: (2025)
di: Gong, Ziren, et al.
Pubblicazione: (2025)
Spider: Any-to-Many Multimodal LLM
di: Lai, Jinxiang, et al.
Pubblicazione: (2024)
di: Lai, Jinxiang, et al.
Pubblicazione: (2024)
AD-DINO: Attention-Dynamic DINO for Distance-Aware Embodied Reference Understanding
di: Guo, Hao, et al.
Pubblicazione: (2024)
di: Guo, Hao, et al.
Pubblicazione: (2024)
Impact of Sunglasses on One-to-Many Facial Identification Accuracy
di: Tian, Sicong, et al.
Pubblicazione: (2024)
di: Tian, Sicong, et al.
Pubblicazione: (2024)
PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training
di: Fu, Weifu, et al.
Pubblicazione: (2026)
di: Fu, Weifu, et al.
Pubblicazione: (2026)
From One-to-One to Many-to-Many: Dynamic Cross-Layer Injection for Deep Vision-Language Fusion
di: Chen, Cheng, et al.
Pubblicazione: (2026)
di: Chen, Cheng, et al.
Pubblicazione: (2026)
Evaluating Stenosis Detection with Grounding DINO, YOLO, and DINO-DETR
di: Ansari, Muhammad Musab
Pubblicazione: (2025)
di: Ansari, Muhammad Musab
Pubblicazione: (2025)
Many-Worlds Inverse Rendering
di: Zhang, Ziyi, et al.
Pubblicazione: (2024)
di: Zhang, Ziyi, et al.
Pubblicazione: (2024)
DINO-Tracker: Taming DINO for Self-Supervised Point Tracking in a Single Video
di: Tumanyan, Narek, et al.
Pubblicazione: (2024)
di: Tumanyan, Narek, et al.
Pubblicazione: (2024)
SegDINO: An Efficient Design for Medical and Natural Image Segmentation with DINO-V3
di: Yang, Sicheng, et al.
Pubblicazione: (2025)
di: Yang, Sicheng, et al.
Pubblicazione: (2025)
Video Individual Counting With Implicit One-to-Many Matching
di: Zhu, Xuhui, et al.
Pubblicazione: (2025)
di: Zhu, Xuhui, et al.
Pubblicazione: (2025)
Beat: Bi-directional One-to-Many Embedding Alignment for Text-based Person Retrieval
di: Ma, Yiwei, et al.
Pubblicazione: (2024)
di: Ma, Yiwei, et al.
Pubblicazione: (2024)
TokenBinder: Text-Video Retrieval with One-to-Many Alignment Paradigm
di: Zhang, Bingqing, et al.
Pubblicazione: (2024)
di: Zhang, Bingqing, et al.
Pubblicazione: (2024)
ManiNeg: Manifestation-guided Multimodal Pretraining for Mammography Classification
di: Li, Xujun, et al.
Pubblicazione: (2024)
di: Li, Xujun, et al.
Pubblicazione: (2024)
Mani-GS: Gaussian Splatting Manipulation with Triangular Mesh
di: Gao, Xiangjun, et al.
Pubblicazione: (2024)
di: Gao, Xiangjun, et al.
Pubblicazione: (2024)
Exploring "Many in Few" and "Few in Many" Properties in Long-Tailed, Highly-Imbalanced IC Defect Classification
di: Shao, Hao-Chiang, et al.
Pubblicazione: (2025)
di: Shao, Hao-Chiang, et al.
Pubblicazione: (2025)
Bayesian Neural Networks for One-to-Many Mapping in Image Enhancement
di: Huang, Guoxi, et al.
Pubblicazione: (2025)
di: Huang, Guoxi, et al.
Pubblicazione: (2025)
Many-to-many Image Generation with Auto-regressive Diffusion Models
di: Shen, Ying, et al.
Pubblicazione: (2024)
di: Shen, Ying, et al.
Pubblicazione: (2024)
BrainDINO: A Brain MRI Foundation Model for Generalizable Clinical Representation Learning
di: Wu, Yizhou, et al.
Pubblicazione: (2026)
di: Wu, Yizhou, et al.
Pubblicazione: (2026)
Memories are One-to-Many Mapping Alleviators in Talking Face Generation
di: Tang, Anni, et al.
Pubblicazione: (2022)
di: Tang, Anni, et al.
Pubblicazione: (2022)
Beyond Subspace Isolation: Many-to-Many Transformer for Light Field Image Super-resolution
di: Hu, Zeke Zexi, et al.
Pubblicazione: (2024)
di: Hu, Zeke Zexi, et al.
Pubblicazione: (2024)
Learning Many-to-Many Mapping for Unpaired Real-World Image Super-resolution and Downscaling
di: Sun, Wanjie, et al.
Pubblicazione: (2023)
di: Sun, Wanjie, et al.
Pubblicazione: (2023)
SUGAR: A Sweeter Spot for Generative Unlearning of Many Identities
di: Nguyen, Dung Thuy, et al.
Pubblicazione: (2025)
di: Nguyen, Dung Thuy, et al.
Pubblicazione: (2025)
Cross-DINO: Cross the Deep MLP and Transformer for Small Object Detection
di: Cao, Guiping, et al.
Pubblicazione: (2025)
di: Cao, Guiping, et al.
Pubblicazione: (2025)
Text-guided Visual Prompt DINO for Generic Segmentation
di: Guan, Yuchen, et al.
Pubblicazione: (2025)
di: Guan, Yuchen, et al.
Pubblicazione: (2025)
Many-MobileNet: Multi-Model Augmentation for Robust Retinal Disease Classification
di: Wang, Hao, et al.
Pubblicazione: (2024)
di: Wang, Hao, et al.
Pubblicazione: (2024)
One Model, Many Budgets: Elastic Latent Interfaces for Diffusion Transformers
di: Haji-Ali, Moayed, et al.
Pubblicazione: (2026)
di: Haji-Ali, Moayed, et al.
Pubblicazione: (2026)
Grounding DINO 1.5: Advance the "Edge" of Open-Set Object Detection
di: Ren, Tianhe, et al.
Pubblicazione: (2024)
di: Ren, Tianhe, et al.
Pubblicazione: (2024)
In Pursuit of Many: A Review of Modern Multiple Object Tracking Systems
di: Bashar, Mk, et al.
Pubblicazione: (2022)
di: Bashar, Mk, et al.
Pubblicazione: (2022)
Galileo: Learning Global & Local Features of Many Remote Sensing Modalities
di: Tseng, Gabriel, et al.
Pubblicazione: (2025)
di: Tseng, Gabriel, et al.
Pubblicazione: (2025)
Many Dialects, Many Languages, One Cultural Lens: Evaluating Multilingual VLMs for Bengali Culture Understanding Across Historically Linked Languages and Regional Dialects
di: Sayeedi, Nurul Labib, et al.
Pubblicazione: (2026)
di: Sayeedi, Nurul Labib, et al.
Pubblicazione: (2026)
Unlocking the Potential of Grounding DINO in Videos: Parameter-Efficient Adaptation for Limited-Data Spatial-Temporal Localization
di: Wang, Zanyi, et al.
Pubblicazione: (2026)
di: Wang, Zanyi, et al.
Pubblicazione: (2026)
From CLIP to DINO: Visual Encoders Shout in Multi-modal Large Language Models
di: Jiang, Dongsheng, et al.
Pubblicazione: (2023)
di: Jiang, Dongsheng, et al.
Pubblicazione: (2023)
One Model, Many Behaviors: Training-Induced Effects on Out-of-Distribution Detection
di: Krumpl, Gerhard, et al.
Pubblicazione: (2026)
di: Krumpl, Gerhard, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Inteval Analysis for two spherical functions arising from robust Perspective-n-Lines problem
di: Zheng, Xiang, et al.
Pubblicazione: (2025) -
FlexGS: Train Once, Deploy Everywhere with Many-in-One Flexible 3D Gaussian Splatting
di: Liu, Hengyu, et al.
Pubblicazione: (2025) -
Many-for-Many: Unify the Training of Multiple Video and Image Generation and Manipulation Tasks
di: Li, Ruibin, et al.
Pubblicazione: (2025) -
DINO-Tok: Adapting DINO for Visual Tokenizers
di: Jia, Mingkai, et al.
Pubblicazione: (2025) -
DINO-Foresight: Looking into the Future with DINO
di: Karypidis, Efstathios, et al.
Pubblicazione: (2024)