MR-GDINO: Efficient Open-World Continual Object Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dong, Bowen, Huang, Zitong, Yang, Guanglei, Zhang, Lei, Zuo, Wangmeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM
von: Dong, Bowen, et al.
Veröffentlicht: (2025)
von: Dong, Bowen, et al.
Veröffentlicht: (2025)
Segmenting Objectiveness and Task-awareness Unknown Region for Autonomous Driving
von: Zheng, Mi, et al.
Veröffentlicht: (2025)
von: Zheng, Mi, et al.
Veröffentlicht: (2025)
LPT++: Efficient Training on Mixture of Long-tailed Experts
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
MetricDepth: Enhancing Monocular Depth Estimation with Deep Metric Learning
von: Liu, Chunpu, et al.
Veröffentlicht: (2024)
von: Liu, Chunpu, et al.
Veröffentlicht: (2024)
IMWA: Iterative Model Weight Averaging Benefits Class-Imbalanced Learning Tasks
von: Huang, Zitong, et al.
Veröffentlicht: (2024)
von: Huang, Zitong, et al.
Veröffentlicht: (2024)
FILP-3D: Enhancing 3D Few-shot Class-incremental Learning with Pre-trained Vision-Language Models
von: Xu, Wan, et al.
Veröffentlicht: (2023)
von: Xu, Wan, et al.
Veröffentlicht: (2023)
Multi-Modality Driven LoRA for Adverse Condition Depth Estimation
von: Yang, Guanglei, et al.
Veröffentlicht: (2024)
von: Yang, Guanglei, et al.
Veröffentlicht: (2024)
Unprejudiced Training Auxiliary Tasks Makes Primary Better: A Multi-Task Learning Perspective
von: Li, Yuanze, et al.
Veröffentlicht: (2024)
von: Li, Yuanze, et al.
Veröffentlicht: (2024)
Grounding-MD: Grounded Video-language Pre-training for Open-World Moment Detection
von: Zhuang, Weijun, et al.
Veröffentlicht: (2025)
von: Zhuang, Weijun, et al.
Veröffentlicht: (2025)
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
von: Huang, Tianyu, et al.
Veröffentlicht: (2023)
von: Huang, Tianyu, et al.
Veröffentlicht: (2023)
Open-World Human-Object Interaction Detection via Multi-modal Prompts
von: Yang, Jie, et al.
Veröffentlicht: (2024)
von: Yang, Jie, et al.
Veröffentlicht: (2024)
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
von: Zou, Jian, et al.
Veröffentlicht: (2023)
von: Zou, Jian, et al.
Veröffentlicht: (2023)
PhysWorld: From Real Videos to World Models of Deformable Objects via Physics-Aware Demonstration Synthesis
von: Yang, Yu, et al.
Veröffentlicht: (2025)
von: Yang, Yu, et al.
Veröffentlicht: (2025)
Self-Supervised Learning for Real-World Super-Resolution from Dual and Multiple Zoomed Observations
von: Zhang, Zhilu, et al.
Veröffentlicht: (2024)
von: Zhang, Zhilu, et al.
Veröffentlicht: (2024)
FedSmoothLoRA: Toward Smoother and Faster Convergence in Federated Low-Rank Adaptation
von: Wang, Zehao, et al.
Veröffentlicht: (2026)
von: Wang, Zehao, et al.
Veröffentlicht: (2026)
NIR-Assisted Image Denoising: A Selective Fusion Approach and A Real-World Benchmark Dataset
von: Xu, Rongjian, et al.
Veröffentlicht: (2024)
von: Xu, Rongjian, et al.
Veröffentlicht: (2024)
Class Balance Matters to Active Class-Incremental Learning
von: Huang, Zitong, et al.
Veröffentlicht: (2024)
von: Huang, Zitong, et al.
Veröffentlicht: (2024)
Responsible Visual Editing
von: Ni, Minheng, et al.
Veröffentlicht: (2024)
von: Ni, Minheng, et al.
Veröffentlicht: (2024)
Visual-O1: Understanding Ambiguous Instructions via Multi-modal Multi-turn Chain-of-thoughts Reasoning
von: Ni, Minheng, et al.
Veröffentlicht: (2024)
von: Ni, Minheng, et al.
Veröffentlicht: (2024)
YOLO-UniOW: Efficient Universal Open-World Object Detection
von: Liu, Lihao, et al.
Veröffentlicht: (2024)
von: Liu, Lihao, et al.
Veröffentlicht: (2024)
CGL: Advancing Continual GUI Learning via Reinforcement Fine-Tuning
von: Yao, Zhenquan, et al.
Veröffentlicht: (2026)
von: Yao, Zhenquan, et al.
Veröffentlicht: (2026)
ArtHOI: Taming Foundation Models for Monocular 4D Reconstruction of Hand-Articulated-Object Interactions
von: Wang, Zikai, et al.
Veröffentlicht: (2026)
von: Wang, Zikai, et al.
Veröffentlicht: (2026)
Semi-supervised Open-World Object Detection
von: Mullappilly, Sahal Shaji, et al.
Veröffentlicht: (2024)
von: Mullappilly, Sahal Shaji, et al.
Veröffentlicht: (2024)
Open World Object Detection: A Survey
von: Li, Yiming, et al.
Veröffentlicht: (2024)
von: Li, Yiming, et al.
Veröffentlicht: (2024)
Beyond Flat Unknown Labels in Open-World Object Detection
von: Zhang, Yuchen, et al.
Veröffentlicht: (2025)
von: Zhang, Yuchen, et al.
Veröffentlicht: (2025)
Dual-Camera Smooth Zoom on Mobile Phones
von: Wu, Renlong, et al.
Veröffentlicht: (2024)
von: Wu, Renlong, et al.
Veröffentlicht: (2024)
Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models
von: Huang, Zitong, et al.
Veröffentlicht: (2026)
von: Huang, Zitong, et al.
Veröffentlicht: (2026)
DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding
von: Ren, Tianhe, et al.
Veröffentlicht: (2024)
von: Ren, Tianhe, et al.
Veröffentlicht: (2024)
OpenAD: Open-World Autonomous Driving Benchmark for 3D Object Detection
von: Xia, Zhongyu, et al.
Veröffentlicht: (2024)
von: Xia, Zhongyu, et al.
Veröffentlicht: (2024)
Detecting Unknown Objects via Energy-based Separation for Open World Object Detection
von: Heo, Jun-Woo, et al.
Veröffentlicht: (2026)
von: Heo, Jun-Woo, et al.
Veröffentlicht: (2026)
Parameter-Efficient Semantic Augmentation for Enhancing Open-Vocabulary Object Detection
von: Cao, Weihao, et al.
Veröffentlicht: (2026)
von: Cao, Weihao, et al.
Veröffentlicht: (2026)
Self-Supervised High Dynamic Range Imaging with Multi-Exposure Images in Dynamic Scenes
von: Zhang, Zhilu, et al.
Veröffentlicht: (2023)
von: Zhang, Zhilu, et al.
Veröffentlicht: (2023)
Lie Flow: Video Dynamic Fields Modeling and Predicting with Lie Algebra as Geometric Physics Principle
von: Qiao, Weidong, et al.
Veröffentlicht: (2026)
von: Qiao, Weidong, et al.
Veröffentlicht: (2026)
COVD: Continual Open-Vocabulary Object Detection with Novel Concept Injection
von: Zhang, Yupeng, et al.
Veröffentlicht: (2026)
von: Zhang, Yupeng, et al.
Veröffentlicht: (2026)
High-Quality Proposal Encoding and Cascade Denoising for Imaginary Supervised Object Detection
von: Chen, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Chen, Zhiyuan, et al.
Veröffentlicht: (2025)
Bridging Geometry-Coherent Text-to-3D Generation with Multi-View Diffusion Priors and Gaussian Splatting
von: Yang, Feng, et al.
Veröffentlicht: (2025)
von: Yang, Feng, et al.
Veröffentlicht: (2025)
YOLO-World: Real-Time Open-Vocabulary Object Detection
von: Cheng, Tianheng, et al.
Veröffentlicht: (2024)
von: Cheng, Tianheng, et al.
Veröffentlicht: (2024)
LLM-Guided Agentic Object Detection for Open-World Understanding
von: Mumcu, Furkan, et al.
Veröffentlicht: (2025)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2025)
Open-Vocabulary Camouflaged Object Segmentation
von: Pang, Youwei, et al.
Veröffentlicht: (2023)
von: Pang, Youwei, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
von: Dong, Bowen, et al.
Veröffentlicht: (2024) -
MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM
von: Dong, Bowen, et al.
Veröffentlicht: (2025) -
Segmenting Objectiveness and Task-awareness Unknown Region for Autonomous Driving
von: Zheng, Mi, et al.
Veröffentlicht: (2025) -
LPT++: Efficient Training on Mixture of Long-tailed Experts
von: Dong, Bowen, et al.
Veröffentlicht: (2024) -
MetricDepth: Enhancing Monocular Depth Estimation with Deep Metric Learning
von: Liu, Chunpu, et al.
Veröffentlicht: (2024)