DART: An Automated End-to-End Object Detection Pipeline with Data Diversification, Open-Vocabulary Bounding Box Annotation, Pseudo-Label Review, and Model Training
Fuente:
arXiv
Saved in:
| Main Authors: | Xin, Chen, Hartel, Andreas, Kasneci, Enkelejda |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OCDet: Object Center Detection via Bounding Box-Aware Heatmap Prediction on Edge Devices with NPUs
by: Xin, Chen, et al.
Published: (2024)
by: Xin, Chen, et al.
Published: (2024)
OVTR: End-to-End Open-Vocabulary Multiple Object Tracking with Transformer
by: Li, Jinyang, et al.
Published: (2025)
by: Li, Jinyang, et al.
Published: (2025)
CoT-PL: Chain-of-Thought Pseudo-Labeling for Open-Vocabulary Object Detection
by: Choi, Hojun, et al.
Published: (2025)
by: Choi, Hojun, et al.
Published: (2025)
DART: Dual Adaptive Refinement Transfer for Open-Vocabulary Multi-Label Recognition
by: Liu, Haijing, et al.
Published: (2025)
by: Liu, Haijing, et al.
Published: (2025)
Leverage Cross-Attention for End-to-End Open-Vocabulary Panoptic Reconstruction
by: Yu, Xuan, et al.
Published: (2025)
by: Yu, Xuan, et al.
Published: (2025)
TEyeD: Over 20 million real-world eye images with Pupil, Eyelid, and Iris 2D and 3D Segmentations, 2D and 3D Landmarks, 3D Eyeball, Gaze Vector, and Eye Movement Types
by: Fuhl, Wolfgang, et al.
Published: (2021)
by: Fuhl, Wolfgang, et al.
Published: (2021)
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
by: Salzmann, Tim, et al.
Published: (2024)
by: Salzmann, Tim, et al.
Published: (2024)
Automated Visual Attention Detection using Mobile Eye Tracking in Behavioral Classroom Studies
by: Bozkir, Efe, et al.
Published: (2025)
by: Bozkir, Efe, et al.
Published: (2025)
Multi-Grid Redundant Bounding Box Annotation for Accurate Object Detection
by: Tesema, Solomon Negussie, et al.
Published: (2022)
by: Tesema, Solomon Negussie, et al.
Published: (2022)
OpenBox: Annotate Any Bounding Boxes in 3D
by: Lee, In-Jae, et al.
Published: (2025)
by: Lee, In-Jae, et al.
Published: (2025)
Taming Self-Training for Open-Vocabulary Object Detection
by: Zhao, Shiyu, et al.
Published: (2023)
by: Zhao, Shiyu, et al.
Published: (2023)
Faster Bounding Box Annotation for Object Detection in Indoor Scenes
by: Adhikari, Bishwo, et al.
Published: (2018)
by: Adhikari, Bishwo, et al.
Published: (2018)
BEEP3D: Box-Supervised End-to-End Pseudo-Mask Generation for 3D Instance Segmentation
by: Yoo, Youngju, et al.
Published: (2025)
by: Yoo, Youngju, et al.
Published: (2025)
Tracking by Detection and Query: An Efficient End-to-End Framework for Multi-Object Tracking
by: Jia, Shukun, et al.
Published: (2024)
by: Jia, Shukun, et al.
Published: (2024)
Future Success Prediction in Open-Vocabulary Object Manipulation Tasks Based on End-Effector Trajectories
by: Kambara, Motonari, et al.
Published: (2024)
by: Kambara, Motonari, et al.
Published: (2024)
ObjectVLA: End-to-End Open-World Object Manipulation Without Demonstration
by: Zhu, Minjie, et al.
Published: (2025)
by: Zhu, Minjie, et al.
Published: (2025)
DEYO: DETR with YOLO for End-to-End Object Detection
by: Ouyang, Haodong
Published: (2024)
by: Ouyang, Haodong
Published: (2024)
An Affordable, Wearable Stereo-Eye-Tracking Platform
by: Zimmer, Alexander, et al.
Published: (2026)
by: Zimmer, Alexander, et al.
Published: (2026)
Trade-offs in Privacy-Preserving Eye Tracking through Iris Obfuscation: A Benchmarking Study
by: Wang, Mengdi, et al.
Published: (2025)
by: Wang, Mengdi, et al.
Published: (2025)
Training an Open-Vocabulary Monocular 3D Object Detection Model without 3D Data
by: Huang, Rui, et al.
Published: (2024)
by: Huang, Rui, et al.
Published: (2024)
Online Pseudo-Label Unified Object Detection for Multiple Datasets Training
by: Tang, XiaoJun, et al.
Published: (2024)
by: Tang, XiaoJun, et al.
Published: (2024)
YOLOv10: Real-Time End-to-End Object Detection
by: Wang, Ao, et al.
Published: (2024)
by: Wang, Ao, et al.
Published: (2024)
SynSeg: Feature Synergy for Multi-Category Contrastive Learning in End-to-End Open-Vocabulary Semantic Segmentation
by: Zhang, Weichen, et al.
Published: (2025)
by: Zhang, Weichen, et al.
Published: (2025)
Training-free Boost for Open-Vocabulary Object Detection with Confidence Aggregation
by: Zheng, Yanhao, et al.
Published: (2024)
by: Zheng, Yanhao, et al.
Published: (2024)
Scaling Open-Vocabulary Object Detection
by: Minderer, Matthias, et al.
Published: (2023)
by: Minderer, Matthias, et al.
Published: (2023)
RQFormer: Rotated Query Transformer for End-to-End Oriented Object Detection
by: Zhao, Jiaqi, et al.
Published: (2023)
by: Zhao, Jiaqi, et al.
Published: (2023)
Collaborative Novel Object Discovery and Box-Guided Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Cao, Yang, et al.
Published: (2024)
by: Cao, Yang, et al.
Published: (2024)
Open-Vocabulary Scene Text Recognition via Pseudo-Image Labeling and Margin Loss
by: Ren, Xuhua, et al.
Published: (2024)
by: Ren, Xuhua, et al.
Published: (2024)
Night Eyes: A Reproducible Framework for Constellation-Based Corneal Reflection Matching
by: Maquiling, Virmarie, et al.
Published: (2026)
by: Maquiling, Virmarie, et al.
Published: (2026)
Cross-View Cross-Modal Unsupervised Domain Adaptation for Driver Monitoring System
by: Bhalla, Aditi, et al.
Published: (2025)
by: Bhalla, Aditi, et al.
Published: (2025)
Iris Style Transfer: Enhancing Iris Recognition with Style Features and Privacy Preservation through Neural Style Transfer
by: Wang, Mengdi, et al.
Published: (2025)
by: Wang, Mengdi, et al.
Published: (2025)
U-DECN: End-to-End Underwater Object Detection ConvNet with Improved DeNoising Training
by: Liu, Zhuoyan, et al.
Published: (2024)
by: Liu, Zhuoyan, et al.
Published: (2024)
DST-Det: Simple Dynamic Self-Training for Open-Vocabulary Object Detection
by: Xu, Shilin, et al.
Published: (2023)
by: Xu, Shilin, et al.
Published: (2023)
BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion
by: Lan, Yuqing, et al.
Published: (2025)
by: Lan, Yuqing, et al.
Published: (2025)
DMAT: An End-to-End Framework for Joint Atmospheric Turbulence Mitigation and Object Detection
by: Hill, Paul, et al.
Published: (2025)
by: Hill, Paul, et al.
Published: (2025)
Learning to Detect and Segment for Open Vocabulary Object Detection
by: Wang, Tao, et al.
Published: (2022)
by: Wang, Tao, et al.
Published: (2022)
Retrieval-Augmented Open-Vocabulary Object Detection
by: Kim, Jooyeon, et al.
Published: (2024)
by: Kim, Jooyeon, et al.
Published: (2024)
Faithful Attention Explainer: Verbalizing Decisions Based on Discriminative Features
by: Rong, Yao, et al.
Published: (2024)
by: Rong, Yao, et al.
Published: (2024)
Open-Vocabulary Object Detection in UAV Imagery: A Review and Future Perspectives
by: Zhou, Yang, et al.
Published: (2025)
by: Zhou, Yang, et al.
Published: (2025)
Learning Humanoid End-Effector Control for Open-Vocabulary Visual Loco-Manipulation
by: Dong, Runpei, et al.
Published: (2026)
by: Dong, Runpei, et al.
Published: (2026)
Similar Items
-
OCDet: Object Center Detection via Bounding Box-Aware Heatmap Prediction on Edge Devices with NPUs
by: Xin, Chen, et al.
Published: (2024) -
OVTR: End-to-End Open-Vocabulary Multiple Object Tracking with Transformer
by: Li, Jinyang, et al.
Published: (2025) -
CoT-PL: Chain-of-Thought Pseudo-Labeling for Open-Vocabulary Object Detection
by: Choi, Hojun, et al.
Published: (2025) -
DART: Dual Adaptive Refinement Transfer for Open-Vocabulary Multi-Label Recognition
by: Liu, Haijing, et al.
Published: (2025) -
Leverage Cross-Attention for End-to-End Open-Vocabulary Panoptic Reconstruction
by: Yu, Xuan, et al.
Published: (2025)