RT-OVAD: Real-Time Open-Vocabulary Aerial Object Detection via Image-Text Collaboration
Fuente:
arXiv
Guardado en:
| Autores principales: | Wei, Guoting, Yuan, Xia, Liu, Yu, Shang, Zhenhao, Xue, Xizhe, Wang, Peng, Yao, Kelu, Zhao, Chunxia, Zhang, Haokui, Xiao, Rong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
OS-W2S: An Automatic Labeling Engine for Language-Guided Open-Set Aerial Object Detection
por: Wei, Guoting, et al.
Publicado: (2025)
por: Wei, Guoting, et al.
Publicado: (2025)
Open-Text Aerial Detection: A Unified Framework For Aerial Visual Grounding And Detection
por: Wei, Guoting, et al.
Publicado: (2026)
por: Wei, Guoting, et al.
Publicado: (2026)
Open-Vocabulary Object Detection in UAV Imagery: A Review and Future Perspectives
por: Zhou, Yang, et al.
Publicado: (2025)
por: Zhou, Yang, et al.
Publicado: (2025)
Rethinking Practical and Efficient Quantization Calibration for Vision-Language Models
por: Shang, Zhenhao, et al.
Publicado: (2026)
por: Shang, Zhenhao, et al.
Publicado: (2026)
REO-VLM: Transforming VLM to Meet Regression Challenges in Earth Observation
por: Xue, Xizhe, et al.
Publicado: (2024)
por: Xue, Xizhe, et al.
Publicado: (2024)
Advancements in Visual Language Models for Remote Sensing: Datasets, Capabilities, and Enhancement Techniques
por: Tao, Lijie, et al.
Publicado: (2024)
por: Tao, Lijie, et al.
Publicado: (2024)
Language Embedding Meets Dynamic Graph: A New Exploration for Neural Architecture Representation Learning
por: Jing, Haizhao, et al.
Publicado: (2025)
por: Jing, Haizhao, et al.
Publicado: (2025)
VK-Det: Visual Knowledge Guided Prototype Learning for Open-Vocabulary Aerial Object Detection
por: Yao, Jianhang, et al.
Publicado: (2025)
por: Yao, Jianhang, et al.
Publicado: (2025)
Cross-View Open-Vocabulary Object Detection in Aerial Imagery
por: Kini, Jyoti, et al.
Publicado: (2025)
por: Kini, Jyoti, et al.
Publicado: (2025)
YOLO-World: Real-Time Open-Vocabulary Object Detection
por: Cheng, Tianheng, et al.
Publicado: (2024)
por: Cheng, Tianheng, et al.
Publicado: (2024)
Toward Open Vocabulary Aerial Object Detection with CLIP-Activated Student-Teacher Learning
por: Li, Yan, et al.
Publicado: (2023)
por: Li, Yan, et al.
Publicado: (2023)
BlabberSeg: Real-Time Embedded Open-Vocabulary Aerial Segmentation
por: Bong, Haechan Mark, et al.
Publicado: (2024)
por: Bong, Haechan Mark, et al.
Publicado: (2024)
3D-RCNet: Learning from Transformer to Build a 3D Relational ConvNet for Hyperspectral Image Classification
por: Jing, Haizhao, et al.
Publicado: (2024)
por: Jing, Haizhao, et al.
Publicado: (2024)
Exploiting Unlabeled Data with Multiple Expert Teachers for Open Vocabulary Aerial Object Detection and Its Orientation Adaptation
por: Li, Yan, et al.
Publicado: (2024)
por: Li, Yan, et al.
Publicado: (2024)
FOLK: Fast Open-Vocabulary 3D Instance Segmentation via Label-guided Knowledge Distillation
por: Wu, Hongrui, et al.
Publicado: (2025)
por: Wu, Hongrui, et al.
Publicado: (2025)
UVLM: Benchmarking Video Language Model for Underwater World Understanding
por: Xue, Xizhe, et al.
Publicado: (2025)
por: Xue, Xizhe, et al.
Publicado: (2025)
Collaborative Vision-Text Representation Optimizing for Open-Vocabulary Segmentation
por: Jiao, Siyu, et al.
Publicado: (2024)
por: Jiao, Siyu, et al.
Publicado: (2024)
RT-DETR++ for UAV Object Detection
por: Shufang, Yuan
Publicado: (2025)
por: Shufang, Yuan
Publicado: (2025)
Open Vocabulary Monocular 3D Object Detection
por: Yao, Jin, et al.
Publicado: (2024)
por: Yao, Jin, et al.
Publicado: (2024)
RTGen: Generating Region-Text Pairs for Open-Vocabulary Object Detection
por: Chen, Fangyi, et al.
Publicado: (2024)
por: Chen, Fangyi, et al.
Publicado: (2024)
Global-Local Collaborative Inference with LLM for Lidar-Based Open-Vocabulary Detection
por: Peng, Xingyu, et al.
Publicado: (2024)
por: Peng, Xingyu, et al.
Publicado: (2024)
Enhancing Maritime Object Detection in Real-Time with RT-DETR and Data Augmentation
por: Nemati, Nader
Publicado: (2025)
por: Nemati, Nader
Publicado: (2025)
Bridging Sensor Gaps via Attention Gated Tuning for Hyperspectral Image Classification
por: Xue, Xizhe, et al.
Publicado: (2023)
por: Xue, Xizhe, et al.
Publicado: (2023)
FBRT-YOLO: Faster and Better for Real-Time Aerial Image Detection
por: Xiao, Yao, et al.
Publicado: (2025)
por: Xiao, Yao, et al.
Publicado: (2025)
Scaling Open-Vocabulary Object Detection
por: Minderer, Matthias, et al.
Publicado: (2023)
por: Minderer, Matthias, et al.
Publicado: (2023)
GLRD: Global-Local Collaborative Reason and Debate with PSL for 3D Open-Vocabulary Detection
por: Peng, Xingyu, et al.
Publicado: (2025)
por: Peng, Xingyu, et al.
Publicado: (2025)
RT-APNN for Solving Gray Radiative Transfer Equations
por: Xie, Xizhe, et al.
Publicado: (2025)
por: Xie, Xizhe, et al.
Publicado: (2025)
Open-Vocabulary Segmentation with Unpaired Mask-Text Supervision
por: Wang, Zhaoqing, et al.
Publicado: (2024)
por: Wang, Zhaoqing, et al.
Publicado: (2024)
Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Object Detection
por: Lu, Yehao, et al.
Publicado: (2025)
por: Lu, Yehao, et al.
Publicado: (2025)
Learning to Detect and Segment for Open Vocabulary Object Detection
por: Wang, Tao, et al.
Publicado: (2022)
por: Wang, Tao, et al.
Publicado: (2022)
Retrieval-Augmented Open-Vocabulary Object Detection
por: Kim, Jooyeon, et al.
Publicado: (2024)
por: Kim, Jooyeon, et al.
Publicado: (2024)
CC-Pan: Channel-wise Compression based Diffusion for Efficient Pan-Sharpening
por: Li, Junjie, et al.
Publicado: (2026)
por: Li, Junjie, et al.
Publicado: (2026)
TCTGNet : A Real‐Time Object Detection Method for Dense Traffic Scenes in Aerial Photography
por: Zeyu Fang, et al.
Publicado: (2026)
por: Zeyu Fang, et al.
Publicado: (2026)
Bilateral Collaboration with Large Vision-Language Models for Open Vocabulary Human-Object Interaction Detection
por: Hu, Yupeng, et al.
Publicado: (2025)
por: Hu, Yupeng, et al.
Publicado: (2025)
Fast-SegSim: Real-Time Open-Vocabulary Segmentation for Robotics in Simulation
por: Yu, Xuan, et al.
Publicado: (2026)
por: Yu, Xuan, et al.
Publicado: (2026)
RT-DETRv3: Real-time End-to-End Object Detection with Hierarchical Dense Positive Supervision
por: Wang, Shuo, et al.
Publicado: (2024)
por: Wang, Shuo, et al.
Publicado: (2024)
RT-DETRv4: Painlessly Furthering Real-Time Object Detection with Vision Foundation Models
por: Liao, Zijun, et al.
Publicado: (2025)
por: Liao, Zijun, et al.
Publicado: (2025)
Open-Vocabulary Object Detection via Neighboring Region Attention Alignment
por: Qiang, Sunyuan, et al.
Publicado: (2024)
por: Qiang, Sunyuan, et al.
Publicado: (2024)
Towards Open-Vocabulary Multimodal 3D Object Detection with Attributes
por: Xiang, Xinhao, et al.
Publicado: (2025)
por: Xiang, Xinhao, et al.
Publicado: (2025)
ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection
por: Zhang, Yupeng, et al.
Publicado: (2025)
por: Zhang, Yupeng, et al.
Publicado: (2025)
Ejemplares similares
-
OS-W2S: An Automatic Labeling Engine for Language-Guided Open-Set Aerial Object Detection
por: Wei, Guoting, et al.
Publicado: (2025) -
Open-Text Aerial Detection: A Unified Framework For Aerial Visual Grounding And Detection
por: Wei, Guoting, et al.
Publicado: (2026) -
Open-Vocabulary Object Detection in UAV Imagery: A Review and Future Perspectives
por: Zhou, Yang, et al.
Publicado: (2025) -
Rethinking Practical and Efficient Quantization Calibration for Vision-Language Models
por: Shang, Zhenhao, et al.
Publicado: (2026) -
REO-VLM: Transforming VLM to Meet Regression Challenges in Earth Observation
por: Xue, Xizhe, et al.
Publicado: (2024)