More Clear, More Flexible, More Precise: A Comprehensive Oriented Object Detection benchmark for UAV
Fuente:
arXiv
Salvato in:
| Autori principali: | Ye, Kai, Tang, Haidi, Liu, Bowen, Dai, Pingyang, Cao, Liujuan, Ji, Rongrong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PartFormer: Awakening Latent Diverse Representation from Vision Transformer for Object Re-Identification
di: Tan, Lei, et al.
Pubblicazione: (2024)
di: Tan, Lei, et al.
Pubblicazione: (2024)
Understanding What Is Not Said:Referring Remote Sensing Image Segmentation with Scarce Expressions
di: Ye, Kai, et al.
Pubblicazione: (2025)
di: Ye, Kai, et al.
Pubblicazione: (2025)
S$^2$Teacher: Step-by-step Teacher for Sparsely Annotated Oriented Object Detection
di: Lin, Yu, et al.
Pubblicazione: (2025)
di: Lin, Yu, et al.
Pubblicazione: (2025)
Knowing Where to Focus: Attention-Guided Alignment for Text-based Person Search
di: Tan, Lei, et al.
Pubblicazione: (2024)
di: Tan, Lei, et al.
Pubblicazione: (2024)
Purifying, Labeling, and Utilizing: A High-Quality Pipeline for Small Object Detection
di: Wang, Siwei, et al.
Pubblicazione: (2025)
di: Wang, Siwei, et al.
Pubblicazione: (2025)
Evolving, Not Training: Zero-Shot Reasoning Segmentation via Evolutionary Prompting
di: Ye, Kai, et al.
Pubblicazione: (2025)
di: Ye, Kai, et al.
Pubblicazione: (2025)
Active-SAOOD: Active Sparsely Annotated Oriented Object Detection in Remote Sensing Images
di: Lin, Yu, et al.
Pubblicazione: (2026)
di: Lin, Yu, et al.
Pubblicazione: (2026)
More Pictures Say More: Visual Intersection Network for Open Set Object Detection
di: Dong, Bingcheng, et al.
Pubblicazione: (2024)
di: Dong, Bingcheng, et al.
Pubblicazione: (2024)
RIS-LAD: A Benchmark and Model for Referring Low-Altitude Drone Image Segmentation
di: Ye, Kai, et al.
Pubblicazione: (2025)
di: Ye, Kai, et al.
Pubblicazione: (2025)
FlexiReID: Adaptive Mixture of Expert for Multi-Modal Person Re-Identification
di: Sun, Zhen, et al.
Pubblicazione: (2025)
di: Sun, Zhen, et al.
Pubblicazione: (2025)
DPM++: Dynamic Masked Metric Learning for Occluded Person Re-identification
di: Tan, Lei, et al.
Pubblicazione: (2026)
di: Tan, Lei, et al.
Pubblicazione: (2026)
Inter2Former: Dynamic Hybrid Attention for Efficient High-Precision Interactive
di: Huang, You, et al.
Pubblicazione: (2025)
di: Huang, You, et al.
Pubblicazione: (2025)
HUWSOD: Holistic Self-training for Unified Weakly Supervised Object Detection
di: Cao, Liujuan, et al.
Pubblicazione: (2024)
di: Cao, Liujuan, et al.
Pubblicazione: (2024)
PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation
di: Ke, Shuyan, et al.
Pubblicazione: (2026)
di: Ke, Shuyan, et al.
Pubblicazione: (2026)
No More Sibling Rivalry: Debiasing Human-Object Interaction Detection
di: Yang, Bin, et al.
Pubblicazione: (2025)
di: Yang, Bin, et al.
Pubblicazione: (2025)
Attention Disturbance and Dual-Path Constraint Network for Occluded Person Re-identification
di: Xia, Jiaer, et al.
Pubblicazione: (2023)
di: Xia, Jiaer, et al.
Pubblicazione: (2023)
AccDiffusion v2: Towards More Accurate Higher-Resolution Diffusion Extrapolation
di: Lin, Zhihang, et al.
Pubblicazione: (2024)
di: Lin, Zhihang, et al.
Pubblicazione: (2024)
Can Unified Generation and Understanding Models Maintain Semantic Equivalence Across Different Output Modalities?
di: Jiang, Hongbo, et al.
Pubblicazione: (2026)
di: Jiang, Hongbo, et al.
Pubblicazione: (2026)
Adapted Center and Scale Prediction: More Stable and More Accurate
di: Wang, Wenhao, et al.
Pubblicazione: (2020)
di: Wang, Wenhao, et al.
Pubblicazione: (2020)
FocSAM: Delving Deeply into Focused Objects in Segmenting Anything
di: Huang, You, et al.
Pubblicazione: (2024)
di: Huang, You, et al.
Pubblicazione: (2024)
Alignment and Adversarial Robustness: Are More Human-Like Models More Secure?
di: Hoak, Blaine, et al.
Pubblicazione: (2025)
di: Hoak, Blaine, et al.
Pubblicazione: (2025)
Do More Details Always Introduce More Hallucinations in LVLM-based Image Captioning?
di: Feng, Mingqian, et al.
Pubblicazione: (2024)
di: Feng, Mingqian, et al.
Pubblicazione: (2024)
DiffusionFace: Towards a Comprehensive Dataset for Diffusion-Based Face Forgery Analysis
di: Chen, Zhongxi, et al.
Pubblicazione: (2024)
di: Chen, Zhongxi, et al.
Pubblicazione: (2024)
Unleashing MLLMs on the Edge: A Unified Framework for Cross-Modal ReID via Adaptive SVD Distillation
di: Jiang, Hongbo, et al.
Pubblicazione: (2026)
di: Jiang, Hongbo, et al.
Pubblicazione: (2026)
CamoTeacher: Dual-Rotation Consistency Learning for Semi-Supervised Camouflaged Object Detection
di: Lai, Xunfa, et al.
Pubblicazione: (2024)
di: Lai, Xunfa, et al.
Pubblicazione: (2024)
Less is More: Token Context-aware Learning for Object Tracking
di: Xu, Chenlong, et al.
Pubblicazione: (2025)
di: Xu, Chenlong, et al.
Pubblicazione: (2025)
Hypergraph Vision Transformers: Images are More than Nodes, More than Edges
di: Fixelle, Joshua
Pubblicazione: (2025)
di: Fixelle, Joshua
Pubblicazione: (2025)
The More You See in 2D, the More You Perceive in 3D
di: Han, Xinyang, et al.
Pubblicazione: (2024)
di: Han, Xinyang, et al.
Pubblicazione: (2024)
More Images, More Problems? A Controlled Analysis of VLM Failure Modes
di: Das, Anurag, et al.
Pubblicazione: (2026)
di: Das, Anurag, et al.
Pubblicazione: (2026)
Samba+: General and Accurate Salient Object Detection via A More Unified Mamba-based Framework
di: Zhao, Wenzhuo, et al.
Pubblicazione: (2026)
di: Zhao, Wenzhuo, et al.
Pubblicazione: (2026)
MAFE R-CNN: Selecting More Samples to Learn Category-aware Features for Small Object Detection
di: Li, Yichen, et al.
Pubblicazione: (2025)
di: Li, Yichen, et al.
Pubblicazione: (2025)
Unlearning the Noisy Correspondence Makes CLIP More Robust
di: Han, Haochen, et al.
Pubblicazione: (2025)
di: Han, Haochen, et al.
Pubblicazione: (2025)
Discriminative Consensus Mining with A Thousand Groups for More Accurate Co-Salient Object Detection
di: Zheng, Peng
Pubblicazione: (2024)
di: Zheng, Peng
Pubblicazione: (2024)
Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors
di: Wang, Xiangchen, et al.
Pubblicazione: (2025)
di: Wang, Xiangchen, et al.
Pubblicazione: (2025)
Less-to-More Generalization: Unlocking More Controllability by In-Context Generation
di: Wu, Shaojin, et al.
Pubblicazione: (2025)
di: Wu, Shaojin, et al.
Pubblicazione: (2025)
One-for-More: Continual Diffusion Model for Anomaly Detection
di: Li, Xiaofan, et al.
Pubblicazione: (2025)
di: Li, Xiaofan, et al.
Pubblicazione: (2025)
Floating No More: Object-Ground Reconstruction from a Single Image
di: Man, Yunze, et al.
Pubblicazione: (2024)
di: Man, Yunze, et al.
Pubblicazione: (2024)
Straightforward Layer-wise Pruning for More Efficient Visual Adaptation
di: Han, Ruizi, et al.
Pubblicazione: (2024)
di: Han, Ruizi, et al.
Pubblicazione: (2024)
FinMMR: Make Financial Numerical Reasoning More Multimodal, Comprehensive, and Challenging
di: Tang, Zichen, et al.
Pubblicazione: (2025)
di: Tang, Zichen, et al.
Pubblicazione: (2025)
Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification
di: Zhu, Lanyun, et al.
Pubblicazione: (2025)
di: Zhu, Lanyun, et al.
Pubblicazione: (2025)
Documenti analoghi
-
PartFormer: Awakening Latent Diverse Representation from Vision Transformer for Object Re-Identification
di: Tan, Lei, et al.
Pubblicazione: (2024) -
Understanding What Is Not Said:Referring Remote Sensing Image Segmentation with Scarce Expressions
di: Ye, Kai, et al.
Pubblicazione: (2025) -
S$^2$Teacher: Step-by-step Teacher for Sparsely Annotated Oriented Object Detection
di: Lin, Yu, et al.
Pubblicazione: (2025) -
Knowing Where to Focus: Attention-Guided Alignment for Text-based Person Search
di: Tan, Lei, et al.
Pubblicazione: (2024) -
Purifying, Labeling, and Utilizing: A High-Quality Pipeline for Small Object Detection
di: Wang, Siwei, et al.
Pubblicazione: (2025)