Heuristic-inspired Reasoning Priors Facilitate Data-Efficient Referring Object Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Xu, Chen, Zhe, Zhang, Jing, Tao, Dacheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LAB-Det: Language as a Domain-Invariant Bridge for Training-Free One-Shot Domain Generalization in Object Detection
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
SimDistill: Simulated Multi-modal Distillation for BEV 3D Object Detection
von: Zhao, Haimei, et al.
Veröffentlicht: (2023)
von: Zhao, Haimei, et al.
Veröffentlicht: (2023)
On Geometry-Enhanced Parameter-Efficient Fine-Tuning for 3D Scene Segmentation
von: Tang, Liyao, et al.
Veröffentlicht: (2025)
von: Tang, Liyao, et al.
Veröffentlicht: (2025)
Multi-Granularity Hand Action Detection
von: Zhe, Ting, et al.
Veröffentlicht: (2023)
von: Zhe, Ting, et al.
Veröffentlicht: (2023)
GoMatching++: Parameter- and Data-Efficient Arbitrary-Shaped Video Text Spotting and Benchmarking
von: He, Haibin, et al.
Veröffentlicht: (2025)
von: He, Haibin, et al.
Veröffentlicht: (2025)
Referring Camouflaged Object Detection
von: Zhang, Xuying, et al.
Veröffentlicht: (2023)
von: Zhang, Xuying, et al.
Veröffentlicht: (2023)
BSDP: Brain-inspired Streaming Dual-level Perturbations for Online Open World Object Detection
von: Chen, Yu, et al.
Veröffentlicht: (2024)
von: Chen, Yu, et al.
Veröffentlicht: (2024)
ClueAegis: Heuristic-to-Reasoning Cognitive-skill Learning for Unified Evidence-based Synthetic Image Detection
von: Cao, Huangsen, et al.
Veröffentlicht: (2026)
von: Cao, Huangsen, et al.
Veröffentlicht: (2026)
Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning
von: Jiang, Qing, et al.
Veröffentlicht: (2025)
von: Jiang, Qing, et al.
Veröffentlicht: (2025)
VideoTIR: Accurate Understanding for Long Videos with Efficient Tool-Integrated Reasoning
von: Gao, Zhe, et al.
Veröffentlicht: (2026)
von: Gao, Zhe, et al.
Veröffentlicht: (2026)
Reasoning-OCR: Can Large Multimodal Models Solve Complex Logical Reasoning Problems from OCR Cues?
von: He, Haibin, et al.
Veröffentlicht: (2025)
von: He, Haibin, et al.
Veröffentlicht: (2025)
Referring Camouflaged Object Detection With Multi-Context Overlapped Windows Cross-Attention
von: Wen, Yu, et al.
Veröffentlicht: (2025)
von: Wen, Yu, et al.
Veröffentlicht: (2025)
EventRR: Event Referential Reasoning for Referring Video Object Segmentation
von: Xu, Huihui, et al.
Veröffentlicht: (2025)
von: Xu, Huihui, et al.
Veröffentlicht: (2025)
FLORA: Formal Language Model Enables Robust Training-free Zero-shot Object Referring Analysis
von: Chen, Zhe, et al.
Veröffentlicht: (2025)
von: Chen, Zhe, et al.
Veröffentlicht: (2025)
MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation
von: Rong, Fu, et al.
Veröffentlicht: (2025)
von: Rong, Fu, et al.
Veröffentlicht: (2025)
Relation DETR: Exploring Explicit Position Relation Prior for Object Detection
von: Hou, Xiuquan, et al.
Veröffentlicht: (2024)
von: Hou, Xiuquan, et al.
Veröffentlicht: (2024)
HandRefiner: Refining Malformed Hands in Generated Images by Diffusion-based Conditional Inpainting
von: Lu, Wenquan, et al.
Veröffentlicht: (2023)
von: Lu, Wenquan, et al.
Veröffentlicht: (2023)
SVAC: Scaling Is All You Need For Referring Video Object Segmentation
von: Zhang, Li, et al.
Veröffentlicht: (2025)
von: Zhang, Li, et al.
Veröffentlicht: (2025)
Event-based Simultaneous Localization and Mapping: A Comprehensive Survey
von: Huang, Kunping, et al.
Veröffentlicht: (2023)
von: Huang, Kunping, et al.
Veröffentlicht: (2023)
OWL: Unsupervised 3D Object Detection by Occupancy Guided Warm-up and Large Model Priors Reasoning
von: Guo, Xusheng, et al.
Veröffentlicht: (2025)
von: Guo, Xusheng, et al.
Veröffentlicht: (2025)
DRIFT: Transferring Reasoning Priors for Efficient MLLM Fine-Tuning
von: Huang, Chao, et al.
Veröffentlicht: (2025)
von: Huang, Chao, et al.
Veröffentlicht: (2025)
On Robust Cross-View Consistency in Self-Supervised Monocular Depth Estimation
von: Zhao, Haimei, et al.
Veröffentlicht: (2022)
von: Zhao, Haimei, et al.
Veröffentlicht: (2022)
UniMix: Towards Domain Adaptive and Generalizable LiDAR Semantic Segmentation in Adverse Weather
von: Zhao, Haimei, et al.
Veröffentlicht: (2024)
von: Zhao, Haimei, et al.
Veröffentlicht: (2024)
Depth as Prior Knowledge for Object Detection
von: Sbeyti, Moussa Kassem, et al.
Veröffentlicht: (2026)
von: Sbeyti, Moussa Kassem, et al.
Veröffentlicht: (2026)
Multi-Object Sketch Animation with Grouping and Motion Trajectory Priors
von: Liang, Guotao, et al.
Veröffentlicht: (2025)
von: Liang, Guotao, et al.
Veröffentlicht: (2025)
Prototype Embedding Optimization for Human-Object Interaction Detection in Livestreaming
von: Zhang, Menghui, et al.
Veröffentlicht: (2025)
von: Zhang, Menghui, et al.
Veröffentlicht: (2025)
Improving Batch Normalization with TTA for Robust Object Detection in Self-Driving
von: Liao, Dacheng, et al.
Veröffentlicht: (2024)
von: Liao, Dacheng, et al.
Veröffentlicht: (2024)
Towards Modality-agnostic Label-efficient Segmentation with Entropy-Regularized Distribution Alignment
von: Tang, Liyao, et al.
Veröffentlicht: (2024)
von: Tang, Liyao, et al.
Veröffentlicht: (2024)
REVEAL: Reference-Grounded Reasoning for Multimodal Manipulation Detection
von: Zhou, Jun, et al.
Veröffentlicht: (2026)
von: Zhou, Jun, et al.
Veröffentlicht: (2026)
Echo-α: Large Agentic Multimodal Reasoning Model for Ultrasound Interpretation
von: Zhang, Jing, et al.
Veröffentlicht: (2026)
von: Zhang, Jing, et al.
Veröffentlicht: (2026)
Parameter-Efficient Semantic Augmentation for Enhancing Open-Vocabulary Object Detection
von: Cao, Weihao, et al.
Veröffentlicht: (2026)
von: Cao, Weihao, et al.
Veröffentlicht: (2026)
HOMEY: Heuristic Object Masking with Enhanced YOLO for Property Insurance Risk Detection
von: Panboonyuen, Teerapong
Veröffentlicht: (2026)
von: Panboonyuen, Teerapong
Veröffentlicht: (2026)
MotionTrack: Learning Motion Predictor for Multiple Object Tracking
von: Xiao, Changcheng, et al.
Veröffentlicht: (2023)
von: Xiao, Changcheng, et al.
Veröffentlicht: (2023)
Rethink Sparse Signals for Pose-guided Text-to-image Generation
von: Xuan, Wenjie, et al.
Veröffentlicht: (2025)
von: Xuan, Wenjie, et al.
Veröffentlicht: (2025)
RefSAM: Efficiently Adapting Segmenting Anything Model for Referring Video Object Segmentation
von: Li, Yonglin, et al.
Veröffentlicht: (2023)
von: Li, Yonglin, et al.
Veröffentlicht: (2023)
Free-Form Scene Editor: Enabling Multi-Round Object Manipulation like in a 3D Engine
von: Shuai, Xincheng, et al.
Veröffentlicht: (2025)
von: Shuai, Xincheng, et al.
Veröffentlicht: (2025)
PSDF: Prior-Driven Neural Implicit Surface Learning for Multi-view Reconstruction
von: Su, Wanjuan, et al.
Veröffentlicht: (2024)
von: Su, Wanjuan, et al.
Veröffentlicht: (2024)
Boosting Image Restoration via Priors from Pre-trained Models
von: Xu, Xiaogang, et al.
Veröffentlicht: (2024)
von: Xu, Xiaogang, et al.
Veröffentlicht: (2024)
Vision-Motion-Reference Alignment for Referring Multi-Object Tracking via Multi-Modal Large Language Models
von: Lv, Weiyi, et al.
Veröffentlicht: (2025)
von: Lv, Weiyi, et al.
Veröffentlicht: (2025)
Contact-aware Human Motion Generation from Textual Descriptions
von: Ma, Sihan, et al.
Veröffentlicht: (2024)
von: Ma, Sihan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
LAB-Det: Language as a Domain-Invariant Bridge for Training-Free One-Shot Domain Generalization in Object Detection
von: Zhang, Xu, et al.
Veröffentlicht: (2026) -
SimDistill: Simulated Multi-modal Distillation for BEV 3D Object Detection
von: Zhao, Haimei, et al.
Veröffentlicht: (2023) -
On Geometry-Enhanced Parameter-Efficient Fine-Tuning for 3D Scene Segmentation
von: Tang, Liyao, et al.
Veröffentlicht: (2025) -
Multi-Granularity Hand Action Detection
von: Zhe, Ting, et al.
Veröffentlicht: (2023) -
GoMatching++: Parameter- and Data-Efficient Arbitrary-Shaped Video Text Spotting and Benchmarking
von: He, Haibin, et al.
Veröffentlicht: (2025)