SORCE: Small Object Retrieval in Complex Environments
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Chunxu, Xie, Chi, Chen, Xiaxu, Li, Wei, Zhu, Feng, Zhao, Rui, Wang, Limin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On the Suitability of Reinforcement Fine-Tuning to Visual Tasks
di: Chen, Xiaxu, et al.
Pubblicazione: (2025)
di: Chen, Xiaxu, et al.
Pubblicazione: (2025)
Reasoning Guided Embeddings: Leveraging MLLM Reasoning for Improved Multimodal Retrieval
di: Liu, Chunxu, et al.
Pubblicazione: (2025)
di: Liu, Chunxu, et al.
Pubblicazione: (2025)
Sparse Global Matching for Video Frame Interpolation with Large Motion
di: Liu, Chunxu, et al.
Pubblicazione: (2024)
di: Liu, Chunxu, et al.
Pubblicazione: (2024)
History-Aware Transformation of ReID Features for Multiple Object Tracking
di: Gao, Ruopeng, et al.
Pubblicazione: (2025)
di: Gao, Ruopeng, et al.
Pubblicazione: (2025)
FreeRet: MLLMs as Training-Free Retrievers
di: Zhu, Yuhan, et al.
Pubblicazione: (2025)
di: Zhu, Yuhan, et al.
Pubblicazione: (2025)
Described Object Detection: Liberating Object Detection with Flexible Expressions
di: Xie, Chi, et al.
Pubblicazione: (2023)
di: Xie, Chi, et al.
Pubblicazione: (2023)
On the Robustness of Human-Object Interaction Detection against Distribution Shift
di: Xie, Chi, et al.
Pubblicazione: (2025)
di: Xie, Chi, et al.
Pubblicazione: (2025)
VFIMamba: Video Frame Interpolation with State Space Models
di: Zhang, Guozhen, et al.
Pubblicazione: (2024)
di: Zhang, Guozhen, et al.
Pubblicazione: (2024)
ADUGS-VINS: Generalized Visual-Inertial Odometry for Robust Navigation in Highly Dynamic and Complex Environments
di: Zhou, Rui, et al.
Pubblicazione: (2024)
di: Zhou, Rui, et al.
Pubblicazione: (2024)
Few-Shot Incremental 3D Object Detection in Dynamic Indoor Environments
di: Zhu, Yun, et al.
Pubblicazione: (2026)
di: Zhu, Yun, et al.
Pubblicazione: (2026)
MeMOTR: Long-Term Memory-Augmented Transformer for Multi-Object Tracking
di: Gao, Ruopeng, et al.
Pubblicazione: (2023)
di: Gao, Ruopeng, et al.
Pubblicazione: (2023)
MotionRAG: Motion Retrieval-Augmented Image-to-Video Generation
di: Zhu, Chenhui, et al.
Pubblicazione: (2025)
di: Zhu, Chenhui, et al.
Pubblicazione: (2025)
AD-Det: Boosting Object Detection in UAV Images with Focused Small Objects and Balanced Tail Classes
di: Li, Zhenteng, et al.
Pubblicazione: (2025)
di: Li, Zhenteng, et al.
Pubblicazione: (2025)
Multiple Object Tracking as ID Prediction
di: Gao, Ruopeng, et al.
Pubblicazione: (2024)
di: Gao, Ruopeng, et al.
Pubblicazione: (2024)
RemDet: Rethinking Efficient Model Design for UAV Object Detection
di: Li, Chen, et al.
Pubblicazione: (2024)
di: Li, Chen, et al.
Pubblicazione: (2024)
StageInteractor: Query-based Object Detector with Cross-stage Interaction
di: Teng, Yao, et al.
Pubblicazione: (2023)
di: Teng, Yao, et al.
Pubblicazione: (2023)
Multi-modal Reference Learning for Fine-grained Text-to-Image Retrieval
di: Ma, Zehong, et al.
Pubblicazione: (2025)
di: Ma, Zehong, et al.
Pubblicazione: (2025)
CycleHOI: Improving Human-Object Interaction Detection with Cycle Consistency of Detection and Generation
di: Wang, Yisen, et al.
Pubblicazione: (2024)
di: Wang, Yisen, et al.
Pubblicazione: (2024)
Composed Object Retrieval: Object-level Retrieval via Composed Expressions
di: Wang, Tong, et al.
Pubblicazione: (2025)
di: Wang, Tong, et al.
Pubblicazione: (2025)
Find your Needle: Small Object Image Retrieval via Multi-Object Attention Optimization
di: Green, Michael, et al.
Pubblicazione: (2025)
di: Green, Michael, et al.
Pubblicazione: (2025)
Object Detectors in the Open Environment: Challenges, Solutions, and Outlook
di: Liang, Siyuan, et al.
Pubblicazione: (2024)
di: Liang, Siyuan, et al.
Pubblicazione: (2024)
EscapeCraft: A 3D Room Escape Environment for Benchmarking Complex Multimodal Reasoning Ability
di: Wang, Ziyue, et al.
Pubblicazione: (2025)
di: Wang, Ziyue, et al.
Pubblicazione: (2025)
Small Object Detection in Complex Backgrounds with Multi-Scale Attention and Global Relation Modeling
di: Tao, Wenguang, et al.
Pubblicazione: (2026)
di: Tao, Wenguang, et al.
Pubblicazione: (2026)
Effective Gaussian Management for High-fidelity Object Reconstruction
di: Liu, Jiateng, et al.
Pubblicazione: (2025)
di: Liu, Jiateng, et al.
Pubblicazione: (2025)
CT-CLIP: A Multi-modal Fusion Framework for Robust Apple Leaf Disease Recognition in Complex Environments
di: Liu, Lemin, et al.
Pubblicazione: (2025)
di: Liu, Lemin, et al.
Pubblicazione: (2025)
Asymmetric Masked Distillation for Pre-Training Small Foundation Models
di: Zhao, Zhiyu, et al.
Pubblicazione: (2023)
di: Zhao, Zhiyu, et al.
Pubblicazione: (2023)
WeDetect: Fast Open-Vocabulary Object Detection as Retrieval
di: Fu, Shenghao, et al.
Pubblicazione: (2025)
di: Fu, Shenghao, et al.
Pubblicazione: (2025)
ZeroI2V: Zero-Cost Adaptation of Pre-trained Transformers from Image to Video
di: Li, Xinhao, et al.
Pubblicazione: (2023)
di: Li, Xinhao, et al.
Pubblicazione: (2023)
Retrieval-Augmented Egocentric Video Captioning
di: Xu, Jilan, et al.
Pubblicazione: (2024)
di: Xu, Jilan, et al.
Pubblicazione: (2024)
SOEDiff: Efficient Distillation for Small Object Editing
di: Wu, Yiming, et al.
Pubblicazione: (2024)
di: Wu, Yiming, et al.
Pubblicazione: (2024)
Re-Aligning Language to Visual Objects with an Agentic Workflow
di: Chen, Yuming, et al.
Pubblicazione: (2025)
di: Chen, Yuming, et al.
Pubblicazione: (2025)
Reminding Multimodal Large Language Models of Object-aware Knowledge with Retrieved Tags
di: Qi, Daiqing, et al.
Pubblicazione: (2024)
di: Qi, Daiqing, et al.
Pubblicazione: (2024)
Boltzmann Attention Sampling for Image Analysis with Small Objects
di: Zhao, Theodore, et al.
Pubblicazione: (2025)
di: Zhao, Theodore, et al.
Pubblicazione: (2025)
GLRT-Based Metric Learning for Remote Sensing Object Retrieval
di: Zhang, Linping, et al.
Pubblicazione: (2024)
di: Zhang, Linping, et al.
Pubblicazione: (2024)
Joint Modeling of Feature, Correspondence, and a Compressed Memory for Video Object Segmentation
di: Zhang, Jiaming, et al.
Pubblicazione: (2023)
di: Zhang, Jiaming, et al.
Pubblicazione: (2023)
Object-Centric Framework for Video Moment Retrieval
di: Li, Zongyao, et al.
Pubblicazione: (2025)
di: Li, Zongyao, et al.
Pubblicazione: (2025)
Small Object Detection Model with Spatial Laplacian Pyramid Attention and Multi-Scale Features Enhancement in Aerial Images
di: Ji, Zhangjian, et al.
Pubblicazione: (2026)
di: Ji, Zhangjian, et al.
Pubblicazione: (2026)
Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval
di: Wang, Zhichuan, et al.
Pubblicazione: (2025)
di: Wang, Zhichuan, et al.
Pubblicazione: (2025)
WinDeskGround: A Benchmark for Robust GUI Grounding in Complex Multi-Window Desktop Environments
di: Zhao, Haoren, et al.
Pubblicazione: (2026)
di: Zhao, Haoren, et al.
Pubblicazione: (2026)
Exploiting Scale-Variant Attention for Segmenting Small Medical Objects
di: Dai, Wei, et al.
Pubblicazione: (2024)
di: Dai, Wei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
On the Suitability of Reinforcement Fine-Tuning to Visual Tasks
di: Chen, Xiaxu, et al.
Pubblicazione: (2025) -
Reasoning Guided Embeddings: Leveraging MLLM Reasoning for Improved Multimodal Retrieval
di: Liu, Chunxu, et al.
Pubblicazione: (2025) -
Sparse Global Matching for Video Frame Interpolation with Large Motion
di: Liu, Chunxu, et al.
Pubblicazione: (2024) -
History-Aware Transformation of ReID Features for Multiple Object Tracking
di: Gao, Ruopeng, et al.
Pubblicazione: (2025) -
FreeRet: MLLMs as Training-Free Retrievers
di: Zhu, Yuhan, et al.
Pubblicazione: (2025)