LLM-Guided Agentic Object Detection for Open-World Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mumcu, Furkan, Jones, Michael J., Cherian, Anoop, Yilmaz, Yasin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Leveraging Multimodal LLM Descriptions of Activity for Explainable Semi-Supervised Video Anomaly Detection
von: Mumcu, Furkan, et al.
Veröffentlicht: (2025)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2025)
Is Video Anomaly Detection Misframed? Evidence from LLM-Based and Multi-Scene Models
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
ComplexVAD: Detecting Interaction Anomalies in Video
von: Mumcu, Furkan, et al.
Veröffentlicht: (2025)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2025)
Universal and Efficient Detection of Adversarial Data through Nonuniform Impact on Network Layers
von: Mumcu, Furkan, et al.
Veröffentlicht: (2025)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2025)
Multimodal Attack Detection for Action Recognition Models
von: Mumcu, Furkan, et al.
Veröffentlicht: (2024)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2024)
Improving Open-World Object Localization by Discovering Background
von: Singh, Ashish, et al.
Veröffentlicht: (2025)
von: Singh, Ashish, et al.
Veröffentlicht: (2025)
Agentic AI-Empowered Dynamic Survey Framework
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
MMHOI: Modeling Complex 3D Multi-Human Multi-Object Interactions
von: Kogashi, Kaen, et al.
Veröffentlicht: (2025)
von: Kogashi, Kaen, et al.
Veröffentlicht: (2025)
QVAD: A Question-Centric Agentic Framework for Efficient and Training-Free Video Anomaly Detection
von: Bekit, Lokman, et al.
Veröffentlicht: (2026)
von: Bekit, Lokman, et al.
Veröffentlicht: (2026)
Socially-Weighted Alignment: A Game-Theoretic Framework for Multi-Agent LLM Systems
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
Robustness of Agentic AI Systems via Adversarially-Aligned Jacobian Regularization
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding
von: Ren, Tianhe, et al.
Veröffentlicht: (2024)
von: Ren, Tianhe, et al.
Veröffentlicht: (2024)
Semi-supervised Open-World Object Detection
von: Mullappilly, Sahal Shaji, et al.
Veröffentlicht: (2024)
von: Mullappilly, Sahal Shaji, et al.
Veröffentlicht: (2024)
Open World Object Detection: A Survey
von: Li, Yiming, et al.
Veröffentlicht: (2024)
von: Li, Yiming, et al.
Veröffentlicht: (2024)
VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection
von: Liu, Chih-Chung, et al.
Veröffentlicht: (2026)
von: Liu, Chih-Chung, et al.
Veröffentlicht: (2026)
Detecting Adversarial Examples
von: Mumcu, Furkan, et al.
Veröffentlicht: (2024)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2024)
Detecting Unknown Objects via Energy-based Separation for Open World Object Detection
von: Heo, Jun-Woo, et al.
Veröffentlicht: (2026)
von: Heo, Jun-Woo, et al.
Veröffentlicht: (2026)
Towards Open-Vocabulary Multimodal 3D Object Detection with Attributes
von: Xiang, Xinhao, et al.
Veröffentlicht: (2025)
von: Xiang, Xinhao, et al.
Veröffentlicht: (2025)
Manual-PA: Learning 3D Part Assembly from Instruction Diagrams
von: Zhang, Jiahao, et al.
Veröffentlicht: (2024)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2024)
Detecting Adversarial Data via Provable Adversarial Noise Amplification
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
Beyond Flat Unknown Labels in Open-World Object Detection
von: Zhang, Yuchen, et al.
Veröffentlicht: (2025)
von: Zhang, Yuchen, et al.
Veröffentlicht: (2025)
MR-GDINO: Efficient Open-World Continual Object Detection
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
YOLO-World: Real-Time Open-Vocabulary Object Detection
von: Cheng, Tianheng, et al.
Veröffentlicht: (2024)
von: Cheng, Tianheng, et al.
Veröffentlicht: (2024)
A Hierarchical Benchmark of Foundation Models for Dermatology
von: Yuceyalcin, Furkan, et al.
Veröffentlicht: (2026)
von: Yuceyalcin, Furkan, et al.
Veröffentlicht: (2026)
OpenAD: Open-World Autonomous Driving Benchmark for 3D Object Detection
von: Xia, Zhongyu, et al.
Veröffentlicht: (2024)
von: Xia, Zhongyu, et al.
Veröffentlicht: (2024)
AssemblyBench: Physics-Aware Assembly of Complex Industrial Objects
von: Li, Danrui, et al.
Veröffentlicht: (2026)
von: Li, Danrui, et al.
Veröffentlicht: (2026)
Beyond Literal Descriptions: Understanding and Locating Open-World Objects Aligned with Human Intentions
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
YOLO-UniOW: Efficient Universal Open-World Object Detection
von: Liu, Lihao, et al.
Veröffentlicht: (2024)
von: Liu, Lihao, et al.
Veröffentlicht: (2024)
Looking Beyond the Known: Towards a Data Discovery Guided Open-World Object Detection
von: Majee, Anay, et al.
Veröffentlicht: (2025)
von: Majee, Anay, et al.
Veröffentlicht: (2025)
Decoupled PROB: Decoupled Query Initialization Tasks and Objectness-Class Learning for Open World Object Detection
von: Inoue, Riku, et al.
Veröffentlicht: (2025)
von: Inoue, Riku, et al.
Veröffentlicht: (2025)
Open-World Human-Object Interaction Detection via Multi-modal Prompts
von: Yang, Jie, et al.
Veröffentlicht: (2024)
von: Yang, Jie, et al.
Veröffentlicht: (2024)
Towards Agentic AI for Multimodal-Guided Video Object Segmentation
von: Tran, Tuyen, et al.
Veröffentlicht: (2025)
von: Tran, Tuyen, et al.
Veröffentlicht: (2025)
Generalized Open-World Semi-Supervised Object Detection
von: Allabadi, Garvita, et al.
Veröffentlicht: (2023)
von: Allabadi, Garvita, et al.
Veröffentlicht: (2023)
GUIDED: Granular Understanding via Identification, Detection, and Discrimination for Fine-Grained Open-Vocabulary Object Detection
von: Li, Jiaming, et al.
Veröffentlicht: (2026)
von: Li, Jiaming, et al.
Veröffentlicht: (2026)
Aligning Step-by-Step Instructional Diagrams to Video Demonstrations
von: Zhang, Jiahao, et al.
Veröffentlicht: (2023)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2023)
OW-Rep: Open World Object Detection with Instance Representation Learning
von: Lee, Sunoh, et al.
Veröffentlicht: (2024)
von: Lee, Sunoh, et al.
Veröffentlicht: (2024)
SpinBench: Perspective and Rotation as a Lens on Spatial Reasoning in VLMs
von: Zhang, Yuyou, et al.
Veröffentlicht: (2025)
von: Zhang, Yuyou, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Open-Vocabulary Object Detection
von: Kim, Jooyeon, et al.
Veröffentlicht: (2024)
von: Kim, Jooyeon, et al.
Veröffentlicht: (2024)
Collaborative Novel Object Discovery and Box-Guided Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
von: Cao, Yang, et al.
Veröffentlicht: (2024)
von: Cao, Yang, et al.
Veröffentlicht: (2024)
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion
von: Lin, Zhiwei, et al.
Veröffentlicht: (2025)
von: Lin, Zhiwei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Leveraging Multimodal LLM Descriptions of Activity for Explainable Semi-Supervised Video Anomaly Detection
von: Mumcu, Furkan, et al.
Veröffentlicht: (2025) -
Is Video Anomaly Detection Misframed? Evidence from LLM-Based and Multi-Scene Models
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026) -
ComplexVAD: Detecting Interaction Anomalies in Video
von: Mumcu, Furkan, et al.
Veröffentlicht: (2025) -
Universal and Efficient Detection of Adversarial Data through Nonuniform Impact on Network Layers
von: Mumcu, Furkan, et al.
Veröffentlicht: (2025) -
Multimodal Attack Detection for Action Recognition Models
von: Mumcu, Furkan, et al.
Veröffentlicht: (2024)