Chain of Visual Perception: Harnessing Multimodal Large Language Models for Zero-shot Camouflaged Object Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tang, Lv, Jiang, Peng-Tao, Shen, Zhihao, Zhang, Hao, Chen, Jinwei, Li, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating SAM2's Role in Camouflaged Object Detection: From SAM to SAM2
von: Tang, Lv, et al.
Veröffentlicht: (2024)
von: Tang, Lv, et al.
Veröffentlicht: (2024)
Scalable Visual State Space Model with Fractal Scanning
von: Tang, Lv, et al.
Veröffentlicht: (2024)
von: Tang, Lv, et al.
Veröffentlicht: (2024)
Synthetic-to-Real Camouflaged Object Detection
von: Luo, Zhihao, et al.
Veröffentlicht: (2025)
von: Luo, Zhihao, et al.
Veröffentlicht: (2025)
Zero-Shot Aerial Object Detection with Visual Description Regularization
von: Zang, Zhengqing, et al.
Veröffentlicht: (2024)
von: Zang, Zhengqing, et al.
Veröffentlicht: (2024)
Frequency Perception Network for Camouflaged Object Detection
von: Cong, Runmin, et al.
Veröffentlicht: (2023)
von: Cong, Runmin, et al.
Veröffentlicht: (2023)
Discover, Segment, and Select: A Progressive Mechanism for Zero-shot Camouflaged Object Segmentation
von: Yang, Yilong, et al.
Veröffentlicht: (2026)
von: Yang, Yilong, et al.
Veröffentlicht: (2026)
Empowering Segmentation Ability to Multi-modal Large Language Models
von: Yang, Yuqi, et al.
Veröffentlicht: (2024)
von: Yang, Yuqi, et al.
Veröffentlicht: (2024)
Exploring Deeper! Segment Anything Model with Depth Perception for Camouflaged Object Detection
von: Yu, Zhenni, et al.
Veröffentlicht: (2024)
von: Yu, Zhenni, et al.
Veröffentlicht: (2024)
Zero-Shot Co-salient Object Detection Framework
von: Xiao, Haoke, et al.
Veröffentlicht: (2023)
von: Xiao, Haoke, et al.
Veröffentlicht: (2023)
Conditional Polarization Guidance for Camouflaged Object Detection
von: Zhang, QIfan, et al.
Veröffentlicht: (2026)
von: Zhang, QIfan, et al.
Veröffentlicht: (2026)
GLCONet: Learning Multi-source Perception Representation for Camouflaged Object Detection
von: Sun, Yanguang, et al.
Veröffentlicht: (2024)
von: Sun, Yanguang, et al.
Veröffentlicht: (2024)
A Survey of Camouflaged Object Detection and Beyond
von: Xiao, Fengyang, et al.
Veröffentlicht: (2024)
von: Xiao, Fengyang, et al.
Veröffentlicht: (2024)
Towards Real Zero-Shot Camouflaged Object Segmentation without Camouflaged Annotations
von: Lei, Cheng, et al.
Veröffentlicht: (2024)
von: Lei, Cheng, et al.
Veröffentlicht: (2024)
FCL-COD: Weakly Supervised Camouflaged Object Detection with Frequency-aware and Contrastive Learning
von: Ni, Jingchen, et al.
Veröffentlicht: (2026)
von: Ni, Jingchen, et al.
Veröffentlicht: (2026)
Dr. Seg: Revisiting GRPO Training for Visual Large Language Models through Perception-Oriented Design
von: Sun, Haoxiang, et al.
Veröffentlicht: (2026)
von: Sun, Haoxiang, et al.
Veröffentlicht: (2026)
Referring Camouflaged Object Detection
von: Zhang, Xuying, et al.
Veröffentlicht: (2023)
von: Zhang, Xuying, et al.
Veröffentlicht: (2023)
Zero-shot Generalizable Incremental Learning for Vision-Language Object Detection
von: Deng, Jieren, et al.
Veröffentlicht: (2024)
von: Deng, Jieren, et al.
Veröffentlicht: (2024)
Mamba-based Spatio-Frequency Motion Perception for Video Camouflaged Object Detection
von: Li, Xin, et al.
Veröffentlicht: (2025)
von: Li, Xin, et al.
Veröffentlicht: (2025)
Harnessing Chain-of-Thought Reasoning in Multimodal Large Language Models for Face Anti-Spoofing
von: Zhang, Honglu, et al.
Veröffentlicht: (2025)
von: Zhang, Honglu, et al.
Veröffentlicht: (2025)
Modality-Agnostic Prompt Learning for Multi-Modal Camouflaged Object Detection
von: Wang, Hao, et al.
Veröffentlicht: (2026)
von: Wang, Hao, et al.
Veröffentlicht: (2026)
VMAD: Visual-enhanced Multimodal Large Language Model for Zero-Shot Anomaly Detection
von: Deng, Huilin, et al.
Veröffentlicht: (2024)
von: Deng, Huilin, et al.
Veröffentlicht: (2024)
Towards Accurate Camouflaged Object Detection with Mixture Convolution and Interactive Fusion
von: Chen, Geng, et al.
Veröffentlicht: (2021)
von: Chen, Geng, et al.
Veröffentlicht: (2021)
LLaVA-RadZ: Can Multimodal Large Language Models Effectively Tackle Zero-shot Radiology Recognition?
von: Li, Bangyan, et al.
Veröffentlicht: (2025)
von: Li, Bangyan, et al.
Veröffentlicht: (2025)
Benchmarking Vision-Language and Multimodal Large Language Models in Zero-shot and Few-shot Scenarios: A study on Christian Iconography
von: Spinaci, Gianmarco, et al.
Veröffentlicht: (2025)
von: Spinaci, Gianmarco, et al.
Veröffentlicht: (2025)
SFGNet: Semantic and Frequency Guided Network for Camouflaged Object Detection
von: Wang, Dezhen, et al.
Veröffentlicht: (2025)
von: Wang, Dezhen, et al.
Veröffentlicht: (2025)
Rethinking Detecting Salient and Camouflaged Objects in Unconstrained Scenes
von: Zhou, Zhangjun, et al.
Veröffentlicht: (2024)
von: Zhou, Zhangjun, et al.
Veröffentlicht: (2024)
CAMotion: A High-Quality Benchmark for Camouflaged Moving Object Detection in the Wild
von: Yao, Siyuan, et al.
Veröffentlicht: (2026)
von: Yao, Siyuan, et al.
Veröffentlicht: (2026)
VoroNav: Voronoi-based Zero-shot Object Navigation with Large Language Model
von: Wu, Pengying, et al.
Veröffentlicht: (2024)
von: Wu, Pengying, et al.
Veröffentlicht: (2024)
Visual Programming for Zero-shot Open-Vocabulary 3D Visual Grounding
von: Yuan, Zhihao, et al.
Veröffentlicht: (2023)
von: Yuan, Zhihao, et al.
Veröffentlicht: (2023)
Understanding Graphical Perception in Data Visualization through Zero-shot Prompting of Vision-Language Models
von: Guo, Grace, et al.
Veröffentlicht: (2024)
von: Guo, Grace, et al.
Veröffentlicht: (2024)
Towards Training-free Open-world Segmentation via Image Prompt Foundation Models
von: Tang, Lv, et al.
Veröffentlicht: (2023)
von: Tang, Lv, et al.
Veröffentlicht: (2023)
Green Video Camouflaged Object Detection
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
Retrospective Memory for Camouflaged Object Detection
von: Zhang, Chenxi, et al.
Veröffentlicht: (2025)
von: Zhang, Chenxi, et al.
Veröffentlicht: (2025)
Boundary-Guided Camouflaged Object Detection
von: Sun, Yujia, et al.
Veröffentlicht: (2022)
von: Sun, Yujia, et al.
Veröffentlicht: (2022)
MagicTryOn: Harnessing Diffusion Transformer for Garment-Preserving Video Virtual Try-on
von: Li, Guangyuan, et al.
Veröffentlicht: (2025)
von: Li, Guangyuan, et al.
Veröffentlicht: (2025)
Seamless Detection: Unifying Salient Object Detection and Camouflaged Object Detection
von: Liu, Yi, et al.
Veröffentlicht: (2024)
von: Liu, Yi, et al.
Veröffentlicht: (2024)
VisTa: Visual-contextual and Text-augmented Zero-shot Object-level OOD Detection
von: Zhang, Bin, et al.
Veröffentlicht: (2025)
von: Zhang, Bin, et al.
Veröffentlicht: (2025)
ZS-VCOS: Zero-Shot Video Camouflaged Object Segmentation By Optical Flow and Open Vocabulary Object Detection
von: Guo, Wenqi, et al.
Veröffentlicht: (2025)
von: Guo, Wenqi, et al.
Veröffentlicht: (2025)
Uncertainty-Masked Bernoulli Diffusion for Camouflaged Object Detection Refinement
von: Shen, Yuqi, et al.
Veröffentlicht: (2025)
von: Shen, Yuqi, et al.
Veröffentlicht: (2025)
Combining Knowledge Graph and LLMs for Enhanced Zero-shot Visual Question Answering
von: Tao, Qian, et al.
Veröffentlicht: (2025)
von: Tao, Qian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Evaluating SAM2's Role in Camouflaged Object Detection: From SAM to SAM2
von: Tang, Lv, et al.
Veröffentlicht: (2024) -
Scalable Visual State Space Model with Fractal Scanning
von: Tang, Lv, et al.
Veröffentlicht: (2024) -
Synthetic-to-Real Camouflaged Object Detection
von: Luo, Zhihao, et al.
Veröffentlicht: (2025) -
Zero-Shot Aerial Object Detection with Visual Description Regularization
von: Zang, Zhengqing, et al.
Veröffentlicht: (2024) -
Frequency Perception Network for Camouflaged Object Detection
von: Cong, Runmin, et al.
Veröffentlicht: (2023)