CoT-PL: Chain-of-Thought Pseudo-Labeling for Open-Vocabulary Object Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Choi, Hojun, Lim, Youngsun, Shin, Jaeyo, Shim, Hyunjung |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Sampling Bag of Views for Open-Vocabulary Object Detection
di: Choi, Hojun, et al.
Pubblicazione: (2024)
di: Choi, Hojun, et al.
Pubblicazione: (2024)
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
di: Lim, Youngsun, et al.
Pubblicazione: (2024)
di: Lim, Youngsun, et al.
Pubblicazione: (2024)
Label-Augmented Dataset Distillation
di: Kang, Seoungyoon, et al.
Pubblicazione: (2024)
di: Kang, Seoungyoon, et al.
Pubblicazione: (2024)
Addressing Image Hallucination in Text-to-Image Generation through Factual Image Retrieval
di: Lim, Youngsun, et al.
Pubblicazione: (2024)
di: Lim, Youngsun, et al.
Pubblicazione: (2024)
Representation Alignment for Just Image Transformers is not Easier than You Think
di: Shin, Jaeyo, et al.
Pubblicazione: (2026)
di: Shin, Jaeyo, et al.
Pubblicazione: (2026)
DreamCatalyst: Fast and High-Quality 3D Editing via Controlling Editability and Identity Preservation
di: Kim, Jiwook, et al.
Pubblicazione: (2024)
di: Kim, Jiwook, et al.
Pubblicazione: (2024)
Understanding Multi-Granularity for Open-Vocabulary Part Segmentation
di: Choi, Jiho, et al.
Pubblicazione: (2024)
di: Choi, Jiho, et al.
Pubblicazione: (2024)
What "Not" to Detect: Negation-Aware VLMs via Structured Reasoning and Token Merging
di: Kang, Inha, et al.
Pubblicazione: (2025)
di: Kang, Inha, et al.
Pubblicazione: (2025)
Robust Driving QA through Metadata-Grounded Context and Task-Specific Prompts
di: Yu, Seungjun, et al.
Pubblicazione: (2025)
di: Yu, Seungjun, et al.
Pubblicazione: (2025)
Fine-Grained Image-Text Correspondence with Cost Aggregation for Open-Vocabulary Part Segmentation
di: Choi, Jiho, et al.
Pubblicazione: (2025)
di: Choi, Jiho, et al.
Pubblicazione: (2025)
VG-CoT: Towards Trustworthy Visual Reasoning via Grounded Chain-of-Thought
di: Lim, Byeonggeuk, et al.
Pubblicazione: (2026)
di: Lim, Byeonggeuk, et al.
Pubblicazione: (2026)
Self-Supervised Vision Transformers Are Efficient Segmentation Learners for Imperfect Labels
di: Lee, Seungho, et al.
Pubblicazione: (2024)
di: Lee, Seungho, et al.
Pubblicazione: (2024)
CoT-RVS: Zero-Shot Chain-of-Thought Reasoning Segmentation for Videos
di: Kao, Shiu-hong, et al.
Pubblicazione: (2025)
di: Kao, Shiu-hong, et al.
Pubblicazione: (2025)
CoT-Seg: Rethinking Segmentation with Chain-of-Thought Reasoning and Self-Correction
di: Kao, Shiu-hong, et al.
Pubblicazione: (2026)
di: Kao, Shiu-hong, et al.
Pubblicazione: (2026)
Decompose and Transfer: CoT-Prompting Enhanced Alignment for Open-Vocabulary Temporal Action Detection
di: Zhu, Sa, et al.
Pubblicazione: (2026)
di: Zhu, Sa, et al.
Pubblicazione: (2026)
CoT-Segmenter: Enhancing OOD Detection in Dense Road Scenes via Chain-of-Thought Reasoning
di: Song, Jeonghyo, et al.
Pubblicazione: (2025)
di: Song, Jeonghyo, et al.
Pubblicazione: (2025)
WaymoQA: A Multi-View Visual Question Answering Dataset for Safety-Critical Reasoning in Autonomous Driving
di: Yu, Seungjun, et al.
Pubblicazione: (2025)
di: Yu, Seungjun, et al.
Pubblicazione: (2025)
MINT-CoT: Enabling Interleaved Visual Tokens in Mathematical Chain-of-Thought Reasoning
di: Chen, Xinyan, et al.
Pubblicazione: (2025)
di: Chen, Xinyan, et al.
Pubblicazione: (2025)
VLM-PL: Advanced Pseudo Labeling Approach for Class Incremental Object Detection via Vision-Language Model
di: Kim, Junsu, et al.
Pubblicazione: (2024)
di: Kim, Junsu, et al.
Pubblicazione: (2024)
Scribble-Guided Diffusion for Training-free Text-to-Image Generation
di: Lee, Seonho, et al.
Pubblicazione: (2024)
di: Lee, Seonho, et al.
Pubblicazione: (2024)
ImageGen-CoT: Enhancing Text-to-Image In-context Learning with Chain-of-Thought Reasoning
di: Liao, Jiaqi, et al.
Pubblicazione: (2025)
di: Liao, Jiaqi, et al.
Pubblicazione: (2025)
Video-CoT: A Comprehensive Dataset for Spatiotemporal Understanding of Videos Based on Chain-of-Thought
di: Zhang, Shuyi, et al.
Pubblicazione: (2025)
di: Zhang, Shuyi, et al.
Pubblicazione: (2025)
AIM-CoT: Active Information-driven Multimodal Chain-of-Thought for Vision-Language Reasoning
di: Li, Xiping, et al.
Pubblicazione: (2025)
di: Li, Xiping, et al.
Pubblicazione: (2025)
CoT3DRef: Chain-of-Thoughts Data-Efficient 3D Visual Grounding
di: Abdelrahman, Eslam, et al.
Pubblicazione: (2023)
di: Abdelrahman, Eslam, et al.
Pubblicazione: (2023)
Memory-Efficient Fine-Tuning for Quantized Diffusion Model
di: Ryu, Hyogon, et al.
Pubblicazione: (2024)
di: Ryu, Hyogon, et al.
Pubblicazione: (2024)
GiPL: Generative augmented iterative Pseudo-Labeling for Cross-Domain Few-Shot Object Detection
di: Liu, Jiacong, et al.
Pubblicazione: (2026)
di: Liu, Jiacong, et al.
Pubblicazione: (2026)
CoT4Det: A Chain-of-Thought Framework for Perception-Oriented Vision-Language Tasks
di: Qi, Yu, et al.
Pubblicazione: (2025)
di: Qi, Yu, et al.
Pubblicazione: (2025)
X-CoT: Explainable Text-to-Video Retrieval via LLM-based Chain-of-Thought Reasoning
di: Pulakurthi, Prasanna Reddy, et al.
Pubblicazione: (2025)
di: Pulakurthi, Prasanna Reddy, et al.
Pubblicazione: (2025)
CoT-Pose: Chain-of-Thought Reasoning for 3D Pose Generation from Abstract Prompts
di: Cha, Junuk, et al.
Pubblicazione: (2025)
di: Cha, Junuk, et al.
Pubblicazione: (2025)
Robust Object Detection with Pseudo Labels from VLMs using Per-Object Co-teaching
di: Bhaskar, Uday, et al.
Pubblicazione: (2025)
di: Bhaskar, Uday, et al.
Pubblicazione: (2025)
Scaling Open-Vocabulary Object Detection
di: Minderer, Matthias, et al.
Pubblicazione: (2023)
di: Minderer, Matthias, et al.
Pubblicazione: (2023)
SUPER-AD: Semantic Uncertainty-aware Planning for End-to-End Robust Autonomous Driving
di: Ryu, Wonjeong, et al.
Pubblicazione: (2025)
di: Ryu, Wonjeong, et al.
Pubblicazione: (2025)
C-CoT: Counterfactual Chain-of-Thought with Vision-Language Models for Safe Autonomous Driving
di: Tian, Kefei, et al.
Pubblicazione: (2026)
di: Tian, Kefei, et al.
Pubblicazione: (2026)
Rethinking Direct Preference Optimization in Diffusion Models
di: Kang, Junyong, et al.
Pubblicazione: (2025)
di: Kang, Junyong, et al.
Pubblicazione: (2025)
Open-Vocabulary Scene Text Recognition via Pseudo-Image Labeling and Margin Loss
di: Ren, Xuhua, et al.
Pubblicazione: (2024)
di: Ren, Xuhua, et al.
Pubblicazione: (2024)
CIR-CoT: Towards Interpretable Composed Image Retrieval via End-to-End Chain-of-Thought Reasoning
di: Lin, Weihuang, et al.
Pubblicazione: (2025)
di: Lin, Weihuang, et al.
Pubblicazione: (2025)
Learning to Detect and Segment for Open Vocabulary Object Detection
di: Wang, Tao, et al.
Pubblicazione: (2022)
di: Wang, Tao, et al.
Pubblicazione: (2022)
Retrieval-Augmented Open-Vocabulary Object Detection
di: Kim, Jooyeon, et al.
Pubblicazione: (2024)
di: Kim, Jooyeon, et al.
Pubblicazione: (2024)
DeCo-DETR: Decoupled Cognition DETR for efficient Open-Vocabulary Object Detection
di: Wang, Siheng, et al.
Pubblicazione: (2026)
di: Wang, Siheng, et al.
Pubblicazione: (2026)
ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection
di: Zhang, Yupeng, et al.
Pubblicazione: (2025)
di: Zhang, Yupeng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Sampling Bag of Views for Open-Vocabulary Object Detection
di: Choi, Hojun, et al.
Pubblicazione: (2024) -
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
di: Lim, Youngsun, et al.
Pubblicazione: (2024) -
Label-Augmented Dataset Distillation
di: Kang, Seoungyoon, et al.
Pubblicazione: (2024) -
Addressing Image Hallucination in Text-to-Image Generation through Factual Image Retrieval
di: Lim, Youngsun, et al.
Pubblicazione: (2024) -
Representation Alignment for Just Image Transformers is not Easier than You Think
di: Shin, Jaeyo, et al.
Pubblicazione: (2026)