VLM-PL: Advanced Pseudo Labeling Approach for Class Incremental Object Detection via Vision-Language Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Junsu, Ku, Yunhoe, Kim, Jihyeon, Cha, Junuk, Baek, Seungryul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can Synthetic Images Conquer Forgetting? Beyond Unexplored Doubts in Few-Shot Class-Incremental Learning
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
Beyond Synthetic Replays: Turning Diffusion Features into Few-Shot Class-Incremental Learning Knowledge
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
SDDGR: Stable Diffusion-based Deep Generative Replay for Class Incremental Object Detection
von: Kim, Junsu, et al.
Veröffentlicht: (2024)
von: Kim, Junsu, et al.
Veröffentlicht: (2024)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
CoT-Pose: Chain-of-Thought Reasoning for 3D Pose Generation from Abstract Prompts
von: Cha, Junuk, et al.
Veröffentlicht: (2025)
von: Cha, Junuk, et al.
Veröffentlicht: (2025)
OpenFS: Multi-Hand-Capable Fingerspelling Recognition with Implicit Signing-Hand Detection and Frame-Wise Letter-Conditioned Synthesis
von: Cha, Junuk, et al.
Veröffentlicht: (2026)
von: Cha, Junuk, et al.
Veröffentlicht: (2026)
EmoTalkingGaussian: Continuous Emotion-conditioned Talking Head Synthesis
von: Cha, Junuk, et al.
Veröffentlicht: (2025)
von: Cha, Junuk, et al.
Veröffentlicht: (2025)
BIGS: Bimanual Category-agnostic Interaction Reconstruction from Monocular Videos via 3D Gaussian Splatting
von: On, Jeongwan, et al.
Veröffentlicht: (2025)
von: On, Jeongwan, et al.
Veröffentlicht: (2025)
3D Reconstruction of Interacting Multi-Person in Clothing from a Single Image
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
B-RIGHT: Benchmark Re-evaluation for Integrity in Generalized Human-Object Interaction Testing
von: Jang, Yoojin, et al.
Veröffentlicht: (2025)
von: Jang, Yoojin, et al.
Veröffentlicht: (2025)
Revisiting Reliability in the Reasoning-based Pose Estimation Benchmark
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
1st Place Solution to the 8th HANDS Workshop Challenge -- ARCTIC Track: 3DGS-based Bimanual Category-agnostic Interaction Reconstruction
von: On, Jeongwan, et al.
Veröffentlicht: (2024)
von: On, Jeongwan, et al.
Veröffentlicht: (2024)
CoT-PL: Chain-of-Thought Pseudo-Labeling for Open-Vocabulary Object Detection
von: Choi, Hojun, et al.
Veröffentlicht: (2025)
von: Choi, Hojun, et al.
Veröffentlicht: (2025)
Leveraging 2D Masked Reconstruction for Domain Adaptation of 3D Pose Estimation
von: Park, Hansoo, et al.
Veröffentlicht: (2025)
von: Park, Hansoo, et al.
Veröffentlicht: (2025)
PLOT: Pseudo-Labeling via Video Object Tracking for Scalable Monocular 3D Object Detection
von: Lee, Seokyeong, et al.
Veröffentlicht: (2025)
von: Lee, Seokyeong, et al.
Veröffentlicht: (2025)
THOM: Generating Physically Plausible Hand-Object Meshes From Text
von: Jeong, Uyoung, et al.
Veröffentlicht: (2026)
von: Jeong, Uyoung, et al.
Veröffentlicht: (2026)
CASA: Class-Agnostic Shared Attributes in Vision-Language Models for Efficient Incremental Object Detection
von: Guo, Mingyi, et al.
Veröffentlicht: (2024)
von: Guo, Mingyi, et al.
Veröffentlicht: (2024)
Exploiting Style Latent Flows for Generalizing Deepfake Video Detection
von: Choi, Jongwook, et al.
Veröffentlicht: (2024)
von: Choi, Jongwook, et al.
Veröffentlicht: (2024)
Learning Adaptive Pseudo-Label Selection for Semi-Supervised 3D Object Detection
von: Kong, Taehun, et al.
Veröffentlicht: (2025)
von: Kong, Taehun, et al.
Veröffentlicht: (2025)
VLM-CPL: Consensus Pseudo Labels from Vision-Language Models for Annotation-Free Pathological Image Classification
von: Zhong, Lanfeng, et al.
Veröffentlicht: (2024)
von: Zhong, Lanfeng, et al.
Veröffentlicht: (2024)
Text2Relight: Creative Portrait Relighting with Text Guidance
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
GiPL: Generative augmented iterative Pseudo-Labeling for Cross-Domain Few-Shot Object Detection
von: Liu, Jiacong, et al.
Veröffentlicht: (2026)
von: Liu, Jiacong, et al.
Veröffentlicht: (2026)
BoIR: Box-Supervised Instance Representation for Multi-Person Pose Estimation
von: Jeong, Uyoung, et al.
Veröffentlicht: (2023)
von: Jeong, Uyoung, et al.
Veröffentlicht: (2023)
Pseudo-Labeling Driven Refinement of Benchmark Object Detection Datasets via Analysis of Learning Patterns
von: Kim, Min Je, et al.
Veröffentlicht: (2025)
von: Kim, Min Je, et al.
Veröffentlicht: (2025)
Improving Weakly-Supervised Object Localization Using Adversarial Erasing and Pseudo Label
von: Kang, Byeongkeun, et al.
Veröffentlicht: (2024)
von: Kang, Byeongkeun, et al.
Veröffentlicht: (2024)
VLM-HOI: Vision Language Models for Interpretable Human-Object Interaction Analysis
von: Kang, Donggoo, et al.
Veröffentlicht: (2024)
von: Kang, Donggoo, et al.
Veröffentlicht: (2024)
Towards Realistic Incremental Scenario in Class Incremental Semantic Segmentation
von: Kwak, Jihwan, et al.
Veröffentlicht: (2024)
von: Kwak, Jihwan, et al.
Veröffentlicht: (2024)
A 2-Stage Model for Vehicle Class and Orientation Detection with Photo-Realistic Image Generation
von: Kim, Youngmin, et al.
Veröffentlicht: (2025)
von: Kim, Youngmin, et al.
Veröffentlicht: (2025)
PoseBH: Prototypical Multi-Dataset Training Beyond Human Pose Estimation
von: Jeong, Uyoung, et al.
Veröffentlicht: (2025)
von: Jeong, Uyoung, et al.
Veröffentlicht: (2025)
Zero-shot Generalizable Incremental Learning for Vision-Language Object Detection
von: Deng, Jieren, et al.
Veröffentlicht: (2024)
von: Deng, Jieren, et al.
Veröffentlicht: (2024)
PL-FSCIL: Harnessing the Power of Prompts for Few-Shot Class-Incremental Learning
von: Tian, Songsong, et al.
Veröffentlicht: (2024)
von: Tian, Songsong, et al.
Veröffentlicht: (2024)
EO-VLM: VLM-Guided Energy Overload Attacks on Vision Models
von: Seo, Minjae, et al.
Veröffentlicht: (2025)
von: Seo, Minjae, et al.
Veröffentlicht: (2025)
Hierarchical Neural Collapse Detection Transformer for Class Incremental Object Detection
von: Pham, Duc Thanh, et al.
Veröffentlicht: (2025)
von: Pham, Duc Thanh, et al.
Veröffentlicht: (2025)
ProPL: Universal Semi-Supervised Ultrasound Image Segmentation via Prompt-Guided Pseudo-Labeling
von: Chen, Yaxiong, et al.
Veröffentlicht: (2025)
von: Chen, Yaxiong, et al.
Veröffentlicht: (2025)
Enhancing Source-Free Domain Adaptive Object Detection with Low-confidence Pseudo Label Distillation
von: Yoon, Ilhoon, et al.
Veröffentlicht: (2024)
von: Yoon, Ilhoon, et al.
Veröffentlicht: (2024)
Beyond Spatial Frequency: Pixel-wise Temporal Frequency-based Deepfake Video Detection
von: Kim, Taehoon, et al.
Veröffentlicht: (2025)
von: Kim, Taehoon, et al.
Veröffentlicht: (2025)
Unsupervised Incremental Learning Using Confidence-Based Pseudo-Labels
von: Rakotoarivony, Lucas
Veröffentlicht: (2025)
von: Rakotoarivony, Lucas
Veröffentlicht: (2025)
Incremental Pseudo-Labeling for Black-Box Unsupervised Domain Adaptation
von: Zou, Yawen, et al.
Veröffentlicht: (2024)
von: Zou, Yawen, et al.
Veröffentlicht: (2024)
Cross Pseudo Labeling For Weakly Supervised Video Anomaly Detection
von: Lee, Dayeon, et al.
Veröffentlicht: (2026)
von: Lee, Dayeon, et al.
Veröffentlicht: (2026)
ST-VLM: Kinematic Instruction Tuning for Spatio-Temporal Reasoning in Vision-Language Models
von: Ko, Dohwan, et al.
Veröffentlicht: (2025)
von: Ko, Dohwan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Can Synthetic Images Conquer Forgetting? Beyond Unexplored Doubts in Few-Shot Class-Incremental Learning
von: Kim, Junsu, et al.
Veröffentlicht: (2025) -
Beyond Synthetic Replays: Turning Diffusion Features into Few-Shot Class-Incremental Learning Knowledge
von: Kim, Junsu, et al.
Veröffentlicht: (2025) -
SDDGR: Stable Diffusion-based Deep Generative Replay for Class Incremental Object Detection
von: Kim, Junsu, et al.
Veröffentlicht: (2024) -
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
von: Cha, Junuk, et al.
Veröffentlicht: (2024) -
CoT-Pose: Chain-of-Thought Reasoning for 3D Pose Generation from Abstract Prompts
von: Cha, Junuk, et al.
Veröffentlicht: (2025)