FlowOVD: Learning Generative Latent Flows for Zero-shot Open-vocabulary Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Yao, Cavallaro, Andrea, Oh, Changjae |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Open-vocabulary object 6D pose estimation
by: Corsetti, Jaime, et al.
Published: (2023)
by: Corsetti, Jaime, et al.
Published: (2023)
High-resolution open-vocabulary object 6D pose estimation
by: Corsetti, Jaime, et al.
Published: (2024)
by: Corsetti, Jaime, et al.
Published: (2024)
Sparse multi-view hand-object reconstruction for unseen environments
by: Pang, Yik Lung, et al.
Published: (2024)
by: Pang, Yik Lung, et al.
Published: (2024)
Improving Generalization of Language-Conditioned Robot Manipulation
by: Cui, Chenglin, et al.
Published: (2025)
by: Cui, Chenglin, et al.
Published: (2025)
Learning human-to-robot handovers through 3D scene reconstruction
by: Wu, Yuekun, et al.
Published: (2025)
by: Wu, Yuekun, et al.
Published: (2025)
The Detector Teaches Itself: Lightweight Self-Supervised Adaptation for Open-Vocabulary Object Detection
by: Wan, Yazhe, et al.
Published: (2026)
by: Wan, Yazhe, et al.
Published: (2026)
Toward Human-Robot Teaming: Learning Handover Behaviors from 3D Scenes
by: Wu, Yuekun, et al.
Published: (2025)
by: Wu, Yuekun, et al.
Published: (2025)
Stereo Hand-Object Reconstruction for Human-to-Robot Handover
by: Pang, Yik Lung, et al.
Published: (2024)
by: Pang, Yik Lung, et al.
Published: (2024)
NoOVD: Novel Category Discovery and Embedding for Open-Vocabulary Object Detection
by: Zhang, Yupeng, et al.
Published: (2026)
by: Zhang, Yupeng, et al.
Published: (2026)
Open-RGBT: Open-vocabulary RGB-T Zero-shot Semantic Segmentation in Open-world Environments
by: Yu, Meng, et al.
Published: (2024)
by: Yu, Meng, et al.
Published: (2024)
PanopticRecon: Leverage Open-vocabulary Instance Segmentation for Zero-shot Panoptic Reconstruction
by: Yu, Xuan, et al.
Published: (2024)
by: Yu, Xuan, et al.
Published: (2024)
DetCLIPv3: Towards Versatile Generative Open-vocabulary Object Detection
by: Yao, Lewei, et al.
Published: (2024)
by: Yao, Lewei, et al.
Published: (2024)
Learning by Erasing: Conditional Entropy based Transferable Out-Of-Distribution Detection
by: Xing, Meng, et al.
Published: (2022)
by: Xing, Meng, et al.
Published: (2022)
MarvelOVD: Marrying Object Recognition and Vision-Language Models for Robust Open-Vocabulary Object Detection
by: Wang, Kuo, et al.
Published: (2024)
by: Wang, Kuo, et al.
Published: (2024)
Zero-shot image privacy classification with Vision-Language Models
by: Baia, Alina Elena, et al.
Published: (2025)
by: Baia, Alina Elena, et al.
Published: (2025)
Chain-of-Caption: Training-free improvement of multimodal large language model on referring expression comprehension
by: Pang, Yik Lung, et al.
Published: (2026)
by: Pang, Yik Lung, et al.
Published: (2026)
Continual Learning in Open-vocabulary Classification with Complementary Memory Systems
by: Zhu, Zhen, et al.
Published: (2023)
by: Zhu, Zhen, et al.
Published: (2023)
Open-vocabulary vs. Closed-set: Best Practice for Few-shot Object Detection Considering Text Describability
by: Hosoya, Yusuke, et al.
Published: (2024)
by: Hosoya, Yusuke, et al.
Published: (2024)
Improving Image De-raining Using Reference-Guided Transformers
by: Ye, Zihao, et al.
Published: (2024)
by: Ye, Zihao, et al.
Published: (2024)
SIA-OVD: Shape-Invariant Adapter for Bridging the Image-Region Gap in Open-Vocabulary Detection
by: Wang, Zishuo, et al.
Published: (2024)
by: Wang, Zishuo, et al.
Published: (2024)
PrivLEX: Detecting legal concepts in images through Vision-Language Models
by: Baranouskaya, Darya, et al.
Published: (2026)
by: Baranouskaya, Darya, et al.
Published: (2026)
Bayesian Prompt Flow Learning for Zero-Shot Anomaly Detection
by: Qu, Zhen, et al.
Published: (2025)
by: Qu, Zhen, et al.
Published: (2025)
Compositional Caching for Training-free Open-vocabulary Attribute Detection
by: Garosi, Marco, et al.
Published: (2025)
by: Garosi, Marco, et al.
Published: (2025)
Diffusion-driven GAN Inversion for Multi-Modal Face Image Generation
by: Kim, Jihyun, et al.
Published: (2024)
by: Kim, Jihyun, et al.
Published: (2024)
Unified Framework for Open-World Compositional Zero-shot Learning
by: Jayasekara, Hirunima, et al.
Published: (2024)
by: Jayasekara, Hirunima, et al.
Published: (2024)
Manifold Induced Biases for Zero-shot and Few-shot Detection of Generated Images
by: Brokman, Jonathan, et al.
Published: (2025)
by: Brokman, Jonathan, et al.
Published: (2025)
FlowComposer: Composable Flows for Compositional Zero-Shot Learning
by: He, Zhenqi, et al.
Published: (2026)
by: He, Zhenqi, et al.
Published: (2026)
PSVMA+: Exploring Multi-granularity Semantic-visual Adaption for Generalized Zero-shot Learning
by: Liu, Man, et al.
Published: (2024)
by: Liu, Man, et al.
Published: (2024)
IGLOSS: Image Generation for Lidar Open-vocabulary Semantic Segmentation
by: Samet, Nermin, et al.
Published: (2026)
by: Samet, Nermin, et al.
Published: (2026)
OpenKD: Opening Prompt Diversity for Zero- and Few-shot Keypoint Detection
by: Lu, Changsheng, et al.
Published: (2024)
by: Lu, Changsheng, et al.
Published: (2024)
SHiNe: Semantic Hierarchy Nexus for Open-vocabulary Object Detection
by: Liu, Mingxuan, et al.
Published: (2024)
by: Liu, Mingxuan, et al.
Published: (2024)
OpenVIS: Open-vocabulary Video Instance Segmentation
by: Guo, Pinxue, et al.
Published: (2023)
by: Guo, Pinxue, et al.
Published: (2023)
DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing
by: Cai, Lingling, et al.
Published: (2025)
by: Cai, Lingling, et al.
Published: (2025)
Extremely Simple Out-of-distribution Detection for Audio-visual Generalized Zero-shot Learning
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
OMG: Towards Open-vocabulary Motion Generation via Mixture of Controllers
by: Liang, Han, et al.
Published: (2023)
by: Liang, Han, et al.
Published: (2023)
From Data to Modeling: Fully Open-vocabulary Scene Graph Generation
by: Chen, Zuyao, et al.
Published: (2025)
by: Chen, Zuyao, et al.
Published: (2025)
Exploiting Style Latent Flows for Generalizing Deepfake Video Detection
by: Choi, Jongwook, et al.
Published: (2024)
by: Choi, Jongwook, et al.
Published: (2024)
ChangeFlow -- Latent Rectified Flow for Change Detection in Remote Sensing
by: Rolih, Blaž, et al.
Published: (2026)
by: Rolih, Blaž, et al.
Published: (2026)
ZS-VCOS: Zero-Shot Video Camouflaged Object Segmentation By Optical Flow and Open Vocabulary Object Detection
by: Guo, Wenqi, et al.
Published: (2025)
by: Guo, Wenqi, et al.
Published: (2025)
Learning Clustering-based Prototypes for Compositional Zero-shot Learning
by: Qu, Hongyu, et al.
Published: (2025)
by: Qu, Hongyu, et al.
Published: (2025)
Similar Items
-
Open-vocabulary object 6D pose estimation
by: Corsetti, Jaime, et al.
Published: (2023) -
High-resolution open-vocabulary object 6D pose estimation
by: Corsetti, Jaime, et al.
Published: (2024) -
Sparse multi-view hand-object reconstruction for unseen environments
by: Pang, Yik Lung, et al.
Published: (2024) -
Improving Generalization of Language-Conditioned Robot Manipulation
by: Cui, Chenglin, et al.
Published: (2025) -
Learning human-to-robot handovers through 3D scene reconstruction
by: Wu, Yuekun, et al.
Published: (2025)