Collaborative Vision-Text Representation Optimizing for Open-Vocabulary Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Jiao, Siyu, Zhu, Hongguang, Huang, Jiannan, Zhao, Yao, Wei, Yunchao, Shi, Humphrey |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Collaborative Feature-Logits Contrastive Learning for Open-Set Semi-Supervised Object Detection
by: Zhong, Xinhao, et al.
Published: (2024)
by: Zhong, Xinhao, et al.
Published: (2024)
CLIP-GS: Unifying Vision-Language Representation with 3D Gaussian Splatting
by: Jiao, Siyu, et al.
Published: (2024)
by: Jiao, Siyu, et al.
Published: (2024)
SAGE: Exploring the Boundaries of Unsafe Concept Domain with Semantic-Augment Erasing
by: Zhu, Hongguang, et al.
Published: (2025)
by: Zhu, Hongguang, et al.
Published: (2025)
Le-DETR: Revisiting Real-Time Detection Transformer with Efficient Encoder Design
by: Huang, Jiannan, et al.
Published: (2026)
by: Huang, Jiannan, et al.
Published: (2026)
Transferable and Principled Efficiency for Open-Vocabulary Segmentation
by: Xu, Jingxuan, et al.
Published: (2024)
by: Xu, Jingxuan, et al.
Published: (2024)
ClassDiffusion: More Aligned Personalization Tuning with Explicit Class Guidance
by: Huang, Jiannan, et al.
Published: (2024)
by: Huang, Jiannan, et al.
Published: (2024)
FlexVAR: Flexible Visual Autoregressive Modeling without Residual Prediction
by: Jiao, Siyu, et al.
Published: (2025)
by: Jiao, Siyu, et al.
Published: (2025)
COS3D: Collaborative Open-Vocabulary 3D Segmentation
by: Zhu, Runsong, et al.
Published: (2025)
by: Zhu, Runsong, et al.
Published: (2025)
Diffusion for Natural Image Matting
by: Hu, Yihan, et al.
Published: (2023)
by: Hu, Yihan, et al.
Published: (2023)
BGM: Background Mixup for X-ray Prohibited Items Detection
by: Liu, Weizhe, et al.
Published: (2024)
by: Liu, Weizhe, et al.
Published: (2024)
Learning Trimaps via Clicks for Image Matting
by: Zhang, Chenyi, et al.
Published: (2024)
by: Zhang, Chenyi, et al.
Published: (2024)
Open-Vocabulary Camouflaged Object Segmentation with Cascaded Vision Language Models
by: Zhao, Kai, et al.
Published: (2025)
by: Zhao, Kai, et al.
Published: (2025)
Decomposed Vision-Language Alignment for Fine-Grained Open-Vocabulary Segmentation
by: Wang, Chenhao, et al.
Published: (2026)
by: Wang, Chenhao, et al.
Published: (2026)
Open-Vocabulary 3D Semantic Segmentation with Text-to-Image Diffusion Models
by: Zhu, Xiaoyu, et al.
Published: (2024)
by: Zhu, Xiaoyu, et al.
Published: (2024)
RT-OVAD: Real-Time Open-Vocabulary Aerial Object Detection via Image-Text Collaboration
by: Wei, Guoting, et al.
Published: (2024)
by: Wei, Guoting, et al.
Published: (2024)
Harnessing Vision Foundation Models for High-Performance, Training-Free Open Vocabulary Segmentation
by: Shi, Yuheng, et al.
Published: (2024)
by: Shi, Yuheng, et al.
Published: (2024)
FGAseg: Fine-Grained Pixel-Text Alignment for Open-Vocabulary Semantic Segmentation
by: Li, Bingyu, et al.
Published: (2025)
by: Li, Bingyu, et al.
Published: (2025)
ROSE: Revolutionizing Open-Set Dense Segmentation with Patch-Wise Perceptual Large Multimodal Model
by: Han, Kunyang, et al.
Published: (2024)
by: Han, Kunyang, et al.
Published: (2024)
Open-Vocabulary Camouflaged Object Segmentation
by: Pang, Youwei, et al.
Published: (2023)
by: Pang, Youwei, et al.
Published: (2023)
PAI-Bench: A Comprehensive Benchmark For Physical AI
by: Zhou, Fengzhe, et al.
Published: (2025)
by: Zhou, Fengzhe, et al.
Published: (2025)
Adapting Vision-Language Model with Fine-grained Semantics for Open-Vocabulary Segmentation
by: Chng, Yong Xien, et al.
Published: (2024)
by: Chng, Yong Xien, et al.
Published: (2024)
FrozenSeg: Harmonizing Frozen Foundation Models for Open-Vocabulary Segmentation
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
IPSeg: Image Posterior Mitigates Semantic Drift in Class-Incremental Segmentation
by: Yu, Xiao, et al.
Published: (2025)
by: Yu, Xiao, et al.
Published: (2025)
Renovating Names in Open-Vocabulary Segmentation Benchmarks
by: Huang, Haiwen, et al.
Published: (2024)
by: Huang, Haiwen, et al.
Published: (2024)
Bilateral Collaboration with Large Vision-Language Models for Open Vocabulary Human-Object Interaction Detection
by: Hu, Yupeng, et al.
Published: (2025)
by: Hu, Yupeng, et al.
Published: (2025)
Leveraging Depth and Language for Open-Vocabulary Domain-Generalized Semantic Segmentation
by: Chen, Siyu, et al.
Published: (2025)
by: Chen, Siyu, et al.
Published: (2025)
Open-Vocabulary Object Detection with Meta Prompt Representation and Instance Contrastive Optimization
by: Wang, Zhao, et al.
Published: (2024)
by: Wang, Zhao, et al.
Published: (2024)
Beyond-Labels: Advancing Open-Vocabulary Segmentation With Vision-Language Models
by: Rahman, Muhammad Atta ur, et al.
Published: (2025)
by: Rahman, Muhammad Atta ur, et al.
Published: (2025)
DreamLCM: Towards High-Quality Text-to-3D Generation via Latent Consistency Model
by: Zhong, Yiming, et al.
Published: (2024)
by: Zhong, Yiming, et al.
Published: (2024)
Test-Time Optimization for Domain Adaptive Open Vocabulary Segmentation
by: De Silva, Ulindu, et al.
Published: (2025)
by: De Silva, Ulindu, et al.
Published: (2025)
From Open-Vocabulary to Vocabulary-Free Semantic Segmentation
by: Reichard, Klara, et al.
Published: (2025)
by: Reichard, Klara, et al.
Published: (2025)
Open-Vocabulary Segmentation with Unpaired Mask-Text Supervision
by: Wang, Zhaoqing, et al.
Published: (2024)
by: Wang, Zhaoqing, et al.
Published: (2024)
DPSeg: Dual-Prompt Cost Volume Learning for Open-Vocabulary Semantic Segmentation
by: Zhao, Ziyu, et al.
Published: (2025)
by: Zhao, Ziyu, et al.
Published: (2025)
Generalization Boosted Adapter for Open-Vocabulary Segmentation
by: Xu, Wenhao, et al.
Published: (2024)
by: Xu, Wenhao, et al.
Published: (2024)
CoMBO: Conflict Mitigation via Branched Optimization for Class Incremental Segmentation
by: Fang, Kai, et al.
Published: (2025)
by: Fang, Kai, et al.
Published: (2025)
Mitigating Objectness Bias and Region-to-Text Misalignment for Open-Vocabulary Panoptic Segmentation
by: Kormushev, Nikolay, et al.
Published: (2026)
by: Kormushev, Nikolay, et al.
Published: (2026)
3D Open-Vocabulary Panoptic Segmentation with 2D-3D Vision-Language Distillation
by: Xiao, Zihao, et al.
Published: (2024)
by: Xiao, Zihao, et al.
Published: (2024)
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation
by: Sun, Lin, et al.
Published: (2024)
by: Sun, Lin, et al.
Published: (2024)
Decouple and Rectify: Semantics-Preserving Structural Enhancement for Open-Vocabulary Remote Sensing Segmentation
by: Feng, Jie, et al.
Published: (2026)
by: Feng, Jie, et al.
Published: (2026)
Taming Generative Synthetic Data for X-ray Prohibited Item Detection
by: Sun, Jialong, et al.
Published: (2025)
by: Sun, Jialong, et al.
Published: (2025)
Similar Items
-
Collaborative Feature-Logits Contrastive Learning for Open-Set Semi-Supervised Object Detection
by: Zhong, Xinhao, et al.
Published: (2024) -
CLIP-GS: Unifying Vision-Language Representation with 3D Gaussian Splatting
by: Jiao, Siyu, et al.
Published: (2024) -
SAGE: Exploring the Boundaries of Unsafe Concept Domain with Semantic-Augment Erasing
by: Zhu, Hongguang, et al.
Published: (2025) -
Le-DETR: Revisiting Real-Time Detection Transformer with Efficient Encoder Design
by: Huang, Jiannan, et al.
Published: (2026) -
Transferable and Principled Efficiency for Open-Vocabulary Segmentation
by: Xu, Jingxuan, et al.
Published: (2024)