FrozenSeg: Harmonizing Frozen Foundation Models for Open-Vocabulary Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Xi, Yang, Haosen, Jin, Sheng, Zhu, Xiatian, Yao, Hongxun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Uncertainty-Aware Pseudo-Label Filtering for Source-Free Unsupervised Domain Adaptation
von: Chen, Xi, et al.
Veröffentlicht: (2024)
von: Chen, Xi, et al.
Veröffentlicht: (2024)
Source-Free Domain Adaptation with Frozen Multimodal Foundation Model
von: Tang, Song, et al.
Veröffentlicht: (2023)
von: Tang, Song, et al.
Veröffentlicht: (2023)
Advancing Generalizable Tumor Segmentation with Anomaly-Aware Open-Vocabulary Attention Maps and Frozen Foundation Diffusion Models
von: Jiang, Yankai, et al.
Veröffentlicht: (2025)
von: Jiang, Yankai, et al.
Veröffentlicht: (2025)
Frozen-DETR: Enhancing DETR with Image Understanding from Frozen Foundation Models
von: Fu, Shenghao, et al.
Veröffentlicht: (2024)
von: Fu, Shenghao, et al.
Veröffentlicht: (2024)
FROSTER: Frozen CLIP Is A Strong Teacher for Open-Vocabulary Action Recognition
von: Huang, Xiaohu, et al.
Veröffentlicht: (2024)
von: Huang, Xiaohu, et al.
Veröffentlicht: (2024)
PCA-Seg: Revisiting Cost Aggregation for Open-Vocabulary Semantic and Part Segmentation
von: Yin, Jianjian, et al.
Veröffentlicht: (2026)
von: Yin, Jianjian, et al.
Veröffentlicht: (2026)
F-LMM: Grounding Frozen Large Multimodal Models
von: Wu, Size, et al.
Veröffentlicht: (2024)
von: Wu, Size, et al.
Veröffentlicht: (2024)
LoGoSeg: Integrating Local and Global Features for Open-Vocabulary Semantic Segmentation
von: Chen, Junyang, et al.
Veröffentlicht: (2026)
von: Chen, Junyang, et al.
Veröffentlicht: (2026)
EOV-Seg: Efficient Open-Vocabulary Panoptic Segmentation
von: Niu, Hongwei, et al.
Veröffentlicht: (2024)
von: Niu, Hongwei, et al.
Veröffentlicht: (2024)
CitySeg: A 3D Open Vocabulary Semantic Segmentation Foundation Model in City-scale Scenarios
von: Xu, Jialei, et al.
Veröffentlicht: (2025)
von: Xu, Jialei, et al.
Veröffentlicht: (2025)
CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation
von: Cho, Seokju, et al.
Veröffentlicht: (2023)
von: Cho, Seokju, et al.
Veröffentlicht: (2023)
DAM: Dual Active Learning with Multimodal Foundation Model for Source-Free Domain Adaptation
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
FreeSeg-Diff: Training-Free Open-Vocabulary Segmentation with Diffusion Models
von: Corradini, Barbara Toniella, et al.
Veröffentlicht: (2024)
von: Corradini, Barbara Toniella, et al.
Veröffentlicht: (2024)
Segmentation as A Plug-and-Play Capability for Frozen Multimodal LLMs
von: Liu, Jiazhen, et al.
Veröffentlicht: (2025)
von: Liu, Jiazhen, et al.
Veröffentlicht: (2025)
Frozen CLIP: A Strong Backbone for Weakly Supervised Semantic Segmentation
von: Zhang, Bingfeng, et al.
Veröffentlicht: (2024)
von: Zhang, Bingfeng, et al.
Veröffentlicht: (2024)
Unsupervised Audio-Visual Segmentation with Modality Alignment
von: Bhosale, Swapnil, et al.
Veröffentlicht: (2024)
von: Bhosale, Swapnil, et al.
Veröffentlicht: (2024)
Single Image, Any Face: Generalisable 3D Face Generation
von: Wang, Wenqing, et al.
Veröffentlicht: (2024)
von: Wang, Wenqing, et al.
Veröffentlicht: (2024)
Dynamic Avatar-Scene Rendering from Human-centric Context
von: Wang, Wenqing, et al.
Veröffentlicht: (2025)
von: Wang, Wenqing, et al.
Veröffentlicht: (2025)
OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning
von: Han, Zongyan, et al.
Veröffentlicht: (2025)
von: Han, Zongyan, et al.
Veröffentlicht: (2025)
Seg2Change: Adapting Open-Vocabulary Semantic Segmentation Model for Remote Sensing Change Detection
von: Su, You, et al.
Veröffentlicht: (2026)
von: Su, You, et al.
Veröffentlicht: (2026)
AttrSeg: Open-Vocabulary Semantic Segmentation via Attribute Decomposition-Aggregation
von: Ma, Chaofan, et al.
Veröffentlicht: (2023)
von: Ma, Chaofan, et al.
Veröffentlicht: (2023)
Towards Data-Efficient Video Pre-training with Frozen Image Foundation Models
von: Orlova, Svetlana, et al.
Veröffentlicht: (2026)
von: Orlova, Svetlana, et al.
Veröffentlicht: (2026)
Seg4Diff: Unveiling Open-Vocabulary Segmentation in Text-to-Image Diffusion Transformers
von: Kim, Chaehyun, et al.
Veröffentlicht: (2025)
von: Kim, Chaehyun, et al.
Veröffentlicht: (2025)
FA-Seg: A Fast and Accurate Diffusion-Based Method for Open-Vocabulary Segmentation
von: Che, Huy, et al.
Veröffentlicht: (2025)
von: Che, Huy, et al.
Veröffentlicht: (2025)
TextRegion: Text-Aligned Region Tokens from Frozen Image-Text Models
von: Xiao, Yao, et al.
Veröffentlicht: (2025)
von: Xiao, Yao, et al.
Veröffentlicht: (2025)
Chat-CBM: Towards Interactive Concept Bottleneck Models with Frozen Large Language Models
von: He, Hangzhou, et al.
Veröffentlicht: (2025)
von: He, Hangzhou, et al.
Veröffentlicht: (2025)
Beyond Text: Frozen Large Language Models in Visual Signal Comprehension
von: Zhu, Lei, et al.
Veröffentlicht: (2024)
von: Zhu, Lei, et al.
Veröffentlicht: (2024)
SynSeg: Feature Synergy for Multi-Category Contrastive Learning in End-to-End Open-Vocabulary Semantic Segmentation
von: Zhang, Weichen, et al.
Veröffentlicht: (2025)
von: Zhang, Weichen, et al.
Veröffentlicht: (2025)
Few-Shot Continual Learning for 3D Brain MRI with Frozen Foundation Models
von: Chen, Chi-Sheng, et al.
Veröffentlicht: (2026)
von: Chen, Chi-Sheng, et al.
Veröffentlicht: (2026)
SegEarth-OV: Towards Training-Free Open-Vocabulary Segmentation for Remote Sensing Images
von: Li, Kaiyu, et al.
Veröffentlicht: (2024)
von: Li, Kaiyu, et al.
Veröffentlicht: (2024)
PMT: Plain Mask Transformer for Image and Video Segmentation with Frozen Vision Encoders
von: Cavagnero, Niccolò, et al.
Veröffentlicht: (2026)
von: Cavagnero, Niccolò, et al.
Veröffentlicht: (2026)
CLIP-VIS: Adapting CLIP for Open-Vocabulary Video Instance Segmentation
von: Zhu, Wenqi, et al.
Veröffentlicht: (2024)
von: Zhu, Wenqi, et al.
Veröffentlicht: (2024)
TIR-Flow: Active Video Search and Reasoning with Frozen VLMs
von: Jin, Hongbo, et al.
Veröffentlicht: (2026)
von: Jin, Hongbo, et al.
Veröffentlicht: (2026)
WOW-Seg: A Word-free Open World Segmentation Model
von: Li, Danyang, et al.
Veröffentlicht: (2026)
von: Li, Danyang, et al.
Veröffentlicht: (2026)
Backbone is All You Need: Assessing Vulnerabilities of Frozen Foundation Models in Synthetic Image Forensics
von: Musso, Chiara, et al.
Veröffentlicht: (2026)
von: Musso, Chiara, et al.
Veröffentlicht: (2026)
HyperTopo-Adapters: Geometry- and Topology-Aware Segmentation of Leaf Lesions on Frozen Encoders
von: Ndubuisi, Chimdi Walter, et al.
Veröffentlicht: (2025)
von: Ndubuisi, Chimdi Walter, et al.
Veröffentlicht: (2025)
Dimensional Coactivation for Representational Consistency in Frozen Vision Foundation Models
von: Saddik, Izaldein Al-Zyoud Abdulmotaleb El
Veröffentlicht: (2026)
von: Saddik, Izaldein Al-Zyoud Abdulmotaleb El
Veröffentlicht: (2026)
RSRefSeg: Referring Remote Sensing Image Segmentation with Foundation Models
von: Chen, Keyan, et al.
Veröffentlicht: (2025)
von: Chen, Keyan, et al.
Veröffentlicht: (2025)
Do Foundation Models Know Geometry? Probing Frozen Features for Continuous Physical Measurement
von: Shkolnikov, Yakov Pyotr
Veröffentlicht: (2026)
von: Shkolnikov, Yakov Pyotr
Veröffentlicht: (2026)
Seed Optimization with Frozen Generator for Superior Zero-shot Low-light Enhancement
von: Gu, Yuxuan, et al.
Veröffentlicht: (2024)
von: Gu, Yuxuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Uncertainty-Aware Pseudo-Label Filtering for Source-Free Unsupervised Domain Adaptation
von: Chen, Xi, et al.
Veröffentlicht: (2024) -
Source-Free Domain Adaptation with Frozen Multimodal Foundation Model
von: Tang, Song, et al.
Veröffentlicht: (2023) -
Advancing Generalizable Tumor Segmentation with Anomaly-Aware Open-Vocabulary Attention Maps and Frozen Foundation Diffusion Models
von: Jiang, Yankai, et al.
Veröffentlicht: (2025) -
Frozen-DETR: Enhancing DETR with Image Understanding from Frozen Foundation Models
von: Fu, Shenghao, et al.
Veröffentlicht: (2024) -
FROSTER: Frozen CLIP Is A Strong Teacher for Open-Vocabulary Action Recognition
von: Huang, Xiaohu, et al.
Veröffentlicht: (2024)