A Simple Recipe for Language-guided Domain Generalized Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Fahes, Mohammad, Vu, Tuan-Hung, Bursuc, Andrei, Pérez, Patrick, de Charette, Raoul |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Domain Adaptation with a Single Vision-Language Embedding
by: Fahes, Mohammad, et al.
Published: (2024)
by: Fahes, Mohammad, et al.
Published: (2024)
CLIP's Visual Embedding Projector is a Few-shot Cornucopia
by: Fahes, Mohammad, et al.
Published: (2024)
by: Fahes, Mohammad, et al.
Published: (2024)
FLOSS: Free Lunch in Open-vocabulary Semantic Segmentation
by: Benigmim, Yasser, et al.
Published: (2025)
by: Benigmim, Yasser, et al.
Published: (2025)
DenseMTL: Cross-task Attention Mechanism for Dense Multi-task Learning
by: Lopes, Ivan, et al.
Published: (2022)
by: Lopes, Ivan, et al.
Published: (2022)
Drive&Segment: Unsupervised Semantic Segmentation of Urban Scenes via Cross-modal Distillation
by: Vobecky, Antonin, et al.
Published: (2022)
by: Vobecky, Antonin, et al.
Published: (2022)
Reliability in Semantic Segmentation: Can We Use Synthetic Data?
by: Loiseau, Thibaut, et al.
Published: (2023)
by: Loiseau, Thibaut, et al.
Published: (2023)
Improving Multimodal Distillation for 3D Semantic Segmentation under Domain Shift
by: Michele, Björn, et al.
Published: (2025)
by: Michele, Björn, et al.
Published: (2025)
Material Transforms from Disentangled NeRF Representations
by: Lopes, Ivan, et al.
Published: (2024)
by: Lopes, Ivan, et al.
Published: (2024)
OccAny: Generalized Unconstrained Urban 3D Occupancy
by: Cao, Anh-Quan, et al.
Published: (2026)
by: Cao, Anh-Quan, et al.
Published: (2026)
BIGFix: Bidirectional Image Generation with Token Fixing
by: Besnier, Victor, et al.
Published: (2025)
by: Besnier, Victor, et al.
Published: (2025)
PaSCo: Urban 3D Panoptic Scene Completion with Uncertainty Awareness
by: Cao, Anh-Quan, et al.
Published: (2023)
by: Cao, Anh-Quan, et al.
Published: (2023)
Test-time Contrastive Concepts for Open-world Semantic Segmentation with Vision-Language Models
by: Wysoczańska, Monika, et al.
Published: (2024)
by: Wysoczańska, Monika, et al.
Published: (2024)
CLIP-DINOiser: Teaching CLIP a few DINO tricks for open-vocabulary semantic segmentation
by: Wysoczańska, Monika, et al.
Published: (2023)
by: Wysoczańska, Monika, et al.
Published: (2023)
MatSwap: Light-aware material transfers in images
by: Lopes, Ivan, et al.
Published: (2025)
by: Lopes, Ivan, et al.
Published: (2025)
OccFeat: Self-supervised Occupancy Feature Prediction for Pretraining BEV Segmentation Networks
by: Sirko-Galouchenko, Sophia, et al.
Published: (2024)
by: Sirko-Galouchenko, Sophia, et al.
Published: (2024)
Learning to Generate Training Datasets for Robust Semantic Segmentation
by: Hariat, Marwane, et al.
Published: (2023)
by: Hariat, Marwane, et al.
Published: (2023)
StableMTL: Repurposing Latent Diffusion Models for Multi-Task Learning from Partially Annotated Synthetic Datasets
by: Cao, Anh-Quan, et al.
Published: (2025)
by: Cao, Anh-Quan, et al.
Published: (2025)
LiDPM: Rethinking Point Diffusion for Lidar Scene Completion
by: Martyniuk, Tetiana, et al.
Published: (2025)
by: Martyniuk, Tetiana, et al.
Published: (2025)
POP-3D: Open-Vocabulary 3D Occupancy Prediction from Images
by: Vobecky, Antonin, et al.
Published: (2024)
by: Vobecky, Antonin, et al.
Published: (2024)
SALUDA: Surface-based Automotive Lidar Unsupervised Domain Adaptation
by: Michele, Björn, et al.
Published: (2023)
by: Michele, Björn, et al.
Published: (2023)
UMBRAE: Unified Multimodal Brain Decoding
by: Xia, Weihao, et al.
Published: (2024)
by: Xia, Weihao, et al.
Published: (2024)
Train Till You Drop: Towards Stable and Robust Source-free Unsupervised 3D Domain Adaptation
by: Michele, Björn, et al.
Published: (2024)
by: Michele, Björn, et al.
Published: (2024)
Three Pillars improving Vision Foundation Model Distillation for Lidar
by: Puy, Gilles, et al.
Published: (2023)
by: Puy, Gilles, et al.
Published: (2023)
ForestMamba: Sparse Mamba with Geometry-guided Queries for 3D Forest Point Cloud Segmentation
by: Nguyen, Trung Thanh, et al.
Published: (2026)
by: Nguyen, Trung Thanh, et al.
Published: (2026)
Bridging the Generalization Gap in Adverse Weather Segmentation: A Training Recipe Perspective
by: Xu, Cong, et al.
Published: (2026)
by: Xu, Cong, et al.
Published: (2026)
Language-driven Grasp Detection with Mask-guided Attention
by: Van Vo, Tuan, et al.
Published: (2024)
by: Van Vo, Tuan, et al.
Published: (2024)
Temporal-Oriented Recipe for Transferring Large Vision-Language Model to Video Understanding
by: Nguyen, Thong, et al.
Published: (2025)
by: Nguyen, Thong, et al.
Published: (2025)
RecipeGen: A Benchmark for Real-World Recipe Image Generation
by: Zhang, Ruoxuan, et al.
Published: (2025)
by: Zhang, Ruoxuan, et al.
Published: (2025)
LiDAS: Lighting-driven Dynamic Active Sensing for Nighttime Perception
by: de Moreau, Simon, et al.
Published: (2025)
by: de Moreau, Simon, et al.
Published: (2025)
Valeo4Cast: A Modular Approach to End-to-End Forecasting
by: Xu, Yihong, et al.
Published: (2024)
by: Xu, Yihong, et al.
Published: (2024)
LatteCLIP: Unsupervised CLIP Fine-Tuning via LMM-Synthetic Texts
by: Cao, Anh-Quan, et al.
Published: (2024)
by: Cao, Anh-Quan, et al.
Published: (2024)
DualProtoSeg: Simple and Efficient Design with Text- and Image-Guided Prototype Learning for Weakly Supervised Histopathology Image Segmentation
by: Vu, Anh M., et al.
Published: (2025)
by: Vu, Anh M., et al.
Published: (2025)
GEM: Boost Simple Network for Glass Surface Segmentation via Segment Anything Model and Data Synthesis
by: Hao, Jing, et al.
Published: (2024)
by: Hao, Jing, et al.
Published: (2024)
DIP: Unsupervised Dense In-Context Post-training of Visual Representations
by: Sirko-Galouchenko, Sophia, et al.
Published: (2025)
by: Sirko-Galouchenko, Sophia, et al.
Published: (2025)
Boosting Visual Instruction Tuning with Self-Supervised Guidance
by: Sirko-Galouchenko, Sophia, et al.
Published: (2026)
by: Sirko-Galouchenko, Sophia, et al.
Published: (2026)
RecipeGen: A Step-Aligned Multimodal Benchmark for Real-World Recipe Generation
by: Zhang, Ruoxuan, et al.
Published: (2025)
by: Zhang, Ruoxuan, et al.
Published: (2025)
Exploiting Domain Properties in Language-Driven Domain Generalization for Semantic Segmentation
by: Jeon, Seogkyu, et al.
Published: (2025)
by: Jeon, Seogkyu, et al.
Published: (2025)
Retrieval Augmented Recipe Generation
by: Liu, Guoshan, et al.
Published: (2024)
by: Liu, Guoshan, et al.
Published: (2024)
3D sans 3D Scans: Scalable Pre-training from Video-Generated Point Clouds
by: Yamada, Ryousuke, et al.
Published: (2025)
by: Yamada, Ryousuke, et al.
Published: (2025)
Language Guided Domain Generalized Medical Image Segmentation
by: Kunhimon, Shahina, et al.
Published: (2024)
by: Kunhimon, Shahina, et al.
Published: (2024)
Similar Items
-
Domain Adaptation with a Single Vision-Language Embedding
by: Fahes, Mohammad, et al.
Published: (2024) -
CLIP's Visual Embedding Projector is a Few-shot Cornucopia
by: Fahes, Mohammad, et al.
Published: (2024) -
FLOSS: Free Lunch in Open-vocabulary Semantic Segmentation
by: Benigmim, Yasser, et al.
Published: (2025) -
DenseMTL: Cross-task Attention Mechanism for Dense Multi-task Learning
by: Lopes, Ivan, et al.
Published: (2022) -
Drive&Segment: Unsupervised Semantic Segmentation of Urban Scenes via Cross-modal Distillation
by: Vobecky, Antonin, et al.
Published: (2022)