Towards Multimodal Domain Generalization with Few Labels
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Hongzhao, Dong, Hao, Wan, Hualei, Li, Shupan, Xu, Mingliang, Khan, Muhammad Haris |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Balancing Multimodal Domain Generalization via Gradient Modulation and Projection
by: Li, Hongzhao, et al.
Published: (2026)
by: Li, Hongzhao, et al.
Published: (2026)
Are We Making Progress in Multimodal Domain Generalization? A Comprehensive Benchmark Study
by: Dong, Hao, et al.
Published: (2026)
by: Dong, Hao, et al.
Published: (2026)
Towards Generalizing to Unseen Domains with Few Labels
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2024)
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2024)
Robust and Label-Efficient Deep Waste Detection
by: Abid, Hassan, et al.
Published: (2025)
by: Abid, Hassan, et al.
Published: (2025)
Improving Pseudo-labelling and Enhancing Robustness for Semi-Supervised Domain Generalization
by: Khan, Adnan, et al.
Published: (2024)
by: Khan, Adnan, et al.
Published: (2024)
Noise-Tolerant Few-Shot Unsupervised Adapter for Vision-Language Models
by: Ali, Eman, et al.
Published: (2023)
by: Ali, Eman, et al.
Published: (2023)
Divergent Domains, Convergent Grading: Enhancing Generalization in Diabetic Retinopathy Grading
by: Chokuwa, Sharon, et al.
Published: (2024)
by: Chokuwa, Sharon, et al.
Published: (2024)
FrogDogNet: Fourier frequency Retained visual prompt Output Guidance for Domain Generalization of CLIP in Remote Sensing
by: Gunduboina, Hariseetharam, et al.
Published: (2025)
by: Gunduboina, Hariseetharam, et al.
Published: (2025)
Domain-Guided Weight Modulation for Semi-Supervised Domain Generalization
by: Galappaththige, Chamuditha Jayanaga, et al.
Published: (2024)
by: Galappaththige, Chamuditha Jayanaga, et al.
Published: (2024)
Improving Single Domain-Generalized Object Detection: A Focus on Diversification and Alignment
by: Danish, Muhammad Sohail, et al.
Published: (2024)
by: Danish, Muhammad Sohail, et al.
Published: (2024)
Judging from Support-set: A New Way to Utilize Few-Shot Segmentation for Segmentation Refinement Process
by: Moon, Seonghyeon, et al.
Published: (2024)
by: Moon, Seonghyeon, et al.
Published: (2024)
Image Classification with Deep Reinforcement Active Learning
by: Jiu, Mingyuan, et al.
Published: (2024)
by: Jiu, Mingyuan, et al.
Published: (2024)
Towards Combating Frequency Simplicity-biased Learning for Domain Generalization
by: He, Xilin, et al.
Published: (2024)
by: He, Xilin, et al.
Published: (2024)
CountZES: Counting via Zero-Shot Exemplar Selection
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
Towards Fine-Grained Adaptation of CLIP via a Self-Trained Alignment Score
by: Ali, Eman, et al.
Published: (2025)
by: Ali, Eman, et al.
Published: (2025)
Towards PerSense++: Advancing Training-Free Personalized Instance Segmentation in Dense Images
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
ReConText3D: Replay-based Continual Text-to-3D Generation
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2026)
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2026)
Adapting In-Domain Few-Shot Segmentation to New Domains without Source Domain Retraining
by: Fan, Qi, et al.
Published: (2025)
by: Fan, Qi, et al.
Published: (2025)
DPA: Dual Prototypes Alignment for Unsupervised Adaptation of Vision-Language Models
by: Ali, Eman, et al.
Published: (2024)
by: Ali, Eman, et al.
Published: (2024)
Not All Modalities Are Equal: Instruction-Aware Gating for Multimodal Videos
by: Ding, Bonan, et al.
Published: (2026)
by: Ding, Bonan, et al.
Published: (2026)
OSLoPrompt: Bridging Low-Supervision Challenges and Open-Set Domain Generalization in CLIP
by: C, Mohamad Hassan N, et al.
Published: (2025)
by: C, Mohamad Hassan N, et al.
Published: (2025)
Few-Shot Object Detection with Sparse Context Transformers
by: Mei, Jie, et al.
Published: (2024)
by: Mei, Jie, et al.
Published: (2024)
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
by: Sharifdeen, Ashshak, et al.
Published: (2025)
by: Sharifdeen, Ashshak, et al.
Published: (2025)
PerSense: Training-Free Personalized Instance Segmentation in Dense Images
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2024)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2024)
Towards Multimodal Open-Set Domain Generalization and Adaptation through Self-supervision
by: Dong, Hao, et al.
Published: (2024)
by: Dong, Hao, et al.
Published: (2024)
Chameleon: Images Are What You Need For Multimodal Learning Robust To Missing Modalities
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
Remedying Target-Domain Astigmatism for Cross-Domain Few-Shot Object Detection
by: Jiang, Yongwei, et al.
Published: (2026)
by: Jiang, Yongwei, et al.
Published: (2026)
DomainGallery: Few-shot Domain-driven Image Generation by Attribute-centric Finetuning
by: Duan, Yuxuan, et al.
Published: (2024)
by: Duan, Yuxuan, et al.
Published: (2024)
Modality Invariant Multimodal Learning to Handle Missing Modalities: A Single-Branch Approach
by: Saeed, Muhammad Saad, et al.
Published: (2024)
by: Saeed, Muhammad Saad, et al.
Published: (2024)
Domain-RAG: Retrieval-Guided Compositional Image Generation for Cross-Domain Few-Shot Object Detection
by: Li, Yu, et al.
Published: (2025)
by: Li, Yu, et al.
Published: (2025)
Random Registers for Cross-Domain Few-Shot Learning
by: Yi, Shuai, et al.
Published: (2025)
by: Yi, Shuai, et al.
Published: (2025)
Vision-aware Multimodal Prompt Tuning for Uploadable Multi-source Few-shot Domain Adaptation
by: Liu, Kuanghong, et al.
Published: (2025)
by: Liu, Kuanghong, et al.
Published: (2025)
Towards Generalized Few-Shot Open-Set Object Detection
by: Su, Binyi, et al.
Published: (2022)
by: Su, Binyi, et al.
Published: (2022)
Pixels Don't Lie (But Your Detector Might): Bootstrapping MLLM-as-a-Judge for Trustworthy Deepfake Detection and Reasoning Supervision
by: Kuckreja, Kartik, et al.
Published: (2026)
by: Kuckreja, Kartik, et al.
Published: (2026)
Pose-Guided Self-Training with Two-Stage Clustering for Unsupervised Landmark Discovery
by: Tourani, Siddharth, et al.
Published: (2024)
by: Tourani, Siddharth, et al.
Published: (2024)
Multimodal Cross-Domain Few-Shot Learning for Egocentric Action Recognition
by: Hatano, Masashi, et al.
Published: (2024)
by: Hatano, Masashi, et al.
Published: (2024)
BioVLM: Routing Prompts, Not Parameters, for Cross-Modality Generalization in Biomedical VLMs
by: Singha, Mainak, et al.
Published: (2026)
by: Singha, Mainak, et al.
Published: (2026)
Domain Adaptation Using Pseudo Labels
by: Chhabra, Sachin, et al.
Published: (2024)
by: Chhabra, Sachin, et al.
Published: (2024)
Few-Shot Domain Adaptation for Learned Image Compression
by: Zhang, Tianyu, et al.
Published: (2024)
by: Zhang, Tianyu, et al.
Published: (2024)
NT-VOT211: A Large-Scale Benchmark for Night-time Visual Object Tracking
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
Similar Items
-
Balancing Multimodal Domain Generalization via Gradient Modulation and Projection
by: Li, Hongzhao, et al.
Published: (2026) -
Are We Making Progress in Multimodal Domain Generalization? A Comprehensive Benchmark Study
by: Dong, Hao, et al.
Published: (2026) -
Towards Generalizing to Unseen Domains with Few Labels
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2024) -
Robust and Label-Efficient Deep Waste Detection
by: Abid, Hassan, et al.
Published: (2025) -
Improving Pseudo-labelling and Enhancing Robustness for Semi-Supervised Domain Generalization
by: Khan, Adnan, et al.
Published: (2024)