Combining inherent knowledge of vision-language models with unsupervised domain adaptation through strong-weak guidance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Westfechtel, Thomas, Zhang, Dexuan, Harada, Tatsuya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unified modality separation: A vision-language framework for unsupervised domain adaptation
von: Li, Xinyao, et al.
Veröffentlicht: (2025)
von: Li, Xinyao, et al.
Veröffentlicht: (2025)
SceneProp: Combining Neural Network and Markov Random Field for Scene-Graph Grounding
von: Otani, Keita, et al.
Veröffentlicht: (2025)
von: Otani, Keita, et al.
Veröffentlicht: (2025)
SADA: Semantic adversarial unsupervised domain adaptation for Temporal Action Localization
von: Pujol-Perich, David, et al.
Veröffentlicht: (2023)
von: Pujol-Perich, David, et al.
Veröffentlicht: (2023)
MIMO: A medical vision language model with visual referring multimodal input and pixel grounding multimodal output
von: Chen, Yanyuan, et al.
Veröffentlicht: (2025)
von: Chen, Yanyuan, et al.
Veröffentlicht: (2025)
SPGen: Stochastic scanpath generation for paintings using unsupervised domain adaptation
von: Kerkouri, Mohamed Amine, et al.
Veröffentlicht: (2026)
von: Kerkouri, Mohamed Amine, et al.
Veröffentlicht: (2026)
Live image-based neurosurgical guidance and roadmap generation using unsupervised embedding
von: Sarwin, Gary, et al.
Veröffentlicht: (2023)
von: Sarwin, Gary, et al.
Veröffentlicht: (2023)
Can language-guided unsupervised adaptation improve medical image classification using unpaired images and texts?
von: Rahman, Umaima, et al.
Veröffentlicht: (2024)
von: Rahman, Umaima, et al.
Veröffentlicht: (2024)
BTMuda: A Bi-level Multi-source unsupervised domain adaptation framework for breast cancer diagnosis
von: Yang, Yuxiang, et al.
Veröffentlicht: (2024)
von: Yang, Yuxiang, et al.
Veröffentlicht: (2024)
Enhancing the vision-language foundation model with key semantic knowledge-emphasized report refinement
von: Huang, Weijian, et al.
Veröffentlicht: (2024)
von: Huang, Weijian, et al.
Veröffentlicht: (2024)
Initialization matters in few-shot adaptation of vision-language models for histopathological image classification
von: Meseguer, Pablo, et al.
Veröffentlicht: (2026)
von: Meseguer, Pablo, et al.
Veröffentlicht: (2026)
RAW-Adapter: Adapting Pre-trained Visual Model to Camera RAW Images
von: Cui, Ziteng, et al.
Veröffentlicht: (2024)
von: Cui, Ziteng, et al.
Veröffentlicht: (2024)
Detection Based Part-level Articulated Object Reconstruction from Single RGBD Image
von: Kawana, Yuki, et al.
Veröffentlicht: (2025)
von: Kawana, Yuki, et al.
Veröffentlicht: (2025)
MaGRITTe: Manipulative and Generative 3D Realization from Image, Topview and Text
von: Hara, Takayuki, et al.
Veröffentlicht: (2024)
von: Hara, Takayuki, et al.
Veröffentlicht: (2024)
Generalizing vision-language models to novel domains: A comprehensive survey
von: Li, Xinyao, et al.
Veröffentlicht: (2025)
von: Li, Xinyao, et al.
Veröffentlicht: (2025)
Explainable artificial intelligence (XAI): from inherent explainability to large language models
von: Mumuni, Fuseini, et al.
Veröffentlicht: (2025)
von: Mumuni, Fuseini, et al.
Veröffentlicht: (2025)
SituationalLLM: Proactive language models with scene awareness for dynamic, contextual task guidance
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
Self-adaptive vision-language model for 3D segmentation of pulmonary artery and vein
von: Guo, Xiaotong, et al.
Veröffentlicht: (2025)
von: Guo, Xiaotong, et al.
Veröffentlicht: (2025)
Generalizable and Animatable Gaussian Head Avatar
von: Chu, Xuangeng, et al.
Veröffentlicht: (2024)
von: Chu, Xuangeng, et al.
Veröffentlicht: (2024)
Exploring selective image matching methods for zero-shot and few-sample unsupervised domain adaptation of urban canopy prediction
von: Francis, John, et al.
Veröffentlicht: (2024)
von: Francis, John, et al.
Veröffentlicht: (2024)
An analysis of vision-language models for fabric retrieval
von: Giuliari, Francesco, et al.
Veröffentlicht: (2025)
von: Giuliari, Francesco, et al.
Veröffentlicht: (2025)
DeViDe: Faceted medical knowledge for improved medical vision-language pre-training
von: Luo, Haozhe, et al.
Veröffentlicht: (2024)
von: Luo, Haozhe, et al.
Veröffentlicht: (2024)
Discovering an Image-Adaptive Coordinate System for Photography Processing
von: Cui, Ziteng, et al.
Veröffentlicht: (2025)
von: Cui, Ziteng, et al.
Veröffentlicht: (2025)
RAW-Adapter: Adapting Pre-trained Visual Model to Camera RAW Images and A Benchmark
von: Cui, Ziteng, et al.
Veröffentlicht: (2025)
von: Cui, Ziteng, et al.
Veröffentlicht: (2025)
Luminance-GS: Adapting 3D Gaussian Splatting to Challenging Lighting Conditions with View-Adaptive Curve Adjustment
von: Cui, Ziteng, et al.
Veröffentlicht: (2025)
von: Cui, Ziteng, et al.
Veröffentlicht: (2025)
Upsampling DINOv2 features for unsupervised vision tasks and weakly supervised materials segmentation
von: Docherty, Ronan, et al.
Veröffentlicht: (2024)
von: Docherty, Ronan, et al.
Veröffentlicht: (2024)
Linking heterogeneous microstructure informatics with expert characterization knowledge through customized and hybrid vision-language representations for industrial qualification
von: Safdar, Mutahar, et al.
Veröffentlicht: (2025)
von: Safdar, Mutahar, et al.
Veröffentlicht: (2025)
Interpreting the linear structure of vision-language model embedding spaces
von: Papadimitriou, Isabel, et al.
Veröffentlicht: (2025)
von: Papadimitriou, Isabel, et al.
Veröffentlicht: (2025)
Improved 3D Scene Stylization via Text-Guided Generative Image Editing with Region-Based Control
von: Fujiwara, Haruo, et al.
Veröffentlicht: (2025)
von: Fujiwara, Haruo, et al.
Veröffentlicht: (2025)
Style-NeRF2NeRF: 3D Style Transfer From Style-Aligned Multi-View Images
von: Fujiwara, Haruo, et al.
Veröffentlicht: (2024)
von: Fujiwara, Haruo, et al.
Veröffentlicht: (2024)
Deep Gaussian mixture model for unsupervised image segmentation
von: Schwab, Matthias, et al.
Veröffentlicht: (2024)
von: Schwab, Matthias, et al.
Veröffentlicht: (2024)
Deep asymmetric mixture model for unsupervised cell segmentation
von: Nan, Yang, et al.
Veröffentlicht: (2024)
von: Nan, Yang, et al.
Veröffentlicht: (2024)
Expert-level vision-language foundation model for real-world radiology and comprehensive evaluation
von: Liu, Xiaohong, et al.
Veröffentlicht: (2024)
von: Liu, Xiaohong, et al.
Veröffentlicht: (2024)
MI-VisionShot: Few-shot adaptation of vision-language models for slide-level classification of histopathological images
von: Meseguer, Pablo, et al.
Veröffentlicht: (2024)
von: Meseguer, Pablo, et al.
Veröffentlicht: (2024)
Do large language vision models understand 3D shapes?
von: Eppel, Sagi
Veröffentlicht: (2024)
von: Eppel, Sagi
Veröffentlicht: (2024)
Visual symbolic mechanisms: Emergent symbol processing in vision language models
von: Assouel, Rim, et al.
Veröffentlicht: (2025)
von: Assouel, Rim, et al.
Veröffentlicht: (2025)
Are vision-language models ready to zero-shot replace supervised classification models in agriculture?
von: Ranario, Earl, et al.
Veröffentlicht: (2025)
von: Ranario, Earl, et al.
Veröffentlicht: (2025)
GeoGuide: Geometric guidance of diffusion models
von: Poleski, Mateusz, et al.
Veröffentlicht: (2024)
von: Poleski, Mateusz, et al.
Veröffentlicht: (2024)
Fine color guidance in diffusion models and its application to image compression at extremely low bitrates
von: Bordin, Tom, et al.
Veröffentlicht: (2024)
von: Bordin, Tom, et al.
Veröffentlicht: (2024)
Enhancing knowledge retention for continual learning with domain-specific adapters and features gating
von: Hedjazi, Mohamed Abbas, et al.
Veröffentlicht: (2025)
von: Hedjazi, Mohamed Abbas, et al.
Veröffentlicht: (2025)
Are vision language models robust to uncertain inputs?
von: Wang, Xi, et al.
Veröffentlicht: (2025)
von: Wang, Xi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Unified modality separation: A vision-language framework for unsupervised domain adaptation
von: Li, Xinyao, et al.
Veröffentlicht: (2025) -
SceneProp: Combining Neural Network and Markov Random Field for Scene-Graph Grounding
von: Otani, Keita, et al.
Veröffentlicht: (2025) -
SADA: Semantic adversarial unsupervised domain adaptation for Temporal Action Localization
von: Pujol-Perich, David, et al.
Veröffentlicht: (2023) -
MIMO: A medical vision language model with visual referring multimodal input and pixel grounding multimodal output
von: Chen, Yanyuan, et al.
Veröffentlicht: (2025) -
SPGen: Stochastic scanpath generation for paintings using unsupervised domain adaptation
von: Kerkouri, Mohamed Amine, et al.
Veröffentlicht: (2026)