Lightweight Unsupervised Federated Learning with Pretrained Vision Language Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yan, Hao, Guo, Yuhong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VLLFL: A Vision-Language Model Based Lightweight Federated Learning Framework for Smart Agriculture
von: Li, Long, et al.
Veröffentlicht: (2025)
von: Li, Long, et al.
Veröffentlicht: (2025)
Federated Learning for Video Violence Detection: Complementary Roles of Lightweight CNNs and Vision-Language Models for Energy-Efficient Use
von: Thuau, Sébastien, et al.
Veröffentlicht: (2025)
von: Thuau, Sébastien, et al.
Veröffentlicht: (2025)
Pretrained Reversible Generation as Unsupervised Visual Representation Learning
von: Xue, Rongkun, et al.
Veröffentlicht: (2024)
von: Xue, Rongkun, et al.
Veröffentlicht: (2024)
Enhancing Generalization in Vision-Language-Action Models by Preserving Pretrained Representations
von: Grover, Shresth, et al.
Veröffentlicht: (2025)
von: Grover, Shresth, et al.
Veröffentlicht: (2025)
Bi-Level Optimization for Single Domain Generalization
von: Heidari, Marzi, et al.
Veröffentlicht: (2026)
von: Heidari, Marzi, et al.
Veröffentlicht: (2026)
Modeling Caption Diversity in Contrastive Vision-Language Pretraining
von: Lavoie, Samuel, et al.
Veröffentlicht: (2024)
von: Lavoie, Samuel, et al.
Veröffentlicht: (2024)
Mordal: Automated Pretrained Model Selection for Vision Language Models
von: He, Shiqi, et al.
Veröffentlicht: (2025)
von: He, Shiqi, et al.
Veröffentlicht: (2025)
From Pixels to Predicates: Learning Symbolic World Models via Pretrained Vision-Language Models
von: Athalye, Ashay, et al.
Veröffentlicht: (2024)
von: Athalye, Ashay, et al.
Veröffentlicht: (2024)
Robust CLIP: Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models
von: Schlarmann, Christian, et al.
Veröffentlicht: (2024)
von: Schlarmann, Christian, et al.
Veröffentlicht: (2024)
Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos
von: Li, Qixiu, et al.
Veröffentlicht: (2025)
von: Li, Qixiu, et al.
Veröffentlicht: (2025)
TinyAlign: Boosting Lightweight Vision-Language Models by Mitigating Modal Alignment Bottlenecks
von: Hu, Yuanze, et al.
Veröffentlicht: (2025)
von: Hu, Yuanze, et al.
Veröffentlicht: (2025)
Renaissance: Investigating the Pretraining of Vision-Language Encoders
von: Fields, Clayton, et al.
Veröffentlicht: (2024)
von: Fields, Clayton, et al.
Veröffentlicht: (2024)
Prompt-Driven Feature Diffusion for Open-World Semi-Supervised Learning
von: Heidari, Marzi, et al.
Veröffentlicht: (2024)
von: Heidari, Marzi, et al.
Veröffentlicht: (2024)
Target-Oriented Single Domain Generalization
von: Heidari, Marzi, et al.
Veröffentlicht: (2025)
von: Heidari, Marzi, et al.
Veröffentlicht: (2025)
Siamese Networks with Soft Labels for Unsupervised Lesion Detection and Patch Pretraining on Screening Mammograms
von: Van Vorst, Kevin, et al.
Veröffentlicht: (2024)
von: Van Vorst, Kevin, et al.
Veröffentlicht: (2024)
DINORANKCLIP: DINOv3 Distillation and Injection for Vision-Language Pretraining with High-Order Ranking Consistency
von: Jiang, Shuyang, et al.
Veröffentlicht: (2026)
von: Jiang, Shuyang, et al.
Veröffentlicht: (2026)
Beyond Human Vision: The Role of Large Vision Language Models in Microscope Image Analysis
von: Verma, Prateek, et al.
Veröffentlicht: (2024)
von: Verma, Prateek, et al.
Veröffentlicht: (2024)
Provable Ordering and Continuity in Vision-Language Pretraining for Generalizable Embodied Agents
von: Zhang, Zhizhen, et al.
Veröffentlicht: (2025)
von: Zhang, Zhizhen, et al.
Veröffentlicht: (2025)
Continual Adaptation of Vision Transformers for Federated Learning
von: Halbe, Shaunak, et al.
Veröffentlicht: (2023)
von: Halbe, Shaunak, et al.
Veröffentlicht: (2023)
To Trust Or Not To Trust Your Vision-Language Model's Prediction
von: Dong, Hao, et al.
Veröffentlicht: (2025)
von: Dong, Hao, et al.
Veröffentlicht: (2025)
Compositional Entailment Learning for Hyperbolic Vision-Language Models
von: Pal, Avik, et al.
Veröffentlicht: (2024)
von: Pal, Avik, et al.
Veröffentlicht: (2024)
Tree of Attributes Prompt Learning for Vision-Language Models
von: Ding, Tong, et al.
Veröffentlicht: (2024)
von: Ding, Tong, et al.
Veröffentlicht: (2024)
Zero-Shot Action Generalization with Limited Observations
von: Alchihabi, Abdullah, et al.
Veröffentlicht: (2025)
von: Alchihabi, Abdullah, et al.
Veröffentlicht: (2025)
Parallel In-context Learning for Large Vision Language Models
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2026)
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2026)
VLSM-Adapter: Finetuning Vision-Language Segmentation Efficiently with Lightweight Blocks
von: Dhakal, Manish, et al.
Veröffentlicht: (2024)
von: Dhakal, Manish, et al.
Veröffentlicht: (2024)
Efficient Few-Shot Learning in Remote Sensing: Fusing Vision and Vision-Language Models
von: Chua, Jia Yun, et al.
Veröffentlicht: (2025)
von: Chua, Jia Yun, et al.
Veröffentlicht: (2025)
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers
von: Roschmann, Simon, et al.
Veröffentlicht: (2025)
von: Roschmann, Simon, et al.
Veröffentlicht: (2025)
Continual Learning in Vision-Language Models via Aligned Model Merging
von: Sokar, Ghada, et al.
Veröffentlicht: (2025)
von: Sokar, Ghada, et al.
Veröffentlicht: (2025)
VisMem: Latent Vision Memory Unlocks Potential of Vision-Language Models
von: Yu, Xinlei, et al.
Veröffentlicht: (2025)
von: Yu, Xinlei, et al.
Veröffentlicht: (2025)
AAPL: Adding Attributes to Prompt Learning for Vision-Language Models
von: Kim, Gahyeon, et al.
Veröffentlicht: (2024)
von: Kim, Gahyeon, et al.
Veröffentlicht: (2024)
Vision-Language Models Provide Promptable Representations for Reinforcement Learning
von: Chen, William, et al.
Veröffentlicht: (2024)
von: Chen, William, et al.
Veröffentlicht: (2024)
Pre-trained Vision-Language Models Learn Discoverable Visual Concepts
von: Zang, Yuan, et al.
Veröffentlicht: (2024)
von: Zang, Yuan, et al.
Veröffentlicht: (2024)
Decoupling Augmentation Bias in Prompt Learning for Vision-Language Models
von: Kim, Gahyeon, et al.
Veröffentlicht: (2025)
von: Kim, Gahyeon, et al.
Veröffentlicht: (2025)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
Effectiveness Assessment of Recent Large Vision-Language Models
von: Jiang, Yao, et al.
Veröffentlicht: (2024)
von: Jiang, Yao, et al.
Veröffentlicht: (2024)
Sim-CLIP: Unsupervised Siamese Adversarial Fine-Tuning for Robust and Semantically-Rich Vision-Language Models
von: Hossain, Md Zarif, et al.
Veröffentlicht: (2024)
von: Hossain, Md Zarif, et al.
Veröffentlicht: (2024)
Adapting Vision-Language Models Without Labels: A Comprehensive Survey
von: Dong, Hao, et al.
Veröffentlicht: (2025)
von: Dong, Hao, et al.
Veröffentlicht: (2025)
RankCLIP: Ranking-Consistent Language-Image Pretraining
von: Zhang, Yiming, et al.
Veröffentlicht: (2024)
von: Zhang, Yiming, et al.
Veröffentlicht: (2024)
Learning Unlabeled Clients Divergence for Federated Semi-Supervised Learning via Anchor Model Aggregation
von: Elbatel, Marawan, et al.
Veröffentlicht: (2024)
von: Elbatel, Marawan, et al.
Veröffentlicht: (2024)
LMFusion: Adapting Pretrained Language Models for Multimodal Generation
von: Shi, Weijia, et al.
Veröffentlicht: (2024)
von: Shi, Weijia, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
VLLFL: A Vision-Language Model Based Lightweight Federated Learning Framework for Smart Agriculture
von: Li, Long, et al.
Veröffentlicht: (2025) -
Federated Learning for Video Violence Detection: Complementary Roles of Lightweight CNNs and Vision-Language Models for Energy-Efficient Use
von: Thuau, Sébastien, et al.
Veröffentlicht: (2025) -
Pretrained Reversible Generation as Unsupervised Visual Representation Learning
von: Xue, Rongkun, et al.
Veröffentlicht: (2024) -
Enhancing Generalization in Vision-Language-Action Models by Preserving Pretrained Representations
von: Grover, Shresth, et al.
Veröffentlicht: (2025) -
Bi-Level Optimization for Single Domain Generalization
von: Heidari, Marzi, et al.
Veröffentlicht: (2026)