Latte: Collaborative Test-Time Adaptation of Vision-Language Models in Federated Learning

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Bao, Wenxuan, Deng, Ruxi, Qiu, Ruizhong, Wei, Tianxin, Tong, Hanghang, He, Jingrui
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866911081583083520
author Bao, Wenxuan
Deng, Ruxi
Qiu, Ruizhong
Wei, Tianxin
Tong, Hanghang
He, Jingrui
author_facet Bao, Wenxuan
Deng, Ruxi
Qiu, Ruizhong
Wei, Tianxin
Tong, Hanghang
He, Jingrui
contents Test-time adaptation with pre-trained vision-language models has gained increasing attention for addressing distribution shifts during testing. Among these approaches, memory-based algorithms stand out due to their training-free nature and ability to leverage historical test data. However, existing test-time adaptation methods are typically designed for a single domain with abundant data. In decentralized settings such as federated learning, applying these methods individually to each client suffers from limited test data, while directly sharing a single global memory via the server prevents proper personalization to each client's unique distribution. To address this, we propose Latte, a novel framework where each client maintains a local memory to store embeddings from its own historical test data and an external memory to store class prototypes from other relevant clients. During communication, each client retrieves prototypes from similar clients under the server's coordination to expand its memory. For local adaptation, Latte utilizes both embedding similarity and uncertainty to enhance model performance. Our theoretical analysis shows that Latte effectively leverages in-distribution clients while remaining robust to out-of-distribution clients. Extensive experiments on domain adaptation and corruption benchmarks validate that Latte achieves superior performance in decentralized settings, while introducing only negligible communication and computation costs. Our code is available at https://github.com/baowenxuan/Latte .
format Preprint
id arxiv_https___arxiv_org_abs_2507_21494
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Latte: Collaborative Test-Time Adaptation of Vision-Language Models in Federated Learning
Bao, Wenxuan
Deng, Ruxi
Qiu, Ruizhong
Wei, Tianxin
Tong, Hanghang
He, Jingrui
Machine Learning
Test-time adaptation with pre-trained vision-language models has gained increasing attention for addressing distribution shifts during testing. Among these approaches, memory-based algorithms stand out due to their training-free nature and ability to leverage historical test data. However, existing test-time adaptation methods are typically designed for a single domain with abundant data. In decentralized settings such as federated learning, applying these methods individually to each client suffers from limited test data, while directly sharing a single global memory via the server prevents proper personalization to each client's unique distribution. To address this, we propose Latte, a novel framework where each client maintains a local memory to store embeddings from its own historical test data and an external memory to store class prototypes from other relevant clients. During communication, each client retrieves prototypes from similar clients under the server's coordination to expand its memory. For local adaptation, Latte utilizes both embedding similarity and uncertainty to enhance model performance. Our theoretical analysis shows that Latte effectively leverages in-distribution clients while remaining robust to out-of-distribution clients. Extensive experiments on domain adaptation and corruption benchmarks validate that Latte achieves superior performance in decentralized settings, while introducing only negligible communication and computation costs. Our code is available at https://github.com/baowenxuan/Latte .
title Latte: Collaborative Test-Time Adaptation of Vision-Language Models in Federated Learning
topic Machine Learning
url https://arxiv.org/abs/2507.21494