SSAM: Self-Supervised Association Modeling for Test-Time Adaption
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yaxiong, Zhang, Zhenqiang, Cheng, Lechao, Zhong, Zhun, Guo, Dan, Wang, Meng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Micro-Action Recognition with Limited Annotations: An Asynchronous Pseudo Labeling and Training Approach
von: Zhang, Yan, et al.
Veröffentlicht: (2025)
von: Zhang, Yan, et al.
Veröffentlicht: (2025)
TDEdit: A Unified Diffusion Framework for Text-Drag Guided Image Manipulation
von: Wang, Qihang, et al.
Veröffentlicht: (2025)
von: Wang, Qihang, et al.
Veröffentlicht: (2025)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
von: Zhang, Zhenxing, et al.
Veröffentlicht: (2024)
von: Zhang, Zhenxing, et al.
Veröffentlicht: (2024)
Knowledge Swapping via Learning and Unlearning
von: Xing, Mingyu, et al.
Veröffentlicht: (2025)
von: Xing, Mingyu, et al.
Veröffentlicht: (2025)
EntityCLIP: Entity-Centric Image-Text Matching via Multimodal Attentive Contrastive Learning
von: Wang, Yaxiong, et al.
Veröffentlicht: (2024)
von: Wang, Yaxiong, et al.
Veröffentlicht: (2024)
Towards Fine-Grained Emotion Understanding via Skeleton-Based Micro-Gesture Recognition
von: Xu, Hao, et al.
Veröffentlicht: (2025)
von: Xu, Hao, et al.
Veröffentlicht: (2025)
FedHPL: Efficient Heterogeneous Federated Learning with Prompt Tuning and Logit Distillation
von: Ma, Yuting, et al.
Veröffentlicht: (2024)
von: Ma, Yuting, et al.
Veröffentlicht: (2024)
Beyond Artificial Misalignment: Detecting and Grounding Semantic-Coordinated Multimodal Manipulations
von: Shen, Jinjie, et al.
Veröffentlicht: (2025)
von: Shen, Jinjie, et al.
Veröffentlicht: (2025)
Towards Unified Multimodal Misinformation Detection in Social Media: A Benchmark Dataset and Baseline
von: Li, Haiyang, et al.
Veröffentlicht: (2025)
von: Li, Haiyang, et al.
Veröffentlicht: (2025)
Motion is the Choreographer: Learning Latent Pose Dynamics for Seamless Sign Language Generation
von: He, Jiayi, et al.
Veröffentlicht: (2025)
von: He, Jiayi, et al.
Veröffentlicht: (2025)
Text-Driven Diffusion Model for Sign Language Production
von: He, Jiayi, et al.
Veröffentlicht: (2025)
von: He, Jiayi, et al.
Veröffentlicht: (2025)
OmniVL-Guard: Towards Unified Vision-Language Forgery Detection and Grounding via Balanced RL
von: Shen, Jinjie, et al.
Veröffentlicht: (2026)
von: Shen, Jinjie, et al.
Veröffentlicht: (2026)
Frequency Decoupling for Motion Magnification via Multi-Level Isomorphic Architecture
von: Wang, Fei, et al.
Veröffentlicht: (2024)
von: Wang, Fei, et al.
Veröffentlicht: (2024)
CanonSLR: Canonical-View Guided Multi-View Continuous Sign Language Recognition
von: Wang, Xu, et al.
Veröffentlicht: (2026)
von: Wang, Xu, et al.
Veröffentlicht: (2026)
OmniVL-Guard Pro: A Tool-Augmented Agent for Omnibus Vision-Language Forensics
von: Shen, Jinjie, et al.
Veröffentlicht: (2026)
von: Shen, Jinjie, et al.
Veröffentlicht: (2026)
Calibrating Undisciplined Over-Smoothing in Transformer for Weakly Supervised Semantic Segmentation
von: Cheng, Lechao, et al.
Veröffentlicht: (2023)
von: Cheng, Lechao, et al.
Veröffentlicht: (2023)
Correspondence as Video: Test-Time Adaption on SAM2 for Reference Segmentation in the Wild
von: Wang, Haoran, et al.
Veröffentlicht: (2025)
von: Wang, Haoran, et al.
Veröffentlicht: (2025)
Prior-Constrained Association Learning for Fine-Grained Generalized Category Discovery
von: Wang, Menglin, et al.
Veröffentlicht: (2025)
von: Wang, Menglin, et al.
Veröffentlicht: (2025)
SSAM: Singular Subspace Alignment for Merging Multimodal Large Language Models
von: Reza, Md Kaykobad, et al.
Veröffentlicht: (2026)
von: Reza, Md Kaykobad, et al.
Veröffentlicht: (2026)
Self-Classification Enhancement and Correction for Weakly Supervised Object Detection
von: Yin, Yufei, et al.
Veröffentlicht: (2025)
von: Yin, Yufei, et al.
Veröffentlicht: (2025)
Advancing Weakly-Supervised Audio-Visual Video Parsing via Segment-wise Pseudo Labeling
von: Zhou, Jinxing, et al.
Veröffentlicht: (2024)
von: Zhou, Jinxing, et al.
Veröffentlicht: (2024)
Adapted-MoE: Mixture of Experts with Test-Time Adaption for Anomaly Detection
von: Lei, Tianwu, et al.
Veröffentlicht: (2024)
von: Lei, Tianwu, et al.
Veröffentlicht: (2024)
Pretrain-then-Adapt: Uncertainty-Aware Test-Time Adaptation for Text-based Person Search
von: Zhang, Jiahao, et al.
Veröffentlicht: (2026)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2026)
Exploiting Intermediate Reconstructions in Optical Coherence Tomography for Test-Time Adaption of Medical Image Segmentation
von: Pinetz, Thomas, et al.
Veröffentlicht: (2026)
von: Pinetz, Thomas, et al.
Veröffentlicht: (2026)
Shaping a Stabilized Video by Mitigating Unintended Changes for Concept-Augmented Video Editing
von: Guo, Mingce, et al.
Veröffentlicht: (2024)
von: Guo, Mingce, et al.
Veröffentlicht: (2024)
Pyramid Pixel Context Adaption Network for Medical Image Classification with Supervised Contrastive Learning
von: Zhang, Xiaoqing, et al.
Veröffentlicht: (2023)
von: Zhang, Xiaoqing, et al.
Veröffentlicht: (2023)
Discrete to Continuous: Generating Smooth Transition Poses from Sign Language Observation
von: Tang, Shengeng, et al.
Veröffentlicht: (2024)
von: Tang, Shengeng, et al.
Veröffentlicht: (2024)
Global-Local Medical SAM Adaptor Based on Full Adaption
von: Wang, Meng, et al.
Veröffentlicht: (2024)
von: Wang, Meng, et al.
Veröffentlicht: (2024)
Infrared and Visible Image Fusion: From Data Compatibility to Task Adaption
von: Liu, Jinyuan, et al.
Veröffentlicht: (2025)
von: Liu, Jinyuan, et al.
Veröffentlicht: (2025)
PSVMA+: Exploring Multi-granularity Semantic-visual Adaption for Generalized Zero-shot Learning
von: Liu, Man, et al.
Veröffentlicht: (2024)
von: Liu, Man, et al.
Veröffentlicht: (2024)
Revisiting the Power of Prompt for Visual Tuning
von: Wang, Yuzhu, et al.
Veröffentlicht: (2024)
von: Wang, Yuzhu, et al.
Veröffentlicht: (2024)
Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering
von: Wang, Xu, et al.
Veröffentlicht: (2025)
von: Wang, Xu, et al.
Veröffentlicht: (2025)
Navigating Semantic Drift in Task-Agnostic Class-Incremental Learning
von: Wu, Fangwen, et al.
Veröffentlicht: (2025)
von: Wu, Fangwen, et al.
Veröffentlicht: (2025)
Decoupled Training with Local Reinforcement Fine-Tuning in Federated Learning
von: Ma, Yuting, et al.
Veröffentlicht: (2026)
von: Ma, Yuting, et al.
Veröffentlicht: (2026)
Noise Adaption Network for Morse Code Image Classification
von: Wang, Xiaxia, et al.
Veröffentlicht: (2024)
von: Wang, Xiaxia, et al.
Veröffentlicht: (2024)
When Test-Time Adaptation Meets Self-Supervised Models
von: Han, Jisu, et al.
Veröffentlicht: (2025)
von: Han, Jisu, et al.
Veröffentlicht: (2025)
Parameter-Efficient Domain Adaption for CSI Crowd-Counting via Self-Supervised Learning with Adapter Modules
von: Custance, Oliver, et al.
Veröffentlicht: (2026)
von: Custance, Oliver, et al.
Veröffentlicht: (2026)
BiPC: Bidirectional Probability Calibration for Unsupervised Domain Adaption
von: Zhou, Wenlve, et al.
Veröffentlicht: (2024)
von: Zhou, Wenlve, et al.
Veröffentlicht: (2024)
Multi-Scale Global-Instance Prompt Tuning for Continual Test-time Adaptation in Medical Image Segmentation
von: Li, Lingrui, et al.
Veröffentlicht: (2026)
von: Li, Lingrui, et al.
Veröffentlicht: (2026)
Distribution Aligned Semantics Adaption for Lifelong Person Re-Identification
von: Wang, Qizao, et al.
Veröffentlicht: (2024)
von: Wang, Qizao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards Micro-Action Recognition with Limited Annotations: An Asynchronous Pseudo Labeling and Training Approach
von: Zhang, Yan, et al.
Veröffentlicht: (2025) -
TDEdit: A Unified Diffusion Framework for Text-Drag Guided Image Manipulation
von: Wang, Qihang, et al.
Veröffentlicht: (2025) -
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
von: Zhang, Zhenxing, et al.
Veröffentlicht: (2024) -
Knowledge Swapping via Learning and Unlearning
von: Xing, Mingyu, et al.
Veröffentlicht: (2025) -
EntityCLIP: Entity-Centric Image-Text Matching via Multimodal Attentive Contrastive Learning
von: Wang, Yaxiong, et al.
Veröffentlicht: (2024)