Disentangling Masked Autoencoders for Unsupervised Domain Generalization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, An, Wang, Han, Wang, Xiang, Chua, Tat-Seng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Compose Your Aesthetics: Empowering Text-to-Image Models with the Principles of Art
von: Jin, Zhe, et al.
Veröffentlicht: (2025)
von: Jin, Zhe, et al.
Veröffentlicht: (2025)
Beyond Surface Artifacts: Capturing Shared Latent Forgery Knowledge Across Modalities
von: Dou, Jingtong, et al.
Veröffentlicht: (2026)
von: Dou, Jingtong, et al.
Veröffentlicht: (2026)
Dysen-VDM: Empowering Dynamics-aware Text-to-Video Diffusion with LLMs
von: Fei, Hao, et al.
Veröffentlicht: (2023)
von: Fei, Hao, et al.
Veröffentlicht: (2023)
Discriminative Probing and Tuning for Text-to-Image Generation
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
Domain-Guided Masked Autoencoders for Unique Player Identification
von: Balaji, Bavesh, et al.
Veröffentlicht: (2024)
von: Balaji, Bavesh, et al.
Veröffentlicht: (2024)
Can I Trust Your Answer? Visually Grounded Video Question Answering
von: Xiao, Junbin, et al.
Veröffentlicht: (2023)
von: Xiao, Junbin, et al.
Veröffentlicht: (2023)
UniFGVC: Universal Training-Free Few-Shot Fine-Grained Vision Classification via Attribute-Aware Multimodal Retrieval
von: Guo, Hongyu, et al.
Veröffentlicht: (2025)
von: Guo, Hongyu, et al.
Veröffentlicht: (2025)
Suppressing Forgery-Specific Shortcuts for Generalizable Deepfake Detection
von: Wang, Yihui, et al.
Veröffentlicht: (2026)
von: Wang, Yihui, et al.
Veröffentlicht: (2026)
Masked Autoencoders Are Effective Tokenizers for Diffusion Models
von: Chen, Hao, et al.
Veröffentlicht: (2025)
von: Chen, Hao, et al.
Veröffentlicht: (2025)
MAGREF: Masked Guidance for Any-Reference Video Generation with Subject Disentanglement
von: Deng, Yufan, et al.
Veröffentlicht: (2025)
von: Deng, Yufan, et al.
Veröffentlicht: (2025)
GrabDAE: An Innovative Framework for Unsupervised Domain Adaptation Utilizing Grab-Mask and Denoise Auto-Encoder
von: Chen, Junzhou, et al.
Veröffentlicht: (2024)
von: Chen, Junzhou, et al.
Veröffentlicht: (2024)
MAPSeg: Unified Unsupervised Domain Adaptation for Heterogeneous Medical Image Segmentation Based on 3D Masked Autoencoding and Pseudo-Labeling
von: Zhang, Xuzhe, et al.
Veröffentlicht: (2023)
von: Zhang, Xuzhe, et al.
Veröffentlicht: (2023)
MUPA: Towards Multi-Path Agentic Reasoning for Grounded Video Question Answering
von: Dang, Jisheng, et al.
Veröffentlicht: (2025)
von: Dang, Jisheng, et al.
Veröffentlicht: (2025)
Gaussian Masked Autoencoders
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025)
von: Rajasegaran, Jathushan, et al.
Veröffentlicht: (2025)
MMDocBench: Benchmarking Large Vision-Language Models for Fine-Grained Visual Document Understanding
von: Zhu, Fengbin, et al.
Veröffentlicht: (2024)
von: Zhu, Fengbin, et al.
Veröffentlicht: (2024)
Revolutionizing Text-to-Image Retrieval as Autoregressive Token-to-Voken Generation
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
Any2Caption:Interpreting Any Condition to Caption for Controllable Video Generation
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
InstructVid2Vid: Controllable Video Editing with Natural Language Instructions
von: Qin, Bosheng, et al.
Veröffentlicht: (2023)
von: Qin, Bosheng, et al.
Veröffentlicht: (2023)
MURE: Hierarchical Multi-Resolution Encoding via Vision-Language Models for Visual Document Retrieval
von: Zhu, Fengbin, et al.
Veröffentlicht: (2026)
von: Zhu, Fengbin, et al.
Veröffentlicht: (2026)
TTOM: Test-Time Optimization and Memorization for Compositional Video Generation
von: Qu, Leigang, et al.
Veröffentlicht: (2025)
von: Qu, Leigang, et al.
Veröffentlicht: (2025)
SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
S4DL: Shift-sensitive Spatial-Spectral Disentangling Learning for Hyperspectral Image Unsupervised Domain Adaptation
von: Feng, Jie, et al.
Veröffentlicht: (2024)
von: Feng, Jie, et al.
Veröffentlicht: (2024)
MaskAdapt: Unsupervised Geometry-Aware Domain Adaptation Using Multimodal Contextual Learning and RGB-Depth Masking
von: Nadeem, Numair, et al.
Veröffentlicht: (2025)
von: Nadeem, Numair, et al.
Veröffentlicht: (2025)
Unsupervised Anomaly Detection in Brain MRI via Disentangled Anatomy Learning
von: Yang, Tao, et al.
Veröffentlicht: (2025)
von: Yang, Tao, et al.
Veröffentlicht: (2025)
Unsupervised Synthetic Image Attribution: Alignment and Disentanglement
von: Liu, Zongfang, et al.
Veröffentlicht: (2026)
von: Liu, Zongfang, et al.
Veröffentlicht: (2026)
Pseudo Labelling for Enhanced Masked Autoencoders
von: Nandam, Srinivasa Rao, et al.
Veröffentlicht: (2024)
von: Nandam, Srinivasa Rao, et al.
Veröffentlicht: (2024)
Generative Cross-Modal Retrieval: Memorizing Images in Multimodal Language Models for Retrieval and Beyond
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
Heterogeneous Generative Knowledge Distillation with Masked Image Modeling
von: Wang, Ziming, et al.
Veröffentlicht: (2023)
von: Wang, Ziming, et al.
Veröffentlicht: (2023)
TIGeR: Unifying Text-to-Image Generation and Retrieval with Large Multimodal Models
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
Hierarchical Semantic Correlation-Aware Masked Autoencoder for Unsupervised Audio-Visual Representation Learning
von: Zeng, Donghuo, et al.
Veröffentlicht: (2026)
von: Zeng, Donghuo, et al.
Veröffentlicht: (2026)
Rethinking Unsupervised Domain Adaptation for Semantic Segmentation
von: Wang, Zhijie, et al.
Veröffentlicht: (2022)
von: Wang, Zhijie, et al.
Veröffentlicht: (2022)
Self Pre-training with Topology- and Spatiality-aware Masked Autoencoders for 3D Medical Image Segmentation
von: Gu, Pengfei, et al.
Veröffentlicht: (2024)
von: Gu, Pengfei, et al.
Veröffentlicht: (2024)
Latent Anomaly Knowledge Excavation: Unveiling Sparse Sensitive Neurons in Vision-Language Models
von: Li, Shaotian, et al.
Veröffentlicht: (2026)
von: Li, Shaotian, et al.
Veröffentlicht: (2026)
MAESIL: Masked Autoencoder for Enhanced Self-supervised Medical Image Learning
von: Kim, Kyeonghun, et al.
Veröffentlicht: (2026)
von: Kim, Kyeonghun, et al.
Veröffentlicht: (2026)
Abductive Ego-View Accident Video Understanding for Safe Driving Perception
von: Fang, Jianwu, et al.
Veröffentlicht: (2024)
von: Fang, Jianwu, et al.
Veröffentlicht: (2024)
Diffusion Autoencoder for Unsupervised Artifact Restoration in Handheld Fundus Images
von: Palani, Mathumetha, et al.
Veröffentlicht: (2026)
von: Palani, Mathumetha, et al.
Veröffentlicht: (2026)
FlowEO: Generative Unsupervised Domain Adaptation for Earth Observation
von: Bellier, Georges Le, et al.
Veröffentlicht: (2025)
von: Bellier, Georges Le, et al.
Veröffentlicht: (2025)
Privacy Protection in MRI Scans Using 3D Masked Autoencoders
von: Van der Goten, Lennart Alexander, et al.
Veröffentlicht: (2023)
von: Van der Goten, Lennart Alexander, et al.
Veröffentlicht: (2023)
Where Culture Fades: Revealing the Cultural Gap in Text-to-Image Generation
von: Shi, Chuancheng, et al.
Veröffentlicht: (2025)
von: Shi, Chuancheng, et al.
Veröffentlicht: (2025)
NExT-OMNI: Towards Any-to-Any Omnimodal Foundation Models with Discrete Flow Matching
von: Luo, Run, et al.
Veröffentlicht: (2025)
von: Luo, Run, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Compose Your Aesthetics: Empowering Text-to-Image Models with the Principles of Art
von: Jin, Zhe, et al.
Veröffentlicht: (2025) -
Beyond Surface Artifacts: Capturing Shared Latent Forgery Knowledge Across Modalities
von: Dou, Jingtong, et al.
Veröffentlicht: (2026) -
Dysen-VDM: Empowering Dynamics-aware Text-to-Video Diffusion with LLMs
von: Fei, Hao, et al.
Veröffentlicht: (2023) -
Discriminative Probing and Tuning for Text-to-Image Generation
von: Qu, Leigang, et al.
Veröffentlicht: (2024) -
Domain-Guided Masked Autoencoders for Unique Player Identification
von: Balaji, Bavesh, et al.
Veröffentlicht: (2024)