Patch-enhanced Mask Encoder Prompt Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Shusong, Liu, Peiye |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Patch-wise Auto-Encoder for Visual Anomaly Detection
di: Cui, Yajie, et al.
Pubblicazione: (2023)
di: Cui, Yajie, et al.
Pubblicazione: (2023)
Cluster and Predict Latent Patches for Improved Masked Image Modeling
di: Darcet, Timothée, et al.
Pubblicazione: (2025)
di: Darcet, Timothée, et al.
Pubblicazione: (2025)
General Purpose Image Encoder DINOv2 for Medical Image Registration
di: Song, Xinrui, et al.
Pubblicazione: (2024)
di: Song, Xinrui, et al.
Pubblicazione: (2024)
HiFA: High-fidelity Text-to-3D Generation with Advanced Diffusion Guidance
di: Zhu, Junzhe, et al.
Pubblicazione: (2023)
di: Zhu, Junzhe, et al.
Pubblicazione: (2023)
Concept-Based Masking: A Patch-Agnostic Defense Against Adversarial Patch Attacks
di: Mehrotra, Ayushi, et al.
Pubblicazione: (2025)
di: Mehrotra, Ayushi, et al.
Pubblicazione: (2025)
Symmetric masking strategy enhances the performance of Masked Image Modeling
di: Nguyen, Khanh-Binh, et al.
Pubblicazione: (2024)
di: Nguyen, Khanh-Binh, et al.
Pubblicazione: (2024)
CertMask: Certifiable Defense Against Adversarial Patches via Theoretically Optimal Mask Coverage
di: Lyu, Xuntao, et al.
Pubblicazione: (2025)
di: Lyu, Xuntao, et al.
Pubblicazione: (2025)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
Diffusion Model Patching via Mixture-of-Prompts
di: Ham, Seokil, et al.
Pubblicazione: (2024)
di: Ham, Seokil, et al.
Pubblicazione: (2024)
GrabDAE: An Innovative Framework for Unsupervised Domain Adaptation Utilizing Grab-Mask and Denoise Auto-Encoder
di: Chen, Junzhou, et al.
Pubblicazione: (2024)
di: Chen, Junzhou, et al.
Pubblicazione: (2024)
One Layer Is Enough: Adapting Pretrained Visual Encoders for Image Generation
di: Gao, Yuan, et al.
Pubblicazione: (2025)
di: Gao, Yuan, et al.
Pubblicazione: (2025)
Refer to Any Segmentation Mask Group With Vision-Language Prompts
di: Cao, Shengcao, et al.
Pubblicazione: (2025)
di: Cao, Shengcao, et al.
Pubblicazione: (2025)
MCGM: Mask Conditional Text-to-Image Generative Model
di: Skaik, Rami, et al.
Pubblicazione: (2024)
di: Skaik, Rami, et al.
Pubblicazione: (2024)
Heterogeneous Generative Knowledge Distillation with Masked Image Modeling
di: Wang, Ziming, et al.
Pubblicazione: (2023)
di: Wang, Ziming, et al.
Pubblicazione: (2023)
MIMIR: Masked Image Modeling for Mutual Information-based Adversarial Robustness
di: Xu, Xiaoyun, et al.
Pubblicazione: (2023)
di: Xu, Xiaoyun, et al.
Pubblicazione: (2023)
Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking
di: Zheng, Zirui, et al.
Pubblicazione: (2025)
di: Zheng, Zirui, et al.
Pubblicazione: (2025)
Car Damage Detection and Patch-to-Patch Self-supervised Image Alignment
di: Chen, Hanxiao
Pubblicazione: (2024)
di: Chen, Hanxiao
Pubblicazione: (2024)
SelfSwapper: Self-Supervised Face Swapping via Shape Agnostic Masked AutoEncoder
di: Lee, Jaeseong, et al.
Pubblicazione: (2024)
di: Lee, Jaeseong, et al.
Pubblicazione: (2024)
Pixel-Aligned Multi-View Generation with Depth Guided Decoder
di: Tang, Zhenggang, et al.
Pubblicazione: (2024)
di: Tang, Zhenggang, et al.
Pubblicazione: (2024)
IPDN: Image-enhanced Prompt Decoding Network for 3D Referring Expression Segmentation
di: Chen, Qi, et al.
Pubblicazione: (2025)
di: Chen, Qi, et al.
Pubblicazione: (2025)
Vanishing Depth: A Depth Adapter with Positional Depth Encoding for Generalized Image Encoders
di: Koch, Paul, et al.
Pubblicazione: (2025)
di: Koch, Paul, et al.
Pubblicazione: (2025)
Dynamic Prompt Optimizing for Text-to-Image Generation
di: Mo, Wenyi, et al.
Pubblicazione: (2024)
di: Mo, Wenyi, et al.
Pubblicazione: (2024)
Dual form Complementary Masking for Domain-Adaptive Image Segmentation
di: Wang, Jiawen, et al.
Pubblicazione: (2025)
di: Wang, Jiawen, et al.
Pubblicazione: (2025)
Near, far: Patch-ordering enhances vision foundation models' scene understanding
di: Pariza, Valentinos, et al.
Pubblicazione: (2024)
di: Pariza, Valentinos, et al.
Pubblicazione: (2024)
AEMIM: Adversarial Examples Meet Masked Image Modeling
di: Xiang, Wenzhao, et al.
Pubblicazione: (2024)
di: Xiang, Wenzhao, et al.
Pubblicazione: (2024)
Patch Progression Masked Autoencoder with Fusion CNN Network for Classifying Evolution Between Two Pairs of 2D OCT Slices
di: Zhang, Philippe, et al.
Pubblicazione: (2025)
di: Zhang, Philippe, et al.
Pubblicazione: (2025)
MarDini: Masked Autoregressive Diffusion for Video Generation at Scale
di: Liu, Haozhe, et al.
Pubblicazione: (2024)
di: Liu, Haozhe, et al.
Pubblicazione: (2024)
HU-based Foreground Masking for 3D Medical Masked Image Modeling
di: Lee, Jin, et al.
Pubblicazione: (2025)
di: Lee, Jin, et al.
Pubblicazione: (2025)
HMAR: Efficient Hierarchical Masked Auto-Regressive Image Generation
di: Kumbong, Hermann, et al.
Pubblicazione: (2025)
di: Kumbong, Hermann, et al.
Pubblicazione: (2025)
TIER: Text-Image Encoder-based Regression for AIGC Image Quality Assessment
di: Yuan, Jiquan, et al.
Pubblicazione: (2024)
di: Yuan, Jiquan, et al.
Pubblicazione: (2024)
Learning Image Priors through Patch-based Diffusion Models for Solving Inverse Problems
di: Hu, Jason, et al.
Pubblicazione: (2024)
di: Hu, Jason, et al.
Pubblicazione: (2024)
Efficient Vision-and-Language Pre-training with Text-Relevant Image Patch Selection
di: Ye, Wei, et al.
Pubblicazione: (2024)
di: Ye, Wei, et al.
Pubblicazione: (2024)
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation
di: Yariv, Guy, et al.
Pubblicazione: (2025)
di: Yariv, Guy, et al.
Pubblicazione: (2025)
Reverse Prompt: Cracking the Recipe Inside Text-to-Image Generation
di: Ren, Zhiyao, et al.
Pubblicazione: (2025)
di: Ren, Zhiyao, et al.
Pubblicazione: (2025)
Long-Text-to-Image Generation via Compositional Prompt Decomposition
di: Huang, Jen-Yuan, et al.
Pubblicazione: (2026)
di: Huang, Jen-Yuan, et al.
Pubblicazione: (2026)
DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer
di: Wu, Yecheng, et al.
Pubblicazione: (2025)
di: Wu, Yecheng, et al.
Pubblicazione: (2025)
PhyPrompt: RL-based Prompt Refinement for Physically Plausible Text-to-Video Generation
di: Wu, Shang, et al.
Pubblicazione: (2026)
di: Wu, Shang, et al.
Pubblicazione: (2026)
MaskAnyNet: Rethinking Masked Image Regions as Valuable Information in Supervised Learning
di: Hong, Jingshan, et al.
Pubblicazione: (2025)
di: Hong, Jingshan, et al.
Pubblicazione: (2025)
GenPilot: A Multi-Agent System for Test-Time Prompt Optimization in Image Generation
di: Ye, Wen, et al.
Pubblicazione: (2025)
di: Ye, Wen, et al.
Pubblicazione: (2025)
TRIPS: Efficient Vision-and-Language Pre-training with Text-Relevant Image Patch Selection
di: Jiang, Chaoya, et al.
Pubblicazione: (2023)
di: Jiang, Chaoya, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Patch-wise Auto-Encoder for Visual Anomaly Detection
di: Cui, Yajie, et al.
Pubblicazione: (2023) -
Cluster and Predict Latent Patches for Improved Masked Image Modeling
di: Darcet, Timothée, et al.
Pubblicazione: (2025) -
General Purpose Image Encoder DINOv2 for Medical Image Registration
di: Song, Xinrui, et al.
Pubblicazione: (2024) -
HiFA: High-fidelity Text-to-3D Generation with Advanced Diffusion Guidance
di: Zhu, Junzhe, et al.
Pubblicazione: (2023) -
Concept-Based Masking: A Patch-Agnostic Defense Against Adversarial Patch Attacks
di: Mehrotra, Ayushi, et al.
Pubblicazione: (2025)