PTDiffusion: Free Lunch for Generating Optical Illusion Hidden Pictures with Phase-Transferred Diffusion Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Xiang, Yang, Shuai, Liu, Jiaying |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Free Lunch for Generating Effective Outlier Supervision
von: Pei, Sen, et al.
Veröffentlicht: (2023)
von: Pei, Sen, et al.
Veröffentlicht: (2023)
FreeCus: Free Lunch Subject-driven Customization in Diffusion Transformers
von: Zhang, Yanbing, et al.
Veröffentlicht: (2025)
von: Zhang, Yanbing, et al.
Veröffentlicht: (2025)
Shortcutting Pre-trained Flow Matching Diffusion Models is Almost Free Lunch
von: Cai, Xu, et al.
Veröffentlicht: (2025)
von: Cai, Xu, et al.
Veröffentlicht: (2025)
FreeInv: Free Lunch for Improving DDIM Inversion
von: Bao, Yuxiang, et al.
Veröffentlicht: (2025)
von: Bao, Yuxiang, et al.
Veröffentlicht: (2025)
Visual Anagrams: Generating Multi-View Optical Illusions with Diffusion Models
von: Geng, Daniel, et al.
Veröffentlicht: (2023)
von: Geng, Daniel, et al.
Veröffentlicht: (2023)
FreeStyle: Free Lunch for Text-guided Style Transfer using Diffusion Models
von: He, Feihong, et al.
Veröffentlicht: (2024)
von: He, Feihong, et al.
Veröffentlicht: (2024)
Free-Lunch Color-Texture Disentanglement for Stylized Image Generation
von: Qin, Jiang, et al.
Veröffentlicht: (2025)
von: Qin, Jiang, et al.
Veröffentlicht: (2025)
Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
FBSDiff: Plug-and-Play Frequency Band Substitution of Diffusion Features for Highly Controllable Text-Driven Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers
von: Zhao, Min, et al.
Veröffentlicht: (2025)
von: Zhao, Min, et al.
Veröffentlicht: (2025)
CineScale: Free Lunch in High-Resolution Cinematic Visual Generation
von: Qiu, Haonan, et al.
Veröffentlicht: (2025)
von: Qiu, Haonan, et al.
Veröffentlicht: (2025)
Free Lunch Alignment of Text-to-Image Diffusion Models without Preference Image Pairs
von: Xian, Jia Jun Cheng, et al.
Veröffentlicht: (2025)
von: Xian, Jia Jun Cheng, et al.
Veröffentlicht: (2025)
FreeCond: Free Lunch in the Input Conditions of Text-Guided Inpainting
von: Hsiao, Teng-Fang, et al.
Veröffentlicht: (2024)
von: Hsiao, Teng-Fang, et al.
Veröffentlicht: (2024)
IllusionVQA: A Challenging Optical Illusion Dataset for Vision Language Models
von: Shahgir, Haz Sameen, et al.
Veröffentlicht: (2024)
von: Shahgir, Haz Sameen, et al.
Veröffentlicht: (2024)
InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
von: Wang, Haofan, et al.
Veröffentlicht: (2024)
von: Wang, Haofan, et al.
Veröffentlicht: (2024)
Sparse-to-Dense: A Free Lunch for Lossless Acceleration of Video Understanding in LLMs
von: Zhang, Xuan, et al.
Veröffentlicht: (2025)
von: Zhang, Xuan, et al.
Veröffentlicht: (2025)
Energy-Calibrated VAE with Test Time Free Lunch
von: Luo, Yihong, et al.
Veröffentlicht: (2023)
von: Luo, Yihong, et al.
Veröffentlicht: (2023)
Leveraging Semantic Attribute Binding for Free-Lunch Color Control in Diffusion Models
von: Laria, Héctor, et al.
Veröffentlicht: (2025)
von: Laria, Héctor, et al.
Veröffentlicht: (2025)
Language-based Image Colorization: A Benchmark and Beyond
von: Li, Yifan, et al.
Veröffentlicht: (2025)
von: Li, Yifan, et al.
Veröffentlicht: (2025)
The Art of Deception: Color Visual Illusions and Diffusion Models
von: Gomez-Villa, Alex, et al.
Veröffentlicht: (2024)
von: Gomez-Villa, Alex, et al.
Veröffentlicht: (2024)
Switch EMA: A Free Lunch for Better Flatness and Sharpness
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
Spatial Reasoning is Not a Free Lunch: A Controlled Study on LLaVA
von: Alam, Nahid, et al.
Veröffentlicht: (2026)
von: Alam, Nahid, et al.
Veröffentlicht: (2026)
Alias-Free Latent Diffusion Models: Improving Fractional Shift Equivariance of Diffusion Latent Space
von: Zhou, Yifan, et al.
Veröffentlicht: (2025)
von: Zhou, Yifan, et al.
Veröffentlicht: (2025)
Free Lunch in Pathology Foundation Model: Task-specific Model Adaptation with Concept-Guided Feature Enhancement
von: Huang, Yanyan, et al.
Veröffentlicht: (2024)
von: Huang, Yanyan, et al.
Veröffentlicht: (2024)
Free Lunch for Stabilizing Rectified Flow Inversion
von: Wang, Chenru, et al.
Veröffentlicht: (2026)
von: Wang, Chenru, et al.
Veröffentlicht: (2026)
Illusion3D: 3D Multiview Illusion with 2D Diffusion Priors
von: Feng, Yue, et al.
Veröffentlicht: (2024)
von: Feng, Yue, et al.
Veröffentlicht: (2024)
FreeBind: Free Lunch in Unified Multimodal Space via Knowledge Fusion
von: Wang, Zehan, et al.
Veröffentlicht: (2024)
von: Wang, Zehan, et al.
Veröffentlicht: (2024)
Free Lunch for Unified Multimodal Models: Enhancing Generation via Reflective Rectification with Inherent Understanding
von: Jiang, Yibo, et al.
Veröffentlicht: (2026)
von: Jiang, Yibo, et al.
Veröffentlicht: (2026)
Factorized Diffusion: Perceptual Illusions by Noise Decomposition
von: Geng, Daniel, et al.
Veröffentlicht: (2024)
von: Geng, Daniel, et al.
Veröffentlicht: (2024)
Illusion-Aware Visual Preprocessing and Anti-Illusion Prompting for Classic Illusion Understanding in Vision-Language Models
von: Zha, Junli, et al.
Veröffentlicht: (2026)
von: Zha, Junli, et al.
Veröffentlicht: (2026)
FLOSS: Free Lunch in Open-vocabulary Semantic Segmentation
von: Benigmim, Yasser, et al.
Veröffentlicht: (2025)
von: Benigmim, Yasser, et al.
Veröffentlicht: (2025)
Exploring Data-Free LoRA Transferability for Video Diffusion Models
von: Wang, Yuchen, et al.
Veröffentlicht: (2026)
von: Wang, Yuchen, et al.
Veröffentlicht: (2026)
Self-Supervised Skeleton-Based Action Representation Learning: A Benchmark and Beyond
von: Zhang, Jiahang, et al.
Veröffentlicht: (2024)
von: Zhang, Jiahang, et al.
Veröffentlicht: (2024)
Intelligent Artistic Typography: A Comprehensive Review of Artistic Text Design and Generation
von: Bai, Yuhang, et al.
Veröffentlicht: (2024)
von: Bai, Yuhang, et al.
Veröffentlicht: (2024)
Freeplane: Unlocking Free Lunch in Triplane-Based Sparse-View Reconstruction Models
von: Sun, Wenqiang, et al.
Veröffentlicht: (2024)
von: Sun, Wenqiang, et al.
Veröffentlicht: (2024)
FreeMorph: Tuning-Free Generalized Image Morphing with Diffusion Model
von: Cao, Yukang, et al.
Veröffentlicht: (2025)
von: Cao, Yukang, et al.
Veröffentlicht: (2025)
IllusionBench+: A Large-scale and Comprehensive Benchmark for Visual Illusion Understanding in Vision-Language Models
von: Zhang, Yiming, et al.
Veröffentlicht: (2025)
von: Zhang, Yiming, et al.
Veröffentlicht: (2025)
Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation
von: Wang, Wenjing, et al.
Veröffentlicht: (2023)
von: Wang, Wenjing, et al.
Veröffentlicht: (2023)
Efficient Generation of Targeted and Transferable Adversarial Examples for Vision-Language Models Via Diffusion Models
von: Guo, Qi, et al.
Veröffentlicht: (2024)
von: Guo, Qi, et al.
Veröffentlicht: (2024)
No Free Lunch in Annotation either: An objective evaluation of foundation models for streamlining annotation in animal tracking
von: Mededovic, Emil, et al.
Veröffentlicht: (2025)
von: Mededovic, Emil, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Free Lunch for Generating Effective Outlier Supervision
von: Pei, Sen, et al.
Veröffentlicht: (2023) -
FreeCus: Free Lunch Subject-driven Customization in Diffusion Transformers
von: Zhang, Yanbing, et al.
Veröffentlicht: (2025) -
Shortcutting Pre-trained Flow Matching Diffusion Models is Almost Free Lunch
von: Cai, Xu, et al.
Veröffentlicht: (2025) -
FreeInv: Free Lunch for Improving DDIM Inversion
von: Bao, Yuxiang, et al.
Veröffentlicht: (2025) -
Visual Anagrams: Generating Multi-View Optical Illusions with Diffusion Models
von: Geng, Daniel, et al.
Veröffentlicht: (2023)