Stable Diffusion Models are Secretly Good at Visual In-Context Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Oorloff, Trevine, Sindagi, Vishwanath, Bandara, Wele Gedara Chaminda, Shafahi, Ali, Ghiasi, Amin, Prakash, Charan, Ardekani, Reza |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Attention Prompt Tuning: Parameter-efficient Adaptation of Pre-trained Models for Spatiotemporal Modeling
by: Bandara, Wele Gedara Chaminda, et al.
Published: (2024)
by: Bandara, Wele Gedara Chaminda, et al.
Published: (2024)
DDPM-CD: Denoising Diffusion Probabilistic Models as Feature Extractors for Change Detection
by: Bandara, Wele Gedara Chaminda, et al.
Published: (2022)
by: Bandara, Wele Gedara Chaminda, et al.
Published: (2022)
$CrowdDiff$: Multi-hypothesis Crowd Density Estimation using Diffusion Models
by: Ranasinghe, Yasiru, et al.
Published: (2023)
by: Ranasinghe, Yasiru, et al.
Published: (2023)
Mitigating Hallucinations in Diffusion Models through Adaptive Attention Modulation
by: Oorloff, Trevine, et al.
Published: (2025)
by: Oorloff, Trevine, et al.
Published: (2025)
$\mathsf{CSMAE~}$:~Cataract Surgical Masked Autoencoder (MAE) based Pre-training
by: Shah, Nisarg A., et al.
Published: (2025)
by: Shah, Nisarg A., et al.
Published: (2025)
AVFF: Audio-Visual Feature Fusion for Video Deepfake Detection
by: Oorloff, Trevine, et al.
Published: (2024)
by: Oorloff, Trevine, et al.
Published: (2024)
Money Recognition for the Visually Impaired: A Case Study on Sri Lankan Banknotes
by: Bandara, Akshaan
Published: (2025)
by: Bandara, Akshaan
Published: (2025)
Good Seed Makes a Good Crop: Discovering Secret Seeds in Text-to-Image Diffusion Models
by: Xu, Katherine, et al.
Published: (2024)
by: Xu, Katherine, et al.
Published: (2024)
Post-Processing Mask-Based Table Segmentation for Structural Coordinate Extraction
by: Bandara, Suren
Published: (2025)
by: Bandara, Suren
Published: (2025)
Is Large-Scale Pretraining the Secret to Good Domain Generalization?
by: Teterwak, Piotr, et al.
Published: (2024)
by: Teterwak, Piotr, et al.
Published: (2024)
Unlocking Visual Secrets: Inverting Features with Diffusion Priors for Image Reconstruction
by: Zhang, Sai Qian, et al.
Published: (2024)
by: Zhang, Sai Qian, et al.
Published: (2024)
Video Summarisation with Incident and Context Information using Generative AI
by: De Silva, Ulindu, et al.
Published: (2025)
by: De Silva, Ulindu, et al.
Published: (2025)
Patch-Wise Self-Supervised Visual Representation Learning: A Fine-Grained Approach
by: Javidani, Ali, et al.
Published: (2023)
by: Javidani, Ali, et al.
Published: (2023)
Analogist: Out-of-the-box Visual In-Context Learning with Image Diffusion Model
by: Gu, Zheng, et al.
Published: (2024)
by: Gu, Zheng, et al.
Published: (2024)
High Dynamic Range Imaging via Visual Attention Modules
by: Omrani, Ali Reza, et al.
Published: (2023)
by: Omrani, Ali Reza, et al.
Published: (2023)
CARD: Correlation Aware Restoration with Diffusion
by: Nezakati, Niki, et al.
Published: (2025)
by: Nezakati, Niki, et al.
Published: (2025)
TDiff: Thermal Plug-And-Play Prior with Patch-Based Diffusion
by: Dashpute, Piyush, et al.
Published: (2025)
by: Dashpute, Piyush, et al.
Published: (2025)
VIRAL: Visual In-Context Reasoning via Analogy in Diffusion Transformers
by: Li, Zhiwen, et al.
Published: (2026)
by: Li, Zhiwen, et al.
Published: (2026)
Diffusion Models are Secretly Zero-Shot 3DGS Harmonizers
by: Skorokhodov, Vsevolod, et al.
Published: (2025)
by: Skorokhodov, Vsevolod, et al.
Published: (2025)
TRACE: Your Diffusion Model is Secretly an Instance Edge Detector
by: Jo, Sanghyun, et al.
Published: (2025)
by: Jo, Sanghyun, et al.
Published: (2025)
Scale-Wise VAR is Secretly Discrete Diffusion
by: Kumar, Amandeep, et al.
Published: (2025)
by: Kumar, Amandeep, et al.
Published: (2025)
mmAnomaly: Leveraging Visual Context for Robust Anomaly Detection in the Non-Visual World with mmWave Radar
by: Toha, Tarik Reza, et al.
Published: (2026)
by: Toha, Tarik Reza, et al.
Published: (2026)
Personal Visual Context Learning in Large Multimodal Models
by: Xue, Zihui, et al.
Published: (2026)
by: Xue, Zihui, et al.
Published: (2026)
Exploring Effective Factors for Improving Visual In-Context Learning
by: Sun, Yanpeng, et al.
Published: (2023)
by: Sun, Yanpeng, et al.
Published: (2023)
Conquering the Retina: Bringing Visual in-Context Learning to OCT
by: Negrini, Alessio, et al.
Published: (2025)
by: Negrini, Alessio, et al.
Published: (2025)
Enhancing Visual In-Context Learning by Multi-Faceted Fusion
by: Liao, Wenwen, et al.
Published: (2026)
by: Liao, Wenwen, et al.
Published: (2026)
Diffusion Model is Secretly a Training-free Open Vocabulary Semantic Segmenter
by: Wang, Jinglong, et al.
Published: (2023)
by: Wang, Jinglong, et al.
Published: (2023)
StableDub: Taming Diffusion Prior for Generalized and Efficient Visual Dubbing
by: Chen, Liyang, et al.
Published: (2025)
by: Chen, Liyang, et al.
Published: (2025)
AICL: Action In-Context Learning for Video Diffusion Model
by: Liu, Jianzhi, et al.
Published: (2024)
by: Liu, Jianzhi, et al.
Published: (2024)
Underlying Semantic Diffusion for Effective and Efficient In-Context Learning
by: Ji, Zhong, et al.
Published: (2025)
by: Ji, Zhong, et al.
Published: (2025)
Learning to Customize Text-to-Image Diffusion In Diverse Context
by: Kim, Taewook, et al.
Published: (2024)
by: Kim, Taewook, et al.
Published: (2024)
True Multimodal In-Context Learning Needs Attention to the Visual Context
by: Chen, Shuo, et al.
Published: (2025)
by: Chen, Shuo, et al.
Published: (2025)
Early detection of diabetes through transfer learning-based eye (vision) screening and improvement of machine learning model performance and advanced parameter setting algorithms
by: Yousefi, Mohammad Reza, et al.
Published: (2025)
by: Yousefi, Mohammad Reza, et al.
Published: (2025)
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks
by: Monsefi, Amin Karimi, et al.
Published: (2024)
by: Monsefi, Amin Karimi, et al.
Published: (2024)
VisualCloze: A Universal Image Generation Framework via Visual In-Context Learning
by: Li, Zhong-Yu, et al.
Published: (2025)
by: Li, Zhong-Yu, et al.
Published: (2025)
Your Diffusion Model is Secretly a Certifiably Robust Classifier
by: Chen, Huanran, et al.
Published: (2024)
by: Chen, Huanran, et al.
Published: (2024)
Your Pre-trained Diffusion Model Secretly Knows Restoration
by: Rajagopalan, Sudarshan, et al.
Published: (2026)
by: Rajagopalan, Sudarshan, et al.
Published: (2026)
Wound3DAssist: A Practical Framework for 3D Wound Assessment
by: Chierchia, Remi, et al.
Published: (2025)
by: Chierchia, Remi, et al.
Published: (2025)
Masked Diffusion Captioning for Visual Feature Learning
by: Feng, Chao, et al.
Published: (2025)
by: Feng, Chao, et al.
Published: (2025)
Towards Reliable and Holistic Visual In-Context Learning Prompt Selection
by: Wu, Wenxiao, et al.
Published: (2025)
by: Wu, Wenxiao, et al.
Published: (2025)
Similar Items
-
Attention Prompt Tuning: Parameter-efficient Adaptation of Pre-trained Models for Spatiotemporal Modeling
by: Bandara, Wele Gedara Chaminda, et al.
Published: (2024) -
DDPM-CD: Denoising Diffusion Probabilistic Models as Feature Extractors for Change Detection
by: Bandara, Wele Gedara Chaminda, et al.
Published: (2022) -
$CrowdDiff$: Multi-hypothesis Crowd Density Estimation using Diffusion Models
by: Ranasinghe, Yasiru, et al.
Published: (2023) -
Mitigating Hallucinations in Diffusion Models through Adaptive Attention Modulation
by: Oorloff, Trevine, et al.
Published: (2025) -
$\mathsf{CSMAE~}$:~Cataract Surgical Masked Autoencoder (MAE) based Pre-training
by: Shah, Nisarg A., et al.
Published: (2025)