Gespeichert in:
| Hauptverfasser: | Yan, Weicai, Lin, Wang, Guo, Zirun, Wang, Ye, Feng, Fangming, Yang, Xiaoda, Wang, Zehan, Jin, Tao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2504.21423 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Low-rank Prompt Interaction for Continual Vision-Language Retrieval
von: Yan, Weicai, et al.
Veröffentlicht: (2025)
von: Yan, Weicai, et al.
Veröffentlicht: (2025)
Efficient Prompting for Continual Adaptation to Missing Modalities
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
Multimodal Prompt Learning with Missing Modalities for Sentiment Analysis and Emotion Recognition
von: Guo, Zirun, et al.
Veröffentlicht: (2024)
von: Guo, Zirun, et al.
Veröffentlicht: (2024)
ConceptGuard: Continual Personalized Text-to-Image Generation with Forgetting and Confusion Mitigation
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
APO: Enhancing Reasoning Ability of MLLMs via Asymmetric Policy Optimization
von: Hong, Minjie, et al.
Veröffentlicht: (2025)
von: Hong, Minjie, et al.
Veröffentlicht: (2025)
PromptRR: Diffusion Models as Prompt Generators for Single Image Reflection Removal
von: Wang, Tao, et al.
Veröffentlicht: (2024)
von: Wang, Tao, et al.
Veröffentlicht: (2024)
Smoothing the Shift: Towards Stable Test-Time Adaptation under Complex Multimodal Noises
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
PromptMRG: Diagnosis-Driven Prompts for Medical Report Generation
von: Jin, Haibo, et al.
Veröffentlicht: (2023)
von: Jin, Haibo, et al.
Veröffentlicht: (2023)
Generalizable Prompt Learning of CLIP: A Brief Overview
von: Cui, Fangming, et al.
Veröffentlicht: (2025)
von: Cui, Fangming, et al.
Veröffentlicht: (2025)
Mask-ControlNet: Higher-Quality Image Generation with An Additional Mask Prompt
von: Huang, Zhiqi, et al.
Veröffentlicht: (2024)
von: Huang, Zhiqi, et al.
Veröffentlicht: (2024)
Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization
von: Yang, Xi, et al.
Veröffentlicht: (2025)
von: Yang, Xi, et al.
Veröffentlicht: (2025)
Determinism of Randomness: Prompt-Residual Seed Shaping for Diffusion Generation
von: Yan, Song, et al.
Veröffentlicht: (2025)
von: Yan, Song, et al.
Veröffentlicht: (2025)
Text-Guided Multi-Scale Frequency Representation Adaptation
von: Yan, Weicai, et al.
Veröffentlicht: (2026)
von: Yan, Weicai, et al.
Veröffentlicht: (2026)
DiffVP: Differential Visual Semantic Prompting for LLM-Based CT Report Generation
von: Tian, Yuhe, et al.
Veröffentlicht: (2026)
von: Tian, Yuhe, et al.
Veröffentlicht: (2026)
LLM-I: LLMs are Naturally Interleaved Multimodal Creators
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
ImVideoEdit: Image-learning Video Editing via 2D Spatial Difference Attention Blocks
von: Xu, Jiayang, et al.
Veröffentlicht: (2026)
von: Xu, Jiayang, et al.
Veröffentlicht: (2026)
DiffAtlas: GenAI-fying Atlas Segmentation via Image-Mask Diffusion
von: Zhang, Hantao, et al.
Veröffentlicht: (2025)
von: Zhang, Hantao, et al.
Veröffentlicht: (2025)
Decoupling Stability and Plasticity for Multi-Modal Test-Time Adaptation
von: He, Yongbo, et al.
Veröffentlicht: (2026)
von: He, Yongbo, et al.
Veröffentlicht: (2026)
Advancing Prompt Learning through an External Layer
von: Cui, Fangming, et al.
Veröffentlicht: (2024)
von: Cui, Fangming, et al.
Veröffentlicht: (2024)
IP-SAM: Prompt-Space Conditioning for Prompt-Absent Camouflaged Object Detection
von: Zhang, Huiyao, et al.
Veröffentlicht: (2026)
von: Zhang, Huiyao, et al.
Veröffentlicht: (2026)
ProGiDiff: Prompt-Guided Diffusion-Based Medical Image Segmentation
von: Lin, Yuan, et al.
Veröffentlicht: (2026)
von: Lin, Yuan, et al.
Veröffentlicht: (2026)
DiffBench Meets DiffAgent: End-to-End LLM-Driven Diffusion Acceleration Code Generation
von: jiao, Jiajun, et al.
Veröffentlicht: (2026)
von: jiao, Jiajun, et al.
Veröffentlicht: (2026)
Sample-specific Masks for Visual Reprogramming-based Prompting
von: Cai, Chengyi, et al.
Veröffentlicht: (2024)
von: Cai, Chengyi, et al.
Veröffentlicht: (2024)
AutoPrompt: Automated Red-Teaming of Text-to-Image Models via LLM-Driven Adversarial Prompts
von: Liu, Yufan, et al.
Veröffentlicht: (2025)
von: Liu, Yufan, et al.
Veröffentlicht: (2025)
Prompt-Driven Feature Diffusion for Open-World Semi-Supervised Learning
von: Heidari, Marzi, et al.
Veröffentlicht: (2024)
von: Heidari, Marzi, et al.
Veröffentlicht: (2024)
Thinking with Programming Vision: Towards a Unified View for Thinking with Images
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
Denoise and Align: Diffusion-Driven Foreground Knowledge Prompting for Open-Vocabulary Temporal Action Detection
von: Zhu, Sa, et al.
Veröffentlicht: (2026)
von: Zhu, Sa, et al.
Veröffentlicht: (2026)
ChangeDiff: A Multi-Temporal Change Detection Data Generator with Flexible Text Prompts via Diffusion Model
von: Zang, Qi, et al.
Veröffentlicht: (2024)
von: Zang, Qi, et al.
Veröffentlicht: (2024)
Prompt Reinjection: Alleviating Prompt Forgetting in Multimodal Diffusion Transformers
von: Yao, Yuxuan, et al.
Veröffentlicht: (2026)
von: Yao, Yuxuan, et al.
Veröffentlicht: (2026)
TextDiff: Mask-Guided Residual Diffusion Models for Scene Text Image Super-Resolution
von: Liu, Baolin, et al.
Veröffentlicht: (2023)
von: Liu, Baolin, et al.
Veröffentlicht: (2023)
Vision-guided and Mask-enhanced Adaptive Denoising for Prompt-based Image Editing
von: Wang, Kejie, et al.
Veröffentlicht: (2024)
von: Wang, Kejie, et al.
Veröffentlicht: (2024)
Pixel-Perfect Depth with Semantics-Prompted Diffusion Transformers
von: Xu, Gangwei, et al.
Veröffentlicht: (2025)
von: Xu, Gangwei, et al.
Veröffentlicht: (2025)
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
Enhancing Prompt Following with Visual Control Through Training-Free Mask-Guided Diffusion
von: Chen, Hongyu, et al.
Veröffentlicht: (2024)
von: Chen, Hongyu, et al.
Veröffentlicht: (2024)
Transitive Vision-Language Prompt Learning for Domain Generalization
von: Wang, Liyuan, et al.
Veröffentlicht: (2024)
von: Wang, Liyuan, et al.
Veröffentlicht: (2024)
Regressor-Segmenter Mutual Prompt Learning for Crowd Counting
von: Guo, Mingyue, et al.
Veröffentlicht: (2023)
von: Guo, Mingyue, et al.
Veröffentlicht: (2023)
CCIS-Diff: A Generative Model with Stable Diffusion Prior for Controlled Colonoscopy Image Synthesis
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
Prompt-SID: Learning Structural Representation Prompt via Latent Diffusion for Single-Image Denoising
von: Li, Huaqiu, et al.
Veröffentlicht: (2025)
von: Li, Huaqiu, et al.
Veröffentlicht: (2025)
Diff-Restorer: Unleashing Visual Prompts for Diffusion-based Universal Image Restoration
von: Zhang, Yuhong, et al.
Veröffentlicht: (2024)
von: Zhang, Yuhong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Low-rank Prompt Interaction for Continual Vision-Language Retrieval
von: Yan, Weicai, et al.
Veröffentlicht: (2025) -
Efficient Prompting for Continual Adaptation to Missing Modalities
von: Guo, Zirun, et al.
Veröffentlicht: (2025) -
Multimodal Prompt Learning with Missing Modalities for Sentiment Analysis and Emotion Recognition
von: Guo, Zirun, et al.
Veröffentlicht: (2024) -
ConceptGuard: Continual Personalized Text-to-Image Generation with Forgetting and Confusion Mitigation
von: Guo, Zirun, et al.
Veröffentlicht: (2025) -
APO: Enhancing Reasoning Ability of MLLMs via Asymmetric Policy Optimization
von: Hong, Minjie, et al.
Veröffentlicht: (2025)