IP-Prompter: Training-Free Theme-Specific Image Generation via Dynamic Visual Prompting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yuxin, Luo, Minyan, Dong, Weiming, Yang, Xiao, Huang, Haibin, Ma, Chongyang, Deussen, Oliver, Lee, Tong-Yee, Xu, Changsheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dance-to-Music Generation with Encoder-based Textual Inversion
von: Li, Sifei, et al.
Veröffentlicht: (2024)
von: Li, Sifei, et al.
Veröffentlicht: (2024)
Self-Reasoning Agentic Framework for Narrative Product Grid-Collage Generation
von: Luo, Minyan, et al.
Veröffentlicht: (2026)
von: Luo, Minyan, et al.
Veröffentlicht: (2026)
CreativeSynth: Cross-Art-Attention for Artistic Image Synthesis with Multimodal Diffusion
von: Huang, Nisha, et al.
Veröffentlicht: (2024)
von: Huang, Nisha, et al.
Veröffentlicht: (2024)
MotionCrafter: One-Shot Motion Customization of Diffusion Models
von: Zhang, Yuxin, et al.
Veröffentlicht: (2023)
von: Zhang, Yuxin, et al.
Veröffentlicht: (2023)
Break-for-Make: Modular Low-Rank Adaptations for Composable Content-Style Customization
von: Xu, Yu, et al.
Veröffentlicht: (2024)
von: Xu, Yu, et al.
Veröffentlicht: (2024)
SongEcho: Towards Cover Song Generation via Instance-Adaptive Element-wise Linear Modulation
von: Li, Sifei, et al.
Veröffentlicht: (2026)
von: Li, Sifei, et al.
Veröffentlicht: (2026)
HeadRouter: A Training-free Image Editing Framework for MM-DiTs by Adaptively Routing Attention Heads
von: Xu, Yu, et al.
Veröffentlicht: (2024)
von: Xu, Yu, et al.
Veröffentlicht: (2024)
VisualPrompter: Semantic-Aware Prompt Optimization with Visual Feedback for Text-to-Image Synthesis
von: Wu, Shiyu, et al.
Veröffentlicht: (2025)
von: Wu, Shiyu, et al.
Veröffentlicht: (2025)
Beyond Pixels: Visual Metaphor Transfer via Schema-Driven Agentic Reasoning
von: Xu, Yu, et al.
Veröffentlicht: (2026)
von: Xu, Yu, et al.
Veröffentlicht: (2026)
Music Style Transfer with Time-Varying Inversion of Diffusion Models
von: Li, Sifei, et al.
Veröffentlicht: (2024)
von: Li, Sifei, et al.
Veröffentlicht: (2024)
DiffPrompter: Differentiable Implicit Visual Prompts for Semantic-Segmentation in Adverse Conditions
von: Kalwar, Sanket, et al.
Veröffentlicht: (2023)
von: Kalwar, Sanket, et al.
Veröffentlicht: (2023)
AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs
von: Paulus, Anselm, et al.
Veröffentlicht: (2024)
von: Paulus, Anselm, et al.
Veröffentlicht: (2024)
A Survey on Cross-Modal Interaction Between Music and Multimodal Data
von: Li, Sifei, et al.
Veröffentlicht: (2025)
von: Li, Sifei, et al.
Veröffentlicht: (2025)
AtelierEval: Agentic Evaluation of Humans & LLMs as Text-to-Image Prompters
von: Luo, Hanjun, et al.
Veröffentlicht: (2026)
von: Luo, Hanjun, et al.
Veröffentlicht: (2026)
GraphPrompter: Multi-stage Adaptive Prompt Optimization for Graph In-Context Learning
von: Lv, Rui, et al.
Veröffentlicht: (2025)
von: Lv, Rui, et al.
Veröffentlicht: (2025)
Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models
von: Wang, Wenbin, et al.
Veröffentlicht: (2024)
von: Wang, Wenbin, et al.
Veröffentlicht: (2024)
CoPrompter: User-Centric Evaluation of LLM Instruction Alignment for Improved Prompt Engineering
von: Joshi, Ishika, et al.
Veröffentlicht: (2024)
von: Joshi, Ishika, et al.
Veröffentlicht: (2024)
AttriPrompter: Auto-Prompting with Attribute Semantics for Zero-shot Nuclei Detection via Visual-Language Pre-trained Models
von: Wu, Yongjian, et al.
Veröffentlicht: (2024)
von: Wu, Yongjian, et al.
Veröffentlicht: (2024)
UADAPy: An Uncertainty-Aware Visualization and Analysis Toolbox
von: Paetzold, Patrick, et al.
Veröffentlicht: (2024)
von: Paetzold, Patrick, et al.
Veröffentlicht: (2024)
GPS: General Per-Sample Prompter
von: Batorski, Pawel, et al.
Veröffentlicht: (2025)
von: Batorski, Pawel, et al.
Veröffentlicht: (2025)
Mixture of In-Context Prompters for Tabular PFNs
von: Xu, Derek, et al.
Veröffentlicht: (2024)
von: Xu, Derek, et al.
Veröffentlicht: (2024)
LLM as Prompter: Low-resource Inductive Reasoning on Arbitrary Knowledge Graphs
von: Wang, Kai, et al.
Veröffentlicht: (2024)
von: Wang, Kai, et al.
Veröffentlicht: (2024)
Identity-Preserving Text-to-Video Generation via Training-Free Prompt, Image, and Guidance Enhancement
von: Gao, Jiayi, et al.
Veröffentlicht: (2025)
von: Gao, Jiayi, et al.
Veröffentlicht: (2025)
Prompter: Utilizing Large Language Model Prompting for a Data Efficient Embodied Instruction Following
von: Inoue, Yuki, et al.
Veröffentlicht: (2022)
von: Inoue, Yuki, et al.
Veröffentlicht: (2022)
VoicePrompter: Robust Zero-Shot Voice Conversion with Voice Prompt and Conditional Flow Matching
von: Choi, Ha-Yeong, et al.
Veröffentlicht: (2025)
von: Choi, Ha-Yeong, et al.
Veröffentlicht: (2025)
DyFo: A Training-Free Dynamic Focus Visual Search for Enhancing LMMs in Fine-Grained Visual Understanding
von: Li, Geng, et al.
Veröffentlicht: (2025)
von: Li, Geng, et al.
Veröffentlicht: (2025)
In-Context Brush: Zero-shot Customized Subject Insertion with Context-Aware Latent Space Manipulation
von: Xu, Yu, et al.
Veröffentlicht: (2025)
von: Xu, Yu, et al.
Veröffentlicht: (2025)
Safeguarding Vision-Language Models Against Patched Visual Prompt Injectors
von: Sun, Jiachen, et al.
Veröffentlicht: (2024)
von: Sun, Jiachen, et al.
Veröffentlicht: (2024)
DynaIP: Dynamic Image Prompt Adapter for Scalable Zero-shot Personalized Text-to-Image Generation
von: Wang, Zhizhong, et al.
Veröffentlicht: (2025)
von: Wang, Zhizhong, et al.
Veröffentlicht: (2025)
Three Heads Are Better Than One: Complementary Experts for Long-Tailed Semi-supervised Learning
von: Ma, Chengcheng, et al.
Veröffentlicht: (2023)
von: Ma, Chengcheng, et al.
Veröffentlicht: (2023)
E-SAM: Training-Free Segment Every Entity Model
von: Zhang, Weiming, et al.
Veröffentlicht: (2025)
von: Zhang, Weiming, et al.
Veröffentlicht: (2025)
ConStyle v2: A Strong Prompter for All-in-One Image Restoration
von: Fan, Dongqi, et al.
Veröffentlicht: (2024)
von: Fan, Dongqi, et al.
Veröffentlicht: (2024)
EchoPilot: Training-Free Ultrasound Video Segmentation via Scale-Space Semantic Prompting and Reliability-Gated Memory
von: Xiao, Ruiqiang, et al.
Veröffentlicht: (2026)
von: Xiao, Ruiqiang, et al.
Veröffentlicht: (2026)
A Comprehensive Review of Few-shot Action Recognition
von: Wanyan, Yuyang, et al.
Veröffentlicht: (2024)
von: Wanyan, Yuyang, et al.
Veröffentlicht: (2024)
Modality-Collaborative Low-Rank Decomposers for Few-Shot Video Domain Adaptation
von: Wanyan, Yuyang, et al.
Veröffentlicht: (2025)
von: Wanyan, Yuyang, et al.
Veröffentlicht: (2025)
Training-Free Layout-to-Image Generation with Marginal Attention Constraints
von: Chen, Huancheng, et al.
Veröffentlicht: (2024)
von: Chen, Huancheng, et al.
Veröffentlicht: (2024)
FlexIP: Dynamic Control of Preservation and Personality for Customized Image Generation
von: Huang, Linyan, et al.
Veröffentlicht: (2025)
von: Huang, Linyan, et al.
Veröffentlicht: (2025)
VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
TCP:Textual-based Class-aware Prompt tuning for Visual-Language Model
von: Yao, Hantao, et al.
Veröffentlicht: (2023)
von: Yao, Hantao, et al.
Veröffentlicht: (2023)
Optimizing 4D Wires for Sparse 3D Abstraction
von: Wu, Dong-Yi, et al.
Veröffentlicht: (2026)
von: Wu, Dong-Yi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Dance-to-Music Generation with Encoder-based Textual Inversion
von: Li, Sifei, et al.
Veröffentlicht: (2024) -
Self-Reasoning Agentic Framework for Narrative Product Grid-Collage Generation
von: Luo, Minyan, et al.
Veröffentlicht: (2026) -
CreativeSynth: Cross-Art-Attention for Artistic Image Synthesis with Multimodal Diffusion
von: Huang, Nisha, et al.
Veröffentlicht: (2024) -
MotionCrafter: One-Shot Motion Customization of Diffusion Models
von: Zhang, Yuxin, et al.
Veröffentlicht: (2023) -
Break-for-Make: Modular Low-Rank Adaptations for Composable Content-Style Customization
von: Xu, Yu, et al.
Veröffentlicht: (2024)