IPAdapter-Instruct: Resolving Ambiguity in Image-based Conditioning using Instruct Prompts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rowles, Ciara, Vainer, Shimon, De Nigris, Dante, Elizarov, Slava, Kutsy, Konstantin, Donné, Simon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Jointly Generating Multi-view Consistent PBR Textures using Collaborative Control
von: Vainer, Shimon, et al.
Veröffentlicht: (2024)
von: Vainer, Shimon, et al.
Veröffentlicht: (2024)
Collaborative Control for Geometry-Conditioned PBR Image Generation
von: Vainer, Shimon, et al.
Veröffentlicht: (2024)
von: Vainer, Shimon, et al.
Veröffentlicht: (2024)
Geometry Image Diffusion: Fast and Data-Efficient Text-to-3D with Image-Based Surface Representation
von: Elizarov, Slava, et al.
Veröffentlicht: (2024)
von: Elizarov, Slava, et al.
Veröffentlicht: (2024)
Foley Control: Aligning a Frozen Latent Text-to-Audio Model to Video
von: Rowles, Ciara, et al.
Veröffentlicht: (2025)
von: Rowles, Ciara, et al.
Veröffentlicht: (2025)
Stylecodes: Encoding Stylistic Information For Image Generation
von: Rowles, Ciara
Veröffentlicht: (2024)
von: Rowles, Ciara
Veröffentlicht: (2024)
VICI: VLM-Instructed Cross-view Image-localisation
von: Zhang, Xiaohan, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaohan, et al.
Veröffentlicht: (2025)
InstructPix2NeRF: Instructed 3D Portrait Editing from a Single Image
von: Li, Jianhui, et al.
Veröffentlicht: (2023)
von: Li, Jianhui, et al.
Veröffentlicht: (2023)
InstructGIE: Towards Generalizable Image Editing
von: Meng, Zichong, et al.
Veröffentlicht: (2024)
von: Meng, Zichong, et al.
Veröffentlicht: (2024)
Scaling Prompt Instructed Zero Shot Composed Image Retrieval with Image-Only Data
von: Duan, Yiqun, et al.
Veröffentlicht: (2025)
von: Duan, Yiqun, et al.
Veröffentlicht: (2025)
InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models
von: Wei, Cong, et al.
Veröffentlicht: (2024)
von: Wei, Cong, et al.
Veröffentlicht: (2024)
InstructEngine: Instruction-driven Text-to-Image Alignment
von: Lu, Xingyu, et al.
Veröffentlicht: (2025)
von: Lu, Xingyu, et al.
Veröffentlicht: (2025)
Batch-Instructed Gradient for Prompt Evolution:Systematic Prompt Optimization for Enhanced Text-to-Image Synthesis
von: Yang, Xinrui, et al.
Veröffentlicht: (2024)
von: Yang, Xinrui, et al.
Veröffentlicht: (2024)
InstructBrush: Learning Attention-based Instruction Optimization for Image Editing
von: Zhao, Ruoyu, et al.
Veröffentlicht: (2024)
von: Zhao, Ruoyu, et al.
Veröffentlicht: (2024)
InstructRestore: Region-Customized Image Restoration with Human Instructions
von: Liu, Shuaizheng, et al.
Veröffentlicht: (2025)
von: Liu, Shuaizheng, et al.
Veröffentlicht: (2025)
LITA: Language Instructed Temporal-Localization Assistant
von: Huang, De-An, et al.
Veröffentlicht: (2024)
von: Huang, De-An, et al.
Veröffentlicht: (2024)
InstructCV: Instruction-Tuned Text-to-Image Diffusion Models as Vision Generalists
von: Gan, Yulu, et al.
Veröffentlicht: (2023)
von: Gan, Yulu, et al.
Veröffentlicht: (2023)
Instruct-IPT: All-in-One Image Processing Transformer via Weight Modulation
von: Tian, Yuchuan, et al.
Veröffentlicht: (2024)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2024)
InstructRL4Pix: Training Diffusion for Image Editing by Reinforcement Learning
von: Li, Tiancheng, et al.
Veröffentlicht: (2024)
von: Li, Tiancheng, et al.
Veröffentlicht: (2024)
InstructUDrag: Joint Text Instructions and Object Dragging for Interactive Image Editing
von: Yu, Haoran, et al.
Veröffentlicht: (2025)
von: Yu, Haoran, et al.
Veröffentlicht: (2025)
Instructing Text-to-Image Diffusion Models via Classifier-Guided Semantic Optimization
von: Chang, Yuanyuan, et al.
Veröffentlicht: (2025)
von: Chang, Yuanyuan, et al.
Veröffentlicht: (2025)
InstructBooth: Instruction-following Personalized Text-to-Image Generation
von: Chae, Daewon, et al.
Veröffentlicht: (2023)
von: Chae, Daewon, et al.
Veröffentlicht: (2023)
InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation
von: Xiao, Jinqi, et al.
Veröffentlicht: (2025)
von: Xiao, Jinqi, et al.
Veröffentlicht: (2025)
InstructSAM: Segment Any Instance with Any Instructions
von: Yuan, Yuqian, et al.
Veröffentlicht: (2026)
von: Yuan, Yuqian, et al.
Veröffentlicht: (2026)
Instruct-Imagen: Image Generation with Multi-modal Instruction
von: Hu, Hexiang, et al.
Veröffentlicht: (2024)
von: Hu, Hexiang, et al.
Veröffentlicht: (2024)
Reward-Instruct: A Reward-Centric Approach to Fast Photo-Realistic Image Generation
von: Luo, Yihong, et al.
Veröffentlicht: (2025)
von: Luo, Yihong, et al.
Veröffentlicht: (2025)
Uncertainty-Instructed Structure Injection for Generalizable HD Map Construction
von: Liu, Xiaolu, et al.
Veröffentlicht: (2025)
von: Liu, Xiaolu, et al.
Veröffentlicht: (2025)
InstructX: Towards Unified Visual Editing with MLLM Guidance
von: Mou, Chong, et al.
Veröffentlicht: (2025)
von: Mou, Chong, et al.
Veröffentlicht: (2025)
InstructVEdit: A Holistic Approach for Instructional Video Editing
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
Seer: Language Instructed Video Prediction with Latent Diffusion Models
von: Gu, Xianfan, et al.
Veröffentlicht: (2023)
von: Gu, Xianfan, et al.
Veröffentlicht: (2023)
VIRST: Video-Instructed Reasoning Assistant for SpatioTemporal Segmentation
von: Hong, Jihwan, et al.
Veröffentlicht: (2026)
von: Hong, Jihwan, et al.
Veröffentlicht: (2026)
InstructAttribute: Fine-grained Object Attributes editing with Instruction
von: Yin, Xingxi, et al.
Veröffentlicht: (2025)
von: Yin, Xingxi, et al.
Veröffentlicht: (2025)
InstructTable: Improving Table Structure Recognition Through Instructions
von: Chen, Boming, et al.
Veröffentlicht: (2026)
von: Chen, Boming, et al.
Veröffentlicht: (2026)
VideoITG: Multimodal Video Understanding with Instructed Temporal Grounding
von: Wang, Shihao, et al.
Veröffentlicht: (2025)
von: Wang, Shihao, et al.
Veröffentlicht: (2025)
Multimodal Self-Instruct: Synthetic Abstract Image and Visual Reasoning Instruction Using Language Model
von: Zhang, Wenqi, et al.
Veröffentlicht: (2024)
von: Zhang, Wenqi, et al.
Veröffentlicht: (2024)
TIDE: Achieving Balanced Subject-Driven Image Generation via Target-Instructed Diffusion Enhancement
von: Lin, Jibai, et al.
Veröffentlicht: (2025)
von: Lin, Jibai, et al.
Veröffentlicht: (2025)
InstructPart: Task-Oriented Part Segmentation with Instruction Reasoning
von: Wan, Zifu, et al.
Veröffentlicht: (2025)
von: Wan, Zifu, et al.
Veröffentlicht: (2025)
Generative Timelines for Instructed Visual Assembly
von: Pardo, Alejandro, et al.
Veröffentlicht: (2024)
von: Pardo, Alejandro, et al.
Veröffentlicht: (2024)
Instruct2See: Learning to Remove Any Obstructions Across Distributions
von: Li, Junhang, et al.
Veröffentlicht: (2025)
von: Li, Junhang, et al.
Veröffentlicht: (2025)
TIGER: Text-Instructed 3D Gaussian Retrieval and Coherent Editing
von: Xu, Teng, et al.
Veröffentlicht: (2024)
von: Xu, Teng, et al.
Veröffentlicht: (2024)
Question-Instructed Visual Descriptions for Zero-Shot Video Question Answering
von: Romero, David, et al.
Veröffentlicht: (2024)
von: Romero, David, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Jointly Generating Multi-view Consistent PBR Textures using Collaborative Control
von: Vainer, Shimon, et al.
Veröffentlicht: (2024) -
Collaborative Control for Geometry-Conditioned PBR Image Generation
von: Vainer, Shimon, et al.
Veröffentlicht: (2024) -
Geometry Image Diffusion: Fast and Data-Efficient Text-to-3D with Image-Based Surface Representation
von: Elizarov, Slava, et al.
Veröffentlicht: (2024) -
Foley Control: Aligning a Frozen Latent Text-to-Audio Model to Video
von: Rowles, Ciara, et al.
Veröffentlicht: (2025) -
Stylecodes: Encoding Stylistic Information For Image Generation
von: Rowles, Ciara
Veröffentlicht: (2024)