HybridBooth: Hybrid Prompt Inversion for Efficient Subject-Driven Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Guan, Shanyan, Ge, Yanhao, Tai, Ying, Yang, Jian, Li, Wei, You, Mingyu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RaPD: Resolution-Agnostic Pixel Diffusion via Semantics-Enriched Implicit Representations
by: Ge, Yanhao, et al.
Published: (2026)
by: Ge, Yanhao, et al.
Published: (2026)
Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent
by: Ci, En, et al.
Published: (2025)
by: Ci, En, et al.
Published: (2025)
UltraHR-100K: Enhancing UHR Image Synthesis with A Large-Scale High-Quality Dataset
by: Zhao, Chen, et al.
Published: (2025)
by: Zhao, Chen, et al.
Published: (2025)
VINS-120K: Ultra High-Resolution Image Editing with A Large-Scale Dataset
by: Chen, Zhizhou, et al.
Published: (2026)
by: Chen, Zhizhou, et al.
Published: (2026)
NeoWorld: Neural Simulation of Explorable Virtual Worlds via Progressive 3D Unfolding
by: Zhao, Yanpeng, et al.
Published: (2025)
by: Zhao, Yanpeng, et al.
Published: (2025)
PostEdit: Posterior Sampling for Efficient Zero-Shot Image Editing
by: Tian, Feng, et al.
Published: (2024)
by: Tian, Feng, et al.
Published: (2024)
Neural Material Adaptor for Visual Grounding of Intrinsic Dynamics
by: Cao, Junyi, et al.
Published: (2024)
by: Cao, Junyi, et al.
Published: (2024)
Guiding a Diffusion Model by Swapping Its Tokens
by: Zhang, Weijia, et al.
Published: (2026)
by: Zhang, Weijia, et al.
Published: (2026)
DisenBooth: Identity-Preserving Disentangled Tuning for Subject-Driven Text-to-Image Generation
by: Chen, Hong, et al.
Published: (2023)
by: Chen, Hong, et al.
Published: (2023)
Octopus: History-Free Gradient Orthogonalization for Continual Learning in Multimodal Large Language Models
by: Liu, Yuehao, et al.
Published: (2026)
by: Liu, Yuehao, et al.
Published: (2026)
ACE-LoRA: Adaptive Orthogonal Decoupling for Continual Image Editing
by: Liu, Yuehao, et al.
Published: (2026)
by: Liu, Yuehao, et al.
Published: (2026)
SceneBooth: Diffusion-based Framework for Subject-preserved Text-to-Image Generation
by: Chai, Shang, et al.
Published: (2025)
by: Chai, Shang, et al.
Published: (2025)
CoDi: Subject-Consistent and Pose-Diverse Text-to-Image Generation
by: Gao, Zhanxin, et al.
Published: (2025)
by: Gao, Zhanxin, et al.
Published: (2025)
AttnDreamBooth: Towards Text-Aligned Personalized Text-to-Image Generation
by: Pang, Lianyu, et al.
Published: (2024)
by: Pang, Lianyu, et al.
Published: (2024)
Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models
by: Ji, Yicheng, et al.
Published: (2026)
by: Ji, Yicheng, et al.
Published: (2026)
PersonaBooth: Personalized Text-to-Motion Generation
by: Kim, Boeun, et al.
Published: (2025)
by: Kim, Boeun, et al.
Published: (2025)
MotionBooth: Motion-Aware Customized Text-to-Video Generation
by: Wu, Jianzong, et al.
Published: (2024)
by: Wu, Jianzong, et al.
Published: (2024)
GroundingBooth: Grounding Text-to-Image Customization
by: Xiong, Zhexiao, et al.
Published: (2024)
by: Xiong, Zhexiao, et al.
Published: (2024)
MultiBooth: Towards Generating All Your Concepts in an Image from Text
by: Zhu, Chenyang, et al.
Published: (2024)
by: Zhu, Chenyang, et al.
Published: (2024)
Hybrid Fourier Score Distillation for Efficient One Image to 3D Object Generation
by: Yang, Shuzhou, et al.
Published: (2024)
by: Yang, Shuzhou, et al.
Published: (2024)
ID-Booth: Identity-consistent Face Generation with Diffusion Models
by: Tomašević, Darian, et al.
Published: (2025)
by: Tomašević, Darian, et al.
Published: (2025)
SEED-Data-Edit Technical Report: A Hybrid Dataset for Instructional Image Editing
by: Ge, Yuying, et al.
Published: (2024)
by: Ge, Yuying, et al.
Published: (2024)
ProEdit: Inversion-based Editing From Prompts Done Right
by: Ouyang, Zhi, et al.
Published: (2025)
by: Ouyang, Zhi, et al.
Published: (2025)
RealTalk: Real-time and Realistic Audio-driven Face Generation with 3D Facial Prior-guided Identity Alignment Network
by: Ji, Xiaozhong, et al.
Published: (2024)
by: Ji, Xiaozhong, et al.
Published: (2024)
OmniBooth: Learning Latent Control for Image Synthesis with Multi-modal Instruction
by: Li, Leheng, et al.
Published: (2024)
by: Li, Leheng, et al.
Published: (2024)
A Generalist FaceX via Learning Unified Facial Representation
by: Han, Yue, et al.
Published: (2023)
by: Han, Yue, et al.
Published: (2023)
StyleDiffusion: Prompt-Embedding Inversion for Text-Based Editing
by: Li, Senmao, et al.
Published: (2023)
by: Li, Senmao, et al.
Published: (2023)
Efficient Image Super-Resolution with Feature Interaction Weighted Hybrid Network
by: Li, Wenjie, et al.
Published: (2022)
by: Li, Wenjie, et al.
Published: (2022)
DiffProxy: Multi-View Human Mesh Recovery via Diffusion-Generated Dense Proxies
by: Wang, Renke, et al.
Published: (2026)
by: Wang, Renke, et al.
Published: (2026)
DreamBoothDPO: Improving Personalized Generation using Direct Preference Optimization
by: Ayupov, Shamil, et al.
Published: (2025)
by: Ayupov, Shamil, et al.
Published: (2025)
AnomalyHybrid: A Domain-agnostic Generative Framework for General Anomaly Detection
by: Zhao, Ying
Published: (2025)
by: Zhao, Ying
Published: (2025)
InvSeg: Test-Time Prompt Inversion for Semantic Segmentation
by: Lin, Jiayi, et al.
Published: (2024)
by: Lin, Jiayi, et al.
Published: (2024)
Dual-Prompt CLIP with Hybrid Visual Encoders for Occluded Person Re-Identification
by: Ji, Zhangjian, et al.
Published: (2026)
by: Ji, Zhangjian, et al.
Published: (2026)
SSR-Encoder: Encoding Selective Subject Representation for Subject-Driven Generation
by: Zhang, Yuxuan, et al.
Published: (2023)
by: Zhang, Yuxuan, et al.
Published: (2023)
StyleBooth: Image Style Editing with Multimodal Instruction
by: Han, Zhen, et al.
Published: (2024)
by: Han, Zhen, et al.
Published: (2024)
AgeBooth: Controllable Facial Aging and Rejuvenation via Diffusion Models
by: Zhu, Shihao, et al.
Published: (2025)
by: Zhu, Shihao, et al.
Published: (2025)
Vamba: Understanding Hour-Long Videos with Hybrid Mamba-Transformers
by: Ren, Weiming, et al.
Published: (2025)
by: Ren, Weiming, et al.
Published: (2025)
InstructBooth: Instruction-following Personalized Text-to-Image Generation
by: Chae, Daewon, et al.
Published: (2023)
by: Chae, Daewon, et al.
Published: (2023)
Improving Subject-Driven Image Synthesis with Subject-Agnostic Guidance
by: Chan, Kelvin C. K., et al.
Published: (2024)
by: Chan, Kelvin C. K., et al.
Published: (2024)
Diff-Prompt: Diffusion-Driven Prompt Generator with Mask Supervision
by: Yan, Weicai, et al.
Published: (2025)
by: Yan, Weicai, et al.
Published: (2025)
Similar Items
-
RaPD: Resolution-Agnostic Pixel Diffusion via Semantics-Enriched Implicit Representations
by: Ge, Yanhao, et al.
Published: (2026) -
Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent
by: Ci, En, et al.
Published: (2025) -
UltraHR-100K: Enhancing UHR Image Synthesis with A Large-Scale High-Quality Dataset
by: Zhao, Chen, et al.
Published: (2025) -
VINS-120K: Ultra High-Resolution Image Editing with A Large-Scale Dataset
by: Chen, Zhizhou, et al.
Published: (2026) -
NeoWorld: Neural Simulation of Explorable Virtual Worlds via Progressive 3D Unfolding
by: Zhao, Yanpeng, et al.
Published: (2025)