Customizing Text-to-Image Diffusion with Object Viewpoint Control
Fuente:
arXiv
Guardado en:
| Autores principales: | Kumari, Nupur, Su, Grace, Zhang, Richard, Park, Taesung, Shechtman, Eli, Zhu, Jun-Yan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Editable Image Elements for Controllable Synthesis
por: Mu, Jiteng, et al.
Publicado: (2024)
por: Mu, Jiteng, et al.
Publicado: (2024)
Generating Multi-Image Synthetic Data for Text-to-Image Customization
por: Kumari, Nupur, et al.
Publicado: (2025)
por: Kumari, Nupur, et al.
Publicado: (2025)
Jump Cut Smoothing for Talking Heads
por: Wang, Xiaojuan, et al.
Publicado: (2024)
por: Wang, Xiaojuan, et al.
Publicado: (2024)
Customizing Text-to-Image Models with a Single Image Pair
por: Jones, Maxwell, et al.
Publicado: (2024)
por: Jones, Maxwell, et al.
Publicado: (2024)
Lazy Diffusion Transformer for Interactive Image Editing
por: Nitzan, Yotam, et al.
Publicado: (2024)
por: Nitzan, Yotam, et al.
Publicado: (2024)
One-step Diffusion with Distribution Matching Distillation
por: Yin, Tianwei, et al.
Publicado: (2023)
por: Yin, Tianwei, et al.
Publicado: (2023)
Identifying Prompted Artist Names from Generated Images
por: Su, Grace, et al.
Publicado: (2025)
por: Su, Grace, et al.
Publicado: (2025)
Improved Distribution Matching Distillation for Fast Image Synthesis
por: Yin, Tianwei, et al.
Publicado: (2024)
por: Yin, Tianwei, et al.
Publicado: (2024)
Distilling Diffusion Models into Conditional GANs
por: Kang, Minguk, et al.
Publicado: (2024)
por: Kang, Minguk, et al.
Publicado: (2024)
Learning an Image Editing Model without Image Editing Pairs
por: Kumari, Nupur, et al.
Publicado: (2025)
por: Kumari, Nupur, et al.
Publicado: (2025)
One-Step Image Translation with Text-to-Image Models
por: Parmar, Gaurav, et al.
Publicado: (2024)
por: Parmar, Gaurav, et al.
Publicado: (2024)
Expressive Text-to-Image Generation with Rich Text
por: Ge, Songwei, et al.
Publicado: (2023)
por: Ge, Songwei, et al.
Publicado: (2023)
Image Neural Field Diffusion Models
por: Chen, Yinbo, et al.
Publicado: (2024)
por: Chen, Yinbo, et al.
Publicado: (2024)
NewMove: Customizing text-to-video models with novel motions
por: Materzynska, Joanna, et al.
Publicado: (2023)
por: Materzynska, Joanna, et al.
Publicado: (2023)
VideoGigaGAN: Towards Detail-rich Video Super-Resolution
por: Xu, Yiran, et al.
Publicado: (2024)
por: Xu, Yiran, et al.
Publicado: (2024)
MotionStream: Real-Time Video Generation with Interactive Motion Controls
por: Shin, Joonghyuk, et al.
Publicado: (2025)
por: Shin, Joonghyuk, et al.
Publicado: (2025)
Generative Photomontage
por: Liu, Sean J., et al.
Publicado: (2024)
por: Liu, Sean J., et al.
Publicado: (2024)
Fine-grained Defocus Blur Control for Generative Image Models
por: Shrivastava, Ayush, et al.
Publicado: (2025)
por: Shrivastava, Ayush, et al.
Publicado: (2025)
From Slow Bidirectional to Fast Autoregressive Video Diffusion Models
por: Yin, Tianwei, et al.
Publicado: (2024)
por: Yin, Tianwei, et al.
Publicado: (2024)
Camera Control for Text-to-Image Generation via Learning Viewpoint Tokens
por: Lu, Xinxuan, et al.
Publicado: (2026)
por: Lu, Xinxuan, et al.
Publicado: (2026)
Infusion: Preventing Customized Text-to-Image Diffusion from Overfitting
por: Zeng, Weili, et al.
Publicado: (2024)
por: Zeng, Weili, et al.
Publicado: (2024)
Removing Distributional Discrepancies in Captions Improves Image-Text Alignment
por: Li, Yuheng, et al.
Publicado: (2024)
por: Li, Yuheng, et al.
Publicado: (2024)
SliderSpace: Decomposing the Visual Capabilities of Diffusion Models
por: Gandikota, Rohit, et al.
Publicado: (2025)
por: Gandikota, Rohit, et al.
Publicado: (2025)
Learning to Customize Text-to-Image Diffusion In Diverse Context
por: Kim, Taewook, et al.
Publicado: (2024)
por: Kim, Taewook, et al.
Publicado: (2024)
Group Diffusion: Enhancing Image Generation by Unlocking Cross-Sample Collaboration
por: Mo, Sicheng, et al.
Publicado: (2025)
por: Mo, Sicheng, et al.
Publicado: (2025)
Zero4D: Training-Free 4D Video Generation From Single Video Using Off-the-Shelf Video Diffusion
por: Park, Jangho, et al.
Publicado: (2025)
por: Park, Jangho, et al.
Publicado: (2025)
Style Customization of Text-to-Vector Generation with Image Diffusion Priors
por: Zhang, Peiying, et al.
Publicado: (2025)
por: Zhang, Peiying, et al.
Publicado: (2025)
DECOR:Decomposition and Projection of Text Embeddings for Text-to-Image Customization
por: Jang, Geonhui, et al.
Publicado: (2024)
por: Jang, Geonhui, et al.
Publicado: (2024)
How to Continually Adapt Text-to-Image Diffusion Models for Flexible Customization?
por: Dong, Jiahua, et al.
Publicado: (2024)
por: Dong, Jiahua, et al.
Publicado: (2024)
Self-Evaluation Unlocks Any-Step Text-to-Image Generation
por: Yu, Xin, et al.
Publicado: (2025)
por: Yu, Xin, et al.
Publicado: (2025)
Structure-Guided Image Completion with Image-level and Object-level Semantic Discriminators
por: Zheng, Haitian, et al.
Publicado: (2022)
por: Zheng, Haitian, et al.
Publicado: (2022)
TurboEdit: Instant text-based image editing
por: Wu, Zongze, et al.
Publicado: (2024)
por: Wu, Zongze, et al.
Publicado: (2024)
TSCnet: A Text-driven Semantic-level Controllable Framework for Customized Low-Light Image Enhancement
por: Zhang, Miao, et al.
Publicado: (2025)
por: Zhang, Miao, et al.
Publicado: (2025)
CustomText: Customized Textual Image Generation using Diffusion Models
por: Paliwal, Shubham, et al.
Publicado: (2024)
por: Paliwal, Shubham, et al.
Publicado: (2024)
Direct Consistency Optimization for Robust Customization of Text-to-Image Diffusion Models
por: Lee, Kyungmin, et al.
Publicado: (2024)
por: Lee, Kyungmin, et al.
Publicado: (2024)
Interact-Custom: Customized Human Object Interaction Image Generation
por: Xu, Zhu, et al.
Publicado: (2025)
por: Xu, Zhu, et al.
Publicado: (2025)
Alignment-Guided Score Matching for Text-to-Image Alignment in Diffusion Models
por: Lee, Jaa-Yeon, et al.
Publicado: (2026)
por: Lee, Jaa-Yeon, et al.
Publicado: (2026)
Improving Viewpoint-Independent Object-Centric Representations through Active Viewpoint Selection
por: Huang, Yinxuan, et al.
Publicado: (2024)
por: Huang, Yinxuan, et al.
Publicado: (2024)
RealCustom: Narrowing Real Text Word for Real-Time Open-Domain Text-to-Image Customization
por: Huang, Mengqi, et al.
Publicado: (2024)
por: Huang, Mengqi, et al.
Publicado: (2024)
GroundingBooth: Grounding Text-to-Image Customization
por: Xiong, Zhexiao, et al.
Publicado: (2024)
por: Xiong, Zhexiao, et al.
Publicado: (2024)
Ejemplares similares
-
Editable Image Elements for Controllable Synthesis
por: Mu, Jiteng, et al.
Publicado: (2024) -
Generating Multi-Image Synthetic Data for Text-to-Image Customization
por: Kumari, Nupur, et al.
Publicado: (2025) -
Jump Cut Smoothing for Talking Heads
por: Wang, Xiaojuan, et al.
Publicado: (2024) -
Customizing Text-to-Image Models with a Single Image Pair
por: Jones, Maxwell, et al.
Publicado: (2024) -
Lazy Diffusion Transformer for Interactive Image Editing
por: Nitzan, Yotam, et al.
Publicado: (2024)