ObjectAdd: Adding Objects into Image via a Training-Free Diffusion Modification Fashion
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Ziyue, Lin, Mingbao, Song, Quanjian, Zhang, Yuxin, Ji, Rongrong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
UniVST: A Unified Framework for Training-free Localized Video Style Transfer
di: Song, Quanjian, et al.
Pubblicazione: (2024)
di: Song, Quanjian, et al.
Pubblicazione: (2024)
DiffusionTrend: A Minimalist Approach to Virtual Fashion Try-On
di: Zhan, Wengyi, et al.
Pubblicazione: (2024)
di: Zhan, Wengyi, et al.
Pubblicazione: (2024)
EasyInv: Toward Fast and Better DDIM Inversion
di: Zhang, Ziyue, et al.
Pubblicazione: (2024)
di: Zhang, Ziyue, et al.
Pubblicazione: (2024)
AccDiffusion: An Accurate Method for Higher-Resolution Image Generation
di: Lin, Zhihang, et al.
Pubblicazione: (2024)
di: Lin, Zhihang, et al.
Pubblicazione: (2024)
LightMotion: A Light and Tuning-free Method for Simulating Camera Motion in Video Generation
di: Song, Quanjian, et al.
Pubblicazione: (2025)
di: Song, Quanjian, et al.
Pubblicazione: (2025)
Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models
di: Tewel, Yoad, et al.
Pubblicazione: (2024)
di: Tewel, Yoad, et al.
Pubblicazione: (2024)
Spatial Re-parameterization for N:M Sparsity
di: Zhang, Yuxin, et al.
Pubblicazione: (2023)
di: Zhang, Yuxin, et al.
Pubblicazione: (2023)
UniPTS: A Unified Framework for Proficient Post-Training Sparsity
di: Xie, Jingjing, et al.
Pubblicazione: (2024)
di: Xie, Jingjing, et al.
Pubblicazione: (2024)
AccDiffusion v2: Towards More Accurate Higher-Resolution Diffusion Extrapolation
di: Lin, Zhihang, et al.
Pubblicazione: (2024)
di: Lin, Zhihang, et al.
Pubblicazione: (2024)
Diffree: Text-Guided Shape Free Object Inpainting with Diffusion Model
di: Zhao, Lirui, et al.
Pubblicazione: (2024)
di: Zhao, Lirui, et al.
Pubblicazione: (2024)
Move and Act: Enhanced Object Manipulation and Background Integrity for Image Editing
di: Jiang, Pengfei, et al.
Pubblicazione: (2024)
di: Jiang, Pengfei, et al.
Pubblicazione: (2024)
CutDiffusion: A Simple, Fast, Cheap, and Strong Diffusion Extrapolation Method
di: Lin, Mingbao, et al.
Pubblicazione: (2024)
di: Lin, Mingbao, et al.
Pubblicazione: (2024)
Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation
di: Xie, Jingjing, et al.
Pubblicazione: (2024)
di: Xie, Jingjing, et al.
Pubblicazione: (2024)
AnySR: Realizing Image Super-Resolution as Any-Scale, Any-Resource
di: Zhan, Wengyi, et al.
Pubblicazione: (2024)
di: Zhan, Wengyi, et al.
Pubblicazione: (2024)
Parallel Vision Token Scheduling for Fast and Accurate Multimodal LMMs Inference
di: Zhan, Wengyi, et al.
Pubblicazione: (2025)
di: Zhan, Wengyi, et al.
Pubblicazione: (2025)
Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference
di: Lin, Zhihang, et al.
Pubblicazione: (2024)
di: Lin, Zhihang, et al.
Pubblicazione: (2024)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
di: Zhong, Yunshan, et al.
Pubblicazione: (2023)
di: Zhong, Yunshan, et al.
Pubblicazione: (2023)
TraDiffusion: Trajectory-Based Training-Free Image Generation
di: Wu, Mingrui, et al.
Pubblicazione: (2024)
di: Wu, Mingrui, et al.
Pubblicazione: (2024)
Towards Accurate Post-Training Quantization of Vision Transformers via Error Reduction
di: Zhong, Yunshan, et al.
Pubblicazione: (2024)
di: Zhong, Yunshan, et al.
Pubblicazione: (2024)
Search and Detect: Training-Free Long Tail Object Detection via Web-Image Retrieval
di: Sidhu, Mankeerat, et al.
Pubblicazione: (2024)
di: Sidhu, Mankeerat, et al.
Pubblicazione: (2024)
TFCounter:Polishing Gems for Training-Free Object Counting
di: Ting, Pan, et al.
Pubblicazione: (2024)
di: Ting, Pan, et al.
Pubblicazione: (2024)
FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization
di: Song, Quanjian, et al.
Pubblicazione: (2026)
di: Song, Quanjian, et al.
Pubblicazione: (2026)
Paint by Inpaint: Learning to Add Image Objects by Removing Them First
di: Wasserman, Navve, et al.
Pubblicazione: (2024)
di: Wasserman, Navve, et al.
Pubblicazione: (2024)
S$^2$Teacher: Step-by-step Teacher for Sparsely Annotated Oriented Object Detection
di: Lin, Yu, et al.
Pubblicazione: (2025)
di: Lin, Yu, et al.
Pubblicazione: (2025)
FocSAM: Delving Deeply into Focused Objects in Segmenting Anything
di: Huang, You, et al.
Pubblicazione: (2024)
di: Huang, You, et al.
Pubblicazione: (2024)
Adding New Categories in Object Detection Using Few-Shot Copy-Paste
di: Deng, Boyang, et al.
Pubblicazione: (2022)
di: Deng, Boyang, et al.
Pubblicazione: (2022)
Training-Free Open-Ended Object Detection and Segmentation via Attention as Prompts
di: Lin, Zhiwei, et al.
Pubblicazione: (2024)
di: Lin, Zhiwei, et al.
Pubblicazione: (2024)
Evaluating Attribute Confusion in Fashion Text-to-Image Generation
di: Liu, Ziyue, et al.
Pubblicazione: (2025)
di: Liu, Ziyue, et al.
Pubblicazione: (2025)
Training-Free Robust Interactive Video Object Segmentation
di: Wei, Xiaoli, et al.
Pubblicazione: (2024)
di: Wei, Xiaoli, et al.
Pubblicazione: (2024)
InsertDiffusion: Identity Preserving Visualization of Objects through a Training-Free Diffusion Architecture
di: Mueller, Phillip, et al.
Pubblicazione: (2024)
di: Mueller, Phillip, et al.
Pubblicazione: (2024)
FACTOR: Counterfactual Training-Free Test-Time Adaptation for Open-Vocabulary Object Detection
di: Zhao, Kaixiang, et al.
Pubblicazione: (2026)
di: Zhao, Kaixiang, et al.
Pubblicazione: (2026)
Object-WIPER : Training-Free Object and Associated Effect Removal in Videos
di: Kushwaha, Saksham Singh, et al.
Pubblicazione: (2026)
di: Kushwaha, Saksham Singh, et al.
Pubblicazione: (2026)
ACTrack: Adding Spatio-Temporal Condition for Visual Object Tracking
di: Han, Yushan, et al.
Pubblicazione: (2024)
di: Han, Yushan, et al.
Pubblicazione: (2024)
FashionComposer: Compositional Fashion Image Generation
di: Ji, Sihui, et al.
Pubblicazione: (2024)
di: Ji, Sihui, et al.
Pubblicazione: (2024)
HUWSOD: Holistic Self-training for Unified Weakly Supervised Object Detection
di: Cao, Liujuan, et al.
Pubblicazione: (2024)
di: Cao, Liujuan, et al.
Pubblicazione: (2024)
FashionMAC: Deformation-Free Fashion Image Generation with Fine-Grained Model Appearance Customization
di: Zhang, Rong, et al.
Pubblicazione: (2025)
di: Zhang, Rong, et al.
Pubblicazione: (2025)
From Objects to Events: Unlocking Complex Visual Understanding in Object Detectors via LLM-guided Symbolic Reasoning
di: Zeng, Yuhui, et al.
Pubblicazione: (2025)
di: Zeng, Yuhui, et al.
Pubblicazione: (2025)
FashionFail: Addressing Failure Cases in Fashion Object Detection and Segmentation
di: Velioglu, Riza, et al.
Pubblicazione: (2024)
di: Velioglu, Riza, et al.
Pubblicazione: (2024)
Purifying, Labeling, and Utilizing: A High-Quality Pipeline for Small Object Detection
di: Wang, Siwei, et al.
Pubblicazione: (2025)
di: Wang, Siwei, et al.
Pubblicazione: (2025)
RORem: Training a Robust Object Remover with Human-in-the-Loop
di: Li, Ruibin, et al.
Pubblicazione: (2025)
di: Li, Ruibin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
UniVST: A Unified Framework for Training-free Localized Video Style Transfer
di: Song, Quanjian, et al.
Pubblicazione: (2024) -
DiffusionTrend: A Minimalist Approach to Virtual Fashion Try-On
di: Zhan, Wengyi, et al.
Pubblicazione: (2024) -
EasyInv: Toward Fast and Better DDIM Inversion
di: Zhang, Ziyue, et al.
Pubblicazione: (2024) -
AccDiffusion: An Accurate Method for Higher-Resolution Image Generation
di: Lin, Zhihang, et al.
Pubblicazione: (2024) -
LightMotion: A Light and Tuning-free Method for Simulating Camera Motion in Video Generation
di: Song, Quanjian, et al.
Pubblicazione: (2025)