FreeInsert: Personalized Object Insertion with Geometric and Style Control
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yuhong, Wang, Han, Wang, Yiwen, Xie, Rong, Song, Li |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FreeInsert: Disentangled Text-Guided Object Insertion in 3D Gaussian Scene without Spatial Priors
by: Li, Chenxi, et al.
Published: (2025)
by: Li, Chenxi, et al.
Published: (2025)
SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion
by: Chen, Xinyu, et al.
Published: (2026)
by: Chen, Xinyu, et al.
Published: (2026)
Point2Insert: Video Object Insertion via Sparse Point Guidance
by: Zhou, Yu, et al.
Published: (2026)
by: Zhou, Yu, et al.
Published: (2026)
Multimodal Semantic-Aware Automatic Colorization with Diffusion Prior
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models
by: Chen, Jinshu, et al.
Published: (2025)
by: Chen, Jinshu, et al.
Published: (2025)
Insert Anything: Image Insertion via In-Context Editing in DiT
by: Song, Wensong, et al.
Published: (2025)
by: Song, Wensong, et al.
Published: (2025)
DreamInsert: Zero-Shot Image-to-Video Object Insertion from A Single Image
by: Zhao, Qi, et al.
Published: (2025)
by: Zhao, Qi, et al.
Published: (2025)
Insert In Style: A Zero-Shot Generative Framework for Harmonious Cross-Domain Object Composition
by: Chittersu, Raghu Vamsi, et al.
Published: (2025)
by: Chittersu, Raghu Vamsi, et al.
Published: (2025)
Style3D: Attention-guided Multi-view Style Transfer for 3D Object Generation
by: Song, Bingjie, et al.
Published: (2024)
by: Song, Bingjie, et al.
Published: (2024)
InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Video Object Insertion
by: Jin, Hoiyeong, et al.
Published: (2025)
by: Jin, Hoiyeong, et al.
Published: (2025)
Beyond Inserting: Learning Identity Embedding for Semantic-Fidelity Personalized Diffusion Generation
by: Li, Yang, et al.
Published: (2024)
by: Li, Yang, et al.
Published: (2024)
PoseTalk: Text-and-Audio-based Pose Control and Motion Refinement for One-Shot Talking Head Generation
by: Ling, Jun, et al.
Published: (2024)
by: Ling, Jun, et al.
Published: (2024)
SceneExpander: Expanding 3D Scenes with Free-Form Inserted Views
by: He, Zijian, et al.
Published: (2026)
by: He, Zijian, et al.
Published: (2026)
InsertDiffusion: Identity Preserving Visualization of Objects through a Training-Free Diffusion Architecture
by: Mueller, Phillip, et al.
Published: (2024)
by: Mueller, Phillip, et al.
Published: (2024)
Semantic and Temporal Integration in Latent Diffusion Space for High-Fidelity Video Super-Resolution
by: Wang, Yiwen, et al.
Published: (2025)
by: Wang, Yiwen, et al.
Published: (2025)
StyleStudio: Text-Driven Style Transfer with Selective Control of Style Elements
by: Lei, Mingkun, et al.
Published: (2024)
by: Lei, Mingkun, et al.
Published: (2024)
Controllable Video Object Insertion via Multiview Priors
by: Qi, Xia, et al.
Published: (2026)
by: Qi, Xia, et al.
Published: (2026)
Anything in Any Scene: Photorealistic Video Object Insertion
by: Bai, Chen, et al.
Published: (2024)
by: Bai, Chen, et al.
Published: (2024)
PIN: Positional Insert Unlocks Object Localisation Abilities in VLMs
by: Dorkenwald, Michael, et al.
Published: (2024)
by: Dorkenwald, Michael, et al.
Published: (2024)
SSP-IR: Semantic and Structure Priors for Diffusion-based Realistic Image Restoration
by: Zhang, Yuhong, et al.
Published: (2024)
by: Zhang, Yuhong, et al.
Published: (2024)
MIRRORTALK: Forging Personalized Avatars Via Disentangled Style and Hierarchical Motion Control
by: Lu, Renjie, et al.
Published: (2026)
by: Lu, Renjie, et al.
Published: (2026)
OmniScaleSR: Unleashing Scale-Controlled Diffusion Prior for Faithful and Realistic Arbitrary-Scale Image Super-Resolution
by: Chai, Xinning, et al.
Published: (2025)
by: Chai, Xinning, et al.
Published: (2025)
StableIdentity: Inserting Anybody into Anywhere at First Sight
by: Wang, Qinghe, et al.
Published: (2024)
by: Wang, Qinghe, et al.
Published: (2024)
Consistent Video Colorization via Palette Guidance
by: Wang, Han, et al.
Published: (2025)
by: Wang, Han, et al.
Published: (2025)
Object Style Diffusion for Generalized Object Detection in Urban Scene
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Style-Adaptive Detection Transformer for Single-Source Domain Generalized Object Detection
by: Han, Jianhong, et al.
Published: (2025)
by: Han, Jianhong, et al.
Published: (2025)
Magic Insert: Style-Aware Drag-and-Drop
by: Ruiz, Nataniel, et al.
Published: (2024)
by: Ruiz, Nataniel, et al.
Published: (2024)
CRAFT-LoRA: Content-Style Personalization via Rank-Constrained Adaptation and Training-Free Fusion
by: Li, Yu, et al.
Published: (2026)
by: Li, Yu, et al.
Published: (2026)
MFP-VTON: Enhancing Mask-Free Person-to-Person Virtual Try-On via Diffusion Transformer
by: Shen, Le, et al.
Published: (2025)
by: Shen, Le, et al.
Published: (2025)
Diff-Restorer: Unleashing Visual Prompts for Diffusion-based Universal Image Restoration
by: Zhang, Yuhong, et al.
Published: (2024)
by: Zhang, Yuhong, et al.
Published: (2024)
VideoAnydoor: High-fidelity Video Object Insertion with Precise Motion Control
by: Tu, Yuanpeng, et al.
Published: (2025)
by: Tu, Yuanpeng, et al.
Published: (2025)
FreeGaussian: Annotation-free Control of Articulated Objects via 3D Gaussian Splats with Flow Derivatives
by: Chen, Qizhi, et al.
Published: (2024)
by: Chen, Qizhi, et al.
Published: (2024)
ObjectDrop: Bootstrapping Counterfactuals for Photorealistic Object Removal and Insertion
by: Winter, Daniel, et al.
Published: (2024)
by: Winter, Daniel, et al.
Published: (2024)
Style Evolving along Chain-of-Thought for Unknown-Domain Object Detection
by: Zhang, Zihao, et al.
Published: (2025)
by: Zhang, Zihao, et al.
Published: (2025)
InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
by: Wang, Haofan, et al.
Published: (2024)
by: Wang, Haofan, et al.
Published: (2024)
TrajRAG: Retrieving Geometric-Semantic Experience for Zero-Shot Object Navigation
by: Wang, Yiyao, et al.
Published: (2026)
by: Wang, Yiyao, et al.
Published: (2026)
A Training-Free Style-Personalization via SVD-Based Feature Decomposition
by: Lee, Kyoungmin, et al.
Published: (2025)
by: Lee, Kyoungmin, et al.
Published: (2025)
Dynamic 2D Gaussians: Geometrically Accurate Radiance Fields for Dynamic Objects
by: Zhang, Shuai, et al.
Published: (2024)
by: Zhang, Shuai, et al.
Published: (2024)
GeoDiffusion: Text-Prompted Geometric Control for Object Detection Data Generation
by: Chen, Kai, et al.
Published: (2023)
by: Chen, Kai, et al.
Published: (2023)
Phrase Grounding-based Style Transfer for Single-Domain Generalized Object Detection
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Similar Items
-
FreeInsert: Disentangled Text-Guided Object Insertion in 3D Gaussian Scene without Spatial Priors
by: Li, Chenxi, et al.
Published: (2025) -
SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion
by: Chen, Xinyu, et al.
Published: (2026) -
Point2Insert: Video Object Insertion via Sparse Point Guidance
by: Zhou, Yu, et al.
Published: (2026) -
Multimodal Semantic-Aware Automatic Colorization with Diffusion Prior
by: Wang, Han, et al.
Published: (2024) -
OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models
by: Chen, Jinshu, et al.
Published: (2025)