BOOTPLACE: Bootstrapped Object Placement with Detection Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Hang, Zuo, Xinxin, Ma, Rui, Cheng, Li |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BootPIG: Bootstrapping Zero-shot Personalized Image Generation Capabilities in Pretrained Diffusion Models
by: Purushwalkam, Senthil, et al.
Published: (2024)
by: Purushwalkam, Senthil, et al.
Published: (2024)
Zero-Shot Human-Object Interaction Synthesis with Multimodal Priors
by: Lou, Yuke, et al.
Published: (2025)
by: Lou, Yuke, et al.
Published: (2025)
ShadowDraw: From Any Object to Shadow-Drawing Compositional Art
by: Luo, Rundong, et al.
Published: (2025)
by: Luo, Rundong, et al.
Published: (2025)
ObjectMover: Generative Object Movement with Video Prior
by: Yu, Xin, et al.
Published: (2025)
by: Yu, Xin, et al.
Published: (2025)
Objectness Similarity: Capturing Object-Level Fidelity in 3D Scene Evaluation
by: Uchida, Yuiko, et al.
Published: (2025)
by: Uchida, Yuiko, et al.
Published: (2025)
Particulate: Feed-Forward 3D Object Articulation
by: Li, Ruining, et al.
Published: (2025)
by: Li, Ruining, et al.
Published: (2025)
Neural Gaffer: Relighting Any Object via Diffusion
by: Jin, Haian, et al.
Published: (2024)
by: Jin, Haian, et al.
Published: (2024)
GaussianAnything: Interactive Point Cloud Flow Matching For 3D Object Generation
by: Lan, Yushi, et al.
Published: (2024)
by: Lan, Yushi, et al.
Published: (2024)
UnCommon Objects in 3D
by: Liu, Xingchen, et al.
Published: (2025)
by: Liu, Xingchen, et al.
Published: (2025)
SpaRP: Fast 3D Object Reconstruction and Pose Estimation from Sparse Views
by: Xu, Chao, et al.
Published: (2024)
by: Xu, Chao, et al.
Published: (2024)
Object-level Visual Prompts for Compositional Image Generation
by: Parmar, Gaurav, et al.
Published: (2025)
by: Parmar, Gaurav, et al.
Published: (2025)
Photorealistic Object Insertion with Diffusion-Guided Inverse Rendering
by: Liang, Ruofan, et al.
Published: (2024)
by: Liang, Ruofan, et al.
Published: (2024)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
by: Zuo, Qi, et al.
Published: (2024)
by: Zuo, Qi, et al.
Published: (2024)
Make It Count: Text-to-Image Generation with an Accurate Number of Objects
by: Binyamin, Lital, et al.
Published: (2024)
by: Binyamin, Lital, et al.
Published: (2024)
Training-Free Text-Guided Color Editing with Multi-Modal Diffusion Transformer
by: Yin, Zixin, et al.
Published: (2025)
by: Yin, Zixin, et al.
Published: (2025)
Auto-Regressive Diffusion for Generating 3D Human-Object Interactions
by: Geng, Zichen, et al.
Published: (2025)
by: Geng, Zichen, et al.
Published: (2025)
Factored-NeuS: Reconstructing Surfaces, Illumination, and Materials of Possibly Glossy Objects
by: Fan, Yue, et al.
Published: (2023)
by: Fan, Yue, et al.
Published: (2023)
Bootstrap-GS: Self-Supervised Augmentation for High-Fidelity Gaussian Splatting
by: Gao, Yifei, et al.
Published: (2024)
by: Gao, Yifei, et al.
Published: (2024)
UniHM: Universal Human Motion Generation with Object Interactions in Indoor Scenes
by: Geng, Zichen, et al.
Published: (2025)
by: Geng, Zichen, et al.
Published: (2025)
Generative Object Insertion in Gaussian Splatting with a Multi-View Diffusion Model
by: Zhong, Hongliang, et al.
Published: (2024)
by: Zhong, Hongliang, et al.
Published: (2024)
PixelMan: Consistent Object Editing with Diffusion Models via Pixel Manipulation and Generation
by: Jiang, Liyao, et al.
Published: (2024)
by: Jiang, Liyao, et al.
Published: (2024)
EgoGrasp: World-Space Hand-Object Interaction Estimation from Egocentric Videos
by: Fu, Hongming, et al.
Published: (2026)
by: Fu, Hongming, et al.
Published: (2026)
Lazy Diffusion Transformer for Interactive Image Editing
by: Nitzan, Yotam, et al.
Published: (2024)
by: Nitzan, Yotam, et al.
Published: (2024)
OmniHands: Towards Robust 4D Hand Mesh Recovery via A Versatile Transformer
by: Lin, Dixuan, et al.
Published: (2024)
by: Lin, Dixuan, et al.
Published: (2024)
LuxDiT: Lighting Estimation with Video Diffusion Transformer
by: Liang, Ruofan, et al.
Published: (2025)
by: Liang, Ruofan, et al.
Published: (2025)
Boosting 3D Object Generation through PBR Materials
by: Wang, Yitong, et al.
Published: (2024)
by: Wang, Yitong, et al.
Published: (2024)
MVTN: Learning Multi-View Transformations for 3D Understanding
by: Hamdi, Abdullah, et al.
Published: (2022)
by: Hamdi, Abdullah, et al.
Published: (2022)
Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control
by: Gu, Zekai, et al.
Published: (2025)
by: Gu, Zekai, et al.
Published: (2025)
SVGBuilder: Component-Based Colored SVG Generation with Text-Guided Autoregressive Transformers
by: Chen, Zehao, et al.
Published: (2024)
by: Chen, Zehao, et al.
Published: (2024)
VividFace: A Diffusion-Based Hybrid Framework for High-Fidelity Video Face Swapping
by: Shao, Hao, et al.
Published: (2024)
by: Shao, Hao, et al.
Published: (2024)
StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal
by: Ye, Chongjie, et al.
Published: (2024)
by: Ye, Chongjie, et al.
Published: (2024)
LAYOUTDREAMER: Physics-guided Layout for Text-to-3D Compositional Scene Generation
by: Zhou, Yang, et al.
Published: (2025)
by: Zhou, Yang, et al.
Published: (2025)
Detection-Driven Object Count Optimization for Text-to-Image Diffusion Models
by: Zafar, Oz, et al.
Published: (2024)
by: Zafar, Oz, et al.
Published: (2024)
HodgeFormer: Transformers for Learnable Operators on Triangular Meshes through Data-Driven Hodge Matrices
by: Nousias, Akis, et al.
Published: (2025)
by: Nousias, Akis, et al.
Published: (2025)
Digital Twin Catalog: A Large-Scale Photorealistic 3D Object Digital Twin Dataset
by: Dong, Zhao, et al.
Published: (2025)
by: Dong, Zhao, et al.
Published: (2025)
Robust Symmetry Detection via Riemannian Langevin Dynamics
by: Je, Jihyeon, et al.
Published: (2024)
by: Je, Jihyeon, et al.
Published: (2024)
Human Action CLIPs: Detecting AI-generated Human Motion
by: Bohacek, Matyas, et al.
Published: (2024)
by: Bohacek, Matyas, et al.
Published: (2024)
Geometric Algebra Meets Large Language Models: Instruction-Based Transformations of Separate Meshes in 3D, Interactive and Controllable Scenes
by: Kolyvakis, Prodromos, et al.
Published: (2024)
by: Kolyvakis, Prodromos, et al.
Published: (2024)
Laplace-Beltrami Operator for Gaussian Splatting
by: Zhou, Hongyu, et al.
Published: (2025)
by: Zhou, Hongyu, et al.
Published: (2025)
Taming Diffusion Probabilistic Models for Character Control
by: Chen, Rui, et al.
Published: (2024)
by: Chen, Rui, et al.
Published: (2024)
Similar Items
-
BootPIG: Bootstrapping Zero-shot Personalized Image Generation Capabilities in Pretrained Diffusion Models
by: Purushwalkam, Senthil, et al.
Published: (2024) -
Zero-Shot Human-Object Interaction Synthesis with Multimodal Priors
by: Lou, Yuke, et al.
Published: (2025) -
ShadowDraw: From Any Object to Shadow-Drawing Compositional Art
by: Luo, Rundong, et al.
Published: (2025) -
ObjectMover: Generative Object Movement with Video Prior
by: Yu, Xin, et al.
Published: (2025) -
Objectness Similarity: Capturing Object-Level Fidelity in 3D Scene Evaluation
by: Uchida, Yuiko, et al.
Published: (2025)