PlaceIt3D: Language-Guided Object Placement in Real 3D Scenes
Fuente:
arXiv
Saved in:
| Main Authors: | Abdelreheem, Ahmed, Aleotti, Filippo, Watson, Jamie, Qureshi, Zawar, Eldesokey, Abdelrahman, Wonka, Peter, Brostow, Gabriel, Vicente, Sara, Garcia-Hernando, Guillermo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DoubleTake: Geometry Guided Depth Estimation
by: Sayed, Mohamed, et al.
Published: (2024)
by: Sayed, Mohamed, et al.
Published: (2024)
AirPlanes: Accurate Plane Estimation via 3D-Consistent Embeddings
by: Watson, Jamie, et al.
Published: (2024)
by: Watson, Jamie, et al.
Published: (2024)
Build-A-Scene: Interactive 3D Layout Control for Diffusion-Based Image Generation
by: Eldesokey, Abdelrahman, et al.
Published: (2024)
by: Eldesokey, Abdelrahman, et al.
Published: (2024)
ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models
by: Gong, Bingchen, et al.
Published: (2024)
by: Gong, Bingchen, et al.
Published: (2024)
LatentMan: Generating Consistent Animated Characters using Image Diffusion Models
by: Eldesokey, Abdelrahman, et al.
Published: (2023)
by: Eldesokey, Abdelrahman, et al.
Published: (2023)
Morpheus: Text-Driven 3D Gaussian Splat Shape and Color Stylization
by: Wynn, Jamie, et al.
Published: (2025)
by: Wynn, Jamie, et al.
Published: (2025)
PartEdit: Fine-Grained Image Editing using Pre-Trained Diffusion Models
by: Cvejic, Aleksandar, et al.
Published: (2025)
by: Cvejic, Aleksandar, et al.
Published: (2025)
AvatarMMC: 3D Head Avatar Generation and Editing with Multi-Modal Conditioning
by: Para, Wamiq Reyaz, et al.
Published: (2024)
by: Para, Wamiq Reyaz, et al.
Published: (2024)
Mind-the-Glitch: Visual Correspondence for Detecting Inconsistencies in Subject-Driven Generation
by: Eldesokey, Abdelrahman, et al.
Published: (2025)
by: Eldesokey, Abdelrahman, et al.
Published: (2025)
EditCLIP: Representation Learning for Image Editing
by: Wang, Qian, et al.
Published: (2025)
by: Wang, Qian, et al.
Published: (2025)
NearID: Identity Representation Learning via Near-identity Distractors
by: Cvejic, Aleksandar, et al.
Published: (2026)
by: Cvejic, Aleksandar, et al.
Published: (2026)
T-3DGS: Removing Transient Objects for 3D Scene Reconstruction
by: Markin, Alexander, et al.
Published: (2024)
by: Markin, Alexander, et al.
Published: (2024)
MonoPlace3D: Learning 3D-Aware Object Placement for 3D Monocular Detection
by: Parihar, Rishubh, et al.
Published: (2025)
by: Parihar, Rishubh, et al.
Published: (2025)
3DCoMPaT$^{++}$: An improved Large-scale 3D Vision Dataset for Compositional Recognition
by: Slim, Habib, et al.
Published: (2023)
by: Slim, Habib, et al.
Published: (2023)
FloorplanQA: A Benchmark for Spatial Reasoning in LLMs using Structured Representations
by: Rodionov, Fedor, et al.
Published: (2025)
by: Rodionov, Fedor, et al.
Published: (2025)
Controllable 3D Placement of Objects with Scene-Aware Diffusion Models
by: Omran, Mohamed, et al.
Published: (2025)
by: Omran, Mohamed, et al.
Published: (2025)
ImaginateAR: AI-Assisted In-Situ Authoring in Augmented Reality
by: Lee, Jaewook, et al.
Published: (2025)
by: Lee, Jaewook, et al.
Published: (2025)
FirePlace: Geometric Refinements of LLM Common Sense Reasoning for 3D Object Placement
by: Huang, Ian, et al.
Published: (2025)
by: Huang, Ian, et al.
Published: (2025)
Zero-Shot Video Semantic Segmentation based on Pre-Trained Diffusion Models
by: Wang, Qian, et al.
Published: (2024)
by: Wang, Qian, et al.
Published: (2024)
Back to 3D: Few-Shot 3D Keypoint Detection with Back-Projected 2D Features
by: Wimmer, Thomas, et al.
Published: (2023)
by: Wimmer, Thomas, et al.
Published: (2023)
Transfer Learning from Simulated to Real Scenes for Monocular 3D Object Detection
by: Mohamed, Sondos, et al.
Published: (2024)
by: Mohamed, Sondos, et al.
Published: (2024)
GroundUp: Rapid Sketch-Based 3D City Massing
by: Unlu, Gizem Esra, et al.
Published: (2024)
by: Unlu, Gizem Esra, et al.
Published: (2024)
NaviNote: Enabling In-situ Spatial Annotation Authoring to Support Exploration and Navigation for Blind and Low Vision People
by: Chen, Ruijia, et al.
Published: (2026)
by: Chen, Ruijia, et al.
Published: (2026)
G-CUT3R: Guided 3D Reconstruction with Camera and Depth Prior Integration
by: Khafizov, Ramil, et al.
Published: (2025)
by: Khafizov, Ramil, et al.
Published: (2025)
MVSAnywhere: Zero-Shot Multi-View Stereo
by: Izquierdo, Sergio, et al.
Published: (2025)
by: Izquierdo, Sergio, et al.
Published: (2025)
Out-of-Distribution Segmentation via Wasserstein-Based Evidential Uncertainty
by: Brosch, Arnold, et al.
Published: (2025)
by: Brosch, Arnold, et al.
Published: (2025)
Is Your LiDAR Placement Optimized for 3D Scene Understanding?
by: Li, Ye, et al.
Published: (2024)
by: Li, Ye, et al.
Published: (2024)
LaGeM: A Large Geometry Model for 3D Representation Learning and Diffusion
by: Zhang, Biao, et al.
Published: (2024)
by: Zhang, Biao, et al.
Published: (2024)
Human Geometry Distribution for 3D Animation Generation
by: Tang, Xiangjun, et al.
Published: (2025)
by: Tang, Xiangjun, et al.
Published: (2025)
From Geometry to Culture: An Iterative VLM Layout Framework for Placing Objects in Complex 3D Scene Contexts
by: Asano, Yuto, et al.
Published: (2025)
by: Asano, Yuto, et al.
Published: (2025)
KITchen: A Real-World Benchmark and Dataset for 6D Object Pose Estimation in Kitchen Environments
by: Younes, Abdelrahman, et al.
Published: (2024)
by: Younes, Abdelrahman, et al.
Published: (2024)
TAPVid-3D: A Benchmark for Tracking Any Point in 3D
by: Koppula, Skanda, et al.
Published: (2024)
by: Koppula, Skanda, et al.
Published: (2024)
Neural Rearrangement Planning for Object Retrieval from Confined Spaces Perceivable by Robot's In-hand RGB-D Sensor
by: Ren, Hanwen, et al.
Published: (2024)
by: Ren, Hanwen, et al.
Published: (2024)
SoccerLens: Grounded Soccer Video Understanding Beyond Accuracy
by: Elsharkawi, Ismael, et al.
Published: (2026)
by: Elsharkawi, Ismael, et al.
Published: (2026)
SHOW3D: Capturing Scenes of 3D Hands and Objects in the Wild
by: Rim, Patrick, et al.
Published: (2026)
by: Rim, Patrick, et al.
Published: (2026)
SceneCraft: Layout-Guided 3D Scene Generation
by: Yang, Xiuyu, et al.
Published: (2024)
by: Yang, Xiuyu, et al.
Published: (2024)
EBNEO Commentary: Mild Hypoxic–Ischemic Encephalopathy ( HIE ): Timing and Pattern of MRI Brain Injury
by: Mahmoud Abdelreheem, et al.
Published: (2025)
by: Mahmoud Abdelreheem, et al.
Published: (2025)
AudioScene: Integrating Object-Event Audio into 3D Scenes
by: Yuan, Shuaihang, et al.
Published: (2025)
by: Yuan, Shuaihang, et al.
Published: (2025)
AnyPlace: Learning Generalized Object Placement for Robot Manipulation
by: Zhao, Yuchi, et al.
Published: (2025)
by: Zhao, Yuchi, et al.
Published: (2025)
Retrieving Objects from 3D Scenes with Box-Guided Open-Vocabulary Instance Segmentation
by: Nguyen, Khanh, et al.
Published: (2025)
by: Nguyen, Khanh, et al.
Published: (2025)
Similar Items
-
DoubleTake: Geometry Guided Depth Estimation
by: Sayed, Mohamed, et al.
Published: (2024) -
AirPlanes: Accurate Plane Estimation via 3D-Consistent Embeddings
by: Watson, Jamie, et al.
Published: (2024) -
Build-A-Scene: Interactive 3D Layout Control for Diffusion-Based Image Generation
by: Eldesokey, Abdelrahman, et al.
Published: (2024) -
ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models
by: Gong, Bingchen, et al.
Published: (2024) -
LatentMan: Generating Consistent Animated Characters using Image Diffusion Models
by: Eldesokey, Abdelrahman, et al.
Published: (2023)