Copy-Trasform-Paste: Zero-Shot Object-Object Alignment Guided by Vision-Language and Geometric Constraints
Fuente:
arXiv
Guardado en:
| Autores principales: | Gatenyo, Rotem, Fried, Ohad |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ImageRAG: Dynamic Image Retrieval for Reference-Guided Image Generation
por: Shalev-Arkushin, Rotem, et al.
Publicado: (2025)
por: Shalev-Arkushin, Rotem, et al.
Publicado: (2025)
Express4D: Expressive, Friendly, and Extensible 4D Facial Motion Generation Benchmark
por: Aloni, Yaron, et al.
Publicado: (2025)
por: Aloni, Yaron, et al.
Publicado: (2025)
DiffUHaul: A Training-Free Method for Object Dragging in Images
por: Avrahami, Omri, et al.
Publicado: (2024)
por: Avrahami, Omri, et al.
Publicado: (2024)
Monkey See, Monkey Do: Harnessing Self-attention in Motion Diffusion for Zero-shot Motion Transfer
por: Raab, Sigal, et al.
Publicado: (2024)
por: Raab, Sigal, et al.
Publicado: (2024)
Next-Scale Autoregressive Models are Zero-Shot Single-Image Object View Synthesizers
por: Yuan, Shiran, et al.
Publicado: (2025)
por: Yuan, Shiran, et al.
Publicado: (2025)
HOI-PAGE: Zero-Shot Human-Object Interaction Generation with Part Affordance Guidance
por: Li, Lei, et al.
Publicado: (2025)
por: Li, Lei, et al.
Publicado: (2025)
InteractEdit: Zero-Shot Editing of Human-Object Interactions in Images
por: Hoe, Jiun Tian, et al.
Publicado: (2025)
por: Hoe, Jiun Tian, et al.
Publicado: (2025)
V-LASIK: Consistent Glasses-Removal from Videos Using Synthetic Data
por: Shalev-Arkushin, Rotem, et al.
Publicado: (2024)
por: Shalev-Arkushin, Rotem, et al.
Publicado: (2024)
REED-VAE: RE-Encode Decode Training for Iterative Image Editing with Diffusion Models
por: Almog, Gal, et al.
Publicado: (2025)
por: Almog, Gal, et al.
Publicado: (2025)
Zero-Shot Human-Object Interaction Synthesis with Multimodal Priors
por: Lou, Yuke, et al.
Publicado: (2025)
por: Lou, Yuke, et al.
Publicado: (2025)
Blended-NeRF: Zero-Shot Object Generation and Blending in Existing Neural Radiance Fields
por: Gordon, Ori, et al.
Publicado: (2023)
por: Gordon, Ori, et al.
Publicado: (2023)
SpotLight: Shadow-Guided Object Relighting via Diffusion
por: Fortier-Chouinard, Frédéric, et al.
Publicado: (2024)
por: Fortier-Chouinard, Frédéric, et al.
Publicado: (2024)
Stability-Driven Motion Generation for Object-Guided Human-Human Co-Manipulation
por: Xu, Jiahao, et al.
Publicado: (2026)
por: Xu, Jiahao, et al.
Publicado: (2026)
Structure-Guided Image Completion with Image-level and Object-level Semantic Discriminators
por: Zheng, Haitian, et al.
Publicado: (2022)
por: Zheng, Haitian, et al.
Publicado: (2022)
RASP: Revisiting 3D Anamorphic Art for Shadow-Guided Packing of Irregular Objects
por: Debnath, Soumyaratna, et al.
Publicado: (2025)
por: Debnath, Soumyaratna, et al.
Publicado: (2025)
Differential Diffusion: Giving Each Pixel Its Strength
por: Levin, Eran, et al.
Publicado: (2023)
por: Levin, Eran, et al.
Publicado: (2023)
Objects With Lighting: A Real-World Dataset for Evaluating Reconstruction and Rendering for Object Relighting
por: Ummenhofer, Benjamin, et al.
Publicado: (2024)
por: Ummenhofer, Benjamin, et al.
Publicado: (2024)
ZeroHSI: Zero-Shot 4D Human-Scene Interaction by Video Generation
por: Li, Hongjie, et al.
Publicado: (2024)
por: Li, Hongjie, et al.
Publicado: (2024)
Handle-based Mesh Deformation Guided By Vision Language Model
por: Sun, Xingpeng, et al.
Publicado: (2025)
por: Sun, Xingpeng, et al.
Publicado: (2025)
ZeroNVS: Zero-Shot 360-Degree View Synthesis from a Single Image
por: Sargent, Kyle, et al.
Publicado: (2023)
por: Sargent, Kyle, et al.
Publicado: (2023)
PLOT: Pseudo-Labeling via Video Object Tracking for Scalable Monocular 3D Object Detection
por: Lee, Seokyeong, et al.
Publicado: (2025)
por: Lee, Seokyeong, et al.
Publicado: (2025)
GaussianObject: High-Quality 3D Object Reconstruction from Four Views with Gaussian Splatting
por: Yang, Chen, et al.
Publicado: (2024)
por: Yang, Chen, et al.
Publicado: (2024)
Towards Geometric-Photometric Joint Alignment for Facial Mesh Registration
por: Wang, Xizhi, et al.
Publicado: (2024)
por: Wang, Xizhi, et al.
Publicado: (2024)
Monocular Human-Object Reconstruction in the Wild
por: Huo, Chaofan, et al.
Publicado: (2024)
por: Huo, Chaofan, et al.
Publicado: (2024)
Stable Flow: Vital Layers for Training-Free Image Editing
por: Avrahami, Omri, et al.
Publicado: (2024)
por: Avrahami, Omri, et al.
Publicado: (2024)
Photorealistic Object Insertion with Diffusion-Guided Inverse Rendering
por: Liang, Ruofan, et al.
Publicado: (2024)
por: Liang, Ruofan, et al.
Publicado: (2024)
Alterbute: Editing Intrinsic Attributes of Objects in Images
por: Reiss, Tal, et al.
Publicado: (2026)
por: Reiss, Tal, et al.
Publicado: (2026)
Image-Guided Geometric Stylization of 3D Meshes
por: Choi, Changwoon, et al.
Publicado: (2026)
por: Choi, Changwoon, et al.
Publicado: (2026)
ClickAIXR: On-Device Multimodal Vision-Language Interaction with Real-World Objects in Extended Reality
por: Khan, Dawar, et al.
Publicado: (2026)
por: Khan, Dawar, et al.
Publicado: (2026)
HOLODECK 2.0: Vision-Language-Guided 3D World Generation with Editing
por: Bian, Zixuan, et al.
Publicado: (2025)
por: Bian, Zixuan, et al.
Publicado: (2025)
CRAFT: Designing Creative and Functional 3D Objects
por: Guo, Michelle, et al.
Publicado: (2024)
por: Guo, Michelle, et al.
Publicado: (2024)
Object-Aware 4D Human Motion Generation
por: Gui, Shurui, et al.
Publicado: (2025)
por: Gui, Shurui, et al.
Publicado: (2025)
Relighting Scenes with Object Insertions in Neural Radiance Fields
por: Zhu, Xuening, et al.
Publicado: (2024)
por: Zhu, Xuening, et al.
Publicado: (2024)
ZeroScene: A Zero-Shot Framework for 3D Scene Generation from a Single Image and Controllable Texture Editing
por: Tang, Xiang, et al.
Publicado: (2025)
por: Tang, Xiang, et al.
Publicado: (2025)
Selfi: Self Improving Reconstruction Engine via 3D Geometric Feature Alignment
por: Deng, Youming, et al.
Publicado: (2025)
por: Deng, Youming, et al.
Publicado: (2025)
NeuROK: Generative 4D Neural Object Kinematics
por: Geng, Chen, et al.
Publicado: (2026)
por: Geng, Chen, et al.
Publicado: (2026)
GausSim: Foreseeing Reality by Gaussian Simulator for Elastic Objects
por: Shao, Yidi, et al.
Publicado: (2024)
por: Shao, Yidi, et al.
Publicado: (2024)
DecoupledGaussian: Object-Scene Decoupling for Physics-Based Interaction
por: Wang, Miaowei, et al.
Publicado: (2025)
por: Wang, Miaowei, et al.
Publicado: (2025)
ROGR: Relightable 3D Objects using Generative Relighting
por: Tang, Jiapeng, et al.
Publicado: (2025)
por: Tang, Jiapeng, et al.
Publicado: (2025)
SOPHY: Learning to Generate Simulation-Ready Objects with Physical Materials
por: Cao, Junyi, et al.
Publicado: (2025)
por: Cao, Junyi, et al.
Publicado: (2025)
Ejemplares similares
-
ImageRAG: Dynamic Image Retrieval for Reference-Guided Image Generation
por: Shalev-Arkushin, Rotem, et al.
Publicado: (2025) -
Express4D: Expressive, Friendly, and Extensible 4D Facial Motion Generation Benchmark
por: Aloni, Yaron, et al.
Publicado: (2025) -
DiffUHaul: A Training-Free Method for Object Dragging in Images
por: Avrahami, Omri, et al.
Publicado: (2024) -
Monkey See, Monkey Do: Harnessing Self-attention in Motion Diffusion for Zero-shot Motion Transfer
por: Raab, Sigal, et al.
Publicado: (2024) -
Next-Scale Autoregressive Models are Zero-Shot Single-Image Object View Synthesizers
por: Yuan, Shiran, et al.
Publicado: (2025)