Bringing Objects to Life: training-free 4D generation from 3D objects through view consistent noise
Fuente:
arXiv
Saved in:
| Main Authors: | Rahamim, Ohad, Malca, Ori, Samuel, Dvir, Chechik, Gal |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Per-Query Visual Concept Learning
by: Malca, Ori, et al.
Published: (2025)
by: Malca, Ori, et al.
Published: (2025)
Lay-A-Scene: Personalized 3D Object Arrangement Using Text-to-Image Priors
by: Rahamim, Ohad, et al.
Published: (2024)
by: Rahamim, Ohad, et al.
Published: (2024)
Fast 4D Mesh Generation by Spatio-Temporal Attention Chains
by: Samuel, Dvir, et al.
Published: (2026)
by: Samuel, Dvir, et al.
Published: (2026)
OmnimatteZero: Fast Training-free Omnimatte with Pre-trained Video Diffusion Models
by: Samuel, Dvir, et al.
Published: (2025)
by: Samuel, Dvir, et al.
Published: (2025)
Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models
by: Tewel, Yoad, et al.
Published: (2024)
by: Tewel, Yoad, et al.
Published: (2024)
Where's Waldo: Diffusion Features for Personalized Segmentation and Retrieval
by: Samuel, Dvir, et al.
Published: (2024)
by: Samuel, Dvir, et al.
Published: (2024)
DiffUHaul: A Training-Free Method for Object Dragging in Images
by: Avrahami, Omri, et al.
Published: (2024)
by: Avrahami, Omri, et al.
Published: (2024)
PatchContrast: Self-Supervised Pre-training for 3D Object Detection
by: Shrout, Oren, et al.
Published: (2023)
by: Shrout, Oren, et al.
Published: (2023)
Fast Autoregressive Video Diffusion and World Models with Temporal Cache Compression and Sparse Attention
by: Samuel, Dvir, et al.
Published: (2026)
by: Samuel, Dvir, et al.
Published: (2026)
Lightning-Fast Image Inversion and Editing for Text-to-Image Diffusion Models
by: Samuel, Dvir, et al.
Published: (2023)
by: Samuel, Dvir, et al.
Published: (2023)
MVD-Fusion: Single-view 3D via Depth-consistent Multi-view Generation
by: Hu, Hanzhe, et al.
Published: (2024)
by: Hu, Hanzhe, et al.
Published: (2024)
Text2Model: Text-based Model Induction for Zero-shot Image Classification
by: Amosy, Ohad, et al.
Published: (2022)
by: Amosy, Ohad, et al.
Published: (2022)
FutrTrack: A Camera-LiDAR Fusion Transformer for 3D Multiple Object Tracking
by: Teye, Martha Teiko, et al.
Published: (2025)
by: Teye, Martha Teiko, et al.
Published: (2025)
Data-Driven Loss Functions for Inference-Time Optimization in Text-to-Image
by: Yiflach, Sapir Esther, et al.
Published: (2025)
by: Yiflach, Sapir Esther, et al.
Published: (2025)
Motion by Queries: Identity-Motion Trade-offs in Text-to-Video Generation
by: Atzmon, Yuval, et al.
Published: (2024)
by: Atzmon, Yuval, et al.
Published: (2024)
Efficient multi-view training for 3D Gaussian Splatting
by: Choi, Minhyuk, et al.
Published: (2025)
by: Choi, Minhyuk, et al.
Published: (2025)
Hybrid bundle-adjusting 3D Gaussians for view consistent rendering with pose optimization
by: Guo, Yanan, et al.
Published: (2024)
by: Guo, Yanan, et al.
Published: (2024)
Single Image Iterative Subject-driven Generation and Editing
by: Shpitzer, Yair, et al.
Published: (2025)
by: Shpitzer, Yair, et al.
Published: (2025)
Efficient4D: Fast Dynamic 3D Object Generation from a Single-view Video
by: Pan, Zijie, et al.
Published: (2024)
by: Pan, Zijie, et al.
Published: (2024)
LiDAR MOT-DETR: A LiDAR-based Two-Stage Transformer for 3D Multiple Object Tracking
by: Teye, Martha Teiko, et al.
Published: (2025)
by: Teye, Martha Teiko, et al.
Published: (2025)
High-fidelity 3D Gaussian Inpainting: preserving multi-view consistency and photorealistic details
by: Zhou, Jun, et al.
Published: (2025)
by: Zhou, Jun, et al.
Published: (2025)
Key-Locked Rank One Editing for Text-to-Image Personalization
by: Tewel, Yoad, et al.
Published: (2023)
by: Tewel, Yoad, et al.
Published: (2023)
Adapting to the Unknown: Training-Free Audio-Visual Event Perception with Dynamic Thresholds
by: Shaar, Eitan, et al.
Published: (2025)
by: Shaar, Eitan, et al.
Published: (2025)
OPEN: Object-wise Position Embedding for Multi-view 3D Object Detection
by: Hou, Jinghua, et al.
Published: (2024)
by: Hou, Jinghua, et al.
Published: (2024)
Bringing Your Portrait to 3D Presence
by: Zhang, Jiawei, et al.
Published: (2025)
by: Zhang, Jiawei, et al.
Published: (2025)
Style3D: Attention-guided Multi-view Style Transfer for 3D Object Generation
by: Song, Bingjie, et al.
Published: (2024)
by: Song, Bingjie, et al.
Published: (2024)
Policy Optimized Text-to-Image Pipeline Design
by: Gadot, Uri, et al.
Published: (2025)
by: Gadot, Uri, et al.
Published: (2025)
SOGDet: Semantic-Occupancy Guided Multi-view 3D Object Detection
by: Zhou, Qiu, et al.
Published: (2023)
by: Zhou, Qiu, et al.
Published: (2023)
HOMER: Homography-Based Efficient Multi-view 3D Object Removal
by: Ni, Jingcheng, et al.
Published: (2025)
by: Ni, Jingcheng, et al.
Published: (2025)
Make It Count: Text-to-Image Generation with an Accurate Number of Objects
by: Binyamin, Lital, et al.
Published: (2024)
by: Binyamin, Lital, et al.
Published: (2024)
Assessing Image Quality Using a Simple Generative Representation
by: Raviv, Simon, et al.
Published: (2024)
by: Raviv, Simon, et al.
Published: (2024)
Muskie: Multi-view Masked Image Modeling for 3D Vision Pre-training
by: Li, Wenyu, et al.
Published: (2025)
by: Li, Wenyu, et al.
Published: (2025)
Find your Needle: Small Object Image Retrieval via Multi-Object Attention Optimization
by: Green, Michael, et al.
Published: (2025)
by: Green, Michael, et al.
Published: (2025)
SR3D: Unleashing Single-view 3D Reconstruction for Transparent and Specular Object Grasping
by: Zhang, Mingxu, et al.
Published: (2025)
by: Zhang, Mingxu, et al.
Published: (2025)
PhyGenHOI: Physically-Aware 4D Generation of Dynamic Human-Object Interactions
by: Benishu, Omer, et al.
Published: (2026)
by: Benishu, Omer, et al.
Published: (2026)
IT$^3$: Idempotent Test-Time Training
by: Durasov, Nikita, et al.
Published: (2024)
by: Durasov, Nikita, et al.
Published: (2024)
FreeSplatter: Pose-free Gaussian Splatting for Sparse-view 3D Reconstruction
by: Xu, Jiale, et al.
Published: (2024)
by: Xu, Jiale, et al.
Published: (2024)
Compositional Video Generation via Inference-Time Guidance
by: Shaulov, Ariel, et al.
Published: (2026)
by: Shaulov, Ariel, et al.
Published: (2026)
FSD-BEV: Foreground Self-Distillation for Multi-view 3D Object Detection
by: Jiang, Zheng, et al.
Published: (2024)
by: Jiang, Zheng, et al.
Published: (2024)
TriTex: Learning Texture from a Single Mesh via Triplane Semantic Features
by: Cohen-Bar, Dana, et al.
Published: (2025)
by: Cohen-Bar, Dana, et al.
Published: (2025)
Similar Items
-
Per-Query Visual Concept Learning
by: Malca, Ori, et al.
Published: (2025) -
Lay-A-Scene: Personalized 3D Object Arrangement Using Text-to-Image Priors
by: Rahamim, Ohad, et al.
Published: (2024) -
Fast 4D Mesh Generation by Spatio-Temporal Attention Chains
by: Samuel, Dvir, et al.
Published: (2026) -
OmnimatteZero: Fast Training-free Omnimatte with Pre-trained Video Diffusion Models
by: Samuel, Dvir, et al.
Published: (2025) -
Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models
by: Tewel, Yoad, et al.
Published: (2024)