Visual Jenga: Discovering Object Dependencies via Counterfactual Inpainting
Fuente:
arXiv
Saved in:
| Main Authors: | Bhattad, Anand, Preechakul, Konpat, Efros, Alexei A. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimizing Diffusion Noise Can Serve As Universal Motion Priors
by: Karunratanakul, Korrawe, et al.
Published: (2023)
by: Karunratanakul, Korrawe, et al.
Published: (2023)
Stronger Semantic Encoders Can Harm Relighting Performance: Probing Visual Priors via Augmented Latent Intrinsics
by: Xing, Xiaoyan, et al.
Published: (2026)
by: Xing, Xiaoyan, et al.
Published: (2026)
ZeroComp: Zero-shot Object Compositing from Image Intrinsics via Diffusion
by: Zhang, Zitian, et al.
Published: (2024)
by: Zhang, Zitian, et al.
Published: (2024)
Interpreting CLIP's Image Representation via Text-Based Decomposition
by: Gandelsman, Yossi, et al.
Published: (2023)
by: Gandelsman, Yossi, et al.
Published: (2023)
Interpreting the Second-Order Effects of Neurons in CLIP
by: Gandelsman, Yossi, et al.
Published: (2024)
by: Gandelsman, Yossi, et al.
Published: (2024)
Objects in Generated Videos Are Slower Than They Appear: Models Suffer Sub-Earth Gravity and Don't Know Galileo's Principle...for now
by: Thozhiyoor, Varun Varma, et al.
Published: (2025)
by: Thozhiyoor, Varun Varma, et al.
Published: (2025)
SpotLight: Shadow-Guided Object Relighting via Diffusion
by: Fortier-Chouinard, Frédéric, et al.
Published: (2024)
by: Fortier-Chouinard, Frédéric, et al.
Published: (2024)
SyncFix: Fixing 3D Reconstructions via Multi-View Synchronization
by: Li, Deming, et al.
Published: (2026)
by: Li, Deming, et al.
Published: (2026)
Rethinking Patch Dependence for Masked Autoencoders
by: Fu, Letian, et al.
Published: (2024)
by: Fu, Letian, et al.
Published: (2024)
Evaluating Multiview Object Consistency in Humans and Image Models
by: Bonnen, Tyler, et al.
Published: (2024)
by: Bonnen, Tyler, et al.
Published: (2024)
GPS as a Control Signal for Image Generation
by: Feng, Chao, et al.
Published: (2025)
by: Feng, Chao, et al.
Published: (2025)
Continuous 3D Perception Model with Persistent State
by: Wang, Qianqian, et al.
Published: (2025)
by: Wang, Qianqian, et al.
Published: (2025)
Name That Part: 3D Part Segmentation and Naming
by: Paul, Soumava, et al.
Published: (2025)
by: Paul, Soumava, et al.
Published: (2025)
Videoshop: Localized Semantic Video Editing with Noise-Extrapolated Diffusion Inversion
by: Fan, Xiang, et al.
Published: (2024)
by: Fan, Xiang, et al.
Published: (2024)
Placing Objects in Context via Inpainting for Out-of-distribution Segmentation
by: de Jorge, Pau, et al.
Published: (2024)
by: de Jorge, Pau, et al.
Published: (2024)
Inpaint360GS: Efficient Object-Aware 3D Inpainting via Gaussian Splatting for 360° Scenes
by: Wang, Shaoxiang, et al.
Published: (2025)
by: Wang, Shaoxiang, et al.
Published: (2025)
Vision Transformers Don't Need Trained Registers
by: Jiang, Nick, et al.
Published: (2025)
by: Jiang, Nick, et al.
Published: (2025)
VisualChef: Generating Visual Aids in Cooking via Mask Inpainting
by: Kuzyk, Oleh, et al.
Published: (2025)
by: Kuzyk, Oleh, et al.
Published: (2025)
Reference-Guided Diffusion Inpainting For Multimodal Counterfactual Generation
by: Buburuzan, Alexandru
Published: (2025)
by: Buburuzan, Alexandru
Published: (2025)
ScribbleLight: Single Image Indoor Relighting with Scribbles
by: Choi, Jun Myeong, et al.
Published: (2024)
by: Choi, Jun Myeong, et al.
Published: (2024)
Generative Blocks World: Moving Things Around in Pictures
by: Vavilala, Vaibhav, et al.
Published: (2025)
by: Vavilala, Vaibhav, et al.
Published: (2025)
Inpainting-Driven Mask Optimization for Object Removal
by: Shimosato, Kodai, et al.
Published: (2024)
by: Shimosato, Kodai, et al.
Published: (2024)
IT$^3$: Idempotent Test-Time Training
by: Durasov, Nikita, et al.
Published: (2024)
by: Durasov, Nikita, et al.
Published: (2024)
COLMAP-Free 3D Gaussian Splatting
by: Fu, Yang, et al.
Published: (2023)
by: Fu, Yang, et al.
Published: (2023)
Toon3D: Seeing Cartoons from New Perspectives
by: Weber, Ethan, et al.
Published: (2024)
by: Weber, Ethan, et al.
Published: (2024)
Latent Intrinsics Emerge from Training to Relight
by: Zhang, Xiao, et al.
Published: (2024)
by: Zhang, Xiao, et al.
Published: (2024)
FreeInpaint: Tuning-free Prompt Alignment and Visual Rationality Enhancement in Image Inpainting
by: Gong, Chao, et al.
Published: (2025)
by: Gong, Chao, et al.
Published: (2025)
Generalizable Sparse-View 3D Reconstruction from Unconstrained Images
by: Gupta, Vinayak, et al.
Published: (2026)
by: Gupta, Vinayak, et al.
Published: (2026)
Improved Convex Decomposition with Ensembling and Negative Primitives
by: Vavilala, Vaibhav, et al.
Published: (2024)
by: Vavilala, Vaibhav, et al.
Published: (2024)
CLII: Visual-Text Inpainting via Cross-Modal Predictive Interaction
by: Zhao, Liang, et al.
Published: (2024)
by: Zhao, Liang, et al.
Published: (2024)
Diffusion-Based Depth Inpainting for Transparent and Reflective Objects
by: Sun, Tianyu, et al.
Published: (2024)
by: Sun, Tianyu, et al.
Published: (2024)
OmniPaint: Mastering Object-Oriented Editing via Disentangled Insertion-Removal Inpainting
by: Yu, Yongsheng, et al.
Published: (2025)
by: Yu, Yongsheng, et al.
Published: (2025)
Disentangled 3D Scene Generation with Layout Learning
by: Epstein, Dave, et al.
Published: (2024)
by: Epstein, Dave, et al.
Published: (2024)
Diffusion Models as Data Mining Tools
by: Siglidis, Ioannis, et al.
Published: (2024)
by: Siglidis, Ioannis, et al.
Published: (2024)
It's Never Too Late: Noise Optimization for Collapse Recovery in Trained Diffusion Models
by: Harrington, Anne, et al.
Published: (2025)
by: Harrington, Anne, et al.
Published: (2025)
MTADiffusion: Mask Text Alignment Diffusion Model for Object Inpainting
by: Huang, Jun, et al.
Published: (2025)
by: Huang, Jun, et al.
Published: (2025)
MObI: Multimodal Object Inpainting Using Diffusion Models
by: Buburuzan, Alexandru, et al.
Published: (2025)
by: Buburuzan, Alexandru, et al.
Published: (2025)
Jenga Stacking Based on 6D Pose Estimation for Architectural Form Finding Process
by: Huang, Zixun
Published: (2023)
by: Huang, Zixun
Published: (2023)
Discovering Concept Directions from Diffusion-based Counterfactuals via Latent Clustering
by: Varshney, Payal, et al.
Published: (2025)
by: Varshney, Payal, et al.
Published: (2025)
Geometry of the Visual Cortex with Applications to Image Inpainting and Enhancement
by: Ballerin, Francesco, et al.
Published: (2023)
by: Ballerin, Francesco, et al.
Published: (2023)
Similar Items
-
Optimizing Diffusion Noise Can Serve As Universal Motion Priors
by: Karunratanakul, Korrawe, et al.
Published: (2023) -
Stronger Semantic Encoders Can Harm Relighting Performance: Probing Visual Priors via Augmented Latent Intrinsics
by: Xing, Xiaoyan, et al.
Published: (2026) -
ZeroComp: Zero-shot Object Compositing from Image Intrinsics via Diffusion
by: Zhang, Zitian, et al.
Published: (2024) -
Interpreting CLIP's Image Representation via Text-Based Decomposition
by: Gandelsman, Yossi, et al.
Published: (2023) -
Interpreting the Second-Order Effects of Neurons in CLIP
by: Gandelsman, Yossi, et al.
Published: (2024)