ZeST: Zero-Shot Material Transfer from a Single Image
Fuente:
arXiv
Saved in:
| Main Authors: | Cheng, Ta-Ying, Sharma, Prafull, Markham, Andrew, Trigoni, Niki, Jampani, Varun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MARBLE: Material Recomposition and Blending in CLIP-Space
by: Cheng, Ta-Ying, et al.
Published: (2025)
by: Cheng, Ta-Ying, et al.
Published: (2025)
ZeST: an LLM-based Zero-Shot Traversability Navigation for Unknown Environments
by: Gummadi, Shreya, et al.
Published: (2025)
by: Gummadi, Shreya, et al.
Published: (2025)
SpatialPIN: Enhancing Spatial Reasoning Capabilities of Vision-Language Models through Prompting and Interacting 3D Priors
by: Ma, Chenyang, et al.
Published: (2024)
by: Ma, Chenyang, et al.
Published: (2024)
Learning Continuous 3D Words for Text-to-Image Generation
by: Cheng, Ta-Ying, et al.
Published: (2024)
by: Cheng, Ta-Ying, et al.
Published: (2024)
WSCLoc: Weakly-Supervised Sparse-View Camera Relocalization
by: Wang, Jialu, et al.
Published: (2024)
by: Wang, Jialu, et al.
Published: (2024)
MambaLoc: Efficient Camera Localisation via State Space Model
by: Wang, Jialu, et al.
Published: (2024)
by: Wang, Jialu, et al.
Published: (2024)
FROMAT: Multiview Material Appearance Transfer via Few-Shot Self-Attention Adaptation
by: Kompanowski, Hubert, et al.
Published: (2025)
by: Kompanowski, Hubert, et al.
Published: (2025)
Spherical Mask: Coarse-to-Fine 3D Point Cloud Instance Segmentation with Spherical Representation
by: Shin, Sangyun, et al.
Published: (2023)
by: Shin, Sangyun, et al.
Published: (2023)
Data Factory with Minimal Human Effort Using VLMs
by: Ye, Jiaojiao, et al.
Published: (2025)
by: Ye, Jiaojiao, et al.
Published: (2025)
Dusk Till Dawn: Self-supervised Nighttime Stereo Depth Estimation using Visual Foundation Models
by: Vankadari, Madhu, et al.
Published: (2024)
by: Vankadari, Madhu, et al.
Published: (2024)
SoundLoc3D: Invisible 3D Sound Source Localization and Classification Using a Multimodal RGB-D Acoustic Camera
by: He, Yuhang, et al.
Published: (2024)
by: He, Yuhang, et al.
Published: (2024)
DynPoint: Dynamic Neural Point For View Synthesis
by: Zhou, Kaichen, et al.
Published: (2023)
by: Zhou, Kaichen, et al.
Published: (2023)
Gen4Gen: Generative Data Pipeline for Generative Multi-Concept Composition
by: Yeh, Chun-Hsiao, et al.
Published: (2024)
by: Yeh, Chun-Hsiao, et al.
Published: (2024)
ZePo: Zero-Shot Portrait Stylization with Faster Sampling
by: Liu, Jin, et al.
Published: (2024)
by: Liu, Jin, et al.
Published: (2024)
Manydepth2: Motion-Aware Self-Supervised Monocular Depth Estimation in Dynamic Scenes
by: Zhou, Kaichen, et al.
Published: (2023)
by: Zhou, Kaichen, et al.
Published: (2023)
HumANDiff: Articulated Noise Diffusion for Motion-Consistent Human Video Generation
by: Hu, Tao, et al.
Published: (2026)
by: Hu, Tao, et al.
Published: (2026)
Computational Tradeoffs in Image Synthesis: Diffusion, Masked-Token, and Next-Token Prediction
by: Kilian, Maciej, et al.
Published: (2024)
by: Kilian, Maciej, et al.
Published: (2024)
VMLoc: Variational Fusion For Learning-Based Multimodal Camera Localization
by: Zhou, Kaichen, et al.
Published: (2020)
by: Zhou, Kaichen, et al.
Published: (2020)
ZeroShape: Regression-based Zero-shot Shape Reconstruction
by: Huang, Zixuan, et al.
Published: (2023)
by: Huang, Zixuan, et al.
Published: (2023)
MVD-Fusion: Single-view 3D via Depth-consistent Multi-view Generation
by: Hu, Hanzhe, et al.
Published: (2024)
by: Hu, Hanzhe, et al.
Published: (2024)
FaceCraft4D: Animated 3D Facial Avatar Generation from a Single Image
by: Yin, Fei, et al.
Published: (2025)
by: Yin, Fei, et al.
Published: (2025)
ZePT: Zero-Shot Pan-Tumor Segmentation via Query-Disentangling and Self-Prompting
by: Jiang, Yankai, et al.
Published: (2023)
by: Jiang, Yankai, et al.
Published: (2023)
SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
by: Huang, Zixuan, et al.
Published: (2025)
by: Huang, Zixuan, et al.
Published: (2025)
Towards Multi-Modal Animal Pose Estimation: A Survey and In-Depth Analysis
by: Deng, Qianyi, et al.
Published: (2024)
by: Deng, Qianyi, et al.
Published: (2024)
Human Video Generation from a Single Image with 3D Pose and View Control
by: Wang, Tiantian, et al.
Published: (2026)
by: Wang, Tiantian, et al.
Published: (2026)
MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion
by: Zhang, Junyi, et al.
Published: (2024)
by: Zhang, Junyi, et al.
Published: (2024)
SViM3D: Stable Video Material Diffusion for Single Image 3D Generation
by: Engelhardt, Andreas, et al.
Published: (2025)
by: Engelhardt, Andreas, et al.
Published: (2025)
ZeD-MAP: Bundle Adjustment Guided Zero-Shot Depth Maps for Real-Time Aerial Imaging
by: Iz, Selim Ahmet, et al.
Published: (2026)
by: Iz, Selim Ahmet, et al.
Published: (2026)
MaterialFusion: High-Quality, Zero-Shot, and Controllable Material Transfer with Diffusion Models
by: Garifullin, Kamil, et al.
Published: (2025)
by: Garifullin, Kamil, et al.
Published: (2025)
WordRobe: Text-Guided Generation of Textured 3D Garments
by: Srivastava, Astitva, et al.
Published: (2024)
by: Srivastava, Astitva, et al.
Published: (2024)
ZeBROD: Zero-Retraining Based Recognition and Object Detection Framework
by: Hidayatullah, Priyanto, et al.
Published: (2025)
by: Hidayatullah, Priyanto, et al.
Published: (2025)
LightHeadEd: Relightable & Editable Head Avatars from a Smartphone
by: Manu, Pranav, et al.
Published: (2025)
by: Manu, Pranav, et al.
Published: (2025)
Understanding the Cross-Domain Capabilities of Video-Based Few-Shot Action Recognition Models
by: Markham, Georgia, et al.
Published: (2024)
by: Markham, Georgia, et al.
Published: (2024)
TripoSR: Fast 3D Object Reconstruction from a Single Image
by: Tochilkin, Dmitry, et al.
Published: (2024)
by: Tochilkin, Dmitry, et al.
Published: (2024)
Learning Action and Reasoning-Centric Image Editing from Videos and Simulations
by: Krojer, Benno, et al.
Published: (2024)
by: Krojer, Benno, et al.
Published: (2024)
EmoCLIP: A Vision-Language Method for Zero-Shot Video Facial Expression Recognition
by: Foteinopoulou, Niki Maria, et al.
Published: (2023)
by: Foteinopoulou, Niki Maria, et al.
Published: (2023)
Zero-Shot Scene Reconstruction from Single Images with Deep Prior Assembly
by: Zhou, Junsheng, et al.
Published: (2024)
by: Zhou, Junsheng, et al.
Published: (2024)
SV3D: Novel Multi-view Synthesis and 3D Generation from a Single Image using Latent Video Diffusion
by: Voleti, Vikram, et al.
Published: (2024)
by: Voleti, Vikram, et al.
Published: (2024)
DreamInsert: Zero-Shot Image-to-Video Object Insertion from A Single Image
by: Zhao, Qi, et al.
Published: (2025)
by: Zhao, Qi, et al.
Published: (2025)
SF3D: Stable Fast 3D Mesh Reconstruction with UV-unwrapping and Illumination Disentanglement
by: Boss, Mark, et al.
Published: (2024)
by: Boss, Mark, et al.
Published: (2024)
Similar Items
-
MARBLE: Material Recomposition and Blending in CLIP-Space
by: Cheng, Ta-Ying, et al.
Published: (2025) -
ZeST: an LLM-based Zero-Shot Traversability Navigation for Unknown Environments
by: Gummadi, Shreya, et al.
Published: (2025) -
SpatialPIN: Enhancing Spatial Reasoning Capabilities of Vision-Language Models through Prompting and Interacting 3D Priors
by: Ma, Chenyang, et al.
Published: (2024) -
Learning Continuous 3D Words for Text-to-Image Generation
by: Cheng, Ta-Ying, et al.
Published: (2024) -
WSCLoc: Weakly-Supervised Sparse-View Camera Relocalization
by: Wang, Jialu, et al.
Published: (2024)