ReasonX: MLLM-Guided Intrinsic Image Decomposition
Fuente:
arXiv
Saved in:
| Main Authors: | Dirik, Alara, Wang, Tuanfeng, Ceylan, Duygu, Zafeiriou, Stefanos, Frühstück, Anna |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PRISM: A Unified Framework for Photorealistic Reconstruction and Intrinsic Scene Modeling
by: Dirik, Alara, et al.
Published: (2025)
by: Dirik, Alara, et al.
Published: (2025)
Geo-ID: Test-Time Geometric Consensus for Cross-View Consistent Intrinsics
by: Dirik, Alara, et al.
Published: (2026)
by: Dirik, Alara, et al.
Published: (2026)
ID-Consistent, Precise Expression Generation with Blendshape-Guided Diffusion
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
SuperGaussian: Repurposing Video Models for 3D Super Resolution
by: Shen, Yuan, et al.
Published: (2024)
by: Shen, Yuan, et al.
Published: (2024)
V-RGBX: Video Editing with Accurate Controls over Intrinsic Properties
by: Fang, Ye, et al.
Published: (2025)
by: Fang, Ye, et al.
Published: (2025)
Neural Garment Dynamics via Manifold-Aware Transformers
by: Li, Peizhuo, et al.
Published: (2024)
by: Li, Peizhuo, et al.
Published: (2024)
DermaFlux: Synthetic Skin Lesion Generation with Rectified Flows for Enhanced Image Classification
by: Galanakis, Stathis, et al.
Published: (2026)
by: Galanakis, Stathis, et al.
Published: (2026)
Dyn-HaMR: Recovering 4D Interacting Hand Motion from a Dynamic Camera
by: Yu, Zhengdi, et al.
Published: (2024)
by: Yu, Zhengdi, et al.
Published: (2024)
Distribution Matching for Multi-Task Learning of Classification Tasks: a Large-Scale Study on Faces & Beyond
by: Kollias, Dimitrios, et al.
Published: (2024)
by: Kollias, Dimitrios, et al.
Published: (2024)
Design2Cloth: 3D Cloth Generation from 2D Masks
by: Zheng, Jiali, et al.
Published: (2024)
by: Zheng, Jiali, et al.
Published: (2024)
FitDiff: Robust monocular 3D facial shape and reflectance estimation using Diffusion Models
by: Galanakis, Stathis, et al.
Published: (2023)
by: Galanakis, Stathis, et al.
Published: (2023)
3D Stylization via Large Reconstruction Model
by: Oztas, Ipek, et al.
Published: (2025)
by: Oztas, Ipek, et al.
Published: (2025)
Reasoning Guided Embeddings: Leveraging MLLM Reasoning for Improved Multimodal Retrieval
by: Liu, Chunxu, et al.
Published: (2025)
by: Liu, Chunxu, et al.
Published: (2025)
UV-free Texture Generation with Denoising and Geodesic Heat Diffusions
by: Foti, Simone, et al.
Published: (2024)
by: Foti, Simone, et al.
Published: (2024)
Pattern Guided UV Recovery for Realistic Video Garment Texturing
by: Zhan, Youyi, et al.
Published: (2024)
by: Zhan, Youyi, et al.
Published: (2024)
ShapeFusion: A 3D diffusion model for localized shape editing
by: Potamias, Rolandos Alexandros, et al.
Published: (2024)
by: Potamias, Rolandos Alexandros, et al.
Published: (2024)
WiLoR: End-to-end 3D Hand Localization and Reconstruction in-the-wild
by: Potamias, Rolandos Alexandros, et al.
Published: (2024)
by: Potamias, Rolandos Alexandros, et al.
Published: (2024)
Context-self contrastive pretraining for crop type semantic segmentation
by: Tarasiou, Michail, et al.
Published: (2021)
by: Tarasiou, Michail, et al.
Published: (2021)
Arc2Avatar: Generating Expressive 3D Avatars from a Single Image via ID Guidance
by: Gerogiannis, Dimitrios, et al.
Published: (2025)
by: Gerogiannis, Dimitrios, et al.
Published: (2025)
SpinMeRound: Consistent Multi-View Identity Generation Using Diffusion Models
by: Galanakis, Stathis, et al.
Published: (2025)
by: Galanakis, Stathis, et al.
Published: (2025)
ID-to-3D: Expressive ID-guided 3D Heads via Score Distillation Sampling
by: Babiloni, Francesca, et al.
Published: (2024)
by: Babiloni, Francesca, et al.
Published: (2024)
Improving face generation quality and prompt following with synthetic captions
by: Tarasiou, Michail, et al.
Published: (2024)
by: Tarasiou, Michail, et al.
Published: (2024)
Deep Face Restoration: A Survey
by: Wang, Tao, et al.
Published: (2022)
by: Wang, Tao, et al.
Published: (2022)
Colorful Diffuse Intrinsic Image Decomposition in the Wild
by: Careaga, Chris, et al.
Published: (2024)
by: Careaga, Chris, et al.
Published: (2024)
ImHead: A Large-scale Implicit Morphable Model for Localized Head Modeling
by: Potamias, Rolandos Alexandros, et al.
Published: (2025)
by: Potamias, Rolandos Alexandros, et al.
Published: (2025)
SAGS: Structure-Aware 3D Gaussian Splatting
by: Ververas, Evangelos, et al.
Published: (2024)
by: Ververas, Evangelos, et al.
Published: (2024)
MonetGPT: Solving Puzzles Enhances MLLMs' Image Retouching Skills
by: Dutt, Niladri Shekhar, et al.
Published: (2025)
by: Dutt, Niladri Shekhar, et al.
Published: (2025)
DISN: Deep Implicit Surface Network for High-quality Single-view 3D Reconstruction
by: Xu, Qiangeng, et al.
Published: (2019)
by: Xu, Qiangeng, et al.
Published: (2019)
Multi-scale Attention-Guided Intrinsic Decomposition and Rendering Pass Prediction for Facial Images
by: Javidnia, Hossein
Published: (2025)
by: Javidnia, Hossein
Published: (2025)
JoIN: Joint GANs Inversion for Intrinsic Image Decomposition
by: Shah, Viraj, et al.
Published: (2023)
by: Shah, Viraj, et al.
Published: (2023)
Intrinsic Image Decomposition Using Point Cloud Representation
by: Xing, Xiaoyan, et al.
Published: (2023)
by: Xing, Xiaoyan, et al.
Published: (2023)
Texture-aware Intrinsic Image Decomposition with Model- and Learning-based Priors
by: Wang, Xiaodong, et al.
Published: (2025)
by: Wang, Xiaodong, et al.
Published: (2025)
GeoFusionLRM: Geometry-Aware Self-Correction for Consistent 3D Reconstruction
by: Yildirim, Ahmet Burak, et al.
Published: (2026)
by: Yildirim, Ahmet Burak, et al.
Published: (2026)
STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025)
Dex2HOI: Dexterous Bimanual Two-Object Interaction Generation
by: Pratikaki, Chrysa, et al.
Published: (2026)
by: Pratikaki, Chrysa, et al.
Published: (2026)
Locally Adaptive Neural 3D Morphable Models
by: Tarasiou, Michail, et al.
Published: (2024)
by: Tarasiou, Michail, et al.
Published: (2024)
Signs as Tokens: A Retrieval-Enhanced Multilingual Sign Language Generator
by: Zuo, Ronglai, et al.
Published: (2024)
by: Zuo, Ronglai, et al.
Published: (2024)
Large Learning Rates Simultaneously Achieve Robustness to Spurious Correlations and Compressibility
by: Barsbey, Melih, et al.
Published: (2025)
by: Barsbey, Melih, et al.
Published: (2025)
Neural Sign Actors: A diffusion model for 3D sign language production from text
by: Baltatzis, Vasileios, et al.
Published: (2023)
by: Baltatzis, Vasileios, et al.
Published: (2023)
Do You See What I Am Pointing At? Gesture-Based Egocentric Video Question Answering
by: Choi, Yura, et al.
Published: (2026)
by: Choi, Yura, et al.
Published: (2026)
Similar Items
-
PRISM: A Unified Framework for Photorealistic Reconstruction and Intrinsic Scene Modeling
by: Dirik, Alara, et al.
Published: (2025) -
Geo-ID: Test-Time Geometric Consensus for Cross-View Consistent Intrinsics
by: Dirik, Alara, et al.
Published: (2026) -
ID-Consistent, Precise Expression Generation with Blendshape-Guided Diffusion
by: Papantoniou, Foivos Paraperas, et al.
Published: (2025) -
SuperGaussian: Repurposing Video Models for 3D Super Resolution
by: Shen, Yuan, et al.
Published: (2024) -
V-RGBX: Video Editing with Accurate Controls over Intrinsic Properties
by: Fang, Ye, et al.
Published: (2025)