SIRR-LMM: Single-image Reflection Removal via Large Multimodal Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Guo, Yu, Lao, Zhiqiang, Song, Xiyun, Zhou, Yubin, Yu, Heather |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ePBR: Extended PBR Materials in Image Synthesis
von: Guo, Yu, et al.
Veröffentlicht: (2025)
von: Guo, Yu, et al.
Veröffentlicht: (2025)
Towards Understanding Graphical Perception in Large Multimodal Models
von: Zhang, Kai, et al.
Veröffentlicht: (2025)
von: Zhang, Kai, et al.
Veröffentlicht: (2025)
Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026)
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026)
Proc3D: Procedural 3D Generation and Parametric Editing of 3D Shapes with Large Language Models
von: Raji, Fadlullah, et al.
Veröffentlicht: (2026)
von: Raji, Fadlullah, et al.
Veröffentlicht: (2026)
Vision-as-Inverse-Graphics Agent via Interleaved Multimodal Reasoning
von: Yin, Shaofeng, et al.
Veröffentlicht: (2026)
von: Yin, Shaofeng, et al.
Veröffentlicht: (2026)
LRM: Large Reconstruction Model for Single Image to 3D
von: Hong, Yicong, et al.
Veröffentlicht: (2023)
von: Hong, Yicong, et al.
Veröffentlicht: (2023)
OT-Talk: Animating 3D Talking Head with Optimal Transportation
von: Wang, Xinmu, et al.
Veröffentlicht: (2025)
von: Wang, Xinmu, et al.
Veröffentlicht: (2025)
Ref-DGS: Reflective Dual Gaussian Splatting
von: Fan, Ningjing, et al.
Veröffentlicht: (2026)
von: Fan, Ningjing, et al.
Veröffentlicht: (2026)
WonderPlay: Dynamic 3D Scene Generation from a Single Image and Actions
von: Li, Zizhang, et al.
Veröffentlicht: (2025)
von: Li, Zizhang, et al.
Veröffentlicht: (2025)
V-LASIK: Consistent Glasses-Removal from Videos Using Synthetic Data
von: Shalev-Arkushin, Rotem, et al.
Veröffentlicht: (2024)
von: Shalev-Arkushin, Rotem, et al.
Veröffentlicht: (2024)
Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control
von: Song, Chenxi, et al.
Veröffentlicht: (2025)
von: Song, Chenxi, et al.
Veröffentlicht: (2025)
SAEdit: Token-level control for continuous image editing via Sparse AutoEncoder
von: Kamenetsky, Ronen, et al.
Veröffentlicht: (2025)
von: Kamenetsky, Ronen, et al.
Veröffentlicht: (2025)
WorldCraft: Photo-Realistic 3D World Creation and Customization via LLM Agents
von: Liu, Xinhang, et al.
Veröffentlicht: (2025)
von: Liu, Xinhang, et al.
Veröffentlicht: (2025)
SRDiffusion: Accelerate Video Diffusion Inference via Sketching-Rendering Cooperation
von: Cheng, Shenggan, et al.
Veröffentlicht: (2025)
von: Cheng, Shenggan, et al.
Veröffentlicht: (2025)
Pygmalion Effect in Vision: Image-to-Clay Translation for Reflective Geometry Reconstruction
von: Lee, Gayoung, et al.
Veröffentlicht: (2025)
von: Lee, Gayoung, et al.
Veröffentlicht: (2025)
TEXGen: a Generative Diffusion Model for Mesh Textures
von: Yu, Xin, et al.
Veröffentlicht: (2024)
von: Yu, Xin, et al.
Veröffentlicht: (2024)
SSR-GS: Separating Specular Reflection in Gaussian Splatting for Glossy Surface Reconstruction
von: Fan, Ningjing, et al.
Veröffentlicht: (2026)
von: Fan, Ningjing, et al.
Veröffentlicht: (2026)
VividFace: A Diffusion-Based Hybrid Framework for High-Fidelity Video Face Swapping
von: Shao, Hao, et al.
Veröffentlicht: (2024)
von: Shao, Hao, et al.
Veröffentlicht: (2024)
Representing Animatable Avatar via Factorized Neural Fields
von: Song, Chunjin, et al.
Veröffentlicht: (2024)
von: Song, Chunjin, et al.
Veröffentlicht: (2024)
Sin3DM: Learning a Diffusion Model from a Single 3D Textured Shape
von: Wu, Rundi, et al.
Veröffentlicht: (2023)
von: Wu, Rundi, et al.
Veröffentlicht: (2023)
StreamME: Simplify 3D Gaussian Avatar within Live Stream
von: Song, Luchuan, et al.
Veröffentlicht: (2025)
von: Song, Luchuan, et al.
Veröffentlicht: (2025)
DiffHDR: Re-Exposing LDR Videos with Video Diffusion Models
von: Yu, Zhengming, et al.
Veröffentlicht: (2026)
von: Yu, Zhengming, et al.
Veröffentlicht: (2026)
UniLat3D: Geometry-Appearance Unified Latents for Single-Stage 3D Generation
von: Wu, Guanjun, et al.
Veröffentlicht: (2025)
von: Wu, Guanjun, et al.
Veröffentlicht: (2025)
Generating by Understanding: Neural Visual Generation with Logical Symbol Groundings
von: Peng, Yifei, et al.
Veröffentlicht: (2023)
von: Peng, Yifei, et al.
Veröffentlicht: (2023)
Enhancing Monocular 3D Scene Completion with Diffusion Model
von: Song, Changlin, et al.
Veröffentlicht: (2025)
von: Song, Changlin, et al.
Veröffentlicht: (2025)
Zero-Shot Human-Object Interaction Synthesis with Multimodal Priors
von: Lou, Yuke, et al.
Veröffentlicht: (2025)
von: Lou, Yuke, et al.
Veröffentlicht: (2025)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
von: Zuo, Qi, et al.
Veröffentlicht: (2024)
von: Zuo, Qi, et al.
Veröffentlicht: (2024)
SeqTex: Generate Mesh Textures in Video Sequence
von: Yuan, Ze, et al.
Veröffentlicht: (2025)
von: Yuan, Ze, et al.
Veröffentlicht: (2025)
Inverse Rendering for High-Genus Surface Meshes from Multi-View Images
von: Gao, Xiang, et al.
Veröffentlicht: (2025)
von: Gao, Xiang, et al.
Veröffentlicht: (2025)
FreeOrbit4D: Training-Free Arbitrary Camera Redirection for Monocular Videos via Foreground-Complete 4D Reconstruction
von: Cao, Wei, et al.
Veröffentlicht: (2026)
von: Cao, Wei, et al.
Veröffentlicht: (2026)
FlexCAD: Unified and Versatile Controllable CAD Generation with Fine-tuned Large Language Models
von: Zhang, Zhanwei, et al.
Veröffentlicht: (2024)
von: Zhang, Zhanwei, et al.
Veröffentlicht: (2024)
Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
von: Liu, Bingchen, et al.
Veröffentlicht: (2024)
von: Liu, Bingchen, et al.
Veröffentlicht: (2024)
SAT: Dynamic Spatial Aptitude Training for Multimodal Language Models
von: Ray, Arijit, et al.
Veröffentlicht: (2024)
von: Ray, Arijit, et al.
Veröffentlicht: (2024)
VibraVerse: A Large-Scale Geometry-Acoustics Alignment Dataset for Physically-Consistent Multimodal Learning
von: Pang, Bo, et al.
Veröffentlicht: (2025)
von: Pang, Bo, et al.
Veröffentlicht: (2025)
Using Gaussian Splats to Create High-Fidelity Facial Geometry and Texture
von: He, Haodi, et al.
Veröffentlicht: (2025)
von: He, Haodi, et al.
Veröffentlicht: (2025)
Materialist: Physically Based Editing Using Single-Image Inverse Rendering
von: Wang, Lezhong, et al.
Veröffentlicht: (2025)
von: Wang, Lezhong, et al.
Veröffentlicht: (2025)
StableMotion: Training Motion Cleanup Models with Unpaired Corrupted Data
von: Mu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Mu, Yuxuan, et al.
Veröffentlicht: (2025)
NeRF-Insert: 3D Local Editing with Multimodal Control Signals
von: Sabat, Benet Oriol, et al.
Veröffentlicht: (2024)
von: Sabat, Benet Oriol, et al.
Veröffentlicht: (2024)
Rigidity-Aware 3D Gaussian Deformation from a Single Image
von: Kim, Jinhyeok, et al.
Veröffentlicht: (2025)
von: Kim, Jinhyeok, et al.
Veröffentlicht: (2025)
WonderZoom: Multi-Scale 3D World Generation
von: Cao, Jin, et al.
Veröffentlicht: (2025)
von: Cao, Jin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ePBR: Extended PBR Materials in Image Synthesis
von: Guo, Yu, et al.
Veröffentlicht: (2025) -
Towards Understanding Graphical Perception in Large Multimodal Models
von: Zhang, Kai, et al.
Veröffentlicht: (2025) -
Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026) -
Proc3D: Procedural 3D Generation and Parametric Editing of 3D Shapes with Large Language Models
von: Raji, Fadlullah, et al.
Veröffentlicht: (2026) -
Vision-as-Inverse-Graphics Agent via Interleaved Multimodal Reasoning
von: Yin, Shaofeng, et al.
Veröffentlicht: (2026)