EscherNet++: Simultaneous Amodal Completion and Scalable View Synthesis through Masked Fine-Tuning and Enhanced Feed-Forward 3D Reconstruction
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Xinan, Irshad, Muhammad Zubair, Yezzi, Anthony, Tsai, Yi-Chang, Kira, Zsolt |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EscherNet: A Generative Model for Scalable View Synthesis
by: Kong, Xin, et al.
Published: (2024)
by: Kong, Xin, et al.
Published: (2024)
EmbodiedSplat: Personalized Real-to-Sim-to-Real Navigation with Gaussian Splats from a Mobile Device
by: Chhablani, Gunjan, et al.
Published: (2025)
by: Chhablani, Gunjan, et al.
Published: (2025)
NeRF-MAE: Masked AutoEncoders for Self-Supervised 3D Representation Learning for Neural Radiance Fields
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
Deep Learning for Crack Detection: A Review of Learning Paradigms, Generalizability, and Datasets
by: Zhang, Xinan, et al.
Published: (2025)
by: Zhang, Xinan, et al.
Published: (2025)
EVE: A Generator-Verifier System for Generative Policies
by: Ali, Yusuf, et al.
Published: (2025)
by: Ali, Yusuf, et al.
Published: (2025)
Learning 3D Robotics Perception using Inductive Priors
by: Irshad, Muhammad Zubair
Published: (2024)
by: Irshad, Muhammad Zubair
Published: (2024)
Zero-Shot Novel View and Depth Synthesis with Multi-View Geometric Diffusion
by: Guizilini, Vitor, et al.
Published: (2025)
by: Guizilini, Vitor, et al.
Published: (2025)
Rethinking Weight Decay for Robust Fine-Tuning of Foundation Models
by: Tian, Junjiao, et al.
Published: (2024)
by: Tian, Junjiao, et al.
Published: (2024)
ICE-G: Image Conditional Editing of 3D Gaussian Splats
by: Jaganathan, Vishnu, et al.
Published: (2024)
by: Jaganathan, Vishnu, et al.
Published: (2024)
Mask Guided Gated Convolution for Amodal Content Completion
by: Saleh, Kaziwa, et al.
Published: (2024)
by: Saleh, Kaziwa, et al.
Published: (2024)
Multi-Agent Amodal Completion: Direct Synthesis with Fine-Grained Semantic Guidance
by: Fan, Hongxing, et al.
Published: (2025)
by: Fan, Hongxing, et al.
Published: (2025)
Advances in Feed-Forward 3D Reconstruction and View Synthesis: A Survey
by: Zhang, Jiahui, et al.
Published: (2025)
by: Zhang, Jiahui, et al.
Published: (2025)
Advances in Feed‐Forward 3D Reconstruction and View Synthesis: A Survey
by: Jiahui Zhang, et al.
Published: (2026)
by: Jiahui Zhang, et al.
Published: (2026)
Formulating Event-based Image Reconstruction as a Linear Inverse Problem with Deep Regularization using Optical Flow
by: Zhang, Zelin, et al.
Published: (2021)
by: Zhang, Zelin, et al.
Published: (2021)
PHAC: Promptable Human Amodal Completion
by: Noh, Seung Young, et al.
Published: (2026)
by: Noh, Seung Young, et al.
Published: (2026)
Directional Gradient Projection for Robust Fine-Tuning of Foundation Models
by: Huang, Chengyue, et al.
Published: (2025)
by: Huang, Chengyue, et al.
Published: (2025)
FRAMES-VQA: Benchmarking Fine-Tuning Robustness across Multi-Modal Shifts in Visual Question Answering
by: Huang, Chengyue, et al.
Published: (2025)
by: Huang, Chengyue, et al.
Published: (2025)
ViewSplat: View-Adaptive Dynamic Gaussian Splatting for Feed-Forward Synthesis
by: Jeong, Moonyeon, et al.
Published: (2026)
by: Jeong, Moonyeon, et al.
Published: (2026)
Hyper-Transformer for Amodal Completion
by: Gao, Jianxiong, et al.
Published: (2024)
by: Gao, Jianxiong, et al.
Published: (2024)
Neural Fields in Robotics: A Survey
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
Training for X-Ray Vision: Amodal Segmentation, Amodal Content Completion, and View-Invariant Object Representation from Multi-Camera Video
by: Moore, Alexander, et al.
Published: (2025)
by: Moore, Alexander, et al.
Published: (2025)
From Rays to Projections: Better Inputs for Feed-Forward View Synthesis
by: Wu, Zirui, et al.
Published: (2026)
by: Wu, Zirui, et al.
Published: (2026)
Open-World Amodal Appearance Completion
by: Ao, Jiayang, et al.
Published: (2024)
by: Ao, Jiayang, et al.
Published: (2024)
Amodal Ground Truth and Completion in the Wild
by: Zhan, Guanqi, et al.
Published: (2023)
by: Zhan, Guanqi, et al.
Published: (2023)
Long-LRM++: Preserving Fine Details in Feed-Forward Wide-Coverage Reconstruction
by: Ziwen, Chen, et al.
Published: (2025)
by: Ziwen, Chen, et al.
Published: (2025)
RIS-Aided Wireless Amodal Sensing for Single-View 3D Reconstruction
by: Wang, Yuhan, et al.
Published: (2026)
by: Wang, Yuhan, et al.
Published: (2026)
Continual Diffusion with STAMINA: STack-And-Mask INcremental Adapters
by: Smith, James Seale, et al.
Published: (2023)
by: Smith, James Seale, et al.
Published: (2023)
MVSplat360: Feed-Forward 360 Scene Synthesis from Sparse Views
by: Chen, Yuedong, et al.
Published: (2024)
by: Chen, Yuedong, et al.
Published: (2024)
Escher's metaphors
Published: (1994)
Published: (1994)
RoboDream: Compositional World Models for Scalable Robot Data Synthesis
by: Ye, Junjie, et al.
Published: (2026)
by: Ye, Junjie, et al.
Published: (2026)
Occlusion-Aware Temporally Consistent Amodal Completion for 3D Human-Object Interaction Reconstruction
by: Doh, Hyungjun, et al.
Published: (2025)
by: Doh, Hyungjun, et al.
Published: (2025)
Generative Escher Meshes
by: Aigerman, Noam, et al.
Published: (2023)
by: Aigerman, Noam, et al.
Published: (2023)
TACO: Taming Diffusion for in-the-wild Video Amodal Completion
by: Lu, Ruijie, et al.
Published: (2025)
by: Lu, Ruijie, et al.
Published: (2025)
VolFill: Single-View Amodal 3D Scene Reconstruction with Volumetric Flow Matching
by: Ngo, Tuan Duc, et al.
Published: (2026)
by: Ngo, Tuan Duc, et al.
Published: (2026)
Amodal3R: Amodal 3D Reconstruction from Occluded 2D Images
by: Wu, Tianhao, et al.
Published: (2025)
by: Wu, Tianhao, et al.
Published: (2025)
Feed-Forward SceneDINO for Unsupervised Semantic Scene Completion
by: Jevtić, Aleksandar, et al.
Published: (2025)
by: Jevtić, Aleksandar, et al.
Published: (2025)
A PDE-based Explanation of Extreme Numerical Sensitivities and Edge of Stability in Training Neural Networks
by: Sun, Yuxin, et al.
Published: (2022)
by: Sun, Yuxin, et al.
Published: (2022)
Feed-Forward Gaussian Splatting from Sparse Aerial Views
by: Wu, Dongli, et al.
Published: (2026)
by: Wu, Dongli, et al.
Published: (2026)
SAGE: Sink-Aware Grounded Decoding for Multimodal Hallucination Mitigation
by: Shukla, Tripti, et al.
Published: (2026)
by: Shukla, Tripti, et al.
Published: (2026)
Integrating Multimodal Large Language Model Knowledge into Amodal Completion
by: Yun, Heecheol, et al.
Published: (2026)
by: Yun, Heecheol, et al.
Published: (2026)
Similar Items
-
EscherNet: A Generative Model for Scalable View Synthesis
by: Kong, Xin, et al.
Published: (2024) -
EmbodiedSplat: Personalized Real-to-Sim-to-Real Navigation with Gaussian Splats from a Mobile Device
by: Chhablani, Gunjan, et al.
Published: (2025) -
NeRF-MAE: Masked AutoEncoders for Self-Supervised 3D Representation Learning for Neural Radiance Fields
by: Irshad, Muhammad Zubair, et al.
Published: (2024) -
Deep Learning for Crack Detection: A Review of Learning Paradigms, Generalizability, and Datasets
by: Zhang, Xinan, et al.
Published: (2025) -
EVE: A Generator-Verifier System for Generative Policies
by: Ali, Yusuf, et al.
Published: (2025)