SEAR: Simple and Efficient Adaptation of Visual Geometric Transformers for RGB+Thermal 3D Reconstruction
Fuente:
arXiv
Saved in:
| Main Authors: | Skorokhodov, Vsevolod, Xu, Chenghao, Sun, Shuo, Fink, Olga, Mielle, Malcolm |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ThermoNeRF: Joint RGB and Thermal Novel View Synthesis for Building Facades using Multimodal Neural Radiance Fields
by: Hassan, Mariam, et al.
Published: (2024)
by: Hassan, Mariam, et al.
Published: (2024)
Exploiting Semantic Scene Reconstruction for Estimating Building Envelope Characteristics
by: Xu, Chenghao, et al.
Published: (2024)
by: Xu, Chenghao, et al.
Published: (2024)
Thermoxels: a voxel-based method to generate simulation-ready 3D thermal models
by: Chassaing, Etienne, et al.
Published: (2025)
by: Chassaing, Etienne, et al.
Published: (2025)
Dense Dynamic Scene Reconstruction and Camera Pose Estimation from Multi-View Videos
by: Sun, Shuo, et al.
Published: (2026)
by: Sun, Shuo, et al.
Published: (2026)
Diffusion Models are Secretly Zero-Shot 3DGS Harmonizers
by: Skorokhodov, Vsevolod, et al.
Published: (2025)
by: Skorokhodov, Vsevolod, et al.
Published: (2025)
High-Fidelity SLAM Using Gaussian Splatting with Rendering-Guided Densification and Regularized Optimization
by: Sun, Shuo, et al.
Published: (2024)
by: Sun, Shuo, et al.
Published: (2024)
Large-scale visual SLAM for in-the-wild videos
by: Sun, Shuo, et al.
Published: (2025)
by: Sun, Shuo, et al.
Published: (2025)
uSF: Learning Neural Semantic Field with Uncertainty
by: Skorokhodov, Vsevolod, et al.
Published: (2023)
by: Skorokhodov, Vsevolod, et al.
Published: (2023)
Geometric Context Transformer for Streaming 3D Reconstruction
by: Chen, Lin-Zhuo, et al.
Published: (2026)
by: Chen, Lin-Zhuo, et al.
Published: (2026)
Recall and Refine: A Simple but Effective Source-free Open-set Domain Adaptation Framework
by: Nejjar, Ismail, et al.
Published: (2024)
by: Nejjar, Ismail, et al.
Published: (2024)
Physics-Informed Neural Networks for Thermophysical Property Retrieval
by: Waseem, Ali, et al.
Published: (2025)
by: Waseem, Ali, et al.
Published: (2025)
UAC: Uncertainty-Aware Calibration of Neural Networks for Gesture Detection
by: Haddad, Farida Al, et al.
Published: (2025)
by: Haddad, Farida Al, et al.
Published: (2025)
Efficient Prediction of Dense Visual Embeddings via Distillation and RGB-D Transformers
by: Fischedick, Söhnke Benedikt, et al.
Published: (2026)
by: Fischedick, Söhnke Benedikt, et al.
Published: (2026)
Unseen Visual Anomaly Generation
by: Sun, Han, et al.
Published: (2024)
by: Sun, Han, et al.
Published: (2024)
AC3D: Analyzing and Improving 3D Camera Control in Video Diffusion Transformers
by: Bahmani, Sherwin, et al.
Published: (2024)
by: Bahmani, Sherwin, et al.
Published: (2024)
CAST: Component-Aligned 3D Scene Reconstruction from an RGB Image
by: Yao, Kaixin, et al.
Published: (2025)
by: Yao, Kaixin, et al.
Published: (2025)
EGP3D: Edge-guided Geometric Preserving 3D Point Cloud Super-resolution for RGB-D camera
by: Fang, Zheng, et al.
Published: (2024)
by: Fang, Zheng, et al.
Published: (2024)
ViGG: Robust RGB-D Point Cloud Registration using Visual-Geometric Mutual Guidance
by: Chen, Congjia, et al.
Published: (2025)
by: Chen, Congjia, et al.
Published: (2025)
Efficient Multi-Task Scene Analysis with RGB-D Transformers
by: Fischedick, Söhnke Benedikt, et al.
Published: (2023)
by: Fischedick, Söhnke Benedikt, et al.
Published: (2023)
SimC3D: A Simple Contrastive 3D Pretraining Framework Using RGB Images
by: Dong, Jiahua, et al.
Published: (2024)
by: Dong, Jiahua, et al.
Published: (2024)
APLA: A Simple Adaptation Method for Vision Transformers
by: Sorkhei, Moein, et al.
Published: (2025)
by: Sorkhei, Moein, et al.
Published: (2025)
TherA: Thermal-Aware Visual-Language Prompting for Controllable RGB-to-Thermal Infrared Translation
by: Lee, Dong-Guw, et al.
Published: (2026)
by: Lee, Dong-Guw, et al.
Published: (2026)
4DRecons: 4D Neural Implicit Deformable Objects Reconstruction from a single RGB-D Camera with Geometrical and Topological Regularizations
by: Cong, Xiaoyan, et al.
Published: (2024)
by: Cong, Xiaoyan, et al.
Published: (2024)
EA-ViT: Efficient Adaptation for Elastic Vision Transformer
by: Zhu, Chen, et al.
Published: (2025)
by: Zhu, Chen, et al.
Published: (2025)
CloudFixer: Test-Time Adaptation for 3D Point Clouds via Diffusion-Guided Geometric Transformation
by: Shim, Hajin, et al.
Published: (2024)
by: Shim, Hajin, et al.
Published: (2024)
Ov3R: Open-Vocabulary Semantic 3D Reconstruction from RGB Videos
by: Gong, Ziren, et al.
Published: (2025)
by: Gong, Ziren, et al.
Published: (2025)
Self-supervised Depth Denoising Using Lower- and Higher-quality RGB-D sensors
by: Shabanov, Akhmedkhan, et al.
Published: (2020)
by: Shabanov, Akhmedkhan, et al.
Published: (2020)
Uncertainty-Guided Alignment for Unsupervised Domain Adaptation in Regression
by: Nejjar, Ismail, et al.
Published: (2024)
by: Nejjar, Ismail, et al.
Published: (2024)
A Unified Structure for Efficient RGB and RGB-D Salient Object Detection
by: Peng, Peng, et al.
Published: (2020)
by: Peng, Peng, et al.
Published: (2020)
Visual Object Tracking on Multi-modal RGB-D Videos: A Review
by: Zhu, Xue-Feng, et al.
Published: (2022)
by: Zhu, Xue-Feng, et al.
Published: (2022)
NOVA3R: Non-pixel-aligned Visual Transformer for Amodal 3D Reconstruction
by: Chen, Weirong, et al.
Published: (2026)
by: Chen, Weirong, et al.
Published: (2026)
Geometric Understanding of Discriminability and Transferability for Visual Domain Adaptation
by: Luo, You-Wei, et al.
Published: (2024)
by: Luo, You-Wei, et al.
Published: (2024)
VD3D: Taming Large Video Diffusion Transformers for 3D Camera Control
by: Bahmani, Sherwin, et al.
Published: (2024)
by: Bahmani, Sherwin, et al.
Published: (2024)
Cross3DVG: Cross-Dataset 3D Visual Grounding on Different RGB-D Scans
by: Miyanishi, Taiki, et al.
Published: (2023)
by: Miyanishi, Taiki, et al.
Published: (2023)
Fus3D: Decoding Consolidated 3D Geometry from Feed-forward Geometry Transformer Latents
by: Fink, Laura, et al.
Published: (2026)
by: Fink, Laura, et al.
Published: (2026)
A Simple Baseline for Efficient Hand Mesh Reconstruction
by: Zhou, Zhishan, et al.
Published: (2024)
by: Zhou, Zhishan, et al.
Published: (2024)
GauSSmart: Enhanced 3D Reconstruction through 2D Foundation Models and Geometric Filtering
by: Valverde, Alexander, et al.
Published: (2025)
by: Valverde, Alexander, et al.
Published: (2025)
FlashSLAM: Accelerated RGB-D SLAM for Real-Time 3D Scene Reconstruction with Gaussian Splatting
by: Pham, Phu, et al.
Published: (2024)
by: Pham, Phu, et al.
Published: (2024)
Towards Multimodal Open-Set Domain Generalization and Adaptation through Self-supervision
by: Dong, Hao, et al.
Published: (2024)
by: Dong, Hao, et al.
Published: (2024)
RemixFusion: Residual-based Mixed Representation for Large-scale Online RGB-D Reconstruction
by: Lan, Yuqing, et al.
Published: (2025)
by: Lan, Yuqing, et al.
Published: (2025)
Similar Items
-
ThermoNeRF: Joint RGB and Thermal Novel View Synthesis for Building Facades using Multimodal Neural Radiance Fields
by: Hassan, Mariam, et al.
Published: (2024) -
Exploiting Semantic Scene Reconstruction for Estimating Building Envelope Characteristics
by: Xu, Chenghao, et al.
Published: (2024) -
Thermoxels: a voxel-based method to generate simulation-ready 3D thermal models
by: Chassaing, Etienne, et al.
Published: (2025) -
Dense Dynamic Scene Reconstruction and Camera Pose Estimation from Multi-View Videos
by: Sun, Shuo, et al.
Published: (2026) -
Diffusion Models are Secretly Zero-Shot 3DGS Harmonizers
by: Skorokhodov, Vsevolod, et al.
Published: (2025)