Pseudo Depth Meets Gaussian: A Feed-forward RGB SLAM Baseline
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Linqing, Xu, Xiuwei, Wang, Yirui, Wang, Hao, Zheng, Wenzhao, Tang, Yansong, Yan, Haibin, Lu, Jiwen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
iGaussian: Real-Time Camera Pose Estimation via Feed-Forward 3D Gaussian Splatting Inversion
von: Wang, Hao, et al.
Veröffentlicht: (2025)
von: Wang, Hao, et al.
Veröffentlicht: (2025)
Anyview: Generalizable Indoor 3D Object Detection with Variable Frames
von: Wu, Zhenyu, et al.
Veröffentlicht: (2023)
von: Wu, Zhenyu, et al.
Veröffentlicht: (2023)
GaussianToken: An Effective Image Tokenizer with 2D Gaussian Splatting
von: Dong, Jiajun, et al.
Veröffentlicht: (2025)
von: Dong, Jiajun, et al.
Veröffentlicht: (2025)
EmbodiedSAM: Online Segment Any 3D Thing in Real Time
von: Xu, Xiuwei, et al.
Veröffentlicht: (2024)
von: Xu, Xiuwei, et al.
Veröffentlicht: (2024)
Towards Accurate Post-training Quantization for Diffusion Models
von: Wang, Changyuan, et al.
Veröffentlicht: (2023)
von: Wang, Changyuan, et al.
Veröffentlicht: (2023)
Q-VLM: Post-training Quantization for Large Vision-Language Models
von: Wang, Changyuan, et al.
Veröffentlicht: (2024)
von: Wang, Changyuan, et al.
Veröffentlicht: (2024)
Hardness-Aware Scene Synthesis for Semi-Supervised 3D Object Detection
von: Zeng, Shuai, et al.
Veröffentlicht: (2024)
von: Zeng, Shuai, et al.
Veröffentlicht: (2024)
Memory-based Adapters for Online 3D Scene Perception
von: Xu, Xiuwei, et al.
Veröffentlicht: (2024)
von: Xu, Xiuwei, et al.
Veröffentlicht: (2024)
OGGSplat: Open Gaussian Growing for Generalizable Reconstruction with Expanded Field-of-View
von: Wang, Yanbo, et al.
Veröffentlicht: (2025)
von: Wang, Yanbo, et al.
Veröffentlicht: (2025)
ManiGaussian: Dynamic Gaussian Splatting for Multi-task Robotic Manipulation
von: Lu, Guanxing, et al.
Veröffentlicht: (2024)
von: Lu, Guanxing, et al.
Veröffentlicht: (2024)
GaussianFormer: Scene as Gaussians for Vision-Based 3D Semantic Occupancy Prediction
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
GaussianWorld: Gaussian World Model for Streaming 3D Occupancy Prediction
von: Zuo, Sicheng, et al.
Veröffentlicht: (2024)
von: Zuo, Sicheng, et al.
Veröffentlicht: (2024)
GauS-SLAM: Dense RGB-D SLAM with Gaussian Surfels
von: Su, Yongxin, et al.
Veröffentlicht: (2025)
von: Su, Yongxin, et al.
Veröffentlicht: (2025)
GlobalMamba: Global Image Serialization for Vision Mamba
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
IGL-Nav: Incremental 3D Gaussian Localization for Image-goal Navigation
von: Guo, Wenxuan, et al.
Veröffentlicht: (2025)
von: Guo, Wenxuan, et al.
Veröffentlicht: (2025)
EfficientLLaVA:Generalizable Auto-Pruning for Large Vision-language Models
von: Liang, Yinan, et al.
Veröffentlicht: (2025)
von: Liang, Yinan, et al.
Veröffentlicht: (2025)
RGB-Only Gaussian Splatting SLAM for Unbounded Outdoor Scenes
von: Yu, Sicheng, et al.
Veröffentlicht: (2025)
von: Yu, Sicheng, et al.
Veröffentlicht: (2025)
GaussianFormer-2: Probabilistic Gaussian Superposition for Efficient 3D Occupancy Prediction
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
UniGoal: Towards Universal Zero-shot Goal-oriented Navigation
von: Yin, Hang, et al.
Veröffentlicht: (2025)
von: Yin, Hang, et al.
Veröffentlicht: (2025)
MonoGS++: Fast and Accurate Monocular RGB Gaussian SLAM
von: Li, Renwu, et al.
Veröffentlicht: (2025)
von: Li, Renwu, et al.
Veröffentlicht: (2025)
Feed-forward Gaussian Registration for Head Avatar Creation and Editing
von: Prinzler, Malte, et al.
Veröffentlicht: (2026)
von: Prinzler, Malte, et al.
Veröffentlicht: (2026)
V2M: Visual 2-Dimensional Mamba for Image Representation Learning
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
von: Wang, Chengkun, et al.
Veröffentlicht: (2024)
GS-SLAM: Dense Visual SLAM with 3D Gaussian Splatting
von: Yan, Chi, et al.
Veröffentlicht: (2023)
von: Yan, Chi, et al.
Veröffentlicht: (2023)
MoManipVLA: Transferring Vision-language-action Models for General Mobile Manipulation
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025)
Path Choice Matters for Clear Attribution in Path Methods
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
Preventing Local Pitfalls in Vector Quantization via Optimal Transport
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
Splat-SLAM: Globally Optimized RGB-only SLAM with 3D Gaussians
von: Sandström, Erik, et al.
Veröffentlicht: (2024)
von: Sandström, Erik, et al.
Veröffentlicht: (2024)
PixelGaussian: Generalizable 3D Gaussian Reconstruction from Arbitrary Views
von: Fei, Xin, et al.
Veröffentlicht: (2024)
von: Fei, Xin, et al.
Veröffentlicht: (2024)
3D Small Object Detection with Dynamic Spatial Pruning
von: Xu, Xiuwei, et al.
Veröffentlicht: (2023)
von: Xu, Xiuwei, et al.
Veröffentlicht: (2023)
GeoLRM: Geometry-Aware Large Reconstruction Model for High-Quality 3D Gaussian Generation
von: Zhang, Chubin, et al.
Veröffentlicht: (2024)
von: Zhang, Chubin, et al.
Veröffentlicht: (2024)
Dropping the D: RGB-D SLAM Without the Depth Sensor
von: Kiray, Mert, et al.
Veröffentlicht: (2025)
von: Kiray, Mert, et al.
Veröffentlicht: (2025)
AnySplat: Feed-forward 3D Gaussian Splatting from Unconstrained Views
von: Jiang, Lihan, et al.
Veröffentlicht: (2025)
von: Jiang, Lihan, et al.
Veröffentlicht: (2025)
Point3R: Streaming 3D Reconstruction with Explicit Spatial Pointer Memory
von: Wu, Yuqi, et al.
Veröffentlicht: (2025)
von: Wu, Yuqi, et al.
Veröffentlicht: (2025)
SG-Nav: Online 3D Scene Graph Prompting for LLM-based Zero-shot Object Navigation
von: Yin, Hang, et al.
Veröffentlicht: (2024)
von: Yin, Hang, et al.
Veröffentlicht: (2024)
Joint 3D Geometry Reconstruction and Motion Generation for 4D Synthesis from a Single Image
von: Zhang, Yanran, et al.
Veröffentlicht: (2025)
von: Zhang, Yanran, et al.
Veröffentlicht: (2025)
AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation
von: Guo, Wenxuan, et al.
Veröffentlicht: (2026)
von: Guo, Wenxuan, et al.
Veröffentlicht: (2026)
TSP3D: Text-guided Sparse Voxel Pruning for Efficient 3D Visual Grounding
von: Guo, Wenxuan, et al.
Veröffentlicht: (2025)
von: Guo, Wenxuan, et al.
Veröffentlicht: (2025)
X-3D: Explicit 3D Structure Modeling for Point Cloud Recognition
von: Sun, Shuofeng, et al.
Veröffentlicht: (2024)
von: Sun, Shuofeng, et al.
Veröffentlicht: (2024)
VGGT-SLAM 2.0: Real-time Dense Feed-forward Scene Reconstruction
von: Maggio, Dominic, et al.
Veröffentlicht: (2026)
von: Maggio, Dominic, et al.
Veröffentlicht: (2026)
ShapeGen: Robotic Data Generation for Category-Level Manipulation
von: Wang, Yirui, et al.
Veröffentlicht: (2026)
von: Wang, Yirui, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
iGaussian: Real-Time Camera Pose Estimation via Feed-Forward 3D Gaussian Splatting Inversion
von: Wang, Hao, et al.
Veröffentlicht: (2025) -
Anyview: Generalizable Indoor 3D Object Detection with Variable Frames
von: Wu, Zhenyu, et al.
Veröffentlicht: (2023) -
GaussianToken: An Effective Image Tokenizer with 2D Gaussian Splatting
von: Dong, Jiajun, et al.
Veröffentlicht: (2025) -
EmbodiedSAM: Online Segment Any 3D Thing in Real Time
von: Xu, Xiuwei, et al.
Veröffentlicht: (2024) -
Towards Accurate Post-training Quantization for Diffusion Models
von: Wang, Changyuan, et al.
Veröffentlicht: (2023)