Point3R: Streaming 3D Reconstruction with Explicit Spatial Pointer Memory
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Yuqi, Zheng, Wenzhao, Zhou, Jie, Lu, Jiwen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Streaming 4D Visual Geometry Transformer
von: Zhuo, Dong, et al.
Veröffentlicht: (2025)
von: Zhuo, Dong, et al.
Veröffentlicht: (2025)
GaussianWorld: Gaussian World Model for Streaming 3D Occupancy Prediction
von: Zuo, Sicheng, et al.
Veröffentlicht: (2024)
von: Zuo, Sicheng, et al.
Veröffentlicht: (2024)
EmbodiedOcc: Embodied 3D Occupancy Prediction for Vision-based Online Scene Understanding
von: Wu, Yuqi, et al.
Veröffentlicht: (2024)
von: Wu, Yuqi, et al.
Veröffentlicht: (2024)
Terra: Explorable Native 3D World Model with Point Latents
von: Huang, Yuanhui, et al.
Veröffentlicht: (2025)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2025)
Driv3R: Learning Dense 4D Reconstruction for Autonomous Driving
von: Fei, Xin, et al.
Veröffentlicht: (2024)
von: Fei, Xin, et al.
Veröffentlicht: (2024)
Hardness-Aware Scene Synthesis for Semi-Supervised 3D Object Detection
von: Zeng, Shuai, et al.
Veröffentlicht: (2024)
von: Zeng, Shuai, et al.
Veröffentlicht: (2024)
PixelGaussian: Generalizable 3D Gaussian Reconstruction from Arbitrary Views
von: Fei, Xin, et al.
Veröffentlicht: (2024)
von: Fei, Xin, et al.
Veröffentlicht: (2024)
GaussianFormer-2: Probabilistic Gaussian Superposition for Efficient 3D Occupancy Prediction
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
SpectralAR: Spectral Autoregressive Visual Generation
von: Huang, Yuanhui, et al.
Veröffentlicht: (2025)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2025)
Doe-1: Closed-Loop Autonomous Driving with Large World Model
von: Zheng, Wenzhao, et al.
Veröffentlicht: (2024)
von: Zheng, Wenzhao, et al.
Veröffentlicht: (2024)
GaussianFormer: Scene as Gaussians for Vision-Based 3D Semantic Occupancy Prediction
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
DriveTok: 3D Driving Scene Tokenization for Unified Multi-View Reconstruction and Understanding
von: Zhuo, Dong, et al.
Veröffentlicht: (2026)
von: Zhuo, Dong, et al.
Veröffentlicht: (2026)
Astra: General Interactive World Model with Autoregressive Denoising
von: Zhu, Yixuan, et al.
Veröffentlicht: (2025)
von: Zhu, Yixuan, et al.
Veröffentlicht: (2025)
Owl-1: Omni World Model for Consistent Long Video Generation
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
von: Huang, Yuanhui, et al.
Veröffentlicht: (2024)
FILT3R: Latent State Adaptive Kalman Filter for Streaming 3D Reconstruction
von: Jin, Seonghyun, et al.
Veröffentlicht: (2026)
von: Jin, Seonghyun, et al.
Veröffentlicht: (2026)
Stag-1: Towards Realistic 4D Driving Simulation with Video Generation Model
von: Wang, Lening, et al.
Veröffentlicht: (2024)
von: Wang, Lening, et al.
Veröffentlicht: (2024)
GPD-1: Generative Pre-training for Driving
von: Xie, Zixun, et al.
Veröffentlicht: (2024)
von: Xie, Zixun, et al.
Veröffentlicht: (2024)
Path Choice Matters for Clear Attribution in Path Methods
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
Preventing Local Pitfalls in Vector Quantization via Optimal Transport
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
von: Zhang, Borui, et al.
Veröffentlicht: (2024)
UniPre3D: Unified Pre-training of 3D Point Cloud Models with Cross-Modal Gaussian Splatting
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
$\bf{D^3}$QE: Learning Discrete Distribution Discrepancy-aware Quantization Error for Autoregressive-Generated Image Detection
von: Zhang, Yanran, et al.
Veröffentlicht: (2025)
von: Zhang, Yanran, et al.
Veröffentlicht: (2025)
SFTok: Bridging the Performance Gap in Discrete Tokenizers
von: Rao, Qihang, et al.
Veröffentlicht: (2025)
von: Rao, Qihang, et al.
Veröffentlicht: (2025)
Quantize-then-Rectify: Efficient VQ-VAE Training
von: Zhang, Borui, et al.
Veröffentlicht: (2025)
von: Zhang, Borui, et al.
Veröffentlicht: (2025)
Fast Shapley Value Estimation: A Unified Approach
von: Zhang, Borui, et al.
Veröffentlicht: (2023)
von: Zhang, Borui, et al.
Veröffentlicht: (2023)
Learning with Noisy Ground Truth: From 2D Classification to 3D Reconstruction
von: Lu, Yangdi, et al.
Veröffentlicht: (2024)
von: Lu, Yangdi, et al.
Veröffentlicht: (2024)
Scalable Object Relation Encoding for Better 3D Spatial Reasoning in Large Language Models
von: Zhou, Shengli, et al.
Veröffentlicht: (2026)
von: Zhou, Shengli, et al.
Veröffentlicht: (2026)
Multi-hypotheses Conditioned Point Cloud Diffusion for 3D Human Reconstruction from Occluded Images
von: Kim, Donghwan, et al.
Veröffentlicht: (2024)
von: Kim, Donghwan, et al.
Veröffentlicht: (2024)
Enhancing 3D Point Cloud Classification with ModelNet-R and Point-SkipNet
von: Saeid, Mohammad, et al.
Veröffentlicht: (2025)
von: Saeid, Mohammad, et al.
Veröffentlicht: (2025)
Evaluating Modern Approaches in 3D Scene Reconstruction: NeRF vs Gaussian-Based Methods
von: Zhou, Yiming, et al.
Veröffentlicht: (2024)
von: Zhou, Yiming, et al.
Veröffentlicht: (2024)
Vega: Learning to Drive with Natural Language Instructions
von: Zuo, Sicheng, et al.
Veröffentlicht: (2026)
von: Zuo, Sicheng, et al.
Veröffentlicht: (2026)
NeuSDFusion: A Spatial-Aware Generative Model for 3D Shape Completion, Reconstruction, and Generation
von: Cui, Ruikai, et al.
Veröffentlicht: (2024)
von: Cui, Ruikai, et al.
Veröffentlicht: (2024)
PointNSP: Autoregressive 3D Point Cloud Generation with Next-Scale Level-of-Detail Prediction
von: Meng, Ziqiao, et al.
Veröffentlicht: (2025)
von: Meng, Ziqiao, et al.
Veröffentlicht: (2025)
TAPVid-3D: A Benchmark for Tracking Any Point in 3D
von: Koppula, Skanda, et al.
Veröffentlicht: (2024)
von: Koppula, Skanda, et al.
Veröffentlicht: (2024)
3D Reconstruction of Objects in Hands without Real World 3D Supervision
von: Prakash, Aditya, et al.
Veröffentlicht: (2023)
von: Prakash, Aditya, et al.
Veröffentlicht: (2023)
XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation
von: Wang, Ziyi, et al.
Veröffentlicht: (2024)
von: Wang, Ziyi, et al.
Veröffentlicht: (2024)
IM-3D: Iterative Multiview Diffusion and Reconstruction for High-Quality 3D Generation
von: Melas-Kyriazi, Luke, et al.
Veröffentlicht: (2024)
von: Melas-Kyriazi, Luke, et al.
Veröffentlicht: (2024)
RayletDF: Raylet Distance Fields for Generalizable 3D Surface Reconstruction from Point Clouds or Gaussians
von: Wei, Shenxing, et al.
Veröffentlicht: (2025)
von: Wei, Shenxing, et al.
Veröffentlicht: (2025)
Measuring 3D Spatial Geometric Consistency in Dynamic Generated Videos
von: Dou, Weijia, et al.
Veröffentlicht: (2026)
von: Dou, Weijia, et al.
Veröffentlicht: (2026)
GaussianToken: An Effective Image Tokenizer with 2D Gaussian Splatting
von: Dong, Jiajun, et al.
Veröffentlicht: (2025)
von: Dong, Jiajun, et al.
Veröffentlicht: (2025)
Geometric Point Attention Transformer for 3D Shape Reassembly
von: Li, Jiahan, et al.
Veröffentlicht: (2024)
von: Li, Jiahan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Streaming 4D Visual Geometry Transformer
von: Zhuo, Dong, et al.
Veröffentlicht: (2025) -
GaussianWorld: Gaussian World Model for Streaming 3D Occupancy Prediction
von: Zuo, Sicheng, et al.
Veröffentlicht: (2024) -
EmbodiedOcc: Embodied 3D Occupancy Prediction for Vision-based Online Scene Understanding
von: Wu, Yuqi, et al.
Veröffentlicht: (2024) -
Terra: Explorable Native 3D World Model with Point Latents
von: Huang, Yuanhui, et al.
Veröffentlicht: (2025) -
Driv3R: Learning Dense 4D Reconstruction for Autonomous Driving
von: Fei, Xin, et al.
Veröffentlicht: (2024)