Saved in:
| Main Authors: | Nie, Chang, Wang, Guangming, Lie, Zhe, Wang, Hesheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2507.17462 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MovSAM: A Single-image Moving Object Segmentation Framework Based on Deep Thinking
by: Nie, Chang, et al.
Published: (2025)
by: Nie, Chang, et al.
Published: (2025)
Towards the Vision-Sound-Language-Action Paradigm: The HEAR Framework for Sound-Centric Manipulation
by: Nie, Chang, et al.
Published: (2026)
by: Nie, Chang, et al.
Published: (2026)
MID: A Self-supervised Multimodal Iterative Denoising Framework
by: Nie, Chang, et al.
Published: (2025)
by: Nie, Chang, et al.
Published: (2025)
MRASfM: Multi-Camera Reconstruction and Aggregation through Structure-from-Motion in Driving Scenes
by: Xuan, Lingfeng, et al.
Published: (2025)
by: Xuan, Lingfeng, et al.
Published: (2025)
End-to-end 2D-3D Registration between Image and LiDAR Point Cloud for Vehicle Localization
by: Wang, Guangming, et al.
Published: (2023)
by: Wang, Guangming, et al.
Published: (2023)
Robot Learning from Human Videos: A Survey
by: Ma, Junyi, et al.
Published: (2026)
by: Ma, Junyi, et al.
Published: (2026)
WeatherCity: Urban Scene Reconstruction with Controllable Multi-Weather Transformation
by: Wu, Wenhua, et al.
Published: (2026)
by: Wu, Wenhua, et al.
Published: (2026)
DifFlow3D: Toward Robust Uncertainty-Aware Scene Flow Estimation with Diffusion Model
by: Liu, Jiuming, et al.
Published: (2023)
by: Liu, Jiuming, et al.
Published: (2023)
Spherical Frustum Sparse Convolution Network for LiDAR Point Cloud Semantic Segmentation
by: Zheng, Yu, et al.
Published: (2023)
by: Zheng, Yu, et al.
Published: (2023)
SemGauss-SLAM: Dense Semantic Gaussian Splatting SLAM
by: Zhu, Siting, et al.
Published: (2024)
by: Zhu, Siting, et al.
Published: (2024)
RegFormer++: An Efficient Large-Scale 3D LiDAR Point Registration Network with Projection-Aware 2D Transformer
by: Liu, Jiuming, et al.
Published: (2026)
by: Liu, Jiuming, et al.
Published: (2026)
Fast Multi-view Consistent 3D Editing with Video Priors
by: Chen, Liyi, et al.
Published: (2025)
by: Chen, Liyi, et al.
Published: (2025)
EMIE-MAP: Large-Scale Road Surface Reconstruction Based on Explicit Mesh and Implicit Encoding
by: Wu, Wenhua, et al.
Published: (2024)
by: Wu, Wenhua, et al.
Published: (2024)
SemAlign3D: Semantic Correspondence between RGB-Images through Aligning 3D Object-Class Representations
by: Wandel, Krispin, et al.
Published: (2025)
by: Wandel, Krispin, et al.
Published: (2025)
DIAL-GS: Dynamic Instance Aware Reconstruction for Label-free Street Scenes with 4D Gaussian Splatting
by: Su, Chenpeng, et al.
Published: (2025)
by: Su, Chenpeng, et al.
Published: (2025)
Geometry-Guided Reinforcement Learning for Multi-view Consistent 3D Scene Editing
by: Wang, Jiyuan, et al.
Published: (2026)
by: Wang, Jiyuan, et al.
Published: (2026)
DGE: Direct Gaussian 3D Editing by Consistent Multi-view Editing
by: Chen, Minghao, et al.
Published: (2024)
by: Chen, Minghao, et al.
Published: (2024)
CAD-SLAM: Consistency-Aware Dynamic SLAM with Dynamic-Static Decoupled Mapping
by: Wu, Wenhua, et al.
Published: (2025)
by: Wu, Wenhua, et al.
Published: (2025)
NeRFs in Robotics: A Survey
by: Wang, Guangming, et al.
Published: (2024)
by: Wang, Guangming, et al.
Published: (2024)
RESBev: Making BEV Perception More Robust
by: Zhuo, Lifeng, et al.
Published: (2026)
by: Zhuo, Lifeng, et al.
Published: (2026)
3DSFLabelling: Boosting 3D Scene Flow Estimation by Pseudo Auto-labelling
by: Jiang, Chaokang, et al.
Published: (2024)
by: Jiang, Chaokang, et al.
Published: (2024)
Fusion4CA: Boosting 3D Object Detection via Comprehensive Image Exploitation
by: Luo, Kang, et al.
Published: (2026)
by: Luo, Kang, et al.
Published: (2026)
Reloc-VGGT: Visual Re-localization with Geometry Grounded Transformer
by: Deng, Tianchen, et al.
Published: (2025)
by: Deng, Tianchen, et al.
Published: (2025)
Compact 3D Gaussian Splatting For Dense Visual SLAM
by: Deng, Tianchen, et al.
Published: (2024)
by: Deng, Tianchen, et al.
Published: (2024)
4Diffusion: Multi-view Video Diffusion Model for 4D Generation
by: Zhang, Haiyu, et al.
Published: (2024)
by: Zhang, Haiyu, et al.
Published: (2024)
MAMBA4D: Efficient Long-Sequence Point Cloud Video Understanding with Disentangled Spatial-Temporal State Space Models
by: Liu, Jiuming, et al.
Published: (2024)
by: Liu, Jiuming, et al.
Published: (2024)
PG-SLAM: Photo-realistic and Geometry-aware RGB-D SLAM in Dynamic Environments
by: Li, Haoang, et al.
Published: (2024)
by: Li, Haoang, et al.
Published: (2024)
DVN-SLAM: Dynamic Visual Neural SLAM Based on Local-Global Encoding
by: Wu, Wenhua, et al.
Published: (2024)
by: Wu, Wenhua, et al.
Published: (2024)
OPEN: Object-wise Position Embedding for Multi-view 3D Object Detection
by: Hou, Jinghua, et al.
Published: (2024)
by: Hou, Jinghua, et al.
Published: (2024)
EditInfinity: Image Editing with Binary-Quantized Generative Models
by: Wang, Jiahuan, et al.
Published: (2025)
by: Wang, Jiahuan, et al.
Published: (2025)
Multi-view Image Prompted Multi-view Diffusion for Improved 3D Generation
by: Kim, Seungwook, et al.
Published: (2024)
by: Kim, Seungwook, et al.
Published: (2024)
Planning from Imagination: Episodic Simulation and Episodic Memory for Vision-and-Language Navigation
by: Pan, Yiyuan, et al.
Published: (2024)
by: Pan, Yiyuan, et al.
Published: (2024)
SAR image segmentation algorithms based on I-divergence-TV model
by: Liu, Guangming
Published: (2023)
by: Liu, Guangming
Published: (2023)
Diff-IP2D: Diffusion-Based Hand-Object Interaction Prediction on Egocentric Videos
by: Ma, Junyi, et al.
Published: (2024)
by: Ma, Junyi, et al.
Published: (2024)
DSLO: Deep Sequence LiDAR Odometry Based on Inconsistent Spatio-temporal Propagation
by: Zhang, Huixin, et al.
Published: (2024)
by: Zhang, Huixin, et al.
Published: (2024)
SCAFusion: A Multimodal 3D Detection Framework for Small Object Detection in Lunar Surface Exploration
by: Chen, Xin, et al.
Published: (2025)
by: Chen, Xin, et al.
Published: (2025)
FLAME: Learning to Navigate with Multimodal LLM in Urban Environments
by: Xu, Yunzhe, et al.
Published: (2024)
by: Xu, Yunzhe, et al.
Published: (2024)
A locally statistical active contour model for SAR image segmentation can be solved by denoising algorithms
by: Liu, Guangming
Published: (2024)
by: Liu, Guangming
Published: (2024)
A global optimization SAR image segmentation model can be easily transformed to a general ROF denoising model
by: Liu, Guangming
Published: (2023)
by: Liu, Guangming
Published: (2023)
Active contours driven by local and global intensity fitting energy with application to SAR image segmentation and its fast solvers
by: Liu, Guangming
Published: (2023)
by: Liu, Guangming
Published: (2023)
Similar Items
-
MovSAM: A Single-image Moving Object Segmentation Framework Based on Deep Thinking
by: Nie, Chang, et al.
Published: (2025) -
Towards the Vision-Sound-Language-Action Paradigm: The HEAR Framework for Sound-Centric Manipulation
by: Nie, Chang, et al.
Published: (2026) -
MID: A Self-supervised Multimodal Iterative Denoising Framework
by: Nie, Chang, et al.
Published: (2025) -
MRASfM: Multi-Camera Reconstruction and Aggregation through Structure-from-Motion in Driving Scenes
by: Xuan, Lingfeng, et al.
Published: (2025) -
End-to-end 2D-3D Registration between Image and LiDAR Point Cloud for Vehicle Localization
by: Wang, Guangming, et al.
Published: (2023)