Scalable Benchmarking and Robust Learning for Noise-Free Ego-Motion and 3D Reconstruction from Noisy Video
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Xiaohao, Zhang, Tianyi, Zhao, Shibo, Li, Xiang, Wang, Sibo, Chen, Yongqi, Li, Ye, Raj, Bhiksha, Johnson-Roberson, Matthew, Scherer, Sebastian, Huang, Xiaonan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Perfect to Noisy World Simulation: Customizable Embodied Multi-modal Perturbations for SLAM Robustness Benchmarking
by: Xu, Xiaohao, et al.
Published: (2024)
by: Xu, Xiaohao, et al.
Published: (2024)
Customizable Perturbation Synthesis for Robust SLAM Benchmarking
by: Xu, Xiaohao, et al.
Published: (2024)
by: Xu, Xiaohao, et al.
Published: (2024)
MAC-Ego3D: Multi-Agent Gaussian Consensus for Real-Time Collaborative Ego-Motion and Photorealistic 3D Reconstruction
by: Xu, Xiaohao, et al.
Published: (2024)
by: Xu, Xiaohao, et al.
Published: (2024)
Towards Ambiguity-Free Spatial Foundation Model: Rethinking and Decoupling Depth Ambiguity
by: Xu, Xiaohao, et al.
Published: (2025)
by: Xu, Xiaohao, et al.
Published: (2025)
Learning Shared RGB-D Fields: Unified Self-supervised Pre-training for Label-efficient LiDAR-Camera 3D Perception
by: Xu, Xiaohao, et al.
Published: (2024)
by: Xu, Xiaohao, et al.
Published: (2024)
$\text{R}^2$-Bench: Benchmarking the Robustness of Referring Perception Models under Perturbations
by: Li, Xiang, et al.
Published: (2024)
by: Li, Xiang, et al.
Published: (2024)
From Single Images to Motion Policies via Video-Generation Environment Representations
by: Zhi, Weiming, et al.
Published: (2025)
by: Zhi, Weiming, et al.
Published: (2025)
QDFormer: Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
Joint Flow Trajectory Optimization For Feasible Robot Motion Generation from Video Demonstrations
by: Dong, Xiaoxiang, et al.
Published: (2025)
by: Dong, Xiaoxiang, et al.
Published: (2025)
Efficient Construction of Implicit Surface Models From a Single Image for Motion Generation
by: Chu, Wei-Teng, et al.
Published: (2025)
by: Chu, Wei-Teng, et al.
Published: (2025)
Instructing Robots by Sketching: Learning from Demonstration via Probabilistic Diagrammatic Teaching
by: Zhi, Weiming, et al.
Published: (2023)
by: Zhi, Weiming, et al.
Published: (2023)
Learning Orbitally Stable Systems for Diagrammatically Teaching
by: Zhi, Weiming, et al.
Published: (2023)
by: Zhi, Weiming, et al.
Published: (2023)
VideoJudge: Bootstrapping Enables Scalable Supervision of MLLM-as-a-Judge for Video Understanding
by: Waheed, Abdul, et al.
Published: (2025)
by: Waheed, Abdul, et al.
Published: (2025)
Revisiting Acoustic Features for Robust ASR
by: Shah, Muhammad A., et al.
Published: (2024)
by: Shah, Muhammad A., et al.
Published: (2024)
Robust Bayesian Scene Reconstruction with Retrieval-Augmented Priors for Precise Grasping and Planning
by: Wright, Herbert, et al.
Published: (2024)
by: Wright, Herbert, et al.
Published: (2024)
Robust Latent Matters: Boosting Image Generation with Sampling Error Synthesis
by: Qiu, Kai, et al.
Published: (2025)
by: Qiu, Kai, et al.
Published: (2025)
Cross-Modal Instructions for Robot Motion Generation
by: Barron, William, et al.
Published: (2025)
by: Barron, William, et al.
Published: (2025)
Speech Robust Bench: A Robustness Benchmark For Speech Recognition
by: Shah, Muhammad A., et al.
Published: (2024)
by: Shah, Muhammad A., et al.
Published: (2024)
DarkGS: Learning Neural Illumination and 3D Gaussians Relighting for Robotic Exploration in the Dark
by: Zhang, Tianyi, et al.
Published: (2024)
by: Zhang, Tianyi, et al.
Published: (2024)
Bi-Manual Joint Camera Calibration and Scene Representation
by: Tang, Haozhan, et al.
Published: (2025)
by: Tang, Haozhan, et al.
Published: (2025)
PhotoReg: Photometrically Registering 3D Gaussian Splatting Models
by: Yuan, Ziwen, et al.
Published: (2024)
by: Yuan, Ziwen, et al.
Published: (2024)
3D Foundation Models Enable Simultaneous Geometry and Pose Estimation of Grasped Objects
by: Zhi, Weiming, et al.
Published: (2024)
by: Zhi, Weiming, et al.
Published: (2024)
Unifying Scene Representation and Hand-Eye Calibration with 3D Foundation Models
by: Zhi, Weiming, et al.
Published: (2024)
by: Zhi, Weiming, et al.
Published: (2024)
SplaTraj: Camera Trajectory Generation with Semantic Gaussian Splatting
by: Liu, Xinyi, et al.
Published: (2024)
by: Liu, Xinyi, et al.
Published: (2024)
Infinite Leagues Under the Sea: Photorealistic 3D Underwater Terrain Generation by Latent Fractal Diffusion Models
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
Teaching Robots Where To Go And How To Act With Human Sketches via Spatial Diagrammatic Instructions
by: Sun, Qilin, et al.
Published: (2024)
by: Sun, Qilin, et al.
Published: (2024)
DeWinder: Single-Channel Wind Noise Reduction using Ultrasound Sensing
by: Yuan, Kuang, et al.
Published: (2024)
by: Yuan, Kuang, et al.
Published: (2024)
EgoForce: Robust Online Egocentric Motion Reconstruction via Diffusion Forcing
by: Hwang, Inwoo, et al.
Published: (2026)
by: Hwang, Inwoo, et al.
Published: (2026)
SVeritas: Benchmark for Robust Speaker Verification under Diverse Conditions
by: Baali, Massa, et al.
Published: (2025)
by: Baali, Massa, et al.
Published: (2025)
EgoTraj-Bench: Towards Robust Trajectory Prediction Under Ego-view Noisy Observations
by: Liu, Jiayi, et al.
Published: (2025)
by: Liu, Jiayi, et al.
Published: (2025)
On the Robust Approximation of ASR Metrics
by: Waheed, Abdul, et al.
Published: (2025)
by: Waheed, Abdul, et al.
Published: (2025)
Human Voice is Unique
by: Singh, Rita, et al.
Published: (2025)
by: Singh, Rita, et al.
Published: (2025)
Synergistic Global-space Camera and Human Reconstruction from Videos
by: Zhao, Yizhou, et al.
Published: (2024)
by: Zhao, Yizhou, et al.
Published: (2024)
ANVIL: Accelerator-Native Video Interpolation via Codec Motion Vector Priors
by: Liu, Shibo
Published: (2026)
by: Liu, Shibo
Published: (2026)
AgentNoiseBench: Benchmarking Robustness of Tool-Using LLM Agents Under Noisy Condition
by: Wang, Ruipeng, et al.
Published: (2026)
by: Wang, Ruipeng, et al.
Published: (2026)
GraphSeg: Segmented 3D Representations via Graph Edge Addition and Contraction
by: Tang, Haozhan, et al.
Published: (2025)
by: Tang, Haozhan, et al.
Published: (2025)
Reasoning about the Unseen for Efficient Outdoor Object Navigation
by: Xie, Quanting, et al.
Published: (2023)
by: Xie, Quanting, et al.
Published: (2023)
DOSE3 : Diffusion-based Out-of-distribution detection on SE(3) trajectories
by: Cheng, Hongzhe, et al.
Published: (2025)
by: Cheng, Hongzhe, et al.
Published: (2025)
Evaluating and Improving Continual Learning in Spoken Language Understanding
by: Yang, Muqiao, et al.
Published: (2024)
by: Yang, Muqiao, et al.
Published: (2024)
When Search Becomes Memory: Turning Robot Design Trials into Transferable Skills
by: Wang, Yunfei, et al.
Published: (2026)
by: Wang, Yunfei, et al.
Published: (2026)
Similar Items
-
From Perfect to Noisy World Simulation: Customizable Embodied Multi-modal Perturbations for SLAM Robustness Benchmarking
by: Xu, Xiaohao, et al.
Published: (2024) -
Customizable Perturbation Synthesis for Robust SLAM Benchmarking
by: Xu, Xiaohao, et al.
Published: (2024) -
MAC-Ego3D: Multi-Agent Gaussian Consensus for Real-Time Collaborative Ego-Motion and Photorealistic 3D Reconstruction
by: Xu, Xiaohao, et al.
Published: (2024) -
Towards Ambiguity-Free Spatial Foundation Model: Rethinking and Decoupling Depth Ambiguity
by: Xu, Xiaohao, et al.
Published: (2025) -
Learning Shared RGB-D Fields: Unified Self-supervised Pre-training for Label-efficient LiDAR-Camera 3D Perception
by: Xu, Xiaohao, et al.
Published: (2024)