Saved in:
| Main Authors: | Xu, Zhen, Zhou, Hongyu, Peng, Sida, Lin, Haotong, Guo, Haoyu, Shao, Jiahao, Yang, Peishan, Yang, Qinglin, Miao, Sheng, He, Xingyi, Wang, Yifan, Wang, Yue, Hu, Ruizhen, Liao, Yiyi, Zhou, Xiaowei, Bao, Hujun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2507.11540 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-view Reconstruction via SfM-guided Monocular Depth Estimation
by: Guo, Haoyu, et al.
Published: (2025)
by: Guo, Haoyu, et al.
Published: (2025)
InfiniDepth: Arbitrary-Resolution and Fine-Grained Depth Estimation with Neural Implicit Fields
by: Yu, Hao, et al.
Published: (2026)
by: Yu, Hao, et al.
Published: (2026)
Prompting Depth Anything for 4K Resolution Accurate Metric Depth Estimation
by: Lin, Haotong, et al.
Published: (2024)
by: Lin, Haotong, et al.
Published: (2024)
FreeTimeGS: Free Gaussian Primitives at Anytime and Anywhere for Dynamic Scene Reconstruction
by: Wang, Yifan, et al.
Published: (2025)
by: Wang, Yifan, et al.
Published: (2025)
Precise Action-to-Video Generation Through Visual Action Prompts
by: Wang, Yuang, et al.
Published: (2025)
by: Wang, Yuang, et al.
Published: (2025)
Learning Temporally Consistent Video Depth from Video Diffusion Priors
by: Shao, Jiahao, et al.
Published: (2024)
by: Shao, Jiahao, et al.
Published: (2024)
The Constant Eye: Benchmarking and Bridging Appearance Robustness in Autonomous Driving
by: Wang, Jiabao, et al.
Published: (2026)
by: Wang, Jiabao, et al.
Published: (2026)
MatchAnything: Universal Cross-Modality Image Matching with Large-Scale Pre-Training
by: He, Xingyi, et al.
Published: (2025)
by: He, Xingyi, et al.
Published: (2025)
StreetCrafter: Street View Synthesis with Controllable Video Diffusion Models
by: Yan, Yunzhi, et al.
Published: (2024)
by: Yan, Yunzhi, et al.
Published: (2024)
Ready-to-React: Online Reaction Policy for Two-Character Interaction Generation
by: Cen, Zhi, et al.
Published: (2025)
by: Cen, Zhi, et al.
Published: (2025)
Efficient LoFTR: Semi-Dense Local Feature Matching with Sparse-Like Speed
by: Wang, Yifan, et al.
Published: (2024)
by: Wang, Yifan, et al.
Published: (2024)
Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models
by: Jin, Yudong, et al.
Published: (2025)
by: Jin, Yudong, et al.
Published: (2025)
Split4D: Decomposed 4D Scene Reconstruction Without Video Segmentation
by: Hu, Yongzhen, et al.
Published: (2025)
by: Hu, Yongzhen, et al.
Published: (2025)
MaPa: Text-driven Photorealistic Material Painting for 3D Shapes
by: Zhang, Shangzan, et al.
Published: (2024)
by: Zhang, Shangzan, et al.
Published: (2024)
World-Grounded Human Motion Recovery via Gravity-View Coordinates
by: Shen, Zehong, et al.
Published: (2024)
by: Shen, Zehong, et al.
Published: (2024)
SAM-guided Graph Cut for 3D Instance Segmentation
by: Guo, Haoyu, et al.
Published: (2023)
by: Guo, Haoyu, et al.
Published: (2023)
Painting 3D Nature in 2D: View Synthesis of Natural Scenes from a Single Semantic Mask
by: Zhang, Shangzan, et al.
Published: (2023)
by: Zhang, Shangzan, et al.
Published: (2023)
Representing Long Volumetric Video with Temporal Gaussian Hierarchy
by: Xu, Zhen, et al.
Published: (2024)
by: Xu, Zhen, et al.
Published: (2024)
ReRoPE: Repurposing RoPE for Relative Camera Control
by: Li, Chunyang, et al.
Published: (2026)
by: Li, Chunyang, et al.
Published: (2026)
UniVerse: Unleashing the Scene Prior of Video Diffusion Models for Robust Radiance Field Reconstruction
by: Cao, Jin, et al.
Published: (2025)
by: Cao, Jin, et al.
Published: (2025)
StreamingTalker: Audio-driven 3D Facial Animation with Autoregressive Diffusion Model
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
BenchDepth: Are We on the Right Way to Evaluate Depth Foundation Models?
by: Li, Zhenyu, et al.
Published: (2025)
by: Li, Zhenyu, et al.
Published: (2025)
BoxDreamer: Dreaming Box Corners for Generalizable Object Pose Estimation
by: Yu, Yuanhong, et al.
Published: (2025)
by: Yu, Yuanhong, et al.
Published: (2025)
Generating Human Motion in 3D Scenes from Text Descriptions
by: Cen, Zhi, et al.
Published: (2024)
by: Cen, Zhi, et al.
Published: (2024)
Glossy Object Reconstruction with Cost-effective Polarized Acquisition
by: Wu, Bojian, et al.
Published: (2025)
by: Wu, Bojian, et al.
Published: (2025)
Habitat-GS: A High-Fidelity Navigation Simulator with Dynamic Gaussian Splatting
by: Xia, Ziyuan, et al.
Published: (2026)
by: Xia, Ziyuan, et al.
Published: (2026)
EnvGS: Modeling View-Dependent Appearance with Environment Gaussian
by: Xie, Tao, et al.
Published: (2024)
by: Xie, Tao, et al.
Published: (2024)
HumanRAM: Feed-forward Human Reconstruction and Animation Model using Transformers
by: Yu, Zhiyuan, et al.
Published: (2025)
by: Yu, Zhiyuan, et al.
Published: (2025)
FreeFix: Boosting 3D Gaussian Splatting via Fine-Tuning-Free Diffusion Models
by: Zhou, Hongyu, et al.
Published: (2026)
by: Zhou, Hongyu, et al.
Published: (2026)
Efficient Depth-Guided Urban View Synthesis
by: Miao, Sheng, et al.
Published: (2024)
by: Miao, Sheng, et al.
Published: (2024)
Pixel-Perfect Depth with Semantics-Prompted Diffusion Transformers
by: Xu, Gangwei, et al.
Published: (2025)
by: Xu, Gangwei, et al.
Published: (2025)
EgoAgent: A Joint Predictive Agent Model in Egocentric Worlds
by: Chen, Lu, et al.
Published: (2025)
by: Chen, Lu, et al.
Published: (2025)
Dyn-E: Local Appearance Editing of Dynamic Neural Radiance Fields
by: Zhang, Shangzan, et al.
Published: (2023)
by: Zhang, Shangzan, et al.
Published: (2023)
Transparent Object Depth Completion
by: Zhou, Yifan, et al.
Published: (2024)
by: Zhou, Yifan, et al.
Published: (2024)
BRIDGE -- Building Reinforcement-Learning Depth-to-Image Data Generation Engine for Monocular Depth Estimation
by: Liu, Dingning, et al.
Published: (2025)
by: Liu, Dingning, et al.
Published: (2025)
HUGS: Holistic Urban 3D Scene Understanding via Gaussian Splatting
by: Zhou, Hongyu, et al.
Published: (2024)
by: Zhou, Hongyu, et al.
Published: (2024)
Street Gaussians: Modeling Dynamic Urban Scenes with Gaussian Splatting
by: Yan, Yunzhi, et al.
Published: (2024)
by: Yan, Yunzhi, et al.
Published: (2024)
PC-Planner: Physics-Constrained Self-Supervised Learning for Robust Neural Motion Planning with Shape-Aware Distance Function
by: Shen, Xujie, et al.
Published: (2024)
by: Shen, Xujie, et al.
Published: (2024)
Unsegment Anything by Simulating Deformation
by: Lu, Jiahao, et al.
Published: (2024)
by: Lu, Jiahao, et al.
Published: (2024)
Propagating Sparse Depth via Depth Foundation Model for Out-of-Distribution Depth Completion
by: Chen, Shenglun, et al.
Published: (2025)
by: Chen, Shenglun, et al.
Published: (2025)
Similar Items
-
Multi-view Reconstruction via SfM-guided Monocular Depth Estimation
by: Guo, Haoyu, et al.
Published: (2025) -
InfiniDepth: Arbitrary-Resolution and Fine-Grained Depth Estimation with Neural Implicit Fields
by: Yu, Hao, et al.
Published: (2026) -
Prompting Depth Anything for 4K Resolution Accurate Metric Depth Estimation
by: Lin, Haotong, et al.
Published: (2024) -
FreeTimeGS: Free Gaussian Primitives at Anytime and Anywhere for Dynamic Scene Reconstruction
by: Wang, Yifan, et al.
Published: (2025) -
Precise Action-to-Video Generation Through Visual Action Prompts
by: Wang, Yuang, et al.
Published: (2025)