CoL3D: Collaborative Learning of Single-view Depth and Camera Intrinsics for Metric 3D Shape Recovery
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Chenghao, Fan, Lubin, Cao, Shen, Wu, Bojian, Ye, Jieping |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PTZ-Calib: Robust Pan-Tilt-Zoom Camera Calibration
von: Guo, Jinhui, et al.
Veröffentlicht: (2025)
von: Guo, Jinhui, et al.
Veröffentlicht: (2025)
Learning Neural Volumetric Pose Features for Camera Localization
von: Lin, Jingyu, et al.
Veröffentlicht: (2024)
von: Lin, Jingyu, et al.
Veröffentlicht: (2024)
HybridGS: Decoupling Transients and Statics with 2D and 3D Gaussian Splatting
von: Lin, Jingyu, et al.
Veröffentlicht: (2024)
von: Lin, Jingyu, et al.
Veröffentlicht: (2024)
TrackGS: Optimizing COLMAP-Free 3D Gaussian Splatting with Global Track Constraints
von: Shi, Dongbo, et al.
Veröffentlicht: (2025)
von: Shi, Dongbo, et al.
Veröffentlicht: (2025)
NoPe-NeRF++: Local-to-Global Optimization of NeRF with No Pose Prior
von: Shi, Dongbo, et al.
Veröffentlicht: (2025)
von: Shi, Dongbo, et al.
Veröffentlicht: (2025)
AddressVLM: Cross-view Alignment Tuning for Image Address Localization using Large Vision-Language Models
von: Xu, Shixiong, et al.
Veröffentlicht: (2025)
von: Xu, Shixiong, et al.
Veröffentlicht: (2025)
AddressCLIP: Empowering Vision-Language Models for City-wide Image Address Localization
von: Xu, Shixiong, et al.
Veröffentlicht: (2024)
von: Xu, Shixiong, et al.
Veröffentlicht: (2024)
SD-VLM: Spatial Measuring and Understanding with Depth-Encoded Vision-Language Models
von: Chen, Pingyi, et al.
Veröffentlicht: (2025)
von: Chen, Pingyi, et al.
Veröffentlicht: (2025)
Re-ranking Reasoning Context with Tree Search Makes Large Vision-Language Models Stronger
von: Yang, Qi, et al.
Veröffentlicht: (2025)
von: Yang, Qi, et al.
Veröffentlicht: (2025)
MVD-Fusion: Single-view 3D via Depth-consistent Multi-view Generation
von: Hu, Hanzhe, et al.
Veröffentlicht: (2024)
von: Hu, Hanzhe, et al.
Veröffentlicht: (2024)
DragScene: Interactive 3D Scene Editing with Single-view Drag Instructions
von: Gu, Chenghao, et al.
Veröffentlicht: (2024)
von: Gu, Chenghao, et al.
Veröffentlicht: (2024)
Collision Avoidance Metric for 3D Camera Evaluation
von: Taamazyan, Vage, et al.
Veröffentlicht: (2024)
von: Taamazyan, Vage, et al.
Veröffentlicht: (2024)
Self-Learning Symmetric Multi-view Probabilistic Clustering
von: Liu, Junjie, et al.
Veröffentlicht: (2023)
von: Liu, Junjie, et al.
Veröffentlicht: (2023)
Depth-discriminative Metric Learning for Monocular 3D Object Detection
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2024)
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2024)
Transformed Multi-view 3D Shape Features with Contrastive Learning
von: Costa, Márcus Vinícius Lobo, et al.
Veröffentlicht: (2025)
von: Costa, Márcus Vinícius Lobo, et al.
Veröffentlicht: (2025)
Coca-Splat: Collaborative Optimization for Camera Parameters and 3D Gaussians
von: Wu, Jiamin, et al.
Veröffentlicht: (2025)
von: Wu, Jiamin, et al.
Veröffentlicht: (2025)
RoScenes: A Large-scale Multi-view 3D Dataset for Roadside Perception
von: Zhu, Xiaosu, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaosu, et al.
Veröffentlicht: (2024)
Animate3D: Animating Any 3D Model with Multi-view Video Diffusion
von: Jiang, Yanqin, et al.
Veröffentlicht: (2024)
von: Jiang, Yanqin, et al.
Veröffentlicht: (2024)
SR3D: Unleashing Single-view 3D Reconstruction for Transparent and Specular Object Grasping
von: Zhang, Mingxu, et al.
Veröffentlicht: (2025)
von: Zhang, Mingxu, et al.
Veröffentlicht: (2025)
L3D-Pose: Lifting Pose for 3D Avatars from a Single Camera in the Wild
von: Debnath, Soumyaratna, et al.
Veröffentlicht: (2025)
von: Debnath, Soumyaratna, et al.
Veröffentlicht: (2025)
Doracamom: Joint 3D Detection and Occupancy Prediction with Multi-view 4D Radars and Cameras for Omnidirectional Perception
von: Zheng, Lianqing, et al.
Veröffentlicht: (2025)
von: Zheng, Lianqing, et al.
Veröffentlicht: (2025)
Single-Shot Metric Depth from Focused Plenoptic Cameras
von: Lasheras-Hernandez, Blanca, et al.
Veröffentlicht: (2024)
von: Lasheras-Hernandez, Blanca, et al.
Veröffentlicht: (2024)
CoRe3D: Collaborative Reasoning as a Foundation for 3D Intelligence
von: Yu, Tianjiao, et al.
Veröffentlicht: (2025)
von: Yu, Tianjiao, et al.
Veröffentlicht: (2025)
RadarCam-Depth: Radar-Camera Fusion for Depth Estimation with Learned Metric Scale
von: Li, Han, et al.
Veröffentlicht: (2024)
von: Li, Han, et al.
Veröffentlicht: (2024)
3D Hand Mesh Recovery from Monocular RGB in Camera Space
von: Li, Haonan, et al.
Veröffentlicht: (2024)
von: Li, Haonan, et al.
Veröffentlicht: (2024)
Learning Image-Adaptive Scale Fields for Metric Depth Recovery
von: Li, Yuanyan, et al.
Veröffentlicht: (2026)
von: Li, Yuanyan, et al.
Veröffentlicht: (2026)
SRAM: Shape-Realism Alignment Metric for No Reference 3D Shape Evaluation
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
DanceCamera3D: 3D Camera Movement Synthesis with Music and Dance
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
Self-learning Canonical Space for Multi-view 3D Human Pose Estimation
von: Li, Xiaoben, et al.
Veröffentlicht: (2024)
von: Li, Xiaoben, et al.
Veröffentlicht: (2024)
Depth Any Camera: Zero-Shot Metric Depth Estimation from Any Camera
von: Guo, Yuliang, et al.
Veröffentlicht: (2025)
von: Guo, Yuliang, et al.
Veröffentlicht: (2025)
Ray Denoising: Depth-aware Hard Negative Sampling for Multi-view 3D Object Detection
von: Liu, Feng, et al.
Veröffentlicht: (2024)
von: Liu, Feng, et al.
Veröffentlicht: (2024)
CT3D++: Improving 3D Object Detection with Keypoint-induced Channel-wise Transformer
von: Sheng, Hualian, et al.
Veröffentlicht: (2024)
von: Sheng, Hualian, et al.
Veröffentlicht: (2024)
CoIn3D: Revisiting Configuration-Invariant Multi-Camera 3D Object Detection
von: Kuang, Zhaonian, et al.
Veröffentlicht: (2026)
von: Kuang, Zhaonian, et al.
Veröffentlicht: (2026)
Distribution and Depth-Aware Transformers for 3D Human Mesh Recovery
von: Bright, Jerrin, et al.
Veröffentlicht: (2024)
von: Bright, Jerrin, et al.
Veröffentlicht: (2024)
Canonical Pose Reconstruction from Single Depth Image for 3D Non-rigid Pose Recovery on Limited Datasets
von: Alhamazani, Fahd, et al.
Veröffentlicht: (2025)
von: Alhamazani, Fahd, et al.
Veröffentlicht: (2025)
CamCtrl3D: Single-Image Scene Exploration with Precise 3D Camera Control
von: Popov, Stefan, et al.
Veröffentlicht: (2025)
von: Popov, Stefan, et al.
Veröffentlicht: (2025)
GeoGS3D: Single-view 3D Reconstruction via Geometric-aware Diffusion Model and Gaussian Splatting
von: Feng, Qijun, et al.
Veröffentlicht: (2024)
von: Feng, Qijun, et al.
Veröffentlicht: (2024)
DepthFusion: Depth-Aware Hybrid Feature Fusion for LiDAR-Camera 3D Object Detection
von: Ji, Mingqian, et al.
Veröffentlicht: (2025)
von: Ji, Mingqian, et al.
Veröffentlicht: (2025)
BLADE: Single-view Body Mesh Learning through Accurate Depth Estimation
von: Wang, Shengze, et al.
Veröffentlicht: (2024)
von: Wang, Shengze, et al.
Veröffentlicht: (2024)
Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs
von: Tong, Jintao, et al.
Veröffentlicht: (2025)
von: Tong, Jintao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PTZ-Calib: Robust Pan-Tilt-Zoom Camera Calibration
von: Guo, Jinhui, et al.
Veröffentlicht: (2025) -
Learning Neural Volumetric Pose Features for Camera Localization
von: Lin, Jingyu, et al.
Veröffentlicht: (2024) -
HybridGS: Decoupling Transients and Statics with 2D and 3D Gaussian Splatting
von: Lin, Jingyu, et al.
Veröffentlicht: (2024) -
TrackGS: Optimizing COLMAP-Free 3D Gaussian Splatting with Global Track Constraints
von: Shi, Dongbo, et al.
Veröffentlicht: (2025) -
NoPe-NeRF++: Local-to-Global Optimization of NeRF with No Pose Prior
von: Shi, Dongbo, et al.
Veröffentlicht: (2025)