Saved in:
| Main Authors: | Lv, Zhen, Long, Yangqi, Huang, Congzhentao, Li, Cao, Lv, Chengfei, Ren, Hao, Zheng, Dian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2411.11934 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SpatialDreamer: Incentivizing Spatial Reasoning via Active Mental Imagery
by: Cao, Meng, et al.
Published: (2025)
by: Cao, Meng, et al.
Published: (2025)
Panorama Generation From NFoV Image Done Right
by: Zheng, Dian, et al.
Published: (2025)
by: Zheng, Dian, et al.
Published: (2025)
StereoWorld: Geometry-Aware Monocular-to-Stereo Video Generation
by: Xing, Ke, et al.
Published: (2025)
by: Xing, Ke, et al.
Published: (2025)
Eye2Eye: A Simple Approach for Monocular-to-Stereo Video Synthesis
by: Geyer, Michal, et al.
Published: (2025)
by: Geyer, Michal, et al.
Published: (2025)
Self-supervised Monocular Depth Estimation with Large Kernel Attention
by: Xiang, Xuezhi, et al.
Published: (2024)
by: Xiang, Xuezhi, et al.
Published: (2024)
4DGT: Learning a 4D Gaussian Transformer Using Real-World Monocular Videos
by: Xu, Zhen, et al.
Published: (2025)
by: Xu, Zhen, et al.
Published: (2025)
Adaptive Discrete Disparity Volume for Self-supervised Monocular Depth Estimation
by: Ren, Jianwei
Published: (2024)
by: Ren, Jianwei
Published: (2024)
Pseudo-Stereo Inputs: A Solution to the Occlusion Challenge in Self-Supervised Stereo Matching
by: Yang, Ruizhi, et al.
Published: (2024)
by: Yang, Ruizhi, et al.
Published: (2024)
DreamStereo: Towards Real-Time Stereo Inpainting for HD Videos
by: Huang, Yuan, et al.
Published: (2026)
by: Huang, Yuan, et al.
Published: (2026)
Self Adaptive Threshold Pseudo-labeling and Unreliable Sample Contrastive Loss for Semi-supervised Image Classification
by: Zhang, Xuerong, et al.
Published: (2024)
by: Zhang, Xuerong, et al.
Published: (2024)
Stereo World Model: Camera-Guided Stereo Video Generation
by: Sun, Yang-Tian, et al.
Published: (2026)
by: Sun, Yang-Tian, et al.
Published: (2026)
Mono2Stereo: Monocular Knowledge Transfer for Enhanced Stereo Matching
by: Wang, Yuran, et al.
Published: (2024)
by: Wang, Yuran, et al.
Published: (2024)
SMFormer: Empowering Self-supervised Stereo Matching via Foundation Models and Data Augmentation
by: Wang, Yun, et al.
Published: (2026)
by: Wang, Yun, et al.
Published: (2026)
Self-Calibrating 4D Novel View Synthesis from Monocular Videos Using Gaussian Splatting
by: Li, Fang, et al.
Published: (2024)
by: Li, Fang, et al.
Published: (2024)
BINO: Encoder Centric Self Supervised Stereo With Native Pair Input
by: Zhou, Haokun
Published: (2026)
by: Zhou, Haokun
Published: (2026)
M2SVid: End-to-End Inpainting and Refinement for Monocular-to-Stereo Video Conversion
by: Shvetsova, Nina, et al.
Published: (2025)
by: Shvetsova, Nina, et al.
Published: (2025)
Self-supervised Pretraining and Finetuning for Monocular Depth and Visual Odometry
by: Chidlovskii, Boris, et al.
Published: (2024)
by: Chidlovskii, Boris, et al.
Published: (2024)
DO3D: Self-supervised Learning of Decomposed Object-aware 3D Motion and Depth from Monocular Videos
by: Wu, Xiuzhe, et al.
Published: (2024)
by: Wu, Xiuzhe, et al.
Published: (2024)
OutDreamer: Video Outpainting with a Diffusion Transformer
by: Zhong, Linhao, et al.
Published: (2025)
by: Zhong, Linhao, et al.
Published: (2025)
Neural Multi-View Self-Calibrated Photometric Stereo without Photometric Stereo Cues
by: Cao, Xu, et al.
Published: (2025)
by: Cao, Xu, et al.
Published: (2025)
DASP: Self-supervised Nighttime Monocular Depth Estimation with Domain Adaptation of Spatiotemporal Priors
by: Huang, Yiheng, et al.
Published: (2025)
by: Huang, Yiheng, et al.
Published: (2025)
Diffusion Priors for Dynamic View Synthesis from Monocular Videos
by: Wang, Chaoyang, et al.
Published: (2024)
by: Wang, Chaoyang, et al.
Published: (2024)
Diving into the Fusion of Monocular Priors for Generalized Stereo Matching
by: Yao, Chengtang, et al.
Published: (2025)
by: Yao, Chengtang, et al.
Published: (2025)
DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation
by: Zhao, Guosheng, et al.
Published: (2024)
by: Zhao, Guosheng, et al.
Published: (2024)
GlossyGS: Inverse Rendering of Glossy Objects with 3D Gaussian Splatting
by: Lai, Shuichang, et al.
Published: (2024)
by: Lai, Shuichang, et al.
Published: (2024)
WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens
by: Wang, Xiaofeng, et al.
Published: (2024)
by: Wang, Xiaofeng, et al.
Published: (2024)
StereoCarla: A High-Fidelity Driving Dataset for Generalizable Stereo
by: Guo, Xianda, et al.
Published: (2025)
by: Guo, Xianda, et al.
Published: (2025)
ClotheDreamer: Text-Guided Garment Generation with 3D Gaussians
by: Liu, Yufei, et al.
Published: (2024)
by: Liu, Yufei, et al.
Published: (2024)
NimbleD: Enhancing Self-supervised Monocular Depth Estimation with Pseudo-labels and Large-scale Video Pre-training
by: Luginov, Albert, et al.
Published: (2024)
by: Luginov, Albert, et al.
Published: (2024)
HumanDreamer: Generating Controllable Human-Motion Videos via Decoupled Generation
by: Wang, Boyuan, et al.
Published: (2025)
by: Wang, Boyuan, et al.
Published: (2025)
FA-Depth: Toward Fast and Accurate Self-supervised Monocular Depth Estimation
by: Wang, Fei, et al.
Published: (2024)
by: Wang, Fei, et al.
Published: (2024)
Adaptive Depth-converted-Scale Convolution for Self-supervised Monocular Depth Estimation
by: Gao, Yanbo, et al.
Published: (2026)
by: Gao, Yanbo, et al.
Published: (2026)
StereoAdapter-2: Globally Structure-Consistent Underwater Stereo Depth Estimation
by: Ren, Zeyu, et al.
Published: (2026)
by: Ren, Zeyu, et al.
Published: (2026)
GaussVideoDreamer: 3D Scene Generation with Video Diffusion and Inconsistency-Aware Gaussian Splatting
by: Hao, Junlin, et al.
Published: (2025)
by: Hao, Junlin, et al.
Published: (2025)
ViDAR: Video Diffusion-Aware 4D Reconstruction From Monocular Inputs
by: Nazarczuk, Michal, et al.
Published: (2025)
by: Nazarczuk, Michal, et al.
Published: (2025)
Alternate Diverse Teaching for Semi-supervised Medical Image Segmentation
by: Zhao, Zhen, et al.
Published: (2023)
by: Zhao, Zhen, et al.
Published: (2023)
Prism: Semi-Supervised Multi-View Stereo with Monocular Structure Priors
by: Rich, Alex, et al.
Published: (2024)
by: Rich, Alex, et al.
Published: (2024)
MonoMVSNet: Monocular Priors Guided Multi-View Stereo Network
by: Jiang, Jianfei, et al.
Published: (2025)
by: Jiang, Jianfei, et al.
Published: (2025)
MVSFormer++: Revealing the Devil in Transformer's Details for Multi-View Stereo
by: Cao, Chenjie, et al.
Published: (2024)
by: Cao, Chenjie, et al.
Published: (2024)
RayZer: A Self-supervised Large View Synthesis Model
by: Jiang, Hanwen, et al.
Published: (2025)
by: Jiang, Hanwen, et al.
Published: (2025)
Similar Items
-
SpatialDreamer: Incentivizing Spatial Reasoning via Active Mental Imagery
by: Cao, Meng, et al.
Published: (2025) -
Panorama Generation From NFoV Image Done Right
by: Zheng, Dian, et al.
Published: (2025) -
StereoWorld: Geometry-Aware Monocular-to-Stereo Video Generation
by: Xing, Ke, et al.
Published: (2025) -
Eye2Eye: A Simple Approach for Monocular-to-Stereo Video Synthesis
by: Geyer, Michal, et al.
Published: (2025) -
Self-supervised Monocular Depth Estimation with Large Kernel Attention
by: Xiang, Xuezhi, et al.
Published: (2024)