VistaDream: Sampling multiview consistent images for single-view scene reconstruction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Haiping, Liu, Yuan, Liu, Ziwei, Wang, Wenping, Dong, Zhen, Yang, Bisheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FreeReg: Image-to-Point Cloud Registration Leveraging Pretrained Diffusion Models and Monocular Depth Estimators
von: Wang, Haiping, et al.
Veröffentlicht: (2023)
von: Wang, Haiping, et al.
Veröffentlicht: (2023)
GAGS: Granularity-Aware Feature Distillation for Language Gaussian Splatting
von: Peng, Yuning, et al.
Veröffentlicht: (2024)
von: Peng, Yuning, et al.
Veröffentlicht: (2024)
ME-CPT: Multi-Task Enhanced Cross-Temporal Point Transformer for Urban 3D Change Detection
von: Zhang, Luqi, et al.
Veröffentlicht: (2025)
von: Zhang, Luqi, et al.
Veröffentlicht: (2025)
SpatialLLM: From Multi-modality Data to Urban Spatial Intelligence
von: Chen, Jiabin, et al.
Veröffentlicht: (2025)
von: Chen, Jiabin, et al.
Veröffentlicht: (2025)
Exploiting Motion Prior for Accurate Pose Estimation of Dashboard Cameras
von: Lu, Yipeng, et al.
Veröffentlicht: (2024)
von: Lu, Yipeng, et al.
Veröffentlicht: (2024)
Reliable-loc: Robust sequential LiDAR global localization in large-scale street scenes based on verifiable cues
von: Zou, Xianghong, et al.
Veröffentlicht: (2024)
von: Zou, Xianghong, et al.
Veröffentlicht: (2024)
Unleashing the Capabilities of Large Vision-Language Models for Intelligent Perception of Roadside Infrastructure
von: Fu, Luxuan, et al.
Veröffentlicht: (2026)
von: Fu, Luxuan, et al.
Veröffentlicht: (2026)
SaliencyI2PLoc: saliency-guided image-point cloud localization using contrastive learning
von: Li, Yuhao, et al.
Veröffentlicht: (2024)
von: Li, Yuhao, et al.
Veröffentlicht: (2024)
SyncDreamer: Generating Multiview-consistent Images from a Single-view Image
von: Liu, Yuan, et al.
Veröffentlicht: (2023)
von: Liu, Yuan, et al.
Veröffentlicht: (2023)
SVII-3D: Advancing Roadside Infrastructure Inventory with Decimeter-level 3D Localization and Comprehension from Sparse Street Imagery
von: Liu, Chong, et al.
Veröffentlicht: (2026)
von: Liu, Chong, et al.
Veröffentlicht: (2026)
Explicitly Guided Information Interaction Network for Cross-modal Point Cloud Completion
von: Xu, Hang, et al.
Veröffentlicht: (2024)
von: Xu, Hang, et al.
Veröffentlicht: (2024)
DPG-CD: Depth-Prior-Guided Cross-Modal Joint 2D-3D Change Detection
von: Zhang, Luqi, et al.
Veröffentlicht: (2026)
von: Zhang, Luqi, et al.
Veröffentlicht: (2026)
WHU-PCPR: A cross-platform heterogeneous point cloud dataset for place recognition in complex urban scenes
von: Zou, Xianghong, et al.
Veröffentlicht: (2026)
von: Zou, Xianghong, et al.
Veröffentlicht: (2026)
Efficient scene text image super-resolution with semantic guidance
von: TomyEnrique, LeoWu, et al.
Veröffentlicht: (2024)
von: TomyEnrique, LeoWu, et al.
Veröffentlicht: (2024)
CADDreamer: CAD Object Generation from Single-view Images
von: Li, Yuan, et al.
Veröffentlicht: (2025)
von: Li, Yuan, et al.
Veröffentlicht: (2025)
EC-Depth: Exploring the consistency of self-supervised monocular depth estimation in challenging scenes
von: Song, Ziyang, et al.
Veröffentlicht: (2023)
von: Song, Ziyang, et al.
Veröffentlicht: (2023)
Teaching in adverse scenes: a statistically feedback-driven threshold and mask adjustment teacher-student framework for object detection in UAV images under adverse scenes
von: Chen, Hongyu, et al.
Veröffentlicht: (2025)
von: Chen, Hongyu, et al.
Veröffentlicht: (2025)
LifelongPR: Lifelong point cloud place recognition based on sample replay and prompt learning
von: Zou, Xianghong, et al.
Veröffentlicht: (2025)
von: Zou, Xianghong, et al.
Veröffentlicht: (2025)
EgoTwin: Dreaming Body and View in First Person
von: Xiu, Jingqiao, et al.
Veröffentlicht: (2025)
von: Xiu, Jingqiao, et al.
Veröffentlicht: (2025)
GeoVista: Web-Augmented Agentic Visual Reasoning for Geolocalization
von: Wang, Yikun, et al.
Veröffentlicht: (2025)
von: Wang, Yikun, et al.
Veröffentlicht: (2025)
ViiNeuS: Volumetric Initialization for Implicit Neural Surface reconstruction of urban scenes with limited image overlap
von: Djeghim, Hala, et al.
Veröffentlicht: (2024)
von: Djeghim, Hala, et al.
Veröffentlicht: (2024)
SpaceVista: All-Scale Visual Spatial Reasoning from mm to km
von: Sun, Peiwen, et al.
Veröffentlicht: (2025)
von: Sun, Peiwen, et al.
Veröffentlicht: (2025)
Expert Knowledge-Guided Decision Calibration for Accurate Fine-Grained Tree Species Classification
von: Long, Chen, et al.
Veröffentlicht: (2026)
von: Long, Chen, et al.
Veröffentlicht: (2026)
Consistent text-to-image generation via scene de-contextualization
von: Tang, Song, et al.
Veröffentlicht: (2025)
von: Tang, Song, et al.
Veröffentlicht: (2025)
DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation
von: Tang, Jiaxiang, et al.
Veröffentlicht: (2023)
von: Tang, Jiaxiang, et al.
Veröffentlicht: (2023)
Part123: Part-aware 3D Reconstruction from a Single-view Image
von: Liu, Anran, et al.
Veröffentlicht: (2024)
von: Liu, Anran, et al.
Veröffentlicht: (2024)
Rethinking Generalizable Infrared Small Target Detection: A Real-scene Benchmark and Cross-view Representation Learning
von: Lu, Yahao, et al.
Veröffentlicht: (2025)
von: Lu, Yahao, et al.
Veröffentlicht: (2025)
HICT: High-precision 3D CBCT reconstruction from a single X-ray
von: Ma, Wen, et al.
Veröffentlicht: (2026)
von: Ma, Wen, et al.
Veröffentlicht: (2026)
UNIC: Neural Garment Deformation Field for Real-time Clothed Character Animation
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2026)
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2026)
Personalized Cell Segmentation: Benchmark and Framework for Reference-Guided Cell Type Segmentation
von: Wang, Bisheng, et al.
Veröffentlicht: (2026)
von: Wang, Bisheng, et al.
Veröffentlicht: (2026)
Multi-view learning for automatic classification of multi-wavelength auroral images
von: Yang, Qiuju, et al.
Veröffentlicht: (2023)
von: Yang, Qiuju, et al.
Veröffentlicht: (2023)
DeepAAT: Deep Automated Aerial Triangulation for Fast UAV-based Mapping
von: Chen, Zequan, et al.
Veröffentlicht: (2024)
von: Chen, Zequan, et al.
Veröffentlicht: (2024)
AnchoredDream: Zero-Shot 360° Indoor Scene Generation from a Single View via Geometric Grounding
von: Yao, Runmao, et al.
Veröffentlicht: (2026)
von: Yao, Runmao, et al.
Veröffentlicht: (2026)
Gamba: Marry Gaussian Splatting with Mamba for single view 3D reconstruction
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
Classifying geospatial objects from multiview aerial imagery using semantic meshes
von: Russell, David, et al.
Veröffentlicht: (2024)
von: Russell, David, et al.
Veröffentlicht: (2024)
3D scene generation from scene graphs and self-attention
von: Bonazzi, Pietro, et al.
Veröffentlicht: (2024)
von: Bonazzi, Pietro, et al.
Veröffentlicht: (2024)
Weakly-supervised 3D coronary artery reconstruction from two-view angiographic images
von: Wang, Lu, et al.
Veröffentlicht: (2020)
von: Wang, Lu, et al.
Veröffentlicht: (2020)
Target-aware Image Editing via Cycle-consistent Constraints
von: Wang, Yanghao, et al.
Veröffentlicht: (2025)
von: Wang, Yanghao, et al.
Veröffentlicht: (2025)
APCoTTA: Continual Test-Time Adaptation for Semantic Segmentation of Airborne LiDAR Point Clouds
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
SERES: Semantic-aware neural reconstruction from sparse views
von: Xu, Bo, et al.
Veröffentlicht: (2025)
von: Xu, Bo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FreeReg: Image-to-Point Cloud Registration Leveraging Pretrained Diffusion Models and Monocular Depth Estimators
von: Wang, Haiping, et al.
Veröffentlicht: (2023) -
GAGS: Granularity-Aware Feature Distillation for Language Gaussian Splatting
von: Peng, Yuning, et al.
Veröffentlicht: (2024) -
ME-CPT: Multi-Task Enhanced Cross-Temporal Point Transformer for Urban 3D Change Detection
von: Zhang, Luqi, et al.
Veröffentlicht: (2025) -
SpatialLLM: From Multi-modality Data to Urban Spatial Intelligence
von: Chen, Jiabin, et al.
Veröffentlicht: (2025) -
Exploiting Motion Prior for Accurate Pose Estimation of Dashboard Cameras
von: Lu, Yipeng, et al.
Veröffentlicht: (2024)