Fake It To Make It: Virtual Multiviews to Enhance Monocular Indoor Semantic Scene Completion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Selvakumar, Anith, Bharadwaj, Manasa |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Getting More for Less: Using Weak Labels and AV-Mixup for Robust Audio-Visual Speaker Verification
von: Selvakumar, Anith, et al.
Veröffentlicht: (2023)
von: Selvakumar, Anith, et al.
Veröffentlicht: (2023)
VROOM - Visual Reconstruction over Onboard Multiview
von: Yadav, Yajat, et al.
Veröffentlicht: (2025)
von: Yadav, Yajat, et al.
Veröffentlicht: (2025)
AdaSFormer: Adaptive Serialized Transformers for Monocular Semantic Scene Completion from Indoor Environments
von: Wang, Xuzhi, et al.
Veröffentlicht: (2026)
von: Wang, Xuzhi, et al.
Veröffentlicht: (2026)
SPARK: Self-supervised Personalized Real-time Monocular Face Capture
von: Baert, Kelian, et al.
Veröffentlicht: (2024)
von: Baert, Kelian, et al.
Veröffentlicht: (2024)
Faster Bounding Box Annotation for Object Detection in Indoor Scenes
von: Adhikari, Bishwo, et al.
Veröffentlicht: (2018)
von: Adhikari, Bishwo, et al.
Veröffentlicht: (2018)
Marigold-DC: Zero-Shot Monocular Depth Completion with Guided Diffusion
von: Viola, Massimiliano, et al.
Veröffentlicht: (2024)
von: Viola, Massimiliano, et al.
Veröffentlicht: (2024)
InSpaceType: Dataset and Benchmark for Reconsidering Cross-Space Type Performance in Indoor Monocular Depth
von: Wu, Cho-Ying, et al.
Veröffentlicht: (2024)
von: Wu, Cho-Ying, et al.
Veröffentlicht: (2024)
InsTex: Indoor Scenes Stylized Texture Synthesis
von: Zhang, Yunfan, et al.
Veröffentlicht: (2025)
von: Zhang, Yunfan, et al.
Veröffentlicht: (2025)
Hierarchical Consensus Network for Multiview Feature Learning
von: Xia, Chengwei, et al.
Veröffentlicht: (2025)
von: Xia, Chengwei, et al.
Veröffentlicht: (2025)
Structure is Supervision: Multiview Masked Autoencoders for Radiology
von: Laguna, Sonia, et al.
Veröffentlicht: (2025)
von: Laguna, Sonia, et al.
Veröffentlicht: (2025)
ConTEXTure: Consistent Multiview Images to Texture
von: Ahn, Jaehoon, et al.
Veröffentlicht: (2024)
von: Ahn, Jaehoon, et al.
Veröffentlicht: (2024)
Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling
von: Wu, Qirui, et al.
Veröffentlicht: (2024)
von: Wu, Qirui, et al.
Veröffentlicht: (2024)
Monocular Occupancy Prediction for Scalable Indoor Scenes
von: Yu, Hongxiao, et al.
Veröffentlicht: (2024)
von: Yu, Hongxiao, et al.
Veröffentlicht: (2024)
EchoScene: Indoor Scene Generation via Information Echo over Scene Graph Diffusion
von: Zhai, Guangyao, et al.
Veröffentlicht: (2024)
von: Zhai, Guangyao, et al.
Veröffentlicht: (2024)
Leveraging Stable Diffusion for Monocular Depth Estimation via Image Semantic Encoding
von: Xia, Jingming, et al.
Veröffentlicht: (2025)
von: Xia, Jingming, et al.
Veröffentlicht: (2025)
Attention over Scene Graphs: Indoor Scene Representations Toward CSAI Classification
von: Barros, Artur, et al.
Veröffentlicht: (2025)
von: Barros, Artur, et al.
Veröffentlicht: (2025)
Global-Aware Monocular Semantic Scene Completion with State Space Models
von: Li, Shijie, et al.
Veröffentlicht: (2025)
von: Li, Shijie, et al.
Veröffentlicht: (2025)
Applying Unsupervised Semantic Segmentation to High-Resolution UAV Imagery for Enhanced Road Scene Parsing
von: Ma, Zihan, et al.
Veröffentlicht: (2024)
von: Ma, Zihan, et al.
Veröffentlicht: (2024)
Monocular Semantic Scene Completion via Masked Recurrent Networks
von: Wang, Xuzhi, et al.
Veröffentlicht: (2025)
von: Wang, Xuzhi, et al.
Veröffentlicht: (2025)
IndoorCrowd: A Multi-Scene Dataset for Human Detection, Segmentation, and Tracking with an Automated Annotation Pipeline
von: Nae, Sebastian-Ion, et al.
Veröffentlicht: (2026)
von: Nae, Sebastian-Ion, et al.
Veröffentlicht: (2026)
PointSSC: A Cooperative Vehicle-Infrastructure Point Cloud Benchmark for Semantic Scene Completion
von: Yan, Yuxiang, et al.
Veröffentlicht: (2023)
von: Yan, Yuxiang, et al.
Veröffentlicht: (2023)
LumiNet: Latent Intrinsics Meets Diffusion Models for Indoor Scene Relighting
von: Xing, Xiaoyan, et al.
Veröffentlicht: (2024)
von: Xing, Xiaoyan, et al.
Veröffentlicht: (2024)
Refining Few-Step Text-to-Multiview Diffusion via Reinforcement Learning
von: Zhang, Ziyi, et al.
Veröffentlicht: (2025)
von: Zhang, Ziyi, et al.
Veröffentlicht: (2025)
Consistency Enhancement-Based Deep Multiview Clustering via Contrastive Learning
von: Yang, Hao, et al.
Veröffentlicht: (2024)
von: Yang, Hao, et al.
Veröffentlicht: (2024)
Fake it till You Make it: Reward Modeling as Discriminative Prediction
von: Liu, Runtao, et al.
Veröffentlicht: (2025)
von: Liu, Runtao, et al.
Veröffentlicht: (2025)
Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes
von: Zhou, Changqing, et al.
Veröffentlicht: (2026)
von: Zhou, Changqing, et al.
Veröffentlicht: (2026)
The Coralscapes Dataset: Semantic Scene Understanding in Coral Reefs
von: Sauder, Jonathan, et al.
Veröffentlicht: (2025)
von: Sauder, Jonathan, et al.
Veröffentlicht: (2025)
TT-BLIP: Enhancing Fake News Detection Using BLIP and Tri-Transformer
von: Choi, Eunjee, et al.
Veröffentlicht: (2024)
von: Choi, Eunjee, et al.
Veröffentlicht: (2024)
Virtually Enriched NYU Depth V2 Dataset for Monocular Depth Estimation: Do We Need Artificial Augmentation?
von: Ignatov, Dmitry, et al.
Veröffentlicht: (2024)
von: Ignatov, Dmitry, et al.
Veröffentlicht: (2024)
GroMo: Plant Growth Modeling with Multiview Images
von: Bhatt, Ruchi, et al.
Veröffentlicht: (2025)
von: Bhatt, Ruchi, et al.
Veröffentlicht: (2025)
Detect Fake with Fake: Leveraging Synthetic Data-driven Representation for Synthetic Image Detection
von: Otake, Hina, et al.
Veröffentlicht: (2024)
von: Otake, Hina, et al.
Veröffentlicht: (2024)
Class-Agnostic Visio-Temporal Scene Sketch Semantic Segmentation
von: Kütük, Aleyna, et al.
Veröffentlicht: (2024)
von: Kütük, Aleyna, et al.
Veröffentlicht: (2024)
OSN: Infinite Representations of Dynamic 3D Scenes from Monocular Videos
von: Song, Ziyang, et al.
Veröffentlicht: (2024)
von: Song, Ziyang, et al.
Veröffentlicht: (2024)
Multiview Scene Graph
von: Zhang, Juexiao, et al.
Veröffentlicht: (2024)
von: Zhang, Juexiao, et al.
Veröffentlicht: (2024)
SCENIR: Visual Semantic Clarity through Unsupervised Scene Graph Retrieval
von: Chaidos, Nikolaos, et al.
Veröffentlicht: (2025)
von: Chaidos, Nikolaos, et al.
Veröffentlicht: (2025)
Benchmarking Federated Learning for Semantic Datasets: Federated Scene Graph Generation
von: Ha, SeungBum, et al.
Veröffentlicht: (2024)
von: Ha, SeungBum, et al.
Veröffentlicht: (2024)
Enhancing Large Vision Model in Street Scene Semantic Understanding through Leveraging Posterior Optimization Trajectory
von: Kou, Wei-Bin, et al.
Veröffentlicht: (2025)
von: Kou, Wei-Bin, et al.
Veröffentlicht: (2025)
Uncertainty Estimation and Out-of-Distribution Detection for LiDAR Scene Semantic Segmentation
von: Shojaei, Hanieh, et al.
Veröffentlicht: (2024)
von: Shojaei, Hanieh, et al.
Veröffentlicht: (2024)
Generative Lifting of Multiview to 3D from Unknown Pose: Wrapping NeRF inside Diffusion
von: Yuan, Xin, et al.
Veröffentlicht: (2024)
von: Yuan, Xin, et al.
Veröffentlicht: (2024)
Skip Mamba Diffusion for Monocular 3D Semantic Scene Completion
von: Liang, Li, et al.
Veröffentlicht: (2025)
von: Liang, Li, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Getting More for Less: Using Weak Labels and AV-Mixup for Robust Audio-Visual Speaker Verification
von: Selvakumar, Anith, et al.
Veröffentlicht: (2023) -
VROOM - Visual Reconstruction over Onboard Multiview
von: Yadav, Yajat, et al.
Veröffentlicht: (2025) -
AdaSFormer: Adaptive Serialized Transformers for Monocular Semantic Scene Completion from Indoor Environments
von: Wang, Xuzhi, et al.
Veröffentlicht: (2026) -
SPARK: Self-supervised Personalized Real-time Monocular Face Capture
von: Baert, Kelian, et al.
Veröffentlicht: (2024) -
Faster Bounding Box Annotation for Object Detection in Indoor Scenes
von: Adhikari, Bishwo, et al.
Veröffentlicht: (2018)