PixCuboid: Room Layout Estimation from Multi-view Featuremetric Alignment
Fuente:
arXiv
Guardado en:
| Autores principales: | Hanning, Gustav, Åström, Kalle, Larsson, Viktor |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Visual Re-Ranking with Non-Visual Side Information
por: Hanning, Gustav, et al.
Publicado: (2025)
por: Hanning, Gustav, et al.
Publicado: (2025)
Sparse Multiview Open-Vocabulary 3D Detection
por: Moliner, Olivier, et al.
Publicado: (2025)
por: Moliner, Olivier, et al.
Publicado: (2025)
Light Future: Multimodal Action Frame Prediction via InstructPix2Pix
por: Zhong, Zesen, et al.
Publicado: (2025)
por: Zhong, Zesen, et al.
Publicado: (2025)
Gaussian Alignment for Relative Camera Pose Estimation via Single-View Reconstruction
por: Li, Yumin, et al.
Publicado: (2025)
por: Li, Yumin, et al.
Publicado: (2025)
Obtaining Favorable Layouts for Multiple Object Generation
por: Battash, Barak, et al.
Publicado: (2024)
por: Battash, Barak, et al.
Publicado: (2024)
Object detection in adverse weather conditions for autonomous vehicles using Instruct Pix2Pix
por: Gurbindo, Unai, et al.
Publicado: (2025)
por: Gurbindo, Unai, et al.
Publicado: (2025)
Gaze Estimation for Human-Robot Interaction: Analysis Using the NICO Platform
por: Palider, Matej, et al.
Publicado: (2025)
por: Palider, Matej, et al.
Publicado: (2025)
Multi-modal On-Device Learning for Monocular Depth Estimation on Ultra-low-power MCUs
por: Nadalini, Davide, et al.
Publicado: (2025)
por: Nadalini, Davide, et al.
Publicado: (2025)
Robust Self-calibration of Focal Lengths from the Fundamental Matrix
por: Kocur, Viktor, et al.
Publicado: (2023)
por: Kocur, Viktor, et al.
Publicado: (2023)
Efficient Vision-based Vehicle Speed Estimation
por: Macko, Andrej, et al.
Publicado: (2025)
por: Macko, Andrej, et al.
Publicado: (2025)
Supersampling of Data from Structured-light Scanner with Deep Learning
por: Melicherčík, Martin, et al.
Publicado: (2023)
por: Melicherčík, Martin, et al.
Publicado: (2023)
Processing and Segmentation of Human Teeth from 2D Images using Weakly Supervised Learning
por: Kunzo, Tomáš, et al.
Publicado: (2023)
por: Kunzo, Tomáš, et al.
Publicado: (2023)
Edge Prediction for Roof Wireframe Reconstruction with Transformers
por: Hanning, Gustav, et al.
Publicado: (2026)
por: Hanning, Gustav, et al.
Publicado: (2026)
Evaluating the Significance of Outdoor Advertising from Driver's Perspective Using Computer Vision
por: Černeková, Zuzana, et al.
Publicado: (2023)
por: Černeková, Zuzana, et al.
Publicado: (2023)
Neuromorphic Monocular Depth Estimation with Uncertainty Modeling
por: Bergkvist, Viktor, et al.
Publicado: (2026)
por: Bergkvist, Viktor, et al.
Publicado: (2026)
A unified Benchmark for Multi-Frame Image Restoration under Severe Refractive Warping
por: Shugaev, Maxim V., et al.
Publicado: (2026)
por: Shugaev, Maxim V., et al.
Publicado: (2026)
Intrinsic Image Diffusion for Indoor Single-view Material Estimation
por: Kocsis, Peter, et al.
Publicado: (2023)
por: Kocsis, Peter, et al.
Publicado: (2023)
Robust Alignment of the Human Embryo in 3D Ultrasound using PCA and an Ensemble of Heuristic, Atlas-based and Learning-based Classifiers Evaluated on the Rotterdam Periconceptional Cohort
por: Herrmann, Nikolai, et al.
Publicado: (2025)
por: Herrmann, Nikolai, et al.
Publicado: (2025)
MdaIF: Robust One-Stop Multi-Degradation-Aware Image Fusion with Language-Driven Semantics
por: Li, Jing, et al.
Publicado: (2025)
por: Li, Jing, et al.
Publicado: (2025)
Removing Cost Volumes from Optical Flow Estimators
por: Kiefhaber, Simon, et al.
Publicado: (2025)
por: Kiefhaber, Simon, et al.
Publicado: (2025)
Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition
por: Wang, Yiming, et al.
Publicado: (2026)
por: Wang, Yiming, et al.
Publicado: (2026)
Learning Fine-to-Coarse Cuboid Shape Abstraction
por: Kobsik, Gregor, et al.
Publicado: (2025)
por: Kobsik, Gregor, et al.
Publicado: (2025)
Robust Multi-Source Covid-19 Detection in CT Images
por: Pritha, Asmita Yuki, et al.
Publicado: (2026)
por: Pritha, Asmita Yuki, et al.
Publicado: (2026)
Detecting 3D Line Segments for 6DoF Pose Estimation with Limited Data
por: Mok, Matej, et al.
Publicado: (2026)
por: Mok, Matej, et al.
Publicado: (2026)
A Multi-purpose Tracking Framework for Salmon Welfare Monitoring in Challenging Environments
por: Høgstedt, Espen Uri, et al.
Publicado: (2025)
por: Høgstedt, Espen Uri, et al.
Publicado: (2025)
Generalized Closed-form Formulae for Feature-based Subpixel Alignment in Patch-based Matching
por: Jospin, Laurent Valentin, et al.
Publicado: (2021)
por: Jospin, Laurent Valentin, et al.
Publicado: (2021)
Hierarchical Deep Learning for Diatom Image Classification: A Multi-Level Taxonomic Approach
por: Ke, Yueying
Publicado: (2025)
por: Ke, Yueying
Publicado: (2025)
Methods and strategies for improving the novel view synthesis quality of neural radiation field
por: Fang, Shun, et al.
Publicado: (2024)
por: Fang, Shun, et al.
Publicado: (2024)
Human Modelling and Pose Estimation Overview
por: Knap, Pawel
Publicado: (2024)
por: Knap, Pawel
Publicado: (2024)
DVLA-RL: Dual-Level Vision-Language Alignment with Reinforcement Learning Gating for Few-Shot Learning
por: Li, Wenhao, et al.
Publicado: (2026)
por: Li, Wenhao, et al.
Publicado: (2026)
Pointing-Based Object Recognition
por: Hajdúch, Lukáš, et al.
Publicado: (2026)
por: Hajdúch, Lukáš, et al.
Publicado: (2026)
UnCageNet: Tracking and Pose Estimation of Caged Animal
por: Dutta, Sayak, et al.
Publicado: (2025)
por: Dutta, Sayak, et al.
Publicado: (2025)
VIAFormer: Voxel-Image Alignment Transformer for High-Fidelity Voxel Refinement
por: Fang, Tiancheng, et al.
Publicado: (2026)
por: Fang, Tiancheng, et al.
Publicado: (2026)
Eye-gaze Guided Multi-modal Alignment for Medical Representation Learning
por: Ma, Chong, et al.
Publicado: (2024)
por: Ma, Chong, et al.
Publicado: (2024)
Scale-Invariant Monocular Depth Estimation via SSI Depth
por: Miangoleh, S. Mahdi H., et al.
Publicado: (2024)
por: Miangoleh, S. Mahdi H., et al.
Publicado: (2024)
MERIT: Multi-view evidential learning for reliable and interpretable liver fibrosis staging
por: Liu, Yuanye, et al.
Publicado: (2024)
por: Liu, Yuanye, et al.
Publicado: (2024)
EE3P: Event-based Estimation of Periodic Phenomena Properties
por: Kolář, Jakub, et al.
Publicado: (2024)
por: Kolář, Jakub, et al.
Publicado: (2024)
SVGS-DSGAT: An IoT-Enabled Innovation in Underwater Robotic Object Detection Technology
por: Wu, Dongli, et al.
Publicado: (2025)
por: Wu, Dongli, et al.
Publicado: (2025)
Low-Cost Tree Crown Dieback Estimation Using Deep Learning-Based Segmentation
por: Allen, M. J., et al.
Publicado: (2024)
por: Allen, M. J., et al.
Publicado: (2024)
Deep Learning-based Depth Estimation Methods from Monocular Image and Videos: A Comprehensive Survey
por: Rajapaksha, Uchitha, et al.
Publicado: (2024)
por: Rajapaksha, Uchitha, et al.
Publicado: (2024)
Ejemplares similares
-
Visual Re-Ranking with Non-Visual Side Information
por: Hanning, Gustav, et al.
Publicado: (2025) -
Sparse Multiview Open-Vocabulary 3D Detection
por: Moliner, Olivier, et al.
Publicado: (2025) -
Light Future: Multimodal Action Frame Prediction via InstructPix2Pix
por: Zhong, Zesen, et al.
Publicado: (2025) -
Gaussian Alignment for Relative Camera Pose Estimation via Single-View Reconstruction
por: Li, Yumin, et al.
Publicado: (2025) -
Obtaining Favorable Layouts for Multiple Object Generation
por: Battash, Barak, et al.
Publicado: (2024)