Beyond Visual Field of View: Perceiving 3D Environment with Echoes and Vision
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhu, Lingyu, Rahtu, Esa, Zhao, Hang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2022
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
3D Gaussian Splatting with Fisheye Images: Field of View Analysis and Depth-Based Initialization
por: Gunes, Ulas, et al.
Publicado: (2025)
por: Gunes, Ulas, et al.
Publicado: (2025)
GS-Pose: Generalizable Segmentation-based 6D Object Pose Estimation with 3D Gaussian Splatting
por: Cai, Dingding, et al.
Publicado: (2024)
por: Cai, Dingding, et al.
Publicado: (2024)
PanDepth: Joint Panoptic Segmentation and Depth Completion
por: Lagos, Juan, et al.
Publicado: (2022)
por: Lagos, Juan, et al.
Publicado: (2022)
The Weighting Game: Evaluating Quality of Explainability Methods
por: Raatikainen, Lassi, et al.
Publicado: (2022)
por: Raatikainen, Lassi, et al.
Publicado: (2022)
MuSHRoom: Multi-Sensor Hybrid Room Dataset for Joint 3D Reconstruction and Novel View Synthesis
por: Ren, Xuqian, et al.
Publicado: (2023)
por: Ren, Xuqian, et al.
Publicado: (2023)
Video Object Segmentation-Aware Audio Generation
por: Viertola, Ilpo, et al.
Publicado: (2025)
por: Viertola, Ilpo, et al.
Publicado: (2025)
SemSegDepth: A Combined Model for Semantic Segmentation and Depth Completion
por: Lagos, Juan Pablo, et al.
Publicado: (2022)
por: Lagos, Juan Pablo, et al.
Publicado: (2022)
HybVIO: Pushing the Limits of Real-time Visual-inertial Odometry
por: Seiskari, Otto, et al.
Publicado: (2021)
por: Seiskari, Otto, et al.
Publicado: (2021)
FIORD: A Fisheye Indoor-Outdoor Dataset with LIDAR Ground Truth for 3D Scene Reconstruction and Benchmarking
por: Gunes, Ulas, et al.
Publicado: (2025)
por: Gunes, Ulas, et al.
Publicado: (2025)
Temporally Aligned Audio for Video with Autoregression
por: Viertola, Ilpo, et al.
Publicado: (2024)
por: Viertola, Ilpo, et al.
Publicado: (2024)
UDGS-SLAM : UniDepth Assisted Gaussian Splatting for Monocular SLAM
por: Mansour, Mostafa, et al.
Publicado: (2024)
por: Mansour, Mostafa, et al.
Publicado: (2024)
DN-Splatter: Depth and Normal Priors for Gaussian Splatting and Meshing
por: Turkulainen, Matias, et al.
Publicado: (2024)
por: Turkulainen, Matias, et al.
Publicado: (2024)
AGS-Mesh: Adaptive Gaussian Splatting and Meshing with Geometric Priors for Indoor Room Reconstruction Using Smartphones
por: Ren, Xuqian, et al.
Publicado: (2024)
por: Ren, Xuqian, et al.
Publicado: (2024)
Boosting Multi-View Indoor 3D Object Detection via Adaptive 3D Volume Construction
por: Zhang, Runmin, et al.
Publicado: (2025)
por: Zhang, Runmin, et al.
Publicado: (2025)
Synchformer: Efficient Synchronization from Sparse Cues
por: Iashin, Vladimir, et al.
Publicado: (2024)
por: Iashin, Vladimir, et al.
Publicado: (2024)
Gaussian Splatting on the Move: Blur and Rolling Shutter Compensation for Natural Camera Motion
por: Seiskari, Otto, et al.
Publicado: (2024)
por: Seiskari, Otto, et al.
Publicado: (2024)
From Flatland to Space: Teaching Vision-Language Models to Perceive and Reason in 3D
por: Zhang, Jiahui, et al.
Publicado: (2025)
por: Zhang, Jiahui, et al.
Publicado: (2025)
V2V3D: View-to-View Denoised 3D Reconstruction for Light-Field Microscopy
por: Zhao, Jiayin, et al.
Publicado: (2025)
por: Zhao, Jiayin, et al.
Publicado: (2025)
View-on-Graph: Zero-shot 3D Visual Grounding via Vision-Language Reasoning on Scene Graphs
por: Liu, Yuanyuan, et al.
Publicado: (2025)
por: Liu, Yuanyuan, et al.
Publicado: (2025)
Neural Radiance and Gaze Fields for Visual Attention Modeling in 3D Environments
por: Chubarau, Andrei, et al.
Publicado: (2025)
por: Chubarau, Andrei, et al.
Publicado: (2025)
3D Reconstruction and New View Synthesis of Indoor Environments based on a Dual Neural Radiance Field
por: Bao, Zhenyu, et al.
Publicado: (2024)
por: Bao, Zhenyu, et al.
Publicado: (2024)
Vision Remember: Recovering Visual Information in Efficient LVLM with Vision Feature Resampling
por: Feng, Ze, et al.
Publicado: (2025)
por: Feng, Ze, et al.
Publicado: (2025)
RCNet: Deep Recurrent Collaborative Network for Multi-View Low-Light Image Enhancement
por: Luo, Hao, et al.
Publicado: (2024)
por: Luo, Hao, et al.
Publicado: (2024)
A Neural Field-Based Approach for View Computation & Data Exploration in 3D Urban Environments
por: Cobeli, Stefan, et al.
Publicado: (2025)
por: Cobeli, Stefan, et al.
Publicado: (2025)
ViewSRD: 3D Visual Grounding via Structured Multi-View Decomposition
por: Huang, Ronggang, et al.
Publicado: (2025)
por: Huang, Ronggang, et al.
Publicado: (2025)
MCGS: Multiview Consistency Enhancement for Sparse-View 3D Gaussian Radiance Fields
por: Xiao, Yuru, et al.
Publicado: (2024)
por: Xiao, Yuru, et al.
Publicado: (2024)
ViPOcc: Leveraging Visual Priors from Vision Foundation Models for Single-View 3D Occupancy Prediction
por: Feng, Yi, et al.
Publicado: (2024)
por: Feng, Yi, et al.
Publicado: (2024)
Reinforced Embodied Active Defense: Exploiting Adaptive Interaction for Robust Visual Perception in Adversarial 3D Environments
por: Yang, Xiao, et al.
Publicado: (2025)
por: Yang, Xiao, et al.
Publicado: (2025)
GS-Occ3D: Scaling Vision-only Occupancy Reconstruction with Gaussian Splatting
por: Ye, Baijun, et al.
Publicado: (2025)
por: Ye, Baijun, et al.
Publicado: (2025)
A Modular Framework for Single-View 3D Reconstruction of Indoor Environments
por: Li, Yuxiao
Publicado: (2025)
por: Li, Yuxiao
Publicado: (2025)
Polar Parametrization for Vision-based Surround-View 3D Detection
por: Chen, Shaoyu, et al.
Publicado: (2022)
por: Chen, Shaoyu, et al.
Publicado: (2022)
BabyVision: Visual Reasoning Beyond Language
por: Chen, Liang, et al.
Publicado: (2026)
por: Chen, Liang, et al.
Publicado: (2026)
Beyond Visual Cues: Synchronously Exploring Target-Centric Semantics for Vision-Language Tracking
por: Ge, Jiawei, et al.
Publicado: (2023)
por: Ge, Jiawei, et al.
Publicado: (2023)
Agent Skills Should Go Beyond Text: The Case for Visual Skills
por: Xu, Binxiao, et al.
Publicado: (2026)
por: Xu, Binxiao, et al.
Publicado: (2026)
Toward Ambulatory Vision: Learning Visually-Grounded Active View Selection
por: Koo, Juil, et al.
Publicado: (2025)
por: Koo, Juil, et al.
Publicado: (2025)
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
por: Huang, Tianyu, et al.
Publicado: (2023)
por: Huang, Tianyu, et al.
Publicado: (2023)
Cog3DMap: Multi-View Vision-Language Reasoning with 3D Cognitive Maps
por: Gwak, Chanyoung, et al.
Publicado: (2026)
por: Gwak, Chanyoung, et al.
Publicado: (2026)
Perceiving Beyond Language Priors: Enhancing Visual Comprehension and Attention in Multimodal Models
por: Ghatkesar, Aarti, et al.
Publicado: (2025)
por: Ghatkesar, Aarti, et al.
Publicado: (2025)
PARTFIELD: Learning 3D Feature Fields for Part Segmentation and Beyond
por: Liu, Minghua, et al.
Publicado: (2025)
por: Liu, Minghua, et al.
Publicado: (2025)
FoVA-Depth: Field-of-View Agnostic Depth Estimation for Cross-Dataset Generalization
por: Lichy, Daniel, et al.
Publicado: (2024)
por: Lichy, Daniel, et al.
Publicado: (2024)
Ejemplares similares
-
3D Gaussian Splatting with Fisheye Images: Field of View Analysis and Depth-Based Initialization
por: Gunes, Ulas, et al.
Publicado: (2025) -
GS-Pose: Generalizable Segmentation-based 6D Object Pose Estimation with 3D Gaussian Splatting
por: Cai, Dingding, et al.
Publicado: (2024) -
PanDepth: Joint Panoptic Segmentation and Depth Completion
por: Lagos, Juan, et al.
Publicado: (2022) -
The Weighting Game: Evaluating Quality of Explainability Methods
por: Raatikainen, Lassi, et al.
Publicado: (2022) -
MuSHRoom: Multi-Sensor Hybrid Room Dataset for Joint 3D Reconstruction and Novel View Synthesis
por: Ren, Xuqian, et al.
Publicado: (2023)