Deep Models for Multi-View 3D Object Recognition: A Review
Fuente:
arXiv
Guardado en:
| Autores principales: | Alzahrani, Mona, Usman, Muhammad, Kammoun, Salma, Anwar, Saeed, Helmy, Tarek |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Bird Eye-View to Street-View: A Survey
por: Bajbaa, Khawlah, et al.
Publicado: (2024)
por: Bajbaa, Khawlah, et al.
Publicado: (2024)
MSRNet: A Multi-Scale Recursive Network for Camouflaged Object Detection
por: Alghamdi, Leena, et al.
Publicado: (2025)
por: Alghamdi, Leena, et al.
Publicado: (2025)
Attention Down-Sampling Transformer, Relative Ranking and Self-Consistency for Blind Image Quality Assessment
por: Alsaafin, Mohammed, et al.
Publicado: (2024)
por: Alsaafin, Mohammed, et al.
Publicado: (2024)
Multistream Network for LiDAR and Camera-based 3D Object Detection in Outdoor Scenes
por: Ibrahim, Muhammad, et al.
Publicado: (2025)
por: Ibrahim, Muhammad, et al.
Publicado: (2025)
DefAn: Definitive Answer Dataset for LLMs Hallucination Evaluation
por: Rahman, A B M Ashikur, et al.
Publicado: (2024)
por: Rahman, A B M Ashikur, et al.
Publicado: (2024)
Hybrid Deep Learning for Hyperspectral Single Image Super-Resolution
por: Muhammad, Usman, et al.
Publicado: (2025)
por: Muhammad, Usman, et al.
Publicado: (2025)
Improved Crop and Weed Detection with Diverse Data Ensemble Learning
por: Asad, Muhammad Hamza, et al.
Publicado: (2023)
por: Asad, Muhammad Hamza, et al.
Publicado: (2023)
From Satellite to Street: A Hybrid Framework Integrating Stable Diffusion and PanoGAN for Consistent Cross-View Synthesis
por: Bajbaa, Khawlah, et al.
Publicado: (2025)
por: Bajbaa, Khawlah, et al.
Publicado: (2025)
BioPose: Biomechanically-accurate 3D Pose Estimation from Monocular Videos
por: Koleini, Farnoosh, et al.
Publicado: (2025)
por: Koleini, Farnoosh, et al.
Publicado: (2025)
DreamComposer: Controllable 3D Object Generation via Multi-View Conditions
por: Yang, Yunhan, et al.
Publicado: (2023)
por: Yang, Yunhan, et al.
Publicado: (2023)
What Matters in Range View 3D Object Detection
por: Wilson, Benjamin, et al.
Publicado: (2024)
por: Wilson, Benjamin, et al.
Publicado: (2024)
Are Deep Learning Models Robust to Partial Object Occlusion in Visual Recognition Tasks?
por: Kassaw, Kaleb, et al.
Publicado: (2024)
por: Kassaw, Kaleb, et al.
Publicado: (2024)
MuM: Multi-View Masked Image Modeling for 3D Vision
por: Nordström, David, et al.
Publicado: (2025)
por: Nordström, David, et al.
Publicado: (2025)
VolTA-3D: Self-Supervised Learning for Brain MRI using 3D Volumetric Token Alignment
por: Makawana, Amy, et al.
Publicado: (2026)
por: Makawana, Amy, et al.
Publicado: (2026)
Crop Pest Classification Using Deep Learning Techniques: A Review
por: Ejaz, Muhammad Hassam, et al.
Publicado: (2025)
por: Ejaz, Muhammad Hassam, et al.
Publicado: (2025)
A Review on Coarse to Fine-Grained Animal Action Recognition
por: Zia, Ali, et al.
Publicado: (2025)
por: Zia, Ali, et al.
Publicado: (2025)
Multi-View 3D Reconstruction using Knowledge Distillation
por: Dutt, Aditya, et al.
Publicado: (2024)
por: Dutt, Aditya, et al.
Publicado: (2024)
3D-Consistent Multi-View Editing by Correspondence Guidance
por: Bengtson, Josef, et al.
Publicado: (2025)
por: Bengtson, Josef, et al.
Publicado: (2025)
RDD4D: 4D Attention-Guided Road Damage Detection And Classification
por: Alkalbani, Asma, et al.
Publicado: (2025)
por: Alkalbani, Asma, et al.
Publicado: (2025)
Hypergraph-based Multi-View Action Recognition using Event Cameras
por: Gao, Yue, et al.
Publicado: (2024)
por: Gao, Yue, et al.
Publicado: (2024)
MaskHand: Generative Masked Modeling for Robust Hand Mesh Reconstruction in the Wild
por: Saleem, Muhammad Usama, et al.
Publicado: (2024)
por: Saleem, Muhammad Usama, et al.
Publicado: (2024)
Ranking vs. Assignment: The Metric Mismatch in Multi-View Object Association
por: Shelukhan, Matvei, et al.
Publicado: (2026)
por: Shelukhan, Matvei, et al.
Publicado: (2026)
Neural Assets: 3D-Aware Multi-Object Scene Synthesis with Image Diffusion Models
por: Wu, Ziyi, et al.
Publicado: (2024)
por: Wu, Ziyi, et al.
Publicado: (2024)
REXO: Indoor Multi-View Radar Object Detection via 3D Bounding Box Diffusion
por: Yataka, Ryoma, et al.
Publicado: (2025)
por: Yataka, Ryoma, et al.
Publicado: (2025)
PolarBEVDet: Exploring Polar Representation for Multi-View 3D Object Detection in Bird's-Eye-View
por: Yu, Zichen, et al.
Publicado: (2024)
por: Yu, Zichen, et al.
Publicado: (2024)
Intelligent Known and Novel Aircraft Recognition -- A Shift from Classification to Similarity Learning for Combat Identification
por: Saeed, Ahmad, et al.
Publicado: (2024)
por: Saeed, Ahmad, et al.
Publicado: (2024)
Reconstruction by Generation: 3D Multi-Object Scene Reconstruction from Sparse Observations
por: Zadaianchuk, Andrii, et al.
Publicado: (2026)
por: Zadaianchuk, Andrii, et al.
Publicado: (2026)
MV-MOS: Multi-View Feature Fusion for 3D Moving Object Segmentation
por: Cheng, Jintao, et al.
Publicado: (2024)
por: Cheng, Jintao, et al.
Publicado: (2024)
Visual Attention Methods in Deep Learning: An In-Depth Survey
por: Hassanin, Mohammed, et al.
Publicado: (2022)
por: Hassanin, Mohammed, et al.
Publicado: (2022)
MaskAdapt: Unsupervised Geometry-Aware Domain Adaptation Using Multimodal Contextual Learning and RGB-Depth Masking
por: Nadeem, Numair, et al.
Publicado: (2025)
por: Nadeem, Numair, et al.
Publicado: (2025)
Segmenting Visuals With Querying Words: Language Anchors For Semi-Supervised Image Segmentation
por: Nadeem, Numair, et al.
Publicado: (2025)
por: Nadeem, Numair, et al.
Publicado: (2025)
ECOR: Explainable CLIP for Object Recognition
por: Rasekh, Ali, et al.
Publicado: (2024)
por: Rasekh, Ali, et al.
Publicado: (2024)
Deep Learning-Based 3D Instance and Semantic Segmentation: A Review
por: Yasir, Siddiqui Muhammad, et al.
Publicado: (2024)
por: Yasir, Siddiqui Muhammad, et al.
Publicado: (2024)
YOLOatr : Deep Learning Based Automatic Target Detection and Localization in Thermal Infrared Imagery
por: Safdar, Aon, et al.
Publicado: (2025)
por: Safdar, Aon, et al.
Publicado: (2025)
Visual Sync: Multi-Camera Synchronization via Cross-View Object Motion
por: Liu, Shaowei, et al.
Publicado: (2025)
por: Liu, Shaowei, et al.
Publicado: (2025)
HOT3D: Hand and Object Tracking in 3D from Egocentric Multi-View Videos
por: Banerjee, Prithviraj, et al.
Publicado: (2024)
por: Banerjee, Prithviraj, et al.
Publicado: (2024)
SPEGNet: Synergistic Perception-Guided Network for Camouflaged Object Detection
por: Jan, Baber, et al.
Publicado: (2025)
por: Jan, Baber, et al.
Publicado: (2025)
FQ-PETR: Fully Quantized Position Embedding Transformation for Multi-View 3D Object Detection
por: Yu, Jiangyong, et al.
Publicado: (2025)
por: Yu, Jiangyong, et al.
Publicado: (2025)
MVSA-Net: Multi-View State-Action Recognition for Robust and Deployable Trajectory Generation
por: Asali, Ehsan, et al.
Publicado: (2023)
por: Asali, Ehsan, et al.
Publicado: (2023)
Viewpoint Textual Inversion: Discovering Scene Representations and 3D View Control in 2D Diffusion Models
por: Burgess, James, et al.
Publicado: (2023)
por: Burgess, James, et al.
Publicado: (2023)
Ejemplares similares
-
Bird Eye-View to Street-View: A Survey
por: Bajbaa, Khawlah, et al.
Publicado: (2024) -
MSRNet: A Multi-Scale Recursive Network for Camouflaged Object Detection
por: Alghamdi, Leena, et al.
Publicado: (2025) -
Attention Down-Sampling Transformer, Relative Ranking and Self-Consistency for Blind Image Quality Assessment
por: Alsaafin, Mohammed, et al.
Publicado: (2024) -
Multistream Network for LiDAR and Camera-based 3D Object Detection in Outdoor Scenes
por: Ibrahim, Muhammad, et al.
Publicado: (2025) -
DefAn: Definitive Answer Dataset for LLMs Hallucination Evaluation
por: Rahman, A B M Ashikur, et al.
Publicado: (2024)