Monocular Per-Object Distance Estimation with Masked Object Modeling
Fuente:
arXiv
Guardado en:
| Autores principales: | Panariello, Aniello, Mancusi, Gianluca, Ali, Fedy Haj, Porrello, Angelo, Calderara, Simone, Cucchiara, Rita |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Is Multiple Object Tracking a Matter of Specialization?
por: Mancusi, Gianluca, et al.
Publicado: (2024)
por: Mancusi, Gianluca, et al.
Publicado: (2024)
DitHub: A Modular Framework for Incremental Open-Vocabulary Object Detection
por: Cappellino, Chiara, et al.
Publicado: (2025)
por: Cappellino, Chiara, et al.
Publicado: (2025)
Mask and Compress: Efficient Skeleton-based Action Recognition in Continual Learning
por: Mosconi, Matteo, et al.
Publicado: (2024)
por: Mosconi, Matteo, et al.
Publicado: (2024)
Zero-Shot Synthetic-to-Real Handwritten Text Recognition via Task Analogies
por: Garrido-Munoz, Carlos, et al.
Publicado: (2026)
por: Garrido-Munoz, Carlos, et al.
Publicado: (2026)
Transporting Task Vectors across Different Architectures without Training
por: Rinaldi, Filippo, et al.
Publicado: (2026)
por: Rinaldi, Filippo, et al.
Publicado: (2026)
Gradient-Sign Masking for Task Vector Transport Across Pre-Trained Models
por: Rinaldi, Filippo, et al.
Publicado: (2025)
por: Rinaldi, Filippo, et al.
Publicado: (2025)
Modular Embedding Recomposition for Incremental Learning
por: Panariello, Aniello, et al.
Publicado: (2025)
por: Panariello, Aniello, et al.
Publicado: (2025)
CLIP with Generative Latent Replay: a Strong Baseline for Incremental Learning
por: Frascaroli, Emanuele, et al.
Publicado: (2024)
por: Frascaroli, Emanuele, et al.
Publicado: (2024)
Trajectory Forecasting through Low-Rank Adaptation of Discrete Latent Codes
por: Benaglia, Riccardo, et al.
Publicado: (2024)
por: Benaglia, Riccardo, et al.
Publicado: (2024)
Accurate and Efficient Low-Rank Model Merging in Core Space
por: Panariello, Aniello, et al.
Publicado: (2025)
por: Panariello, Aniello, et al.
Publicado: (2025)
ABRA: Teleporting Fine-Tuned Knowledge Across Domains for Open-Vocabulary Object Detection
por: Bernardi, Mattia, et al.
Publicado: (2026)
por: Bernardi, Mattia, et al.
Publicado: (2026)
How to Train Your Metamorphic Deep Neural Network
por: Sommariva, Thomas, et al.
Publicado: (2025)
por: Sommariva, Thomas, et al.
Publicado: (2025)
An Attention-based Representation Distillation Baseline for Multi-Label Continual Learning
por: Menabue, Martin, et al.
Publicado: (2024)
por: Menabue, Martin, et al.
Publicado: (2024)
Selective Attention-based Modulation for Continual Learning
por: Bellitto, Giovanni, et al.
Publicado: (2024)
por: Bellitto, Giovanni, et al.
Publicado: (2024)
Rethinking Transparent Object Grasping: Depth Completion with Monocular Depth Estimation and Instance Mask
por: Cheng, Yaofeng, et al.
Publicado: (2025)
por: Cheng, Yaofeng, et al.
Publicado: (2025)
May the Forgetting Be with You: Alternate Replay for Learning with Noisy Labels
por: Millunzi, Monica, et al.
Publicado: (2024)
por: Millunzi, Monica, et al.
Publicado: (2024)
Personalized Instance-based Navigation Toward User-Specific Objects in Realistic Environments
por: Barsellotti, Luca, et al.
Publicado: (2024)
por: Barsellotti, Luca, et al.
Publicado: (2024)
DMODE: Differential Monocular Object Distance Estimation Module without Class Specific Information
por: Agand, Pedram, et al.
Publicado: (2022)
por: Agand, Pedram, et al.
Publicado: (2022)
Challenges for Monocular 6D Object Pose Estimation in Robotics
por: Thalhammer, Stefan, et al.
Publicado: (2023)
por: Thalhammer, Stefan, et al.
Publicado: (2023)
Object Gaussian for Monocular 6D Pose Estimation from Sparse Views
por: Luo, Luqing, et al.
Publicado: (2024)
por: Luo, Luqing, et al.
Publicado: (2024)
CVAM-Pose: Conditional Variational Autoencoder for Multi-Object Monocular Pose Estimation
por: Zhao, Jianyu, et al.
Publicado: (2024)
por: Zhao, Jianyu, et al.
Publicado: (2024)
Monocular Depth Estimation and Segmentation for Transparent Object with Iterative Semantic and Geometric Fusion
por: Liu, Jiangyuan, et al.
Publicado: (2025)
por: Liu, Jiangyuan, et al.
Publicado: (2025)
Mapping High-level Semantic Regions in Indoor Environments without Object Recognition
por: Bigazzi, Roberto, et al.
Publicado: (2024)
por: Bigazzi, Roberto, et al.
Publicado: (2024)
Mask6D: Masked Pose Priors For 6D Object Pose Estimation
por: Xie, Yuechen, et al.
Publicado: (2025)
por: Xie, Yuechen, et al.
Publicado: (2025)
Active Event Alignment for Monocular Distance Estimation
por: Cai, Nan, et al.
Publicado: (2024)
por: Cai, Nan, et al.
Publicado: (2024)
Simultaneous Multiple Object Detection and Pose Estimation using 3D Model Infusion with Monocular Vision
por: Li, Congliang, et al.
Publicado: (2022)
por: Li, Congliang, et al.
Publicado: (2022)
Generalizing Monocular 3D Object Detection
por: Kumar, Abhinav
Publicado: (2025)
por: Kumar, Abhinav
Publicado: (2025)
A Unified Masked Jigsaw Puzzle Framework for Vision and Language Models
por: Ye, Weixin, et al.
Publicado: (2026)
por: Ye, Weixin, et al.
Publicado: (2026)
Monocular Human-Object Reconstruction in the Wild
por: Huo, Chaofan, et al.
Publicado: (2024)
por: Huo, Chaofan, et al.
Publicado: (2024)
Robust Object Detection with Pseudo Labels from VLMs using Per-Object Co-teaching
por: Bhaskar, Uday, et al.
Publicado: (2025)
por: Bhaskar, Uday, et al.
Publicado: (2025)
MonoCLUE : Object-Aware Clustering Enhances Monocular 3D Object Detection
por: Yang, Sunghun, et al.
Publicado: (2025)
por: Yang, Sunghun, et al.
Publicado: (2025)
OBMO: One Bounding Box Multiple Objects for Monocular 3D Object Detection
por: Huang, Chenxi, et al.
Publicado: (2022)
por: Huang, Chenxi, et al.
Publicado: (2022)
MaskHOI: Robust 3D Hand-Object Interaction Estimation via Masked Pre-training
por: Xie, Yuechen, et al.
Publicado: (2025)
por: Xie, Yuechen, et al.
Publicado: (2025)
Open Vocabulary Monocular 3D Object Detection
por: Yao, Jin, et al.
Publicado: (2024)
por: Yao, Jin, et al.
Publicado: (2024)
HORT: Monocular Hand-held Objects Reconstruction with Transformers
por: Chen, Zerui, et al.
Publicado: (2025)
por: Chen, Zerui, et al.
Publicado: (2025)
Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization
por: Compagnoni, Alberto, et al.
Publicado: (2025)
por: Compagnoni, Alberto, et al.
Publicado: (2025)
Mask Consistency Regularization in Object Removal
por: Yuan, Hua, et al.
Publicado: (2025)
por: Yuan, Hua, et al.
Publicado: (2025)
Representation Based Regression for Object Distance Estimation
por: Ahishali, Mete, et al.
Publicado: (2021)
por: Ahishali, Mete, et al.
Publicado: (2021)
SingRef6D: Monocular Novel Object Pose Estimation with a Single RGB Reference
por: Wang, Jiahui, et al.
Publicado: (2025)
por: Wang, Jiahui, et al.
Publicado: (2025)
TexHOI: Reconstructing Textures of 3D Unknown Objects in Monocular Hand-Object Interaction Scenes
por: Aggarwal, Alakh, et al.
Publicado: (2025)
por: Aggarwal, Alakh, et al.
Publicado: (2025)
Ejemplares similares
-
Is Multiple Object Tracking a Matter of Specialization?
por: Mancusi, Gianluca, et al.
Publicado: (2024) -
DitHub: A Modular Framework for Incremental Open-Vocabulary Object Detection
por: Cappellino, Chiara, et al.
Publicado: (2025) -
Mask and Compress: Efficient Skeleton-based Action Recognition in Continual Learning
por: Mosconi, Matteo, et al.
Publicado: (2024) -
Zero-Shot Synthetic-to-Real Handwritten Text Recognition via Task Analogies
por: Garrido-Munoz, Carlos, et al.
Publicado: (2026) -
Transporting Task Vectors across Different Architectures without Training
por: Rinaldi, Filippo, et al.
Publicado: (2026)