DADO: A Depth-Attention framework for Object Discovery
Fuente:
arXiv
Salvato in:
| Autori principali: | Gonzalez, Federico, Talavera, Estefania, Radeva, Petia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Slot Attention-based Feature Filtering for Few-Shot Learning
di: Rodenas, Javier, et al.
Pubblicazione: (2025)
di: Rodenas, Javier, et al.
Pubblicazione: (2025)
Multi-label out-of-distribution detection via evidential learning
di: Aguilar, Eduardo, et al.
Pubblicazione: (2025)
di: Aguilar, Eduardo, et al.
Pubblicazione: (2025)
Stochastic-based Patch Filtering for Few-Shot Learning
di: Rodenas, Javier, et al.
Pubblicazione: (2025)
di: Rodenas, Javier, et al.
Pubblicazione: (2025)
MomentsNeRF: Leveraging Orthogonal Moments for Few-Shot Neural Rendering
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2024)
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2024)
VolETA: One- and Few-shot Food Volume Estimation
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2024)
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2024)
MVSBoost: An Efficient Point Cloud-based 3D Reconstruction
di: Haroon, Umair, et al.
Pubblicazione: (2024)
di: Haroon, Umair, et al.
Pubblicazione: (2024)
FoodMem: Near Real-time and Precise Food Video Segmentation
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2024)
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2024)
VolTex: Food Volume Estimation using Text-Guided Segmentation and Neural Surface Reconstruction
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2025)
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2025)
VolE: A Point-cloud Framework for Food 3D Reconstruction and Volume Estimation
di: Haroon, Umair, et al.
Pubblicazione: (2025)
di: Haroon, Umair, et al.
Pubblicazione: (2025)
Advancing Image-Based Grapevine Variety Classification with a New Benchmark and Evaluation of Masked Autoencoders
di: Carneiro, Gabriel A., et al.
Pubblicazione: (2025)
di: Carneiro, Gabriel A., et al.
Pubblicazione: (2025)
All4One: Symbiotic Neighbour Contrastive Learning via Self-Attention and Redundancy Reduction
di: Estepa, Imanol G., et al.
Pubblicazione: (2023)
di: Estepa, Imanol G., et al.
Pubblicazione: (2023)
PerBite: A Curated Diagnostic Workflow for Bite-Aware Food Volume Estimation
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2026)
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2026)
Explainable AI model reveals disease-related mechanisms in single-cell RNA-seq data
di: Usman, Mohammad, et al.
Pubblicazione: (2025)
di: Usman, Mohammad, et al.
Pubblicazione: (2025)
LAND: Lung and Nodule Diffusion for 3D Chest CT Synthesis with Anatomical Guidance
di: Oliveras, Anna, et al.
Pubblicazione: (2025)
di: Oliveras, Anna, et al.
Pubblicazione: (2025)
BenchSeg: A Large-Scale Dataset and Benchmark for Multi-View Food Video Segmentation
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2026)
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2026)
Hybrid deep learning-based strategy for the hepatocellular carcinoma cancer grade classification of H&E stained liver histopathology images
di: Deshpande, Ajinkya, et al.
Pubblicazione: (2024)
di: Deshpande, Ajinkya, et al.
Pubblicazione: (2024)
Indoor scene recognition from images under visual corruptions
di: Costa, Willams de Lima, et al.
Pubblicazione: (2024)
di: Costa, Willams de Lima, et al.
Pubblicazione: (2024)
Representation Discrepancy Bridging Method for Remote Sensing Image-Text Retrieval
di: Ning, Hailong, et al.
Pubblicazione: (2025)
di: Ning, Hailong, et al.
Pubblicazione: (2025)
Transparent Object Depth Completion
di: Zhou, Yifan, et al.
Pubblicazione: (2024)
di: Zhou, Yifan, et al.
Pubblicazione: (2024)
Adaptive Slot Attention: Object Discovery with Dynamic Slot Number
di: Fan, Ke, et al.
Pubblicazione: (2024)
di: Fan, Ke, et al.
Pubblicazione: (2024)
ST-Gait++: Leveraging spatio-temporal convolutions for gait-based emotion recognition on videos
di: Lima, Maria Luísa, et al.
Pubblicazione: (2024)
di: Lima, Maria Luísa, et al.
Pubblicazione: (2024)
Self-Supervised Learning for Transparent Object Depth Completion Using Depth from Non-Transparent Objects
di: Fan, Xianghui, et al.
Pubblicazione: (2025)
di: Fan, Xianghui, et al.
Pubblicazione: (2025)
Depth Awakens: A Depth-perceptual Attention Fusion Network for RGB-D Camouflaged Object Detection
di: Liua, Xinran, et al.
Pubblicazione: (2024)
di: Liua, Xinran, et al.
Pubblicazione: (2024)
Depth as Prior Knowledge for Object Detection
di: Sbeyti, Moussa Kassem, et al.
Pubblicazione: (2026)
di: Sbeyti, Moussa Kassem, et al.
Pubblicazione: (2026)
Learning from Semantic Dictionaries: Discriminative Codebook Contrastive Learning for Unified Visual Representation and Generation
di: Estepa, Imanol G., et al.
Pubblicazione: (2026)
di: Estepa, Imanol G., et al.
Pubblicazione: (2026)
DepthMOT: Depth Cues Lead to a Strong Multi-Object Tracker
di: Wu, Jiapeng, et al.
Pubblicazione: (2024)
di: Wu, Jiapeng, et al.
Pubblicazione: (2024)
Integrating Object Detection, LiDAR-Enhanced Depth Estimation, and Segmentation Models for Railway Environments
di: Giannico, Enrico Francesco, et al.
Pubblicazione: (2026)
di: Giannico, Enrico Francesco, et al.
Pubblicazione: (2026)
Gated Cross-Attention Network for Depth Completion
di: Jia, Xiaogang, et al.
Pubblicazione: (2023)
di: Jia, Xiaogang, et al.
Pubblicazione: (2023)
Enhancing Object Discovery for Unsupervised Instance Segmentation and Object Detection
di: Feng, Xingyu, et al.
Pubblicazione: (2025)
di: Feng, Xingyu, et al.
Pubblicazione: (2025)
Rethinking Transparent Object Grasping: Depth Completion with Monocular Depth Estimation and Instance Mask
di: Cheng, Yaofeng, et al.
Pubblicazione: (2025)
di: Cheng, Yaofeng, et al.
Pubblicazione: (2025)
DepthFlow: Exploiting Depth-Flow Structural Correlations for Unsupervised Video Object Segmentation
di: Cho, Suhwan, et al.
Pubblicazione: (2025)
di: Cho, Suhwan, et al.
Pubblicazione: (2025)
Learning Object Focused Attention
di: Trivedy, Vivek, et al.
Pubblicazione: (2025)
di: Trivedy, Vivek, et al.
Pubblicazione: (2025)
Attention Is All You Need For Mixture-of-Depths Routing
di: Gadhikar, Advait, et al.
Pubblicazione: (2024)
di: Gadhikar, Advait, et al.
Pubblicazione: (2024)
Pyramid Feature Attention Network for Monocular Depth Prediction
di: Xu, Yifang, et al.
Pubblicazione: (2024)
di: Xu, Yifang, et al.
Pubblicazione: (2024)
Diffusion-Based Depth Inpainting for Transparent and Reflective Objects
di: Sun, Tianyu, et al.
Pubblicazione: (2024)
di: Sun, Tianyu, et al.
Pubblicazione: (2024)
Ensemble Foreground Management for Unsupervised Object Discovery
di: Wu, Ziling, et al.
Pubblicazione: (2025)
di: Wu, Ziling, et al.
Pubblicazione: (2025)
Unsupervised Discovery of Object-Centric Neural Fields
di: Luo, Rundong, et al.
Pubblicazione: (2024)
di: Luo, Rundong, et al.
Pubblicazione: (2024)
Unsupervised Object Discovery: A Comprehensive Survey and Unified Taxonomy
di: Villa-Vásquez, José-Fabian, et al.
Pubblicazione: (2024)
di: Villa-Vásquez, José-Fabian, et al.
Pubblicazione: (2024)
CylinderDepth: Cylindrical Spatial Attention for Multi-View Consistent Self-Supervised Surround Depth Estimation
di: Abualhanud, Samer, et al.
Pubblicazione: (2025)
di: Abualhanud, Samer, et al.
Pubblicazione: (2025)
Masked Multi-Query Slot Attention for Unsupervised Object Discovery
di: Pramanik, Rishav, et al.
Pubblicazione: (2024)
di: Pramanik, Rishav, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Slot Attention-based Feature Filtering for Few-Shot Learning
di: Rodenas, Javier, et al.
Pubblicazione: (2025) -
Multi-label out-of-distribution detection via evidential learning
di: Aguilar, Eduardo, et al.
Pubblicazione: (2025) -
Stochastic-based Patch Filtering for Few-Shot Learning
di: Rodenas, Javier, et al.
Pubblicazione: (2025) -
MomentsNeRF: Leveraging Orthogonal Moments for Few-Shot Neural Rendering
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2024) -
VolETA: One- and Few-shot Food Volume Estimation
di: AlMughrabi, Ahmad, et al.
Pubblicazione: (2024)