Spatially Optimized Compact Deep Metric Learning Model for Similarity Search
Fuente:
arXiv
Guardado en:
| Autores principales: | Islam, Md. Farhadul, Reza, Md. Tanzim, Manab, Meem Arafat, Mahin, Mohammad Rakibul Hasan, Zabeen, Sarah, Noor, Jannatun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Med-IC: Fusing a Single Layer Involution with Convolutions for Enhanced Medical Image Classification and Segmentation
por: Islam, Md. Farhadul, et al.
Publicado: (2024)
por: Islam, Md. Farhadul, et al.
Publicado: (2024)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
por: Raoufi, Behnam, et al.
Publicado: (2025)
por: Raoufi, Behnam, et al.
Publicado: (2025)
Trapped in texture bias? A large scale comparison of deep instance segmentation
por: Theodoridis, Johannes, et al.
Publicado: (2024)
por: Theodoridis, Johannes, et al.
Publicado: (2024)
Combined Hyperbolic and Euclidean Soft Triple Loss Beyond the Single Space Deep Metric Learning
por: Saeki, Shozo, et al.
Publicado: (2025)
por: Saeki, Shozo, et al.
Publicado: (2025)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
por: Gupta, Sunny, et al.
Publicado: (2024)
por: Gupta, Sunny, et al.
Publicado: (2024)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
por: Kashyap, Pankhi, et al.
Publicado: (2024)
por: Kashyap, Pankhi, et al.
Publicado: (2024)
Quantized FCA: Efficient Zero-Shot Texture Anomaly Detection
por: Ardelean, Andrei-Timotei, et al.
Publicado: (2025)
por: Ardelean, Andrei-Timotei, et al.
Publicado: (2025)
DSER: Spectral Epipolar Representation for Efficient Light Field Depth Estimation
por: Mohammad, Noor Islam S., et al.
Publicado: (2025)
por: Mohammad, Noor Islam S., et al.
Publicado: (2025)
Normalizing Flow-Based Metric for Image Generation
por: Jeevan, Pranav, et al.
Publicado: (2024)
por: Jeevan, Pranav, et al.
Publicado: (2024)
Hierarchical Spatial Algorithms for High-Resolution Image Quantization and Feature Extraction
por: Mohammad, Noor Islam S.
Publicado: (2025)
por: Mohammad, Noor Islam S.
Publicado: (2025)
FLD+: Data-efficient Evaluation Metric for Generative Models
por: Jeevan, Pranav, et al.
Publicado: (2024)
por: Jeevan, Pranav, et al.
Publicado: (2024)
CNN-based local features for navigation near an asteroid
por: Knuuttila, Olli, et al.
Publicado: (2023)
por: Knuuttila, Olli, et al.
Publicado: (2023)
MaSC: A Masked Similarity Metric for Evaluating Concept-Driven Generation
por: Bartkowiak, Patryk, et al.
Publicado: (2026)
por: Bartkowiak, Patryk, et al.
Publicado: (2026)
Sign language recognition based on deep learning and low-cost handcrafted descriptors
por: Carneiro, Alvaro Leandro Cavalcante, et al.
Publicado: (2024)
por: Carneiro, Alvaro Leandro Cavalcante, et al.
Publicado: (2024)
EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents
por: Riva, Paolo, et al.
Publicado: (2026)
por: Riva, Paolo, et al.
Publicado: (2026)
CytoNet: A Foundation Model for the Human Cerebral Cortex at Cellular Resolution
por: Schiffer, Christian, et al.
Publicado: (2025)
por: Schiffer, Christian, et al.
Publicado: (2025)
Spatially-Aware Speaker for Vision-and-Language Navigation Instruction Generation
por: Gopinathan, Muraleekrishna, et al.
Publicado: (2024)
por: Gopinathan, Muraleekrishna, et al.
Publicado: (2024)
WaveMix: A Resource-efficient Neural Network for Image Analysis
por: Jeevan, Pranav, et al.
Publicado: (2022)
por: Jeevan, Pranav, et al.
Publicado: (2022)
Which Backbone to Use: A Resource-efficient Domain Specific Comparison for Computer Vision
por: Jeevan, Pranav, et al.
Publicado: (2024)
por: Jeevan, Pranav, et al.
Publicado: (2024)
Quantifying and Inducing Shape Bias in CNNs via Max-Pool Dilation
por: Sawada, Takito, et al.
Publicado: (2026)
por: Sawada, Takito, et al.
Publicado: (2026)
Interpreting Video Representations with Spatio-Temporal Sparse Autoencoders
por: Dokme, Atahan, et al.
Publicado: (2026)
por: Dokme, Atahan, et al.
Publicado: (2026)
Vision transformers in domain adaptation and domain generalization: a study of robustness
por: Alijani, Shadi, et al.
Publicado: (2024)
por: Alijani, Shadi, et al.
Publicado: (2024)
U-Net-Like Spiking Neural Networks for Single Image Dehazing
por: Li, Huibin, et al.
Publicado: (2025)
por: Li, Huibin, et al.
Publicado: (2025)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
por: Li, Danyang, et al.
Publicado: (2025)
por: Li, Danyang, et al.
Publicado: (2025)
Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight
por: Romero, Angel, et al.
Publicado: (2025)
por: Romero, Angel, et al.
Publicado: (2025)
WayFASTER: a Self-Supervised Traversability Prediction for Increased Navigation Awareness
por: Gasparino, Mateus Valverde, et al.
Publicado: (2024)
por: Gasparino, Mateus Valverde, et al.
Publicado: (2024)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
por: Lim, Shoon Kit, et al.
Publicado: (2025)
por: Lim, Shoon Kit, et al.
Publicado: (2025)
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
por: Gopinathan, Muraleekrishna, et al.
Publicado: (2024)
por: Gopinathan, Muraleekrishna, et al.
Publicado: (2024)
Deep Probabilistic Traversability with Test-time Adaptation for Uncertainty-aware Planetary Rover Navigation
por: Endo, Masafumi, et al.
Publicado: (2024)
por: Endo, Masafumi, et al.
Publicado: (2024)
CoMoCAVs: Cohesive Decision-Guided Motion Planning for Connected and Autonomous Vehicles with Multi-Policy Reinforcement Learning
por: Hu, Pan
Publicado: (2025)
por: Hu, Pan
Publicado: (2025)
Single-Shot Metric Depth from Focused Plenoptic Cameras
por: Lasheras-Hernandez, Blanca, et al.
Publicado: (2024)
por: Lasheras-Hernandez, Blanca, et al.
Publicado: (2024)
Motion Perceiver: Real-Time Occupancy Forecasting for Embedded Systems
por: Ferenczi, Bryce, et al.
Publicado: (2023)
por: Ferenczi, Bryce, et al.
Publicado: (2023)
WaveMixSR-V2: Enhancing Super-resolution with Higher Efficiency
por: Jeevan, Pranav, et al.
Publicado: (2024)
por: Jeevan, Pranav, et al.
Publicado: (2024)
SUN Team's Contribution to ABAW 2024 Competition: Audio-visual Valence-Arousal Estimation and Expression Recognition
por: Dresvyanskiy, Denis, et al.
Publicado: (2024)
por: Dresvyanskiy, Denis, et al.
Publicado: (2024)
Learning the meanings of function words from grounded language using a visual question answering model
por: Portelance, Eva, et al.
Publicado: (2023)
por: Portelance, Eva, et al.
Publicado: (2023)
Closed-Loop Neural Activation Control in Vision-Language-Action Models
por: Babu, Abhijith, et al.
Publicado: (2026)
por: Babu, Abhijith, et al.
Publicado: (2026)
Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
por: Yang, Shan
Publicado: (2026)
por: Yang, Shan
Publicado: (2026)
A Segmented Robot Grasping Perception Neural Network for Edge AI
por: Bröcheler, Casper, et al.
Publicado: (2025)
por: Bröcheler, Casper, et al.
Publicado: (2025)
Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention
por: Durrani, Hamza Ahmed, et al.
Publicado: (2026)
por: Durrani, Hamza Ahmed, et al.
Publicado: (2026)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
por: Tourani, Ali, et al.
Publicado: (2023)
por: Tourani, Ali, et al.
Publicado: (2023)
Ejemplares similares
-
Med-IC: Fusing a Single Layer Involution with Convolutions for Enhanced Medical Image Classification and Segmentation
por: Islam, Md. Farhadul, et al.
Publicado: (2024) -
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
por: Raoufi, Behnam, et al.
Publicado: (2025) -
Trapped in texture bias? A large scale comparison of deep instance segmentation
por: Theodoridis, Johannes, et al.
Publicado: (2024) -
Combined Hyperbolic and Euclidean Soft Triple Loss Beyond the Single Space Deep Metric Learning
por: Saeki, Shozo, et al.
Publicado: (2025) -
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
por: Gupta, Sunny, et al.
Publicado: (2024)