LostPaw: Finding Lost Pets using a Contrastive Learning-based Transformer with Visual Input
Fuente:
arXiv
Guardado en:
| Autores principales: | Voinea, Andrei, Kock, Robin, Dhali, Maruf A. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Oil Spill Segmentation using Deep Encoder-Decoder models
por: Satyanarayana, Abhishek Ramanathapura, et al.
Publicado: (2023)
por: Satyanarayana, Abhishek Ramanathapura, et al.
Publicado: (2023)
Large Vision-Language Models Get Lost in Attention
por: Xi, Gongli, et al.
Publicado: (2026)
por: Xi, Gongli, et al.
Publicado: (2026)
StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning
por: He, Xixiang, et al.
Publicado: (2026)
por: He, Xixiang, et al.
Publicado: (2026)
Lost in OCR Translation? Vision-Based Approaches to Robust Document Retrieval
por: Most, Alexander, et al.
Publicado: (2025)
por: Most, Alexander, et al.
Publicado: (2025)
Lost in UNet: Improving Infrared Small Target Detection by Underappreciated Local Features
por: Quan, Wuzhou, et al.
Publicado: (2024)
por: Quan, Wuzhou, et al.
Publicado: (2024)
Lost in Space: Probing Fine-grained Spatial Understanding in Vision and Language Resamplers
por: Pantazopoulos, Georgios, et al.
Publicado: (2024)
por: Pantazopoulos, Georgios, et al.
Publicado: (2024)
Lost in Time: Clock and Calendar Understanding Challenges in Multimodal LLMs
por: Saxena, Rohit, et al.
Publicado: (2025)
por: Saxena, Rohit, et al.
Publicado: (2025)
From Pixels to People: Satellite-Based Mapping and Quantification of Riverbank Erosion and Lost Villages in Bangladesh
por: Rafat, M Saifuzzaman, et al.
Publicado: (2025)
por: Rafat, M Saifuzzaman, et al.
Publicado: (2025)
Lost in Space? Vision-Language Models Struggle with Relative Camera Pose Estimation
por: Deng, Ken, et al.
Publicado: (2026)
por: Deng, Ken, et al.
Publicado: (2026)
Linear Differential Vision Transformer: Learning Visual Contrasts via Pairwise Differentials
por: Pu, Yifan, et al.
Publicado: (2025)
por: Pu, Yifan, et al.
Publicado: (2025)
Have the VLMs Lost Confidence? A Study of Sycophancy in VLMs
por: Li, Shuo, et al.
Publicado: (2024)
por: Li, Shuo, et al.
Publicado: (2024)
Lost in Edits? A $λ$-Compass for AIGC Provenance
por: You, Wenhao, et al.
Publicado: (2025)
por: You, Wenhao, et al.
Publicado: (2025)
Lost in Translation and Noise: A Deep Dive into the Failure Modes of VLMs on Real-World Tables
por: Singh, Anshul, et al.
Publicado: (2025)
por: Singh, Anshul, et al.
Publicado: (2025)
BiNet: Degraded-Manuscript Binarization in Diverse Document Textures and Layouts using Deep Encoder-Decoder Networks
por: Dhali, Maruf A., et al.
Publicado: (2019)
por: Dhali, Maruf A., et al.
Publicado: (2019)
Lost in the Hype: Revealing and Dissecting the Performance Degradation of Medical Multimodal Large Language Models in Image Classification
por: Zhu, Xun, et al.
Publicado: (2026)
por: Zhu, Xun, et al.
Publicado: (2026)
Mutual Information guided Visual Contrastive Learning
por: Chen, Hanyang, et al.
Publicado: (2025)
por: Chen, Hanyang, et al.
Publicado: (2025)
Cluster Contrast for Unsupervised Visual Representation Learning
por: Giakoumoglou, Nikolaos, et al.
Publicado: (2025)
por: Giakoumoglou, Nikolaos, et al.
Publicado: (2025)
Discovering Hidden Visual Concepts Beyond Linguistic Input in Infant Learning
por: Ke, Xueyi, et al.
Publicado: (2025)
por: Ke, Xueyi, et al.
Publicado: (2025)
VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models
por: Balakrishnan, Ravikumar, et al.
Publicado: (2025)
por: Balakrishnan, Ravikumar, et al.
Publicado: (2025)
VISOR: Visual Input-based Steering for Output Redirection in Vision-Language Models
por: Phute, Mansi, et al.
Publicado: (2025)
por: Phute, Mansi, et al.
Publicado: (2025)
A Contrastive Learning Scheme with Transformer Innate Patches
por: Jyhne, Sander Riisøen, et al.
Publicado: (2023)
por: Jyhne, Sander Riisøen, et al.
Publicado: (2023)
Enhancing Multimodal Entity Linking with Jaccard Distance-based Conditional Contrastive Learning and Contextual Visual Augmentation
por: Nguyen, Cong-Duy, et al.
Publicado: (2025)
por: Nguyen, Cong-Duy, et al.
Publicado: (2025)
Dating ancient manuscripts using radiocarbon and AI-based writing style analysis
por: Popović, Mladen, et al.
Publicado: (2024)
por: Popović, Mladen, et al.
Publicado: (2024)
Survey on Hand Gesture Recognition from Visual Input
por: Linardakis, Manousos, et al.
Publicado: (2025)
por: Linardakis, Manousos, et al.
Publicado: (2025)
Learning Transformer-based World Models with Contrastive Predictive Coding
por: Burchi, Maxime, et al.
Publicado: (2025)
por: Burchi, Maxime, et al.
Publicado: (2025)
Predictive Reasoning with Augmented Anomaly Contrastive Learning for Compositional Visual Relations
por: Li, Chengtai, et al.
Publicado: (2026)
por: Li, Chengtai, et al.
Publicado: (2026)
Variational Contrastive Learning for Skeleton-based Action Recognition
por: Nguyen, Dang Dinh, et al.
Publicado: (2026)
por: Nguyen, Dang Dinh, et al.
Publicado: (2026)
Unbiased Visual Reasoning with Controlled Visual Inputs
por: Li, Zhaonan, et al.
Publicado: (2025)
por: Li, Zhaonan, et al.
Publicado: (2025)
A Unified Framework for Microscopy Defocus Deblur with Multi-Pyramid Transformer and Contrastive Learning
por: Zhang, Yuelin, et al.
Publicado: (2024)
por: Zhang, Yuelin, et al.
Publicado: (2024)
ViTaS: Visual Tactile Soft Fusion Contrastive Learning for Visuomotor Learning
por: Tian, Yufeng, et al.
Publicado: (2026)
por: Tian, Yufeng, et al.
Publicado: (2026)
Focusing by Contrastive Attention: Enhancing VLMs' Visual Reasoning
por: Ge, Yuyao, et al.
Publicado: (2025)
por: Ge, Yuyao, et al.
Publicado: (2025)
RingMo-Aerial: An Aerial Remote Sensing Foundation Model With Affine Transformation Contrastive Learning
por: Diao, Wenhui, et al.
Publicado: (2024)
por: Diao, Wenhui, et al.
Publicado: (2024)
Evaluating Visual Explanations of Attention Maps for Transformer-based Medical Imaging
por: Chung, Minjae, et al.
Publicado: (2025)
por: Chung, Minjae, et al.
Publicado: (2025)
Gaze on the Prize: Shaping Visual Attention with Return-Guided Contrastive Learning
por: Lee, Andrew, et al.
Publicado: (2025)
por: Lee, Andrew, et al.
Publicado: (2025)
Explainable Melanoma Diagnosis with Contrastive Learning and LLM-based Report Generation
por: Zheng, Junwen, et al.
Publicado: (2025)
por: Zheng, Junwen, et al.
Publicado: (2025)
Multi-Label Contrastive Learning for Abstract Visual Reasoning
por: Małkiński, Mikołaj, et al.
Publicado: (2020)
por: Małkiński, Mikołaj, et al.
Publicado: (2020)
ConPro: Learning Severity Representation for Medical Images using Contrastive Learning and Preference Optimization
por: Nguyen, Hong, et al.
Publicado: (2024)
por: Nguyen, Hong, et al.
Publicado: (2024)
BEAT: Visual Backdoor Attacks on VLM-based Embodied Agents via Contrastive Trigger Learning
por: Zhan, Qiusi, et al.
Publicado: (2025)
por: Zhan, Qiusi, et al.
Publicado: (2025)
ConVQG: Contrastive Visual Question Generation with Multimodal Guidance
por: Mi, Li, et al.
Publicado: (2024)
por: Mi, Li, et al.
Publicado: (2024)
HCVP: Leveraging Hierarchical Contrastive Visual Prompt for Domain Generalization
por: Zhou, Guanglin, et al.
Publicado: (2024)
por: Zhou, Guanglin, et al.
Publicado: (2024)
Ejemplares similares
-
Oil Spill Segmentation using Deep Encoder-Decoder models
por: Satyanarayana, Abhishek Ramanathapura, et al.
Publicado: (2023) -
Large Vision-Language Models Get Lost in Attention
por: Xi, Gongli, et al.
Publicado: (2026) -
StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning
por: He, Xixiang, et al.
Publicado: (2026) -
Lost in OCR Translation? Vision-Based Approaches to Robust Document Retrieval
por: Most, Alexander, et al.
Publicado: (2025) -
Lost in UNet: Improving Infrared Small Target Detection by Underappreciated Local Features
por: Quan, Wuzhou, et al.
Publicado: (2024)