Jewelry Recognition via Encoder-Decoder Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Alcalde-Llergo, José M., Yeguas-Bolívar, Enrique, Zingoni, Andrea, Fuerte-Jurado, Alejandro |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Automatic Identification and Description of Jewelry Through Computer Vision and Neural Networks for Translators and Interpreters
di: Alcalde-Llergo, Jose Manuel, et al.
Pubblicazione: (2025)
di: Alcalde-Llergo, Jose Manuel, et al.
Pubblicazione: (2025)
Determining the Difficulties of Students With Dyslexia via Virtual Reality and Artificial Intelligence: An Exploratory Analysis
di: Yeguas-Bolívar, Enrique, et al.
Pubblicazione: (2024)
di: Yeguas-Bolívar, Enrique, et al.
Pubblicazione: (2024)
Fostering Inclusion: A Virtual Reality Experience to Raise Awareness of Dyslexia-Related Barriers in University Settings
di: Alcalde-Llergo, José Manuel, et al.
Pubblicazione: (2025)
di: Alcalde-Llergo, José Manuel, et al.
Pubblicazione: (2025)
A VR Serious Game to Increase Empathy towards Students with Phonological Dyslexia
di: Alcalde-Llergo, José M., et al.
Pubblicazione: (2024)
di: Alcalde-Llergo, José M., et al.
Pubblicazione: (2024)
Use of recommendation models to provide support to dyslexic students
di: Morciano, Gianluca, et al.
Pubblicazione: (2024)
di: Morciano, Gianluca, et al.
Pubblicazione: (2024)
Behavioral Engagement in VR-Based Sign Language Learning: Visual Attention as a Predictor of Performance and Temporal Dynamics
di: Traini, Davide, et al.
Pubblicazione: (2026)
di: Traini, Davide, et al.
Pubblicazione: (2026)
Leveraging Machine Learning Techniques to Investigate Media and Information Literacy Competence in Tackling Disinformation
di: Alcalde-Llergo, José Manuel, et al.
Pubblicazione: (2026)
di: Alcalde-Llergo, José Manuel, et al.
Pubblicazione: (2026)
Variational Encoder--Multi-Decoder (VE-MD) for Privacy-by-functional-design (Group) Emotion Recognition
di: Augusma, Anderson, et al.
Pubblicazione: (2026)
di: Augusma, Anderson, et al.
Pubblicazione: (2026)
Oil Spill Segmentation using Deep Encoder-Decoder models
di: Satyanarayana, Abhishek Ramanathapura, et al.
Pubblicazione: (2023)
di: Satyanarayana, Abhishek Ramanathapura, et al.
Pubblicazione: (2023)
Teacher Encoder-Student Decoder Denoising Guided Segmentation Network for Anomaly Detection
di: Song, Shixuan, et al.
Pubblicazione: (2025)
di: Song, Shixuan, et al.
Pubblicazione: (2025)
EDIT: Enhancing Vision Transformers by Mitigating Attention Sink through an Encoder-Decoder Architecture
di: Feng, Wenfeng, et al.
Pubblicazione: (2025)
di: Feng, Wenfeng, et al.
Pubblicazione: (2025)
SEDEG:Sequential Enhancement of Decoder and Encoder's Generality for Class Incremental Learning with Small Memory
di: Chen, Hongyang, et al.
Pubblicazione: (2025)
di: Chen, Hongyang, et al.
Pubblicazione: (2025)
Training program on sign language: social inclusion through Virtual Reality in ISENSE project
di: Bisio, Alessia, et al.
Pubblicazione: (2024)
di: Bisio, Alessia, et al.
Pubblicazione: (2024)
SIEDD: Shared-Implicit Encoder with Discrete Decoders
di: Rangarajan, Vikram, et al.
Pubblicazione: (2025)
di: Rangarajan, Vikram, et al.
Pubblicazione: (2025)
Real-Time Manipulation Action Recognition with a Factorized Graph Sequence Encoder
di: Erdogan, Enes, et al.
Pubblicazione: (2025)
di: Erdogan, Enes, et al.
Pubblicazione: (2025)
Analysing the Needs of Homeless People Using Feature Selection and Mining Association Rules
di: Alcalde-Llergo, José M., et al.
Pubblicazione: (2024)
di: Alcalde-Llergo, José M., et al.
Pubblicazione: (2024)
Deformable Image Registration with Multi-scale Feature Fusion from Shared Encoder, Auxiliary and Pyramid Decoders
di: Zhou, Hongchao, et al.
Pubblicazione: (2024)
di: Zhou, Hongchao, et al.
Pubblicazione: (2024)
Road Traffic Sign Recognition method using Siamese network Combining Efficient-CNN based Encoder
di: Xi, Zhenghao, et al.
Pubblicazione: (2025)
di: Xi, Zhenghao, et al.
Pubblicazione: (2025)
Bidirectional Trained Tree-Structured Decoder for Handwritten Mathematical Expression Recognition
di: Cheng, Hanbo, et al.
Pubblicazione: (2023)
di: Cheng, Hanbo, et al.
Pubblicazione: (2023)
Exploring Camera Encoder Designs for Autonomous Driving Perception
di: Lakshmanan, Barath, et al.
Pubblicazione: (2024)
di: Lakshmanan, Barath, et al.
Pubblicazione: (2024)
Design and evaluation of a serious game in virtual reality to increase empathy towards students with phonological dyslexia
di: Alcalde-Llergo, Jose Manuel, et al.
Pubblicazione: (2025)
di: Alcalde-Llergo, Jose Manuel, et al.
Pubblicazione: (2025)
Agglomerating Large Vision Encoders via Distillation for VFSS Segmentation
di: Zeng, Chengxi, et al.
Pubblicazione: (2025)
di: Zeng, Chengxi, et al.
Pubblicazione: (2025)
Automatic dental superimposition of 3D intraorals and 2D photographs for human identification
di: Villegas-Yeguas, Antonio D., et al.
Pubblicazione: (2026)
di: Villegas-Yeguas, Antonio D., et al.
Pubblicazione: (2026)
Surgical Triplet Recognition via Diffusion Model
di: Liu, Daochang, et al.
Pubblicazione: (2024)
di: Liu, Daochang, et al.
Pubblicazione: (2024)
Enhancing Polyp Segmentation via Encoder Attention and Dynamic Kernel Update
di: Chashmi, Fatemeh Salahi, et al.
Pubblicazione: (2025)
di: Chashmi, Fatemeh Salahi, et al.
Pubblicazione: (2025)
SHIELD: Suppressing Hallucinations In LVLM Encoders via Bias and Vulnerability Defense
di: Huang, Yiyang, et al.
Pubblicazione: (2025)
di: Huang, Yiyang, et al.
Pubblicazione: (2025)
Compound Expression Recognition via Multi Model Ensemble
di: Yu, Jun, et al.
Pubblicazione: (2024)
di: Yu, Jun, et al.
Pubblicazione: (2024)
Improving Image Coding for Machines through Optimizing Encoder via Auxiliary Loss
di: Iino, Kei, et al.
Pubblicazione: (2024)
di: Iino, Kei, et al.
Pubblicazione: (2024)
Fusion to Enhance: Fusion Visual Encoder to Enhance Multimodal Language Model
di: She, Yifei, et al.
Pubblicazione: (2025)
di: She, Yifei, et al.
Pubblicazione: (2025)
EVEv2: Improved Baselines for Encoder-Free Vision-Language Models
di: Diao, Haiwen, et al.
Pubblicazione: (2025)
di: Diao, Haiwen, et al.
Pubblicazione: (2025)
Investigating Redundancy in Multimodal Large Language Models with Multiple Vision Encoders
di: Wang, Yizhou, et al.
Pubblicazione: (2025)
di: Wang, Yizhou, et al.
Pubblicazione: (2025)
Compound Expression Recognition via Large Vision-Language Models
di: Yu, Jun, et al.
Pubblicazione: (2025)
di: Yu, Jun, et al.
Pubblicazione: (2025)
MagicVL-2B: Empowering Vision-Language Models on Mobile Devices with Lightweight Visual Encoders via Curriculum Learning
di: Liu, Yi, et al.
Pubblicazione: (2025)
di: Liu, Yi, et al.
Pubblicazione: (2025)
Explainable Face Recognition via Improved Localization
di: Shadman, Rashik, et al.
Pubblicazione: (2025)
di: Shadman, Rashik, et al.
Pubblicazione: (2025)
Mitigating Hallucinations in Video Large Language Models via Spatiotemporal-Semantic Contrastive Decoding
di: Gao, Yuansheng, et al.
Pubblicazione: (2026)
di: Gao, Yuansheng, et al.
Pubblicazione: (2026)
Efficient Encoder-Free Fourier-based 3D Large Multimodal Model
di: Mei, Guofeng, et al.
Pubblicazione: (2026)
di: Mei, Guofeng, et al.
Pubblicazione: (2026)
SelfSwapper: Self-Supervised Face Swapping via Shape Agnostic Masked AutoEncoder
di: Lee, Jaeseong, et al.
Pubblicazione: (2024)
di: Lee, Jaeseong, et al.
Pubblicazione: (2024)
Image Based Character Recognition, Documentation System To Decode Inscription From Temple
di: G, Velmathi, et al.
Pubblicazione: (2024)
di: G, Velmathi, et al.
Pubblicazione: (2024)
Octopus: Alleviating Hallucination via Dynamic Contrastive Decoding
di: Suo, Wei, et al.
Pubblicazione: (2025)
di: Suo, Wei, et al.
Pubblicazione: (2025)
Disentanglement and Compositionality of Letter Identity and Letter Position in Variational Auto-Encoder Vision Models
di: Bianchi, Bruno, et al.
Pubblicazione: (2024)
di: Bianchi, Bruno, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Automatic Identification and Description of Jewelry Through Computer Vision and Neural Networks for Translators and Interpreters
di: Alcalde-Llergo, Jose Manuel, et al.
Pubblicazione: (2025) -
Determining the Difficulties of Students With Dyslexia via Virtual Reality and Artificial Intelligence: An Exploratory Analysis
di: Yeguas-Bolívar, Enrique, et al.
Pubblicazione: (2024) -
Fostering Inclusion: A Virtual Reality Experience to Raise Awareness of Dyslexia-Related Barriers in University Settings
di: Alcalde-Llergo, José Manuel, et al.
Pubblicazione: (2025) -
A VR Serious Game to Increase Empathy towards Students with Phonological Dyslexia
di: Alcalde-Llergo, José M., et al.
Pubblicazione: (2024) -
Use of recommendation models to provide support to dyslexic students
di: Morciano, Gianluca, et al.
Pubblicazione: (2024)