Studying Illustrations in Manuscripts: An Efficient Deep-Learning Approach
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Evron, Yoav, Siegal, Michal Bar-Asher, Fire, Michael |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Resource-Efficient Streaming of Large-Scale Medical Image Datasets for Deep Learning
von: Kulkarni, Pranav, et al.
Veröffentlicht: (2023)
von: Kulkarni, Pranav, et al.
Veröffentlicht: (2023)
Deep Learning for Technical Document Classification
von: Jiang, Shuo, et al.
Veröffentlicht: (2021)
von: Jiang, Shuo, et al.
Veröffentlicht: (2021)
AdaTask: A Task-aware Adaptive Learning Rate Approach to Multi-task Learning
von: Yang, Enneng, et al.
Veröffentlicht: (2022)
von: Yang, Enneng, et al.
Veröffentlicht: (2022)
Hierarchy-of-Visual-Words: a Learning-based Approach for Trademark Image Retrieval
von: Lourenço, Vítor N., et al.
Veröffentlicht: (2019)
von: Lourenço, Vítor N., et al.
Veröffentlicht: (2019)
A Guide to Similarity Measures
von: Levy, Avivit, et al.
Veröffentlicht: (2024)
von: Levy, Avivit, et al.
Veröffentlicht: (2024)
Semantic-Cohesive Knowledge Distillation for Deep Cross-modal Hashing
von: Sun, Changchang, et al.
Veröffentlicht: (2025)
von: Sun, Changchang, et al.
Veröffentlicht: (2025)
iRAG: Advancing RAG for Videos with an Incremental Approach
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2024)
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2024)
Data-Centric Approach to Constrained Machine Learning: A Case Study on Conway's Game of Life
von: Bibin, Anton, et al.
Veröffentlicht: (2024)
von: Bibin, Anton, et al.
Veröffentlicht: (2024)
Self-Supervised Learning as Discrete Communication
von: Zaher, Kawtar, et al.
Veröffentlicht: (2026)
von: Zaher, Kawtar, et al.
Veröffentlicht: (2026)
Generalized Contrastive Learning for Multi-Modal Retrieval and Ranking
von: Zhu, Tianyu, et al.
Veröffentlicht: (2024)
von: Zhu, Tianyu, et al.
Veröffentlicht: (2024)
Region-Point Joint Representation for Effective Trajectory Similarity Learning
von: Long, Hao, et al.
Veröffentlicht: (2025)
von: Long, Hao, et al.
Veröffentlicht: (2025)
Decoupled Training: Return of Frustratingly Easy Multi-Domain Learning
von: Wang, Ximei, et al.
Veröffentlicht: (2023)
von: Wang, Ximei, et al.
Veröffentlicht: (2023)
A Dataset and Framework for Learning State-invariant Object Representations
von: Sarkar, Rohan, et al.
Veröffentlicht: (2024)
von: Sarkar, Rohan, et al.
Veröffentlicht: (2024)
Online Learning via Memory: Retrieval-Augmented Detector Adaptation
von: Jian, Yanan, et al.
Veröffentlicht: (2024)
von: Jian, Yanan, et al.
Veröffentlicht: (2024)
Transfer Learning with Self-Supervised Vision Transformers for Snake Identification
von: Miyaguchi, Anthony, et al.
Veröffentlicht: (2024)
von: Miyaguchi, Anthony, et al.
Veröffentlicht: (2024)
Transfer Learning and Mixup for Fine-Grained Few-Shot Fungi Classification
von: Tam, Jason Kahei, et al.
Veröffentlicht: (2025)
von: Tam, Jason Kahei, et al.
Veröffentlicht: (2025)
LotusFilter: Fast Diverse Nearest Neighbor Search via a Learned Cutoff Table
von: Matsui, Yusuke
Veröffentlicht: (2025)
von: Matsui, Yusuke
Veröffentlicht: (2025)
Dual Pose-invariant Embeddings: Learning Category and Object-specific Discriminative Representations for Recognition and Retrieval
von: Sarkar, Rohan, et al.
Veröffentlicht: (2024)
von: Sarkar, Rohan, et al.
Veröffentlicht: (2024)
CoopHash: Cooperative Learning of Multipurpose Descriptor and Contrastive Pair Generator via Variational MCMC Teaching for Supervised Image Hashing
von: Doan, Khoa D., et al.
Veröffentlicht: (2022)
von: Doan, Khoa D., et al.
Veröffentlicht: (2022)
Low-Data Classification of Historical Music Manuscripts: A Few-Shot Learning Approach
von: Shatri, Elona, et al.
Veröffentlicht: (2024)
von: Shatri, Elona, et al.
Veröffentlicht: (2024)
Revisiting Uncertainty: On Evidential Learning for Partially Relevant Video Retrieval
von: Li, Jun, et al.
Veröffentlicht: (2026)
von: Li, Jun, et al.
Veröffentlicht: (2026)
AutoPP: Towards Automated Product Poster Generation and Optimization
von: Fan, Jiahao, et al.
Veröffentlicht: (2025)
von: Fan, Jiahao, et al.
Veröffentlicht: (2025)
Tile-Based ViT Inference with Visual-Cluster Priors for Zero-Shot Multi-Species Plant Identification
von: Gustineli, Murilo, et al.
Veröffentlicht: (2025)
von: Gustineli, Murilo, et al.
Veröffentlicht: (2025)
Improving Visual Recommendation on E-commerce Platforms Using Vision-Language Models
von: Yada, Yuki, et al.
Veröffentlicht: (2025)
von: Yada, Yuki, et al.
Veröffentlicht: (2025)
Uni-Layout: Integrating Human Feedback in Unified Layout Generation and Evaluation
von: Lu, Shuo, et al.
Veröffentlicht: (2025)
von: Lu, Shuo, et al.
Veröffentlicht: (2025)
Patent Figure Classification using Large Vision-language Models
von: Awale, Sushil, et al.
Veröffentlicht: (2025)
von: Awale, Sushil, et al.
Veröffentlicht: (2025)
Image Hashing via Cross-View Code Alignment in the Age of Foundation Models
von: Moummad, Ilyass, et al.
Veröffentlicht: (2025)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2025)
$\texttt{InfoHier}$: Hierarchical Information Extraction via Encoding and Embedding
von: Zhang, Tianru, et al.
Veröffentlicht: (2025)
von: Zhang, Tianru, et al.
Veröffentlicht: (2025)
LD-DETR: Loop Decoder DEtection TRansformer for Video Moment Retrieval and Highlight Detection
von: Zhao, Pengcheng, et al.
Veröffentlicht: (2025)
von: Zhao, Pengcheng, et al.
Veröffentlicht: (2025)
Maximal Matching Matters: Preventing Representation Collapse for Robust Cross-Modal Retrieval
von: Alomari, Hani, et al.
Veröffentlicht: (2025)
von: Alomari, Hani, et al.
Veröffentlicht: (2025)
Multi-Label Plant Species Prediction with Metadata-Enhanced Multi-Head Vision Transformers
von: Herasimchyk, Hanna, et al.
Veröffentlicht: (2025)
von: Herasimchyk, Hanna, et al.
Veröffentlicht: (2025)
GBSK: Skeleton Clustering via Granular-ball Computing and Multi-Sampling for Large-Scale Data
von: Chen, Yewang, et al.
Veröffentlicht: (2025)
von: Chen, Yewang, et al.
Veröffentlicht: (2025)
Embedding-based Retrieval in Multimodal Content Moderation
von: Liang, Hanzhong, et al.
Veröffentlicht: (2025)
von: Liang, Hanzhong, et al.
Veröffentlicht: (2025)
Domain-invariant feature learning in brain MR imaging for content-based image retrieval
von: Tobari, Shuya, et al.
Veröffentlicht: (2025)
von: Tobari, Shuya, et al.
Veröffentlicht: (2025)
When & How to Write for Personalized Demand-aware Query Rewriting in Video Search
von: cheng, Cheng, et al.
Veröffentlicht: (2025)
von: cheng, Cheng, et al.
Veröffentlicht: (2025)
VQPP: Video Query Performance Prediction Benchmark
von: Lutu, Adrian Catalin, et al.
Veröffentlicht: (2026)
von: Lutu, Adrian Catalin, et al.
Veröffentlicht: (2026)
SOLAR: SVD-Optimized Lifelong Attention for Recommendation
von: Zhang, Chenghao, et al.
Veröffentlicht: (2026)
von: Zhang, Chenghao, et al.
Veröffentlicht: (2026)
PixRec: Leveraging Visual Context for Next-Item Prediction in Sequential Recommendation
von: Chakrabarty, Sayak, et al.
Veröffentlicht: (2026)
von: Chakrabarty, Sayak, et al.
Veröffentlicht: (2026)
FOR: Finetuning for Object Level Open Vocabulary Image Retrieval
von: Levi, Hila, et al.
Veröffentlicht: (2024)
von: Levi, Hila, et al.
Veröffentlicht: (2024)
CaReBench: A Fine-Grained Benchmark for Video Captioning and Retrieval
von: Xu, Yifan, et al.
Veröffentlicht: (2024)
von: Xu, Yifan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards Resource-Efficient Streaming of Large-Scale Medical Image Datasets for Deep Learning
von: Kulkarni, Pranav, et al.
Veröffentlicht: (2023) -
Deep Learning for Technical Document Classification
von: Jiang, Shuo, et al.
Veröffentlicht: (2021) -
AdaTask: A Task-aware Adaptive Learning Rate Approach to Multi-task Learning
von: Yang, Enneng, et al.
Veröffentlicht: (2022) -
Hierarchy-of-Visual-Words: a Learning-based Approach for Trademark Image Retrieval
von: Lourenço, Vítor N., et al.
Veröffentlicht: (2019) -
A Guide to Similarity Measures
von: Levy, Avivit, et al.
Veröffentlicht: (2024)