A comprehensive survey of oracle character recognition: challenges, benchmarks, and beyond
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Jing, Chi, Xueke, Wang, Qiufeng, Wang, Dahan, Huang, Kaizhu, Liu, Yongge, Liu, Cheng-lin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
IDEA: Image Description Enhanced CLIP-Adapter
by: Ye, Zhipeng, et al.
Published: (2025)
by: Ye, Zhipeng, et al.
Published: (2025)
The Demon is in Ambiguity: Revisiting Situation Recognition with Single Positive Multi-Label Learning
by: Lin, Yiming, et al.
Published: (2025)
by: Lin, Yiming, et al.
Published: (2025)
Interpret Your Decision: Logical Reasoning Regularization for Generalization in Visual Classification
by: Tan, Zhaorui, et al.
Published: (2024)
by: Tan, Zhaorui, et al.
Published: (2024)
An open dataset for oracle bone script recognition and decipherment
by: Wang, Pengjie, et al.
Published: (2024)
by: Wang, Pengjie, et al.
Published: (2024)
Towards a Universal 3D Medical Multi-modality Generalization via Learning Personalized Invariant Representation
by: Tan, Zhaorui, et al.
Published: (2024)
by: Tan, Zhaorui, et al.
Published: (2024)
Rethinking Information Loss in Medical Image Segmentation with Various-sized Targets
by: Liu, Tianyi, et al.
Published: (2024)
by: Liu, Tianyi, et al.
Published: (2024)
Towards Faithful Reasoning in Comics for Small MLLMs
by: Feng, Chengcheng, et al.
Published: (2026)
by: Feng, Chengcheng, et al.
Published: (2026)
PixelArena: A benchmark for Pixel-Precision Visual Intelligence
by: Liang, Feng, et al.
Published: (2025)
by: Liang, Feng, et al.
Published: (2025)
MedMAP: Promoting Incomplete Multi-modal Brain Tumor Segmentation with Alignment
by: Liu, Tianyi, et al.
Published: (2024)
by: Liu, Tianyi, et al.
Published: (2024)
How well are open sourced AI-generated image detection models out-of-the-box: A comprehensive benchmark study
by: Ren, Simiao, et al.
Published: (2026)
by: Ren, Simiao, et al.
Published: (2026)
Few Shot Semantic Segmentation: a review of methodologies, benchmarks, and open challenges
by: Catalano, Nico, et al.
Published: (2023)
by: Catalano, Nico, et al.
Published: (2023)
Covariance-based Space Regularization for Few-shot Class Incremental Learning
by: Hu, Yijie, et al.
Published: (2024)
by: Hu, Yijie, et al.
Published: (2024)
BFANet: Revisiting 3D Semantic Segmentation with Boundary Feature Analysis
by: Zhao, Weiguang, et al.
Published: (2025)
by: Zhao, Weiguang, et al.
Published: (2025)
Rethinking Multi-domain Generalization with A General Learning Objective
by: Tan, Zhaorui, et al.
Published: (2024)
by: Tan, Zhaorui, et al.
Published: (2024)
A comprehensive multimodal dataset and benchmark for ulcerative colitis scoring in endoscopy
by: Ghatwary, Noha, et al.
Published: (2026)
by: Ghatwary, Noha, et al.
Published: (2026)
A generalizable framework for low-rank tensor completion with numerical priors
by: Yuan, Shiran, et al.
Published: (2023)
by: Yuan, Shiran, et al.
Published: (2023)
Generalizing vision-language models to novel domains: A comprehensive survey
by: Li, Xinyao, et al.
Published: (2025)
by: Li, Xinyao, et al.
Published: (2025)
HCR-Net: A deep learning based script independent handwritten character recognition network
by: Chauhan, Vinod Kumar, et al.
Published: (2021)
by: Chauhan, Vinod Kumar, et al.
Published: (2021)
Divide-and-Conquer Inference for Large-Scale Visual Recognition with Multimodal Large Language Models
by: Ye, Zhipeng, et al.
Published: (2026)
by: Ye, Zhipeng, et al.
Published: (2026)
Face recognition on point cloud with cgan-top for denoising
by: Liu, Junyu, et al.
Published: (2025)
by: Liu, Junyu, et al.
Published: (2025)
Mind the Gap: Promoting Missing Modality Brain Tumor Segmentation with Alignment
by: Liu, Tianyi, et al.
Published: (2024)
by: Liu, Tianyi, et al.
Published: (2024)
Less yet robust: crucial region selection for scene recognition
by: Zhang, Jianqi, et al.
Published: (2024)
by: Zhang, Jianqi, et al.
Published: (2024)
LiteAttention: A Temporal Sparse Attention for Diffusion Transformers
by: Shmilovich, Dor, et al.
Published: (2025)
by: Shmilovich, Dor, et al.
Published: (2025)
Hybrid guided variational autoencoder for visual place recognition
by: Wang, Ni, et al.
Published: (2026)
by: Wang, Ni, et al.
Published: (2026)
Intelligent recognition of GPR road hidden defect images based on feature fusion and attention mechanism
by: Lv, Haotian, et al.
Published: (2025)
by: Lv, Haotian, et al.
Published: (2025)
DiverseGRPO: Mitigating Mode Collapse in Image Generation via Diversity-Aware GRPO
by: Liu, Henglin, et al.
Published: (2025)
by: Liu, Henglin, et al.
Published: (2025)
BCFPL: Binary classification ConvNet based Fast Parking space recognition with Low resolution image
by: Zhang, Shuo, et al.
Published: (2024)
by: Zhang, Shuo, et al.
Published: (2024)
NTIRE 2025 challenge on Text to Image Generation Model Quality Assessment
by: Han, Shuhao, et al.
Published: (2025)
by: Han, Shuhao, et al.
Published: (2025)
GeoSDF: Plane Geometry Diagram Synthesis via Signed Distance Field
by: Zhang, Chengrui, et al.
Published: (2025)
by: Zhang, Chengrui, et al.
Published: (2025)
CAMEL2: Enhancing weakly supervised learning for histopathology images by incorporating the significance ratio
by: Xu, Gang, et al.
Published: (2023)
by: Xu, Gang, et al.
Published: (2023)
BENet: A Cross-domain Robust Network for Detecting Face Forgeries via Bias Expansion and Latent-space Attention
by: Liu, Weihua, et al.
Published: (2024)
by: Liu, Weihua, et al.
Published: (2024)
BenchReAD: A systematic benchmark for retinal anomaly detection
by: Lian, Chenyu, et al.
Published: (2025)
by: Lian, Chenyu, et al.
Published: (2025)
DvD: Unleashing a Generative Paradigm for Document Dewarping via Coordinates-based Diffusion Model
by: Zhang, Weiguang, et al.
Published: (2025)
by: Zhang, Weiguang, et al.
Published: (2025)
SImpHAR: Advancing impedance-based human activity recognition using 3D simulation and text-to-motion models
by: Ray, Lala Shakti Swarup, et al.
Published: (2025)
by: Ray, Lala Shakti Swarup, et al.
Published: (2025)
Continual-learning-based framework for structural damage recognition
by: Shu, Jiangpeng, et al.
Published: (2024)
by: Shu, Jiangpeng, et al.
Published: (2024)
HetScene: Heterogeneity-Aware Diffusion for Dense Indoor Scene Generation
by: Chen, Zini, et al.
Published: (2026)
by: Chen, Zini, et al.
Published: (2026)
A benchmark dataset for deep learning-based airplane detection: HRPlanes
by: Bakirman, Tolga, et al.
Published: (2022)
by: Bakirman, Tolga, et al.
Published: (2022)
APTOS-2024 challenge report: Generation of synthetic 3D OCT images from fundus photographs
by: Liu, Bowen, et al.
Published: (2025)
by: Liu, Bowen, et al.
Published: (2025)
Seeing Symbols, Missing Cultures: Probing Vision-Language Models' Reasoning on Fire Imagery and Cultural Meaning
by: Yu, Haorui, et al.
Published: (2025)
by: Yu, Haorui, et al.
Published: (2025)
Lightweight framework for underground pipeline recognition and spatial localization based on multi-view 2D GPR images
by: Lv, Haotian, et al.
Published: (2025)
by: Lv, Haotian, et al.
Published: (2025)
Similar Items
-
IDEA: Image Description Enhanced CLIP-Adapter
by: Ye, Zhipeng, et al.
Published: (2025) -
The Demon is in Ambiguity: Revisiting Situation Recognition with Single Positive Multi-Label Learning
by: Lin, Yiming, et al.
Published: (2025) -
Interpret Your Decision: Logical Reasoning Regularization for Generalization in Visual Classification
by: Tan, Zhaorui, et al.
Published: (2024) -
An open dataset for oracle bone script recognition and decipherment
by: Wang, Pengjie, et al.
Published: (2024) -
Towards a Universal 3D Medical Multi-modality Generalization via Learning Personalized Invariant Representation
by: Tan, Zhaorui, et al.
Published: (2024)