Zero-Shot Chinese Character Recognition with Hierarchical Multi-Granularity Image-Text Aligning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Yinglian, Yu, Haiyang, Wang, Qizao, Lu, Wei, Xue, Xiangyang, Li, Bin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AMR-CCR: Anchored Modular Retrieval for Continual Chinese Character Recognition
by: Wu, Yuchuan, et al.
Published: (2026)
by: Wu, Yuchuan, et al.
Published: (2026)
Distribution Aligned Semantics Adaption for Lifelong Person Re-Identification
by: Wang, Qizao, et al.
Published: (2024)
by: Wang, Qizao, et al.
Published: (2024)
CoLa: Chinese Character Decomposition with Compositional Latent Components
by: Shi, Fan, et al.
Published: (2025)
by: Shi, Fan, et al.
Published: (2025)
Image-Text-Image Knowledge Transfer for Lifelong Person Re-Identification with Hybrid Clothing States
by: Wang, Qizao, et al.
Published: (2024)
by: Wang, Qizao, et al.
Published: (2024)
When Large Vision-Language Models Meet Person Re-Identification
by: Wang, Qizao, et al.
Published: (2024)
by: Wang, Qizao, et al.
Published: (2024)
Unleashing the Potential of Tracklets for Unsupervised Video Person Re-Identification
by: Meng, Nanxing, et al.
Published: (2024)
by: Meng, Nanxing, et al.
Published: (2024)
Zero-Shot Chinese Character Recognition via Global-Local Dual-Branch Alignment and Hierarchical Inference
by: Cao, Wei, et al.
Published: (2026)
by: Cao, Wei, et al.
Published: (2026)
Exploring Fine-Grained Representation and Recomposition for Cloth-Changing Person Re-Identification
by: Wang, Qizao, et al.
Published: (2023)
by: Wang, Qizao, et al.
Published: (2023)
EAFormer: Scene Text Segmentation with Edge-Aware Transformers
by: Yu, Haiyang, et al.
Published: (2024)
by: Yu, Haiyang, et al.
Published: (2024)
Complementary Text-Guided Attention for Zero-Shot Adversarial Robustness
by: Yu, Lu, et al.
Published: (2026)
by: Yu, Lu, et al.
Published: (2026)
Content and Salient Semantics Collaboration for Cloth-Changing Person Re-Identification
by: Wang, Qizao, et al.
Published: (2024)
by: Wang, Qizao, et al.
Published: (2024)
Multi-Modal Character Localization and Extraction for Chinese Text Recognition
by: Li, Qilong, et al.
Published: (2026)
by: Li, Qilong, et al.
Published: (2026)
Entropy-Aware Structural Alignment for Zero-Shot Handwritten Chinese Character Recognition
by: Luo, Qiuming, et al.
Published: (2026)
by: Luo, Qiuming, et al.
Published: (2026)
Multi-Granularity Mutual Refinement Network for Zero-Shot Learning
by: Wang, Ning, et al.
Published: (2025)
by: Wang, Ning, et al.
Published: (2025)
HierCode: A Lightweight Hierarchical Codebook for Zero-shot Chinese Text Recognition
by: Zhang, Yuyi, et al.
Published: (2024)
by: Zhang, Yuyi, et al.
Published: (2024)
Text-Guided Attention is All You Need for Zero-Shot Robustness in Vision-Language Models
by: Yu, Lu, et al.
Published: (2024)
by: Yu, Lu, et al.
Published: (2024)
Subject-Aware Multi-Granularity Alignment for Zero-Shot EEG-to-Image Retrieval
by: Jiang, Lin, et al.
Published: (2026)
by: Jiang, Lin, et al.
Published: (2026)
PHT-CAD: Efficient CAD Parametric Primitive Analysis with Progressive Hierarchical Tuning
by: Niu, Ke, et al.
Published: (2025)
by: Niu, Ke, et al.
Published: (2025)
GranAlign: Granularity-Aware Alignment Framework for Zero-Shot Video Moment Retrieval
by: Jeon, Mingyu, et al.
Published: (2026)
by: Jeon, Mingyu, et al.
Published: (2026)
ChatReID: Open-ended Interactive Person Retrieval via Hierarchical Progressive Tuning for Vision Language Models
by: Niu, Ke, et al.
Published: (2025)
by: Niu, Ke, et al.
Published: (2025)
RaCMC: Residual-Aware Compensation Network with Multi-Granularity Constraints for Fake News Detection
by: Yu, Xinquan, et al.
Published: (2024)
by: Yu, Xinquan, et al.
Published: (2024)
Benchmarking Zero-Shot Recognition with Vision-Language Models: Challenges on Granularity and Specificity
by: Xu, Zhenlin, et al.
Published: (2023)
by: Xu, Zhenlin, et al.
Published: (2023)
Skeleton and Font Generation Network for Zero-shot Chinese Character Generation
by: Xue, Mobai, et al.
Published: (2025)
by: Xue, Mobai, et al.
Published: (2025)
MGHFT: Multi-Granularity Hierarchical Fusion Transformer for Cross-Modal Sticker Emotion Recognition
by: Chen, Jian, et al.
Published: (2025)
by: Chen, Jian, et al.
Published: (2025)
FedDEO: Description-Enhanced One-Shot Federated Learning with Diffusion Models
by: Yang, Mingzhao, et al.
Published: (2024)
by: Yang, Mingzhao, et al.
Published: (2024)
Focus on the Whole Character: Discriminative Character Modeling for Scene Text Recognition
by: Zhou, Bangbang, et al.
Published: (2024)
by: Zhou, Bangbang, et al.
Published: (2024)
Learning to Align: Addressing Character Frequency Distribution Shifts in Handwritten Text Recognition
by: Kaliosis, Panagiotis, et al.
Published: (2025)
by: Kaliosis, Panagiotis, et al.
Published: (2025)
Linguistics-aware Masked Image Modeling for Self-supervised Scene Text Recognition
by: Zhang, Yifei, et al.
Published: (2025)
by: Zhang, Yifei, et al.
Published: (2025)
Distributed Zero-Shot Learning for Visual Recognition
by: Chen, Zhi, et al.
Published: (2025)
by: Chen, Zhi, et al.
Published: (2025)
Image Over Text: Transforming Formula Recognition Evaluation with Character Detection Matching
by: Wang, Bin, et al.
Published: (2024)
by: Wang, Bin, et al.
Published: (2024)
MTGA: Multi-View Temporal Granularity Aligned Aggregation for Event-Based Lip-Reading
by: Zhang, Wenhao, et al.
Published: (2024)
by: Zhang, Wenhao, et al.
Published: (2024)
VFM-Loc: Zero-Shot Cross-View Geo-Localization via Aligning Discriminative Visual Hierarchies
by: Lu, Jun, et al.
Published: (2026)
by: Lu, Jun, et al.
Published: (2026)
Performance Analysis of Few-Shot Learning Approaches for Bangla Handwritten Character and Digit Recognition
by: Ahamed, Mehedi, et al.
Published: (2025)
by: Ahamed, Mehedi, et al.
Published: (2025)
Text as Any-Modality for Zero-Shot Classification by Consistent Prompt Tuning
by: Wu, Xiangyu, et al.
Published: (2025)
by: Wu, Xiangyu, et al.
Published: (2025)
One-Shot Heterogeneous Federated Learning with Local Model-Guided Diffusion Models
by: Yang, Mingzhao, et al.
Published: (2023)
by: Yang, Mingzhao, et al.
Published: (2023)
Zero-Shot Skeleton-based Action Recognition with Dual Visual-Text Alignment
by: Kuang, Jidong, et al.
Published: (2024)
by: Kuang, Jidong, et al.
Published: (2024)
Text-Enhanced Zero-Shot Action Recognition: A training-free approach
by: Bosetti, Massimo, et al.
Published: (2024)
by: Bosetti, Massimo, et al.
Published: (2024)
Zero-Shot Synthetic-to-Real Handwritten Text Recognition via Task Analogies
by: Garrido-Munoz, Carlos, et al.
Published: (2026)
by: Garrido-Munoz, Carlos, et al.
Published: (2026)
MuSc: Zero-Shot Industrial Anomaly Classification and Segmentation with Mutual Scoring of the Unlabeled Images
by: Li, Xurui, et al.
Published: (2024)
by: Li, Xurui, et al.
Published: (2024)
DocCogito: Aligning Layout Cognition and Step-Level Grounded Reasoning for Document Understanding
by: Wu, Yuchuan, et al.
Published: (2026)
by: Wu, Yuchuan, et al.
Published: (2026)
Similar Items
-
AMR-CCR: Anchored Modular Retrieval for Continual Chinese Character Recognition
by: Wu, Yuchuan, et al.
Published: (2026) -
Distribution Aligned Semantics Adaption for Lifelong Person Re-Identification
by: Wang, Qizao, et al.
Published: (2024) -
CoLa: Chinese Character Decomposition with Compositional Latent Components
by: Shi, Fan, et al.
Published: (2025) -
Image-Text-Image Knowledge Transfer for Lifelong Person Re-Identification with Hybrid Clothing States
by: Wang, Qizao, et al.
Published: (2024) -
When Large Vision-Language Models Meet Person Re-Identification
by: Wang, Qizao, et al.
Published: (2024)