Col-OLHTR: A Novel Framework for Multimodal Online Handwritten Text Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Chenyu, Hu, Jinshui, Yin, Baocai, Pan, Jia, Yin, Bing, Du, Jun, Liu, Qingfeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NAMER: Non-Autoregressive Modeling for Handwritten Mathematical Expression Recognition
by: Liu, Chenyu, et al.
Published: (2024)
by: Liu, Chenyu, et al.
Published: (2024)
Binary-Gaussian: Compact and Progressive Representation for 3D Gaussian Segmentation
by: Yang, An, et al.
Published: (2025)
by: Yang, An, et al.
Published: (2025)
Introducing Multimodal Paradigm for Learning Sleep Staging PSG via General-Purpose Model
by: Zhou, Jianheng, et al.
Published: (2025)
by: Zhou, Jianheng, et al.
Published: (2025)
TDATR: Improving End-to-End Table Recognition via Table Detail-Aware Learning and Cell-Level Visual Alignment
by: Qin, Chunxia, et al.
Published: (2026)
by: Qin, Chunxia, et al.
Published: (2026)
OpenMarcie: Dataset for Multimodal Action Recognition in Industrial Environments
by: Bello, Hymalai, et al.
Published: (2026)
by: Bello, Hymalai, et al.
Published: (2026)
WiFi-based Cross-Domain Gesture Recognition Using Attention Mechanism
by: Liu, Ruijing, et al.
Published: (2025)
by: Liu, Ruijing, et al.
Published: (2025)
Multipath Interference Suppression in Indirect Time-of-Flight Imaging via a Novel Compressed Sensing Framework
by: Du, Yansong, et al.
Published: (2025)
by: Du, Yansong, et al.
Published: (2025)
Exploring Part-Informed Visual-Language Learning for Person Re-Identification
by: Lin, Yin, et al.
Published: (2023)
by: Lin, Yin, et al.
Published: (2023)
Cross-Domain Multi-Person Human Activity Recognition via Near-Field Wi-Fi Sensing
by: Li, Xin, et al.
Published: (2025)
by: Li, Xin, et al.
Published: (2025)
WDMIR: Wavelet-Driven Multimodal Intent Recognition
by: Gong, Weiyin, et al.
Published: (2025)
by: Gong, Weiyin, et al.
Published: (2025)
Task-Oriented Feature Compression for Multimodal Understanding via Device-Edge Co-Inference
by: Yuan, Cheng, et al.
Published: (2025)
by: Yuan, Cheng, et al.
Published: (2025)
OG-PCL: Efficient Sparse Point Cloud Processing for Human Activity Recognition
by: Yan, Jiuqi, et al.
Published: (2025)
by: Yan, Jiuqi, et al.
Published: (2025)
Towards Open-Set Myoelectric Gesture Recognition via Dual-Perspective Inconsistency Learning
by: Liu, Chen, et al.
Published: (2024)
by: Liu, Chen, et al.
Published: (2024)
Body-Area Capacitive or Electric Field Sensing for Human Activity Recognition and Human-Computer Interaction: A Comprehensive Survey
by: Bian, Sizhen, et al.
Published: (2024)
by: Bian, Sizhen, et al.
Published: (2024)
1DFormer: a Transformer Architecture Learning 1D Landmark Representations for Facial Landmark Tracking
by: Yin, Shi, et al.
Published: (2023)
by: Yin, Shi, et al.
Published: (2023)
Two-Stage Hierarchical and Explainable Feature Selection Framework for Dimensionality Reduction in Sleep Staging
by: Deng, Yangfan, et al.
Published: (2024)
by: Deng, Yangfan, et al.
Published: (2024)
Generative Video Semantic Communication via Multimodal Semantic Fusion with Large Model
by: Yin, Hang, et al.
Published: (2025)
by: Yin, Hang, et al.
Published: (2025)
RFL: Simplifying Chemical Structure Recognition with Ring-Free Language
by: Chang, Qikai, et al.
Published: (2024)
by: Chang, Qikai, et al.
Published: (2024)
A PolSAR Scattering Power Factorization Framework and Novel Roll-Invariant Parameters Based Unsupervised Classification Scheme Using a Geodesic Distance
by: Ratha, Debanshu, et al.
Published: (2019)
by: Ratha, Debanshu, et al.
Published: (2019)
Task-Oriented Communication for Human Action Understanding via Edge-Cloud Co-Inference
by: Liu, Jingyi, et al.
Published: (2026)
by: Liu, Jingyi, et al.
Published: (2026)
Surface Recognition for e-Scooter Using Smartphone IMU Sensor
by: Eweida, Areej, et al.
Published: (2023)
by: Eweida, Areej, et al.
Published: (2023)
Integrated Image Reconstruction and Target Recognition based on Deep Learning Technique
by: Zhang, Cien, et al.
Published: (2025)
by: Zhang, Cien, et al.
Published: (2025)
Radar-Based Recognition of Static Hand Gestures in American Sign Language
by: Schuessler, Christian, et al.
Published: (2024)
by: Schuessler, Christian, et al.
Published: (2024)
Multimodal Indoor Localization Using Crowdsourced Radio Maps
by: Yi, Zhaoguang, et al.
Published: (2023)
by: Yi, Zhaoguang, et al.
Published: (2023)
Radar-Camera Fused Multi-Object Tracking: Online Calibration and Common Feature
by: Cheng, Lei, et al.
Published: (2025)
by: Cheng, Lei, et al.
Published: (2025)
BenchHAR: Benchmarking Self-Supervised Learning for Generalizable Sensor-based Activity Recognition
by: Cai, Yize, et al.
Published: (2026)
by: Cai, Yize, et al.
Published: (2026)
Open-Set Gait Recognition from Sparse mmWave Radar Point Clouds
by: Mazzieri, Riccardo, et al.
Published: (2025)
by: Mazzieri, Riccardo, et al.
Published: (2025)
Enabling Visual Recognition at Radio Frequency
by: Lai, Haowen, et al.
Published: (2024)
by: Lai, Haowen, et al.
Published: (2024)
Wi-CBR: Salient-aware Adaptive WiFi Sensing for Cross-domain Behavior Recognition
by: Zhang, Ruobei, et al.
Published: (2025)
by: Zhang, Ruobei, et al.
Published: (2025)
Enhancing Automatic Modulation Recognition With a Reconstruction-Driven Vision Transformer Under Limited Labels
by: Ahmadi, Hossein, et al.
Published: (2025)
by: Ahmadi, Hossein, et al.
Published: (2025)
DoRF: Doppler Radiance Fields for Robust Human Activity Recognition Using Wi-Fi
by: Hasanzadeh, Navid, et al.
Published: (2025)
by: Hasanzadeh, Navid, et al.
Published: (2025)
mmID: High-Resolution mmWave Imaging for Human Identification
by: Jayaweera, Sakila S., et al.
Published: (2024)
by: Jayaweera, Sakila S., et al.
Published: (2024)
Multimodal Power Outage Prediction for Rapid Disaster Response and Resource Allocation
by: Aparcedo, Alejandro, et al.
Published: (2024)
by: Aparcedo, Alejandro, et al.
Published: (2024)
Neural-HAR: A Dimension-Gated CNN Accelerator for Real-Time Radar Human Activity Recognition
by: Wu, Yizhuo, et al.
Published: (2025)
by: Wu, Yizhuo, et al.
Published: (2025)
Generative AI Empowered LiDAR Point Cloud Generation with Multimodal Transformer
by: Farzanullah, Mohammad, et al.
Published: (2024)
by: Farzanullah, Mohammad, et al.
Published: (2024)
X-Fi: A Modality-Invariant Foundation Model for Multimodal Human Sensing
by: Chen, Xinyan, et al.
Published: (2024)
by: Chen, Xinyan, et al.
Published: (2024)
A Foundation Model for DAS Signal Recognition and Visual Prompt Tuning of the Pre-trained Model for Downstream Tasks
by: Gui, Kun, et al.
Published: (2025)
by: Gui, Kun, et al.
Published: (2025)
Large Model for Small Data: Foundation Model for Cross-Modal RF Human Activity Recognition
by: Weng, Yuxuan, et al.
Published: (2024)
by: Weng, Yuxuan, et al.
Published: (2024)
Bidirectional Trained Tree-Structured Decoder for Handwritten Mathematical Expression Recognition
by: Cheng, Hanbo, et al.
Published: (2023)
by: Cheng, Hanbo, et al.
Published: (2023)
WiFi-TCN: Temporal Convolution for Human Interaction Recognition based on WiFi signal
by: Lin, Chih-Yang, et al.
Published: (2023)
by: Lin, Chih-Yang, et al.
Published: (2023)
Similar Items
-
NAMER: Non-Autoregressive Modeling for Handwritten Mathematical Expression Recognition
by: Liu, Chenyu, et al.
Published: (2024) -
Binary-Gaussian: Compact and Progressive Representation for 3D Gaussian Segmentation
by: Yang, An, et al.
Published: (2025) -
Introducing Multimodal Paradigm for Learning Sleep Staging PSG via General-Purpose Model
by: Zhou, Jianheng, et al.
Published: (2025) -
TDATR: Improving End-to-End Table Recognition via Table Detail-Aware Learning and Cell-Level Visual Alignment
by: Qin, Chunxia, et al.
Published: (2026) -
OpenMarcie: Dataset for Multimodal Action Recognition in Industrial Environments
by: Bello, Hymalai, et al.
Published: (2026)