InkFM: A Foundational Model for Full-Page Online Handwritten Note Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Fadeeva, Anastasiia, Coriou, Vincent, Antognini, Diego, Musat, Claudiu, Maksai, Andrii |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Representing Online Handwriting for Recognition in Large Vision-Language Models
by: Fadeeva, Anastasiia, et al.
Published: (2024)
by: Fadeeva, Anastasiia, et al.
Published: (2024)
MathWriting: A Dataset For Handwritten Mathematical Expression Recognition
by: Gervais, Philippe, et al.
Published: (2024)
by: Gervais, Philippe, et al.
Published: (2024)
InkSight: Offline-to-Online Handwriting Conversion by Teaching Vision-Language Models to Read and Write
by: Mitrevski, Blagoj, et al.
Published: (2024)
by: Mitrevski, Blagoj, et al.
Published: (2024)
Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation
by: Brioschi, Riccardo, et al.
Published: (2025)
by: Brioschi, Riccardo, et al.
Published: (2025)
Diffusion-Based Ukrainian Handwritten Text Generation with Cross-Domain Style Transfer
by: Ahitoliev, Andrii, et al.
Published: (2026)
by: Ahitoliev, Andrii, et al.
Published: (2026)
FM-OSD: Foundation Model-Enabled One-Shot Detection of Anatomical Landmarks
by: Miao, Juzheng, et al.
Published: (2024)
by: Miao, Juzheng, et al.
Published: (2024)
Render-FM: A Foundation Model for Real-time Photorealistic Volumetric Rendering
by: Gao, Zhongpai, et al.
Published: (2025)
by: Gao, Zhongpai, et al.
Published: (2025)
FusionFM: Fusing Eye-specific Foundational Models for Optimized Ophthalmic Diagnosis
by: Zou, Ke, et al.
Published: (2025)
by: Zou, Ke, et al.
Published: (2025)
Online Handwritten Signature Verification Based on Temporal-Spatial Graph Attention Transformer
by: Yuan, Hai-jie, et al.
Published: (2025)
by: Yuan, Hai-jie, et al.
Published: (2025)
MapFM: Foundation Model-Driven HD Mapping with Multi-Task Contextual Learning
by: Ivanov, Leonid, et al.
Published: (2025)
by: Ivanov, Leonid, et al.
Published: (2025)
SHRUG-FM: Reliability-Aware Foundation Models for Earth Observation
by: Gonzalez-Calabuig, Maria, et al.
Published: (2025)
by: Gonzalez-Calabuig, Maria, et al.
Published: (2025)
Pan-FM: A Pan-Organ Foundation Model with Saliency-Guided Masking for Missing Robustness
by: Wu, Qiangqiang, et al.
Published: (2026)
by: Wu, Qiangqiang, et al.
Published: (2026)
Judge a Book by its Cover: Investigating Multi-Modal LLMs for Multi-Page Handwritten Document Transcription
by: Gutteridge, Benjamin, et al.
Published: (2025)
by: Gutteridge, Benjamin, et al.
Published: (2025)
HighFM: Towards a Foundation Model for Learning Representations from High-Frequency Earth Observation Data
by: Girtsou, Stella, et al.
Published: (2026)
by: Girtsou, Stella, et al.
Published: (2026)
Handwritten Text Recognition: A Survey
by: Garrido-Munoz, Carlos, et al.
Published: (2025)
by: Garrido-Munoz, Carlos, et al.
Published: (2025)
PLATTER: A Page-Level Handwritten Text Recognition System for Indic Scripts
by: Kasuba, Badri Vishal, et al.
Published: (2025)
by: Kasuba, Badri Vishal, et al.
Published: (2025)
An End-to-End, Segmentation-Free, Arabic Handwritten Recognition Model on KHATT
by: Aabed, Sondos, et al.
Published: (2024)
by: Aabed, Sondos, et al.
Published: (2024)
Multimodal, Multi-Disease Medical Imaging Foundation Model (MerMED-FM)
by: Zhou, Yang, et al.
Published: (2025)
by: Zhou, Yang, et al.
Published: (2025)
Understanding the Transfer Limits of Vision Foundation Models
by: Huang, Shiqi, et al.
Published: (2026)
by: Huang, Shiqi, et al.
Published: (2026)
Towards Scalable Training for Handwritten Mathematical Expression Recognition
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
Learning to Learn from Language Feedback with Social Meta-Learning
by: Cook, Jonathan, et al.
Published: (2026)
by: Cook, Jonathan, et al.
Published: (2026)
Benchmarking Pathology Foundation Models for Spatial Domain Understanding
by: Zhao, Bokai, et al.
Published: (2026)
by: Zhao, Bokai, et al.
Published: (2026)
EchoFM: Foundation Model for Generalizable Echocardiogram Analysis
by: Kim, Sekeun, et al.
Published: (2024)
by: Kim, Sekeun, et al.
Published: (2024)
VATr++: Choose Your Words Wisely for Handwritten Text Generation
by: Vanherle, Bram, et al.
Published: (2024)
by: Vanherle, Bram, et al.
Published: (2024)
Can AI Assistance Aid in the Grading of Handwritten Answer Sheets?
by: Sil, Pritam, et al.
Published: (2024)
by: Sil, Pritam, et al.
Published: (2024)
SimpleDoc: Multi-Modal Document Understanding with Dual-Cue Page Retrieval and Iterative Refinement
by: Jain, Chelsi, et al.
Published: (2025)
by: Jain, Chelsi, et al.
Published: (2025)
Pisces: An Auto-regressive Foundation Model for Image Understanding and Generation
by: Xu, Zhiyang, et al.
Published: (2025)
by: Xu, Zhiyang, et al.
Published: (2025)
SoccerMaster: A Vision Foundation Model for Soccer Understanding
by: Yang, Haolin, et al.
Published: (2025)
by: Yang, Haolin, et al.
Published: (2025)
EVLF-FM: Explainable Vision Language Foundation Model for Medicine
by: Bai, Yang, et al.
Published: (2025)
by: Bai, Yang, et al.
Published: (2025)
MedFM-Robust: Benchmarking Robustness of Medical Foundation Models
by: Cui, Xiangxiang, et al.
Published: (2026)
by: Cui, Xiangxiang, et al.
Published: (2026)
Mask & Match: Learning to Recognize Handwritten Math with Self-Supervised Attention
by: Mitra, Shree, et al.
Published: (2025)
by: Mitra, Shree, et al.
Published: (2025)
Bidirectional Trained Tree-Structured Decoder for Handwritten Mathematical Expression Recognition
by: Cheng, Hanbo, et al.
Published: (2023)
by: Cheng, Hanbo, et al.
Published: (2023)
Learning to Align: Addressing Character Frequency Distribution Shifts in Handwritten Text Recognition
by: Kaliosis, Panagiotis, et al.
Published: (2025)
by: Kaliosis, Panagiotis, et al.
Published: (2025)
SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery
by: Ma, Qiwei, et al.
Published: (2025)
by: Ma, Qiwei, et al.
Published: (2025)
SynthFM: Training Modality-agnostic Foundation Models for Medical Image Segmentation without Real Medical Data
by: Sengupta, Sourya, et al.
Published: (2025)
by: Sengupta, Sourya, et al.
Published: (2025)
FairMedFM: Fairness Benchmarking for Medical Imaging Foundation Models
by: Jin, Ruinan, et al.
Published: (2024)
by: Jin, Ruinan, et al.
Published: (2024)
SemiHMER: Semi-supervised Handwritten Mathematical Expression Recognition using pseudo-labels
by: Chen, Kehua, et al.
Published: (2025)
by: Chen, Kehua, et al.
Published: (2025)
RVAFM: Re-parameterizing Vertical Attention Fusion Module for Handwritten Paragraph Text Recognition
by: Zheng, Jinhui, et al.
Published: (2025)
by: Zheng, Jinhui, et al.
Published: (2025)
PyPotteryInk: One-Step Diffusion Model for Sketch to Publication-ready Archaeological Drawings
by: Cardarelli, Lorenzo
Published: (2025)
by: Cardarelli, Lorenzo
Published: (2025)
Seeing the Big Picture: Evaluating Multimodal LLMs' Ability to Interpret and Grade Handwritten Student Work
by: Henkel, Owen, et al.
Published: (2025)
by: Henkel, Owen, et al.
Published: (2025)
Similar Items
-
Representing Online Handwriting for Recognition in Large Vision-Language Models
by: Fadeeva, Anastasiia, et al.
Published: (2024) -
MathWriting: A Dataset For Handwritten Mathematical Expression Recognition
by: Gervais, Philippe, et al.
Published: (2024) -
InkSight: Offline-to-Online Handwriting Conversion by Teaching Vision-Language Models to Read and Write
by: Mitrevski, Blagoj, et al.
Published: (2024) -
Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation
by: Brioschi, Riccardo, et al.
Published: (2025) -
Diffusion-Based Ukrainian Handwritten Text Generation with Cross-Domain Style Transfer
by: Ahitoliev, Andrii, et al.
Published: (2026)