TC-OCR: TableCraft OCR for Efficient Detection & Recognition of Table Structure & Content
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Anand, Avinash, Jaiswal, Raj, Bhuyan, Pijush, Gupta, Mohit, Bangar, Siddhesh, Imam, Md. Modassir, Shah, Rajiv Ratn, Satoh, Shin'ichi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RanLayNet: A Dataset for Document Layout Detection used for Domain Adaptation and Generalization
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
Certified Zeroth-order Black-Box Defense with Robust UNet Denoiser
von: Verma, Astha, et al.
Veröffentlicht: (2023)
von: Verma, Astha, et al.
Veröffentlicht: (2023)
Advancements in Scientific Controllable Text Generation Methods
von: Goel, Arnav, et al.
Veröffentlicht: (2023)
von: Goel, Arnav, et al.
Veröffentlicht: (2023)
Su-RoBERTa: A Semi-supervised Approach to Predicting Suicide Risk through Social Media using Base Language Models
von: Tank, Chayan, et al.
Veröffentlicht: (2024)
von: Tank, Chayan, et al.
Veröffentlicht: (2024)
Improving Physics Reasoning in Large Language Models Using Mixture of Refinement Agents
von: Jaiswal, Raj, et al.
Veröffentlicht: (2024)
von: Jaiswal, Raj, et al.
Veröffentlicht: (2024)
Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
Med-CoDE: Medical Critique based Disagreement Evaluation Framework
von: Gupta, Mohit, et al.
Veröffentlicht: (2025)
von: Gupta, Mohit, et al.
Veröffentlicht: (2025)
Mathify: Evaluating Large Language Models on Mathematical Problem Solving Tasks
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
KG-CTG: Citation Generation through Knowledge Graph-guided Large Language Models
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
Context-Enhanced Language Models for Generating Multi-Paper Citations
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
Improving Multimodal LLMs Ability In Geometry Problem Solving, Reasoning, And Multistep Scoring
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
Depression Detection and Analysis using Large Language Models on Textual and Audio-Visual Modalities
von: Tank, Chayan, et al.
Veröffentlicht: (2024)
von: Tank, Chayan, et al.
Veröffentlicht: (2024)
ReSeDis: A Dataset for Referring-based Object Search across Large-Scale Image Collections
von: Huang, Ziling, et al.
Veröffentlicht: (2025)
von: Huang, Ziling, et al.
Veröffentlicht: (2025)
Uplifting Table Tennis: A Robust, Real-World Application for 3D Trajectory and Spin Estimation
von: Kienzle, Daniel, et al.
Veröffentlicht: (2025)
von: Kienzle, Daniel, et al.
Veröffentlicht: (2025)
On Optimal Steering to Achieve Exact Fairness
von: Sharma, Mohit, et al.
Veröffentlicht: (2025)
von: Sharma, Mohit, et al.
Veröffentlicht: (2025)
Keystroke Dynamics Against Academic Dishonesty in the Age of LLMs
von: Kundu, Debnath, et al.
Veröffentlicht: (2024)
von: Kundu, Debnath, et al.
Veröffentlicht: (2024)
Enhancing LLMs for Physics Problem-Solving using Reinforcement Learning with Human-AI Feedback
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
OCR-Agent: Agentic OCR with Capability and Memory Reflection
von: Wen, Shimin, et al.
Veröffentlicht: (2026)
von: Wen, Shimin, et al.
Veröffentlicht: (2026)
OmniOCR: Generalist OCR for Ethnic Minority Languages
von: Liu, Bonan, et al.
Veröffentlicht: (2026)
von: Liu, Bonan, et al.
Veröffentlicht: (2026)
Speech Representation Learning Revisited: The Necessity of Separate Learnable Parameters and Robust Data Augmentation
von: Yadav, Hemant, et al.
Veröffentlicht: (2024)
von: Yadav, Hemant, et al.
Veröffentlicht: (2024)
LittiChoQA: Literary Texts in Indic Languages Chosen for Question Answering
von: Khandelwal, Aarya, et al.
Veröffentlicht: (2026)
von: Khandelwal, Aarya, et al.
Veröffentlicht: (2026)
RConE: Rough Cone Embedding for Multi-Hop Logical Query Answering on Multi-Modal Knowledge Graphs
von: Kharbanda, Mayank, et al.
Veröffentlicht: (2024)
von: Kharbanda, Mayank, et al.
Veröffentlicht: (2024)
Analysing the Masked predictive coding training criterion for pre-training a Speech Representation Model
von: Yadav, Hemant, et al.
Veröffentlicht: (2023)
von: Yadav, Hemant, et al.
Veröffentlicht: (2023)
Long-context Non-factoid Question Answering in Indic Languages
von: Mishra, Ritwik, et al.
Veröffentlicht: (2025)
von: Mishra, Ritwik, et al.
Veröffentlicht: (2025)
MS-HuBERT: Mitigating Pre-training and Inference Mismatch in Masked Language Modelling methods for learning Speech Representations
von: Yadav, Hemant, et al.
Veröffentlicht: (2024)
von: Yadav, Hemant, et al.
Veröffentlicht: (2024)
JOOCI: a Framework for Learning Comprehensive Speech Representations
von: Yadav, Hemant, et al.
Veröffentlicht: (2024)
von: Yadav, Hemant, et al.
Veröffentlicht: (2024)
The Effects of Short Video-Sharing Services on Video Copy Detection
von: Yanagi, Rintaro, et al.
Veröffentlicht: (2024)
von: Yanagi, Rintaro, et al.
Veröffentlicht: (2024)
olmOCR 2: Unit Test Rewards for Document OCR
von: Poznanski, Jake, et al.
Veröffentlicht: (2025)
von: Poznanski, Jake, et al.
Veröffentlicht: (2025)
PubMed-OCR: PMC Open Access OCR Annotations
von: Heidenreich, Hunter, et al.
Veröffentlicht: (2026)
von: Heidenreich, Hunter, et al.
Veröffentlicht: (2026)
Chapter Слушая Сад: переосмысление звуковых аспектов чеховской пьесы на примере японской и итальянской постановок
von: Murata, Shin’ichi
Veröffentlicht: (2024)
von: Murata, Shin’ichi
Veröffentlicht: (2024)
Probabilistic Online Event Downsampling
von: Girbau-Xalabarder, Andreu, et al.
Veröffentlicht: (2025)
von: Girbau-Xalabarder, Andreu, et al.
Veröffentlicht: (2025)
OCR-Quality: A Human-Annotated Dataset for OCR Quality Assessment
von: Zhang, Yulong
Veröffentlicht: (2025)
von: Zhang, Yulong
Veröffentlicht: (2025)
Design and Implementation of an OCR-Powered Pipeline for Table Extraction from Invoices
von: Patel, Parshva Dhilankumar
Veröffentlicht: (2025)
von: Patel, Parshva Dhilankumar
Veröffentlicht: (2025)
Fairness Without Labels: Pseudo-Balancing for Bias Mitigation in Face Gender Classification
von: Dong, Haohua, et al.
Veröffentlicht: (2025)
von: Dong, Haohua, et al.
Veröffentlicht: (2025)
MM-PhyRLHF: Reinforcement Learning Framework for Multimodal Physics Question-Answering
von: Kapuriya, Janak, et al.
Veröffentlicht: (2024)
von: Kapuriya, Janak, et al.
Veröffentlicht: (2024)
Steps are all you need: Rethinking STEM Education with Prompt Engineering
von: Addala, Krishnasai, et al.
Veröffentlicht: (2024)
von: Addala, Krishnasai, et al.
Veröffentlicht: (2024)
Better and Worse with Scale: How Contextual Entrainment Diverges with Model Size
von: Kukreja, Dikshant, et al.
Veröffentlicht: (2026)
von: Kukreja, Dikshant, et al.
Veröffentlicht: (2026)
Agentar-Fin-OCR
von: Qian, Siyi, et al.
Veröffentlicht: (2026)
von: Qian, Siyi, et al.
Veröffentlicht: (2026)
Guidelines for External Disturbance Factors in the Use of OCR in Real-World Environments
von: Iwata, Kenji, et al.
Veröffentlicht: (2025)
von: Iwata, Kenji, et al.
Veröffentlicht: (2025)
When Good OCR Is Not Enough: Benchmarking OCR Robustness for Retrieval-Augmented Generation
von: Sun, Lin, et al.
Veröffentlicht: (2026)
von: Sun, Lin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
RanLayNet: A Dataset for Document Layout Detection used for Domain Adaptation and Generalization
von: Anand, Avinash, et al.
Veröffentlicht: (2024) -
Certified Zeroth-order Black-Box Defense with Robust UNet Denoiser
von: Verma, Astha, et al.
Veröffentlicht: (2023) -
Advancements in Scientific Controllable Text Generation Methods
von: Goel, Arnav, et al.
Veröffentlicht: (2023) -
Su-RoBERTa: A Semi-supervised Approach to Predicting Suicide Risk through Social Media using Base Language Models
von: Tank, Chayan, et al.
Veröffentlicht: (2024) -
Improving Physics Reasoning in Large Language Models Using Mixture of Refinement Agents
von: Jaiswal, Raj, et al.
Veröffentlicht: (2024)