Seeing Justice Clearly: Handwritten Legal Document Translation with OCR and Vision-Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Nigam, Shubham Kumar, Shukla, Parjanya Aditya, Shallum, Noel, Bhattacharya, Arnab |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
NyayaMind- A Framework for Transparent Legal Reasoning and Judgment Prediction in the Indian Legal System
por: Shukla, Parjanya Aditya, et al.
Publicado: (2026)
por: Shukla, Parjanya Aditya, et al.
Publicado: (2026)
Segment First, Retrieve Better: Realistic Legal Search via Rhetorical Role-Based Queries
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
Structured Legal Document Generation in India: A Model-Agnostic Wrapper Approach with VidhikDastaavej
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
LegalSeg: Unlocking the Structure of Indian Legal Judgments Through Rhetorical Role Classification
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
Legal Judgment Reimagined: PredEx and the Rise of Intelligent AI Interpretation in Indian Courts
por: Nigam, Shubham Kumar, et al.
Publicado: (2024)
por: Nigam, Shubham Kumar, et al.
Publicado: (2024)
NyayaAnumana & INLegalLlama: The Largest Indian Legal Judgment Prediction Dataset and Specialized Language Model for Enhanced Decision Analysis
por: Nigam, Shubham Kumar, et al.
Publicado: (2024)
por: Nigam, Shubham Kumar, et al.
Publicado: (2024)
Seeing Straight: Document Orientation Detection for Efficient OCR
por: Goswami, Suranjan, et al.
Publicado: (2025)
por: Goswami, Suranjan, et al.
Publicado: (2025)
TathyaNyaya and FactLegalLlama: Advancing Factual Judgment Prediction and Explanation in the Indian Legal Context
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
ReGal: A First Look at PPO-based Legal AI for Judgment Prediction and Summarization in India
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
IBPS: Indian Bail Prediction System
por: Srivastava, Puspesh Kumar, et al.
Publicado: (2025)
por: Srivastava, Puspesh Kumar, et al.
Publicado: (2025)
Baseer: A Vision-Language Model for Arabic Document-to-Markdown OCR
por: Hennara, Khalil, et al.
Publicado: (2025)
por: Hennara, Khalil, et al.
Publicado: (2025)
Can Vision-Language Models Evaluate Handwritten Math?
por: Nath, Oikantik, et al.
Publicado: (2025)
por: Nath, Oikantik, et al.
Publicado: (2025)
NyayaRAG: Realistic Legal Judgment Prediction with RAG under the Indian Common Law System
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
por: Nigam, Shubham Kumar, et al.
Publicado: (2025)
Rethinking Legal Judgement Prediction in a Realistic Scenario in the Era of Large Language Models
por: Nigam, Shubham Kumar, et al.
Publicado: (2024)
por: Nigam, Shubham Kumar, et al.
Publicado: (2024)
olmOCR 2: Unit Test Rewards for Document OCR
por: Poznanski, Jake, et al.
Publicado: (2025)
por: Poznanski, Jake, et al.
Publicado: (2025)
Analyzing Images of Legal Documents: Toward Multi-Modal LLMs for Access to Justice
por: Westermann, Hannes, et al.
Publicado: (2024)
por: Westermann, Hannes, et al.
Publicado: (2024)
GutenOCR: A Grounded Vision-Language Front-End for Documents
por: Heidenreich, Hunter, et al.
Publicado: (2026)
por: Heidenreich, Hunter, et al.
Publicado: (2026)
Inference-Time Structural Reasoning for Compositional Vision-Language Understanding
por: Bhattacharya, Amartya
Publicado: (2026)
por: Bhattacharya, Amartya
Publicado: (2026)
PreP-OCR: A Complete Pipeline for Document Image Restoration and Enhanced OCR Accuracy
por: Guan, Shuhao, et al.
Publicado: (2025)
por: Guan, Shuhao, et al.
Publicado: (2025)
Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models
por: Khan, Mohammed Safi Ur Rahman, et al.
Publicado: (2026)
por: Khan, Mohammed Safi Ur Rahman, et al.
Publicado: (2026)
Show, don't tell -- Providing Visual Error Feedback for Handwritten Documents
por: Yasin, Said, et al.
Publicado: (2026)
por: Yasin, Said, et al.
Publicado: (2026)
Unlocking the Archives: Using Large Language Models to Transcribe Handwritten Historical Documents
por: Humphries, Mark, et al.
Publicado: (2024)
por: Humphries, Mark, et al.
Publicado: (2024)
"See the World, Discover Knowledge": A Chinese Factuality Evaluation for Large Vision Language Models
por: Gu, Jihao, et al.
Publicado: (2025)
por: Gu, Jihao, et al.
Publicado: (2025)
From Seeing to Thinking: Decoupling Perception and Reasoning Improves Post-Training of Vision-Language Models
por: Wu, Juncheng, et al.
Publicado: (2026)
por: Wu, Juncheng, et al.
Publicado: (2026)
Zero-Shot Product Attribute Labeling with Vision-Language Models: A Three-Tier Evaluation Framework
por: Shukla, Shubham, et al.
Publicado: (2026)
por: Shukla, Shubham, et al.
Publicado: (2026)
Improving MLLM's Document Image Machine Translation via Synchronously Self-reviewing Its OCR Proficiency
por: Liang, Yupu, et al.
Publicado: (2025)
por: Liang, Yupu, et al.
Publicado: (2025)
Improving OCR for Historical Texts of Multiple Languages
por: Westerdijk, Hylke, et al.
Publicado: (2025)
por: Westerdijk, Hylke, et al.
Publicado: (2025)
Seeing No Evil: Blinding Large Vision-Language Models to Safety Instructions via Adversarial Attention Hijacking
por: Li, Jingru, et al.
Publicado: (2026)
por: Li, Jingru, et al.
Publicado: (2026)
Seeing Syntax: Uncovering Syntactic Learning Limitations in Vision-Language Models
por: Dumpala, Sri Harsha, et al.
Publicado: (2024)
por: Dumpala, Sri Harsha, et al.
Publicado: (2024)
The Illusion-Illusion: Vision Language Models See Illusions Where There are None
por: Ullman, Tomer
Publicado: (2024)
por: Ullman, Tomer
Publicado: (2024)
GlotOCR Bench: OCR Models Still Struggle Beyond a Handful of Unicode Scripts
por: Kargaran, Amir Hossein, et al.
Publicado: (2026)
por: Kargaran, Amir Hossein, et al.
Publicado: (2026)
Seeing Clearly, Reasoning Confidently: Plug-and-Play Remedies for Vision Language Model Blindness
por: Hu, Xin, et al.
Publicado: (2026)
por: Hu, Xin, et al.
Publicado: (2026)
Automated LaTeX Code Generation from Handwritten Math Expressions Using Vision Transformer
por: Sundararaj, Jayaprakash, et al.
Publicado: (2024)
por: Sundararaj, Jayaprakash, et al.
Publicado: (2024)
Confidence-Aware Document OCR Error Detection
por: Hemmer, Arthur, et al.
Publicado: (2024)
por: Hemmer, Arthur, et al.
Publicado: (2024)
DharmaOCR: Specialized Small Language Models for Structured OCR that outperform Open-Source and Commercial Baselines
por: Cardoso, Gabriel Pimenta de Freitas, et al.
Publicado: (2026)
por: Cardoso, Gabriel Pimenta de Freitas, et al.
Publicado: (2026)
KazakhOCR: A Synthetic Benchmark for Evaluating Multimodal Models in Low-Resource Kazakh Script OCR
por: Gagnier, Henry, et al.
Publicado: (2026)
por: Gagnier, Henry, et al.
Publicado: (2026)
Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models
por: Wu, Jiaying, et al.
Publicado: (2025)
por: Wu, Jiaying, et al.
Publicado: (2025)
Seeing Through Their Eyes: Evaluating Visual Perspective Taking in Vision Language Models
por: Góral, Gracjan, et al.
Publicado: (2024)
por: Góral, Gracjan, et al.
Publicado: (2024)
RoundTripOCR: A Data Generation Technique for Enhancing Post-OCR Error Correction in Low-Resource Devanagari Languages
por: Kashid, Harshvivek, et al.
Publicado: (2024)
por: Kashid, Harshvivek, et al.
Publicado: (2024)
BHDD: A Burmese Handwritten Digit Dataset
por: Aung, Swan Htet, et al.
Publicado: (2026)
por: Aung, Swan Htet, et al.
Publicado: (2026)
Ejemplares similares
-
NyayaMind- A Framework for Transparent Legal Reasoning and Judgment Prediction in the Indian Legal System
por: Shukla, Parjanya Aditya, et al.
Publicado: (2026) -
Segment First, Retrieve Better: Realistic Legal Search via Rhetorical Role-Based Queries
por: Nigam, Shubham Kumar, et al.
Publicado: (2025) -
Structured Legal Document Generation in India: A Model-Agnostic Wrapper Approach with VidhikDastaavej
por: Nigam, Shubham Kumar, et al.
Publicado: (2025) -
LegalSeg: Unlocking the Structure of Indian Legal Judgments Through Rhetorical Role Classification
por: Nigam, Shubham Kumar, et al.
Publicado: (2025) -
Legal Judgment Reimagined: PredEx and the Rise of Intelligent AI Interpretation in Indian Courts
por: Nigam, Shubham Kumar, et al.
Publicado: (2024)