Fine-Tuning Vision-Language Models for Markdown Conversion of Financial Tables in Malaysian Audited Financial Reports
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tan, Jin Khye, Choong, En Jun, Chitty, Ethan Jeremiah, Choo, Yan Pheng, Wong, John Hsin Yang, Cheah, Chern Eu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Graphemic Normalization of the Perso-Arabic Script
von: Doctor, Raiomond, et al.
Veröffentlicht: (2022)
von: Doctor, Raiomond, et al.
Veröffentlicht: (2022)
Beyond Arabic: Software for Perso-Arabic Script Manipulation
von: Gutkin, Alexander, et al.
Veröffentlicht: (2023)
von: Gutkin, Alexander, et al.
Veröffentlicht: (2023)
"The Data Says Otherwise"-Towards Automated Fact-checking and Communication of Data Claims
von: Fu, Yu, et al.
Veröffentlicht: (2024)
von: Fu, Yu, et al.
Veröffentlicht: (2024)
A Semantic Approach to Negation Detection and Word Disambiguation with Natural Language Processing
von: Okpala, Izunna, et al.
Veröffentlicht: (2023)
von: Okpala, Izunna, et al.
Veröffentlicht: (2023)
S2Doc -- Spatial-Semantic Document Format
von: Kempf, Sebastian, et al.
Veröffentlicht: (2025)
von: Kempf, Sebastian, et al.
Veröffentlicht: (2025)
Co-Writing with AI, on Human Terms: Aligning Research with User Demands Across the Writing Process
von: Reza, Mohi, et al.
Veröffentlicht: (2025)
von: Reza, Mohi, et al.
Veröffentlicht: (2025)
HiPS: Hierarchical PDF Segmentation of Textbooks
von: Wehnert, Sabine, et al.
Veröffentlicht: (2025)
von: Wehnert, Sabine, et al.
Veröffentlicht: (2025)
Evaluating LLM Prompts for Data Augmentation in Multi-label Classification of Ecological Texts
von: Glazkova, Anna, et al.
Veröffentlicht: (2024)
von: Glazkova, Anna, et al.
Veröffentlicht: (2024)
AI-Friendly LaTeX: Using LaTeX Code as a Knowledge Source for Retrieval-Augmented Generation
von: Verhoeff, Tom
Veröffentlicht: (2026)
von: Verhoeff, Tom
Veröffentlicht: (2026)
Fine-Tuning LLMs with Fine-Grained Human Feedback on Text Spans
von: CH-Wang, Sky, et al.
Veröffentlicht: (2025)
von: CH-Wang, Sky, et al.
Veröffentlicht: (2025)
Extracting Structured Insights from Financial News: An Augmented LLM Driven Approach
von: Dolphin, Rian, et al.
Veröffentlicht: (2024)
von: Dolphin, Rian, et al.
Veröffentlicht: (2024)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
Can AI Read Between The Lines? Benchmarking LLMs On Financial Nuance
von: Kubica, Dominick, et al.
Veröffentlicht: (2025)
von: Kubica, Dominick, et al.
Veröffentlicht: (2025)
Intent Classification for Bank Chatbots through LLM Fine-Tuning
von: Lajčinová, Bibiána, et al.
Veröffentlicht: (2024)
von: Lajčinová, Bibiána, et al.
Veröffentlicht: (2024)
The Data Efficiency Frontier of Financial Foundation Models: Scaling Laws from Continued Pretraining
von: Ponnock, Jesse
Veröffentlicht: (2025)
von: Ponnock, Jesse
Veröffentlicht: (2025)
Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
von: Yang, Shan
Veröffentlicht: (2026)
von: Yang, Shan
Veröffentlicht: (2026)
Fin-ExBERT: User Intent based Text Extraction in Financial Context using Graph-Augmented BERT and trainable Plugin
von: Sarker, Soumick, et al.
Veröffentlicht: (2025)
von: Sarker, Soumick, et al.
Veröffentlicht: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
Enhancing the Reasoning Capabilities of Small Language Models via Solution Guidance Fine-Tuning
von: Bi, Jing, et al.
Veröffentlicht: (2024)
von: Bi, Jing, et al.
Veröffentlicht: (2024)
Analog Circuit Sizing Using Machine Learning Based Transistor Circuit Model
von: Rajeoni, Alireza Bagheri
Veröffentlicht: (2023)
von: Rajeoni, Alireza Bagheri
Veröffentlicht: (2023)
EmoS: A High-Fidelity Multimodal Benchmark for Fine-grained Streaming Emotional Understanding
von: Guo, Pengze, et al.
Veröffentlicht: (2026)
von: Guo, Pengze, et al.
Veröffentlicht: (2026)
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
von: Gonzalez, Alberto Andres Valdes
Veröffentlicht: (2026)
von: Gonzalez, Alberto Andres Valdes
Veröffentlicht: (2026)
Key-Value Means: Transformers with Expandable Block-Recurrent Compressed Memory
von: Goldstein, Daniel, et al.
Veröffentlicht: (2026)
von: Goldstein, Daniel, et al.
Veröffentlicht: (2026)
Annotation Entropy Predicts Per-Example Learning Dynamics in LoRA Fine-Tuning
von: Steele, Brady
Veröffentlicht: (2026)
von: Steele, Brady
Veröffentlicht: (2026)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
von: Han, Xudong, et al.
Veröffentlicht: (2025)
von: Han, Xudong, et al.
Veröffentlicht: (2025)
Kastor: Fine-tuned Small Language Models for Shape-based Active Relation Extraction
von: Celian, Ringwald, et al.
Veröffentlicht: (2025)
von: Celian, Ringwald, et al.
Veröffentlicht: (2025)
ReaderLM-v2: Small Language Model for HTML to Markdown and JSON
von: Wang, Feng, et al.
Veröffentlicht: (2025)
von: Wang, Feng, et al.
Veröffentlicht: (2025)
Bridging the Gap: An Intermediate Language for Enhanced and Cost-Effective Grapheme-to-Phoneme Conversion with Homographs with Multiple Pronunciations Disambiguation
von: Bertina, Abbas, et al.
Veröffentlicht: (2025)
von: Bertina, Abbas, et al.
Veröffentlicht: (2025)
Improving the Capabilities of Large Language Model Based Marketing Analytics Copilots With Semantic Search And Fine-Tuning
von: Gao, Yilin, et al.
Veröffentlicht: (2024)
von: Gao, Yilin, et al.
Veröffentlicht: (2024)
KTBox: A Modular LaTeX Framework for Semantic Color, Structured Highlighting, and Scholarly Communication
von: Mangal, Bhaskar, et al.
Veröffentlicht: (2025)
von: Mangal, Bhaskar, et al.
Veröffentlicht: (2025)
Robustness of Large Language Models to Perturbations in Text
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
UM_FHS at the CLEF 2025 SimpleText Track: Comparing No-Context and Fine-Tune Approaches for GPT-4.1 Models in Sentence and Document-Level Text Simplification
von: Kocbek, Primoz, et al.
Veröffentlicht: (2025)
von: Kocbek, Primoz, et al.
Veröffentlicht: (2025)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
von: Saji, Alan, et al.
Veröffentlicht: (2025)
von: Saji, Alan, et al.
Veröffentlicht: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
VLEU: a Method for Automatic Evaluation for Generalizability of Text-to-Image Models
von: Cao, Jingtao, et al.
Veröffentlicht: (2024)
von: Cao, Jingtao, et al.
Veröffentlicht: (2024)
Conversations with Andrea: Visitors' Opinions on Android Robots in a Museum
von: Heisler, Marcel, et al.
Veröffentlicht: (2025)
von: Heisler, Marcel, et al.
Veröffentlicht: (2025)
Measurement Risk in Supervised Financial NLP: Rubric and Metric Sensitivity on JF-ICR
von: Chang, Sidi, et al.
Veröffentlicht: (2026)
von: Chang, Sidi, et al.
Veröffentlicht: (2026)
Photonic AI: A Hybrid Diffractive Holographic Neural System for Passive Optical Real-Time Image Classification
von: Hiremath, Prakul Sunil
Veröffentlicht: (2026)
von: Hiremath, Prakul Sunil
Veröffentlicht: (2026)
Ähnliche Einträge
-
Graphemic Normalization of the Perso-Arabic Script
von: Doctor, Raiomond, et al.
Veröffentlicht: (2022) -
Beyond Arabic: Software for Perso-Arabic Script Manipulation
von: Gutkin, Alexander, et al.
Veröffentlicht: (2023) -
"The Data Says Otherwise"-Towards Automated Fact-checking and Communication of Data Claims
von: Fu, Yu, et al.
Veröffentlicht: (2024) -
A Semantic Approach to Negation Detection and Word Disambiguation with Natural Language Processing
von: Okpala, Izunna, et al.
Veröffentlicht: (2023) -
S2Doc -- Spatial-Semantic Document Format
von: Kempf, Sebastian, et al.
Veröffentlicht: (2025)