Enhancing Large Vision-Language Models with Layout Modality for Table Question Answering on Japanese Annual Securities Reports
Fuente:
arXiv
Salvato in:
| Autori principali: | Aida, Hayato, Takahashi, Kosuke, Omi, Takahiro |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Pretraining and Updates of Domain-Specific LLM: A Case Study in the Japanese Business Domain
di: Takahashi, Kosuke, et al.
Pubblicazione: (2024)
di: Takahashi, Kosuke, et al.
Pubblicazione: (2024)
Towards Probabilistic Question Answering Over Tabular Data
di: Shen, Chen, et al.
Pubblicazione: (2025)
di: Shen, Chen, et al.
Pubblicazione: (2025)
Enhancing OCR for Sino-Vietnamese Language Processing via Fine-tuned PaddleOCRv5
di: Nguyen, Minh Hoang, et al.
Pubblicazione: (2025)
di: Nguyen, Minh Hoang, et al.
Pubblicazione: (2025)
A Graph-based Approach for Multi-Modal Question Answering from Flowcharts in Telecom Documents
di: Soman, Sumit, et al.
Pubblicazione: (2025)
di: Soman, Sumit, et al.
Pubblicazione: (2025)
PerkwE_COQA: Enhanced Persian Conversational Question Answering by combining contextual keyword extraction with Large Language Models
di: Moradbeiki, Pardis, et al.
Pubblicazione: (2024)
di: Moradbeiki, Pardis, et al.
Pubblicazione: (2024)
MORQA: Benchmarking Evaluation Metrics for Medical Open-Ended Question Answering
di: Yim, Wen-wai, et al.
Pubblicazione: (2025)
di: Yim, Wen-wai, et al.
Pubblicazione: (2025)
Can Out-of-Distribution Evaluations Uncover Reliance on Shortcuts? A Case Study in Question Answering
di: Štefánik, Michal, et al.
Pubblicazione: (2025)
di: Štefánik, Michal, et al.
Pubblicazione: (2025)
Evaluating Prompting Strategies for Chart Question Answering with Large Language Models
di: Naikar, Ruthuparna, et al.
Pubblicazione: (2026)
di: Naikar, Ruthuparna, et al.
Pubblicazione: (2026)
SUGARCREPE++ Dataset: Vision-Language Model Sensitivity to Semantic and Lexical Alterations
di: Dumpala, Sri Harsha, et al.
Pubblicazione: (2024)
di: Dumpala, Sri Harsha, et al.
Pubblicazione: (2024)
The Superalignment of Superhuman Intelligence with Large Language Models
di: Huang, Minlie, et al.
Pubblicazione: (2024)
di: Huang, Minlie, et al.
Pubblicazione: (2024)
An Unforgeable Publicly Verifiable Watermark for Large Language Models
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
Exploring State Tracking Capabilities of Large Language Models
di: Rezaee, Kiamehr, et al.
Pubblicazione: (2025)
di: Rezaee, Kiamehr, et al.
Pubblicazione: (2025)
A Survey of Text Watermarking in the Era of Large Language Models
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
Distilling Large Language Models for Efficient Clinical Information Extraction
di: Vedula, Karthik S., et al.
Pubblicazione: (2024)
di: Vedula, Karthik S., et al.
Pubblicazione: (2024)
Quantifying Genuine Awareness in Hallucination Prediction Beyond Question-Side Shortcuts
di: Seo, Yeongbin, et al.
Pubblicazione: (2025)
di: Seo, Yeongbin, et al.
Pubblicazione: (2025)
jina-vlm: Small Multilingual Vision Language Model
di: Koukounas, Andreas, et al.
Pubblicazione: (2025)
di: Koukounas, Andreas, et al.
Pubblicazione: (2025)
Triad: A Framework Leveraging a Multi-Role LLM-based Agent to Solve Knowledge Base Question Answering
di: Zong, Chang, et al.
Pubblicazione: (2024)
di: Zong, Chang, et al.
Pubblicazione: (2024)
Towards Effective and Efficient Continual Pre-training of Large Language Models
di: Chen, Jie, et al.
Pubblicazione: (2024)
di: Chen, Jie, et al.
Pubblicazione: (2024)
Unifying Uniform and Binary-coding Quantization for Accurate Compression of Large Language Models
di: Park, Seungcheol, et al.
Pubblicazione: (2025)
di: Park, Seungcheol, et al.
Pubblicazione: (2025)
Accurate Sublayer Pruning for Large Language Models by Exploiting Latency and Tunability Information
di: Park, Seungcheol, et al.
Pubblicazione: (2025)
di: Park, Seungcheol, et al.
Pubblicazione: (2025)
Setting Standards in Turkish NLP: TR-MMLU for Large Language Model Evaluation
di: Bayram, M. Ali, et al.
Pubblicazione: (2024)
di: Bayram, M. Ali, et al.
Pubblicazione: (2024)
Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation
di: Liu, Aiwei, et al.
Pubblicazione: (2024)
di: Liu, Aiwei, et al.
Pubblicazione: (2024)
Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey
di: Fang, Xi, et al.
Pubblicazione: (2024)
di: Fang, Xi, et al.
Pubblicazione: (2024)
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
di: Pan, Leyi, et al.
Pubblicazione: (2025)
di: Pan, Leyi, et al.
Pubblicazione: (2025)
The Paradox of Poetic Intent in Back-Translation: Evaluating the Quality of Large Language Models in Chinese Translation
di: Weigang, Li, et al.
Pubblicazione: (2025)
di: Weigang, Li, et al.
Pubblicazione: (2025)
Unveiling Attractor Cycles in Large Language Models: A Dynamical Systems View of Successive Paraphrasing
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
NurValues: Real-World Nursing Values Evaluation for Large Language Models in Clinical Context
di: Yao, Ben, et al.
Pubblicazione: (2025)
di: Yao, Ben, et al.
Pubblicazione: (2025)
Large Language Models Report Subjective Experience Under Self-Referential Processing
di: Berg, Cameron, et al.
Pubblicazione: (2025)
di: Berg, Cameron, et al.
Pubblicazione: (2025)
Trusted Uncertainty in Large Language Models: A Unified Framework for Confidence Calibration and Risk-Controlled Refusal
di: Oehri, Markus, et al.
Pubblicazione: (2025)
di: Oehri, Markus, et al.
Pubblicazione: (2025)
"As Eastern Powers, I will veto." : An Investigation of Nation-level Bias of Large Language Models in International Relations
di: Choi, Jonghyeon, et al.
Pubblicazione: (2025)
di: Choi, Jonghyeon, et al.
Pubblicazione: (2025)
Knowledge Distillation of Domain-adapted LLMs for Question-Answering in Telecom
di: Sen, Rishika, et al.
Pubblicazione: (2025)
di: Sen, Rishika, et al.
Pubblicazione: (2025)
Clarification as Supervision: Reinforcement Learning for Vision-Language Interfaces
di: Gkountouras, John, et al.
Pubblicazione: (2025)
di: Gkountouras, John, et al.
Pubblicazione: (2025)
Transforming and Combining Rewards for Aligning Large Language Models
di: Wang, Zihao, et al.
Pubblicazione: (2024)
di: Wang, Zihao, et al.
Pubblicazione: (2024)
Evaluating Pixel Language Models on Non-Standardized Languages
di: Muñoz-Ortiz, Alberto, et al.
Pubblicazione: (2024)
di: Muñoz-Ortiz, Alberto, et al.
Pubblicazione: (2024)
A Survey on Vision-Language-Action Models for Embodied AI
di: Ma, Yueen, et al.
Pubblicazione: (2024)
di: Ma, Yueen, et al.
Pubblicazione: (2024)
Towards Ontology-Enhanced Representation Learning for Large Language Models
di: Ronzano, Francesco, et al.
Pubblicazione: (2024)
di: Ronzano, Francesco, et al.
Pubblicazione: (2024)
IRONIC: Coherence-Aware Reasoning Chains for Multi-Modal Sarcasm Detection
di: Ramakrishnan, Aashish Anantha, et al.
Pubblicazione: (2025)
di: Ramakrishnan, Aashish Anantha, et al.
Pubblicazione: (2025)
NERCat: Fine-Tuning for Enhanced Named Entity Recognition in Catalan
di: Ferreres, Guillem Cadevall, et al.
Pubblicazione: (2025)
di: Ferreres, Guillem Cadevall, et al.
Pubblicazione: (2025)
Prompt Engineering and the Effectiveness of Large Language Models in Enhancing Human Productivity
di: Anam, Rizal Khoirul
Pubblicazione: (2025)
di: Anam, Rizal Khoirul
Pubblicazione: (2025)
Continuous Latent Diffusion Language Model
di: Guo, Hongcan, et al.
Pubblicazione: (2026)
di: Guo, Hongcan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Pretraining and Updates of Domain-Specific LLM: A Case Study in the Japanese Business Domain
di: Takahashi, Kosuke, et al.
Pubblicazione: (2024) -
Towards Probabilistic Question Answering Over Tabular Data
di: Shen, Chen, et al.
Pubblicazione: (2025) -
Enhancing OCR for Sino-Vietnamese Language Processing via Fine-tuned PaddleOCRv5
di: Nguyen, Minh Hoang, et al.
Pubblicazione: (2025) -
A Graph-based Approach for Multi-Modal Question Answering from Flowcharts in Telecom Documents
di: Soman, Sumit, et al.
Pubblicazione: (2025) -
PerkwE_COQA: Enhanced Persian Conversational Question Answering by combining contextual keyword extraction with Large Language Models
di: Moradbeiki, Pardis, et al.
Pubblicazione: (2024)