TALENT: Table VQA via Augmented Language-Enhanced Natural-text Transcription
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yutong, Guo, Wang, Wanying, Wu, Yue, Miao, Zichen, Wang, Haoyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Embedding Domain Knowledge for Large Language Models via Reinforcement Learning from Augmented Generation
von: Nie, Chaojun, et al.
Veröffentlicht: (2025)
von: Nie, Chaojun, et al.
Veröffentlicht: (2025)
ActTraitBench: Quantifying the Knowledge-Decision Gap in Large Language Models via Human-Grounded Behavioral Validation
von: Yang, Yutong, et al.
Veröffentlicht: (2026)
von: Yang, Yutong, et al.
Veröffentlicht: (2026)
Layer Importance and Hallucination Analysis in Large Language Models via Enhanced Activation Variance-Sparsity
von: Song, Zichen, et al.
Veröffentlicht: (2024)
von: Song, Zichen, et al.
Veröffentlicht: (2024)
Multimodal Retrieval-Augmented Generation with Large Language Models for Medical VQA
von: Karim, A H M Rezaul, et al.
Veröffentlicht: (2025)
von: Karim, A H M Rezaul, et al.
Veröffentlicht: (2025)
Benchmarking Multimodal Retrieval Augmented Generation with Dynamic VQA Dataset and Self-adaptive Planning Agent
von: Li, Yangning, et al.
Veröffentlicht: (2024)
von: Li, Yangning, et al.
Veröffentlicht: (2024)
Memory-Augmented Multimodal LLMs for Surgical VQA via Self-Contained Inquiry
von: Hou, Wenjun, et al.
Veröffentlicht: (2024)
von: Hou, Wenjun, et al.
Veröffentlicht: (2024)
Muffin or Chihuahua? Challenging Multimodal Large Language Models with Multipanel VQA
von: Fan, Yue, et al.
Veröffentlicht: (2024)
von: Fan, Yue, et al.
Veröffentlicht: (2024)
SimpleVQA: Multimodal Factuality Evaluation for Multimodal Large Language Models
von: Cheng, Xianfu, et al.
Veröffentlicht: (2025)
von: Cheng, Xianfu, et al.
Veröffentlicht: (2025)
AVSS: Layer Importance Evaluation in Large Language Models via Activation Variance-Sparsity Analysis
von: Song, Zichen, et al.
Veröffentlicht: (2024)
von: Song, Zichen, et al.
Veröffentlicht: (2024)
Retrieval-Augmented Generation for Natural Language Processing: A Survey
von: Wu, Shangyu, et al.
Veröffentlicht: (2024)
von: Wu, Shangyu, et al.
Veröffentlicht: (2024)
BlendFilter: Advancing Retrieval-Augmented Large Language Models via Query Generation Blending and Knowledge Filtering
von: Wang, Haoyu, et al.
Veröffentlicht: (2024)
von: Wang, Haoyu, et al.
Veröffentlicht: (2024)
Filling the Image Information Gap for VQA: Prompting Large Language Models to Proactively Ask Questions
von: Wang, Ziyue, et al.
Veröffentlicht: (2023)
von: Wang, Ziyue, et al.
Veröffentlicht: (2023)
IllusionVQA: A Challenging Optical Illusion Dataset for Vision Language Models
von: Shahgir, Haz Sameen, et al.
Veröffentlicht: (2024)
von: Shahgir, Haz Sameen, et al.
Veröffentlicht: (2024)
Profiling Patient Transcript Using Large Language Model Reasoning Augmentation for Alzheimer's Disease Detection
von: Chen, Chin-Po, et al.
Veröffentlicht: (2024)
von: Chen, Chin-Po, et al.
Veröffentlicht: (2024)
$C^3$: Confidence Calibration Model Cascade for Inference-Efficient Cross-Lingual Natural Language Understanding
von: Lu, Taixi, et al.
Veröffentlicht: (2024)
von: Lu, Taixi, et al.
Veröffentlicht: (2024)
Unlocking the Power of LLM Uncertainty for Active In-Context Example Selection
von: Huang, Hsiu-Yuan, et al.
Veröffentlicht: (2024)
von: Huang, Hsiu-Yuan, et al.
Veröffentlicht: (2024)
Beyond Spurious Signals: Debiasing Multimodal Large Language Models via Counterfactual Inference and Adaptive Expert Routing
von: Wu, Zichen, et al.
Veröffentlicht: (2025)
von: Wu, Zichen, et al.
Veröffentlicht: (2025)
ELICIT: LLM Augmentation via External In-Context Capability
von: Wang, Futing, et al.
Veröffentlicht: (2024)
von: Wang, Futing, et al.
Veröffentlicht: (2024)
LLM-Generated Natural Language Meets Scaling Laws: New Explorations and Data Augmentation Methods
von: Wang, Zhenhua, et al.
Veröffentlicht: (2024)
von: Wang, Zhenhua, et al.
Veröffentlicht: (2024)
NLKI: A lightweight Natural Language Knowledge Integration Framework for Improving Small VLMs in Commonsense VQA Tasks
von: Dutta, Aritra, et al.
Veröffentlicht: (2025)
von: Dutta, Aritra, et al.
Veröffentlicht: (2025)
Enhancing Chain of Thought Prompting in Large Language Models via Reasoning Patterns
von: Zhang, Yufeng, et al.
Veröffentlicht: (2024)
von: Zhang, Yufeng, et al.
Veröffentlicht: (2024)
Augment before You Try: Knowledge-Enhanced Table Question Answering via Table Expansion
von: Liu, Yujian, et al.
Veröffentlicht: (2024)
von: Liu, Yujian, et al.
Veröffentlicht: (2024)
Clarify or Answer: Reinforcement Learning for Agentic VQA with Context Under-specification
von: Cao, Zongwan, et al.
Veröffentlicht: (2026)
von: Cao, Zongwan, et al.
Veröffentlicht: (2026)
Illusory VQA: Benchmarking and Enhancing Multimodal Models on Visual Illusions
von: Rostamkhani, Mohammadmostafa, et al.
Veröffentlicht: (2024)
von: Rostamkhani, Mohammadmostafa, et al.
Veröffentlicht: (2024)
SK-VQA: Synthetic Knowledge Generation at Scale for Training Context-Augmented Multimodal LLMs
von: Su, Xin, et al.
Veröffentlicht: (2024)
von: Su, Xin, et al.
Veröffentlicht: (2024)
Aligning Language Models with Real-time Knowledge Editing
von: Tang, Chenming, et al.
Veröffentlicht: (2025)
von: Tang, Chenming, et al.
Veröffentlicht: (2025)
Enhancing Financial VQA in Vision Language Models using Intermediate Structured Representations
von: Srivastava, Archita, et al.
Veröffentlicht: (2025)
von: Srivastava, Archita, et al.
Veröffentlicht: (2025)
$\textit{LinkPrompt}$: Natural and Universal Adversarial Attacks on Prompt-based Language Models
von: Xu, Yue, et al.
Veröffentlicht: (2024)
von: Xu, Yue, et al.
Veröffentlicht: (2024)
PAGE: Prompt Augmentation for text Generation Enhancement
von: Pacchiotti, Mauro Jose, et al.
Veröffentlicht: (2025)
von: Pacchiotti, Mauro Jose, et al.
Veröffentlicht: (2025)
Improving Implicit Discourse Relation Recognition with Natural Language Explanations from LLMs
von: Wang, Heng, et al.
Veröffentlicht: (2026)
von: Wang, Heng, et al.
Veröffentlicht: (2026)
A Survey on Data Augmentation in Large Model Era
von: Zhou, Yue, et al.
Veröffentlicht: (2024)
von: Zhou, Yue, et al.
Veröffentlicht: (2024)
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence
von: Ma, Dongsheng, et al.
Veröffentlicht: (2026)
von: Ma, Dongsheng, et al.
Veröffentlicht: (2026)
Natural Language Fine-Tuning
von: Liu, Jia, et al.
Veröffentlicht: (2024)
von: Liu, Jia, et al.
Veröffentlicht: (2024)
RAG in the Wild: On the (In)effectiveness of LLMs with Mixture-of-Knowledge Retrieval Augmentation
von: Xu, Ran, et al.
Veröffentlicht: (2025)
von: Xu, Ran, et al.
Veröffentlicht: (2025)
MedFrameQA: A Multi-Image Medical VQA Benchmark for Clinical Reasoning
von: Yu, Suhao, et al.
Veröffentlicht: (2025)
von: Yu, Suhao, et al.
Veröffentlicht: (2025)
mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA
von: Zhang, Tao, et al.
Veröffentlicht: (2024)
von: Zhang, Tao, et al.
Veröffentlicht: (2024)
Selective Augmentation: Improving Universal Automatic Phonetic Transcription via G2P Bootstrapping
von: Bystrich, Tobias, et al.
Veröffentlicht: (2026)
von: Bystrich, Tobias, et al.
Veröffentlicht: (2026)
Building A Coding Assistant via the Retrieval-Augmented Language Model
von: Li, Xinze, et al.
Veröffentlicht: (2024)
von: Li, Xinze, et al.
Veröffentlicht: (2024)
Enhancing Cross-lingual Sentence Embedding for Low-resource Languages with Word Alignment
von: Miao, Zhongtao, et al.
Veröffentlicht: (2024)
von: Miao, Zhongtao, et al.
Veröffentlicht: (2024)
Labeling Free-text Data using Language Model Ensembles
von: Qiu, Jiaxing, et al.
Veröffentlicht: (2025)
von: Qiu, Jiaxing, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Embedding Domain Knowledge for Large Language Models via Reinforcement Learning from Augmented Generation
von: Nie, Chaojun, et al.
Veröffentlicht: (2025) -
ActTraitBench: Quantifying the Knowledge-Decision Gap in Large Language Models via Human-Grounded Behavioral Validation
von: Yang, Yutong, et al.
Veröffentlicht: (2026) -
Layer Importance and Hallucination Analysis in Large Language Models via Enhanced Activation Variance-Sparsity
von: Song, Zichen, et al.
Veröffentlicht: (2024) -
Multimodal Retrieval-Augmented Generation with Large Language Models for Medical VQA
von: Karim, A H M Rezaul, et al.
Veröffentlicht: (2025) -
Benchmarking Multimodal Retrieval Augmented Generation with Dynamic VQA Dataset and Self-adaptive Planning Agent
von: Li, Yangning, et al.
Veröffentlicht: (2024)