Multimodal Table Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Mingyu, Feng, Xinwei, Si, Qingyi, She, Qiaoqiao, Lin, Zheng, Jiang, Wenbin, Wang, Weiping |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Closer Look into LLMs for Table Understanding
von: Wang, Jia, et al.
Veröffentlicht: (2026)
von: Wang, Jia, et al.
Veröffentlicht: (2026)
Are Large Language Models Table-based Fact-Checkers?
von: Zhang, Hanwen, et al.
Veröffentlicht: (2024)
von: Zhang, Hanwen, et al.
Veröffentlicht: (2024)
Test-time Prompt Intervention
von: Yang, Chenxu, et al.
Veröffentlicht: (2025)
von: Yang, Chenxu, et al.
Veröffentlicht: (2025)
Breaking the Trade-Off Between Faithfulness and Expressiveness for Large Language Models
von: Yang, Chenxu, et al.
Veröffentlicht: (2025)
von: Yang, Chenxu, et al.
Veröffentlicht: (2025)
TableDreamer: Progressive and Weakness-guided Data Synthesis from Scratch for Table Instruction Tuning
von: Zheng, Mingyu, et al.
Veröffentlicht: (2025)
von: Zheng, Mingyu, et al.
Veröffentlicht: (2025)
A Multi-Task Role-Playing Agent Capable of Imitating Character Linguistic Styles
von: Chen, Siyuan, et al.
Veröffentlicht: (2024)
von: Chen, Siyuan, et al.
Veröffentlicht: (2024)
System 1&2 Synergy via Dynamic Model Interpolation
von: Yang, Chenxu, et al.
Veröffentlicht: (2026)
von: Yang, Chenxu, et al.
Veröffentlicht: (2026)
Dynamic Early Exit in Reasoning Models
von: Yang, Chenxu, et al.
Veröffentlicht: (2025)
von: Yang, Chenxu, et al.
Veröffentlicht: (2025)
Weights-Rotated Preference Optimization for Large Language Models
von: Yang, Chenxu, et al.
Veröffentlicht: (2025)
von: Yang, Chenxu, et al.
Veröffentlicht: (2025)
Understanding the Role of LLMs in Multimodal Evaluation Benchmarks
von: Jiang, Botian, et al.
Veröffentlicht: (2024)
von: Jiang, Botian, et al.
Veröffentlicht: (2024)
Pruning Large Language Models to Intra-module Low-rank Architecture with Transitional Activations
von: Shen, Bowen, et al.
Veröffentlicht: (2024)
von: Shen, Bowen, et al.
Veröffentlicht: (2024)
MS2SL: Multimodal Spoken Data-Driven Continuous Sign Language Production
von: Ma, Jian, et al.
Veröffentlicht: (2024)
von: Ma, Jian, et al.
Veröffentlicht: (2024)
Emotion and Intent Joint Understanding in Multimodal Conversation: A Benchmarking Dataset
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
Towards Robust Multimodal Sentiment Analysis with Incomplete Data
von: Zhang, Haoyu, et al.
Veröffentlicht: (2024)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2024)
Benchmarking Multimodal LLMs on Recognition and Understanding over Chemical Tables
von: Zhou, Yitong, et al.
Veröffentlicht: (2025)
von: Zhou, Yitong, et al.
Veröffentlicht: (2025)
MAC: A Live Benchmark for Multimodal Large Language Models in Scientific Understanding
von: Jiang, Mohan, et al.
Veröffentlicht: (2025)
von: Jiang, Mohan, et al.
Veröffentlicht: (2025)
LongDocURL: a Comprehensive Multimodal Long Document Benchmark Integrating Understanding, Reasoning, and Locating
von: Deng, Chao, et al.
Veröffentlicht: (2024)
von: Deng, Chao, et al.
Veröffentlicht: (2024)
When Language Overrules: Revealing Text Dominance in Multimodal Large Language Models
von: Wu, Huyu, et al.
Veröffentlicht: (2025)
von: Wu, Huyu, et al.
Veröffentlicht: (2025)
Multimodal Transformers are Hierarchical Modal-wise Heterogeneous Graphs
von: Jin, Yijie, et al.
Veröffentlicht: (2025)
von: Jin, Yijie, et al.
Veröffentlicht: (2025)
Towards Analyzing and Understanding the Limitations of DPO: A Theoretical Perspective
von: Feng, Duanyu, et al.
Veröffentlicht: (2024)
von: Feng, Duanyu, et al.
Veröffentlicht: (2024)
Efficient LLM-Jailbreaking via Multimodal-LLM Jailbreak
von: Ji, Haoxuan, et al.
Veröffentlicht: (2024)
von: Ji, Haoxuan, et al.
Veröffentlicht: (2024)
InFi-Check: Interpretable and Fine-Grained Fact-Checking of LLMs
von: Bai, Yuzhuo, et al.
Veröffentlicht: (2026)
von: Bai, Yuzhuo, et al.
Veröffentlicht: (2026)
Enhancing Semantics in Multimodal Chain of Thought via Soft Negative Sampling
von: Zheng, Guangmin, et al.
Veröffentlicht: (2024)
von: Zheng, Guangmin, et al.
Veröffentlicht: (2024)
LooGLE: Can Long-Context Language Models Understand Long Contexts?
von: Li, Jiaqi, et al.
Veröffentlicht: (2023)
von: Li, Jiaqi, et al.
Veröffentlicht: (2023)
Graph2Eval: Automatic Multimodal Task Generation for Agents via Knowledge Graphs
von: Chen, Yurun, et al.
Veröffentlicht: (2025)
von: Chen, Yurun, et al.
Veröffentlicht: (2025)
Leveraging Computerized Adaptive Testing for Cost-effective Evaluation of Large Language Models in Medical Benchmarking
von: Zheng, Tianpeng, et al.
Veröffentlicht: (2026)
von: Zheng, Tianpeng, et al.
Veröffentlicht: (2026)
Problem-Solving Logic Guided Curriculum In-Context Learning for LLMs Complex Reasoning
von: Ma, Xuetao, et al.
Veröffentlicht: (2025)
von: Ma, Xuetao, et al.
Veröffentlicht: (2025)
UniOQA: A Unified Framework for Knowledge Graph Question Answering with Large Language Models
von: Li, Zhuoyang, et al.
Veröffentlicht: (2024)
von: Li, Zhuoyang, et al.
Veröffentlicht: (2024)
Understanding Textual Capability Degradation in Speech LLMs via Parameter Importance Analysis
von: Wang, Chao, et al.
Veröffentlicht: (2025)
von: Wang, Chao, et al.
Veröffentlicht: (2025)
Sparsity Induction for Accurate Post-Training Pruning of Large Language Models
von: Jiang, Minhao, et al.
Veröffentlicht: (2026)
von: Jiang, Minhao, et al.
Veröffentlicht: (2026)
Aligning to Illusions: Choice Blindness in Human and AI Feedback
von: Wu, Wenbin
Veröffentlicht: (2026)
von: Wu, Wenbin
Veröffentlicht: (2026)
WebNovelBench: Placing LLM Novelists on the Web Novel Distribution
von: Lin, Leon, et al.
Veröffentlicht: (2025)
von: Lin, Leon, et al.
Veröffentlicht: (2025)
TableRAG: Million-Token Table Understanding with Language Models
von: Chen, Si-An, et al.
Veröffentlicht: (2024)
von: Chen, Si-An, et al.
Veröffentlicht: (2024)
Elastic MoE: Unlocking the Inference-Time Scalability of Mixture-of-Experts
von: Gu, Naibin, et al.
Veröffentlicht: (2025)
von: Gu, Naibin, et al.
Veröffentlicht: (2025)
AviationLMM: A Large Multimodal Foundation Model for Civil Aviation
von: Li, Wenbin, et al.
Veröffentlicht: (2026)
von: Li, Wenbin, et al.
Veröffentlicht: (2026)
Enhancing Contextual Understanding in Large Language Models through Contrastive Decoding
von: Zhao, Zheng, et al.
Veröffentlicht: (2024)
von: Zhao, Zheng, et al.
Veröffentlicht: (2024)
Quantization-Aware and Tensor-Compressed Training of Transformers for Natural Language Understanding
von: Yang, Zi, et al.
Veröffentlicht: (2023)
von: Yang, Zi, et al.
Veröffentlicht: (2023)
MMUnlearner: Reformulating Multimodal Machine Unlearning in the Era of Multimodal Large Language Models
von: Huo, Jiahao, et al.
Veröffentlicht: (2025)
von: Huo, Jiahao, et al.
Veröffentlicht: (2025)
UniCode: Learning a Unified Codebook for Multimodal Large Language Models
von: Zheng, Sipeng, et al.
Veröffentlicht: (2024)
von: Zheng, Sipeng, et al.
Veröffentlicht: (2024)
MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning
von: Liu, Fuxiao, et al.
Veröffentlicht: (2023)
von: Liu, Fuxiao, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
A Closer Look into LLMs for Table Understanding
von: Wang, Jia, et al.
Veröffentlicht: (2026) -
Are Large Language Models Table-based Fact-Checkers?
von: Zhang, Hanwen, et al.
Veröffentlicht: (2024) -
Test-time Prompt Intervention
von: Yang, Chenxu, et al.
Veröffentlicht: (2025) -
Breaking the Trade-Off Between Faithfulness and Expressiveness for Large Language Models
von: Yang, Chenxu, et al.
Veröffentlicht: (2025) -
TableDreamer: Progressive and Weakness-guided Data Synthesis from Scratch for Table Instruction Tuning
von: Zheng, Mingyu, et al.
Veröffentlicht: (2025)