HighlightBench: Benchmarking Markup-Driven Table Reasoning in Scientific Documents
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Lexin, Liu, Shenghua, Wang, Yiwei, Cai, Yujun, Ge, Yuyao, Yao, Jiayu, Guo, Jiafeng, Cheng, Xueqi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Focusing by Contrastive Attention: Enhancing VLMs' Visual Reasoning
di: Ge, Yuyao, et al.
Pubblicazione: (2025)
di: Ge, Yuyao, et al.
Pubblicazione: (2025)
Prism-$Δ$: Differential Subspace Steering for Prompt Highlighting in Large Language Models
di: Ge, Yuyao, et al.
Pubblicazione: (2026)
di: Ge, Yuyao, et al.
Pubblicazione: (2026)
A Survey of Vibe Coding with Large Language Models
di: Ge, Yuyao, et al.
Pubblicazione: (2025)
di: Ge, Yuyao, et al.
Pubblicazione: (2025)
Reward and Guidance through Rubrics: Promoting Exploration to Improve Multi-Domain Reasoning
di: Bi, Baolong, et al.
Pubblicazione: (2025)
di: Bi, Baolong, et al.
Pubblicazione: (2025)
Gated Differentiable Working Memory for Long-Context Language Modeling
di: Mei, Lingrui, et al.
Pubblicazione: (2026)
di: Mei, Lingrui, et al.
Pubblicazione: (2026)
Innate Reasoning is Not Enough: In-Context Learning Enhances Reasoning Large Language Models with Less Overthinking
di: Ge, Yuyao, et al.
Pubblicazione: (2025)
di: Ge, Yuyao, et al.
Pubblicazione: (2025)
Who is in the Spotlight: The Hidden Bias Undermining Multimodal Retrieval-Augmented Generation
di: Yao, Jiayu, et al.
Pubblicazione: (2025)
di: Yao, Jiayu, et al.
Pubblicazione: (2025)
AudioMotionBench: Evaluating Auditory Motion Perception in Audio LLMs
di: Sun, Zhe, et al.
Pubblicazione: (2025)
di: Sun, Zhe, et al.
Pubblicazione: (2025)
ChainMPQ: Interleaved Text-Image Reasoning Chains for Mitigating Relation Hallucinations
di: Wu, Yike, et al.
Pubblicazione: (2025)
di: Wu, Yike, et al.
Pubblicazione: (2025)
FrameMind: Frame-Interleaved Video Reasoning via Reinforcement Learning
di: Ge, Haonan, et al.
Pubblicazione: (2025)
di: Ge, Haonan, et al.
Pubblicazione: (2025)
CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning
di: Wu, Hang, et al.
Pubblicazione: (2026)
di: Wu, Hang, et al.
Pubblicazione: (2026)
MVAM: Multi-View Attention Method for Fine-grained Image-Text Matching
di: Cui, Wanqing, et al.
Pubblicazione: (2024)
di: Cui, Wanqing, et al.
Pubblicazione: (2024)
Cure or Poison? Embedding Instructions Visually Alters Hallucination in Vision-Language Models
di: Wang, Zhaochen, et al.
Pubblicazione: (2025)
di: Wang, Zhaochen, et al.
Pubblicazione: (2025)
MRFD: Multi-Region Fusion Decoding with Self-Consistency for Mitigating Hallucinations in LVLMs
di: Ge, Haonan, et al.
Pubblicazione: (2025)
di: Ge, Haonan, et al.
Pubblicazione: (2025)
SemVink: Advancing VLMs' Semantic Understanding of Optical Illusions via Visual Global Thinking
di: Li, Sifan, et al.
Pubblicazione: (2025)
di: Li, Sifan, et al.
Pubblicazione: (2025)
ViewFusion: Structured Spatial Thinking Chains for Multi-View Reasoning
di: Tao, Xingjian, et al.
Pubblicazione: (2026)
di: Tao, Xingjian, et al.
Pubblicazione: (2026)
Can Graph Descriptive Order Affect Solving Graph Problems with LLMs?
di: Ge, Yuyao, et al.
Pubblicazione: (2024)
di: Ge, Yuyao, et al.
Pubblicazione: (2024)
a1: Steep Test-time Scaling Law via Environment Augmented Generation
di: Mei, Lingrui, et al.
Pubblicazione: (2025)
di: Mei, Lingrui, et al.
Pubblicazione: (2025)
How does Watermarking Affect Visual Language Models in Document Understanding?
di: Xu, Chunxue, et al.
Pubblicazione: (2025)
di: Xu, Chunxue, et al.
Pubblicazione: (2025)
SciVideoBench: Benchmarking Scientific Video Reasoning in Large Multimodal Models
di: Deng, Andong, et al.
Pubblicazione: (2025)
di: Deng, Andong, et al.
Pubblicazione: (2025)
Multimodal Markup Document Models for Graphic Design Completion
di: Kikuchi, Kotaro, et al.
Pubblicazione: (2024)
di: Kikuchi, Kotaro, et al.
Pubblicazione: (2024)
Classifier Guidance Enhances Diffusion-based Adversarial Purification by Preserving Predictive Information
di: Zhang, Mingkun, et al.
Pubblicazione: (2024)
di: Zhang, Mingkun, et al.
Pubblicazione: (2024)
Adaptive Markup Language Generation for Contextually-Grounded Visual Document Understanding
di: Xiao, Han, et al.
Pubblicazione: (2025)
di: Xiao, Han, et al.
Pubblicazione: (2025)
GIR-Bench: Versatile Benchmark for Generating Images with Reasoning
di: Li, Hongxiang, et al.
Pubblicazione: (2025)
di: Li, Hongxiang, et al.
Pubblicazione: (2025)
GDI-Bench: A Benchmark for General Document Intelligence with Vision and Reasoning Decoupling
di: Li, Siqi, et al.
Pubblicazione: (2025)
di: Li, Siqi, et al.
Pubblicazione: (2025)
CLIPure: Purification in Latent Space via CLIP for Adversarially Robust Zero-Shot Classification
di: Zhang, Mingkun, et al.
Pubblicazione: (2025)
di: Zhang, Mingkun, et al.
Pubblicazione: (2025)
R2I-Bench: Benchmarking Reasoning-Driven Text-to-Image Generation
di: Chen, Kaijie, et al.
Pubblicazione: (2025)
di: Chen, Kaijie, et al.
Pubblicazione: (2025)
Not in Sync: Unveiling Temporal Bias in Audio Chat Models
di: Yao, Jiayu, et al.
Pubblicazione: (2025)
di: Yao, Jiayu, et al.
Pubblicazione: (2025)
A Survey of Context Engineering for Large Language Models
di: Mei, Lingrui, et al.
Pubblicazione: (2025)
di: Mei, Lingrui, et al.
Pubblicazione: (2025)
Benchmarking Scientific Understanding and Reasoning for Video Generation using VideoScience-Bench
di: Hu, Lanxiang, et al.
Pubblicazione: (2025)
di: Hu, Lanxiang, et al.
Pubblicazione: (2025)
PromptCD: Test-Time Behavior Enhancement via Polarity-Prompt Contrastive Decoding
di: Bi, Baolong, et al.
Pubblicazione: (2026)
di: Bi, Baolong, et al.
Pubblicazione: (2026)
$A^2R^2$: Advancing Img2LaTeX Conversion via Visual Reasoning with Attention-Guided Refinement
di: Li, Zhecheng, et al.
Pubblicazione: (2025)
di: Li, Zhecheng, et al.
Pubblicazione: (2025)
CausalDiff: Causality-Inspired Disentanglement via Diffusion Model for Adversarial Defense
di: Zhang, Mingkun, et al.
Pubblicazione: (2024)
di: Zhang, Mingkun, et al.
Pubblicazione: (2024)
Visual Transformation Telling
di: Cui, Wanqing, et al.
Pubblicazione: (2023)
di: Cui, Wanqing, et al.
Pubblicazione: (2023)
PresentBench: A Fine-Grained Rubric-Based Benchmark for Slide Generation
di: Chen, Xin-Sheng, et al.
Pubblicazione: (2026)
di: Chen, Xin-Sheng, et al.
Pubblicazione: (2026)
VLRS-Bench: A Vision-Language Reasoning Benchmark for Remote Sensing
di: Luo, Zhiming, et al.
Pubblicazione: (2026)
di: Luo, Zhiming, et al.
Pubblicazione: (2026)
TiViBench: Benchmarking Think-in-Video Reasoning for Video Generative Models
di: Chen, Harold Haodong, et al.
Pubblicazione: (2025)
di: Chen, Harold Haodong, et al.
Pubblicazione: (2025)
UniREditBench: A Unified Reasoning-based Image Editing Benchmark
di: Han, Feng, et al.
Pubblicazione: (2025)
di: Han, Feng, et al.
Pubblicazione: (2025)
SLANG: New Concept Comprehension of Large Language Models
di: Mei, Lingrui, et al.
Pubblicazione: (2024)
di: Mei, Lingrui, et al.
Pubblicazione: (2024)
LPNL: Scalable Link Prediction with Large Language Models
di: Bi, Baolong, et al.
Pubblicazione: (2024)
di: Bi, Baolong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Focusing by Contrastive Attention: Enhancing VLMs' Visual Reasoning
di: Ge, Yuyao, et al.
Pubblicazione: (2025) -
Prism-$Δ$: Differential Subspace Steering for Prompt Highlighting in Large Language Models
di: Ge, Yuyao, et al.
Pubblicazione: (2026) -
A Survey of Vibe Coding with Large Language Models
di: Ge, Yuyao, et al.
Pubblicazione: (2025) -
Reward and Guidance through Rubrics: Promoting Exploration to Improve Multi-Domain Reasoning
di: Bi, Baolong, et al.
Pubblicazione: (2025) -
Gated Differentiable Working Memory for Long-Context Language Modeling
di: Mei, Lingrui, et al.
Pubblicazione: (2026)