T2R-bench: A Benchmark for Generating Article-Level Reports from Real World Industrial Tables
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Jie, Pan, Changzai, Wei, Kaiwen, Xiong, Sishi, Zhao, Yu, Li, Xiangyu, Peng, Jiaxin, Gu, Xiaoyan, Yang, Jian, Chang, Wenhan, Wu, Zhenhe, Zhong, Jiang, Song, Shuangyong, Li, Yongxiang, Li, Xuelong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TableZoomer: A Collaborative Agent Framework for Large-scale Table Question Answering
by: Xiong, Sishi, et al.
Published: (2025)
by: Xiong, Sishi, et al.
Published: (2025)
ReasonTabQA: A Comprehensive Benchmark for Table Question Answering from Real World Industrial Scenarios
by: Pan, Changzai, et al.
Published: (2026)
by: Pan, Changzai, et al.
Published: (2026)
TableReasoner: Advancing Table Reasoning Framework with Large Language Models
by: Xiong, Sishi, et al.
Published: (2025)
by: Xiong, Sishi, et al.
Published: (2025)
Table-R1: Region-based Reinforcement Learning for Table Understanding
by: Wu, Zhenhe, et al.
Published: (2025)
by: Wu, Zhenhe, et al.
Published: (2025)
MR-UIE: Multi-Perspective Reasoning with Reinforcement Learning for Universal Information Extraction
by: Li, Zhongqiu, et al.
Published: (2025)
by: Li, Zhongqiu, et al.
Published: (2025)
Prompt-Level Reward Specifications for Open-Ended Post-Training
by: Weng, Zijun, et al.
Published: (2026)
by: Weng, Zijun, et al.
Published: (2026)
RB-SQL: A Retrieval-based LLM Framework for Text-to-SQL
by: Wu, Zhenhe, et al.
Published: (2024)
by: Wu, Zhenhe, et al.
Published: (2024)
D-QRELO: Training- and Data-Free Delta Compression for Large Language Models via Quantization and Residual Low-Rank Approximation
by: Li, Junlin, et al.
Published: (2026)
by: Li, Junlin, et al.
Published: (2026)
Pairwise Preference Reward and Group-Based Diversity Enhancement for Superior Open-Ended Generation
by: Cao, Guining, et al.
Published: (2026)
by: Cao, Guining, et al.
Published: (2026)
Emotional Support with LLM-based Empathetic Dialogue Generation
by: Wang, Shiquan, et al.
Published: (2025)
by: Wang, Shiquan, et al.
Published: (2025)
RealSec-bench: A Benchmark for Evaluating Secure Code Generation in Real-World Repositories
by: Wang, Yanlin, et al.
Published: (2026)
by: Wang, Yanlin, et al.
Published: (2026)
HPSU: A Benchmark for Human-Level Perception in Real-World Spoken Speech Understanding
by: Li, Chen, et al.
Published: (2025)
by: Li, Chen, et al.
Published: (2025)
$τ$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
by: Yao, Shunyu, et al.
Published: (2024)
by: Yao, Shunyu, et al.
Published: (2024)
SEC-bench: Automated Benchmarking of LLM Agents on Real-World Software Security Tasks
by: Lee, Hwiwon, et al.
Published: (2025)
by: Lee, Hwiwon, et al.
Published: (2025)
Comment on “Bariatric Surgery Prior to Hip and Knee Arthroplasty: A Systematic Review and Meta‐Analysis of Postoperative Outcomes”
by: Zheng Han, et al.
Published: (2026)
by: Zheng Han, et al.
Published: (2026)
Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models
by: Nie, Shuo, et al.
Published: (2026)
by: Nie, Shuo, et al.
Published: (2026)
A modular model for reliability analysis model of PMS with multiple K/N subsystems and mixed shocks
by: Weijie Wang, et al.
Published: (2024)
by: Weijie Wang, et al.
Published: (2024)
Towards Robustness and Diversity: Continual Learning in Dialog Generation with Text-Mixup and Batch Nuclear-Norm Maximization
by: Wang, Zihan, et al.
Published: (2024)
by: Wang, Zihan, et al.
Published: (2024)
Chain-of-Lure: A Universal Jailbreak Attack Framework using Unconstrained Synthetic Narratives
by: Chang, Wenhan, et al.
Published: (2025)
by: Chang, Wenhan, et al.
Published: (2025)
Asymmetric Synthesis of Seven‐Membered Lactams: Recent Advances and Future Perspectives
by: Danyang Xie, et al.
Published: (2024)
by: Danyang Xie, et al.
Published: (2024)
Lemur: Log Parsing with Entropy Sampling and Chain-of-Thought Merging
by: Zhang, Wei, et al.
Published: (2024)
by: Zhang, Wei, et al.
Published: (2024)
Aligning MLLM Benchmark With Human Preferences via Structural Equation Modeling
by: Xiong, Shengwu., et al.
Published: (2025)
by: Xiong, Shengwu., et al.
Published: (2025)
Introducing Visual Scenes and Reasoning: A More Realistic Benchmark for Spoken Language Understanding
by: Wu, Di, et al.
Published: (2025)
by: Wu, Di, et al.
Published: (2025)
Multi-SWE-bench: A Multilingual Benchmark for Issue Resolving
by: Zan, Daoguang, et al.
Published: (2025)
by: Zan, Daoguang, et al.
Published: (2025)
Clustering-Oriented Generative Attribute Graph Imputation
by: Chen, Mulin, et al.
Published: (2025)
by: Chen, Mulin, et al.
Published: (2025)
From Laboratory to Real-World Applications: Benchmarking Agentic Code Reasoning at the Repository Level
by: Li, Jia, et al.
Published: (2026)
by: Li, Jia, et al.
Published: (2026)
Real-IAD: A Real-World Multi-View Dataset for Benchmarking Versatile Industrial Anomaly Detection
by: Wang, Chengjie, et al.
Published: (2024)
by: Wang, Chengjie, et al.
Published: (2024)
Patient‐derived xenograft models in pan‐cancer: From bench to clinic
by: Jiatong Li, et al.
Published: (2025)
by: Jiatong Li, et al.
Published: (2025)
TELEVAL: A Dynamic Benchmark Designed for Spoken Language Models in Chinese Interactive Scenarios
by: Li, Zehan, et al.
Published: (2025)
by: Li, Zehan, et al.
Published: (2025)
GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness
by: Chen, Hongjie, et al.
Published: (2025)
by: Chen, Hongjie, et al.
Published: (2025)
A Unified Spoken Language Model with Injected Emotional-Attribution Thinking for Human-like Interaction
by: Wang, Qing, et al.
Published: (2026)
by: Wang, Qing, et al.
Published: (2026)
SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
by: Jimenez, Carlos E., et al.
Published: (2023)
by: Jimenez, Carlos E., et al.
Published: (2023)
RealBench: A Repo-Level Code Generation Benchmark Aligned with Real-World Software Development Practices
by: Li, Jia, et al.
Published: (2026)
by: Li, Jia, et al.
Published: (2026)
TimeMachine-bench: A Benchmark for Evaluating Model Capabilities in Repository-Level Migration Tasks
by: Fujii, Ryo, et al.
Published: (2026)
by: Fujii, Ryo, et al.
Published: (2026)
CL-bench: A Benchmark for Context Learning
by: Dou, Shihan, et al.
Published: (2026)
by: Dou, Shihan, et al.
Published: (2026)
CFVBench: A Comprehensive Video Benchmark for Fine-grained Multimodal Retrieval-Augmented Generation
by: Wei, Kaiwen, et al.
Published: (2025)
by: Wei, Kaiwen, et al.
Published: (2025)
The maximum storage capacity of open-loop written RRAM is around 4 bits
by: Li, Yongxiang, et al.
Published: (2024)
by: Li, Yongxiang, et al.
Published: (2024)
RSA-Bench: Benchmarking Audio Large Models in Real-World Acoustic Scenarios
by: Zhang, Yibo, et al.
Published: (2026)
by: Zhang, Yibo, et al.
Published: (2026)
A Linkage Between 25‐Hydroxyvitamin D and Post‐Stroke Cognitive Impairment, as Well as the Duration of Hospitalization After a Stroke
by: Jianrong Xiong, et al.
Published: (2025)
by: Jianrong Xiong, et al.
Published: (2025)
Respiratory Failure Associated With Mutations in the RYR1 Gene: A Case Report
by: Chenliang Zhao, et al.
Published: (2025)
by: Chenliang Zhao, et al.
Published: (2025)
Similar Items
-
TableZoomer: A Collaborative Agent Framework for Large-scale Table Question Answering
by: Xiong, Sishi, et al.
Published: (2025) -
ReasonTabQA: A Comprehensive Benchmark for Table Question Answering from Real World Industrial Scenarios
by: Pan, Changzai, et al.
Published: (2026) -
TableReasoner: Advancing Table Reasoning Framework with Large Language Models
by: Xiong, Sishi, et al.
Published: (2025) -
Table-R1: Region-based Reinforcement Learning for Table Understanding
by: Wu, Zhenhe, et al.
Published: (2025) -
MR-UIE: Multi-Perspective Reasoning with Reinforcement Learning for Universal Information Extraction
by: Li, Zhongqiu, et al.
Published: (2025)