SportsMetrics: Blending Text and Numerical Data to Understand Information Fusion in LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Hu, Yebowen, Song, Kaiqiang, Cho, Sangwoo, Wang, Xiaoyang, Foroosh, Hassan, Yu, Dong, Liu, Fei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
When Reasoning Meets Information Aggregation: A Case Study with Sports Narratives
por: Hu, Yebowen, et al.
Publicado: (2024)
por: Hu, Yebowen, et al.
Publicado: (2024)
Can Large Language Models do Analytical Reasoning?
por: Hu, Yebowen, et al.
Publicado: (2024)
por: Hu, Yebowen, et al.
Publicado: (2024)
STRUX: An LLM for Decision-Making with Structured Explanations
por: Lu, Yiming, et al.
Publicado: (2024)
por: Lu, Yiming, et al.
Publicado: (2024)
DeFine: Decision-Making with Analogical Reasoning over Factor Profiles
por: Hu, Yebowen, et al.
Publicado: (2024)
por: Hu, Yebowen, et al.
Publicado: (2024)
InFoBench: Evaluating Instruction Following Ability in Large Language Models
por: Qin, Yiwei, et al.
Publicado: (2024)
por: Qin, Yiwei, et al.
Publicado: (2024)
SPECTRUM: Speaker-Enhanced Pre-Training for Long Dialogue Summarization
por: Cho, Sangwoo, et al.
Publicado: (2024)
por: Cho, Sangwoo, et al.
Publicado: (2024)
Polarity Calibration for Opinion Summarization
por: Lei, Yuanyuan, et al.
Publicado: (2024)
por: Lei, Yuanyuan, et al.
Publicado: (2024)
MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning
por: Liu, Fuxiao, et al.
Publicado: (2023)
por: Liu, Fuxiao, et al.
Publicado: (2023)
Skills-in-Context Prompting: Unlocking Compositionality in Large Language Models
por: Chen, Jiaao, et al.
Publicado: (2023)
por: Chen, Jiaao, et al.
Publicado: (2023)
Complex Logical Instruction Generation
por: Zhang, Mian, et al.
Publicado: (2025)
por: Zhang, Mian, et al.
Publicado: (2025)
Sports and Women's Sports: Gender Bias in Text Generation with Olympic Data
por: Biester, Laura
Publicado: (2025)
por: Biester, Laura
Publicado: (2025)
ToBlend: Token-Level Blending With an Ensemble of LLMs to Attack AI-Generated Text Detection
por: Huang, Fan, et al.
Publicado: (2024)
por: Huang, Fan, et al.
Publicado: (2024)
LiveMCP-101: Stress Testing and Diagnosing MCP-enabled Agents on Challenging Queries
por: Yin, Ming, et al.
Publicado: (2025)
por: Yin, Ming, et al.
Publicado: (2025)
BadRAG: Identifying Vulnerabilities in Retrieval Augmented Generation of Large Language Models
por: Xue, Jiaqi, et al.
Publicado: (2024)
por: Xue, Jiaqi, et al.
Publicado: (2024)
Sports Intelligence: Assessing the Sports Understanding Capabilities of Language Models through Question Answering from Text to Video
por: Yang, Zhengbang, et al.
Publicado: (2024)
por: Yang, Zhengbang, et al.
Publicado: (2024)
From Language Modeling to Instruction Following: Understanding the Behavior Shift in LLMs after Instruction Tuning
por: Wu, Xuansheng, et al.
Publicado: (2023)
por: Wu, Xuansheng, et al.
Publicado: (2023)
Understanding the Therapeutic Relationship between Counselors and Clients in Online Text-based Counseling using LLMs
por: Li, Anqi, et al.
Publicado: (2024)
por: Li, Anqi, et al.
Publicado: (2024)
Symbolic or Numerical? Understanding Physics Problem Solving in Reasoning LLMs
por: Dan, Nifu, et al.
Publicado: (2025)
por: Dan, Nifu, et al.
Publicado: (2025)
SQLForge: Synthesizing Reliable and Diverse Data to Enhance Text-to-SQL Reasoning in LLMs
por: Guo, Yu, et al.
Publicado: (2025)
por: Guo, Yu, et al.
Publicado: (2025)
OpenCharacter: Training Customizable Role-Playing LLMs with Large-Scale Synthetic Personas
por: Wang, Xiaoyang, et al.
Publicado: (2025)
por: Wang, Xiaoyang, et al.
Publicado: (2025)
SportQA: A Benchmark for Sports Understanding in Large Language Models
por: Xia, Haotian, et al.
Publicado: (2024)
por: Xia, Haotian, et al.
Publicado: (2024)
When LLMs Meet Cunning Texts: A Fallacy Understanding Benchmark for Large Language Models
por: Li, Yinghui, et al.
Publicado: (2024)
por: Li, Yinghui, et al.
Publicado: (2024)
Do Language Models Understand the Cognitive Tasks Given to Them? Investigations with the N-Back Paradigm
por: Hu, Xiaoyang, et al.
Publicado: (2024)
por: Hu, Xiaoyang, et al.
Publicado: (2024)
Improved Evidence Extraction and Metrics for Document Inconsistency Detection with LLMs
por: Tan, Nelvin, et al.
Publicado: (2026)
por: Tan, Nelvin, et al.
Publicado: (2026)
Understanding and Mitigating Numerical Sources of Nondeterminism in LLM Inference
por: Yuan, Jiayi, et al.
Publicado: (2025)
por: Yuan, Jiayi, et al.
Publicado: (2025)
Multi-Granular Multimodal Clue Fusion for Meme Understanding
por: Zheng, Li, et al.
Publicado: (2025)
por: Zheng, Li, et al.
Publicado: (2025)
Lowest Span Confidence: A Zero-Shot Metric for Efficient and Black-Box Hallucination Detection in LLMs
por: Qiao, Yitong, et al.
Publicado: (2026)
por: Qiao, Yitong, et al.
Publicado: (2026)
Evaluation Metrics for Text Data Augmentation in NLP
por: Amadeus, Marcellus, et al.
Publicado: (2024)
por: Amadeus, Marcellus, et al.
Publicado: (2024)
Can LLMs Understand Unvoiced Speech? Exploring EMG-to-Text Conversion with LLMs
por: Mohapatra, Payal, et al.
Publicado: (2025)
por: Mohapatra, Payal, et al.
Publicado: (2025)
ReHear: Iterative Pseudo-Label Refinement for Semi-Supervised Speech Recognition via Audio Large Language Models
por: Liu, Zefang, et al.
Publicado: (2026)
por: Liu, Zefang, et al.
Publicado: (2026)
Enhancing Character-Level Understanding in LLMs through Token Internal Structure Learning
por: Xu, Zhu, et al.
Publicado: (2024)
por: Xu, Zhu, et al.
Publicado: (2024)
Towards Automatic Evaluation for LLMs' Clinical Capabilities: Metric, Data, and Algorithm
por: Liu, Lei, et al.
Publicado: (2024)
por: Liu, Lei, et al.
Publicado: (2024)
Multimodal Magic Elevating Depression Detection with a Fusion of Text and Audio Intelligence
por: Gan, Lindy, et al.
Publicado: (2025)
por: Gan, Lindy, et al.
Publicado: (2025)
A Versatile Multimodal Agent for Multimedia Content Generation
por: Zhang, Daoan, et al.
Publicado: (2026)
por: Zhang, Daoan, et al.
Publicado: (2026)
Moneyball with LLMs: Analyzing Tabular Summarization in Sports Narratives
por: Upadhyay, Ritam, et al.
Publicado: (2025)
por: Upadhyay, Ritam, et al.
Publicado: (2025)
Understanding the Influence of Synthetic Data for Text Embedders
por: Springer, Jacob Mitchell, et al.
Publicado: (2025)
por: Springer, Jacob Mitchell, et al.
Publicado: (2025)
CapsFusion: Rethinking Image-Text Data at Scale
por: Yu, Qiying, et al.
Publicado: (2023)
por: Yu, Qiying, et al.
Publicado: (2023)
Leveraging Parameter Space Symmetries for Reasoning Skill Transfer in LLMs
por: Horoi, Stefan, et al.
Publicado: (2025)
por: Horoi, Stefan, et al.
Publicado: (2025)
SCOUT: Active Information Foraging for Long-Text Understanding with Decoupled Epistemic States
por: Zhang, Zhenliang, et al.
Publicado: (2026)
por: Zhang, Zhenliang, et al.
Publicado: (2026)
ProtT3: Protein-to-Text Generation for Text-based Protein Understanding
por: Liu, Zhiyuan, et al.
Publicado: (2024)
por: Liu, Zhiyuan, et al.
Publicado: (2024)
Ejemplares similares
-
When Reasoning Meets Information Aggregation: A Case Study with Sports Narratives
por: Hu, Yebowen, et al.
Publicado: (2024) -
Can Large Language Models do Analytical Reasoning?
por: Hu, Yebowen, et al.
Publicado: (2024) -
STRUX: An LLM for Decision-Making with Structured Explanations
por: Lu, Yiming, et al.
Publicado: (2024) -
DeFine: Decision-Making with Analogical Reasoning over Factor Profiles
por: Hu, Yebowen, et al.
Publicado: (2024) -
InFoBench: Evaluating Instruction Following Ability in Large Language Models
por: Qin, Yiwei, et al.
Publicado: (2024)