LitBench: A Graph-Centric Large Language Model Benchmarking Tool For Literature Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Varvarigos, Andreas, Maatouk, Ali, Zhang, Jiasheng, Bui, Ngoc, Chen, Jialin, Tassiulas, Leandros, Ying, Rex |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LitFM: A Retrieval Augmented Structure-aware Foundation Model For Citation Graphs
by: Zhang, Jiasheng, et al.
Published: (2024)
by: Zhang, Jiasheng, et al.
Published: (2024)
Tele-LLMs: A Series of Specialized Large Language Models for Telecommunications
by: Maatouk, Ali, et al.
Published: (2024)
by: Maatouk, Ali, et al.
Published: (2024)
LitSearch: A Retrieval Benchmark for Scientific Literature Search
by: Ajith, Anirudh, et al.
Published: (2024)
by: Ajith, Anirudh, et al.
Published: (2024)
Multi-Modal Time Series Prediction via Mixture of Modulated Experts
by: Zhang, Lige, et al.
Published: (2026)
by: Zhang, Lige, et al.
Published: (2026)
LitLLMs, LLMs for Literature Review: Are we there yet?
by: Agarwal, Shubham, et al.
Published: (2024)
by: Agarwal, Shubham, et al.
Published: (2024)
LLAssist: Simple Tools for Automating Literature Review Using Large Language Models
by: Haryanto, Christoforus Yoga
Published: (2024)
by: Haryanto, Christoforus Yoga
Published: (2024)
Oignon: Citation Graph Tool
by: Ballington, Harry
Published: (2025)
by: Ballington, Harry
Published: (2025)
Optimizing Data Extraction from Materials Science Literature: A Study of Tools Using Large Language Models
by: Ning, Wenkai, et al.
Published: (2025)
by: Ning, Wenkai, et al.
Published: (2025)
Aletheia-Probe: A Tool for Automated Journal Assessment
by: Florath, Andreas
Published: (2026)
by: Florath, Andreas
Published: (2026)
PST-Bench: Tracing and Benchmarking the Source of Publications
by: Zhang, Fanjin, et al.
Published: (2024)
by: Zhang, Fanjin, et al.
Published: (2024)
OAG-Bench: A Human-Curated Benchmark for Academic Graph Mining
by: Zhang, Fanjin, et al.
Published: (2024)
by: Zhang, Fanjin, et al.
Published: (2024)
LitBench: A Benchmark and Dataset for Reliable Evaluation of Creative Writing
by: Fein, Daniel, et al.
Published: (2025)
by: Fein, Daniel, et al.
Published: (2025)
Loci Similes: A Benchmark for Extracting Intertextualities in Latin Literature
by: Schelb, Julian, et al.
Published: (2026)
by: Schelb, Julian, et al.
Published: (2026)
Cool URIs for FAIR Knowledge Graphs
by: Thalhammer, Andreas
Published: (2024)
by: Thalhammer, Andreas
Published: (2024)
Interactive Graph Visualization and TeamingRecommendation in an Interdisciplinary Project'sTalent Knowledge Graph
by: Xu, Jiawei, et al.
Published: (2025)
by: Xu, Jiawei, et al.
Published: (2025)
More Parameters Than Populations: A Systematic Literature Review of Large Language Models within Survey Research
by: Buskirk, Trent D., et al.
Published: (2025)
by: Buskirk, Trent D., et al.
Published: (2025)
The QIC-Index: A Novel, Data-Centric Metric for Quantifying the Impact of Research Data Sharing
by: Frasch, Martin G.
Published: (2025)
by: Frasch, Martin G.
Published: (2025)
Towards the relationship between AIGC in manuscript writing and author profiles: evidence from preprints in LLMs
by: Liu, Jialin, et al.
Published: (2024)
by: Liu, Jialin, et al.
Published: (2024)
GLiSE: A Prompt-Driven and ML-Powered Tool for Automated Grey Literature Extraction in Software Engineering
by: Cherief, Houcine Abdelkader, et al.
Published: (2025)
by: Cherief, Houcine Abdelkader, et al.
Published: (2025)
Large Language Models for Departmental Expert Review Quality Scores
by: Langfeldt, Liv, et al.
Published: (2026)
by: Langfeldt, Liv, et al.
Published: (2026)
Exploring the applicability of Large Language Models to citation context analysis
by: Nishikawa, Kai, et al.
Published: (2024)
by: Nishikawa, Kai, et al.
Published: (2024)
Leveraging Large Language Models for Realizing Truly Intelligent User Interfaces
by: Oelen, Allard, et al.
Published: (2025)
by: Oelen, Allard, et al.
Published: (2025)
Streamlining the Selection Phase of Systematic Literature Reviews (SLRs) Using AI-Enabled GPT-4 Assistant API
by: Jafari, Seyed Mohammad Ali
Published: (2024)
by: Jafari, Seyed Mohammad Ali
Published: (2024)
Application of Module to Coding Theory: A Systematic Literature Review
by: Faldiyan, Muhammad, et al.
Published: (2024)
by: Faldiyan, Muhammad, et al.
Published: (2024)
Prompt perturbation and fraction facilitation sometimes strengthen Large Language Model scores
by: Thelwall, Mike
Published: (2025)
by: Thelwall, Mike
Published: (2025)
pyKCN: A Python Tool for Bridging Scientific Knowledge
by: Lu, Zhenyuan, et al.
Published: (2024)
by: Lu, Zhenyuan, et al.
Published: (2024)
Expertise Indices: Variants, Modifications, Advancements, and Computational Tools in R
by: Nandy, Abhirup, et al.
Published: (2026)
by: Nandy, Abhirup, et al.
Published: (2026)
Do Large Language Models know Which Published Articles have been Retracted?
by: Thelwall, Mike
Published: (2026)
by: Thelwall, Mike
Published: (2026)
Decoding Patterns of Data Generation Teams for Clinical and Scientific Success: Insights from the Bridge2AI Talent Knowledge Graph
by: Xu, Jiawei, et al.
Published: (2025)
by: Xu, Jiawei, et al.
Published: (2025)
LLM-Based Information Extraction to Support Scientific Literature Research and Publication Workflows
by: Ateia, Samy, et al.
Published: (2025)
by: Ateia, Samy, et al.
Published: (2025)
Network Analysis, Plot Theory: Revisiting French Literature through Character Networks
by: Chen, Newman, et al.
Published: (2024)
by: Chen, Newman, et al.
Published: (2024)
Towards Development of Automated Knowledge Maps and Databases for Materials Engineering using Large Language Models
by: Prasad, Deepak, et al.
Published: (2024)
by: Prasad, Deepak, et al.
Published: (2024)
Research quality evaluation by AI in the era of Large Language Models: Advantages, disadvantages, and systemic effects
by: Thelwall, Mike
Published: (2025)
by: Thelwall, Mike
Published: (2025)
SWARM-SLR AIssistant: A Unified Framework for Scalable Systematic Literature Review Automation
by: Wittenborg, Tim, et al.
Published: (2026)
by: Wittenborg, Tim, et al.
Published: (2026)
Trends in Equal-Contribution Authorship: A Large-Scale Bibliometric Analysis of Biomedical Literature
by: Xu, Binbin
Published: (2026)
by: Xu, Binbin
Published: (2026)
From Coverage to Prestige: A Comprehensive Assessment of Large-Scale Scientometric Data
by: Rong, Guoyang, et al.
Published: (2025)
by: Rong, Guoyang, et al.
Published: (2025)
Can Large Language Models Evaluate Grant Proposal Quality? Revisiting the Wennerås and Wold Peer Review Data
by: Sandström, Ulf, et al.
Published: (2026)
by: Sandström, Ulf, et al.
Published: (2026)
Large Language Models for Web Accessibility: A Systematic Literature Review
by: Aljedaani, Wajdi, et al.
Published: (2026)
by: Aljedaani, Wajdi, et al.
Published: (2026)
How good is the h-index?
by: Borji, Ali
Published: (2025)
by: Borji, Ali
Published: (2025)
Rising Prevalence of Detected AI-Generated Text in Medical Literature: Longitudinal Analysis in Open Access Articles
by: Wolfrath, Nathan, et al.
Published: (2026)
by: Wolfrath, Nathan, et al.
Published: (2026)
Similar Items
-
LitFM: A Retrieval Augmented Structure-aware Foundation Model For Citation Graphs
by: Zhang, Jiasheng, et al.
Published: (2024) -
Tele-LLMs: A Series of Specialized Large Language Models for Telecommunications
by: Maatouk, Ali, et al.
Published: (2024) -
LitSearch: A Retrieval Benchmark for Scientific Literature Search
by: Ajith, Anirudh, et al.
Published: (2024) -
Multi-Modal Time Series Prediction via Mixture of Modulated Experts
by: Zhang, Lige, et al.
Published: (2026) -
LitLLMs, LLMs for Literature Review: Are we there yet?
by: Agarwal, Shubham, et al.
Published: (2024)