Evaluating the Predictive Capacity of ChatGPT for Academic Peer Review Outcomes Across Multiple Platforms
Fuente:
arXiv
Saved in:
| Main Authors: | Thelwall, Mike, Yaghi, Abdullah |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
In which fields can ChatGPT detect journal article quality? An evaluation of REF2021 results
by: Thelwall, Mike, et al.
Published: (2024)
by: Thelwall, Mike, et al.
Published: (2024)
Can ChatGPT evaluate research quality?
by: Thelwall, Mike
Published: (2024)
by: Thelwall, Mike
Published: (2024)
A Global South Strategy for Evaluating Research Value with ChatGPT
by: Nunkoo, Robin, et al.
Published: (2025)
by: Nunkoo, Robin, et al.
Published: (2025)
In which fields do ChatGPT scores align better than citations with research quality?
by: Thelwall, Mike
Published: (2025)
by: Thelwall, Mike
Published: (2025)
Which stylistic features fool ChatGPT research evaluations?
by: Kousha, Kayvan, et al.
Published: (2026)
by: Kousha, Kayvan, et al.
Published: (2026)
Estimating the quality of academic books from their descriptions with ChatGPT
by: Thelwall, Mike, et al.
Published: (2025)
by: Thelwall, Mike, et al.
Published: (2025)
Evaluating Research Quality with Large Language Models: An Analysis of ChatGPT's Effectiveness with Different Settings and Inputs
by: Thelwall, Mike
Published: (2024)
by: Thelwall, Mike
Published: (2024)
Research evaluation with ChatGPT: Is it age, country, length, or field biased?
by: Thelwall, Mike, et al.
Published: (2024)
by: Thelwall, Mike, et al.
Published: (2024)
Implicit and Explicit Research Quality Score Probabilities from ChatGPT
by: Thelwall, Mike, et al.
Published: (2025)
by: Thelwall, Mike, et al.
Published: (2025)
Journal Quality Factors from ChatGPT: More meaningful than Impact Factors?
by: Thelwall, Mike, et al.
Published: (2024)
by: Thelwall, Mike, et al.
Published: (2024)
Evaluating the quality of published medical research with ChatGPT
by: Thelwall, Mike, et al.
Published: (2024)
by: Thelwall, Mike, et al.
Published: (2024)
Designing large language model prompts to extract scores from messy text: A shared dataset and challenge
by: Thelwall, Mike
Published: (2026)
by: Thelwall, Mike
Published: (2026)
Can ChatGPT evaluate research environments? Evidence from REF2021
by: Kousha, Kayvan, et al.
Published: (2025)
by: Kousha, Kayvan, et al.
Published: (2025)
Assessing the societal influence of academic research with ChatGPT: Impact case study evaluations
by: Kousha, Kayvan, et al.
Published: (2024)
by: Kousha, Kayvan, et al.
Published: (2024)
How much are LLMs changing the language of academic papers after ChatGPT? A multi-database and full text analysis
by: Kousha, Kayvan, et al.
Published: (2025)
by: Kousha, Kayvan, et al.
Published: (2025)
Can ChatGPT be a good follower of academic paradigms? Research quality evaluations in conflicting areas of sociology
by: Thelwall, Mike, et al.
Published: (2025)
by: Thelwall, Mike, et al.
Published: (2025)
Is ChatGPT Transforming Academics' Writing Style?
by: Geng, Mingmeng, et al.
Published: (2024)
by: Geng, Mingmeng, et al.
Published: (2024)
Can Large Language Models Evaluate Grant Proposal Quality? Revisiting the Wennerås and Wold Peer Review Data
by: Sandström, Ulf, et al.
Published: (2026)
by: Sandström, Ulf, et al.
Published: (2026)
Sentiment Analysis of Citations in Scientific Articles Using ChatGPT: Identifying Potential Biases and Conflicts of Interest
by: Hariri, Walid
Published: (2024)
by: Hariri, Walid
Published: (2024)
Delving into the Utilisation of ChatGPT in Scientific Publications in Astronomy
by: Astarita, Simone, et al.
Published: (2024)
by: Astarita, Simone, et al.
Published: (2024)
Quantitative Methods in Research Evaluation Citation Indicators, Altmetrics, and Artificial Intelligence
by: Thelwall, Mike
Published: (2024)
by: Thelwall, Mike
Published: (2024)
Quantifying Similarity: Text-Mining Approaches to Evaluate ChatGPT and Google Bard Content in Relation to BioMedical Literature
by: Klimczak, Jakub, et al.
Published: (2024)
by: Klimczak, Jakub, et al.
Published: (2024)
Can Smaller Large Language Models Evaluate Research Quality?
by: Thelwall, Mike
Published: (2025)
by: Thelwall, Mike
Published: (2025)
Research quality evaluation by AI in the era of Large Language Models: Advantages, disadvantages, and systemic effects
by: Thelwall, Mike
Published: (2025)
by: Thelwall, Mike
Published: (2025)
Prompt perturbation and fraction facilitation sometimes strengthen Large Language Model scores
by: Thelwall, Mike
Published: (2025)
by: Thelwall, Mike
Published: (2025)
Do Large Language Models know Which Published Articles have been Retracted?
by: Thelwall, Mike
Published: (2026)
by: Thelwall, Mike
Published: (2026)
Is OpenAlex Suitable for Research Quality Evaluation and Which Citation Indicator is Best?
by: Thelwall, Mike, et al.
Published: (2025)
by: Thelwall, Mike, et al.
Published: (2025)
Automated Peer Reviewing in Paper SEA: Standardization, Evaluation, and Analysis
by: Yu, Jianxiang, et al.
Published: (2024)
by: Yu, Jianxiang, et al.
Published: (2024)
From ChatGPT, DALL-E 3 to Sora: How has Generative AI Changed Digital Humanities Research and Services?
by: Liu, Jiangfeng, et al.
Published: (2024)
by: Liu, Jiangfeng, et al.
Published: (2024)
Will AI be overconfident about academic research findings when reliant on abstracts? (v1)
by: Thelwall, Mike
Published: (2026)
by: Thelwall, Mike
Published: (2026)
Can social media provide early warning of retraction? Evidence from critical tweets identified by human annotation and large language models
by: Zheng, Er-Te, et al.
Published: (2024)
by: Zheng, Er-Te, et al.
Published: (2024)
Commitment Checklist: Auditing Author Commitments in Peer Review
by: Chen, Chung-Chi, et al.
Published: (2026)
by: Chen, Chung-Chi, et al.
Published: (2026)
Have LLM-associated terms increased in article full texts in all fields?
by: Thelwall, Mike, et al.
Published: (2026)
by: Thelwall, Mike, et al.
Published: (2026)
Causal Effect of Group Diversity on Redundancy and Coverage in Peer-Reviewing
by: Goyal, Navita, et al.
Published: (2024)
by: Goyal, Navita, et al.
Published: (2024)
An Evaluation of GPT-4V for Transcribing the Urban Renewal Hand-Written Collection
by: Lee, Myeong, et al.
Published: (2024)
by: Lee, Myeong, et al.
Published: (2024)
ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature
by: Gray, Andrew
Published: (2024)
by: Gray, Andrew
Published: (2024)
Where there's a will there's a way: ChatGPT is used more for science in countries where it is prohibited
by: Bao, Honglin, et al.
Published: (2024)
by: Bao, Honglin, et al.
Published: (2024)
Large Language Models for Departmental Expert Review Quality Scores
by: Langfeldt, Liv, et al.
Published: (2026)
by: Langfeldt, Liv, et al.
Published: (2026)
Can Small and Reasoning Large Language Models Score Journal Articles for Research Quality and Do Averaging and Few-shot Help?
by: Thelwall, Mike, et al.
Published: (2025)
by: Thelwall, Mike, et al.
Published: (2025)
AnalyticsGPT: An LLM Workflow for Scientometric Question Answering
by: Ly, Khang, et al.
Published: (2026)
by: Ly, Khang, et al.
Published: (2026)
Similar Items
-
In which fields can ChatGPT detect journal article quality? An evaluation of REF2021 results
by: Thelwall, Mike, et al.
Published: (2024) -
Can ChatGPT evaluate research quality?
by: Thelwall, Mike
Published: (2024) -
A Global South Strategy for Evaluating Research Value with ChatGPT
by: Nunkoo, Robin, et al.
Published: (2025) -
In which fields do ChatGPT scores align better than citations with research quality?
by: Thelwall, Mike
Published: (2025) -
Which stylistic features fool ChatGPT research evaluations?
by: Kousha, Kayvan, et al.
Published: (2026)