Can Smaller Large Language Models Evaluate Research Quality?
Fuente:
arXiv
Salvato in:
| Autore principale: | Thelwall, Mike |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Evaluating Research Quality with Large Language Models: An Analysis of ChatGPT's Effectiveness with Different Settings and Inputs
di: Thelwall, Mike
Pubblicazione: (2024)
di: Thelwall, Mike
Pubblicazione: (2024)
Can Small and Reasoning Large Language Models Score Journal Articles for Research Quality and Do Averaging and Few-shot Help?
di: Thelwall, Mike, et al.
Pubblicazione: (2025)
di: Thelwall, Mike, et al.
Pubblicazione: (2025)
Can ChatGPT evaluate research quality?
di: Thelwall, Mike
Pubblicazione: (2024)
di: Thelwall, Mike
Pubblicazione: (2024)
Can Large Language Models Evaluate Grant Proposal Quality? Revisiting the Wennerås and Wold Peer Review Data
di: Sandström, Ulf, et al.
Pubblicazione: (2026)
di: Sandström, Ulf, et al.
Pubblicazione: (2026)
Assessing the societal influence of academic research with ChatGPT: Impact case study evaluations
di: Kousha, Kayvan, et al.
Pubblicazione: (2024)
di: Kousha, Kayvan, et al.
Pubblicazione: (2024)
Evaluating the quality of published medical research with ChatGPT
di: Thelwall, Mike, et al.
Pubblicazione: (2024)
di: Thelwall, Mike, et al.
Pubblicazione: (2024)
Research quality evaluation by AI in the era of Large Language Models: Advantages, disadvantages, and systemic effects
di: Thelwall, Mike
Pubblicazione: (2025)
di: Thelwall, Mike
Pubblicazione: (2025)
Prompt perturbation and fraction facilitation sometimes strengthen Large Language Model scores
di: Thelwall, Mike
Pubblicazione: (2025)
di: Thelwall, Mike
Pubblicazione: (2025)
Do Large Language Models know Which Published Articles have been Retracted?
di: Thelwall, Mike
Pubblicazione: (2026)
di: Thelwall, Mike
Pubblicazione: (2026)
Is OpenAlex Suitable for Research Quality Evaluation and Which Citation Indicator is Best?
di: Thelwall, Mike, et al.
Pubblicazione: (2025)
di: Thelwall, Mike, et al.
Pubblicazione: (2025)
Quantitative Methods in Research Evaluation Citation Indicators, Altmetrics, and Artificial Intelligence
di: Thelwall, Mike
Pubblicazione: (2024)
di: Thelwall, Mike
Pubblicazione: (2024)
Implicit and Explicit Research Quality Score Probabilities from ChatGPT
di: Thelwall, Mike, et al.
Pubblicazione: (2025)
di: Thelwall, Mike, et al.
Pubblicazione: (2025)
Large Language Models for Departmental Expert Review Quality Scores
di: Langfeldt, Liv, et al.
Pubblicazione: (2026)
di: Langfeldt, Liv, et al.
Pubblicazione: (2026)
A Global South Strategy for Evaluating Research Value with ChatGPT
di: Nunkoo, Robin, et al.
Pubblicazione: (2025)
di: Nunkoo, Robin, et al.
Pubblicazione: (2025)
Can social media provide early warning of retraction? Evidence from critical tweets identified by human annotation and large language models
di: Zheng, Er-Te, et al.
Pubblicazione: (2024)
di: Zheng, Er-Te, et al.
Pubblicazione: (2024)
Journal Quality Factors from ChatGPT: More meaningful than Impact Factors?
di: Thelwall, Mike, et al.
Pubblicazione: (2024)
di: Thelwall, Mike, et al.
Pubblicazione: (2024)
Can ChatGPT be a good follower of academic paradigms? Research quality evaluations in conflicting areas of sociology
di: Thelwall, Mike, et al.
Pubblicazione: (2025)
di: Thelwall, Mike, et al.
Pubblicazione: (2025)
In which fields do ChatGPT scores align better than citations with research quality?
di: Thelwall, Mike
Pubblicazione: (2025)
di: Thelwall, Mike
Pubblicazione: (2025)
Research evaluation with ChatGPT: Is it age, country, length, or field biased?
di: Thelwall, Mike, et al.
Pubblicazione: (2024)
di: Thelwall, Mike, et al.
Pubblicazione: (2024)
Can ChatGPT evaluate research environments? Evidence from REF2021
di: Kousha, Kayvan, et al.
Pubblicazione: (2025)
di: Kousha, Kayvan, et al.
Pubblicazione: (2025)
Will AI be overconfident about academic research findings when reliant on abstracts? (v1)
di: Thelwall, Mike
Pubblicazione: (2026)
di: Thelwall, Mike
Pubblicazione: (2026)
Designing large language model prompts to extract scores from messy text: A shared dataset and challenge
di: Thelwall, Mike
Pubblicazione: (2026)
di: Thelwall, Mike
Pubblicazione: (2026)
Estimating the quality of academic books from their descriptions with ChatGPT
di: Thelwall, Mike, et al.
Pubblicazione: (2025)
di: Thelwall, Mike, et al.
Pubblicazione: (2025)
How much are LLMs changing the language of academic papers after ChatGPT? A multi-database and full text analysis
di: Kousha, Kayvan, et al.
Pubblicazione: (2025)
di: Kousha, Kayvan, et al.
Pubblicazione: (2025)
Which stylistic features fool ChatGPT research evaluations?
di: Kousha, Kayvan, et al.
Pubblicazione: (2026)
di: Kousha, Kayvan, et al.
Pubblicazione: (2026)
In which fields can ChatGPT detect journal article quality? An evaluation of REF2021 results
di: Thelwall, Mike, et al.
Pubblicazione: (2024)
di: Thelwall, Mike, et al.
Pubblicazione: (2024)
Have LLM-associated terms increased in article full texts in all fields?
di: Thelwall, Mike, et al.
Pubblicazione: (2026)
di: Thelwall, Mike, et al.
Pubblicazione: (2026)
Evaluating the Predictive Capacity of ChatGPT for Academic Peer Review Outcomes Across Multiple Platforms
di: Thelwall, Mike, et al.
Pubblicazione: (2024)
di: Thelwall, Mike, et al.
Pubblicazione: (2024)
Bridging the Evaluation Gap: Leveraging Large Language Models for Topic Model Evaluation
di: Tan, Zhiyin, et al.
Pubblicazione: (2025)
di: Tan, Zhiyin, et al.
Pubblicazione: (2025)
Toward Purpose-oriented Topic Model Evaluation enabled by Large Language Models
di: Tan, Zhiyin, et al.
Pubblicazione: (2025)
di: Tan, Zhiyin, et al.
Pubblicazione: (2025)
Matching Game Preferences Through Dialogical Large Language Models: A Perspective
di: Fabre, Renaud, et al.
Pubblicazione: (2025)
di: Fabre, Renaud, et al.
Pubblicazione: (2025)
Towards Large Language Models for Lunar Mission Planning and In Situ Resource Utilization
di: Pekala, Michael, et al.
Pubblicazione: (2025)
di: Pekala, Michael, et al.
Pubblicazione: (2025)
LLAssist: Simple Tools for Automating Literature Review Using Large Language Models
di: Haryanto, Christoforus Yoga
Pubblicazione: (2024)
di: Haryanto, Christoforus Yoga
Pubblicazione: (2024)
Fairness Evaluation of Large Language Models in Academic Library Reference Services
di: Wang, Haining, et al.
Pubblicazione: (2025)
di: Wang, Haining, et al.
Pubblicazione: (2025)
The emergence of Large Language Models (LLM) as a tool in literature reviews: an LLM automated systematic review
di: Scherbakov, Dmitry, et al.
Pubblicazione: (2024)
di: Scherbakov, Dmitry, et al.
Pubblicazione: (2024)
Can news and social media attention reduce the influence of problematic research?
di: Zheng, Er-Te, et al.
Pubblicazione: (2025)
di: Zheng, Er-Te, et al.
Pubblicazione: (2025)
Do Large Language Models Reduce Research Novelty? Evidence from Information Systems Journals
di: Safari, Ali
Pubblicazione: (2026)
di: Safari, Ali
Pubblicazione: (2026)
Named Entity Recognition of Historical Texts via Large Language Model
di: Zhang, Shibingfeng, et al.
Pubblicazione: (2025)
di: Zhang, Shibingfeng, et al.
Pubblicazione: (2025)
LLMs4Synthesis: Leveraging Large Language Models for Scientific Synthesis
di: Giglou, Hamed Babaei, et al.
Pubblicazione: (2024)
di: Giglou, Hamed Babaei, et al.
Pubblicazione: (2024)
Agreement Between Large Language Models, Human Reviewers, and Authors in Evaluating STROBE Checklists for Observational Studies in Rheumatology
di: Bilgin, Emre, et al.
Pubblicazione: (2026)
di: Bilgin, Emre, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Evaluating Research Quality with Large Language Models: An Analysis of ChatGPT's Effectiveness with Different Settings and Inputs
di: Thelwall, Mike
Pubblicazione: (2024) -
Can Small and Reasoning Large Language Models Score Journal Articles for Research Quality and Do Averaging and Few-shot Help?
di: Thelwall, Mike, et al.
Pubblicazione: (2025) -
Can ChatGPT evaluate research quality?
di: Thelwall, Mike
Pubblicazione: (2024) -
Can Large Language Models Evaluate Grant Proposal Quality? Revisiting the Wennerås and Wold Peer Review Data
di: Sandström, Ulf, et al.
Pubblicazione: (2026) -
Assessing the societal influence of academic research with ChatGPT: Impact case study evaluations
di: Kousha, Kayvan, et al.
Pubblicazione: (2024)