Daily and Weekly Periodicity in Large Language Model Performance and Its Implications for Research
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tschisgale, Paul, Wulff, Peter |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating NLP Embedding Models for Handling Science-Specific Symbolic Expressions in Student Texts
von: Bleckmann, Tom, et al.
Veröffentlicht: (2025)
von: Bleckmann, Tom, et al.
Veröffentlicht: (2025)
Evaluating GPT- and Reasoning-based Large Language Models on Physics Olympiad Problems: Surpassing Human Performance and Implications for Educational Assessment
von: Tschisgale, Paul, et al.
Veröffentlicht: (2025)
von: Tschisgale, Paul, et al.
Veröffentlicht: (2025)
Developing and Evaluating a Large Language Model-Based Automated Feedback System Grounded in Evidence-Centered Design for Supporting Physics Problem Solving
von: Maus, Holger, et al.
Veröffentlicht: (2025)
von: Maus, Holger, et al.
Veröffentlicht: (2025)
Metacognitive Myopia in Large Language Models
von: Scholten, Florian, et al.
Veröffentlicht: (2024)
von: Scholten, Florian, et al.
Veröffentlicht: (2024)
Beyond Words: How Large Language Models Perform in Quantitative Management Problem-Solving
von: Kuzmanko, Jonathan
Veröffentlicht: (2025)
von: Kuzmanko, Jonathan
Veröffentlicht: (2025)
TransitGPT: A Generative AI-based framework for interacting with GTFS data using Large Language Models
von: Devunuri, Saipraneeth, et al.
Veröffentlicht: (2024)
von: Devunuri, Saipraneeth, et al.
Veröffentlicht: (2024)
The GPT Surprise: Offering Large Language Model Chat in a Massive Coding Class Reduced Engagement but Increased Adopters Exam Performances
von: Nie, Allen, et al.
Veröffentlicht: (2024)
von: Nie, Allen, et al.
Veröffentlicht: (2024)
Large Language Model-Based Agents for Automated Research Reproducibility: An Exploratory Study in Alzheimer's Disease
von: Dobbins, Nic, et al.
Veröffentlicht: (2025)
von: Dobbins, Nic, et al.
Veröffentlicht: (2025)
Limits of Large Language Models in Debating Humans
von: Flamino, James, et al.
Veröffentlicht: (2024)
von: Flamino, James, et al.
Veröffentlicht: (2024)
Domain-Shift-Aware Conformal Prediction for Large Language Models
von: Lin, Zhexiao, et al.
Veröffentlicht: (2025)
von: Lin, Zhexiao, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Adaptation of Large Language Models for Protein-Protein Interaction Analysis
von: Jantre, Sanket, et al.
Veröffentlicht: (2025)
von: Jantre, Sanket, et al.
Veröffentlicht: (2025)
Augmented Risk Prediction for the Onset of Alzheimer's Disease from Electronic Health Records with Large Language Models
von: Wang, Jiankun, et al.
Veröffentlicht: (2024)
von: Wang, Jiankun, et al.
Veröffentlicht: (2024)
Assessing Large Language Models in Mechanical Engineering Education: A Study on Mechanics-Focused Conceptual Understanding
von: Tian, Jie, et al.
Veröffentlicht: (2024)
von: Tian, Jie, et al.
Veröffentlicht: (2024)
From Traditional Taggers to LLMs: A Comparative Study of POS Tagging for Medieval Romance Languages
von: Schöffel, Matthias, et al.
Veröffentlicht: (2026)
von: Schöffel, Matthias, et al.
Veröffentlicht: (2026)
"All that Glitters": Approaches to Evaluations with Unreliable Model and Human Annotations
von: Hardy, Michael
Veröffentlicht: (2024)
von: Hardy, Michael
Veröffentlicht: (2024)
ImplicitRM: Unbiased Reward Modeling from Implicit Preference Data for LLM alignment
von: Wang, Hao, et al.
Veröffentlicht: (2026)
von: Wang, Hao, et al.
Veröffentlicht: (2026)
Chitchat with AI: Understand the supply chain carbon disclosure of companies worldwide through Large Language Model
von: Hang, Haotian, et al.
Veröffentlicht: (2025)
von: Hang, Haotian, et al.
Veröffentlicht: (2025)
Contextual Phenotyping of Pediatric Sepsis Cohort Using Large Language Models
von: Nagori, Aditya, et al.
Veröffentlicht: (2025)
von: Nagori, Aditya, et al.
Veröffentlicht: (2025)
Language Models as Causal Effect Generators
von: Bynum, Lucius E. J., et al.
Veröffentlicht: (2024)
von: Bynum, Lucius E. J., et al.
Veröffentlicht: (2024)
The Use of a Large Language Model for Cyberbullying Detection
von: Ogunleye, Bayode, et al.
Veröffentlicht: (2024)
von: Ogunleye, Bayode, et al.
Veröffentlicht: (2024)
Dissecting Physics Reasoning in Small Language Models: A Multi-Dimensional Analysis from an Educational Perspective
von: Scaria, Nicy, et al.
Veröffentlicht: (2025)
von: Scaria, Nicy, et al.
Veröffentlicht: (2025)
RJUA-MedDQA: A Multimodal Benchmark for Medical Document Question Answering and Clinical Reasoning
von: Jin, Congyun, et al.
Veröffentlicht: (2024)
von: Jin, Congyun, et al.
Veröffentlicht: (2024)
DeepScore: A Comprehensive Approach to Measuring Quality in AI-Generated Clinical Documentation
von: Oleson, Jon
Veröffentlicht: (2024)
von: Oleson, Jon
Veröffentlicht: (2024)
Improving LLM Leaderboards with Psychometrical Methodology
von: Federiakin, Denis
Veröffentlicht: (2025)
von: Federiakin, Denis
Veröffentlicht: (2025)
Collective Reasoning Among LLMs: A Framework for Answer Validation Without Ground Truth
von: Davoudi, Seyed Pouyan Mousavi, et al.
Veröffentlicht: (2025)
von: Davoudi, Seyed Pouyan Mousavi, et al.
Veröffentlicht: (2025)
A Bayesian Approach to Harnessing the Power of LLMs in Authorship Attribution
von: Hu, Zhengmian, et al.
Veröffentlicht: (2024)
von: Hu, Zhengmian, et al.
Veröffentlicht: (2024)
Language-Dependent Political Bias in AI: A Study of ChatGPT and Gemini
von: Yuksel, Dogus, et al.
Veröffentlicht: (2025)
von: Yuksel, Dogus, et al.
Veröffentlicht: (2025)
Performance Evaluation of Large Language Models in Statistical Programming
von: Song, Xinyi, et al.
Veröffentlicht: (2025)
von: Song, Xinyi, et al.
Veröffentlicht: (2025)
Reliable and Efficient Amortized Model-based Evaluation
von: Truong, Sang, et al.
Veröffentlicht: (2025)
von: Truong, Sang, et al.
Veröffentlicht: (2025)
United in Diversity? Contextual Biases in LLM-Based Predictions of the 2024 European Parliament Elections
von: von der Heyde, Leah, et al.
Veröffentlicht: (2024)
von: von der Heyde, Leah, et al.
Veröffentlicht: (2024)
LLM4ED: Large Language Models for Automatic Equation Discovery
von: Du, Mengge, et al.
Veröffentlicht: (2024)
von: Du, Mengge, et al.
Veröffentlicht: (2024)
Addressing Longstanding Challenges in Cognitive Science with Language Models
von: Wulff, Dirk U., et al.
Veröffentlicht: (2025)
von: Wulff, Dirk U., et al.
Veröffentlicht: (2025)
StatLLM: A Dataset for Evaluating the Performance of Large Language Models in Statistical Analysis
von: Song, Xinyi, et al.
Veröffentlicht: (2025)
von: Song, Xinyi, et al.
Veröffentlicht: (2025)
Evaluating the Use of Large Language Models as Synthetic Social Agents in Social Science Research
von: Madden, Emma Rose
Veröffentlicht: (2025)
von: Madden, Emma Rose
Veröffentlicht: (2025)
Online Reasoning Calibration: Test-Time Training Enables Generalizable Conformal LLM Reasoning
von: Zhou, Cai, et al.
Veröffentlicht: (2026)
von: Zhou, Cai, et al.
Veröffentlicht: (2026)
Beyond the Hype: Embeddings vs. Prompting for Multiclass Classification Tasks
von: Kokkodis, Marios, et al.
Veröffentlicht: (2025)
von: Kokkodis, Marios, et al.
Veröffentlicht: (2025)
ICE-ID: A Novel Historical Census Dataset for Longitudinal Identity Resolution
von: de Carvalho, Gonçalo Hora, et al.
Veröffentlicht: (2025)
von: de Carvalho, Gonçalo Hora, et al.
Veröffentlicht: (2025)
Reinforcement Learning from Human Feedback with High-Confidence Safety Constraints
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2025)
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2025)
Unified Representation of Genomic and Biomedical Concepts through Multi-Task, Multi-Source Contrastive Learning
von: Yuan, Hongyi, et al.
Veröffentlicht: (2024)
von: Yuan, Hongyi, et al.
Veröffentlicht: (2024)
A Rational Analysis of the Speech-to-Song Illusion
von: Marjieh, Raja, et al.
Veröffentlicht: (2024)
von: Marjieh, Raja, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Evaluating NLP Embedding Models for Handling Science-Specific Symbolic Expressions in Student Texts
von: Bleckmann, Tom, et al.
Veröffentlicht: (2025) -
Evaluating GPT- and Reasoning-based Large Language Models on Physics Olympiad Problems: Surpassing Human Performance and Implications for Educational Assessment
von: Tschisgale, Paul, et al.
Veröffentlicht: (2025) -
Developing and Evaluating a Large Language Model-Based Automated Feedback System Grounded in Evidence-Centered Design for Supporting Physics Problem Solving
von: Maus, Holger, et al.
Veröffentlicht: (2025) -
Metacognitive Myopia in Large Language Models
von: Scholten, Florian, et al.
Veröffentlicht: (2024) -
Beyond Words: How Large Language Models Perform in Quantitative Management Problem-Solving
von: Kuzmanko, Jonathan
Veröffentlicht: (2025)