Evaluating the capability of large language models to personalize science texts for diverse middle-school-age learners
Fuente:
arXiv
Saved in:
| Main Authors: | Vaccaro Jr, Michael, Friday, Mikayla, Zaghi, Arash |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Large-scale moral machine experiment on large language models
by: Ahmad, Muhammad Shahrul Zaim bin, et al.
Published: (2024)
by: Ahmad, Muhammad Shahrul Zaim bin, et al.
Published: (2024)
Assessing the nature of large language models: A caution against anthropocentrism
by: Speed, Ann
Published: (2023)
by: Speed, Ann
Published: (2023)
Recourse for reclamation: Chatting with generative language models
by: Chien, Jennifer, et al.
Published: (2024)
by: Chien, Jennifer, et al.
Published: (2024)
A validity-guided workflow for robust large language model research in psychology
by: Lin, Zhicheng
Published: (2025)
by: Lin, Zhicheng
Published: (2025)
Evidence of a log scaling law for political persuasion with large language models
by: Hackenburg, Kobi, et al.
Published: (2024)
by: Hackenburg, Kobi, et al.
Published: (2024)
Malinowski in the Age of AI: Can large language models create a text game based on an anthropological classic?
by: Hoffmann, Michael Peter, et al.
Published: (2024)
by: Hoffmann, Michael Peter, et al.
Published: (2024)
The opportunities and risks of large language models in mental health
by: Lawrence, Hannah R., et al.
Published: (2024)
by: Lawrence, Hannah R., et al.
Published: (2024)
How do datasets, developers, and models affect biases in a low-resourced language?: The Case of the Bengali Language
by: Das, Dipto, et al.
Published: (2025)
by: Das, Dipto, et al.
Published: (2025)
PhDGPT: Introducing a psychometric and linguistic dataset about how large language models perceive graduate students and professors in psychology
by: De Duro, Edoardo Sebastiano, et al.
Published: (2024)
by: De Duro, Edoardo Sebastiano, et al.
Published: (2024)
Acceptance of cybernetic avatars for capability enhancement: a large-scale survey
by: Aymerich-Franch, Laura, et al.
Published: (2025)
by: Aymerich-Franch, Laura, et al.
Published: (2025)
Artificial intelligence language technologies in multilingual healthcare: Grand challenges ahead
by: Briva-Iglesias, Vicent
Published: (2026)
by: Briva-Iglesias, Vicent
Published: (2026)
Generics in science communication: Misaligned interpretations across laypeople, scientists, and large language models
by: Peters, Uwe, et al.
Published: (2026)
by: Peters, Uwe, et al.
Published: (2026)
Enhancing behavioral nudges with large language model-based iterative personalization: A field experiment on electricity and hot-water conservation
by: Li, Zonghan, et al.
Published: (2026)
by: Li, Zonghan, et al.
Published: (2026)
NARRA-Gym for Evaluating Interactive Narrative Agents
by: Huang, Yue, et al.
Published: (2026)
by: Huang, Yue, et al.
Published: (2026)
Evaluating Large Language Models in Theory of Mind Tasks
by: Kosinski, Michal
Published: (2023)
by: Kosinski, Michal
Published: (2023)
Multi-turn Evaluation of Anthropomorphic Behaviours in Large Language Models
by: Ibrahim, Lujain, et al.
Published: (2025)
by: Ibrahim, Lujain, et al.
Published: (2025)
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
by: Arita, Takaya, et al.
Published: (2025)
by: Arita, Takaya, et al.
Published: (2025)
A Systematic Review on the Evaluation of Large Language Models in Theory of Mind Tasks
by: Sarıtaş, Karahan, et al.
Published: (2025)
by: Sarıtaş, Karahan, et al.
Published: (2025)
DaKultur: Evaluating the Cultural Awareness of Language Models for Danish with Native Speakers
by: Müller-Eberstein, Max, et al.
Published: (2025)
by: Müller-Eberstein, Max, et al.
Published: (2025)
SOTOPIA-$Ω$: Dynamic Strategy Injection Learning and Social Instruction Following Evaluation for Social Agents
by: Zhang, Wenyuan, et al.
Published: (2025)
by: Zhang, Wenyuan, et al.
Published: (2025)
Automatic deductive coding in discourse analysis: an application of large language models in learning analytics
by: Zhang, Lishan, et al.
Published: (2024)
by: Zhang, Lishan, et al.
Published: (2024)
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
by: Badawi, Abeer, et al.
Published: (2025)
by: Badawi, Abeer, et al.
Published: (2025)
Threefold model for AI Readiness: A Case Study with Finnish Healthcare SMEs
by: Alnajjar, Mohammed, et al.
Published: (2025)
by: Alnajjar, Mohammed, et al.
Published: (2025)
Medical large language models are easily distracted
by: Vishwanath, Krithik, et al.
Published: (2025)
by: Vishwanath, Krithik, et al.
Published: (2025)
To what extent is ChatGPT useful for language teacher lesson plan creation?
by: Dornburg, Alex, et al.
Published: (2024)
by: Dornburg, Alex, et al.
Published: (2024)
Evaluating Human-AI Safety: A Framework for Measuring Harmful Capability Uplift
by: Vaccaro, Michelle, et al.
Published: (2026)
by: Vaccaro, Michelle, et al.
Published: (2026)
Fact-checking information from large language models can decrease headline discernment
by: DeVerna, Matthew R., et al.
Published: (2023)
by: DeVerna, Matthew R., et al.
Published: (2023)
Do AI tutors empower or enslave learners? Toward a critical use of AI in education
by: Favero, Lucile, et al.
Published: (2025)
by: Favero, Lucile, et al.
Published: (2025)
Virtual Agent-Based Communication Skills Training to Facilitate Health Persuasion Among Peers
by: Nouraei, Farnaz, et al.
Published: (2024)
by: Nouraei, Farnaz, et al.
Published: (2024)
Embarrassed to observe: The effects of directive language in brand conversation
by: Andriuzzi, Andria, et al.
Published: (2025)
by: Andriuzzi, Andria, et al.
Published: (2025)
Medication counseling with large language models: balancing flexibility and rigidity
by: Sabel, Joar, et al.
Published: (2025)
by: Sabel, Joar, et al.
Published: (2025)
When combinations of humans and AI are useful: A systematic review and meta-analysis
by: Vaccaro, Michelle, et al.
Published: (2024)
by: Vaccaro, Michelle, et al.
Published: (2024)
Creativity Benchmark: A benchmark for marketing creativity for large language models
by: Bhat, Ninad, et al.
Published: (2025)
by: Bhat, Ninad, et al.
Published: (2025)
Exploring the Ethical Concerns in User Reviews of Mental Health Apps using Topic Modeling and Sentiment Analysis
by: Rahman, Mohammad Masudur, et al.
Published: (2026)
by: Rahman, Mohammad Masudur, et al.
Published: (2026)
The Hidden Language of Harm: Examining the Role of Emojis in Harmful Online Communication and Content Moderation
by: Zhou, Yuhang, et al.
Published: (2025)
by: Zhou, Yuhang, et al.
Published: (2025)
Designing Computational Tools for Exploring Causal Relationships in Qualitative Data
by: Meng, Han, et al.
Published: (2026)
by: Meng, Han, et al.
Published: (2026)
What is Stigma Attributed to? A Theory-Grounded, Expert-Annotated Interview Corpus for Demystifying Mental-Health Stigma
by: Meng, Han, et al.
Published: (2025)
by: Meng, Han, et al.
Published: (2025)
Epistemological Fault Lines Between Human and Artificial Intelligence
by: Quattrociocchi, Walter, et al.
Published: (2025)
by: Quattrociocchi, Walter, et al.
Published: (2025)
"Would You Want an AI Tutor?" Understanding Stakeholder Perceptions of LLM-based Systems in the Classroom
by: Fuligni, Caterina, et al.
Published: (2025)
by: Fuligni, Caterina, et al.
Published: (2025)
When Algorithms Meet Artists: Semantic Compression of Artists' Concerns in the Public AI-Art Debate
by: Mukherjee-Gandhi, Ariya, et al.
Published: (2025)
by: Mukherjee-Gandhi, Ariya, et al.
Published: (2025)
Similar Items
-
Large-scale moral machine experiment on large language models
by: Ahmad, Muhammad Shahrul Zaim bin, et al.
Published: (2024) -
Assessing the nature of large language models: A caution against anthropocentrism
by: Speed, Ann
Published: (2023) -
Recourse for reclamation: Chatting with generative language models
by: Chien, Jennifer, et al.
Published: (2024) -
A validity-guided workflow for robust large language model research in psychology
by: Lin, Zhicheng
Published: (2025) -
Evidence of a log scaling law for political persuasion with large language models
by: Hackenburg, Kobi, et al.
Published: (2024)