Exploring AI-Enabled Test Practice, Affect, and Test Outcomes in Language Assessment
Fuente:
arXiv
Saved in:
| Main Authors: | Burstein, Jill, Cardwell, Ramsey, Chuang, Ping-Ling, Michalowski, Allison, Nydick, Steven |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Where Assessment Validation and Responsible AI Meet
by: Burstein, Jill, et al.
Published: (2024)
by: Burstein, Jill, et al.
Published: (2024)
Responsible AI for Test Equity and Quality: The Duolingo English Test as a Case Study
by: Burstein, Jill, et al.
Published: (2024)
by: Burstein, Jill, et al.
Published: (2024)
Assessing AI Utility: The Random Guesser Test for Sequential Decision-Making Systems
by: Ide, Shun, et al.
Published: (2024)
by: Ide, Shun, et al.
Published: (2024)
Introducing AI to an Online Petition Platform Changed Outputs but not Outcomes
by: Corpus, Isabel, et al.
Published: (2025)
by: Corpus, Isabel, et al.
Published: (2025)
Empowering Affected Individuals to Shape AI Fairness Assessments: Processes, Criteria, and Tools
by: Luo, Lin, et al.
Published: (2026)
by: Luo, Lin, et al.
Published: (2026)
Testing network clustering algorithms with Natural Language Processing
by: Achitouv, Ixandra, et al.
Published: (2024)
by: Achitouv, Ixandra, et al.
Published: (2024)
When Testing AI Tests Us: Safeguarding Mental Health on the Digital Frontlines
by: Pendse, Sachin R., et al.
Published: (2025)
by: Pendse, Sachin R., et al.
Published: (2025)
Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies
by: Brundage, Miles, et al.
Published: (2026)
by: Brundage, Miles, et al.
Published: (2026)
Consumer UAV Cybersecurity Vulnerability Assessment Using Fuzzing Tests
by: Rudo, David, et al.
Published: (2020)
by: Rudo, David, et al.
Published: (2020)
Bridging Psychometric and Content Development Practices with AI: A Community-Based Workflow for Augmenting Hawaiian Language Assessments
by: Kūkea-Shultz, Pōhai, et al.
Published: (2025)
by: Kūkea-Shultz, Pōhai, et al.
Published: (2025)
The Lovelace Test of Intelligence: Can Humans Recognise and Esteem AI-Generated Art?
by: Gajewska, Ewelina
Published: (2025)
by: Gajewska, Ewelina
Published: (2025)
Drawing Your Programs: Exploring the Applications of Visual-Prompting with GenAI for Teaching and Assessment
by: Smith IV, David H., et al.
Published: (2026)
by: Smith IV, David H., et al.
Published: (2026)
TestAgent: An Adaptive and Intelligent Expert for Human Assessment
by: Yu, Junhao, et al.
Published: (2025)
by: Yu, Junhao, et al.
Published: (2025)
Toward Third-Party Assurance of AI Systems: Design Requirements, Prototype, and Early Testing
by: Kim, Rachel M., et al.
Published: (2026)
by: Kim, Rachel M., et al.
Published: (2026)
A Practical Guide for Supporting Formative Assessment and Feedback Using Generative AI
by: Prompiengchai, Sapolnach, et al.
Published: (2025)
by: Prompiengchai, Sapolnach, et al.
Published: (2025)
Towards Effective Discrimination Testing for Generative AI
by: Zollo, Thomas P., et al.
Published: (2024)
by: Zollo, Thomas P., et al.
Published: (2024)
A Risk Assessment Framework for Digital Identification Systems
by: Woodruff, Allison, et al.
Published: (2025)
by: Woodruff, Allison, et al.
Published: (2025)
Fairness Testing of Large Language Models in Role-Playing
by: Li, Xinyue, et al.
Published: (2024)
by: Li, Xinyue, et al.
Published: (2024)
Start Making Sense: Practical Approaches to Outcomes Assessment for Libraries.
by: Rabine, Julie, et al.
Published: (2000)
by: Rabine, Julie, et al.
Published: (2000)
Automated Question Generation for Science Tests in Arabic Language Using NLP Techniques
by: Tami, Mohammad, et al.
Published: (2024)
by: Tami, Mohammad, et al.
Published: (2024)
Exploring Factors Affecting Student Learning Satisfaction during COVID-19 in South Korea
by: Han, Jiwon, et al.
Published: (2024)
by: Han, Jiwon, et al.
Published: (2024)
Personalized AI Practice Replicates Learning Rate Regularity at Scale
by: Beauchesne, Jocelyn, et al.
Published: (2026)
by: Beauchesne, Jocelyn, et al.
Published: (2026)
Exploring Student-AI Interactions in Vibe Coding
by: Geng, Francis, et al.
Published: (2025)
by: Geng, Francis, et al.
Published: (2025)
The Ideological Turing Test for Moderation of Outgroup Affective Animosity
by: Gamba, David, et al.
Published: (2025)
by: Gamba, David, et al.
Published: (2025)
Exploring Teacher-Chatbot Interaction and Affect in Block-Based Programming
by: Riahi, Bahare, et al.
Published: (2026)
by: Riahi, Bahare, et al.
Published: (2026)
Paths to Testing: Why Women Enter and Remain in Software Testing?
by: Silva, Kleice, et al.
Published: (2024)
by: Silva, Kleice, et al.
Published: (2024)
Using Games to Learn How Large Language Models Work
by: Chen, Allison, et al.
Published: (2026)
by: Chen, Allison, et al.
Published: (2026)
A Canary in the AI Coal Mine: American Jews May Be Disproportionately Harmed by Intellectual Property Dispossession in Large Language Model Training
by: Precel, Heila, et al.
Published: (2024)
by: Precel, Heila, et al.
Published: (2024)
Large Language Models as Students Who Think Aloud: Overly Coherent, Verbose, and Confident
by: Borchers, Conrad, et al.
Published: (2026)
by: Borchers, Conrad, et al.
Published: (2026)
AI-rays: Exploring Bias in the Gaze of AI Through a Multimodal Interactive Installation
by: Gao, Ziyao, et al.
Published: (2024)
by: Gao, Ziyao, et al.
Published: (2024)
Enabling the AI Revolution in Healthcare
by: Singh, Mona, et al.
Published: (2025)
by: Singh, Mona, et al.
Published: (2025)
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
by: Morabito, Robert, et al.
Published: (2024)
by: Morabito, Robert, et al.
Published: (2024)
AI Data Centers Need Pioneers to Deliver Scalable Power via Offgrid AI
by: Reinhardt, Steven P.
Published: (2025)
by: Reinhardt, Steven P.
Published: (2025)
AI-Generated Letters from the Future: A Randomized Test of Personalized Climate Communication
by: Powdthavee, Nattavudh, et al.
Published: (2026)
by: Powdthavee, Nattavudh, et al.
Published: (2026)
Assessment Twins: A Protocol for AI-Vulnerable Summative Assessment
by: Roe, Jasper, et al.
Published: (2025)
by: Roe, Jasper, et al.
Published: (2025)
The AI Assessment Scale Revisited: A Framework for Educational Assessment
by: Perkins, Mike, et al.
Published: (2024)
by: Perkins, Mike, et al.
Published: (2024)
How to Assess Trustworthy AI in Practice
by: Zicari, Roberto V., et al.
Published: (2022)
by: Zicari, Roberto V., et al.
Published: (2022)
Scrutinizing Index-Based Risk Assessments: A Case Study in NYC Decision-making for Heat Emergency Management
by: Gosciak, Jennah, et al.
Published: (2026)
by: Gosciak, Jennah, et al.
Published: (2026)
Exploring LLMs for Predicting Tutor Strategy and Student Outcomes in Dialogues
by: Ikram, Fareya, et al.
Published: (2025)
by: Ikram, Fareya, et al.
Published: (2025)
Who Gets Seen in the Age of AI? Adoption Patterns of Large Language Models in Scholarly Writing and Citation Outcomes
by: Khan, Farhan Kamrul, et al.
Published: (2025)
by: Khan, Farhan Kamrul, et al.
Published: (2025)
Similar Items
-
Where Assessment Validation and Responsible AI Meet
by: Burstein, Jill, et al.
Published: (2024) -
Responsible AI for Test Equity and Quality: The Duolingo English Test as a Case Study
by: Burstein, Jill, et al.
Published: (2024) -
Assessing AI Utility: The Random Guesser Test for Sequential Decision-Making Systems
by: Ide, Shun, et al.
Published: (2024) -
Introducing AI to an Online Petition Platform Changed Outputs but not Outcomes
by: Corpus, Isabel, et al.
Published: (2025) -
Empowering Affected Individuals to Shape AI Fairness Assessments: Processes, Criteria, and Tools
by: Luo, Lin, et al.
Published: (2026)