Simulation as Reality? The Effectiveness of LLM-Generated Data in Open-ended Question Assessment
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Long, Zhang, Meng, Wang, Wei Lin, Luo, Yu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Simulated Adoption: Decoupling Magnitude and Direction in LLM In-Context Conflict Resolution
por: Zhang, Long, et al.
Publicado: (2026)
por: Zhang, Long, et al.
Publicado: (2026)
Augmenting Human-Annotated Training Data with Large Language Model Generation and Distillation in Open-Response Assessment
por: Borchers, Conrad, et al.
Publicado: (2025)
por: Borchers, Conrad, et al.
Publicado: (2025)
Beyond Static Question Banks: Dynamic Knowledge Expansion via LLM-Automated Graph Construction and Adaptive Generation
por: Wang, Yingquan, et al.
Publicado: (2026)
por: Wang, Yingquan, et al.
Publicado: (2026)
PersonalityScanner: Exploring the Validity of Personality Assessment Based on Multimodal Signals in Virtual Reality
por: Zhang, Xintong, et al.
Publicado: (2024)
por: Zhang, Xintong, et al.
Publicado: (2024)
Comparing Rationality Between Large Language Models and Humans: Insights and Open Questions
por: Alsagheer, Dana, et al.
Publicado: (2024)
por: Alsagheer, Dana, et al.
Publicado: (2024)
Law in Silico: Simulating Legal Society with LLM-Based Agents
por: Wang, Yiding, et al.
Publicado: (2025)
por: Wang, Yiding, et al.
Publicado: (2025)
Explainable Ethical Assessment on Human Behaviors by Generating Conflicting Social Norms
por: Sun, Yuxi, et al.
Publicado: (2025)
por: Sun, Yuxi, et al.
Publicado: (2025)
EcoLANG: Efficient and Effective Agent Communication Language Induction for Social Simulation
por: Mou, Xinyi, et al.
Publicado: (2025)
por: Mou, Xinyi, et al.
Publicado: (2025)
What Can Student-AI Dialogues Tell Us About Students' Self-Regulated Learning? An exploratory framework
por: Zhang, Long, et al.
Publicado: (2025)
por: Zhang, Long, et al.
Publicado: (2025)
Integrating LLM and Diffusion-Based Agents for Social Simulation
por: Li, Xinyi, et al.
Publicado: (2025)
por: Li, Xinyi, et al.
Publicado: (2025)
An LLM -Powered Assessment Retrieval-Augmented Generation (RAG) For Higher Education
por: Barenji, Reza Vatankhah, et al.
Publicado: (2026)
por: Barenji, Reza Vatankhah, et al.
Publicado: (2026)
Chinese Court Simulation with LLM-Based Agent System
por: Zhang, Kaiyuan, et al.
Publicado: (2025)
por: Zhang, Kaiyuan, et al.
Publicado: (2025)
Controlling Cloze-test Question Item Difficulty with PLM-based Surrogate Models for IRT Assessment
por: Zhang, Jingshen, et al.
Publicado: (2024)
por: Zhang, Jingshen, et al.
Publicado: (2024)
Question Type, Cognitive Load, and CEFR Alignment: Evaluating LLM-Generated EFL Grammar Drill Exercises
por: Woollaston, Steve, et al.
Publicado: (2026)
por: Woollaston, Steve, et al.
Publicado: (2026)
SocioVerse: A World Model for Social Simulation Powered by LLM Agents and A Pool of 10 Million Real-World Users
por: Zhang, Xinnong, et al.
Publicado: (2025)
por: Zhang, Xinnong, et al.
Publicado: (2025)
MADS: Multi-Agent Dialogue Simulation for Diverse Persuasion Data Generation
por: Li, Mingjin, et al.
Publicado: (2025)
por: Li, Mingjin, et al.
Publicado: (2025)
IDEAlign: Comparing Large Language Models to Human Experts in Open-ended Interpretive Annotations
por: Nam, Hyunji, et al.
Publicado: (2025)
por: Nam, Hyunji, et al.
Publicado: (2025)
Using LLMs for Knowledge Component-level Correctness Labeling in Open-ended Coding Problems
por: Duan, Zhangqi, et al.
Publicado: (2026)
por: Duan, Zhangqi, et al.
Publicado: (2026)
Effective Data Stewardship in Higher Education: Skills, Competences, and the Emerging Role of Open Data Stewards
por: Fitsilis, Panos, et al.
Publicado: (2024)
por: Fitsilis, Panos, et al.
Publicado: (2024)
Predicting and Understanding Turn-Taking Behavior in Open-Ended Group Activities in Virtual Reality
por: Wang, Portia, et al.
Publicado: (2024)
por: Wang, Portia, et al.
Publicado: (2024)
Test Case-Informed Knowledge Tracing for Open-ended Coding Tasks
por: Duan, Zhangqi, et al.
Publicado: (2024)
por: Duan, Zhangqi, et al.
Publicado: (2024)
Examining and Comparing the Effectiveness of Virtual Reality Serious Games and LEGO Serious Play for Learning Scrum
por: Gordillo, Aldo, et al.
Publicado: (2024)
por: Gordillo, Aldo, et al.
Publicado: (2024)
Nudging the Somas: Exploring How Live-Configurable Mixed Reality Objects Shape Open-Ended Intercorporeal Movements
por: Hu, Botao Amber, et al.
Publicado: (2025)
por: Hu, Botao Amber, et al.
Publicado: (2025)
A Multi-agent Simulation for the Mass School Shootings
por: Dai, Wei, et al.
Publicado: (2024)
por: Dai, Wei, et al.
Publicado: (2024)
Smoke Screens and Scapegoats: The Reality of General Data Protection Regulation Compliance -- Privacy and Ethics in the Case of Replika AI
por: Piispanen, Joni-Roy, et al.
Publicado: (2024)
por: Piispanen, Joni-Roy, et al.
Publicado: (2024)
Human-like Social Compliance in Large Language Models: Unifying Sycophancy and Conformity through Signal Competition Dynamics
por: Zhang, Long, et al.
Publicado: (2025)
por: Zhang, Long, et al.
Publicado: (2025)
Another Body in the World: Flusserian Freedom in Mixed Reality
por: Zhou, Aven Le, et al.
Publicado: (2024)
por: Zhou, Aven Le, et al.
Publicado: (2024)
On the Trustworthiness of Generative Foundation Models: Guideline, Assessment, and Perspective
por: Huang, Yue, et al.
Publicado: (2025)
por: Huang, Yue, et al.
Publicado: (2025)
Beyond Prompting: An Efficient Embedding Framework for Open-Domain Question Answering
por: Hu, Zhanghao, et al.
Publicado: (2025)
por: Hu, Zhanghao, et al.
Publicado: (2025)
Improving Socratic Question Generation using Data Augmentation and Preference Optimization
por: Kumar, Nischal Ashok, et al.
Publicado: (2024)
por: Kumar, Nischal Ashok, et al.
Publicado: (2024)
Scaling Equitable Reflection Assessment in Education via Large Language Models and Role-Based Feedback Agents
por: Zhang, Chenyu, et al.
Publicado: (2025)
por: Zhang, Chenyu, et al.
Publicado: (2025)
TIM: A Large-Scale Dataset and large Timeline Intelligence Model for Open-domain Timeline Summarization
por: Hu, Chuanrui, et al.
Publicado: (2025)
por: Hu, Chuanrui, et al.
Publicado: (2025)
Synergy: A Next-Generation General-Purpose Agent for Open Agentic Web
por: Nie, Xiaohang, et al.
Publicado: (2026)
por: Nie, Xiaohang, et al.
Publicado: (2026)
Answering Students' Questions on Course Forums Using Multiple Chain-of-Thought Reasoning and Finetuning RAG-Enabled LLM
por: Wang, Neo, et al.
Publicado: (2025)
por: Wang, Neo, et al.
Publicado: (2025)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
por: Ma, Mingyu Derek, et al.
Publicado: (2023)
por: Ma, Mingyu Derek, et al.
Publicado: (2023)
Licensing Open Government Data
por: Lee, Jyh-An
Publicado: (2025)
por: Lee, Jyh-An
Publicado: (2025)
Designing Effective LLM-Assisted Interfaces for Curriculum Development
por: Faraji, Abdolali, et al.
Publicado: (2025)
por: Faraji, Abdolali, et al.
Publicado: (2025)
Reimagining Assessment in the Age of Generative AI: Lessons from Open-Book Exams with ChatGPT
por: Mahmoud, Qusay H.
Publicado: (2026)
por: Mahmoud, Qusay H.
Publicado: (2026)
Ethical Risk Assessment of the Data Harnessing Process of LLM supported on Consensus of Well-known Multi-Ethical Frameworks
por: Khan, Javed I., et al.
Publicado: (2026)
por: Khan, Javed I., et al.
Publicado: (2026)
Quantifying the Persona Effect in LLM Simulations
por: Hu, Tiancheng, et al.
Publicado: (2024)
por: Hu, Tiancheng, et al.
Publicado: (2024)
Ejemplares similares
-
Simulated Adoption: Decoupling Magnitude and Direction in LLM In-Context Conflict Resolution
por: Zhang, Long, et al.
Publicado: (2026) -
Augmenting Human-Annotated Training Data with Large Language Model Generation and Distillation in Open-Response Assessment
por: Borchers, Conrad, et al.
Publicado: (2025) -
Beyond Static Question Banks: Dynamic Knowledge Expansion via LLM-Automated Graph Construction and Adaptive Generation
por: Wang, Yingquan, et al.
Publicado: (2026) -
PersonalityScanner: Exploring the Validity of Personality Assessment Based on Multimodal Signals in Virtual Reality
por: Zhang, Xintong, et al.
Publicado: (2024) -
Comparing Rationality Between Large Language Models and Humans: Insights and Open Questions
por: Alsagheer, Dana, et al.
Publicado: (2024)