Orchestrating LLM Agents for Scientific Research: A Pilot Study of Multiple Choice Question (MCQ) Generation and Evaluation
Fuente:
arXiv
Guardado en:
| Autor principal: | An, Yuan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CODE-GEN: A Human-in-the-Loop RAG-Based Agentic AI System for Multiple-Choice Question Generation
por: Duan, Xiaojing, et al.
Publicado: (2026)
por: Duan, Xiaojing, et al.
Publicado: (2026)
The Basic B*** Effect: The Use of LLM-based Agents Reduces the Distinctiveness and Diversity of People's Choices
por: Matz, Sandra C., et al.
Publicado: (2025)
por: Matz, Sandra C., et al.
Publicado: (2025)
Assessing the Prevalence of AI-assisted Cheating in Programming Courses: A Pilot Study
por: Delphino, Kaléu
Publicado: (2025)
por: Delphino, Kaléu
Publicado: (2025)
Pilot Study on Generative AI and Critical Thinking in Higher Education Classrooms
por: Lamberti, W. F., et al.
Publicado: (2025)
por: Lamberti, W. F., et al.
Publicado: (2025)
Building Better AI Agents: A Provocation on the Utilisation of Persona in LLM-based Conversational Agents
por: Sun, Guangzhi, et al.
Publicado: (2024)
por: Sun, Guangzhi, et al.
Publicado: (2024)
LegalWebAgent: Empowering Access to Justice via LLM-Based Web Agents
por: Tan, Jinzhe, et al.
Publicado: (2025)
por: Tan, Jinzhe, et al.
Publicado: (2025)
Responding to Generative AI Technologies with Research-through-Design: The Ryelands AI Lab as an Exploratory Study
por: Benjamin, Jesse Josua, et al.
Publicado: (2024)
por: Benjamin, Jesse Josua, et al.
Publicado: (2024)
When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis
por: Najera, Aisha, et al.
Publicado: (2026)
por: Najera, Aisha, et al.
Publicado: (2026)
Examining and Addressing Barriers to Diversity in LLM-Generated Ideas
por: Deng, Yuting, et al.
Publicado: (2026)
por: Deng, Yuting, et al.
Publicado: (2026)
Exploring User Acceptance and Concerns toward LLM-powered Conversational Agents in Immersive Extended Reality
por: Bozkir, Efe, et al.
Publicado: (2025)
por: Bozkir, Efe, et al.
Publicado: (2025)
Exploring How Multiple Levels of GPT-Generated Programming Hints Support or Disappoint Novices
por: Xiao, Ruiwei, et al.
Publicado: (2024)
por: Xiao, Ruiwei, et al.
Publicado: (2024)
Who Defines "Best"? Towards Interactive, User-Defined Evaluation of LLM Leaderboards
por: Jung, Minji, et al.
Publicado: (2026)
por: Jung, Minji, et al.
Publicado: (2026)
From Pilots to Practices: A Scoping Review of GenAI-Enabled Personalization in Computer Science Education
por: Reihanian, Iman, et al.
Publicado: (2025)
por: Reihanian, Iman, et al.
Publicado: (2025)
Generating the Modal Worker: A Cross-Model Audit of Race and Gender in LLM-Generated Personas Across 41 Occupations
por: van der Linden, Ilona, et al.
Publicado: (2025)
por: van der Linden, Ilona, et al.
Publicado: (2025)
Longitudinal Study on Social and Emotional Use of AI Conversational Agent
por: Chandra, Mohit, et al.
Publicado: (2025)
por: Chandra, Mohit, et al.
Publicado: (2025)
Multi-Agent Comedy Club: Investigating Community Discussion Effects on LLM Humor Generation
por: Hong, Shiwei, et al.
Publicado: (2026)
por: Hong, Shiwei, et al.
Publicado: (2026)
AI and Consciousness: Shifting Focus Towards Tractable Questions
por: Comsa, Iulia-Maria
Publicado: (2026)
por: Comsa, Iulia-Maria
Publicado: (2026)
A Study about Distribution and Acceptance of Conversational Agents for Mental Health in Germany: Keep the Human in the Loop?
por: Lukas, Christina
Publicado: (2025)
por: Lukas, Christina
Publicado: (2025)
LLM Agents for Education: Advances and Applications
por: Chu, Zhendong, et al.
Publicado: (2025)
por: Chu, Zhendong, et al.
Publicado: (2025)
Who's Asking? Simulating Role-Based Questions for Conversational AI Evaluation
por: Kaur, Navreet, et al.
Publicado: (2025)
por: Kaur, Navreet, et al.
Publicado: (2025)
Evaluating Contextually Personalized Programming Exercises Created with Generative AI
por: Logacheva, Evanfiya, et al.
Publicado: (2024)
por: Logacheva, Evanfiya, et al.
Publicado: (2024)
The Impact of Big Five Personality Traits on AI Agent Decision-Making in Public Spaces: A Social Simulation Study
por: Ren, Mingjun, et al.
Publicado: (2025)
por: Ren, Mingjun, et al.
Publicado: (2025)
How to Capture and Study Conversations Between Research Participants and ChatGPT: GPT for Researchers (g4r.org)
por: Kim, Jin
Publicado: (2025)
por: Kim, Jin
Publicado: (2025)
Evaluating Generative AI as an Educational Tool for Radiology Resident Report Drafting
por: Verdone, Antonio, et al.
Publicado: (2025)
por: Verdone, Antonio, et al.
Publicado: (2025)
Comparing LLM-Based Conversational and Graphical Interfaces for Industrial Decision Tasks: An Exploratory Mixed-Methods Study
por: Figliè, Roberto, et al.
Publicado: (2026)
por: Figliè, Roberto, et al.
Publicado: (2026)
The RIGID Framework: Research-Integrated, Generative AI-Mediated Instructional Design
por: Kwak, Yerin, et al.
Publicado: (2026)
por: Kwak, Yerin, et al.
Publicado: (2026)
Addressing Bias in Generative AI: Challenges and Research Opportunities in Information Management
por: Wei, Xiahua, et al.
Publicado: (2025)
por: Wei, Xiahua, et al.
Publicado: (2025)
LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation
por: Ghosh, Himel, et al.
Publicado: (2026)
por: Ghosh, Himel, et al.
Publicado: (2026)
A Comparative Study of Technical Writing Feedback Quality: Evaluating LLMs, SLMs, and Humans in Computer Science Topics
por: Liu, Suqing, et al.
Publicado: (2025)
por: Liu, Suqing, et al.
Publicado: (2025)
AgentSUMO: An Agentic Framework for Interactive Simulation Scenario Generation in SUMO via Large Language Models
por: Jeong, Minwoo, et al.
Publicado: (2025)
por: Jeong, Minwoo, et al.
Publicado: (2025)
Investigating Collaborative Data Practices: a Case Study on Artificial Intelligence for Healthcare Research
por: Henkin, Rafael, et al.
Publicado: (2023)
por: Henkin, Rafael, et al.
Publicado: (2023)
Exploring Human-AI Collaboration Using Mental Models of Early Adopters of Multi-Agent Generative AI Tools
por: Naik, Suchismita, et al.
Publicado: (2025)
por: Naik, Suchismita, et al.
Publicado: (2025)
Would You Rely on an Eerie Agent? A Systematic Review of the Impact of the Uncanny Valley Effect on Trust in Human-Agent Interaction
por: Alipour, Ahdiyeh, et al.
Publicado: (2025)
por: Alipour, Ahdiyeh, et al.
Publicado: (2025)
WaLLM -- Insights from an LLM-Powered Chatbot deployment via WhatsApp
por: Eltigani, Hiba, et al.
Publicado: (2025)
por: Eltigani, Hiba, et al.
Publicado: (2025)
Exploring Communication Strategies for Collaborative LLM Agents in Mathematical Problem-Solving
por: Zhang, Liang, et al.
Publicado: (2025)
por: Zhang, Liang, et al.
Publicado: (2025)
Enhancing Selection of Climate Tech Startups with AI -- A Case Study on Integrating Human and AI Evaluations in the ClimaTech Great Global Innovation Challenge
por: Turliuk, Jennifer, et al.
Publicado: (2025)
por: Turliuk, Jennifer, et al.
Publicado: (2025)
LLM Social Simulations Are a Promising Research Method
por: Anthis, Jacy Reese, et al.
Publicado: (2025)
por: Anthis, Jacy Reese, et al.
Publicado: (2025)
From Prompts to Constructs: A Dual-Validity Framework for LLM Research in Psychology
por: Lin, Zhicheng
Publicado: (2025)
por: Lin, Zhicheng
Publicado: (2025)
Social Welfare Function Leaderboard: When LLM Agents Allocate Social Welfare
por: Shi, Zhengliang, et al.
Publicado: (2025)
por: Shi, Zhengliang, et al.
Publicado: (2025)
A Meta-Analysis of LLM Effects on Students across Qualification, Socialisation, and Subjectification
por: Huang, Jiayu, et al.
Publicado: (2025)
por: Huang, Jiayu, et al.
Publicado: (2025)
Ejemplares similares
-
CODE-GEN: A Human-in-the-Loop RAG-Based Agentic AI System for Multiple-Choice Question Generation
por: Duan, Xiaojing, et al.
Publicado: (2026) -
The Basic B*** Effect: The Use of LLM-based Agents Reduces the Distinctiveness and Diversity of People's Choices
por: Matz, Sandra C., et al.
Publicado: (2025) -
Assessing the Prevalence of AI-assisted Cheating in Programming Courses: A Pilot Study
por: Delphino, Kaléu
Publicado: (2025) -
Pilot Study on Generative AI and Critical Thinking in Higher Education Classrooms
por: Lamberti, W. F., et al.
Publicado: (2025) -
Building Better AI Agents: A Provocation on the Utilisation of Persona in LLM-based Conversational Agents
por: Sun, Guangzhi, et al.
Publicado: (2024)