Orchestrating LLM Agents for Scientific Research: A Pilot Study of Multiple Choice Question (MCQ) Generation and Evaluation
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | An, Yuan |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
CODE-GEN: A Human-in-the-Loop RAG-Based Agentic AI System for Multiple-Choice Question Generation
par: Duan, Xiaojing, et autres
Publié: (2026)
par: Duan, Xiaojing, et autres
Publié: (2026)
The Basic B*** Effect: The Use of LLM-based Agents Reduces the Distinctiveness and Diversity of People's Choices
par: Matz, Sandra C., et autres
Publié: (2025)
par: Matz, Sandra C., et autres
Publié: (2025)
Assessing the Prevalence of AI-assisted Cheating in Programming Courses: A Pilot Study
par: Delphino, Kaléu
Publié: (2025)
par: Delphino, Kaléu
Publié: (2025)
Pilot Study on Generative AI and Critical Thinking in Higher Education Classrooms
par: Lamberti, W. F., et autres
Publié: (2025)
par: Lamberti, W. F., et autres
Publié: (2025)
Building Better AI Agents: A Provocation on the Utilisation of Persona in LLM-based Conversational Agents
par: Sun, Guangzhi, et autres
Publié: (2024)
par: Sun, Guangzhi, et autres
Publié: (2024)
LegalWebAgent: Empowering Access to Justice via LLM-Based Web Agents
par: Tan, Jinzhe, et autres
Publié: (2025)
par: Tan, Jinzhe, et autres
Publié: (2025)
Responding to Generative AI Technologies with Research-through-Design: The Ryelands AI Lab as an Exploratory Study
par: Benjamin, Jesse Josua, et autres
Publié: (2024)
par: Benjamin, Jesse Josua, et autres
Publié: (2024)
When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis
par: Najera, Aisha, et autres
Publié: (2026)
par: Najera, Aisha, et autres
Publié: (2026)
Examining and Addressing Barriers to Diversity in LLM-Generated Ideas
par: Deng, Yuting, et autres
Publié: (2026)
par: Deng, Yuting, et autres
Publié: (2026)
Exploring User Acceptance and Concerns toward LLM-powered Conversational Agents in Immersive Extended Reality
par: Bozkir, Efe, et autres
Publié: (2025)
par: Bozkir, Efe, et autres
Publié: (2025)
Exploring How Multiple Levels of GPT-Generated Programming Hints Support or Disappoint Novices
par: Xiao, Ruiwei, et autres
Publié: (2024)
par: Xiao, Ruiwei, et autres
Publié: (2024)
Who Defines "Best"? Towards Interactive, User-Defined Evaluation of LLM Leaderboards
par: Jung, Minji, et autres
Publié: (2026)
par: Jung, Minji, et autres
Publié: (2026)
From Pilots to Practices: A Scoping Review of GenAI-Enabled Personalization in Computer Science Education
par: Reihanian, Iman, et autres
Publié: (2025)
par: Reihanian, Iman, et autres
Publié: (2025)
Generating the Modal Worker: A Cross-Model Audit of Race and Gender in LLM-Generated Personas Across 41 Occupations
par: van der Linden, Ilona, et autres
Publié: (2025)
par: van der Linden, Ilona, et autres
Publié: (2025)
Longitudinal Study on Social and Emotional Use of AI Conversational Agent
par: Chandra, Mohit, et autres
Publié: (2025)
par: Chandra, Mohit, et autres
Publié: (2025)
Multi-Agent Comedy Club: Investigating Community Discussion Effects on LLM Humor Generation
par: Hong, Shiwei, et autres
Publié: (2026)
par: Hong, Shiwei, et autres
Publié: (2026)
AI and Consciousness: Shifting Focus Towards Tractable Questions
par: Comsa, Iulia-Maria
Publié: (2026)
par: Comsa, Iulia-Maria
Publié: (2026)
A Study about Distribution and Acceptance of Conversational Agents for Mental Health in Germany: Keep the Human in the Loop?
par: Lukas, Christina
Publié: (2025)
par: Lukas, Christina
Publié: (2025)
LLM Agents for Education: Advances and Applications
par: Chu, Zhendong, et autres
Publié: (2025)
par: Chu, Zhendong, et autres
Publié: (2025)
Who's Asking? Simulating Role-Based Questions for Conversational AI Evaluation
par: Kaur, Navreet, et autres
Publié: (2025)
par: Kaur, Navreet, et autres
Publié: (2025)
Evaluating Contextually Personalized Programming Exercises Created with Generative AI
par: Logacheva, Evanfiya, et autres
Publié: (2024)
par: Logacheva, Evanfiya, et autres
Publié: (2024)
The Impact of Big Five Personality Traits on AI Agent Decision-Making in Public Spaces: A Social Simulation Study
par: Ren, Mingjun, et autres
Publié: (2025)
par: Ren, Mingjun, et autres
Publié: (2025)
How to Capture and Study Conversations Between Research Participants and ChatGPT: GPT for Researchers (g4r.org)
par: Kim, Jin
Publié: (2025)
par: Kim, Jin
Publié: (2025)
Evaluating Generative AI as an Educational Tool for Radiology Resident Report Drafting
par: Verdone, Antonio, et autres
Publié: (2025)
par: Verdone, Antonio, et autres
Publié: (2025)
Comparing LLM-Based Conversational and Graphical Interfaces for Industrial Decision Tasks: An Exploratory Mixed-Methods Study
par: Figliè, Roberto, et autres
Publié: (2026)
par: Figliè, Roberto, et autres
Publié: (2026)
The RIGID Framework: Research-Integrated, Generative AI-Mediated Instructional Design
par: Kwak, Yerin, et autres
Publié: (2026)
par: Kwak, Yerin, et autres
Publié: (2026)
Addressing Bias in Generative AI: Challenges and Research Opportunities in Information Management
par: Wei, Xiahua, et autres
Publié: (2025)
par: Wei, Xiahua, et autres
Publié: (2025)
LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation
par: Ghosh, Himel, et autres
Publié: (2026)
par: Ghosh, Himel, et autres
Publié: (2026)
A Comparative Study of Technical Writing Feedback Quality: Evaluating LLMs, SLMs, and Humans in Computer Science Topics
par: Liu, Suqing, et autres
Publié: (2025)
par: Liu, Suqing, et autres
Publié: (2025)
AgentSUMO: An Agentic Framework for Interactive Simulation Scenario Generation in SUMO via Large Language Models
par: Jeong, Minwoo, et autres
Publié: (2025)
par: Jeong, Minwoo, et autres
Publié: (2025)
Investigating Collaborative Data Practices: a Case Study on Artificial Intelligence for Healthcare Research
par: Henkin, Rafael, et autres
Publié: (2023)
par: Henkin, Rafael, et autres
Publié: (2023)
Exploring Human-AI Collaboration Using Mental Models of Early Adopters of Multi-Agent Generative AI Tools
par: Naik, Suchismita, et autres
Publié: (2025)
par: Naik, Suchismita, et autres
Publié: (2025)
Would You Rely on an Eerie Agent? A Systematic Review of the Impact of the Uncanny Valley Effect on Trust in Human-Agent Interaction
par: Alipour, Ahdiyeh, et autres
Publié: (2025)
par: Alipour, Ahdiyeh, et autres
Publié: (2025)
WaLLM -- Insights from an LLM-Powered Chatbot deployment via WhatsApp
par: Eltigani, Hiba, et autres
Publié: (2025)
par: Eltigani, Hiba, et autres
Publié: (2025)
Exploring Communication Strategies for Collaborative LLM Agents in Mathematical Problem-Solving
par: Zhang, Liang, et autres
Publié: (2025)
par: Zhang, Liang, et autres
Publié: (2025)
Enhancing Selection of Climate Tech Startups with AI -- A Case Study on Integrating Human and AI Evaluations in the ClimaTech Great Global Innovation Challenge
par: Turliuk, Jennifer, et autres
Publié: (2025)
par: Turliuk, Jennifer, et autres
Publié: (2025)
LLM Social Simulations Are a Promising Research Method
par: Anthis, Jacy Reese, et autres
Publié: (2025)
par: Anthis, Jacy Reese, et autres
Publié: (2025)
From Prompts to Constructs: A Dual-Validity Framework for LLM Research in Psychology
par: Lin, Zhicheng
Publié: (2025)
par: Lin, Zhicheng
Publié: (2025)
Social Welfare Function Leaderboard: When LLM Agents Allocate Social Welfare
par: Shi, Zhengliang, et autres
Publié: (2025)
par: Shi, Zhengliang, et autres
Publié: (2025)
A Meta-Analysis of LLM Effects on Students across Qualification, Socialisation, and Subjectification
par: Huang, Jiayu, et autres
Publié: (2025)
par: Huang, Jiayu, et autres
Publié: (2025)
Documents similaires
-
CODE-GEN: A Human-in-the-Loop RAG-Based Agentic AI System for Multiple-Choice Question Generation
par: Duan, Xiaojing, et autres
Publié: (2026) -
The Basic B*** Effect: The Use of LLM-based Agents Reduces the Distinctiveness and Diversity of People's Choices
par: Matz, Sandra C., et autres
Publié: (2025) -
Assessing the Prevalence of AI-assisted Cheating in Programming Courses: A Pilot Study
par: Delphino, Kaléu
Publié: (2025) -
Pilot Study on Generative AI and Critical Thinking in Higher Education Classrooms
par: Lamberti, W. F., et autres
Publié: (2025) -
Building Better AI Agents: A Provocation on the Utilisation of Persona in LLM-based Conversational Agents
par: Sun, Guangzhi, et autres
Publié: (2024)