D-GEN: Automatic Distractor Generation and Evaluation for Reliable Assessment of Generative Model
Fuente:
arXiv
Saved in:
| Main Authors: | Byun, Grace, Choi, Jinho D. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLM-as-a-Grader: Practical Insights from Large Language Model for Short-Answer and Report Evaluation
by: Byun, Grace, et al.
Published: (2025)
by: Byun, Grace, et al.
Published: (2025)
Secure Multifaceted-RAG for Enterprise: Hybrid Knowledge Retrieval with Security Filtering
by: Byun, Grace, et al.
Published: (2025)
by: Byun, Grace, et al.
Published: (2025)
Measuring Sycophancy of Language Models in Multi-turn Dialogues
by: Hong, Jiseung, et al.
Published: (2025)
by: Hong, Jiseung, et al.
Published: (2025)
CRADLE Bench: A Clinician-Annotated Benchmark for Multi-Faceted Mental Health Crisis and Safety Risk Detection
by: Byun, Grace, et al.
Published: (2025)
by: Byun, Grace, et al.
Published: (2025)
CDGP: Automatic Cloze Distractor Generation based on Pre-trained Language Model
by: Chiang, Shang-Hsuan, et al.
Published: (2024)
by: Chiang, Shang-Hsuan, et al.
Published: (2024)
Diverse and Effective Synthetic Data Generation for Adaptable Zero-Shot Dialogue State Tracking
by: Finch, James D., et al.
Published: (2024)
by: Finch, James D., et al.
Published: (2024)
Distractor Generation in Multiple-Choice Tasks: A Survey of Methods, Datasets, and Evaluation
by: Alhazmi, Elaf, et al.
Published: (2024)
by: Alhazmi, Elaf, et al.
Published: (2024)
Transforming Slot Schema Induction with Generative Dialogue State Inference
by: Finch, James D., et al.
Published: (2024)
by: Finch, James D., et al.
Published: (2024)
Difficulty-Controllable Cloze Question Distractor Generation
by: Kang, Seokhoon, et al.
Published: (2025)
by: Kang, Seokhoon, et al.
Published: (2025)
Generative Induction of Dialogue Task Schemas with Streaming Refinement and Simulated Interactions
by: Finch, James D., et al.
Published: (2025)
by: Finch, James D., et al.
Published: (2025)
Finding A Voice: Exploring the Potential of African American Dialect and Voice Generation for Chatbots
by: Finch, Sarah E., et al.
Published: (2025)
by: Finch, Sarah E., et al.
Published: (2025)
What is Your Favorite Gender, MLM? Gender Bias Evaluation in Multilingual Masked Language Models
by: Yu, Jeongrok, et al.
Published: (2024)
by: Yu, Jeongrok, et al.
Published: (2024)
TarGEN: Targeted Data Generation with Large Language Models
by: Gupta, Himanshu, et al.
Published: (2023)
by: Gupta, Himanshu, et al.
Published: (2023)
Calibrating Verbalized Confidence with Self-Generated Distractors
by: Wang, Victor, et al.
Published: (2025)
by: Wang, Victor, et al.
Published: (2025)
Do We Still Need Audio? Rethinking Speaker Diarization with a Text-Based Approach Using Multiple Prediction Models
by: Wu, Peilin, et al.
Published: (2025)
by: Wu, Peilin, et al.
Published: (2025)
Unsupervised Distractor Generation via Large Language Model Distilling and Counterfactual Contrastive Decoding
by: Qu, Fanyi, et al.
Published: (2024)
by: Qu, Fanyi, et al.
Published: (2024)
Leveraging Explicit Reasoning for Inference Integration in Commonsense-Augmented Dialogue Models
by: Finch, Sarah E., et al.
Published: (2024)
by: Finch, Sarah E., et al.
Published: (2024)
ISSR: Iterative Selection with Self-Review for Vocabulary Test Distractor Generation
by: Liu, Yu-Cheng, et al.
Published: (2025)
by: Liu, Yu-Cheng, et al.
Published: (2025)
ETM: Modern Insights into Perspective on Text-to-SQL Evaluation in the Age of Large Language Models
by: Ascoli, Benjamin G., et al.
Published: (2024)
by: Ascoli, Benjamin G., et al.
Published: (2024)
SS-GEN: A Social Story Generation Framework with Large Language Models
by: Feng, Yi, et al.
Published: (2024)
by: Feng, Yi, et al.
Published: (2024)
Exploring Automated Distractor Generation for Math Multiple-choice Questions via Large Language Models
by: Feng, Wanyong, et al.
Published: (2024)
by: Feng, Wanyong, et al.
Published: (2024)
Beyond Fine-Tuning: In-Context Learning and Chain-of-Thought for Reasoned Distractor Generation
by: Alhazmi, Elaf, et al.
Published: (2026)
by: Alhazmi, Elaf, et al.
Published: (2026)
Tailoring Diagnostic Modeling to Individual Learners: Personalized Distractor Generation via MCTS-Guided Reasoning Reconstruction
by: Wu, Tao, et al.
Published: (2025)
by: Wu, Tao, et al.
Published: (2025)
TRUST: An LLM-Based Dialogue System for Trauma Understanding and Structured Assessments
by: Tu, Sichang, et al.
Published: (2025)
by: Tu, Sichang, et al.
Published: (2025)
Are Language Models Sensitive to Morally Irrelevant Distractors?
by: Shaw, Andrew, et al.
Published: (2026)
by: Shaw, Andrew, et al.
Published: (2026)
Can LLMs Model Incorrect Student Reasoning? A Case Study on Distractor Generation
by: Zengaffinen, Yanick, et al.
Published: (2026)
by: Zengaffinen, Yanick, et al.
Published: (2026)
†DAGGER: Distractor-Aware Graph Generation for Executable Reasoning in Math Problems
by: Nazi, Zabir Al, et al.
Published: (2026)
by: Nazi, Zabir Al, et al.
Published: (2026)
Enhancing Clinical Multiple-Choice Questions Benchmarks with Knowledge Graph Guided Distractor Generation
by: Yang, Running, et al.
Published: (2025)
by: Yang, Running, et al.
Published: (2025)
DualReward: A Dynamic Reinforcement Learning Framework for Cloze Tests Distractor Generation
by: Huang, Tianyou, et al.
Published: (2025)
by: Huang, Tianyou, et al.
Published: (2025)
Automated Distractor and Feedback Generation for Math Multiple-choice Questions via In-context Learning
by: McNichols, Hunter, et al.
Published: (2023)
by: McNichols, Hunter, et al.
Published: (2023)
Evaluating Diversity in Automatic Poetry Generation
by: Chen, Yanran, et al.
Published: (2024)
by: Chen, Yanran, et al.
Published: (2024)
Automatic Answerability Evaluation for Question Generation
by: Wang, Zifan, et al.
Published: (2023)
by: Wang, Zifan, et al.
Published: (2023)
ConvoSense: Overcoming Monotonous Commonsense Inferences for Conversational AI
by: Finch, Sarah E., et al.
Published: (2024)
by: Finch, Sarah E., et al.
Published: (2024)
Automating PTSD Diagnostics in Clinical Interviews: Leveraging Large Language Models for Trauma Assessments
by: Tu, Sichang, et al.
Published: (2024)
by: Tu, Sichang, et al.
Published: (2024)
Enhancing Distractor Generation for Multiple-Choice Questions with Retrieval Augmented Pretraining and Knowledge Graph Integration
by: Yu, Han-Cheng, et al.
Published: (2024)
by: Yu, Han-Cheng, et al.
Published: (2024)
DisGeM: Distractor Generation for Multiple Choice Questions with Span Masking
by: Cavusoglu, Devrim, et al.
Published: (2024)
by: Cavusoglu, Devrim, et al.
Published: (2024)
Learning to Correction: Explainable Feedback Generation for Visual Commonsense Reasoning Distractor
by: Chen, Jiali, et al.
Published: (2024)
by: Chen, Jiali, et al.
Published: (2024)
PersonaKit (PK): A Plug-and-Play Platform for User Testing Diverse Roles in Full-Duplex Dialogue
by: Jeon, Hyunbae, et al.
Published: (2026)
by: Jeon, Hyunbae, et al.
Published: (2026)
DYCP: Dynamic Context Pruning for Long-Form Dialogue with LLMs
by: Choi, Nayoung, et al.
Published: (2026)
by: Choi, Nayoung, et al.
Published: (2026)
Evaluation of Automatic Speech Recognition Using Generative Large Language Models
by: Bañeras-Roux, Thibault, et al.
Published: (2026)
by: Bañeras-Roux, Thibault, et al.
Published: (2026)
Similar Items
-
LLM-as-a-Grader: Practical Insights from Large Language Model for Short-Answer and Report Evaluation
by: Byun, Grace, et al.
Published: (2025) -
Secure Multifaceted-RAG for Enterprise: Hybrid Knowledge Retrieval with Security Filtering
by: Byun, Grace, et al.
Published: (2025) -
Measuring Sycophancy of Language Models in Multi-turn Dialogues
by: Hong, Jiseung, et al.
Published: (2025) -
CRADLE Bench: A Clinician-Annotated Benchmark for Multi-Faceted Mental Health Crisis and Safety Risk Detection
by: Byun, Grace, et al.
Published: (2025) -
CDGP: Automatic Cloze Distractor Generation based on Pre-trained Language Model
by: Chiang, Shang-Hsuan, et al.
Published: (2024)