Reference-based Metrics Disprove Themselves in Question Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Nguyen, Bang, Yu, Mengxia, Huang, Yun, Jiang, Meng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
QG-SMS: Enhancing Test Item Analysis via Student Modeling and Simulation
di: Nguyen, Bang, et al.
Pubblicazione: (2025)
di: Nguyen, Bang, et al.
Pubblicazione: (2025)
Rephrase and Respond: Let Large Language Models Ask Better Questions for Themselves
di: Deng, Yihe, et al.
Pubblicazione: (2023)
di: Deng, Yihe, et al.
Pubblicazione: (2023)
Context Selection and Rewriting for Video-based Educational Question Generation
di: Yu, Mengxia, et al.
Pubblicazione: (2025)
di: Yu, Mengxia, et al.
Pubblicazione: (2025)
When Models Examine Themselves: Vocabulary-Activation Correspondence in Self-Referential Processing
di: Dadfar, Zachary Pedram
Pubblicazione: (2026)
di: Dadfar, Zachary Pedram
Pubblicazione: (2026)
Quiet-STaR: Language Models Can Teach Themselves to Think Before Speaking
di: Zelikman, Eric, et al.
Pubblicazione: (2024)
di: Zelikman, Eric, et al.
Pubblicazione: (2024)
QOG:Question and Options Generation based on Language Model
di: Zhou, Jincheng
Pubblicazione: (2024)
di: Zhou, Jincheng
Pubblicazione: (2024)
Can Large Language Models Still Explain Themselves? Investigating the Impact of Quantization on Self-Explanations
di: Wang, Qianli, et al.
Pubblicazione: (2026)
di: Wang, Qianli, et al.
Pubblicazione: (2026)
Graph-based Uncertainty Metrics for Long-form Language Model Outputs
di: Jiang, Mingjian, et al.
Pubblicazione: (2024)
di: Jiang, Mingjian, et al.
Pubblicazione: (2024)
Synthetic Multimodal Question Generation
di: Wu, Ian, et al.
Pubblicazione: (2024)
di: Wu, Ian, et al.
Pubblicazione: (2024)
Clinically Grounded Agent-based Report Evaluation: An Interpretable Metric for Radiology Report Generation
di: Dua, Radhika, et al.
Pubblicazione: (2025)
di: Dua, Radhika, et al.
Pubblicazione: (2025)
Beyond Independent Passages: Adaptive Passage Combination Retrieval for Retrieval Augmented Open-Domain Question Answering
di: Ko, Ting-Wen, et al.
Pubblicazione: (2025)
di: Ko, Ting-Wen, et al.
Pubblicazione: (2025)
Towards Verifiable Text Generation with Symbolic References
di: Hennigen, Lucas Torroba, et al.
Pubblicazione: (2023)
di: Hennigen, Lucas Torroba, et al.
Pubblicazione: (2023)
Rhetorical Questions in LLM Representations: A Linear Probing Study
di: Yao, Louie Hong, et al.
Pubblicazione: (2026)
di: Yao, Louie Hong, et al.
Pubblicazione: (2026)
RadioRAG: Online Retrieval-augmented Generation for Radiology Question Answering
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2024)
di: Arasteh, Soroosh Tayebi, et al.
Pubblicazione: (2024)
Machine Unlearning in Generative AI: A Survey
di: Liu, Zheyuan, et al.
Pubblicazione: (2024)
di: Liu, Zheyuan, et al.
Pubblicazione: (2024)
$C^2$: Scalable Auto-Feedback for LLM-based Chart Generation
di: Koh, Woosung, et al.
Pubblicazione: (2024)
di: Koh, Woosung, et al.
Pubblicazione: (2024)
Beyond LLM-as-a-Judge: Deterministic Metrics for Multilingual Generative Text Evaluation
di: Alam, Firoj, et al.
Pubblicazione: (2026)
di: Alam, Firoj, et al.
Pubblicazione: (2026)
Is Conformal Factuality for RAG-based LLMs Robust? Novel Metrics and Systematic Insights
di: Chen, Yi, et al.
Pubblicazione: (2026)
di: Chen, Yi, et al.
Pubblicazione: (2026)
An Examination of the Robustness of Reference-Free Image Captioning Evaluation Metrics
di: Ahmadi, Saba, et al.
Pubblicazione: (2023)
di: Ahmadi, Saba, et al.
Pubblicazione: (2023)
Explicit Diversity Conditions for Effective Question Answer Generation with Large Language Models
di: Yadav, Vikas, et al.
Pubblicazione: (2024)
di: Yadav, Vikas, et al.
Pubblicazione: (2024)
Generating Plausible Distractors for Multiple-Choice Questions via Student Choice Prediction
di: Lee, Yooseop, et al.
Pubblicazione: (2025)
di: Lee, Yooseop, et al.
Pubblicazione: (2025)
The Challenge of Achieving Attributability in Multilingual Table-to-Text Generation with Question-Answer Blueprints
di: Haussmann, Aden
Pubblicazione: (2025)
di: Haussmann, Aden
Pubblicazione: (2025)
Evaluating Evaluation Metrics -- The Mirage of Hallucination Detection
di: Kulkarni, Atharva, et al.
Pubblicazione: (2025)
di: Kulkarni, Atharva, et al.
Pubblicazione: (2025)
Towards Enriched Controllability for Educational Question Generation
di: Leite, Bernardo, et al.
Pubblicazione: (2023)
di: Leite, Bernardo, et al.
Pubblicazione: (2023)
A Dataset for Evaluating LLM-based Evaluation Functions for Research Question Extraction Task
di: Fujisaki, Yuya, et al.
Pubblicazione: (2024)
di: Fujisaki, Yuya, et al.
Pubblicazione: (2024)
ReALM: Reference Resolution As Language Modeling
di: Moniz, Joel Ruben Antony, et al.
Pubblicazione: (2024)
di: Moniz, Joel Ruben Antony, et al.
Pubblicazione: (2024)
TopBench: A Benchmark for Implicit Prediction and Reasoning over Tabular Question Answering
di: Ji, An-Yang, et al.
Pubblicazione: (2026)
di: Ji, An-Yang, et al.
Pubblicazione: (2026)
Large Language Models in Fire Engineering: An Examination of Technical Questions Against Domain Knowledge
di: Hostetter, Haley, et al.
Pubblicazione: (2024)
di: Hostetter, Haley, et al.
Pubblicazione: (2024)
MeDiSumQA: Patient-Oriented Question-Answer Generation from Discharge Letters
di: Dada, Amin, et al.
Pubblicazione: (2025)
di: Dada, Amin, et al.
Pubblicazione: (2025)
Listen to the Context: Towards Faithful Large Language Models for Retrieval Augmented Generation on Climate Questions
di: Thulke, David, et al.
Pubblicazione: (2025)
di: Thulke, David, et al.
Pubblicazione: (2025)
On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization
di: Singh, Janvijay, et al.
Pubblicazione: (2025)
di: Singh, Janvijay, et al.
Pubblicazione: (2025)
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
di: Anugraha, David, et al.
Pubblicazione: (2024)
di: Anugraha, David, et al.
Pubblicazione: (2024)
Multi-hop Question Answering under Temporal Knowledge Editing
di: Cheng, Keyuan, et al.
Pubblicazione: (2024)
di: Cheng, Keyuan, et al.
Pubblicazione: (2024)
Deep Learning Approaches for Improving Question Answering Systems in Hepatocellular Carcinoma Research
di: Huo, Shuning, et al.
Pubblicazione: (2024)
di: Huo, Shuning, et al.
Pubblicazione: (2024)
When Answers Stray from Questions: Hallucination Detection via Question-Answer Orthogonal Decomposition
di: Yao, Siyang, et al.
Pubblicazione: (2026)
di: Yao, Siyang, et al.
Pubblicazione: (2026)
AutoLibra: Agent Metric Induction from Open-Ended Human Feedback
di: Zhu, Hao, et al.
Pubblicazione: (2025)
di: Zhu, Hao, et al.
Pubblicazione: (2025)
QuIM-RAG: Advancing Retrieval-Augmented Generation with Inverted Question Matching for Enhanced QA Performance
di: Saha, Binita, et al.
Pubblicazione: (2025)
di: Saha, Binita, et al.
Pubblicazione: (2025)
Transformer Circuit Faithfulness Metrics are not Robust
di: Miller, Joseph, et al.
Pubblicazione: (2024)
di: Miller, Joseph, et al.
Pubblicazione: (2024)
To be Continuous, or to be Discrete, Those are Bits of Questions
di: Wang, Yiran, et al.
Pubblicazione: (2024)
di: Wang, Yiran, et al.
Pubblicazione: (2024)
Generating Multiple-Choice Knowledge Questions with Interpretable Difficulty Estimation using Knowledge Graphs and Large Language Models
di: Şakiroğlu, Mehmet Can, et al.
Pubblicazione: (2026)
di: Şakiroğlu, Mehmet Can, et al.
Pubblicazione: (2026)
Documenti analoghi
-
QG-SMS: Enhancing Test Item Analysis via Student Modeling and Simulation
di: Nguyen, Bang, et al.
Pubblicazione: (2025) -
Rephrase and Respond: Let Large Language Models Ask Better Questions for Themselves
di: Deng, Yihe, et al.
Pubblicazione: (2023) -
Context Selection and Rewriting for Video-based Educational Question Generation
di: Yu, Mengxia, et al.
Pubblicazione: (2025) -
When Models Examine Themselves: Vocabulary-Activation Correspondence in Self-Referential Processing
di: Dadfar, Zachary Pedram
Pubblicazione: (2026) -
Quiet-STaR: Language Models Can Teach Themselves to Think Before Speaking
di: Zelikman, Eric, et al.
Pubblicazione: (2024)