Saved in:
| Main Authors: | Zhang, Qin, Ge, Hao, Chen, Xiaojun, Fang, Meng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2402.17333 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can multiple-choice questions really be useful in detecting the abilities of LLMs?
by: Li, Wangyue, et al.
Published: (2024)
by: Li, Wangyue, et al.
Published: (2024)
TANQ: An open domain dataset of table answered questions
by: Akhtar, Mubashara, et al.
Published: (2024)
by: Akhtar, Mubashara, et al.
Published: (2024)
CMRAG: Co-modality-based visual document retrieval and question answering
by: Chen, Wang, et al.
Published: (2025)
by: Chen, Wang, et al.
Published: (2025)
Query pipeline optimization for cancer patient question answering systems
by: He, Maolin, et al.
Published: (2024)
by: He, Maolin, et al.
Published: (2024)
FusionMind -- Improving question and answering with external context fusion
by: Verma, Shreyas, et al.
Published: (2023)
by: Verma, Shreyas, et al.
Published: (2023)
Agribot: agriculture-specific question answer system
by: Jain, Naman, et al.
Published: (2025)
by: Jain, Naman, et al.
Published: (2025)
Which questions should I answer? Salience Prediction of Inquisitive Questions
by: Wu, Yating, et al.
Published: (2024)
by: Wu, Yating, et al.
Published: (2024)
ACL-Verbatim: hallucination-free question answering for research
by: Recski, Gábor, et al.
Published: (2026)
by: Recski, Gábor, et al.
Published: (2026)
ChatGPT for automated grading of short answer questions in mechanical ventilation
by: Jade, Tejas, et al.
Published: (2025)
by: Jade, Tejas, et al.
Published: (2025)
keqing: knowledge-based question answering is a nature chain-of-thought mentor of LLM
by: Wang, Chaojie, et al.
Published: (2023)
by: Wang, Chaojie, et al.
Published: (2023)
RealMedQA: A pilot biomedical question answering dataset containing realistic clinical questions
by: Kell, Gregory, et al.
Published: (2024)
by: Kell, Gregory, et al.
Published: (2024)
70B-parameter large language models in Japanese medical question-answering
by: Sukeda, Issey, et al.
Published: (2024)
by: Sukeda, Issey, et al.
Published: (2024)
ConSens: Assessing context grounding in open-book question answering
by: Vankov, Ivan, et al.
Published: (2025)
by: Vankov, Ivan, et al.
Published: (2025)
Large language models provide unsafe answers to patient-posed medical questions
by: Draelos, Rachel L., et al.
Published: (2025)
by: Draelos, Rachel L., et al.
Published: (2025)
Retrieval augmented text-to-SQL generation for epidemiological question answering using electronic health records
by: Ziletti, Angelo, et al.
Published: (2024)
by: Ziletti, Angelo, et al.
Published: (2024)
Enhancing textual textbook question answering with large language models and retrieval augmented generation
by: Alawwad, Hessa Abdulrahman, et al.
Published: (2024)
by: Alawwad, Hessa Abdulrahman, et al.
Published: (2024)
From text to multimodal: a survey of adversarial example generation in question answering systems
by: Yigit, Gulsum, et al.
Published: (2023)
by: Yigit, Gulsum, et al.
Published: (2023)
Few shot chain-of-thought driven reasoning to prompt LLMs for open ended medical question answering
by: Nachane, Saeel Sandeep, et al.
Published: (2024)
by: Nachane, Saeel Sandeep, et al.
Published: (2024)
KeyKnowledgeRAG (K^2RAG): An Enhanced RAG method for improved LLM question-answering capabilities
by: Markondapatnaikuni, Hruday, et al.
Published: (2025)
by: Markondapatnaikuni, Hruday, et al.
Published: (2025)
CaLMQA: Exploring culturally specific long-form question answering across 23 languages
by: Arora, Shane, et al.
Published: (2024)
by: Arora, Shane, et al.
Published: (2024)
What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering
by: Maar, Jim, et al.
Published: (2026)
by: Maar, Jim, et al.
Published: (2026)
Multi-step retrieval and reasoning improves radiology question answering with large language models
by: Wind, Sebastian, et al.
Published: (2025)
by: Wind, Sebastian, et al.
Published: (2025)
Reinforcement learning for question answering in programming domain using public community scoring as a human feedback
by: Gorbatovski, Alexey, et al.
Published: (2024)
by: Gorbatovski, Alexey, et al.
Published: (2024)
Inter-Passage Verification for Multi-evidence Multi-answer QA
by: Chen, Bingsen, et al.
Published: (2025)
by: Chen, Bingsen, et al.
Published: (2025)
Can we repurpose multiple-choice question-answering models to rerank retrieved documents?
by: Catapang, Jasper Kyle
Published: (2025)
by: Catapang, Jasper Kyle
Published: (2025)
Overview of the MedHopQA track at BioCreative IX: track description, participation and evaluation of systems for multi-hop medical question answering
by: Islamaj, Rezarta, et al.
Published: (2026)
by: Islamaj, Rezarta, et al.
Published: (2026)
Learning the meanings of function words from grounded language using a visual question answering model
by: Portelance, Eva, et al.
Published: (2023)
by: Portelance, Eva, et al.
Published: (2023)
DCR: Divide-and-Conquer Reasoning for Multi-choice Question Answering with LLMs
by: Meng, Zijie, et al.
Published: (2024)
by: Meng, Zijie, et al.
Published: (2024)
C-Mining: Unsupervised Discovery of Seeds for Cultural Data Synthesis via Geometric Misalignment
by: Zeng, Pufan, et al.
Published: (2026)
by: Zeng, Pufan, et al.
Published: (2026)
OpenMedLM: Prompt engineering can out-perform fine-tuning in medical question-answering with open-source large language models
by: Maharjan, Jenish, et al.
Published: (2024)
by: Maharjan, Jenish, et al.
Published: (2024)
Predictions from language models for multiple-choice tasks are not robust under variation of scoring methods
by: Tsvilodub, Polina, et al.
Published: (2024)
by: Tsvilodub, Polina, et al.
Published: (2024)
CrossICL: Cross-Task In-Context Learning via Unsupervised Demonstration Transfer
by: Gao, Jinglong, et al.
Published: (2025)
by: Gao, Jinglong, et al.
Published: (2025)
Jochre 3 and the Yiddish OCR corpus
by: Urieli, Assaf, et al.
Published: (2025)
by: Urieli, Assaf, et al.
Published: (2025)
Robustness assessment of large audio language models in multiple-choice evaluation
by: López, Fernando, et al.
Published: (2025)
by: López, Fernando, et al.
Published: (2025)
Question answering system of bridge design specification based on large language model
by: Zhang, Leye, et al.
Published: (2024)
by: Zhang, Leye, et al.
Published: (2024)
Detecting (Un)answerability in Large Language Models with Linear Directions
by: Lavi, Maor Juliet, et al.
Published: (2025)
by: Lavi, Maor Juliet, et al.
Published: (2025)
Where is the answer? Investigating Positional Bias in Language Model Knowledge Extraction
by: Saito, Kuniaki, et al.
Published: (2024)
by: Saito, Kuniaki, et al.
Published: (2024)
Unsupervised Distractor Generation via Large Language Model Distilling and Counterfactual Contrastive Decoding
by: Qu, Fanyi, et al.
Published: (2024)
by: Qu, Fanyi, et al.
Published: (2024)
A Report on the llms evaluating the high school questions
by: Jiawei, Zhu, et al.
Published: (2025)
by: Jiawei, Zhu, et al.
Published: (2025)
Large Language Models for Predictive Analysis: How Far Are They?
by: Chen, Qin, et al.
Published: (2025)
by: Chen, Qin, et al.
Published: (2025)
Similar Items
-
Can multiple-choice questions really be useful in detecting the abilities of LLMs?
by: Li, Wangyue, et al.
Published: (2024) -
TANQ: An open domain dataset of table answered questions
by: Akhtar, Mubashara, et al.
Published: (2024) -
CMRAG: Co-modality-based visual document retrieval and question answering
by: Chen, Wang, et al.
Published: (2025) -
Query pipeline optimization for cancer patient question answering systems
by: He, Maolin, et al.
Published: (2024) -
FusionMind -- Improving question and answering with external context fusion
by: Verma, Shreyas, et al.
Published: (2023)