Anchored Answers: Unravelling Positional Bias in GPT-2's Multiple-Choice Questions
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Li, Ruizhe, Gao, Yanjun |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Answer Matching Outperforms Multiple Choice for Language Model Evaluation
par: Chandak, Nikhil, et autres
Publié: (2025)
par: Chandak, Nikhil, et autres
Publié: (2025)
Generating Plausible Distractors for Multiple-Choice Questions via Student Choice Prediction
par: Lee, Yooseop, et autres
Publié: (2025)
par: Lee, Yooseop, et autres
Publié: (2025)
Biomedical Entity Linking as Multiple Choice Question Answering
par: Lin, Zhenxi, et autres
Publié: (2024)
par: Lin, Zhenxi, et autres
Publié: (2024)
Option-ID Based Elimination For Multiple Choice Questions
par: Zhu, Zhenhao, et autres
Publié: (2025)
par: Zhu, Zhenhao, et autres
Publié: (2025)
Pattern Recognition or Medical Knowledge? The Problem with Multiple-Choice Questions in Medicine
par: Griot, Maxime, et autres
Publié: (2024)
par: Griot, Maxime, et autres
Publié: (2024)
When Answers Stray from Questions: Hallucination Detection via Question-Answer Orthogonal Decomposition
par: Yao, Siyang, et autres
Publié: (2026)
par: Yao, Siyang, et autres
Publié: (2026)
A Study on Large Language Models' Limitations in Multiple-Choice Question Answering
par: Khatun, Aisha, et autres
Publié: (2024)
par: Khatun, Aisha, et autres
Publié: (2024)
Attributing Response to Context: A Jensen-Shannon Divergence Driven Mechanistic Study of Context Attribution in Retrieval-Augmented Generation
par: Li, Ruizhe, et autres
Publié: (2025)
par: Li, Ruizhe, et autres
Publié: (2025)
Answer, Assemble, Ace: Understanding How LMs Answer Multiple Choice Questions
par: Wiegreffe, Sarah, et autres
Publié: (2024)
par: Wiegreffe, Sarah, et autres
Publié: (2024)
Race, Ethnicity and Their Implication on Bias in Large Language Models
par: Hu, Shiyue, et autres
Publié: (2026)
par: Hu, Shiyue, et autres
Publié: (2026)
UnibucLLM: Harnessing LLMs for Automated Prediction of Item Difficulty and Response Time for Multiple-Choice Questions
par: Rogoz, Ana-Cristina, et autres
Publié: (2024)
par: Rogoz, Ana-Cristina, et autres
Publié: (2024)
Generating Multiple-Choice Knowledge Questions with Interpretable Difficulty Estimation using Knowledge Graphs and Large Language Models
par: Şakiroğlu, Mehmet Can, et autres
Publié: (2026)
par: Şakiroğlu, Mehmet Can, et autres
Publié: (2026)
Explicit Diversity Conditions for Effective Question Answer Generation with Large Language Models
par: Yadav, Vikas, et autres
Publié: (2024)
par: Yadav, Vikas, et autres
Publié: (2024)
The Challenge of Achieving Attributability in Multilingual Table-to-Text Generation with Question-Answer Blueprints
par: Haussmann, Aden
Publié: (2025)
par: Haussmann, Aden
Publié: (2025)
Token Homogenization under Positional Bias
par: Yusupov, Viacheslav, et autres
Publié: (2025)
par: Yusupov, Viacheslav, et autres
Publié: (2025)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
par: Xu, Xin, et autres
Publié: (2025)
par: Xu, Xin, et autres
Publié: (2025)
DLP-LoRA: Efficient Task-Specific LoRA Fusion with a Dynamic, Lightweight Plugin for Large Language Models
par: Zhang, Yuxuan, et autres
Publié: (2024)
par: Zhang, Yuxuan, et autres
Publié: (2024)
MultiQ&A: An Analysis in Measuring Robustness via Automated Crowdsourcing of Question Perturbations and Answers
par: Cho, Nicole, et autres
Publié: (2025)
par: Cho, Nicole, et autres
Publié: (2025)
MeDiSumQA: Patient-Oriented Question-Answer Generation from Discharge Letters
par: Dada, Amin, et autres
Publié: (2025)
par: Dada, Amin, et autres
Publié: (2025)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
par: Ma, Mingyu Derek, et autres
Publié: (2023)
par: Ma, Mingyu Derek, et autres
Publié: (2023)
More Thinking, More Bias: Length-Driven Position Bias in Reasoning Models
par: Wang, Xiao
Publié: (2026)
par: Wang, Xiao
Publié: (2026)
Multiple Choice Learning of Low-Rank Adapters for Language Modeling
par: Letzelter, Victor, et autres
Publié: (2025)
par: Letzelter, Victor, et autres
Publié: (2025)
Position-Aware Parameter Efficient Fine-Tuning Approach for Reducing Positional Bias in LLMs
par: Zhang, Zheng, et autres
Publié: (2024)
par: Zhang, Zheng, et autres
Publié: (2024)
Multilingual Non-Factoid Question Answering with Answer Paragraph Selection
par: Mishra, Ritwik, et autres
Publié: (2024)
par: Mishra, Ritwik, et autres
Publié: (2024)
How Prevalent is Gender Bias in ChatGPT? -- Exploring German and English ChatGPT Responses
par: Urchs, Stefanie, et autres
Publié: (2023)
par: Urchs, Stefanie, et autres
Publié: (2023)
Automated Generation of Challenging Multiple-Choice Questions for Vision Language Model Evaluation
par: Zhang, Yuhui, et autres
Publié: (2025)
par: Zhang, Yuhui, et autres
Publié: (2025)
Lost in the Middle at Birth: An Exact Theory of Transformer Position Bias
par: Chowdhury, Borun D
Publié: (2026)
par: Chowdhury, Borun D
Publié: (2026)
Found in the Middle: Calibrating Positional Attention Bias Improves Long Context Utilization
par: Hsieh, Cheng-Yu, et autres
Publié: (2024)
par: Hsieh, Cheng-Yu, et autres
Publié: (2024)
Probabilities of Chat LLMs Are Miscalibrated but Still Predict Correctness on Multiple-Choice Q&A
par: Plaut, Benjamin, et autres
Publié: (2024)
par: Plaut, Benjamin, et autres
Publié: (2024)
Can GPT Improve the State of Prior Authorization via Guideline Based Automated Question Answering?
par: Vatsal, Shubham, et autres
Publié: (2024)
par: Vatsal, Shubham, et autres
Publié: (2024)
Assessing Modality Bias in Video Question Answering Benchmarks with Multimodal Large Language Models
par: Park, Jean, et autres
Publié: (2024)
par: Park, Jean, et autres
Publié: (2024)
Diagnosing Retrieval Bias Under Multiple In-Context Knowledge Updates in Large Language Models
par: Qiao, Boyu, et autres
Publié: (2026)
par: Qiao, Boyu, et autres
Publié: (2026)
Single layer tiny Co$^4$ outpaces GPT-2 and GPT-BERT
par: Zain, Noor Ul, et autres
Publié: (2025)
par: Zain, Noor Ul, et autres
Publié: (2025)
Unraveling the cognitive patterns of Large Language Models through module communities
par: Bhandari, Kushal Raj, et autres
Publié: (2025)
par: Bhandari, Kushal Raj, et autres
Publié: (2025)
Clinical QA 2.0: Multi-Task Learning for Answer Extraction and Categorization
par: Pattnayak, Priyaranjan, et autres
Publié: (2025)
par: Pattnayak, Priyaranjan, et autres
Publié: (2025)
Listen Again and Choose the Right Answer: A New Paradigm for Automatic Speech Recognition with Large Language Models
par: Hu, Yuchen, et autres
Publié: (2024)
par: Hu, Yuchen, et autres
Publié: (2024)
Proving that Cryptic Crossword Clue Answers are Correct
par: Andrews, Martin, et autres
Publié: (2024)
par: Andrews, Martin, et autres
Publié: (2024)
Verif.ai: Towards an Open-Source Scientific Generative Question-Answering System with Referenced and Verifiable Answers
par: Košprdić, Miloš, et autres
Publié: (2024)
par: Košprdić, Miloš, et autres
Publié: (2024)
Inference-Time Intervention: Eliciting Truthful Answers from a Language Model
par: Li, Kenneth, et autres
Publié: (2023)
par: Li, Kenneth, et autres
Publié: (2023)
Universal Neurons in GPT2 Language Models
par: Gurnee, Wes, et autres
Publié: (2024)
par: Gurnee, Wes, et autres
Publié: (2024)
Documents similaires
-
Answer Matching Outperforms Multiple Choice for Language Model Evaluation
par: Chandak, Nikhil, et autres
Publié: (2025) -
Generating Plausible Distractors for Multiple-Choice Questions via Student Choice Prediction
par: Lee, Yooseop, et autres
Publié: (2025) -
Biomedical Entity Linking as Multiple Choice Question Answering
par: Lin, Zhenxi, et autres
Publié: (2024) -
Option-ID Based Elimination For Multiple Choice Questions
par: Zhu, Zhenhao, et autres
Publié: (2025) -
Pattern Recognition or Medical Knowledge? The Problem with Multiple-Choice Questions in Medicine
par: Griot, Maxime, et autres
Publié: (2024)