Do LLMs and Humans Find the Same Questions Difficult? A Case Study on Japanese Quiz Answering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sugiura, Naoya, Yamada, Kosuke, Ogawa, Yasuhiro, Toyama, Katsuhiko, Sasano, Ryohei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Ruri: Japanese General Text Embeddings
von: Tsukagoshi, Hayato, et al.
Veröffentlicht: (2024)
von: Tsukagoshi, Hayato, et al.
Veröffentlicht: (2024)
To Drop or Not to Drop? Predicting Argument Ellipsis Judgments: A Case Study in Japanese
von: Ishizuki, Yukiko, et al.
Veröffentlicht: (2024)
von: Ishizuki, Yukiko, et al.
Veröffentlicht: (2024)
How Do Language Models Acquire Character-Level Information?
von: Sato, Soma, et al.
Veröffentlicht: (2026)
von: Sato, Soma, et al.
Veröffentlicht: (2026)
FrameEOL: Semantic Frame Induction using Causal Language Models
von: Yano, Chihiro, et al.
Veröffentlicht: (2025)
von: Yano, Chihiro, et al.
Veröffentlicht: (2025)
Are Social Sentiments Inherent in LLMs? An Empirical Study on Extraction of Inter-demographic Sentiments
von: Tanaka, Kunitomo, et al.
Veröffentlicht: (2024)
von: Tanaka, Kunitomo, et al.
Veröffentlicht: (2024)
Do LLMs Find Human Answers To Fact-Driven Questions Perplexing? A Case Study on Reddit
von: Seegmiller, Parker, et al.
Veröffentlicht: (2024)
von: Seegmiller, Parker, et al.
Veröffentlicht: (2024)
Redundancy, Isotropy, and Intrinsic Dimensionality of Prompt-based Text Embeddings
von: Tsukagoshi, Hayato, et al.
Veröffentlicht: (2025)
von: Tsukagoshi, Hayato, et al.
Veröffentlicht: (2025)
Can We Still Hear the Accent? Investigating the Resilience of Native Language Signals in the LLM Era
von: Utami, Nabelanita, et al.
Veröffentlicht: (2026)
von: Utami, Nabelanita, et al.
Veröffentlicht: (2026)
Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering
von: Keluskar, Aryan, et al.
Veröffentlicht: (2024)
von: Keluskar, Aryan, et al.
Veröffentlicht: (2024)
Difficult for Whom? A Study of Japanese Lexical Complexity
von: Nohejl, Adam, et al.
Veröffentlicht: (2024)
von: Nohejl, Adam, et al.
Veröffentlicht: (2024)
On Representational Dissociation of Language and Arithmetic in Large Language Models
von: Kisako, Riku, et al.
Veröffentlicht: (2025)
von: Kisako, Riku, et al.
Veröffentlicht: (2025)
CiMaTe: Citation Count Prediction Effectively Leveraging the Main Text
von: Hirako, Jun, et al.
Veröffentlicht: (2024)
von: Hirako, Jun, et al.
Veröffentlicht: (2024)
When Is 0.1% Enough? Analyzing the Combined Effects of Dimensionality Reduction and Quantization on Text Embedding Compression
von: Kisako, Riku, et al.
Veröffentlicht: (2026)
von: Kisako, Riku, et al.
Veröffentlicht: (2026)
Verifying Claims About Metaphors with Large-Scale Automatic Metaphor Identification
von: Aono, Kotaro, et al.
Veröffentlicht: (2024)
von: Aono, Kotaro, et al.
Veröffentlicht: (2024)
Simplifying Translations for Children: Iterative Simplification Considering Age of Acquisition with LLMs
von: Oshika, Masashi, et al.
Veröffentlicht: (2024)
von: Oshika, Masashi, et al.
Veröffentlicht: (2024)
Enhancing Large Vision-Language Models with Layout Modality for Table Question Answering on Japanese Annual Securities Reports
von: Aida, Hayato, et al.
Veröffentlicht: (2025)
von: Aida, Hayato, et al.
Veröffentlicht: (2025)
DiscoTrace: Representing and Comparing Answering Strategies of Humans and LLMs in Information-Seeking Question Answering
von: Srikanth, Neha, et al.
Veröffentlicht: (2026)
von: Srikanth, Neha, et al.
Veröffentlicht: (2026)
Sentence Representations via Gaussian Embedding
von: Yoda, Shohei, et al.
Veröffentlicht: (2023)
von: Yoda, Shohei, et al.
Veröffentlicht: (2023)
Do LLMs Understand Romanian Driving Laws? A Study on Multimodal and Fine-Tuned Question Answering
von: Barbu, Eduard, et al.
Veröffentlicht: (2025)
von: Barbu, Eduard, et al.
Veröffentlicht: (2025)
Hidden in the Haystack: Smaller Needles are More Difficult for LLMs to Find
von: Bianchi, Owen, et al.
Veröffentlicht: (2025)
von: Bianchi, Owen, et al.
Veröffentlicht: (2025)
EDINET-Bench: Evaluating LLMs on Complex Financial Tasks using Japanese Financial Statements
von: Sugiura, Issa, et al.
Veröffentlicht: (2025)
von: Sugiura, Issa, et al.
Veröffentlicht: (2025)
Coal Mining Question Answering with LLMs
von: Rivera, Antonio Carlos, et al.
Veröffentlicht: (2024)
von: Rivera, Antonio Carlos, et al.
Veröffentlicht: (2024)
SAND-Math: Using LLMs to Generate Novel, Difficult and Useful Mathematics Questions and Answers
von: Manem, Chaitanya, et al.
Veröffentlicht: (2025)
von: Manem, Chaitanya, et al.
Veröffentlicht: (2025)
Context Quality Matters in Training Fusion-in-Decoder for Extractive Open-Domain Question Answering
von: Akimoto, Kosuke, et al.
Veröffentlicht: (2024)
von: Akimoto, Kosuke, et al.
Veröffentlicht: (2024)
A Gaze-grounded Visual Question Answering Dataset for Clarifying Ambiguous Japanese Questions
von: Inadumi, Shun, et al.
Veröffentlicht: (2024)
von: Inadumi, Shun, et al.
Veröffentlicht: (2024)
Pretraining and Updates of Domain-Specific LLM: A Case Study in the Japanese Business Domain
von: Takahashi, Kosuke, et al.
Veröffentlicht: (2024)
von: Takahashi, Kosuke, et al.
Veröffentlicht: (2024)
Agri-Query: A Case Study on RAG vs. Long-Context LLMs for Cross-Lingual Technical Question Answering
von: Gun, Julius, et al.
Veröffentlicht: (2025)
von: Gun, Julius, et al.
Veröffentlicht: (2025)
Exploring the Role of Knowledge Graph-Based RAG in Japanese Medical Question Answering with Small-Scale LLMs
von: Chen, Yingjian, et al.
Veröffentlicht: (2025)
von: Chen, Yingjian, et al.
Veröffentlicht: (2025)
Improving Sentence Embeddings with Automatic Generation of Training Data Using Few-shot Examples
von: Sato, Soma, et al.
Veröffentlicht: (2024)
von: Sato, Soma, et al.
Veröffentlicht: (2024)
Do LLMs Understand Your Translations? Evaluating Paragraph-level MT with Question Answering
von: Fernandes, Patrick, et al.
Veröffentlicht: (2025)
von: Fernandes, Patrick, et al.
Veröffentlicht: (2025)
ChatGPT as a Translation Engine: A Case Study on Japanese-English
von: Sutanto, Vincent Michael, et al.
Veröffentlicht: (2025)
von: Sutanto, Vincent Michael, et al.
Veröffentlicht: (2025)
On the Calibration of Multilingual Question Answering LLMs
von: Yang, Yahan, et al.
Veröffentlicht: (2023)
von: Yang, Yahan, et al.
Veröffentlicht: (2023)
JDocQA: Japanese Document Question Answering Dataset for Generative Language Models
von: Onami, Eri, et al.
Veröffentlicht: (2024)
von: Onami, Eri, et al.
Veröffentlicht: (2024)
LLMs Encode How Difficult Problems Are
von: Lugoloobi, William, et al.
Veröffentlicht: (2025)
von: Lugoloobi, William, et al.
Veröffentlicht: (2025)
Accurate Table Question Answering with Accessible LLMs
von: Jiang, Yangfan, et al.
Veröffentlicht: (2026)
von: Jiang, Yangfan, et al.
Veröffentlicht: (2026)
How Do Humans Write Code? Large Models Do It the Same Way Too
von: Li, Long, et al.
Veröffentlicht: (2024)
von: Li, Long, et al.
Veröffentlicht: (2024)
Agentic LLMs for Question Answering over Tabular Data
von: Tyagi, Rishit, et al.
Veröffentlicht: (2025)
von: Tyagi, Rishit, et al.
Veröffentlicht: (2025)
None of the Above, Less of the Right: Parallel Patterns between Humans and LLMs on Multi-Choice Questions Answering
von: Tam, Zhi Rui, et al.
Veröffentlicht: (2025)
von: Tam, Zhi Rui, et al.
Veröffentlicht: (2025)
Disambiguation in Conversational Question Answering in the Era of LLMs and Agents: A Survey
von: Tanjim, Md Mehrab, et al.
Veröffentlicht: (2025)
von: Tanjim, Md Mehrab, et al.
Veröffentlicht: (2025)
Combining LLMs and Knowledge Graphs to Reduce Hallucinations in Question Answering
von: Pusch, Larissa, et al.
Veröffentlicht: (2024)
von: Pusch, Larissa, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Ruri: Japanese General Text Embeddings
von: Tsukagoshi, Hayato, et al.
Veröffentlicht: (2024) -
To Drop or Not to Drop? Predicting Argument Ellipsis Judgments: A Case Study in Japanese
von: Ishizuki, Yukiko, et al.
Veröffentlicht: (2024) -
How Do Language Models Acquire Character-Level Information?
von: Sato, Soma, et al.
Veröffentlicht: (2026) -
FrameEOL: Semantic Frame Induction using Causal Language Models
von: Yano, Chihiro, et al.
Veröffentlicht: (2025) -
Are Social Sentiments Inherent in LLMs? An Empirical Study on Extraction of Inter-demographic Sentiments
von: Tanaka, Kunitomo, et al.
Veröffentlicht: (2024)