Difficult for Whom? A Study of Japanese Lexical Complexity
Fuente:
arXiv
Saved in:
| Main Authors: | Nohejl, Adam, Hayakawa, Akio, Ide, Yusuke, Watanabe, Taro |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dispersion Measures as Predictors of Lexical Decision Time, Word Familiarity, and Lexical Complexity
by: Nohejl, Adam, et al.
Published: (2025)
by: Nohejl, Adam, et al.
Published: (2025)
Towards Automated Lexicography: Generating and Evaluating Definitions for Learner's Dictionaries
by: Ide, Yusuke, et al.
Published: (2026)
by: Ide, Yusuke, et al.
Published: (2026)
Sakura at BEA 2026 Shared Task 1: What Makes Vocabulary Difficult?
by: Nohejl, Adam, et al.
Published: (2026)
by: Nohejl, Adam, et al.
Published: (2026)
CoAM: Corpus of All-Type Multiword Expressions
by: Ide, Yusuke, et al.
Published: (2024)
by: Ide, Yusuke, et al.
Published: (2024)
Measuring the Robustness of Reference-Free Dialogue Evaluation Systems
by: Vasselli, Justin, et al.
Published: (2025)
by: Vasselli, Justin, et al.
Published: (2025)
Toward the Evaluation of Large Language Models Considering Score Variance across Instruction Templates
by: Sakai, Yusuke, et al.
Published: (2024)
by: Sakai, Yusuke, et al.
Published: (2024)
Towards Trustworthy Lexical Simplification: Exploring Safety and Efficiency with Small LLMs
by: Hayakawa, Akio, et al.
Published: (2025)
by: Hayakawa, Akio, et al.
Published: (2025)
How to Make the Most of LLMs' Grammatical Knowledge for Acceptability Judgments
by: Ide, Yusuke, et al.
Published: (2024)
by: Ide, Yusuke, et al.
Published: (2024)
Dictionaries to the Rescue: Cross-Lingual Vocabulary Transfer for Low-Resource Languages Using Bilingual Dictionaries
by: Sakajo, Haruki, et al.
Published: (2025)
by: Sakajo, Haruki, et al.
Published: (2025)
Beyond Film Subtitles: Is YouTube the Best Approximation of Spoken Vocabulary?
by: Nohejl, Adam, et al.
Published: (2024)
by: Nohejl, Adam, et al.
Published: (2024)
gec-metrics: A Unified Library for Grammatical Error Correction Evaluation
by: Goto, Takumi, et al.
Published: (2025)
by: Goto, Takumi, et al.
Published: (2025)
Do LLMs and Humans Find the Same Questions Difficult? A Case Study on Japanese Quiz Answering
by: Sugiura, Naoya, et al.
Published: (2025)
by: Sugiura, Naoya, et al.
Published: (2025)
Edit-level Majority Voting Mitigates Over-Correction in LLM-based Grammatical Error Correction
by: Goto, Takumi, et al.
Published: (2026)
by: Goto, Takumi, et al.
Published: (2026)
Reliability Crisis of Reference-free Metrics for Grammatical Error Correction
by: Goto, Takumi, et al.
Published: (2025)
by: Goto, Takumi, et al.
Published: (2025)
Rethinking Evaluation Metrics for Grammatical Error Correction: Why Use a Different Evaluation Process than Human?
by: Goto, Takumi, et al.
Published: (2025)
by: Goto, Takumi, et al.
Published: (2025)
Grammatical Error Correction Evaluation by Optimally Transporting Edit Representation
by: Goto, Takumi, et al.
Published: (2026)
by: Goto, Takumi, et al.
Published: (2026)
JDocQA: Japanese Document Question Answering Dataset for Generative Language Models
by: Onami, Eri, et al.
Published: (2024)
by: Onami, Eri, et al.
Published: (2024)
mbrs: A Library for Minimum Bayes Risk Decoding
by: Deguchi, Hiroyuki, et al.
Published: (2024)
by: Deguchi, Hiroyuki, et al.
Published: (2024)
Revisiting Compositional Generalization Capability of Large Language Models Considering Instruction Following Ability
by: Sakai, Yusuke, et al.
Published: (2025)
by: Sakai, Yusuke, et al.
Published: (2025)
IMPARA-GED: Grammatical Error Detection is Boosting Reference-free Grammatical Error Quality Estimator
by: Sakai, Yusuke, et al.
Published: (2025)
by: Sakai, Yusuke, et al.
Published: (2025)
Diversity Explains Inference Scaling Laws: Through a Case Study of Minimum Bayes Risk Decoding
by: Kamigaito, Hidetaka, et al.
Published: (2024)
by: Kamigaito, Hidetaka, et al.
Published: (2024)
HalluCiteChecker: A Lightweight Toolkit for Hallucinated Citation Detection and Verification in the Era of AI Scientists
by: Sakai, Yusuke, et al.
Published: (2026)
by: Sakai, Yusuke, et al.
Published: (2026)
Lexical Complexity Prediction and Lexical Simplification for Catalan and Spanish: Resource Creation, Quality Assessment, and Ethical Considerations
by: Bott, Stefan, et al.
Published: (2024)
by: Bott, Stefan, et al.
Published: (2024)
mCSQA: Multilingual Commonsense Reasoning Dataset with Unified Creation Strategy by Language Models and Humans
by: Sakai, Yusuke, et al.
Published: (2024)
by: Sakai, Yusuke, et al.
Published: (2024)
HalluCitation Matters: Revealing the Impact of Hallucinated References with 300 Hallucinated Papers in ACL Conferences
by: Sakai, Yusuke, et al.
Published: (2026)
by: Sakai, Yusuke, et al.
Published: (2026)
Estimating Lexical Complexity from Document-Level Distributions
by: Wold, Sondre, et al.
Published: (2024)
by: Wold, Sondre, et al.
Published: (2024)
AdTEC: A Unified Benchmark for Evaluating Text Quality in Search Engine Advertising
by: Zhang, Peinan, et al.
Published: (2024)
by: Zhang, Peinan, et al.
Published: (2024)
Multilingual Dialogue Generation and Localization with Dialogue Act Scripting
by: Vasselli, Justin, et al.
Published: (2025)
by: Vasselli, Justin, et al.
Published: (2025)
IRR: Image Review Ranking Framework for Evaluating Vision-Language Models
by: Hayashi, Kazuki, et al.
Published: (2024)
by: Hayashi, Kazuki, et al.
Published: (2024)
Tonguescape: Exploring Language Models Understanding of Vowel Articulation
by: Sakajo, Haruki, et al.
Published: (2025)
by: Sakajo, Haruki, et al.
Published: (2025)
An Empirical Study of LLM-as-a-Judge: How Design Choices Impact Evaluation Reliability
by: Yamauchi, Yusuke, et al.
Published: (2025)
by: Yamauchi, Yusuke, et al.
Published: (2025)
Generating Difficult-to-Translate Texts
by: Zouhar, Vilém, et al.
Published: (2025)
by: Zouhar, Vilém, et al.
Published: (2025)
StructLens: A Structural Lens for Language Models via Maximum Spanning Trees
by: Sakajo, Haruki, et al.
Published: (2026)
by: Sakajo, Haruki, et al.
Published: (2026)
Introducing Syllable Tokenization for Low-resource Languages: A Case Study with Swahili
by: Atuhurra, Jesse, et al.
Published: (2024)
by: Atuhurra, Jesse, et al.
Published: (2024)
Common to Whom? Regional Cultural Commonsense and LLM Bias in India
by: Madhusudan, Sangmitra, et al.
Published: (2026)
by: Madhusudan, Sangmitra, et al.
Published: (2026)
Does Pre-trained Language Model Actually Infer Unseen Links in Knowledge Graph Completion?
by: Sakai, Yusuke, et al.
Published: (2023)
by: Sakai, Yusuke, et al.
Published: (2023)
Multilinguality of Large Language Models From a Structural Perspective
by: Sakajo, Haruki, et al.
Published: (2026)
by: Sakajo, Haruki, et al.
Published: (2026)
LLMs Encode How Difficult Problems Are
by: Lugoloobi, William, et al.
Published: (2025)
by: Lugoloobi, William, et al.
Published: (2025)
Diagnosing Vision Language Models' Perception by Leveraging Human Methods for Color Vision Deficiencies
by: Hayashi, Kazuki, et al.
Published: (2025)
by: Hayashi, Kazuki, et al.
Published: (2025)
BQA: Body Language Question Answering Dataset for Video Large Language Models
by: Ozaki, Shintaro, et al.
Published: (2024)
by: Ozaki, Shintaro, et al.
Published: (2024)
Similar Items
-
Dispersion Measures as Predictors of Lexical Decision Time, Word Familiarity, and Lexical Complexity
by: Nohejl, Adam, et al.
Published: (2025) -
Towards Automated Lexicography: Generating and Evaluating Definitions for Learner's Dictionaries
by: Ide, Yusuke, et al.
Published: (2026) -
Sakura at BEA 2026 Shared Task 1: What Makes Vocabulary Difficult?
by: Nohejl, Adam, et al.
Published: (2026) -
CoAM: Corpus of All-Type Multiword Expressions
by: Ide, Yusuke, et al.
Published: (2024) -
Measuring the Robustness of Reference-Free Dialogue Evaluation Systems
by: Vasselli, Justin, et al.
Published: (2025)