Sakura at BEA 2026 Shared Task 1: What Makes Vocabulary Difficult?
Fuente:
arXiv
Guardado en:
| Autores principales: | Nohejl, Adam, Wu, Xuanxin, Ide, Yusuke, Machin, Maria Angelica Riera, Chang, Yi-Ning, Yanaka, Hitomi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards Automated Lexicography: Generating and Evaluating Definitions for Learner's Dictionaries
por: Ide, Yusuke, et al.
Publicado: (2026)
por: Ide, Yusuke, et al.
Publicado: (2026)
Difficult for Whom? A Study of Japanese Lexical Complexity
por: Nohejl, Adam, et al.
Publicado: (2024)
por: Nohejl, Adam, et al.
Publicado: (2024)
Fine-Grained Analysis of Shared Syntactic Mechanisms in Language Models
por: Kumon, Ryoma, et al.
Publicado: (2026)
por: Kumon, Ryoma, et al.
Publicado: (2026)
Beyond Film Subtitles: Is YouTube the Best Approximation of Spoken Vocabulary?
por: Nohejl, Adam, et al.
Publicado: (2024)
por: Nohejl, Adam, et al.
Publicado: (2024)
What Do Vision-Language Models Encode for Personalized Image Aesthetics Assessment?
por: Ryu, Koki, et al.
Publicado: (2026)
por: Ryu, Koki, et al.
Publicado: (2026)
Analyzing the Inner Workings of Transformers in Compositional Generalization
por: Kumon, Ryoma, et al.
Publicado: (2025)
por: Kumon, Ryoma, et al.
Publicado: (2025)
Enhancing Rating Prediction with Off-the-Shelf LLMs Using In-Context User Reviews
por: Ryu, Koki, et al.
Publicado: (2025)
por: Ryu, Koki, et al.
Publicado: (2025)
NeuronMoE: Neuron-Guided Mixture-of-Experts for Efficient Multilingual LLM Extension
por: Li, Rongzhi, et al.
Publicado: (2026)
por: Li, Rongzhi, et al.
Publicado: (2026)
RETUYT-INCO at BEA 2026 Shared Task 2: Meta-prompting in Rubric-based Scoring for German
por: Sastre, Ignacio, et al.
Publicado: (2026)
por: Sastre, Ignacio, et al.
Publicado: (2026)
Can Large Language Models Robustly Perform Natural Language Inference for Japanese Comparatives?
por: Mikami, Yosuke, et al.
Publicado: (2025)
por: Mikami, Yosuke, et al.
Publicado: (2025)
LLMs Struggle with NLI for Perfect Aspect: A Cross-Linguistic Study in Chinese and Japanese
por: Lu, Jie, et al.
Publicado: (2025)
por: Lu, Jie, et al.
Publicado: (2025)
Implementing a Logical Inference System for Japanese Comparatives
por: Mikami, Yosuke, et al.
Publicado: (2025)
por: Mikami, Yosuke, et al.
Publicado: (2025)
Evaluating Structural Generalization in Neural Machine Translation
por: Kumon, Ryoma, et al.
Publicado: (2024)
por: Kumon, Ryoma, et al.
Publicado: (2024)
Investigating Training and Generalization in Faithful Self-Explanations of Large Language Models
por: Doi, Tomoki, et al.
Publicado: (2025)
por: Doi, Tomoki, et al.
Publicado: (2025)
Comprehensive Evaluation of Large Language Models for Topic Modeling
por: Doi, Tomoki, et al.
Publicado: (2024)
por: Doi, Tomoki, et al.
Publicado: (2024)
CoAM: Corpus of All-Type Multiword Expressions
por: Ide, Yusuke, et al.
Publicado: (2024)
por: Ide, Yusuke, et al.
Publicado: (2024)
On the Multilingual Ability of Decoder-based Pre-trained Language Models: Finding and Controlling Language-Specific Neurons
por: Kojima, Takeshi, et al.
Publicado: (2024)
por: Kojima, Takeshi, et al.
Publicado: (2024)
Exploring Intra and Inter-language Consistency in Embeddings with ICA
por: Li, Rongzhi, et al.
Publicado: (2024)
por: Li, Rongzhi, et al.
Publicado: (2024)
Dispersion Measures as Predictors of Lexical Decision Time, Word Familiarity, and Lexical Complexity
por: Nohejl, Adam, et al.
Publicado: (2025)
por: Nohejl, Adam, et al.
Publicado: (2025)
Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset
por: Yamamoto, Taisei, et al.
Publicado: (2025)
por: Yamamoto, Taisei, et al.
Publicado: (2025)
Neuron-Level Analysis of Cultural Understanding in Large Language Models
por: Yamamoto, Taisei, et al.
Publicado: (2025)
por: Yamamoto, Taisei, et al.
Publicado: (2025)
Derivational Probing: Unveiling the Layer-wise Derivation of Syntactic Structures in Neural Language Models
por: Someya, Taiga, et al.
Publicado: (2025)
por: Someya, Taiga, et al.
Publicado: (2025)
Findings of the BEA 2025 Shared Task on Pedagogical Ability Assessment of AI-powered Tutors
por: Kochmar, Ekaterina, et al.
Publicado: (2025)
por: Kochmar, Ekaterina, et al.
Publicado: (2025)
Measuring the Robustness of Reference-Free Dialogue Evaluation Systems
por: Vasselli, Justin, et al.
Publicado: (2025)
por: Vasselli, Justin, et al.
Publicado: (2025)
BD at BEA 2025 Shared Task: MPNet Ensembles for Pedagogical Mistake Identification and Localization in AI Tutor Responses
por: Rohan, Shadman, et al.
Publicado: (2025)
por: Rohan, Shadman, et al.
Publicado: (2025)
The ADAIO System at the BEA-2023 Shared Task on Generating AI Teacher Responses in Educational Dialogues
por: Adigwe, Adaeze, et al.
Publicado: (2023)
por: Adigwe, Adaeze, et al.
Publicado: (2023)
Toward the Evaluation of Large Language Models Considering Score Variance across Instruction Templates
por: Sakai, Yusuke, et al.
Publicado: (2024)
por: Sakai, Yusuke, et al.
Publicado: (2024)
MSA at BEA 2025 Shared Task: Disagreement-Aware Instruction Tuning for Multi-Dimensional Evaluation of LLMs as Math Tutors
por: Hikal, Baraa, et al.
Publicado: (2025)
por: Hikal, Baraa, et al.
Publicado: (2025)
An In-depth Evaluation of Large Language Models in Sentence Simplification with Error-based Human Assessment
por: Wu, Xuanxin, et al.
Publicado: (2024)
por: Wu, Xuanxin, et al.
Publicado: (2024)
NeuralNexus at BEA 2025 Shared Task: Retrieval-Augmented Prompting for Mistake Identification in AI Tutors
por: Naeem, Numaan, et al.
Publicado: (2025)
por: Naeem, Numaan, et al.
Publicado: (2025)
RETUYT-INCO at BEA 2025 Shared Task: How Far Can Lightweight Models Go in AI-powered Tutor Evaluation?
por: Góngora, Santiago, et al.
Publicado: (2025)
por: Góngora, Santiago, et al.
Publicado: (2025)
Do Large Vision-Language Models Distinguish between the Actual and Apparent Features of Illusions?
por: Shinozaki, Taiga, et al.
Publicado: (2025)
por: Shinozaki, Taiga, et al.
Publicado: (2025)
Dictionaries to the Rescue: Cross-Lingual Vocabulary Transfer for Low-Resource Languages Using Bilingual Dictionaries
por: Sakajo, Haruki, et al.
Publicado: (2025)
por: Sakajo, Haruki, et al.
Publicado: (2025)
Policy-based Sentence Simplification: Replacing Parallel Corpora with LLM-as-a-Judge
por: Wu, Xuanxin, et al.
Publicado: (2025)
por: Wu, Xuanxin, et al.
Publicado: (2025)
How to Make the Most of LLMs' Grammatical Knowledge for Acceptability Judgments
por: Ide, Yusuke, et al.
Publicado: (2024)
por: Ide, Yusuke, et al.
Publicado: (2024)
J-CHAT: Japanese Large-scale Spoken Dialogue Corpus for Spoken Dialogue Language Modeling
por: Nakata, Wataru, et al.
Publicado: (2024)
por: Nakata, Wataru, et al.
Publicado: (2024)
Toward Conversational Hungarian Speech Recognition: Introducing the BEA-Large and BEA-Dialogue Datasets
por: Gedeon, Máté, et al.
Publicado: (2025)
por: Gedeon, Máté, et al.
Publicado: (2025)
Bridging Perception and Language: A Systematic Benchmark for LVLMs' Understanding of Amodal Completion Reports
por: Watahiki, Amane, et al.
Publicado: (2025)
por: Watahiki, Amane, et al.
Publicado: (2025)
Reasoning Model Is Superior LLM-Judge, Yet Suffers from Biases
por: Huang, Hui, et al.
Publicado: (2026)
por: Huang, Hui, et al.
Publicado: (2026)
Difficult Task Yes but Simple Task No: Unveiling the Laziness in Multimodal LLMs
por: Zhao, Sihang, et al.
Publicado: (2024)
por: Zhao, Sihang, et al.
Publicado: (2024)
Ejemplares similares
-
Towards Automated Lexicography: Generating and Evaluating Definitions for Learner's Dictionaries
por: Ide, Yusuke, et al.
Publicado: (2026) -
Difficult for Whom? A Study of Japanese Lexical Complexity
por: Nohejl, Adam, et al.
Publicado: (2024) -
Fine-Grained Analysis of Shared Syntactic Mechanisms in Language Models
por: Kumon, Ryoma, et al.
Publicado: (2026) -
Beyond Film Subtitles: Is YouTube the Best Approximation of Spoken Vocabulary?
por: Nohejl, Adam, et al.
Publicado: (2024) -
What Do Vision-Language Models Encode for Personalized Image Aesthetics Assessment?
por: Ryu, Koki, et al.
Publicado: (2026)