ProLex: A Benchmark for Language Proficiency-oriented Lexical Substitution
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Xuanming, Chen, Zixun, Yu, Zhou |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
MultiLexNorm++: A Unified Benchmark and a Generative Model for Lexical Normalization for Asian Languages
par: Buaphet, Weerayut, et autres
Publié: (2026)
par: Buaphet, Weerayut, et autres
Publié: (2026)
DECOR: Improving Coherence in L2 English Writing with a Novel Benchmark for Incoherence Detection, Reasoning, and Rewriting
par: Zhang, Xuanming, et autres
Publié: (2024)
par: Zhang, Xuanming, et autres
Publié: (2024)
LexInstructEval: Lexical Instruction Following Evaluation for Large Language Models
par: Ren, Huimin, et autres
Publié: (2025)
par: Ren, Huimin, et autres
Publié: (2025)
LexPro-1.0 Technical Report
par: Chen, Haotian, et autres
Publié: (2025)
par: Chen, Haotian, et autres
Publié: (2025)
NeLLCom-Lex: A Neural-agent Framework to Study the Interplay between Lexical Systems and Language Use
par: Zhang, Yuqing, et autres
Publié: (2025)
par: Zhang, Yuqing, et autres
Publié: (2025)
FastLexRank: Efficient Lexical Ranking for Structuring Social Media Posts
par: Li, Mao, et autres
Publié: (2024)
par: Li, Mao, et autres
Publié: (2024)
SetLexSem Challenge: Using Set Operations to Evaluate the Lexical and Semantic Robustness of Language Models
par: Akhbari, Bardiya, et autres
Publié: (2024)
par: Akhbari, Bardiya, et autres
Publié: (2024)
ViLexNorm: A Lexical Normalization Corpus for Vietnamese Social Media Text
par: Nguyen, Thanh-Nhi, et autres
Publié: (2024)
par: Nguyen, Thanh-Nhi, et autres
Publié: (2024)
LexEval: A Comprehensive Chinese Legal Benchmark for Evaluating Large Language Models
par: Li, Haitao, et autres
Publié: (2024)
par: Li, Haitao, et autres
Publié: (2024)
VarBench: Robust Language Model Benchmarking Through Dynamic Variable Perturbation
par: Qian, Kun, et autres
Publié: (2024)
par: Qian, Kun, et autres
Publié: (2024)
Lexical Substitution is not Synonym Substitution: On the Importance of Producing Contextually Relevant Word Substitutes
par: Vladika, Juraj, et autres
Publié: (2025)
par: Vladika, Juraj, et autres
Publié: (2025)
LexSumm and LexT5: Benchmarking and Modeling Legal Summarization Tasks in English
par: Santosh, T. Y. S. S., et autres
Publié: (2024)
par: Santosh, T. Y. S. S., et autres
Publié: (2024)
ViSoLex: An Open-Source Repository for Vietnamese Social Media Lexical Normalization
par: Nguyen, Anh Thi-Hoang, et autres
Publié: (2025)
par: Nguyen, Anh Thi-Hoang, et autres
Publié: (2025)
LexGenius: An Expert-Level Benchmark for Large Language Models in Legal General Intelligence
par: Liu, Wenjin, et autres
Publié: (2025)
par: Liu, Wenjin, et autres
Publié: (2025)
Redefining Simplicity: Benchmarking Large Language Models from Lexical to Document Simplification
par: Qiang, Jipeng, et autres
Publié: (2025)
par: Qiang, Jipeng, et autres
Publié: (2025)
Generalization or Memorization: Dynamic Decoding for Mode Steering
par: Zhang, Xuanming
Publié: (2025)
par: Zhang, Xuanming
Publié: (2025)
SpeciaLex: A Benchmark for In-Context Specialized Lexicon Learning
par: Imperial, Joseph Marvin, et autres
Publié: (2024)
par: Imperial, Joseph Marvin, et autres
Publié: (2024)
LexTime: A Benchmark for Temporal Ordering of Legal Events
par: Barale, Claire, et autres
Publié: (2025)
par: Barale, Claire, et autres
Publié: (2025)
Unsupervised Candidate Ranking for Lexical Substitution via Holistic Sentence Semantics
par: Hu, Zhongyang, et autres
Publié: (2025)
par: Hu, Zhongyang, et autres
Publié: (2025)
The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks
par: Pomerenke, David, et autres
Publié: (2025)
par: Pomerenke, David, et autres
Publié: (2025)
MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark
par: Liu, Hongwei, et autres
Publié: (2024)
par: Liu, Hongwei, et autres
Publié: (2024)
Cognition-of-Thought Elicits Social-Aligned Reasoning in Large Language Models
par: Zhang, Xuanming, et autres
Publié: (2025)
par: Zhang, Xuanming, et autres
Publié: (2025)
LexRel: Benchmarking Legal Relation Extraction for Chinese Civil Cases
par: Cai, Yida, et autres
Publié: (2025)
par: Cai, Yida, et autres
Publié: (2025)
SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis
par: Cai, Hengxing, et autres
Publié: (2024)
par: Cai, Hengxing, et autres
Publié: (2024)
ProBench: Benchmarking Large Language Models in Competitive Programming
par: Yang, Lei, et autres
Publié: (2025)
par: Yang, Lei, et autres
Publié: (2025)
PsychoLex: Unveiling the Psychological Mind of Large Language Models
par: Abbasi, Mohammad Amin, et autres
Publié: (2024)
par: Abbasi, Mohammad Amin, et autres
Publié: (2024)
LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation
par: Li, Haitao, et autres
Publié: (2025)
par: Li, Haitao, et autres
Publié: (2025)
Defining and Assessing Lexical Proficiency
par: Leńko-Szymańska, Agnieszka
Publié: (2025)
par: Leńko-Szymańska, Agnieszka
Publié: (2025)
MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark
par: Wang, Yubo, et autres
Publié: (2024)
par: Wang, Yubo, et autres
Publié: (2024)
MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
par: Xuan, Weihao, et autres
Publié: (2025)
par: Xuan, Weihao, et autres
Publié: (2025)
ChiKhaPo: A Large-Scale Multilingual Benchmark for Evaluating Lexical Comprehension and Generation in Large Language Models
par: Chang, Emily, et autres
Publié: (2025)
par: Chang, Emily, et autres
Publié: (2025)
Unraveling the Complexities of Second Language Lexical Stress Processing: The Impact of First Language Transfer, Second Language Proficiency, and Exposure
par: Nuria Sagarra, et autres
Publié: (2024)
par: Nuria Sagarra, et autres
Publié: (2024)
TriLex: A Framework for Multilingual Sentiment Analysis in Low-Resource South African Languages
par: Nkongolo, Mike, et autres
Publié: (2025)
par: Nkongolo, Mike, et autres
Publié: (2025)
MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark
par: Yue, Xiang, et autres
Publié: (2024)
par: Yue, Xiang, et autres
Publié: (2024)
MetaMind: Modeling Human Social Thoughts with Metacognitive Multi-Agent Systems
par: Zhang, Xuanming, et autres
Publié: (2025)
par: Zhang, Xuanming, et autres
Publié: (2025)
PrunePath: Towards Highly Structured Sparse Language Models
par: Gu, Zhexuan, et autres
Publié: (2026)
par: Gu, Zhexuan, et autres
Publié: (2026)
Towards Controllable Natural Language Inference through Lexical Inference Types
par: Zhang, Yingji, et autres
Publié: (2023)
par: Zhang, Yingji, et autres
Publié: (2023)
Bringing Pedagogy into Focus: Evaluating Virtual Teaching Assistants' Question-Answering in Asynchronous Learning Environments
par: Siyan, Li, et autres
Publié: (2025)
par: Siyan, Li, et autres
Publié: (2025)
Disce aut Deficere: Evaluating LLMs Proficiency on the INVALSI Italian Benchmark
par: Mercorio, Fabio, et autres
Publié: (2024)
par: Mercorio, Fabio, et autres
Publié: (2024)
VISLA Benchmark: Evaluating Embedding Sensitivity to Semantic and Lexical Alterations
par: Dumpala, Sri Harsha, et autres
Publié: (2024)
par: Dumpala, Sri Harsha, et autres
Publié: (2024)
Documents similaires
-
MultiLexNorm++: A Unified Benchmark and a Generative Model for Lexical Normalization for Asian Languages
par: Buaphet, Weerayut, et autres
Publié: (2026) -
DECOR: Improving Coherence in L2 English Writing with a Novel Benchmark for Incoherence Detection, Reasoning, and Rewriting
par: Zhang, Xuanming, et autres
Publié: (2024) -
LexInstructEval: Lexical Instruction Following Evaluation for Large Language Models
par: Ren, Huimin, et autres
Publié: (2025) -
LexPro-1.0 Technical Report
par: Chen, Haotian, et autres
Publié: (2025) -
NeLLCom-Lex: A Neural-agent Framework to Study the Interplay between Lexical Systems and Language Use
par: Zhang, Yuqing, et autres
Publié: (2025)