ERAS: Evaluating the Robustness of Chinese NLP Models to Morphological Garden Path Errors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Qinchan, Hao, Sophie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CEC-Zero: Chinese Error Correction Solution Based on LLM
von: Zhang, Sophie, et al.
Veröffentlicht: (2025)
von: Zhang, Sophie, et al.
Veröffentlicht: (2025)
Loss-Aware Curriculum Learning for Chinese Grammatical Error Correction
von: Zhang, Ding, et al.
Veröffentlicht: (2024)
von: Zhang, Ding, et al.
Veröffentlicht: (2024)
Evaluation Metrics for Text Data Augmentation in NLP
von: Amadeus, Marcellus, et al.
Veröffentlicht: (2024)
von: Amadeus, Marcellus, et al.
Veröffentlicht: (2024)
Select, Label, Evaluate: Active Testing in NLP
von: Purificato, Antonio, et al.
Veröffentlicht: (2026)
von: Purificato, Antonio, et al.
Veröffentlicht: (2026)
Large Language Models Meet NLP: A Survey
von: Qin, Libo, et al.
Veröffentlicht: (2024)
von: Qin, Libo, et al.
Veröffentlicht: (2024)
Alirector: Alignment-Enhanced Chinese Grammatical Error Corrector
von: Yang, Haihui, et al.
Veröffentlicht: (2024)
von: Yang, Haihui, et al.
Veröffentlicht: (2024)
GREEN: Generative Radiology Report Evaluation and Error Notation
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024)
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024)
An Empirical Study on Large Language Models in Accuracy and Robustness under Chinese Industrial Scenarios
von: Li, Zongjie, et al.
Veröffentlicht: (2024)
von: Li, Zongjie, et al.
Veröffentlicht: (2024)
ChineseErrorCorrector3-4B: State-of-the-Art Chinese Spelling and Grammar Corrector
von: Tian, Wei, et al.
Veröffentlicht: (2025)
von: Tian, Wei, et al.
Veröffentlicht: (2025)
Evaluating Morphological Compositional Generalization in Large Language Models
von: Ismayilzada, Mete, et al.
Veröffentlicht: (2024)
von: Ismayilzada, Mete, et al.
Veröffentlicht: (2024)
CIF-Bench: A Chinese Instruction-Following Benchmark for Evaluating the Generalizability of Large Language Models
von: LI, Yizhi, et al.
Veröffentlicht: (2024)
von: LI, Yizhi, et al.
Veröffentlicht: (2024)
Enhancing Steganographic Text Extraction: Evaluating the Impact of NLP Models on Accuracy and Semantic Coherence
von: Li, Mingyang, et al.
Veröffentlicht: (2024)
von: Li, Mingyang, et al.
Veröffentlicht: (2024)
Deanthropomorphising NLP: Can a Language Model Be Conscious?
von: Shardlow, Matthew, et al.
Veröffentlicht: (2022)
von: Shardlow, Matthew, et al.
Veröffentlicht: (2022)
The Garden of Forking Paths: Observing Dynamic Parameters Distribution in Large Language Models
von: Nicolini, Carlo, et al.
Veröffentlicht: (2024)
von: Nicolini, Carlo, et al.
Veröffentlicht: (2024)
Safety Evaluation of DeepSeek Models in Chinese Contexts
von: Zhang, Wenjing, et al.
Veröffentlicht: (2025)
von: Zhang, Wenjing, et al.
Veröffentlicht: (2025)
Leveraging LLMs for Bangla Grammar Error Correction:Error Categorization, Synthetic Data, and Model Evaluation
von: Bhattacharyya, Pramit, et al.
Veröffentlicht: (2024)
von: Bhattacharyya, Pramit, et al.
Veröffentlicht: (2024)
MASE: Interpretable NLP Models via Model-Agnostic Saliency Estimation
von: Yang, Zhou, et al.
Veröffentlicht: (2025)
von: Yang, Zhou, et al.
Veröffentlicht: (2025)
Advancing Prompt Recovery in NLP: A Deep Dive into the Integration of Gemma-2b-it and Phi2 Models
von: Chen, Jianlong, et al.
Veröffentlicht: (2024)
von: Chen, Jianlong, et al.
Veröffentlicht: (2024)
Benchmarking Large Language Models on Multiple Tasks in Bioinformatics NLP with Prompting
von: Jiang, Jiyue, et al.
Veröffentlicht: (2025)
von: Jiang, Jiyue, et al.
Veröffentlicht: (2025)
AncientBench: Towards Comprehensive Evaluation on Excavated and Transmitted Chinese Corpora
von: Zhou, Zhihan, et al.
Veröffentlicht: (2025)
von: Zhou, Zhihan, et al.
Veröffentlicht: (2025)
CL$^2$GEC: A Multi-Discipline Benchmark for Continual Learning in Chinese Literature Grammatical Error Correction
von: Qin, Shang, et al.
Veröffentlicht: (2025)
von: Qin, Shang, et al.
Veröffentlicht: (2025)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs
von: Calderon, Nitay, et al.
Veröffentlicht: (2024)
von: Calderon, Nitay, et al.
Veröffentlicht: (2024)
Modeling Orthographic Variation Improves NLP Performance for Nigerian Pidgin
von: Lin, Pin-Jie, et al.
Veröffentlicht: (2024)
von: Lin, Pin-Jie, et al.
Veröffentlicht: (2024)
Tokenization and Representation Biases in Multilingual Models on Dialectal NLP Tasks
von: Kanjirangat, Vani, et al.
Veröffentlicht: (2025)
von: Kanjirangat, Vani, et al.
Veröffentlicht: (2025)
FoundaBench: Evaluating Chinese Fundamental Knowledge Capabilities of Large Language Models
von: Li, Wei, et al.
Veröffentlicht: (2024)
von: Li, Wei, et al.
Veröffentlicht: (2024)
NLP Security and Ethics, in the Wild
von: Lent, Heather, et al.
Veröffentlicht: (2025)
von: Lent, Heather, et al.
Veröffentlicht: (2025)
Reasoning Robustness of LLMs to Adversarial Typographical Errors
von: Gan, Esther, et al.
Veröffentlicht: (2024)
von: Gan, Esther, et al.
Veröffentlicht: (2024)
EssayBench: Evaluating Large Language Models in Multi-Genre Chinese Essay Writing
von: Gao, Fan, et al.
Veröffentlicht: (2025)
von: Gao, Fan, et al.
Veröffentlicht: (2025)
Benchmarking Large Language Models on CFLUE -- A Chinese Financial Language Understanding Evaluation Dataset
von: Zhu, Jie, et al.
Veröffentlicht: (2024)
von: Zhu, Jie, et al.
Veröffentlicht: (2024)
Batayan: A Filipino NLP benchmark for evaluating Large Language Models
von: Montalan, Jann Railey, et al.
Veröffentlicht: (2025)
von: Montalan, Jann Railey, et al.
Veröffentlicht: (2025)
EvalxNLP: A Framework for Benchmarking Post-Hoc Explainability Methods on NLP Models
von: Dhaini, Mahdi, et al.
Veröffentlicht: (2025)
von: Dhaini, Mahdi, et al.
Veröffentlicht: (2025)
NLP Verification: Towards a General Methodology for Certifying Robustness
von: Casadio, Marco, et al.
Veröffentlicht: (2024)
von: Casadio, Marco, et al.
Veröffentlicht: (2024)
Cheems: A Practical Guidance for Building and Evaluating Chinese Reward Models from Scratch
von: Wen, Xueru, et al.
Veröffentlicht: (2025)
von: Wen, Xueru, et al.
Veröffentlicht: (2025)
Evaluating Modern Large Language Models on Low-Resource and Morphologically Rich Languages:A Cross-Lingual Benchmark Across Cantonese, Japanese, and Turkish
von: Xia, Chengxuan, et al.
Veröffentlicht: (2025)
von: Xia, Chengxuan, et al.
Veröffentlicht: (2025)
HSKBenchmark: Modeling and Benchmarking Chinese Second Language Acquisition in Large Language Models through Curriculum Tuning
von: Yang, Qihao, et al.
Veröffentlicht: (2025)
von: Yang, Qihao, et al.
Veröffentlicht: (2025)
An Audit on the Perspectives and Challenges of Hallucinations in NLP
von: Venkit, Pranav Narayanan, et al.
Veröffentlicht: (2024)
von: Venkit, Pranav Narayanan, et al.
Veröffentlicht: (2024)
State of NLP in Kenya: A Survey
von: Amol, Cynthia Jayne, et al.
Veröffentlicht: (2024)
von: Amol, Cynthia Jayne, et al.
Veröffentlicht: (2024)
On Evaluating Explanation Utility for Human-AI Decision Making in NLP
von: Chaleshtori, Fateme Hashemi, et al.
Veröffentlicht: (2024)
von: Chaleshtori, Fateme Hashemi, et al.
Veröffentlicht: (2024)
From Word Sequences to Behavioral Sequences: Adapting Modeling and Evaluation Paradigms for Longitudinal NLP
von: Ganesan, Adithya V, et al.
Veröffentlicht: (2026)
von: Ganesan, Adithya V, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CEC-Zero: Chinese Error Correction Solution Based on LLM
von: Zhang, Sophie, et al.
Veröffentlicht: (2025) -
Loss-Aware Curriculum Learning for Chinese Grammatical Error Correction
von: Zhang, Ding, et al.
Veröffentlicht: (2024) -
Evaluation Metrics for Text Data Augmentation in NLP
von: Amadeus, Marcellus, et al.
Veröffentlicht: (2024) -
Select, Label, Evaluate: Active Testing in NLP
von: Purificato, Antonio, et al.
Veröffentlicht: (2026) -
Large Language Models Meet NLP: A Survey
von: Qin, Libo, et al.
Veröffentlicht: (2024)