Error Diversity Matters: An Error-Resistant Ensemble Method for Unsupervised Dependency Parsing
Fuente:
arXiv
Guardado en:
| Autores principales: | Shayegh, Behzad, Lee, Hobie H. -B., Zhu, Xiaodan, Cheung, Jackie Chi Kit, Mou, Lili |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Ensemble Distillation for Unsupervised Constituency Parsing
por: Shayegh, Behzad, et al.
Publicado: (2023)
por: Shayegh, Behzad, et al.
Publicado: (2023)
Tree-Averaging Algorithms for Ensemble-Based Unsupervised Discontinuous Constituency Parsing
por: Shayegh, Behzad, et al.
Publicado: (2024)
por: Shayegh, Behzad, et al.
Publicado: (2024)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
por: Wen, Yuqiao, et al.
Publicado: (2024)
por: Wen, Yuqiao, et al.
Publicado: (2024)
PreSumm: Predicting Summarization Performance Without Summarizing
por: Koniaev, Steven, et al.
Publicado: (2025)
por: Koniaev, Steven, et al.
Publicado: (2025)
Feeding Two Birds or Favoring One? Adequacy-Fluency Tradeoffs in Evaluation and Meta-Evaluation of Machine Translation
por: Shayegh, Behzad, et al.
Publicado: (2025)
por: Shayegh, Behzad, et al.
Publicado: (2025)
Mechanistic Understanding and Mitigation of Language Model Non-Factual Hallucinations
por: Yu, Lei, et al.
Publicado: (2024)
por: Yu, Lei, et al.
Publicado: (2024)
Improving the Calibration of Confidence Scores in Text Generation Using the Output Distribution's Characteristics
por: Flores, Lorenzo Jaime Yu, et al.
Publicado: (2025)
por: Flores, Lorenzo Jaime Yu, et al.
Publicado: (2025)
$\texttt{COSMIC}$: Mutual Information for Task-Agnostic Summarization Evaluation
por: Darrin, Maxime, et al.
Publicado: (2024)
por: Darrin, Maxime, et al.
Publicado: (2024)
Solving the Challenge Set without Solving the Task: On Winograd Schemas as a Test of Pronominal Coreference Resolution
por: Porada, Ian, et al.
Publicado: (2024)
por: Porada, Ian, et al.
Publicado: (2024)
$(RSA)^2$: A Rhetorical-Strategy-Aware Rational Speech Act Framework for Figurative Language Understanding
por: Piano, Cesare Spinoso-Di, et al.
Publicado: (2025)
por: Piano, Cesare Spinoso-Di, et al.
Publicado: (2025)
Can Vision Language Models Be Adaptive in Mathematics Education? A Learner Model-based Rubric Study
por: Gao, Jie, et al.
Publicado: (2026)
por: Gao, Jie, et al.
Publicado: (2026)
A Controlled Reevaluation of Coreference Resolution Models
por: Porada, Ian, et al.
Publicado: (2024)
por: Porada, Ian, et al.
Publicado: (2024)
Detecting Errors through Ensembling Prompts (DEEP): An End-to-End LLM Framework for Detecting Factual Errors
por: Chandler, Alex, et al.
Publicado: (2024)
por: Chandler, Alex, et al.
Publicado: (2024)
Improving LLM Classification of Logical Errors by Integrating Error Relationship into Prompts
por: Lee, Yanggyu, et al.
Publicado: (2024)
por: Lee, Yanggyu, et al.
Publicado: (2024)
Unsupervised Layer-wise Score Aggregation for Textual OOD Detection
por: Darrin, Maxime, et al.
Publicado: (2023)
por: Darrin, Maxime, et al.
Publicado: (2023)
Does This Summary Answer My Question? Modeling Query-Focused Summary Readers with Rational Speech Acts
por: Piano, Cesare Spinoso-Di, et al.
Publicado: (2024)
por: Piano, Cesare Spinoso-Di, et al.
Publicado: (2024)
AceParse: A Comprehensive Dataset with Diverse Structured Texts for Academic Literature Parsing
por: Ji, Huawei, et al.
Publicado: (2024)
por: Ji, Huawei, et al.
Publicado: (2024)
Identifying the Achilles' Heel: An Iterative Method for Dynamically Uncovering Factual Errors in Large Language Models
por: Wang, Wenxuan, et al.
Publicado: (2024)
por: Wang, Wenxuan, et al.
Publicado: (2024)
Attention with Dependency Parsing Augmentation for Fine-Grained Attribution
por: Ding, Qiang, et al.
Publicado: (2024)
por: Ding, Qiang, et al.
Publicado: (2024)
Can LLMs Take Retrieved Information with a Grain of Salt?
por: Shayegh, Behzad, et al.
Publicado: (2026)
por: Shayegh, Behzad, et al.
Publicado: (2026)
Stochastic Chameleons: Irrelevant Context Hallucinations Reveal Class-Based (Mis)Generalization in LLMs
por: Cheng, Ziling, et al.
Publicado: (2025)
por: Cheng, Ziling, et al.
Publicado: (2025)
Neither Valid nor Reliable? Investigating the Use of LLMs as Judges
por: Chehbouni, Khaoula, et al.
Publicado: (2025)
por: Chehbouni, Khaoula, et al.
Publicado: (2025)
Revisiting Intermediate-Layer Matching in Knowledge Distillation: Layer-Selection Strategy Doesn't Matter (Much)
por: Yu, Zony, et al.
Publicado: (2025)
por: Yu, Zony, et al.
Publicado: (2025)
Revisiting Structured Sentiment Analysis as Latent Dependency Graph Parsing
por: Zhou, Chengjie, et al.
Publicado: (2024)
por: Zhou, Chengjie, et al.
Publicado: (2024)
Self-Error-Instruct: Generalizing from Errors for LLMs Mathematical Reasoning
por: Yu, Erxin, et al.
Publicado: (2025)
por: Yu, Erxin, et al.
Publicado: (2025)
Zero-Shot Continuous Prompt Transfer: Generalizing Task Semantics Across Language Models
por: Wu, Zijun, et al.
Publicado: (2023)
por: Wu, Zijun, et al.
Publicado: (2023)
LLMR: Knowledge Distillation with a Large Language Model-Induced Reward
por: Li, Dongheng, et al.
Publicado: (2024)
por: Li, Dongheng, et al.
Publicado: (2024)
Towards Automatic Error Recovery in Parsing Expression
por: de Medeiros, Sérgio Queiroz, et al.
Publicado: (2025)
por: de Medeiros, Sérgio Queiroz, et al.
Publicado: (2025)
ErrorMap and ErrorAtlas: Charting the Failure Landscape of Large Language Models
por: Ashury-Tahan, Shir, et al.
Publicado: (2026)
por: Ashury-Tahan, Shir, et al.
Publicado: (2026)
Subtle Errors in Reasoning: Preference Learning via Error-injected Self-editing
por: Xu, Kaishuai, et al.
Publicado: (2024)
por: Xu, Kaishuai, et al.
Publicado: (2024)
Confident in a Confidence Score: Investigating the Sensitivity of Confidence Scores to Supervised Fine-Tuning
por: Flores, Lorenzo Jaime Yu, et al.
Publicado: (2026)
por: Flores, Lorenzo Jaime Yu, et al.
Publicado: (2026)
Challenges to Evaluating the Generalization of Coreference Resolution Models: A Measurement Modeling Perspective
por: Porada, Ian, et al.
Publicado: (2023)
por: Porada, Ian, et al.
Publicado: (2023)
Can LLMs Reason Abstractly Over Math Word Problems Without CoT? Disentangling Abstract Formulation From Arithmetic Computation
por: Cheng, Ziling, et al.
Publicado: (2025)
por: Cheng, Ziling, et al.
Publicado: (2025)
Mind the Error! Detection and Localization of Instruction Errors in Vision-and-Language Navigation
por: Taioli, Francesco, et al.
Publicado: (2024)
por: Taioli, Francesco, et al.
Publicado: (2024)
Empirical Analysis for Unsupervised Universal Dependency Parse Tree Aggregation
por: Kulkarni, Adithya, et al.
Publicado: (2024)
por: Kulkarni, Adithya, et al.
Publicado: (2024)
Concurrent Linguistic Error Detection (CLED): a New Methodology for Error Detection in Large Language Models
por: Zhu, Jinhua, et al.
Publicado: (2024)
por: Zhu, Jinhua, et al.
Publicado: (2024)
Speak & Spell: LLM-Driven Controllable Phonetic Error Augmentation for Robust Dialogue State Tracking
por: Lee, Jihyun, et al.
Publicado: (2024)
por: Lee, Jihyun, et al.
Publicado: (2024)
Cross-lingual Back-Parsing: Utterance Synthesis from Meaning Representation for Zero-Resource Semantic Parsing
por: Kang, Deokhyung, et al.
Publicado: (2024)
por: Kang, Deokhyung, et al.
Publicado: (2024)
Error Typing for Smarter Rewards: Improving Process Reward Models with Error-Aware Hierarchical Supervision
por: Pala, Tej Deep, et al.
Publicado: (2025)
por: Pala, Tej Deep, et al.
Publicado: (2025)
IMPARA-GED: Grammatical Error Detection is Boosting Reference-free Grammatical Error Quality Estimator
por: Sakai, Yusuke, et al.
Publicado: (2025)
por: Sakai, Yusuke, et al.
Publicado: (2025)
Ejemplares similares
-
Ensemble Distillation for Unsupervised Constituency Parsing
por: Shayegh, Behzad, et al.
Publicado: (2023) -
Tree-Averaging Algorithms for Ensemble-Based Unsupervised Discontinuous Constituency Parsing
por: Shayegh, Behzad, et al.
Publicado: (2024) -
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
por: Wen, Yuqiao, et al.
Publicado: (2024) -
PreSumm: Predicting Summarization Performance Without Summarizing
por: Koniaev, Steven, et al.
Publicado: (2025) -
Feeding Two Birds or Favoring One? Adequacy-Fluency Tradeoffs in Evaluation and Meta-Evaluation of Machine Translation
por: Shayegh, Behzad, et al.
Publicado: (2025)