GUMBridge: a Corpus for Varieties of Bridging Anaphora
Fuente:
arXiv
Saved in:
| Main Authors: | Levine, Lauren, Zeldes, Amir |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Subjectivity in the Annotation of Bridging Anaphora
by: Levine, Lauren, et al.
Published: (2025)
by: Levine, Lauren, et al.
Published: (2025)
Unifying the Scope of Bridging Anaphora Types in English: Bridging Annotations in ARRAU and GUM
by: Levine, Lauren, et al.
Published: (2024)
by: Levine, Lauren, et al.
Published: (2024)
LLMBridge: An LLM Pipeline for End-to-end Referential Bridging Resolution in English
by: Levine, Lauren, et al.
Published: (2026)
by: Levine, Lauren, et al.
Published: (2026)
Lacuna Language Learning: Leveraging RNNs for Ranked Text Completion in Digitized Coptic Manuscripts
by: Levine, Lauren, et al.
Published: (2024)
by: Levine, Lauren, et al.
Published: (2024)
Dialect Matters: Cross-Lingual ASR Transfer for Low-Resource Indic Language Varieties
by: Dhasmana, Akriti, et al.
Published: (2026)
by: Dhasmana, Akriti, et al.
Published: (2026)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
by: Ashuach, Tomer, et al.
Published: (2025)
by: Ashuach, Tomer, et al.
Published: (2025)
Is Textual Similarity Invariant under Machine Translation? Evidence Based on the Political Manifesto Corpus
by: Boratyn, Daria, et al.
Published: (2026)
by: Boratyn, Daria, et al.
Published: (2026)
The GDN-CC Dataset: Automatic Corpus Clarification for AI-enhanced Democratic Citizen Consultations
by: Lequeu, Pierre-Antoine, et al.
Published: (2026)
by: Lequeu, Pierre-Antoine, et al.
Published: (2026)
Character-Level Transformer for Tajik-Persian Transliteration with a Parallel Lexical Corpus
by: Arabov, Mullosharaf K.
Published: (2026)
by: Arabov, Mullosharaf K.
Published: (2026)
IWLV-Ramayana: A Sarga-Aligned Parallel Corpus of Valmiki's Ramayana Across Indian Languages
by: VP, Sumesh
Published: (2026)
by: VP, Sumesh
Published: (2026)
ELCC: the Emergent Language Corpus Collection
by: Boldt, Brendon, et al.
Published: (2024)
by: Boldt, Brendon, et al.
Published: (2024)
The Knesset Corpus: An Annotated Corpus of Hebrew Parliamentary Proceedings
by: Goldin, Gili, et al.
Published: (2024)
by: Goldin, Gili, et al.
Published: (2024)
Clinical Document Corpora -- Real Ones, Translated and Synthetic Substitutes, and Assorted Domain Proxies: A Survey of Diversity in Corpus Design, with Focus on German Text Data
by: Hahn, Udo
Published: (2024)
by: Hahn, Udo
Published: (2024)
COSTAR-A: A prompting framework for enhancing Large Language Model performance on Point-of-View questions
by: Ohalete, Nzubechukwu C., et al.
Published: (2025)
by: Ohalete, Nzubechukwu C., et al.
Published: (2025)
Corpus Considerations for Annotator Modeling and Scaling
by: Sarumi, Olufunke O., et al.
Published: (2024)
by: Sarumi, Olufunke O., et al.
Published: (2024)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
by: Saji, Alan, et al.
Published: (2025)
by: Saji, Alan, et al.
Published: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
by: Peters, Sydney, et al.
Published: (2025)
by: Peters, Sydney, et al.
Published: (2025)
Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
by: Yang, Shan
Published: (2026)
by: Yang, Shan
Published: (2026)
Charting a Decade of Computational Linguistics in Italy: The CLiC-it Corpus
by: Alzetta, Chiara, et al.
Published: (2025)
by: Alzetta, Chiara, et al.
Published: (2025)
Bridging the Gap: An Intermediate Language for Enhanced and Cost-Effective Grapheme-to-Phoneme Conversion with Homographs with Multiple Pronunciations Disambiguation
by: Bertina, Abbas, et al.
Published: (2025)
by: Bertina, Abbas, et al.
Published: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
by: Collado-Montañez, Jaime, et al.
Published: (2025)
by: Collado-Montañez, Jaime, et al.
Published: (2025)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
by: Smădu, Răzvan-Alexandru, et al.
Published: (2025)
by: Smădu, Răzvan-Alexandru, et al.
Published: (2025)
Historical Ink: 19th Century Latin American Spanish Newspaper Corpus with LLM OCR Correction
by: Manrique-Gómez, Laura, et al.
Published: (2024)
by: Manrique-Gómez, Laura, et al.
Published: (2024)
Graphemic Normalization of the Perso-Arabic Script
by: Doctor, Raiomond, et al.
Published: (2022)
by: Doctor, Raiomond, et al.
Published: (2022)
Beyond Arabic: Software for Perso-Arabic Script Manipulation
by: Gutkin, Alexander, et al.
Published: (2023)
by: Gutkin, Alexander, et al.
Published: (2023)
CorpusStudio: Surfacing Emergent Patterns in a Corpus of Prior Work while Writing
by: Dang, Hai, et al.
Published: (2025)
by: Dang, Hai, et al.
Published: (2025)
Recent Trends in Linear Text Segmentation: a Survey
by: Ghinassi, Iacopo, et al.
Published: (2024)
by: Ghinassi, Iacopo, et al.
Published: (2024)
Morphological Analysis for the Maltese Language: The Challenges of a Hybrid System
by: Borg, Claudia, et al.
Published: (2017)
by: Borg, Claudia, et al.
Published: (2017)
KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation
by: Nzeyimana, Antoine, et al.
Published: (2025)
by: Nzeyimana, Antoine, et al.
Published: (2025)
Low-resource neural machine translation with morphological modeling
by: Nzeyimana, Antoine
Published: (2024)
by: Nzeyimana, Antoine
Published: (2024)
Linguistic Interpretability of Transformer-based Language Models: a systematic review
by: López-Otal, Miguel, et al.
Published: (2025)
by: López-Otal, Miguel, et al.
Published: (2025)
Rubrik's Cube: Testing a New Rubric for Evaluating Explanations on the CUBE dataset
by: Galvan-Sosa, Diana, et al.
Published: (2025)
by: Galvan-Sosa, Diana, et al.
Published: (2025)
MALT: Mechanistic Ablation of Lossy Translation in LLMs for a Low-Resource Language: Urdu
by: Bajwa, Taaha Saleem
Published: (2025)
by: Bajwa, Taaha Saleem
Published: (2025)
MIX : a Multi-task Learning Approach to Solve Open-Domain Question Answering
by: Chaybouti, Sofian, et al.
Published: (2020)
by: Chaybouti, Sofian, et al.
Published: (2020)
EfficientQA : a RoBERTa Based Phrase-Indexed Question-Answering System
by: Chaybouti, Sofian, et al.
Published: (2021)
by: Chaybouti, Sofian, et al.
Published: (2021)
I run as fast as a rabbit, can you? A Multilingual Simile Dialogue Dataset
by: Ma, Longxuan, et al.
Published: (2023)
by: Ma, Longxuan, et al.
Published: (2023)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
by: Oketunji, Abiodun Finbarrs
Published: (2023)
by: Oketunji, Abiodun Finbarrs
Published: (2023)
What makes a language easy to deep-learn? Deep neural networks and humans similarly benefit from compositional structure
by: Galke, Lukas, et al.
Published: (2023)
by: Galke, Lukas, et al.
Published: (2023)
The Pragmatic Persona: Discovering LLM Persona through Bridging Inference
by: Yang, Jisoo, et al.
Published: (2026)
by: Yang, Jisoo, et al.
Published: (2026)
DESS: DeBERTa Enhanced Syntactic-Semantic Aspect Sentiment Triplet Extraction
by: Thenuwara, Vishal, et al.
Published: (2025)
by: Thenuwara, Vishal, et al.
Published: (2025)
Similar Items
-
Subjectivity in the Annotation of Bridging Anaphora
by: Levine, Lauren, et al.
Published: (2025) -
Unifying the Scope of Bridging Anaphora Types in English: Bridging Annotations in ARRAU and GUM
by: Levine, Lauren, et al.
Published: (2024) -
LLMBridge: An LLM Pipeline for End-to-end Referential Bridging Resolution in English
by: Levine, Lauren, et al.
Published: (2026) -
Lacuna Language Learning: Leveraging RNNs for Ranked Text Completion in Digitized Coptic Manuscripts
by: Levine, Lauren, et al.
Published: (2024) -
Dialect Matters: Cross-Lingual ASR Transfer for Low-Resource Indic Language Varieties
by: Dhasmana, Akriti, et al.
Published: (2026)