Coconstructions in spoken data: UD annotation guidelines and first results
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pannitto, Ludovica, Kahane, Sylvain, Dobrovoljc, Kaja, Battaglia, Elena, Guillaume, Bruno, Mauri, Caterina, Zucchini, Eleonora |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The KIPARLA Forest treebank of spoken Italian: an overview of initial design choices
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2024)
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2024)
Towards the first UD Treebank of Spoken Italian: the KIParla forest
von: Pannitto, Ludovica
Veröffentlicht: (2024)
von: Pannitto, Ludovica
Veröffentlicht: (2024)
Is Semi-Automatic Transcription Useful in Corpus Creation? Preliminary Considerations on the KIParla Corpus
von: Simonotti, Martina, et al.
Veröffentlicht: (2026)
von: Simonotti, Martina, et al.
Veröffentlicht: (2026)
Counting trees: A treebank-driven exploration of syntactic variation in speech and writing across languages
von: Dobrovoljc, Kaja
Veröffentlicht: (2025)
von: Dobrovoljc, Kaja
Veröffentlicht: (2025)
Annotating Constructions with UD: the experience of the Italian Constructicon
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2024)
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2024)
Recurrent babbling: evaluating the acquisition of grammar from limited input data
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2020)
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2020)
Linguistic Characteristics of AI-Generated Text: A Survey
von: Terčon, Luka, et al.
Veröffentlicht: (2025)
von: Terčon, Luka, et al.
Veröffentlicht: (2025)
CALaMo: a Constructionist Assessment of Language Models
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2023)
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2023)
Constraining constructions with WordNet: pros and cons for the semantic annotation of fillers in the Italian Constructicon
von: Pisciotta, Flavio, et al.
Veröffentlicht: (2025)
von: Pisciotta, Flavio, et al.
Veröffentlicht: (2025)
Did somebody say "Gest-IT"? A pilot exploration of multimodal data management
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2024)
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2024)
'Layer su Layer': Identifying and Disambiguating the Italian NPN Construction in BERT's family
von: Gorzoni, Greta, et al.
Veröffentlicht: (2026)
von: Gorzoni, Greta, et al.
Veröffentlicht: (2026)
Evaluating Metalinguistic Knowledge in Large Language Models across the World's Languages
von: Arčon, Tjaša, et al.
Veröffentlicht: (2026)
von: Arčon, Tjaša, et al.
Veröffentlicht: (2026)
Tracking Semantic Change in Slovene: A Novel Dataset and Optimal Transport-Based Distance
von: Pranjić, Marko, et al.
Veröffentlicht: (2024)
von: Pranjić, Marko, et al.
Veröffentlicht: (2024)
Towards Corpus-Grounded Agentic LLMs for Multilingual Grammatical Analysis
von: Klemen, Matej, et al.
Veröffentlicht: (2025)
von: Klemen, Matej, et al.
Veröffentlicht: (2025)
Sparse Logistic Regression with High-order Features for Automatic Grammar Rule Extraction from Treebanks
von: Herrera, Santiago, et al.
Veröffentlicht: (2024)
von: Herrera, Santiago, et al.
Veröffentlicht: (2024)
Parser agreement and disagreement in L2 Korean UD: Implications for human-in-the-loop annotation
von: Sung, Hakyung, et al.
Veröffentlicht: (2026)
von: Sung, Hakyung, et al.
Veröffentlicht: (2026)
Syntaxe théorique et formelle
von: Kahane, Sylvain, et al.
Veröffentlicht: (2023)
von: Kahane, Sylvain, et al.
Veröffentlicht: (2023)
A UD Treebank for Bohairic Coptic
von: Zeldes, Amir, et al.
Veröffentlicht: (2025)
von: Zeldes, Amir, et al.
Veröffentlicht: (2025)
Building UD Cairo for Old English in the Classroom
von: Levine, Lauren, et al.
Veröffentlicht: (2025)
von: Levine, Lauren, et al.
Veröffentlicht: (2025)
K-UD: Revising Korean Universal Dependencies Guidelines
von: Kim, Kyuwon, et al.
Veröffentlicht: (2024)
von: Kim, Kyuwon, et al.
Veröffentlicht: (2024)
Aligning the Norwegian UD Treebank with Entity and Coreference Information
von: Jørgensen, Tollef Emil, et al.
Veröffentlicht: (2023)
von: Jørgensen, Tollef Emil, et al.
Veröffentlicht: (2023)
Universal NER v2: Towards a Massively Multilingual Named Entity Recognition Benchmark
von: Blevins, Terra, et al.
Veröffentlicht: (2026)
von: Blevins, Terra, et al.
Veröffentlicht: (2026)
Out-of-distribution generalisation in spoken language understanding
von: Porjazovski, Dejan, et al.
Veröffentlicht: (2024)
von: Porjazovski, Dejan, et al.
Veröffentlicht: (2024)
Exploring Multiple Strategies to Improve Multilingual Coreference Resolution in CorefUD
von: Pražák, Ondřej, et al.
Veröffentlicht: (2024)
von: Pražák, Ondřej, et al.
Veröffentlicht: (2024)
Lost in Speech: Benchmarking, Evaluation, and Parsing of Spoken Code-Switching Beyond Standard UD Assumptions
von: Tyagi, Nemika, et al.
Veröffentlicht: (2026)
von: Tyagi, Nemika, et al.
Veröffentlicht: (2026)
Encoding of lexical tone in self-supervised models of spoken language
von: Shen, Gaofei, et al.
Veröffentlicht: (2024)
von: Shen, Gaofei, et al.
Veröffentlicht: (2024)
Empirical evidence of Large Language Model's influence on human spoken communication
von: Yakura, Hiromu, et al.
Veröffentlicht: (2024)
von: Yakura, Hiromu, et al.
Veröffentlicht: (2024)
Scaling few-shot spoken word classification with generative meta-continual learning
von: Beyers, Louise, et al.
Veröffentlicht: (2026)
von: Beyers, Louise, et al.
Veröffentlicht: (2026)
Extracting accent features in spoken Brazilian Portuguese without sociolinguistic labels
von: Leite, Pedro H. L., et al.
Veröffentlicht: (2026)
von: Leite, Pedro H. L., et al.
Veröffentlicht: (2026)
The UD-NewsCrawl Treebank: Reflections and Challenges from a Large-scale Tagalog Syntactic Annotation Project
von: Aquino, Angelina A., et al.
Veröffentlicht: (2025)
von: Aquino, Angelina A., et al.
Veröffentlicht: (2025)
Exposing propaganda: an analysis of stylistic cues comparing human annotations and machine classification
von: Faye, Géraud, et al.
Veröffentlicht: (2024)
von: Faye, Géraud, et al.
Veröffentlicht: (2024)
Tailoring AI-Driven Reading Scaffolds to the Distinct Needs of Neurodiverse Learners
von: Jhilal, Soufiane, et al.
Veröffentlicht: (2026)
von: Jhilal, Soufiane, et al.
Veröffentlicht: (2026)
The realization of tones in spontaneous spoken Taiwan Mandarin: a corpus-based survey and theory-driven computational modeling
von: Lu, Yuxin, et al.
Veröffentlicht: (2025)
von: Lu, Yuxin, et al.
Veröffentlicht: (2025)
Parsing the Switch: LLM-Based UD Annotation for Complex Code-Switched and Low-Resource Languages
von: Kellert, Olga, et al.
Veröffentlicht: (2025)
von: Kellert, Olga, et al.
Veröffentlicht: (2025)
LLM_annotate: A Python package for annotating and analyzing fiction characters
von: Rosenbusch, Hannes
Veröffentlicht: (2025)
von: Rosenbusch, Hannes
Veröffentlicht: (2025)
Overview of MWE history, challenges, and horizons: standing at the 20th anniversary of the MWE workshop series via MWE-UD2024
von: Han, Lifeng, et al.
Veröffentlicht: (2024)
von: Han, Lifeng, et al.
Veröffentlicht: (2024)
UD-KSL Treebank v1.3: A semi-automated framework for aligning XPOS-extracted units with UPOS tags
von: Sung, Hakyung, et al.
Veröffentlicht: (2025)
von: Sung, Hakyung, et al.
Veröffentlicht: (2025)
Does language matter for spoken word classification? A multilingual generative meta-learning approach
von: Ziki, Batsirayi Mupamhi, et al.
Veröffentlicht: (2026)
von: Ziki, Batsirayi Mupamhi, et al.
Veröffentlicht: (2026)
LLMs for automatic annotation of Mandarin narrative transcripts
von: Zhao, Qingwen, et al.
Veröffentlicht: (2026)
von: Zhao, Qingwen, et al.
Veröffentlicht: (2026)
Recovering document annotations for sentence-level bitext
von: Wicks, Rachel, et al.
Veröffentlicht: (2024)
von: Wicks, Rachel, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
The KIPARLA Forest treebank of spoken Italian: an overview of initial design choices
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2024) -
Towards the first UD Treebank of Spoken Italian: the KIParla forest
von: Pannitto, Ludovica
Veröffentlicht: (2024) -
Is Semi-Automatic Transcription Useful in Corpus Creation? Preliminary Considerations on the KIParla Corpus
von: Simonotti, Martina, et al.
Veröffentlicht: (2026) -
Counting trees: A treebank-driven exploration of syntactic variation in speech and writing across languages
von: Dobrovoljc, Kaja
Veröffentlicht: (2025) -
Annotating Constructions with UD: the experience of the Italian Constructicon
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2024)