The KIPARLA Forest treebank of spoken Italian: an overview of initial design choices
Fuente:
arXiv
Guardado en:
| Autores principales: | Pannitto, Ludovica, Mauri, Caterina |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards the first UD Treebank of Spoken Italian: the KIParla forest
por: Pannitto, Ludovica
Publicado: (2024)
por: Pannitto, Ludovica
Publicado: (2024)
Coconstructions in spoken data: UD annotation guidelines and first results
por: Pannitto, Ludovica, et al.
Publicado: (2026)
por: Pannitto, Ludovica, et al.
Publicado: (2026)
Is Semi-Automatic Transcription Useful in Corpus Creation? Preliminary Considerations on the KIParla Corpus
por: Simonotti, Martina, et al.
Publicado: (2026)
por: Simonotti, Martina, et al.
Publicado: (2026)
'Layer su Layer': Identifying and Disambiguating the Italian NPN Construction in BERT's family
por: Gorzoni, Greta, et al.
Publicado: (2026)
por: Gorzoni, Greta, et al.
Publicado: (2026)
Recurrent babbling: evaluating the acquisition of grammar from limited input data
por: Pannitto, Ludovica, et al.
Publicado: (2020)
por: Pannitto, Ludovica, et al.
Publicado: (2020)
CALaMo: a Constructionist Assessment of Language Models
por: Pannitto, Ludovica, et al.
Publicado: (2023)
por: Pannitto, Ludovica, et al.
Publicado: (2023)
Constraining constructions with WordNet: pros and cons for the semantic annotation of fillers in the Italian Constructicon
por: Pisciotta, Flavio, et al.
Publicado: (2025)
por: Pisciotta, Flavio, et al.
Publicado: (2025)
Annotating Constructions with UD: the experience of the Italian Constructicon
por: Pannitto, Ludovica, et al.
Publicado: (2024)
por: Pannitto, Ludovica, et al.
Publicado: (2024)
Did somebody say "Gest-IT"? A pilot exploration of multimodal data management
por: Pannitto, Ludovica, et al.
Publicado: (2024)
por: Pannitto, Ludovica, et al.
Publicado: (2024)
Punctuation-aware treebank tree binarization
por: Klinger, Eitan, et al.
Publicado: (2025)
por: Klinger, Eitan, et al.
Publicado: (2025)
Construction and educational application of a linguistically grounded dependency treebank for Uyghur
por: Zuo, Jiaxin, et al.
Publicado: (2025)
por: Zuo, Jiaxin, et al.
Publicado: (2025)
Counting trees: A treebank-driven exploration of syntactic variation in speech and writing across languages
por: Dobrovoljc, Kaja
Publicado: (2025)
por: Dobrovoljc, Kaja
Publicado: (2025)
Second language Korean Universal Dependency treebank v1.2: Focus on data augmentation and annotation scheme refinement
por: Sung, Hakyung, et al.
Publicado: (2025)
por: Sung, Hakyung, et al.
Publicado: (2025)
How Humans and LLMs Organize Conceptual Knowledge: Exploring Subordinate Categories in Italian
por: Pedrotti, Andrea, et al.
Publicado: (2025)
por: Pedrotti, Andrea, et al.
Publicado: (2025)
Out-of-distribution generalisation in spoken language understanding
por: Porjazovski, Dejan, et al.
Publicado: (2024)
por: Porjazovski, Dejan, et al.
Publicado: (2024)
Encoding of lexical tone in self-supervised models of spoken language
por: Shen, Gaofei, et al.
Publicado: (2024)
por: Shen, Gaofei, et al.
Publicado: (2024)
Empirical evidence of Large Language Model's influence on human spoken communication
por: Yakura, Hiromu, et al.
Publicado: (2024)
por: Yakura, Hiromu, et al.
Publicado: (2024)
Scaling few-shot spoken word classification with generative meta-continual learning
por: Beyers, Louise, et al.
Publicado: (2026)
por: Beyers, Louise, et al.
Publicado: (2026)
Towards an empirical understanding of MoE design choices
por: Fan, Dongyang, et al.
Publicado: (2024)
por: Fan, Dongyang, et al.
Publicado: (2024)
Extracting accent features in spoken Brazilian Portuguese without sociolinguistic labels
por: Leite, Pedro H. L., et al.
Publicado: (2026)
por: Leite, Pedro H. L., et al.
Publicado: (2026)
The realization of tones in spontaneous spoken Taiwan Mandarin: a corpus-based survey and theory-driven computational modeling
por: Lu, Yuxin, et al.
Publicado: (2025)
por: Lu, Yuxin, et al.
Publicado: (2025)
Layers of technology in pluriversal design. Decolonising language technology with the LiveLanguage initiative
por: Koch, Gertraud, et al.
Publicado: (2024)
por: Koch, Gertraud, et al.
Publicado: (2024)
Does language matter for spoken word classification? A multilingual generative meta-learning approach
por: Ziki, Batsirayi Mupamhi, et al.
Publicado: (2026)
por: Ziki, Batsirayi Mupamhi, et al.
Publicado: (2026)
PejorativITy: Disambiguating Pejorative Epithets to Improve Misogyny Detection in Italian Tweets
por: Muti, Arianna, et al.
Publicado: (2024)
por: Muti, Arianna, et al.
Publicado: (2024)
Quantizing Whisper-small: How design choices affect ASR performance
por: Söhler, Arthur, et al.
Publicado: (2025)
por: Söhler, Arthur, et al.
Publicado: (2025)
Gujarati-English Code-Switching Speech Recognition using ensemble prediction of spoken language
por: Sharma, Yash, et al.
Publicado: (2024)
por: Sharma, Yash, et al.
Publicado: (2024)
Detecting the terminality of speech-turn boundary for spoken interactions in French TV and Radio content
por: Uro, Rémi, et al.
Publicado: (2024)
por: Uro, Rémi, et al.
Publicado: (2024)
Transfer Learning Enhanced Single-choice Decision for Multi-choice Question Answering
por: Cui, Chenhao, et al.
Publicado: (2024)
por: Cui, Chenhao, et al.
Publicado: (2024)
Is one brick enough to break the wall of spoken dialogue state tracking?
por: Druart, Lucas, et al.
Publicado: (2023)
por: Druart, Lucas, et al.
Publicado: (2023)
Optimizing the role of human evaluation in LLM-based spoken document summarization systems
por: Kroll, Margaret, et al.
Publicado: (2024)
por: Kroll, Margaret, et al.
Publicado: (2024)
BabySLM: language-acquisition-friendly benchmark of self-supervised spoken language models
por: Lavechin, Marvin, et al.
Publicado: (2023)
por: Lavechin, Marvin, et al.
Publicado: (2023)
BAMBI: Developing Baby Language Models for Italian
por: Suozzi, Alice, et al.
Publicado: (2025)
por: Suozzi, Alice, et al.
Publicado: (2025)
IMB: An Italian Medical Benchmark for Question Answering
por: Romano, Antonio, et al.
Publicado: (2025)
por: Romano, Antonio, et al.
Publicado: (2025)
An overview of artificial intelligence in computer-assisted language learning
por: Katinskaia, Anisia
Publicado: (2025)
por: Katinskaia, Anisia
Publicado: (2025)
EventNet-ITA: Italian Frame Parsing for Events
por: Rovera, Marco
Publicado: (2023)
por: Rovera, Marco
Publicado: (2023)
Evalita-LLM: Benchmarking Large Language Models on Italian
por: Magnini, Bernardo, et al.
Publicado: (2025)
por: Magnini, Bernardo, et al.
Publicado: (2025)
Lost in the Pipeline: How Well Do Large Language Models Handle Data Preparation?
por: Spreafico, Matteo, et al.
Publicado: (2025)
por: Spreafico, Matteo, et al.
Publicado: (2025)
Domain Embeddings for Generating Complex Descriptions of Concepts in Italian Language
por: Maisto, Alessandro
Publicado: (2024)
por: Maisto, Alessandro
Publicado: (2024)
Addressing Hallucinations with RAG and NMISS in Italian Healthcare LLM Chatbots
por: Priola, Maria Paola
Publicado: (2024)
por: Priola, Maria Paola
Publicado: (2024)
IT5: Text-to-text Pretraining for Italian Language Understanding and Generation
por: Sarti, Gabriele, et al.
Publicado: (2022)
por: Sarti, Gabriele, et al.
Publicado: (2022)
Ejemplares similares
-
Towards the first UD Treebank of Spoken Italian: the KIParla forest
por: Pannitto, Ludovica
Publicado: (2024) -
Coconstructions in spoken data: UD annotation guidelines and first results
por: Pannitto, Ludovica, et al.
Publicado: (2026) -
Is Semi-Automatic Transcription Useful in Corpus Creation? Preliminary Considerations on the KIParla Corpus
por: Simonotti, Martina, et al.
Publicado: (2026) -
'Layer su Layer': Identifying and Disambiguating the Italian NPN Construction in BERT's family
por: Gorzoni, Greta, et al.
Publicado: (2026) -
Recurrent babbling: evaluating the acquisition of grammar from limited input data
por: Pannitto, Ludovica, et al.
Publicado: (2020)