TalkTag: Fine-Grained Morphosyntactic Error Annotation for Transcribed Speech
Fuente:
arXiv
Saved in:
| Main Authors: | Venturini, Shamira, Hennhöfer, Oliver, Kinkel, Steffen, Strötgen, Jannik |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Arabic Morphosyntactic Tagging and Dependency Parsing with Large Language Models
by: Adel, Mohamed, et al.
Published: (2026)
by: Adel, Mohamed, et al.
Published: (2026)
Discourse-Aware In-Context Learning for Temporal Expression Normalization
by: Gautam, Akash Kumar, et al.
Published: (2024)
by: Gautam, Akash Kumar, et al.
Published: (2024)
Speech LLMs are Contextual Reasoning Transcribers
by: Deng, Keqi, et al.
Published: (2026)
by: Deng, Keqi, et al.
Published: (2026)
NLNDE at SemEval-2023 Task 12: Adaptive Pretraining and Source Language Selection for Low-Resource Multilingual Sentiment Analysis
by: Wang, Mingyang, et al.
Published: (2023)
by: Wang, Mingyang, et al.
Published: (2023)
Morphosyntactic Analysis for CHILDES
by: Liu, Houjun, et al.
Published: (2024)
by: Liu, Houjun, et al.
Published: (2024)
Killkan: The Automatic Speech Recognition Dataset for Kichwa with Morphosyntactic Information
by: Taguchi, Chihiro, et al.
Published: (2024)
by: Taguchi, Chihiro, et al.
Published: (2024)
Better Call SAUL: Fluent and Consistent Language Model Editing with Generation Regularization
by: Wang, Mingyang, et al.
Published: (2024)
by: Wang, Mingyang, et al.
Published: (2024)
Learn it or Leave it: Module Composition and Pruning for Continual Learning
by: Wang, Mingyang, et al.
Published: (2024)
by: Wang, Mingyang, et al.
Published: (2024)
Rehearsal-Free Modular and Compositional Continual Learning for Language Models
by: Wang, Mingyang, et al.
Published: (2024)
by: Wang, Mingyang, et al.
Published: (2024)
Language Mixing in Reasoning Language Models: Patterns, Impact, and Internal Causes
by: Wang, Mingyang, et al.
Published: (2025)
by: Wang, Mingyang, et al.
Published: (2025)
Bring Your Own Knowledge: A Survey of Methods for LLM Knowledge Expansion
by: Wang, Mingyang, et al.
Published: (2025)
by: Wang, Mingyang, et al.
Published: (2025)
Morphosyntactic probing of multilingual BERT models
by: Acs, Judit, et al.
Published: (2023)
by: Acs, Judit, et al.
Published: (2023)
TagSpeech: End-to-End Multi-Speaker ASR and Diarization with Fine-Grained Temporal Grounding
by: Huo, Mingyue, et al.
Published: (2026)
by: Huo, Mingyue, et al.
Published: (2026)
TOGGL: Transcribing Overlapping Speech with Staggered Labeling
by: Li, Chak-Fai, et al.
Published: (2024)
by: Li, Chak-Fai, et al.
Published: (2024)
Lost in Multilinguality: Dissecting Cross-lingual Factual Inconsistency in Transformer Language Models
by: Wang, Mingyang, et al.
Published: (2025)
by: Wang, Mingyang, et al.
Published: (2025)
Enhancing Korean Dependency Parsing with Morphosyntactic Features
by: Park, Jungyeul, et al.
Published: (2025)
by: Park, Jungyeul, et al.
Published: (2025)
Transcribing and Translating, Fast and Slow: Joint Speech Translation and Recognition
by: Moritz, Niko, et al.
Published: (2024)
by: Moritz, Niko, et al.
Published: (2024)
MedErrBench: A Fine-Grained Multilingual Benchmark for Medical Error Detection and Correction with Clinical Expert Annotations
by: Ma, Congbo, et al.
Published: (2026)
by: Ma, Congbo, et al.
Published: (2026)
ARCADE: A City-Scale Corpus for Fine-Grained Arabic Dialect Tagging
by: Nacar, Omer, et al.
Published: (2026)
by: Nacar, Omer, et al.
Published: (2026)
A State-of-the-Art Morphosyntactic Parser and Lemmatizer for Ancient Greek
by: Celano, Giuseppe G. A.
Published: (2024)
by: Celano, Giuseppe G. A.
Published: (2024)
Large Language Models as Automatic Annotators and Annotation Adjudicators for Fine-Grained Opinion Analysis
by: Negi, Gaurav, et al.
Published: (2026)
by: Negi, Gaurav, et al.
Published: (2026)
Kallaama: A Transcribed Speech Dataset about Agriculture in the Three Most Widely Spoken Languages in Senegal
by: Gauthier, Elodie, et al.
Published: (2024)
by: Gauthier, Elodie, et al.
Published: (2024)
ÚFAL LatinPipe at EvaLatin 2024: Morphosyntactic Analysis of Latin
by: Straka, Milan, et al.
Published: (2024)
by: Straka, Milan, et al.
Published: (2024)
From Tags to Trees: Structuring Fine-Grained Knowledge for Controllable Data Selection in LLM Instruction Tuning
by: Niu, Zihan, et al.
Published: (2026)
by: Niu, Zihan, et al.
Published: (2026)
DiDOTS: Knowledge Distillation from Large-Language-Models for Dementia Obfuscation in Transcribed Speech
by: Woszczyk, Dominika, et al.
Published: (2024)
by: Woszczyk, Dominika, et al.
Published: (2024)
Talk, Snap, Complain: Validation-Aware Multimodal Expert Framework for Fine-Grained Customer Grievances
by: Singh, Rishu Kumar, et al.
Published: (2025)
by: Singh, Rishu Kumar, et al.
Published: (2025)
To Diverge or Not to Diverge: A Morphosyntactic Perspective on Machine Translation vs Human Translation
by: Luo, Jiaming, et al.
Published: (2024)
by: Luo, Jiaming, et al.
Published: (2024)
Fine-Grained Perspectives: Modeling Explanations with Annotator-Specific Rationales
by: Sarumi, Olufunke O., et al.
Published: (2026)
by: Sarumi, Olufunke O., et al.
Published: (2026)
Intent Matters: Enhancing AI Tutoring with Fine-Grained Pedagogical Intent Annotation
by: Petukhova, Kseniia, et al.
Published: (2025)
by: Petukhova, Kseniia, et al.
Published: (2025)
Rubato: Transcribing Piano Music with Timestamps
by: Tamer, Nazif Can, et al.
Published: (2026)
by: Tamer, Nazif Can, et al.
Published: (2026)
Acquiring Pronunciation Knowledge from Transcribed Speech Audio via Multi-task Learning
by: Sun, Siqi, et al.
Published: (2024)
by: Sun, Siqi, et al.
Published: (2024)
Morphosyntactic Variation in Medieval Celtic Languages
Published: (2021)
Published: (2021)
Zero Resource Cross-Lingual Part Of Speech Tagging
by: Chopra, Sahil
Published: (2024)
by: Chopra, Sahil
Published: (2024)
Multi-head Sequence Tagging Model for Grammatical Error Correction
by: Al-Sabahi, Kamal, et al.
Published: (2024)
by: Al-Sabahi, Kamal, et al.
Published: (2024)
Fine-Grained Reward Optimization for Machine Translation using Error Severity Mappings
by: Ramos, Miguel Moura, et al.
Published: (2024)
by: Ramos, Miguel Moura, et al.
Published: (2024)
DeFine: A Decomposed and Fine-Grained Annotated Dataset for Long-form Article Generation
by: Wang, Ming, et al.
Published: (2025)
by: Wang, Ming, et al.
Published: (2025)
GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio
by: Chen, Guoguo, et al.
Published: (2021)
by: Chen, Guoguo, et al.
Published: (2021)
Open-Source Web Service with Morphological Dictionary-Supplemented Deep Learning for Morphosyntactic Analysis of Czech
by: Straka, Milan, et al.
Published: (2024)
by: Straka, Milan, et al.
Published: (2024)
Large Language Model Can Transcribe Speech in Multi-Talker Scenarios with Versatile Instructions
by: Meng, Lingwei, et al.
Published: (2024)
by: Meng, Lingwei, et al.
Published: (2024)
FEANEL: A Benchmark for Fine-Grained Error Analysis in K-12 English Writing
by: Ye, Jingheng, et al.
Published: (2025)
by: Ye, Jingheng, et al.
Published: (2025)
Similar Items
-
Arabic Morphosyntactic Tagging and Dependency Parsing with Large Language Models
by: Adel, Mohamed, et al.
Published: (2026) -
Discourse-Aware In-Context Learning for Temporal Expression Normalization
by: Gautam, Akash Kumar, et al.
Published: (2024) -
Speech LLMs are Contextual Reasoning Transcribers
by: Deng, Keqi, et al.
Published: (2026) -
NLNDE at SemEval-2023 Task 12: Adaptive Pretraining and Source Language Selection for Low-Resource Multilingual Sentiment Analysis
by: Wang, Mingyang, et al.
Published: (2023) -
Morphosyntactic Analysis for CHILDES
by: Liu, Houjun, et al.
Published: (2024)