Universal-2-TF: Robust All-Neural Text Formatting for ASR
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Khare, Yash, Peyash, Taufiquzzaman, Vanzo, Andrea, Yoshioka, Takuya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Benchmark of French ASR Systems Based on Error Severity
von: Tholly, Antoine, et al.
Veröffentlicht: (2025)
von: Tholly, Antoine, et al.
Veröffentlicht: (2025)
Dialect Matters: Cross-Lingual ASR Transfer for Low-Resource Indic Language Varieties
von: Dhasmana, Akriti, et al.
Veröffentlicht: (2026)
von: Dhasmana, Akriti, et al.
Veröffentlicht: (2026)
RAID: A Shared Benchmark for Robust Evaluation of Machine-Generated Text Detectors
von: Dugan, Liam, et al.
Veröffentlicht: (2024)
von: Dugan, Liam, et al.
Veröffentlicht: (2024)
Hard to Be Heard: Phoneme-Level ASR Analysis of Phonologically Complex, Low-Resource Endangered Languages
von: Akavarapu, V. S. D. S. Mahesh, et al.
Veröffentlicht: (2026)
von: Akavarapu, V. S. D. S. Mahesh, et al.
Veröffentlicht: (2026)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
Robustness of Large Language Models to Perturbations in Text
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
Revisiting Word Embeddings in the LLM Era
von: Mahajan, Yash, et al.
Veröffentlicht: (2024)
von: Mahajan, Yash, et al.
Veröffentlicht: (2024)
Emergent Lexical Semantics in Neural Language Models: Testing Martin's Law on LLM-Generated Text
von: Kugler, Kai
Veröffentlicht: (2025)
von: Kugler, Kai
Veröffentlicht: (2025)
Text Summarization With Graph Attention Networks
von: Ardestani, Mohammadreza, et al.
Veröffentlicht: (2026)
von: Ardestani, Mohammadreza, et al.
Veröffentlicht: (2026)
Active Few-Shot Learning for Text Classification
von: Ahmadnia, Saeed, et al.
Veröffentlicht: (2025)
von: Ahmadnia, Saeed, et al.
Veröffentlicht: (2025)
Investigating the Impact of Text Summarization on Topic Modeling
von: Khandelwal, Trishia
Veröffentlicht: (2024)
von: Khandelwal, Trishia
Veröffentlicht: (2024)
Normalization of Lithuanian Text Using Regular Expressions
von: Kasparaitis, Pijus
Veröffentlicht: (2023)
von: Kasparaitis, Pijus
Veröffentlicht: (2023)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
Recent Trends in Linear Text Segmentation: a Survey
von: Ghinassi, Iacopo, et al.
Veröffentlicht: (2024)
von: Ghinassi, Iacopo, et al.
Veröffentlicht: (2024)
SW-ASR: A Context-Aware Hybrid ASR Pipeline for Robust Single Word Speech Recognition
von: Sharma, Manali, et al.
Veröffentlicht: (2026)
von: Sharma, Manali, et al.
Veröffentlicht: (2026)
Enhancing Emotion Prediction in News Headlines: Insights from ChatGPT and Seq2Seq Models for Free-Text Generation
von: Gao, Ge, et al.
Veröffentlicht: (2024)
von: Gao, Ge, et al.
Veröffentlicht: (2024)
Comprehensiveness Metrics for Automatic Evaluation of Factual Recall in Text Generation
von: Dejl, Adam, et al.
Veröffentlicht: (2025)
von: Dejl, Adam, et al.
Veröffentlicht: (2025)
VertAttack: Taking advantage of Text Classifiers' horizontal vision
von: Rusert, Jonathan
Veröffentlicht: (2024)
von: Rusert, Jonathan
Veröffentlicht: (2024)
All for One: LLMs Solve Mental Math at the Last Token With Information Transferred From Other Tokens
von: Mamidanna, Siddarth, et al.
Veröffentlicht: (2025)
von: Mamidanna, Siddarth, et al.
Veröffentlicht: (2025)
UM_FHS at the CLEF 2025 SimpleText Track: Comparing No-Context and Fine-Tune Approaches for GPT-4.1 Models in Sentence and Document-Level Text Simplification
von: Kocbek, Primoz, et al.
Veröffentlicht: (2025)
von: Kocbek, Primoz, et al.
Veröffentlicht: (2025)
Yes-MT's Submission to the Low-Resource Indic Language Translation Shared Task in WMT 2024
von: Bhaskar, Yash, et al.
Veröffentlicht: (2025)
von: Bhaskar, Yash, et al.
Veröffentlicht: (2025)
Fine-Tuning LLMs with Fine-Grained Human Feedback on Text Spans
von: CH-Wang, Sky, et al.
Veröffentlicht: (2025)
von: CH-Wang, Sky, et al.
Veröffentlicht: (2025)
A RoBERTa-Based Functional Syntax Annotation Model for Chinese Texts
von: Xiaohui, Han, et al.
Veröffentlicht: (2025)
von: Xiaohui, Han, et al.
Veröffentlicht: (2025)
Algorithm for Semantic Network Generation from Texts of Low Resource Languages Such as Kiswahili
von: Wanjawa, Barack Wamkaya, et al.
Veröffentlicht: (2025)
von: Wanjawa, Barack Wamkaya, et al.
Veröffentlicht: (2025)
Unstructured Text Enhanced Open-domain Dialogue System: A Systematic Survey
von: Ma, Longxuan, et al.
Veröffentlicht: (2024)
von: Ma, Longxuan, et al.
Veröffentlicht: (2024)
Vocabulary Transfer for Biomedical Texts: Add Tokens if You Can Not Add Data
von: Singh, Priyanka, et al.
Veröffentlicht: (2022)
von: Singh, Priyanka, et al.
Veröffentlicht: (2022)
German Text Simplification: Finetuning Large Language Models with Semi-Synthetic Data
von: Klöser, Lars, et al.
Veröffentlicht: (2024)
von: Klöser, Lars, et al.
Veröffentlicht: (2024)
Ensembling Multilingual Transformers for Robust Sentiment Analysis of Tweets
von: Bilehsavar, Meysam Shirdel, et al.
Veröffentlicht: (2025)
von: Bilehsavar, Meysam Shirdel, et al.
Veröffentlicht: (2025)
Schema as Parameterized Tools for Universal Information Extraction
von: Liang, Sheng, et al.
Veröffentlicht: (2025)
von: Liang, Sheng, et al.
Veröffentlicht: (2025)
RuOpinionNE-2024: Extraction of Opinion Tuples from Russian News Texts
von: Loukachevitch, Natalia, et al.
Veröffentlicht: (2025)
von: Loukachevitch, Natalia, et al.
Veröffentlicht: (2025)
SEED: Enhancing Text-to-SQL Performance and Practical Usability Through Automatic Evidence Generation
von: Yun, Janghyeon, et al.
Veröffentlicht: (2025)
von: Yun, Janghyeon, et al.
Veröffentlicht: (2025)
The Curious Case of Visual Grounding: Different Effects for Speech- and Text-based Language Encoders
von: Sauter, Adrian, et al.
Veröffentlicht: (2025)
von: Sauter, Adrian, et al.
Veröffentlicht: (2025)
Comparing Complex Concepts with Transformers: Matching Patent Claims Against Natural Language Text
von: Blume, Matthias, et al.
Veröffentlicht: (2024)
von: Blume, Matthias, et al.
Veröffentlicht: (2024)
Named Entity Recognition for Address Extraction in Speech-to-Text Transcriptions Using Synthetic Data
von: Lajčinová, Bibiána, et al.
Veröffentlicht: (2024)
von: Lajčinová, Bibiána, et al.
Veröffentlicht: (2024)
Lacuna Language Learning: Leveraging RNNs for Ranked Text Completion in Digitized Coptic Manuscripts
von: Levine, Lauren, et al.
Veröffentlicht: (2024)
von: Levine, Lauren, et al.
Veröffentlicht: (2024)
On the Influence of Discourse Relations in Persuasive Texts
von: Turk, Nawar, et al.
Veröffentlicht: (2025)
von: Turk, Nawar, et al.
Veröffentlicht: (2025)
Beyond Demographics: Fine-tuning Large Language Models to Predict Individuals' Subjective Text Perceptions
von: Orlikowski, Matthias, et al.
Veröffentlicht: (2025)
von: Orlikowski, Matthias, et al.
Veröffentlicht: (2025)
Y-NQ: English-Yorùbá Evaluation dataset for Open-Book Reading Comprehension and Text Generation
von: Costa-jussà, Marta R., et al.
Veröffentlicht: (2024)
von: Costa-jussà, Marta R., et al.
Veröffentlicht: (2024)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
von: Saji, Alan, et al.
Veröffentlicht: (2025)
von: Saji, Alan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Benchmark of French ASR Systems Based on Error Severity
von: Tholly, Antoine, et al.
Veröffentlicht: (2025) -
Dialect Matters: Cross-Lingual ASR Transfer for Low-Resource Indic Language Varieties
von: Dhasmana, Akriti, et al.
Veröffentlicht: (2026) -
RAID: A Shared Benchmark for Robust Evaluation of Machine-Generated Text Detectors
von: Dugan, Liam, et al.
Veröffentlicht: (2024) -
Hard to Be Heard: Phoneme-Level ASR Analysis of Phonologically Complex, Low-Resource Endangered Languages
von: Akavarapu, V. S. D. S. Mahesh, et al.
Veröffentlicht: (2026) -
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)