Selective Augmentation: Improving Universal Automatic Phonetic Transcription via G2P Bootstrapping
Fuente:
arXiv
Guardado en:
| Autores principales: | Bystrich, Tobias, Pritzen, Julia M., Schmidt, Christoph A., Wich-Reif, Claudia |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Impact of Automatic Speech Transcription on Speaker Attribution
por: Aggazzotti, Cristina, et al.
Publicado: (2025)
por: Aggazzotti, Cristina, et al.
Publicado: (2025)
Cross-Domain Data Selection and Augmentation for Automatic Compliance Detection
por: Ikhwantri, Fariz, et al.
Publicado: (2026)
por: Ikhwantri, Fariz, et al.
Publicado: (2026)
Multitask Learning for Grapheme-to-Phoneme Conversion of Anglicisms in German Speech Recognition
por: Pritzen, Julia, et al.
Publicado: (2021)
por: Pritzen, Julia, et al.
Publicado: (2021)
Majority of the Bests: Improving Best-of-N via Bootstrapping
por: Rakhsha, Amin, et al.
Publicado: (2025)
por: Rakhsha, Amin, et al.
Publicado: (2025)
Auto-GDA: Automatic Domain Adaptation for Efficient Grounding Verification in Retrieval-Augmented Generation
por: Leemann, Tobias, et al.
Publicado: (2024)
por: Leemann, Tobias, et al.
Publicado: (2024)
Data Augmentations for Improved (Large) Language Model Generalization
por: Feder, Amir, et al.
Publicado: (2023)
por: Feder, Amir, et al.
Publicado: (2023)
I Have No Mouth, and I Must Rhyme: Uncovering Internal Phonetic Representations in LLaMA 3.2
por: McLaughlin, Oliver, et al.
Publicado: (2025)
por: McLaughlin, Oliver, et al.
Publicado: (2025)
PAST: Phonetic-Acoustic Speech Tokenizer
por: Har-Tuv, Nadav, et al.
Publicado: (2025)
por: Har-Tuv, Nadav, et al.
Publicado: (2025)
Automatic Prompt Selection for Large Language Models
por: Do, Viet-Tung, et al.
Publicado: (2024)
por: Do, Viet-Tung, et al.
Publicado: (2024)
VISPA: Pluralistic Alignment via Automatic Value Selection and Activation
por: Zheng, Shenyan, et al.
Publicado: (2026)
por: Zheng, Shenyan, et al.
Publicado: (2026)
Self-Supervised Speech Representations are More Phonetic than Semantic
por: Choi, Kwanghee, et al.
Publicado: (2024)
por: Choi, Kwanghee, et al.
Publicado: (2024)
Bootstrapping Language Models with DPO Implicit Rewards
por: Chen, Changyu, et al.
Publicado: (2024)
por: Chen, Changyu, et al.
Publicado: (2024)
AutoBaxBuilder: Bootstrapping Code Security Benchmarking
por: von Arx, Tobias, et al.
Publicado: (2025)
por: von Arx, Tobias, et al.
Publicado: (2025)
Self-Supervised Speech Models Encode Phonetic Context via Position-dependent Orthogonal Subspaces
por: Choi, Kwanghee, et al.
Publicado: (2026)
por: Choi, Kwanghee, et al.
Publicado: (2026)
ISPA: Inter-Species Phonetic Alphabet for Transcribing Animal Sounds
por: Hagiwara, Masato, et al.
Publicado: (2024)
por: Hagiwara, Masato, et al.
Publicado: (2024)
Forging the Forger: An Attempt to Improve Authorship Verification via Data Augmentation
por: Corbara, Silvia, et al.
Publicado: (2024)
por: Corbara, Silvia, et al.
Publicado: (2024)
Informed Bootstrap Augmentation Improves EEG Decoding
por: Jeong, Woojae, et al.
Publicado: (2025)
por: Jeong, Woojae, et al.
Publicado: (2025)
ReEval: Automatic Hallucination Evaluation for Retrieval-Augmented Large Language Models via Transferable Adversarial Attacks
por: Yu, Xiaodong, et al.
Publicado: (2023)
por: Yu, Xiaodong, et al.
Publicado: (2023)
Phonetic and Lexical Discovery of a Canine Language using HuBERT
por: Li, Xingyuan, et al.
Publicado: (2024)
por: Li, Xingyuan, et al.
Publicado: (2024)
Improving Sentence Embeddings with Automatic Generation of Training Data Using Few-shot Examples
por: Sato, Soma, et al.
Publicado: (2024)
por: Sato, Soma, et al.
Publicado: (2024)
Bootstrapping Post-training Signals for Open-ended Tasks via Rubric-based Self-play on Pre-training Text
por: Huang, Chengyu, et al.
Publicado: (2026)
por: Huang, Chengyu, et al.
Publicado: (2026)
Less is More for Improving Automatic Evaluation of Factual Consistency
por: Wang, Tong, et al.
Publicado: (2024)
por: Wang, Tong, et al.
Publicado: (2024)
Huntington Disease Automatic Speech Recognition with Biomarker Supervision
por: Wang, Charles L., et al.
Publicado: (2026)
por: Wang, Charles L., et al.
Publicado: (2026)
Joint Unsupervised and Supervised Training for Automatic Speech Recognition via Bilevel Optimization
por: Saif, A F M, et al.
Publicado: (2024)
por: Saif, A F M, et al.
Publicado: (2024)
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards
por: Nica, Andreea, et al.
Publicado: (2025)
por: Nica, Andreea, et al.
Publicado: (2025)
The ART of Conversation: Measuring Phonetic Convergence and Deliberate Imitation in L2-Speech with a Siamese RNN
por: Yuan, Zheng, et al.
Publicado: (2023)
por: Yuan, Zheng, et al.
Publicado: (2023)
Data Augmentation via Causal-Residual Bootstrapping
por: Gajewski, Mateusz, et al.
Publicado: (2026)
por: Gajewski, Mateusz, et al.
Publicado: (2026)
Automatic Combination of Sample Selection Strategies for Few-Shot Learning
por: Pecher, Branislav, et al.
Publicado: (2024)
por: Pecher, Branislav, et al.
Publicado: (2024)
Less is More: Improving LLM Alignment via Preference Data Selection
por: Deng, Xun, et al.
Publicado: (2025)
por: Deng, Xun, et al.
Publicado: (2025)
Can Authorship Attribution Models Distinguish Speakers in Speech Transcripts?
por: Aggazzotti, Cristina, et al.
Publicado: (2023)
por: Aggazzotti, Cristina, et al.
Publicado: (2023)
Modeling Real-Time Interactive Conversations as Timed Diarized Transcripts
por: Tanzer, Garrett, et al.
Publicado: (2024)
por: Tanzer, Garrett, et al.
Publicado: (2024)
Improving Customer Service with Automatic Topic Detection in User Emails
por: Bašaragin, Bojana, et al.
Publicado: (2025)
por: Bašaragin, Bojana, et al.
Publicado: (2025)
Reasoning in Trees: Improving Retrieval-Augmented Generation for Multi-Hop Question Answering
por: Shi, Yuling, et al.
Publicado: (2026)
por: Shi, Yuling, et al.
Publicado: (2026)
Dissecting Linear Recurrent Models: How Different Gating Strategies Drive Selectivity and Generalization
por: Bouhadjar, Younes, et al.
Publicado: (2026)
por: Bouhadjar, Younes, et al.
Publicado: (2026)
Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-Semantic Embeddings
por: Taylor, Russell, et al.
Publicado: (2025)
por: Taylor, Russell, et al.
Publicado: (2025)
Error Analysis in a Modular Meeting Transcription System
por: Vieting, Peter, et al.
Publicado: (2025)
por: Vieting, Peter, et al.
Publicado: (2025)
Selective Attention Improves Transformer
por: Leviathan, Yaniv, et al.
Publicado: (2024)
por: Leviathan, Yaniv, et al.
Publicado: (2024)
Automatic Functional Differentiation in JAX
por: Lin, Min
Publicado: (2023)
por: Lin, Min
Publicado: (2023)
Uncertainty-Aware Answer Selection for Improved Reasoning in Multi-LLM Systems
por: Agrawal, Aakriti, et al.
Publicado: (2025)
por: Agrawal, Aakriti, et al.
Publicado: (2025)
Nostra Domina at EvaLatin 2024: Improving Latin Polarity Detection through Data Augmentation
por: Bothwell, Stephen, et al.
Publicado: (2024)
por: Bothwell, Stephen, et al.
Publicado: (2024)
Ejemplares similares
-
The Impact of Automatic Speech Transcription on Speaker Attribution
por: Aggazzotti, Cristina, et al.
Publicado: (2025) -
Cross-Domain Data Selection and Augmentation for Automatic Compliance Detection
por: Ikhwantri, Fariz, et al.
Publicado: (2026) -
Multitask Learning for Grapheme-to-Phoneme Conversion of Anglicisms in German Speech Recognition
por: Pritzen, Julia, et al.
Publicado: (2021) -
Majority of the Bests: Improving Best-of-N via Bootstrapping
por: Rakhsha, Amin, et al.
Publicado: (2025) -
Auto-GDA: Automatic Domain Adaptation for Efficient Grounding Verification in Retrieval-Augmented Generation
por: Leemann, Tobias, et al.
Publicado: (2024)