HebDB: a Weakly Supervised Dataset for Hebrew Speech Processing
Fuente:
arXiv
Salvato in:
| Autori principali: | Turetzky, Arnon, Tal, Or, Segal-Feldman, Yael, Dissen, Yehoshua, Zeldes, Ella, Roth, Amit, Cohen, Eyal, Shrem, Yosi, Chernyak, Bronya R., Seleznova, Olga, Keshet, Joseph, Adi, Yossi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PatchDSU: Uncertainty Modeling for Out of Distribution Generalization in Keyword Spotting
di: Chernyak, Bronya Roni, et al.
Pubblicazione: (2025)
di: Chernyak, Bronya Roni, et al.
Pubblicazione: (2025)
A Language Modeling Approach to Diacritic-Free Hebrew TTS
di: Roth, Amit, et al.
Pubblicazione: (2024)
di: Roth, Amit, et al.
Pubblicazione: (2024)
Enhancing TTS Stability in Hebrew using Discrete Semantic Units
di: Zeldes, Ella, et al.
Pubblicazione: (2024)
di: Zeldes, Ella, et al.
Pubblicazione: (2024)
LAST: Language Model Aware Speech Tokenization
di: Turetzky, Arnon, et al.
Pubblicazione: (2024)
di: Turetzky, Arnon, et al.
Pubblicazione: (2024)
Enhanced ASR Robustness to Packet Loss with a Front-End Adaptation Network
di: Dissen, Yehoshua, et al.
Pubblicazione: (2024)
di: Dissen, Yehoshua, et al.
Pubblicazione: (2024)
Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark
di: Turetzky, Arnon, et al.
Pubblicazione: (2026)
di: Turetzky, Arnon, et al.
Pubblicazione: (2026)
HebID: Detecting Social Identities in Hebrew-language Political Text
di: Mor-Lan, Guy, et al.
Pubblicazione: (2025)
di: Mor-Lan, Guy, et al.
Pubblicazione: (2025)
Joint Enhancement and Classification using Coupled Diffusion Models of Signals and Logits
di: Nurko, Gilad, et al.
Pubblicazione: (2026)
di: Nurko, Gilad, et al.
Pubblicazione: (2026)
Keyword Spotting with Hyper-Matched Filters for Small Footprint Devices
di: Segal-Feldman, Yael, et al.
Pubblicazione: (2025)
di: Segal-Feldman, Yael, et al.
Pubblicazione: (2025)
Salmon: A Suite for Acoustic Language Model Evaluation
di: Maimon, Gallil, et al.
Pubblicazione: (2024)
di: Maimon, Gallil, et al.
Pubblicazione: (2024)
Speech Synthesis From Continuous Features Using Per-Token Latent Diffusion
di: Turetzky, Arnon, et al.
Pubblicazione: (2024)
di: Turetzky, Arnon, et al.
Pubblicazione: (2024)
Whisper in Medusa's Ear: Multi-head Efficient Decoding for Transformer-based ASR
di: Segal-Feldman, Yael, et al.
Pubblicazione: (2024)
di: Segal-Feldman, Yael, et al.
Pubblicazione: (2024)
Auto-Regressive vs Flow-Matching: a Comparative Study of Modeling Paradigms for Text-to-Music Generation
di: Tal, Or, et al.
Pubblicazione: (2025)
di: Tal, Or, et al.
Pubblicazione: (2025)
Scaling Analysis of Interleaved Speech-Text Language Models
di: Maimon, Gallil, et al.
Pubblicazione: (2025)
di: Maimon, Gallil, et al.
Pubblicazione: (2025)
FlowTSE: Target Speaker Extraction with Flow Matching
di: Navon, Aviv, et al.
Pubblicazione: (2025)
di: Navon, Aviv, et al.
Pubblicazione: (2025)
Marriage and the theology of Hebrews. A theological reading of Heb 12:28–13:6 with a focus on marriage
di: Adriani Milli Rodrigues
Pubblicazione: (2018)
di: Adriani Milli Rodrigues
Pubblicazione: (2018)
PAST: Phonetic-Acoustic Speech Tokenizer
di: Har-Tuv, Nadav, et al.
Pubblicazione: (2025)
di: Har-Tuv, Nadav, et al.
Pubblicazione: (2025)
Drax: Speech Recognition with Discrete Flow Matching
di: Navon, Aviv, et al.
Pubblicazione: (2025)
di: Navon, Aviv, et al.
Pubblicazione: (2025)
FrackyFrac: A Standalone UniFrac Calculator
di: Lavon, Amit, et al.
Pubblicazione: (2024)
di: Lavon, Amit, et al.
Pubblicazione: (2024)
How Long Does Infinite Width Last? Signal Propagation in Long-Range Linear Recurrences
di: Seleznova, Mariia
Pubblicazione: (2026)
di: Seleznova, Mariia
Pubblicazione: (2026)
Learning to Read and Developmental Dyslexia in Hebrew
di: Adi Shechter, et al.
Pubblicazione: (2024)
di: Adi Shechter, et al.
Pubblicazione: (2024)
JudgeMeNot: Personalizing Large Language Models to Emulate Judicial Reasoning in Hebrew
di: Razumenko, Itay, et al.
Pubblicazione: (2026)
di: Razumenko, Itay, et al.
Pubblicazione: (2026)
Joint Audio and Symbolic Conditioning for Temporally Controlled Text-to-Music Generation
di: Tal, Or, et al.
Pubblicazione: (2024)
di: Tal, Or, et al.
Pubblicazione: (2024)
D-Nikud: Enhancing Hebrew Diacritization with LSTM and Pretrained Models
di: Rosenthal, Adi, et al.
Pubblicazione: (2024)
di: Rosenthal, Adi, et al.
Pubblicazione: (2024)
Co‐creation: Pioneering progress in cerebral palsy research
di: Rachel Byrne,, et al.
Pubblicazione: (2024)
di: Rachel Byrne,, et al.
Pubblicazione: (2024)
Large Language Model Guided Decoding for Self-Supervised Speech Recognition
di: Cohen, Eyal, et al.
Pubblicazione: (2025)
di: Cohen, Eyal, et al.
Pubblicazione: (2025)
Controlled Loop Expansion for Strained Twisted Bilayer Graphene
di: Keshet, Eyal, et al.
Pubblicazione: (2026)
di: Keshet, Eyal, et al.
Pubblicazione: (2026)
The ringdown-Hawking radiation connection in real and analogue black holes
di: Keshet, Eyal, et al.
Pubblicazione: (2024)
di: Keshet, Eyal, et al.
Pubblicazione: (2024)
Lowering the Horizon on Dark Energy: A Late-Time Response to Early Solutions for the Hubble Tension
di: Adi, Tal
Pubblicazione: (2025)
di: Adi, Tal
Pubblicazione: (2025)
Beyond Transcription: Mechanistic Interpretability in ASR
di: Glazer, Neta, et al.
Pubblicazione: (2025)
di: Glazer, Neta, et al.
Pubblicazione: (2025)
The Larger the Better? Improved LLM Code-Generation via Budget Reallocation
di: Hassid, Michael, et al.
Pubblicazione: (2024)
di: Hassid, Michael, et al.
Pubblicazione: (2024)
Resilient Biosecurity in the Era of AI-Enabled Bioweapons
di: Feldman, Jonathan, et al.
Pubblicazione: (2025)
di: Feldman, Jonathan, et al.
Pubblicazione: (2025)
In planta genome editing in citrus facilitated by co‐expression of CRISPR / Cas and developmental regulators
di: Gilor Kelly, et al.
Pubblicazione: (2025)
di: Gilor Kelly, et al.
Pubblicazione: (2025)
Formal Language Knowledge Corpus for Retrieval Augmented Generation
di: Zayyad, Majd, et al.
Pubblicazione: (2024)
di: Zayyad, Majd, et al.
Pubblicazione: (2024)
NAST: Noise Aware Speech Tokenization for Speech Language Models
di: Messica, Shoval, et al.
Pubblicazione: (2024)
di: Messica, Shoval, et al.
Pubblicazione: (2024)
Pointer Chasing with Unlimited Interaction
di: Fischer, Orr, et al.
Pubblicazione: (2025)
di: Fischer, Orr, et al.
Pubblicazione: (2025)
The Perks of Life as a Coral‐Boring Bivalve
di: Tal Amit, et al.
Pubblicazione: (2024)
di: Tal Amit, et al.
Pubblicazione: (2024)
The Knesset Corpus: An Annotated Corpus of Hebrew Parliamentary Proceedings
di: Goldin, Gili, et al.
Pubblicazione: (2024)
di: Goldin, Gili, et al.
Pubblicazione: (2024)
Probing the Dark Sector using the Fornax Satellite
di: Zamlung, Eyal, et al.
Pubblicazione: (2025)
di: Zamlung, Eyal, et al.
Pubblicazione: (2025)
Tradition or Innovation: A Comparison of Modern ASR Methods for Forced Alignment
di: Rousso, Rotem, et al.
Pubblicazione: (2024)
di: Rousso, Rotem, et al.
Pubblicazione: (2024)
Documenti analoghi
-
PatchDSU: Uncertainty Modeling for Out of Distribution Generalization in Keyword Spotting
di: Chernyak, Bronya Roni, et al.
Pubblicazione: (2025) -
A Language Modeling Approach to Diacritic-Free Hebrew TTS
di: Roth, Amit, et al.
Pubblicazione: (2024) -
Enhancing TTS Stability in Hebrew using Discrete Semantic Units
di: Zeldes, Ella, et al.
Pubblicazione: (2024) -
LAST: Language Model Aware Speech Tokenization
di: Turetzky, Arnon, et al.
Pubblicazione: (2024) -
Enhanced ASR Robustness to Packet Loss with a Front-End Adaptation Network
di: Dissen, Yehoshua, et al.
Pubblicazione: (2024)