Infant Cry Detection Using Causal Temporal Representation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fu, Minghao, Li, Danning, Gadhiya, Aryan, Lambright, Benjamin, Alowais, Mohamed, Bahnassy, Mohab, Elletter, Saad El Dine, Toyin, Hawau Olamide, Jiang, Haiyan, Zhang, Kun, Aldarmaki, Hanan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Are LLMs Good Text Diacritizers? An Arabic and Yoruba Case Study
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2025)
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2025)
STTATTS: Unified Speech-To-Text And Text-To-Speech Model
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2024)
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2024)
Dialectal Coverage And Generalization in Arabic Speech Recognition
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2024)
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2024)
ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2025)
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2025)
Voice of a Continent: Mapping Africa's Speech Technology Frontier
von: Elmadany, AbdelRahim, et al.
Veröffentlicht: (2025)
von: Elmadany, AbdelRahim, et al.
Veröffentlicht: (2025)
Aligning Stuttered-Speech Research with End-User Needs: Scoping Review, Survey, and Guidelines
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2026)
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2026)
Clinical Annotations for Automatic Stuttering Severity Assessment
von: Valente, Ana Rita, et al.
Veröffentlicht: (2025)
von: Valente, Ana Rita, et al.
Veröffentlicht: (2025)
NADI 2025: The First Multidialectal Arabic Speech Processing Shared Task
von: Talafha, Bashar, et al.
Veröffentlicht: (2025)
von: Talafha, Bashar, et al.
Veröffentlicht: (2025)
Where Are We? Evaluating LLM Performance on African Languages
von: Adebara, Ife, et al.
Veröffentlicht: (2025)
von: Adebara, Ife, et al.
Veröffentlicht: (2025)
Exploring the Limitations of Detecting Machine-Generated Text
von: Doughman, Jad, et al.
Veröffentlicht: (2024)
von: Doughman, Jad, et al.
Veröffentlicht: (2024)
SparQLe: Speech Queries to Text Translation Through LLMs
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2025)
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2025)
RelUNet: Relative Channel Fusion U-Net for Multichannel Speech Enhancement
von: Aldarmaki, Ibrahim, et al.
Veröffentlicht: (2024)
von: Aldarmaki, Ibrahim, et al.
Veröffentlicht: (2024)
Station list and links to master tracks in different resolutions of POLARSTERN cruise ANT-XXII/1, Bremerhaven - Cape Town, 2004-10-12 - 2004-11-04
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
Station list and links to master tracks in different resolutions of POLARSTERN cruise ANT-XV/1, Bremerhaven - Cape Town, 1997-10-15 - 1997-11-07
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
Station list and links to master tracks in different resolutions of POLARSTERN cruise ANT-XXVII/4, Cape Town - Bremerhaven, 2011-04-20 - 2011-05-20
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
Station list and links to master tracks in different resolutions of POLARSTERN cruise ANT-XXVIII/1, Bremerhaven - Cape Town, 2011-10-28 - 2011-11-30
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
Station list and links to master tracks in different resolutions of POLARSTERN cruise ANT-XVIII/1, Bremerhaven - Cape Town, 2000-09-29 - 2000-10-24
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
Station list and links to master tracks in different resolutions of POLARSTERN cruise ANT-XVI/1, Bremerhaven - Cape Town, 1998-12-15 - 1999-01-06
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
von: El Naggar, Saad El Dine
Veröffentlicht: (2015)
Mixat: A Data Set of Bilingual Emirati-English Speech
von: Ali, Maryam Al, et al.
Veröffentlicht: (2024)
von: Ali, Maryam Al, et al.
Veröffentlicht: (2024)
Spoken Word2Vec: Learning Skipgram Embeddings from Speech
von: Sayeed, Mohammad Amaan, et al.
Veröffentlicht: (2023)
von: Sayeed, Mohammad Amaan, et al.
Veröffentlicht: (2023)
Automatic Restoration of Diacritics for Speech Data Sets
von: Shatnawi, Sara, et al.
Veröffentlicht: (2023)
von: Shatnawi, Sara, et al.
Veröffentlicht: (2023)
Domain-Agnostic Causal-Aware Audio Transformer for Infant Cry Classification
von: Owino, Geofrey, et al.
Veröffentlicht: (2025)
von: Owino, Geofrey, et al.
Veröffentlicht: (2025)
InfantCryNet: A Data-driven Framework for Intelligent Analysis of Infant Cries
von: Hong, Mengze, et al.
Veröffentlicht: (2024)
von: Hong, Mengze, et al.
Veröffentlicht: (2024)
Continuous thermosalinograph oceanography along POLARSTERN cruise track ANT-XXVII/4
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (2011)
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (2011)
Continuous thermosalinograph oceanography along POLARSTERN cruise track ANT-XV/1
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (1998)
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (1998)
Continuous thermosalinograph oceanography along POLARSTERN cruise track ANT-XXII/1
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (2007)
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (2007)
Continuous thermosalinograph oceanography along POLARSTERN cruise track ANT-XXVIII/1
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (2012)
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (2012)
Analysis of UVB irradiance during cruise ANT-XXII/1 (Table 10.1)
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (2007)
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (2007)
Continuous thermosalinograph oceanography along POLARSTERN cruise track ANT-XVI/1
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (2007)
von: El Naggar, Saad El Dine, et al.
Veröffentlicht: (2007)
On Exact Reznick, Hilbert-Artin and Putinar's Representations
von: Magron, Victor, et al.
Veröffentlicht: (2018)
von: Magron, Victor, et al.
Veröffentlicht: (2018)
Towards Identifiability of Hierarchical Temporal Causal Representation Learning
von: Li, Zijian, et al.
Veröffentlicht: (2025)
von: Li, Zijian, et al.
Veröffentlicht: (2025)
CryCeleb: A Speaker Verification Dataset Based on Infant Cry Sounds
von: Budaghyan, David, et al.
Veröffentlicht: (2023)
von: Budaghyan, David, et al.
Veröffentlicht: (2023)
Unrequited Emotions: Investigating the Gaps in Motivation and Practice in Speech Emotion Recognition Research
von: Wong, Taryn, et al.
Veröffentlicht: (2026)
von: Wong, Taryn, et al.
Veröffentlicht: (2026)
Morphemes Without Borders: Evaluating Root-Pattern Morphology in Arabic Tokenizers and LLMs
von: Alakeel, Yara, et al.
Veröffentlicht: (2026)
von: Alakeel, Yara, et al.
Veröffentlicht: (2026)
Linear Semantic Segmentation for Low-Resource Spoken Dialects
von: Chirkunov, Kirill, et al.
Veröffentlicht: (2026)
von: Chirkunov, Kirill, et al.
Veröffentlicht: (2026)
SPIRIT: Patching Speech Language Models against Jailbreak Attacks
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2025)
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2025)
PALM: Few-Shot Prompt Learning for Audio Language Models
von: Hanif, Asif, et al.
Veröffentlicht: (2024)
von: Hanif, Asif, et al.
Veröffentlicht: (2024)
ICSD: An Open-source Dataset for Infant Cry and Snoring Detection
von: Liu, Qingyu, et al.
Veröffentlicht: (2024)
von: Liu, Qingyu, et al.
Veröffentlicht: (2024)
Father Trait Anger and Exposure to Infant Cry: Effects on Emotion, Appraisals of Infants, and Cognitive Performance
von: Lauren M. Francis, et al.
Veröffentlicht: (2025)
von: Lauren M. Francis, et al.
Veröffentlicht: (2025)
Privacy-Enhancing Infant Cry Classification with Federated Transformers and Denoising Regularization
von: Owino, Geofrey, et al.
Veröffentlicht: (2025)
von: Owino, Geofrey, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Are LLMs Good Text Diacritizers? An Arabic and Yoruba Case Study
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2025) -
STTATTS: Unified Speech-To-Text And Text-To-Speech Model
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2024) -
Dialectal Coverage And Generalization in Arabic Speech Recognition
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2024) -
ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2025) -
Voice of a Continent: Mapping Africa's Speech Technology Frontier
von: Elmadany, AbdelRahim, et al.
Veröffentlicht: (2025)