A Reliable and Efficient Detection Pipeline for Rodent Ultrasonic Vocalizations
Fuente:
arXiv
Guardado en:
| Autores principales: | Anis, Sabah Shahnoor, Kellis, Devin M., Kaigler, Kris Ford, Wilson, Marlene A., O'Reilly, Christian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Speaker Embeddings With Weakly Supervised Voice Activity Detection For Efficient Speaker Diarization
por: Thienpondt, Jenthe, et al.
Publicado: (2024)
por: Thienpondt, Jenthe, et al.
Publicado: (2024)
Deep Audio Watermarks are Shallow: Limitations of Post-Hoc Watermarking Techniques for Speech
por: O'Reilly, Patrick, et al.
Publicado: (2025)
por: O'Reilly, Patrick, et al.
Publicado: (2025)
Text2FX: Harnessing CLAP Embeddings for Text-Guided Audio Effects
por: Chu, Annie, et al.
Publicado: (2024)
por: Chu, Annie, et al.
Publicado: (2024)
Code Drift: Towards Idempotent Neural Audio Codecs
por: O'Reilly, Patrick, et al.
Publicado: (2024)
por: O'Reilly, Patrick, et al.
Publicado: (2024)
Mel-RoFormer for Vocal Separation and Vocal Melody Transcription
por: Wang, Ju-Chiang, et al.
Publicado: (2024)
por: Wang, Ju-Chiang, et al.
Publicado: (2024)
CVSM: Contrastive Vocal Similarity Modeling
por: Garoufis, Christos, et al.
Publicado: (2025)
por: Garoufis, Christos, et al.
Publicado: (2025)
The Rhythm In Anything: Audio-Prompted Drums Generation with Masked Language Modeling
por: O'Reilly, Patrick, et al.
Publicado: (2025)
por: O'Reilly, Patrick, et al.
Publicado: (2025)
Harmonic Summation-Based Robust Pitch Estimation in Noisy and Reverberant Environments
por: Singh, Anup, et al.
Publicado: (2025)
por: Singh, Anup, et al.
Publicado: (2025)
Melodic and Metrical Elements of Expressiveness in Hindustani Vocal Music
por: Bhake, Yash, et al.
Publicado: (2025)
por: Bhake, Yash, et al.
Publicado: (2025)
Auditory Representation Effective for Estimating Vocal Tract Information
por: Irino, Toshio, et al.
Publicado: (2023)
por: Irino, Toshio, et al.
Publicado: (2023)
Biodenoising: Animal Vocalization Denoising without Access to Clean Data
por: Miron, Marius, et al.
Publicado: (2024)
por: Miron, Marius, et al.
Publicado: (2024)
Drum-to-Vocal Percussion Sound Conversion and Its Evaluation Methodology
por: Nobukawa, Rinka, et al.
Publicado: (2025)
por: Nobukawa, Rinka, et al.
Publicado: (2025)
Analysis of Self-Supervised Speech Models on Children's Speech and Infant Vocalizations
por: Li, Jialu, et al.
Publicado: (2024)
por: Li, Jialu, et al.
Publicado: (2024)
voc2vec: A Foundation Model for Non-Verbal Vocalization
por: Koudounas, Alkis, et al.
Publicado: (2025)
por: Koudounas, Alkis, et al.
Publicado: (2025)
RMVPE: A Robust Model for Vocal Pitch Estimation in Polyphonic Music
por: Wei, Haojie, et al.
Publicado: (2023)
por: Wei, Haojie, et al.
Publicado: (2023)
Attention-Based Audio Embeddings for Query-by-Example
por: Singh, Anup, et al.
Publicado: (2022)
por: Singh, Anup, et al.
Publicado: (2022)
Learning Vocal-Tract Area and Radiation with a Physics-Informed Webster Model
por: Lu, Minhui, et al.
Publicado: (2026)
por: Lu, Minhui, et al.
Publicado: (2026)
Enhancing Child Vocalization Classification with Phonetically-Tuned Embeddings for Assisting Autism Diagnosis
por: Li, Jialu, et al.
Publicado: (2023)
por: Li, Jialu, et al.
Publicado: (2023)
DiffVox: A Differentiable Model for Capturing and Analysing Vocal Effects Distributions
por: Yu, Chin-Yun, et al.
Publicado: (2025)
por: Yu, Chin-Yun, et al.
Publicado: (2025)
DJCM: A Deep Joint Cascade Model for Singing Voice Separation and Vocal Pitch Estimation
por: Wei, Haojie, et al.
Publicado: (2024)
por: Wei, Haojie, et al.
Publicado: (2024)
Hearing Health in Home Healthcare: Leveraging LLMs for Illness Scoring and ALMs for Vocal Biomarker Extraction
por: Chen, Yu-Wen, et al.
Publicado: (2025)
por: Chen, Yu-Wen, et al.
Publicado: (2025)
Computational Extraction of Intonation and Tuning Systems from Multiple Microtonal Monophonic Vocal Recordings with Diverse Modes
por: Shafiei, Sepideh, et al.
Publicado: (2025)
por: Shafiei, Sepideh, et al.
Publicado: (2025)
Blind Source Separation in Biomedical Signals Using Variational Methods
por: Torabi, Yasaman, et al.
Publicado: (2025)
por: Torabi, Yasaman, et al.
Publicado: (2025)
Large Language Models and Non-Negative Matrix Factorization for Bioacoustic Signal Decomposition
por: Torabi, Yasaman, et al.
Publicado: (2025)
por: Torabi, Yasaman, et al.
Publicado: (2025)
Live Vocal Extraction from K-pop Performances
por: Kim, Yujin, et al.
Publicado: (2025)
por: Kim, Yujin, et al.
Publicado: (2025)
Relating the Neural Representations of Vocalized, Mimed, and Imagined Speech
por: Maghsoudi, Maryam, et al.
Publicado: (2026)
por: Maghsoudi, Maryam, et al.
Publicado: (2026)
VocalAgent: Large Language Models for Vocal Health Diagnostics with Safety-Aware Evaluation
por: Kim, Yubin, et al.
Publicado: (2025)
por: Kim, Yubin, et al.
Publicado: (2025)
A Scalable Pipeline for Enabling Non-Verbal Speech Generation and Understanding
por: Ye, Runchuan, et al.
Publicado: (2025)
por: Ye, Runchuan, et al.
Publicado: (2025)
MIDI-Informed Singing Accompaniment Generation in a Compositional Song Pipeline
por: Tsai, Fang-Duo, et al.
Publicado: (2026)
por: Tsai, Fang-Duo, et al.
Publicado: (2026)
Bird Vocalization Embedding Extraction Using Self-Supervised Disentangled Representation Learning
por: Shi, Runwu, et al.
Publicado: (2024)
por: Shi, Runwu, et al.
Publicado: (2024)
Sound-Based Spin Estimation in Table Tennis: Dataset and Real-Time Classification Pipeline
por: Gossard, Thomas, et al.
Publicado: (2024)
por: Gossard, Thomas, et al.
Publicado: (2024)
A Detailed Audio-Text Data Simulation Pipeline using Single-Event Sounds
por: Xu, Xuenan, et al.
Publicado: (2024)
por: Xu, Xuenan, et al.
Publicado: (2024)
DiariZen Explained: A Tutorial for the Open Source State-of-the-Art Speaker Diarization Pipeline
por: Raghav, Nikhil
Publicado: (2026)
por: Raghav, Nikhil
Publicado: (2026)
SOA: Reducing Domain Mismatch in SSL Pipeline by Speech Only Adaptation for Low Resource ASR
por: Shankar, Natarajan Balaji, et al.
Publicado: (2024)
por: Shankar, Natarajan Balaji, et al.
Publicado: (2024)
Towards the Synthesis of Non-speech Vocalizations
por: Hoq, Enjamamul, et al.
Publicado: (2024)
por: Hoq, Enjamamul, et al.
Publicado: (2024)
Recurrence-Based Nonlinear Vocal Dynamics as Digital Biomarkers for Depression Detection from Conversational Speech
por: Samanta, Himadri S
Publicado: (2026)
por: Samanta, Himadri S
Publicado: (2026)
Echoes of Ideology: Toward an Audio Analysis Pipeline to Unveil Character Traits in Historical Nazi Propaganda Films
por: Ruth, Nicolas, et al.
Publicado: (2026)
por: Ruth, Nicolas, et al.
Publicado: (2026)
Feature Representations for Automatic Meerkat Vocalization Classification
por: Mahmoud, Imen Ben, et al.
Publicado: (2024)
por: Mahmoud, Imen Ben, et al.
Publicado: (2024)
Feasibility of Mental Health Triage Call Priority Prediction Using Machine Learning
por: Rana, Rajib, et al.
Publicado: (2024)
por: Rana, Rajib, et al.
Publicado: (2024)
Large Language Model-based Nonnegative Matrix Factorization For Cardiorespiratory Sound Separation
por: Torabi, Yasaman, et al.
Publicado: (2025)
por: Torabi, Yasaman, et al.
Publicado: (2025)
Ejemplares similares
-
Speaker Embeddings With Weakly Supervised Voice Activity Detection For Efficient Speaker Diarization
por: Thienpondt, Jenthe, et al.
Publicado: (2024) -
Deep Audio Watermarks are Shallow: Limitations of Post-Hoc Watermarking Techniques for Speech
por: O'Reilly, Patrick, et al.
Publicado: (2025) -
Text2FX: Harnessing CLAP Embeddings for Text-Guided Audio Effects
por: Chu, Annie, et al.
Publicado: (2024) -
Code Drift: Towards Idempotent Neural Audio Codecs
por: O'Reilly, Patrick, et al.
Publicado: (2024) -
Mel-RoFormer for Vocal Separation and Vocal Melody Transcription
por: Wang, Ju-Chiang, et al.
Publicado: (2024)