Evaluation of Deep Audio Representations for Hearables
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gröger, Fabian, Baumann, Pascal, Amruthalingam, Ludovic, Simon, Laurent, Giurda, Ruksana, Lionetti, Simone |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Representation-Based Data Quality Audits for Audio
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025)
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025)
Clinical Uncertainty Impacts Machine Learning Evaluations
von: Lionetti, Simone, et al.
Veröffentlicht: (2025)
von: Lionetti, Simone, et al.
Veröffentlicht: (2025)
Is Hyperbolic Space All You Need for Medical Anomaly Detection?
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025)
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025)
AudioMosaic: Contrastive Masked Audio Representation Learning
von: Huang, Hanxun, et al.
Veröffentlicht: (2026)
von: Huang, Hanxun, et al.
Veröffentlicht: (2026)
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
AudioCodecBench: A Comprehensive Benchmark for Audio Codec Evaluation
von: Wang, Lu, et al.
Veröffentlicht: (2025)
von: Wang, Lu, et al.
Veröffentlicht: (2025)
HEAR: Holistic Evaluation of Audio Representations
von: Turian, Joseph, et al.
Veröffentlicht: (2022)
von: Turian, Joseph, et al.
Veröffentlicht: (2022)
A Human-Inspired Decoupled Architecture for Efficient Audio Representation Learning
von: Kawano, Harunori, et al.
Veröffentlicht: (2026)
von: Kawano, Harunori, et al.
Veröffentlicht: (2026)
Phase-Aware Deep Learning with Complex-Valued CNNs for Audio Signal Applications
von: Agrawal, Naman
Veröffentlicht: (2025)
von: Agrawal, Naman
Veröffentlicht: (2025)
Explainable Multi-Modal Deep Learning for Automatic Detection of Lung Diseases from Respiratory Audio Signals
von: Saky, S M Asiful Islam, et al.
Veröffentlicht: (2025)
von: Saky, S M Asiful Islam, et al.
Veröffentlicht: (2025)
Unsupervised Evaluation of Deep Audio Embeddings for Music Structure Analysis
von: Marmoret, Axel
Veröffentlicht: (2026)
von: Marmoret, Axel
Veröffentlicht: (2026)
Structured-Noise Masked Modeling for Video, Audio and Beyond
von: Bhowmik, Aritra, et al.
Veröffentlicht: (2025)
von: Bhowmik, Aritra, et al.
Veröffentlicht: (2025)
Low-Resource Guidance for Controllable Latent Audio Diffusion
von: Novack, Zachary, et al.
Veröffentlicht: (2026)
von: Novack, Zachary, et al.
Veröffentlicht: (2026)
Exploring Token-Space Manipulation in Latent Audio Tokenizers
von: Paissan, Francesco, et al.
Veröffentlicht: (2026)
von: Paissan, Francesco, et al.
Veröffentlicht: (2026)
Preference-Based Learning in Audio Applications: A Systematic Analysis
von: Broukhim, Aaron, et al.
Veröffentlicht: (2025)
von: Broukhim, Aaron, et al.
Veröffentlicht: (2025)
AUDETER: A Large-scale Dataset for Deepfake Audio Detection in Open Worlds
von: Wang, Qizhou, et al.
Veröffentlicht: (2025)
von: Wang, Qizhou, et al.
Veröffentlicht: (2025)
When Denoising Hinders: Revisiting Zero-Shot ASR with SAM-Audio and Whisper
von: Islam, Akif, et al.
Veröffentlicht: (2026)
von: Islam, Akif, et al.
Veröffentlicht: (2026)
PoDAR: Power-Disentangled Audio Representation for Generative Modeling
von: Luebs, Alejandro, et al.
Veröffentlicht: (2026)
von: Luebs, Alejandro, et al.
Veröffentlicht: (2026)
Real-Time Voicemail Detection in Telephony Audio Using Temporal Speech Activity Features
von: Saurav, Kumar
Veröffentlicht: (2026)
von: Saurav, Kumar
Veröffentlicht: (2026)
A$^2$-LLM: An End-to-end Conversational Audio Avatar Large Language Model
von: Hu, Xiaolin, et al.
Veröffentlicht: (2026)
von: Hu, Xiaolin, et al.
Veröffentlicht: (2026)
A Survey of Deep Learning Audio Generation Methods
von: Božić, Matej, et al.
Veröffentlicht: (2024)
von: Božić, Matej, et al.
Veröffentlicht: (2024)
Boosting ASR Robustness via Test-Time Reinforcement Learning with Audio-Text Semantic Rewards
von: Fang, Linghan, et al.
Veröffentlicht: (2026)
von: Fang, Linghan, et al.
Veröffentlicht: (2026)
CoDiCodec: Unifying Continuous and Discrete Compressed Representations of Audio
von: Pasini, Marco, et al.
Veröffentlicht: (2025)
von: Pasini, Marco, et al.
Veröffentlicht: (2025)
QAMRO: Quality-aware Adaptive Margin Ranking Optimization for Human-aligned Assessment of Audio Generation Systems
von: Wang, Chien-Chun, et al.
Veröffentlicht: (2025)
von: Wang, Chien-Chun, et al.
Veröffentlicht: (2025)
Linguistic and Audio Embedding-Based Machine Learning for Alzheimer's Dementia and Mild Cognitive Impairment Detection: Insights from the PROCESS Challenge
von: Devahi, Adharsha Sam Edwin Sam, et al.
Veröffentlicht: (2025)
von: Devahi, Adharsha Sam Edwin Sam, et al.
Veröffentlicht: (2025)
Myna: Masking-Based Contrastive Learning of Musical Representations
von: Yonay, Ori, et al.
Veröffentlicht: (2025)
von: Yonay, Ori, et al.
Veröffentlicht: (2025)
Investigating Design Choices in Joint-Embedding Predictive Architectures for General Audio Representation Learning
von: Riou, Alain, et al.
Veröffentlicht: (2024)
von: Riou, Alain, et al.
Veröffentlicht: (2024)
Evaluating Fake Music Detection Performance Under Audio Augmentations
von: Sroka, Tomasz, et al.
Veröffentlicht: (2025)
von: Sroka, Tomasz, et al.
Veröffentlicht: (2025)
Lyrics Matter: Exploiting the Power of Learnt Representations for Music Popularity Prediction
von: Choudhary, Yash, et al.
Veröffentlicht: (2025)
von: Choudhary, Yash, et al.
Veröffentlicht: (2025)
Aria-MIDI: A Dataset of Piano MIDI Files for Symbolic Music Modeling
von: Bradshaw, Louis, et al.
Veröffentlicht: (2025)
von: Bradshaw, Louis, et al.
Veröffentlicht: (2025)
KAD: No More FAD! An Effective and Efficient Evaluation Metric for Audio Generation
von: Chung, Yoonjin, et al.
Veröffentlicht: (2025)
von: Chung, Yoonjin, et al.
Veröffentlicht: (2025)
AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs
von: Nguyen, Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Hoang, et al.
Veröffentlicht: (2025)
Synthetic Data Augmentation for Medical Audio Classification: A Preliminary Evaluation
von: McShannon, David, et al.
Veröffentlicht: (2026)
von: McShannon, David, et al.
Veröffentlicht: (2026)
aTENNuate: Optimized Real-time Speech Enhancement with Deep SSMs on Raw Audio
von: Pei, Yan Ru, et al.
Veröffentlicht: (2024)
von: Pei, Yan Ru, et al.
Veröffentlicht: (2024)
Jailbreak-AudioBench: In-Depth Evaluation and Analysis of Jailbreak Threats for Large Audio Language Models
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
Guiding Audio Editing with Audio Language Model
von: Lan, Zitong, et al.
Veröffentlicht: (2025)
von: Lan, Zitong, et al.
Veröffentlicht: (2025)
MAEB: Massive Audio Embedding Benchmark
von: Assadi, Adnan El, et al.
Veröffentlicht: (2026)
von: Assadi, Adnan El, et al.
Veröffentlicht: (2026)
Survey on the Evaluation of Generative Models in Music
von: Lerch, Alexander, et al.
Veröffentlicht: (2025)
von: Lerch, Alexander, et al.
Veröffentlicht: (2025)
SUBARU: A Practical Approach to Power Saving in Hearables Using SUB-Nyquist Audio Resolution Upsampling
von: Tamiti, Tarikul Islam, et al.
Veröffentlicht: (2025)
von: Tamiti, Tarikul Islam, et al.
Veröffentlicht: (2025)
Synthetic Audio Helps for Cognitive State Tasks
von: Soubki, Adil, et al.
Veröffentlicht: (2025)
von: Soubki, Adil, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Representation-Based Data Quality Audits for Audio
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025) -
Clinical Uncertainty Impacts Machine Learning Evaluations
von: Lionetti, Simone, et al.
Veröffentlicht: (2025) -
Is Hyperbolic Space All You Need for Medical Anomaly Detection?
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025) -
AudioMosaic: Contrastive Masked Audio Representation Learning
von: Huang, Hanxun, et al.
Veröffentlicht: (2026) -
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)