Gespeichert in:
| Hauptverfasser: | Rabuge, Miguel, Lourenço, Nuno |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2502.08785 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Single Microphone Own Voice Detection based on Simulated Transfer Functions for Hearing Aids
von: Mayuravaani, Mathuranathan, et al.
Veröffentlicht: (2026)
von: Mayuravaani, Mathuranathan, et al.
Veröffentlicht: (2026)
Fine-grained Soundscape Control for Augmented Hearing
von: Oh, Seunghyun, et al.
Veröffentlicht: (2026)
von: Oh, Seunghyun, et al.
Veröffentlicht: (2026)
Developing an AI-Guided Assistant Device for the Deaf and Hearing Impaired
von: Jiayu, et al.
Veröffentlicht: (2025)
von: Jiayu, et al.
Veröffentlicht: (2025)
Non-Intrusive Speech Intelligibility Prediction for Hearing Aids using Whisper and Metadata
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2023)
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2023)
Improving Machine Hearing on Limited Data Sets
von: Harar, Pavol, et al.
Veröffentlicht: (2019)
von: Harar, Pavol, et al.
Veröffentlicht: (2019)
Text-Independent Speaker Identification Using Audio Looping With Margin Based Loss Functions
von: Garcia, Elliot Q C, et al.
Veröffentlicht: (2025)
von: Garcia, Elliot Q C, et al.
Veröffentlicht: (2025)
HAAQI-Net: A Non-intrusive Neural Music Audio Quality Assessment Model for Hearing Aids
von: Wisnu, Dyah A. M. G., et al.
Veröffentlicht: (2024)
von: Wisnu, Dyah A. M. G., et al.
Veröffentlicht: (2024)
AuditoryBench++: Can Language Models Understand Auditory Knowledge without Hearing?
von: Ok, Hyunjong, et al.
Veröffentlicht: (2025)
von: Ok, Hyunjong, et al.
Veröffentlicht: (2025)
Remixing Music for Hearing Aids Using Ensemble of Fine-Tuned Source Separators
von: Daly, Matthew
Veröffentlicht: (2024)
von: Daly, Matthew
Veröffentlicht: (2024)
Hearing Your Blood Sugar: Non-Invasive Glucose Measurement Through Simple Vocal Signals, Transforming any Speech into a Sensor with Machine Learning
von: Ahmadli, Nihat, et al.
Veröffentlicht: (2024)
von: Ahmadli, Nihat, et al.
Veröffentlicht: (2024)
Count The Notes: Histogram-Based Supervision for Automatic Music Transcription
von: Yaffe, Jonathan, et al.
Veröffentlicht: (2025)
von: Yaffe, Jonathan, et al.
Veröffentlicht: (2025)
Hearing Anywhere in Any Environment
von: Liu, Xiulong, et al.
Veröffentlicht: (2025)
von: Liu, Xiulong, et al.
Veröffentlicht: (2025)
Transformer Based Machine Fault Detection From Audio Input
von: Holla, Kiran Voderhobli
Veröffentlicht: (2026)
von: Holla, Kiran Voderhobli
Veröffentlicht: (2026)
EmoHRNet: High-Resolution Neural Network Based Speech Emotion Recognition
von: Muppidi, Akshay, et al.
Veröffentlicht: (2025)
von: Muppidi, Akshay, et al.
Veröffentlicht: (2025)
Audio Question Answering with GRPO-Based Fine-Tuning and Calibrated Segment-Level Predictions
von: Gibier, Marcel, et al.
Veröffentlicht: (2025)
von: Gibier, Marcel, et al.
Veröffentlicht: (2025)
Revisit Modality Imbalance at the Decision Layer
von: Ma, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Ma, Xiaoyu, et al.
Veröffentlicht: (2025)
How to Count Coughs: An Event-Based Framework for Evaluating Automatic Cough Detection Algorithm Performance
von: Orlandic, Lara, et al.
Veröffentlicht: (2024)
von: Orlandic, Lara, et al.
Veröffentlicht: (2024)
Improving Perceptual Audio Aesthetic Assessment via Triplet Loss and Self-Supervised Embeddings
von: Wisnu, Dyah A. M. G., et al.
Veröffentlicht: (2025)
von: Wisnu, Dyah A. M. G., et al.
Veröffentlicht: (2025)
GE2E-AC: Generalized End-to-End Loss Training for Accent Classification
von: Watanabe, Chihiro, et al.
Veröffentlicht: (2024)
von: Watanabe, Chihiro, et al.
Veröffentlicht: (2024)
Improving Out-of-Domain Audio Deepfake Detection via Layer Selection and Fusion of SSL-Based Countermeasures
von: Serrano, Pierre, et al.
Veröffentlicht: (2025)
von: Serrano, Pierre, et al.
Veröffentlicht: (2025)
Automatic Identification of Samples in Hip-Hop Music via Multi-Loss Training and an Artificial Dataset
von: Cheston, Huw, et al.
Veröffentlicht: (2025)
von: Cheston, Huw, et al.
Veröffentlicht: (2025)
Monaural Speech Enhancement with Complex Convolutional Block Attention Module and Joint Time Frequency Losses
von: Zhao, Shengkui, et al.
Veröffentlicht: (2021)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2021)
Morse Code-Enabled Speech Recognition for Individuals with Visual and Hearing Impairments
von: Choudhury, Ritabrata Roy
Veröffentlicht: (2024)
von: Choudhury, Ritabrata Roy
Veröffentlicht: (2024)
Hybrid Losses for Hierarchical Embedding Learning
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
Patient-Aware Feature Alignment for Robust Lung Sound Classification:Cohesion-Separation and Global Alignment Losses
von: Jeong, Seung Gyu, et al.
Veröffentlicht: (2025)
von: Jeong, Seung Gyu, et al.
Veröffentlicht: (2025)
Losses Can Be Blessings: Routing Self-Supervised Speech Representations Towards Efficient Multilingual and Multitask Speech Processing
von: Fu, Yonggan, et al.
Veröffentlicht: (2022)
von: Fu, Yonggan, et al.
Veröffentlicht: (2022)
Imagine to Hear: Auditory Knowledge Generation can be an Effective Assistant for Language Models
von: Yoo, Suho, et al.
Veröffentlicht: (2025)
von: Yoo, Suho, et al.
Veröffentlicht: (2025)
What Do Language Models Hear? Probing for Auditory Representations in Language Models
von: Ngo, Jerry, et al.
Veröffentlicht: (2024)
von: Ngo, Jerry, et al.
Veröffentlicht: (2024)
On the Condition Monitoring of Bolted Joints through Acoustic Emission and Deep Transfer Learning: Generalization, Ordinal Loss and Super-Convergence
von: Ramasso, Emmanuel, et al.
Veröffentlicht: (2024)
von: Ramasso, Emmanuel, et al.
Veröffentlicht: (2024)
Pruning-aware Loss Functions for STOI-Optimized Pruned Recurrent Autoencoders for the Compression of the Stimulation Patterns of Cochlear Implants at Zero Delay
von: Hinrichs, Reemt, et al.
Veröffentlicht: (2025)
von: Hinrichs, Reemt, et al.
Veröffentlicht: (2025)
An Independence-promoting Loss for Music Generation with Language Models
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
Hear What Matters! Text-conditioned Selective Video-to-Audio Generation
von: Lee, Junwon, et al.
Veröffentlicht: (2025)
von: Lee, Junwon, et al.
Veröffentlicht: (2025)
Testing chatbots on the creation of encoders for audio conditioned image generation
von: León, Jorge E., et al.
Veröffentlicht: (2025)
von: León, Jorge E., et al.
Veröffentlicht: (2025)
What You Read Isn't What You Hear: Linguistic Sensitivity in Deepfake Speech Detection
von: Nguyen, Binh, et al.
Veröffentlicht: (2025)
von: Nguyen, Binh, et al.
Veröffentlicht: (2025)
An Attention Long Short-Term Memory based system for automatic classification of speech intelligibility
von: Fernández-Díaz, Miguel, et al.
Veröffentlicht: (2024)
von: Fernández-Díaz, Miguel, et al.
Veröffentlicht: (2024)
Improving Membership Inference in ASR Model Auditing with Perturbed Loss Features
von: Teixeira, Francisco, et al.
Veröffentlicht: (2024)
von: Teixeira, Francisco, et al.
Veröffentlicht: (2024)
Representation-Based Data Quality Audits for Audio
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025)
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025)
Respiratory Disease Classification and Biometric Analysis Using Biosignals from Digital Stethoscopes
von: Casado, Constantino Álvarez, et al.
Veröffentlicht: (2023)
von: Casado, Constantino Álvarez, et al.
Veröffentlicht: (2023)
Enhanced ASR Robustness to Packet Loss with a Front-End Adaptation Network
von: Dissen, Yehoshua, et al.
Veröffentlicht: (2024)
von: Dissen, Yehoshua, et al.
Veröffentlicht: (2024)
Focal Loss based Residual Convolutional Neural Network for Speech Emotion Recognition
von: Tripathi, Suraj, et al.
Veröffentlicht: (2019)
von: Tripathi, Suraj, et al.
Veröffentlicht: (2019)
Ähnliche Einträge
-
Single Microphone Own Voice Detection based on Simulated Transfer Functions for Hearing Aids
von: Mayuravaani, Mathuranathan, et al.
Veröffentlicht: (2026) -
Fine-grained Soundscape Control for Augmented Hearing
von: Oh, Seunghyun, et al.
Veröffentlicht: (2026) -
Developing an AI-Guided Assistant Device for the Deaf and Hearing Impaired
von: Jiayu, et al.
Veröffentlicht: (2025) -
Non-Intrusive Speech Intelligibility Prediction for Hearing Aids using Whisper and Metadata
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2023) -
Improving Machine Hearing on Limited Data Sets
von: Harar, Pavol, et al.
Veröffentlicht: (2019)