Saved in:
| Main Authors: | Milling, Manuel, Triantafyllopoulos, Andreas, Gebhard, Alexander, Rampp, Simon, Schuller, Björn W. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2603.25476 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Does the Definition of Difficulty Matter? Scoring Functions and their Role for Curriculum Learning
by: Rampp, Simon, et al.
Published: (2024)
by: Rampp, Simon, et al.
Published: (2024)
autrainer: A Modular and Extensible Deep Learning Toolkit for Computer Audition Tasks
by: Rampp, Simon, et al.
Published: (2024)
by: Rampp, Simon, et al.
Published: (2024)
An automatic analysis of ultrasound vocalisations for the prediction of interaction context in captive Egyptian fruit bats
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
Bringing the Discussion of Minima Sharpness to the Audio Domain: a Filter-Normalised Evaluation for Acoustic Scene Classification
by: Milling, Manuel, et al.
Published: (2023)
by: Milling, Manuel, et al.
Published: (2023)
INTERSPEECH 2009 Emotion Challenge Revisited: Benchmarking 15 Years of Progress in Speech Emotion Recognition
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
Exploring Meta Information for Audio-based Zero-shot Bird Classification
by: Gebhard, Alexander, et al.
Published: (2023)
by: Gebhard, Alexander, et al.
Published: (2023)
Audio-based Step-count Estimation for Running -- Windowing and Neural Network Baselines
by: Wagner, Philipp, et al.
Published: (2024)
by: Wagner, Philipp, et al.
Published: (2024)
Audio Enhancement for Computer Audition -- An Iterative Training Paradigm Using Sample Importance
by: Milling, Manuel, et al.
Published: (2024)
by: Milling, Manuel, et al.
Published: (2024)
Computer Audition: From Task-Specific Machine Learning to Foundation Models
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
Expressivity and Speech Synthesis
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
Enrolment-based personalisation for improving individual-level fairness in speech emotion recognition
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
Enrolment-based personalisation for improving individual-level fairness in speech emotion recognition
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)
DFingerNet: Noise-Adaptive Speech Enhancement for Hearing Aids
by: Tsangko, Iosif, et al.
Published: (2025)
by: Tsangko, Iosif, et al.
Published: (2025)
Neuroplasticity in Artificial Intelligence -- An Overview and Inspirations on Drop In & Out Learning
by: Li, Yupei, et al.
Published: (2025)
by: Li, Yupei, et al.
Published: (2025)
From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview
by: Li, Yupei, et al.
Published: (2024)
by: Li, Yupei, et al.
Published: (2024)
Enhancing Efficiency and Performance in Deepfake Audio Detection through Neuron-level Dropin & Neuroplasticity Mechanisms
by: Li, Yupei, et al.
Published: (2026)
by: Li, Yupei, et al.
Published: (2026)
Detecting COPD Through Speech Analysis: A Dataset of Danish Speech and Machine Learning Approach
by: Sankey-Olsen, Cuno, et al.
Published: (2025)
by: Sankey-Olsen, Cuno, et al.
Published: (2025)
Affect and Effect: Limitations of regularisation-based continual learning in EEG-based emotion classification
by: Peire, Nina, et al.
Published: (2026)
by: Peire, Nina, et al.
Published: (2026)
Audio-based Kinship Verification Using Age Domain Conversion
by: Sun, Qiyang, et al.
Published: (2024)
by: Sun, Qiyang, et al.
Published: (2024)
Normalise for Fairness: A Simple Normalisation Technique for Fairness in Regression Machine Learning Problems
by: Amin, Mostafa M., et al.
Published: (2022)
by: Amin, Mostafa M., et al.
Published: (2022)
Charting 15 years of progress in deep learning for speech emotion recognition: A replication study
by: Triantafyllopoulos, Andreas, et al.
Published: (2025)
by: Triantafyllopoulos, Andreas, et al.
Published: (2025)
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
by: Jing, Xin, et al.
Published: (2024)
by: Jing, Xin, et al.
Published: (2024)
Learning to Pay Attention: Unsupervised Modeling of Attentive and Inattentive Respondents in Survey Data
by: Triantafyllopoulos, Ilias, et al.
Published: (2026)
by: Triantafyllopoulos, Ilias, et al.
Published: (2026)
Enhancing Emotional Text-to-Speech Controllability with Natural Language Guidance through Contrastive Learning and Diffusion Models
by: Jing, Xin, et al.
Published: (2024)
by: Jing, Xin, et al.
Published: (2024)
Discourse Features Enhance Detection of Document-Level Machine-Generated Content
by: Li, Yupei, et al.
Published: (2024)
by: Li, Yupei, et al.
Published: (2024)
Large Language Models for Depression Recognition in Spoken Language Integrating Psychological Knowledge
by: Li, Yupei, et al.
Published: (2025)
by: Li, Yupei, et al.
Published: (2025)
How Intermodal Interaction Affects the Performance of Deep Multimodal Fusion for Mixed-Type Time Series
by: Dietz, Simon, et al.
Published: (2024)
by: Dietz, Simon, et al.
Published: (2024)
Modeling Emotional Trajectories in Written Stories Utilizing Transformers and Weakly-Supervised Learning
by: Christ, Lukas, et al.
Published: (2024)
by: Christ, Lukas, et al.
Published: (2024)
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
by: Rajapakshe, Thejan, et al.
Published: (2022)
by: Rajapakshe, Thejan, et al.
Published: (2022)
Explainable Artificial Intelligence for Medical Applications: A Review
by: Sun, Qiyang, et al.
Published: (2024)
by: Sun, Qiyang, et al.
Published: (2024)
MELT: Towards Automated Multimodal Emotion Data Annotation by Leveraging LLM Embedded Knowledge
by: Jing, Xin, et al.
Published: (2025)
by: Jing, Xin, et al.
Published: (2025)
Representation Learning with Parameterised Quantum Circuits for Advancing Speech Emotion Recognition
by: Rajapakshe, Thejan, et al.
Published: (2025)
by: Rajapakshe, Thejan, et al.
Published: (2025)
Automatic Emotion Modelling in Written Stories
by: Christ, Lukas, et al.
Published: (2022)
by: Christ, Lukas, et al.
Published: (2022)
Abusive Speech Detection in Indic Languages Using Acoustic Features
by: Spiesberger, Anika A., et al.
Published: (2024)
by: Spiesberger, Anika A., et al.
Published: (2024)
How False Data Affects Machine Learning Models in Electrochemistry?
by: Deshsorna, Krittapong, et al.
Published: (2023)
by: Deshsorna, Krittapong, et al.
Published: (2023)
Are you sure? Analysing Uncertainty Quantification Approaches for Real-world Speech Emotion Recognition
by: Schrüfer, Oliver, et al.
Published: (2024)
by: Schrüfer, Oliver, et al.
Published: (2024)
Towards Multimodal Prediction of Spontaneous Humour: A Novel Dataset and First Results
by: Christ, Lukas, et al.
Published: (2022)
by: Christ, Lukas, et al.
Published: (2022)
SmoothCLAP: Soft-Target Enhanced Contrastive Language\--Audio Pretraining for Affective Computing
by: Jing, Xin, et al.
Published: (2026)
by: Jing, Xin, et al.
Published: (2026)
Large-Scale Dataset Pruning in Adversarial Training through Data Importance Extrapolation
by: Nieth, Björn, et al.
Published: (2024)
by: Nieth, Björn, et al.
Published: (2024)
A Pilot Study on Curator-Guided Multilingual Art Description for Blind and Low-Vision Audiences with Small Vision-Language Models
by: Tsangko, Iosif, et al.
Published: (2026)
by: Tsangko, Iosif, et al.
Published: (2026)
Similar Items
-
Does the Definition of Difficulty Matter? Scoring Functions and their Role for Curriculum Learning
by: Rampp, Simon, et al.
Published: (2024) -
autrainer: A Modular and Extensible Deep Learning Toolkit for Computer Audition Tasks
by: Rampp, Simon, et al.
Published: (2024) -
An automatic analysis of ultrasound vocalisations for the prediction of interaction context in captive Egyptian fruit bats
by: Triantafyllopoulos, Andreas, et al.
Published: (2024) -
Bringing the Discussion of Minima Sharpness to the Audio Domain: a Filter-Normalised Evaluation for Acoustic Scene Classification
by: Milling, Manuel, et al.
Published: (2023) -
INTERSPEECH 2009 Emotion Challenge Revisited: Benchmarking 15 Years of Progress in Speech Emotion Recognition
by: Triantafyllopoulos, Andreas, et al.
Published: (2024)