Emovectors: assessing emotional content in jazz improvisations for creativity evaluation
Fuente:
arXiv
Guardado en:
| Autor principal: | Jordanous, Anna |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Beyond saliency: enhancing explanation of speech emotion recognition with expert-referenced acoustic cues
por: Nasr, Seham, et al.
Publicado: (2025)
por: Nasr, Seham, et al.
Publicado: (2025)
Heterogeneous bimodal attention fusion for speech emotion recognition
por: Luo, Jiachen, et al.
Publicado: (2025)
por: Luo, Jiachen, et al.
Publicado: (2025)
Improving speaker verification robustness with synthetic emotional utterances
por: Koditala, Nikhil Kumar, et al.
Publicado: (2024)
por: Koditala, Nikhil Kumar, et al.
Publicado: (2024)
Latent Fourier Transform
por: Wang, Mason, et al.
Publicado: (2026)
por: Wang, Mason, et al.
Publicado: (2026)
learning discriminative features from spectrograms using center loss for speech emotion recognition
por: Dai, Dongyang, et al.
Publicado: (2025)
por: Dai, Dongyang, et al.
Publicado: (2025)
Ensemble of classifiers for speech evaluation
por: Belokrylov, G., et al.
Publicado: (2024)
por: Belokrylov, G., et al.
Publicado: (2024)
Harmonic Reasoning in Large Language Models
por: Kruspe, Anna
Publicado: (2024)
por: Kruspe, Anna
Publicado: (2024)
Adaptive Accompaniment with ReaLchords
por: Wu, Yusong, et al.
Publicado: (2025)
por: Wu, Yusong, et al.
Publicado: (2025)
Automated evaluation of children's speech fluency for low-resource languages
por: Zhang, Bowen, et al.
Publicado: (2025)
por: Zhang, Bowen, et al.
Publicado: (2025)
Deep learning for music generation. Four approaches and their comparative evaluation
por: Paroiu, Razvan, et al.
Publicado: (2025)
por: Paroiu, Razvan, et al.
Publicado: (2025)
Modeling speech emotion with label variance and analyzing performance across speakers and unseen acoustic conditions
por: Mitra, Vikramjit, et al.
Publicado: (2025)
por: Mitra, Vikramjit, et al.
Publicado: (2025)
DialogGraph-LLM: Graph-Informed LLMs for End-to-End Audio Dialogue Intent Recognition
por: Liu, HongYu, et al.
Publicado: (2025)
por: Liu, HongYu, et al.
Publicado: (2025)
Listen Like a Teacher: Mitigating Whisper Hallucinations using Adaptive Layer Attention and Knowledge Distillation
por: Tripathi, Kumud, et al.
Publicado: (2025)
por: Tripathi, Kumud, et al.
Publicado: (2025)
ERIS: Evolutionary Real-world Interference Scheme for Jailbreaking Audio Large Models
por: Zhang, Yibo, et al.
Publicado: (2025)
por: Zhang, Yibo, et al.
Publicado: (2025)
JamendoMaxCaps: A Large Scale Music-caption Dataset with Imputed Metadata
por: Roy, Abhinaba, et al.
Publicado: (2025)
por: Roy, Abhinaba, et al.
Publicado: (2025)
LoopGen: Training-Free Loopable Music Generation
por: Marincione, Davide, et al.
Publicado: (2025)
por: Marincione, Davide, et al.
Publicado: (2025)
Improving Anomalous Sound Detection with Attribute-aware Representation from Domain-adaptive Pre-training
por: Fang, Xin, et al.
Publicado: (2025)
por: Fang, Xin, et al.
Publicado: (2025)
MATPAC++: Enhanced Masked Latent Prediction for Self-Supervised Audio Representation Learning
por: Quelennec, Aurian, et al.
Publicado: (2025)
por: Quelennec, Aurian, et al.
Publicado: (2025)
LSZone: A Lightweight Spatial Information Modeling Architecture for Real-time In-car Multi-zone Speech Separation
por: Chen, Jun, et al.
Publicado: (2025)
por: Chen, Jun, et al.
Publicado: (2025)
ArtiFree: Detecting and Reducing Generative Artifacts in Diffusion-based Speech Enhancement
por: Chhaglani, Bhawana, et al.
Publicado: (2025)
por: Chhaglani, Bhawana, et al.
Publicado: (2025)
Eliminating stability hallucinations in llm-based tts models via attention guidance
por: Wang, ShiMing, et al.
Publicado: (2025)
por: Wang, ShiMing, et al.
Publicado: (2025)
More Than A Shortcut: A Hyperbolic Approach To Early-Exit Networks
por: Bhosale, Swapnil, et al.
Publicado: (2025)
por: Bhosale, Swapnil, et al.
Publicado: (2025)
YingMusic-Singer: Zero-shot Singing Voice Synthesis and Editing with Annotation-free Melody Guidance
por: Zheng, Junjie, et al.
Publicado: (2025)
por: Zheng, Junjie, et al.
Publicado: (2025)
Diffusion-based Surrogate Model for Time-varying Underwater Acoustic Channels
por: Li, Kexin, et al.
Publicado: (2025)
por: Li, Kexin, et al.
Publicado: (2025)
Masked Latent Prediction and Classification for Self-Supervised Audio Representation Learning
por: Quelennec, Aurian, et al.
Publicado: (2025)
por: Quelennec, Aurian, et al.
Publicado: (2025)
AudioMoG: Guiding Audio Generation with Mixture-of-Guidance
por: Wang, Junyou, et al.
Publicado: (2025)
por: Wang, Junyou, et al.
Publicado: (2025)
State Space Models for Bioacoustics: A Comparative Evaluation with Transformers
por: Tang, Chengyu, et al.
Publicado: (2025)
por: Tang, Chengyu, et al.
Publicado: (2025)
Graph Embedding with Mel-spectrograms for Underwater Acoustic Target Recognition
por: Feng, Sheng, et al.
Publicado: (2025)
por: Feng, Sheng, et al.
Publicado: (2025)
AnalysisGNN: Unified Music Analysis with Graph Neural Networks
por: Karystinaios, Emmanouil, et al.
Publicado: (2025)
por: Karystinaios, Emmanouil, et al.
Publicado: (2025)
Empowering Global Voices: A Data-Efficient, Phoneme-Tone Adaptive Approach to High-Fidelity Speech Synthesis
por: Geng, Yizhong, et al.
Publicado: (2025)
por: Geng, Yizhong, et al.
Publicado: (2025)
Audio-Maestro: Enhancing Large Audio-Language Models with Tool-Augmented Reasoning
por: Lee, Kuan-Yi, et al.
Publicado: (2025)
por: Lee, Kuan-Yi, et al.
Publicado: (2025)
Twenty-Five Years of MIR Research: Achievements, Practices, Evaluations, and Future Challenges
por: Peeters, Geoffroy, et al.
Publicado: (2025)
por: Peeters, Geoffroy, et al.
Publicado: (2025)
'Studies for': A Human-AI Co-Creative Sound Artwork Using a Real-time Multi-channel Sound Generation Model
por: Nagashima, Chihiro, et al.
Publicado: (2025)
por: Nagashima, Chihiro, et al.
Publicado: (2025)
DDSC: Dynamic Dual-Signal Curriculum for Data-Efficient Acoustic Scene Classification under Domain Shift
por: Zhang, Peihong, et al.
Publicado: (2025)
por: Zhang, Peihong, et al.
Publicado: (2025)
Formula-Supervised Sound Event Detection: Pre-Training Without Real Data
por: Shibata, Yuto, et al.
Publicado: (2025)
por: Shibata, Yuto, et al.
Publicado: (2025)
MUSE-Explainer: Counterfactual Explanations for Symbolic Music Graph Classification Models
por: Hilaire, Baptiste, et al.
Publicado: (2025)
por: Hilaire, Baptiste, et al.
Publicado: (2025)
An Agent-Based Framework for Automated Higher-Voice Harmony Generation
por: Ganapathy, Nia D'Souza, et al.
Publicado: (2025)
por: Ganapathy, Nia D'Souza, et al.
Publicado: (2025)
PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation
por: Xie, Tianxin, et al.
Publicado: (2025)
por: Xie, Tianxin, et al.
Publicado: (2025)
RFM-Editing: Rectified Flow Matching for Text-guided Audio Editing
por: Gao, Liting, et al.
Publicado: (2025)
por: Gao, Liting, et al.
Publicado: (2025)
ABC-Eval: Benchmarking Large Language Models on Symbolic Music Understanding and Instruction Following
por: Zhao, Jiahao, et al.
Publicado: (2025)
por: Zhao, Jiahao, et al.
Publicado: (2025)
Ejemplares similares
-
Beyond saliency: enhancing explanation of speech emotion recognition with expert-referenced acoustic cues
por: Nasr, Seham, et al.
Publicado: (2025) -
Heterogeneous bimodal attention fusion for speech emotion recognition
por: Luo, Jiachen, et al.
Publicado: (2025) -
Improving speaker verification robustness with synthetic emotional utterances
por: Koditala, Nikhil Kumar, et al.
Publicado: (2024) -
Latent Fourier Transform
por: Wang, Mason, et al.
Publicado: (2026) -
learning discriminative features from spectrograms using center loss for speech emotion recognition
por: Dai, Dongyang, et al.
Publicado: (2025)