Spatio-temporal Latent Representations for the Analysis of Acoustic Scenes in-the-wild
Fuente:
arXiv
Guardado en:
| Autores principales: | Montero-Ramírez, Claudia, Rituerto-González, Esther, Peláez-Moreno, Carmen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Decoding Vocal Articulations from Acoustic Latent Representations
por: Cámara, Mateo, et al.
Publicado: (2024)
por: Cámara, Mateo, et al.
Publicado: (2024)
Sub-band Domain Multi-Hypothesis Acoustic Echo Canceler Based Acoustic Scene Analysis
por: Southwell, Benjamin J, et al.
Publicado: (2025)
por: Southwell, Benjamin J, et al.
Publicado: (2025)
Scene-wide Acoustic Parameter Estimation
por: Falcon-Perez, Ricardo, et al.
Publicado: (2024)
por: Falcon-Perez, Ricardo, et al.
Publicado: (2024)
Leveraging Self-supervised Audio Representations for Data-Efficient Acoustic Scene Classification
por: Cai, Yiqiang, et al.
Publicado: (2024)
por: Cai, Yiqiang, et al.
Publicado: (2024)
Acoustic Scene Classification Using CNN-GRU Model Without Knowledge Distillation
por: Tan, Ee-Leng, et al.
Publicado: (2025)
por: Tan, Ee-Leng, et al.
Publicado: (2025)
Leveraging Content and Acoustic Representations for Speech Emotion Recognition
por: Dutta, Soumya, et al.
Publicado: (2024)
por: Dutta, Soumya, et al.
Publicado: (2024)
Speak in the Scene: Diffusion-based Acoustic Scene Transfer toward Immersive Speech Generation
por: Kim, Miseul, et al.
Publicado: (2024)
por: Kim, Miseul, et al.
Publicado: (2024)
Improving Acoustic Scene Classification in Low-Resource Conditions
por: Chen, Zhi, et al.
Publicado: (2024)
por: Chen, Zhi, et al.
Publicado: (2024)
Acoustic Teleportation via Disentangled Neural Audio Codec Representations
por: Grundhuber, Philipp, et al.
Publicado: (2025)
por: Grundhuber, Philipp, et al.
Publicado: (2025)
Blind Acoustic Parameter Estimation Through Task-Agnostic Embeddings Using Latent Approximations
por: Götz, Philipp, et al.
Publicado: (2024)
por: Götz, Philipp, et al.
Publicado: (2024)
Joint Analysis of Acoustic Scenes and Sound Events Based on Semi-Supervised Training of Sound Events With Partial Labels
por: Imoto, Keisuke
Publicado: (2025)
por: Imoto, Keisuke
Publicado: (2025)
HASRD: Hierarchical Acoustic and Semantic Representation Disentanglement
por: Hussein, Amir, et al.
Publicado: (2025)
por: Hussein, Amir, et al.
Publicado: (2025)
XANE Background Acoustic Embeddings: Ablation and Clustering Analysis
por: Sharma, Dushyant, et al.
Publicado: (2024)
por: Sharma, Dushyant, et al.
Publicado: (2024)
Evaluation of preprocessing pipelines in the creation of in-the-wild TTS datasets
por: Di Bernardo, Matías, et al.
Publicado: (2025)
por: Di Bernardo, Matías, et al.
Publicado: (2025)
Low-Complexity Acoustic Scene Classification Using Parallel Attention-Convolution Network
por: Li, Yanxiong, et al.
Publicado: (2024)
por: Li, Yanxiong, et al.
Publicado: (2024)
Data-Efficient Low-Complexity Acoustic Scene Classification in the DCASE 2024 Challenge
por: Schmid, Florian, et al.
Publicado: (2024)
por: Schmid, Florian, et al.
Publicado: (2024)
Online Domain-Incremental Learning Approach to Classify Acoustic Scenes in All Locations
por: Mulimani, Manjunath, et al.
Publicado: (2024)
por: Mulimani, Manjunath, et al.
Publicado: (2024)
Low-Complexity Acoustic Scene Classification with Device Information in the DCASE 2025 Challenge
por: Schmid, Florian, et al.
Publicado: (2025)
por: Schmid, Florian, et al.
Publicado: (2025)
Leveraging Discriminative Latent Representations for Conditioning GAN-Based Speech Enhancement
por: Shetu, Shrishti Saha, et al.
Publicado: (2025)
por: Shetu, Shrishti Saha, et al.
Publicado: (2025)
Semantic-VAE: Semantic-Alignment Latent Representation for Better Speech Synthesis
por: Niu, Zhikang, et al.
Publicado: (2025)
por: Niu, Zhikang, et al.
Publicado: (2025)
On the Relation Between Speech Quality and Quantized Latent Representations of Neural Codecs
por: Halimeh, Mhd Modar, et al.
Publicado: (2025)
por: Halimeh, Mhd Modar, et al.
Publicado: (2025)
Variational Bayesian Adaptive Learning of Deep Latent Variables for Acoustic Knowledge Transfer
por: Hu, Hu, et al.
Publicado: (2025)
por: Hu, Hu, et al.
Publicado: (2025)
Lightweight and Generalizable Acoustic Scene Representations via Contrastive Fine-Tuning and Distillation
por: Yuan, Kuang, et al.
Publicado: (2025)
por: Yuan, Kuang, et al.
Publicado: (2025)
Disentangled Acoustic Fields For Multimodal Physical Scene Understanding
por: Yin, Jie, et al.
Publicado: (2024)
por: Yin, Jie, et al.
Publicado: (2024)
Data-Efficient Low-Complexity Acoustic Scene Classification via Distilling and Progressive Pruning
por: Han, Bing, et al.
Publicado: (2024)
por: Han, Bing, et al.
Publicado: (2024)
Perceptual evaluation of Acoustic Level of Detail in Virtual Acoustic Environments
por: Fichna, Stefan, et al.
Publicado: (2025)
por: Fichna, Stefan, et al.
Publicado: (2025)
Investigation of Speech and Noise Latent Representations in Single-channel VAE-based Speech Enhancement
por: Li, Jiatong, et al.
Publicado: (2025)
por: Li, Jiatong, et al.
Publicado: (2025)
Towards Perception-Informed Latent HRTF Representations
por: Zhang, You, et al.
Publicado: (2025)
por: Zhang, You, et al.
Publicado: (2025)
Benchmarking Representations for Speech, Music, and Acoustic Events
por: La Quatra, Moreno, et al.
Publicado: (2024)
por: La Quatra, Moreno, et al.
Publicado: (2024)
NAT: Neural Acoustic Transfer for Interactive Scenes in Real Time
por: Jin, Xutong, et al.
Publicado: (2025)
por: Jin, Xutong, et al.
Publicado: (2025)
Unsupervised Acoustic Scene Mapping Based on Acoustic Features and Dimensionality Reduction
por: Cohen, Idan, et al.
Publicado: (2023)
por: Cohen, Idan, et al.
Publicado: (2023)
One-Shot Distributed Node-Specific Signal Estimation with Non-Overlapping Latent Subspaces in Acoustic Sensor Networks
por: Didier, Paul, et al.
Publicado: (2024)
por: Didier, Paul, et al.
Publicado: (2024)
Listen through the Sound: Generative Speech Restoration Leveraging Acoustic Context Representation
por: Chung, Soo-Whan, et al.
Publicado: (2025)
por: Chung, Soo-Whan, et al.
Publicado: (2025)
TF-SepNet: An Efficient 1D Kernel Design in CNNs for Low-Complexity Acoustic Scene Classification
por: Cai, Yiqiang, et al.
Publicado: (2023)
por: Cai, Yiqiang, et al.
Publicado: (2023)
Causal Spatio-Temporal Sound Field Reconstruction
por: Sundström, David, et al.
Publicado: (2026)
por: Sundström, David, et al.
Publicado: (2026)
SAMOS: A Neural MOS Prediction Model Leveraging Semantic Representations and Acoustic Features
por: Shi, Yu-Fei, et al.
Publicado: (2024)
por: Shi, Yu-Fei, et al.
Publicado: (2024)
Why Can't They Remember? Uncovering Representation and Retrieval Bottlenecks in Multi-Turn Acoustic Memory
por: Xiao, Yang, et al.
Publicado: (2026)
por: Xiao, Yang, et al.
Publicado: (2026)
Categorical Unsupervised Variational Acoustic Clustering
por: Fiorio, Luan Vinícius, et al.
Publicado: (2025)
por: Fiorio, Luan Vinícius, et al.
Publicado: (2025)
On the Use of Dereverberation for Acoustic Feedback Cancellation
por: Liekens, Basil, et al.
Publicado: (2026)
por: Liekens, Basil, et al.
Publicado: (2026)
Acoustic and Semantic Modeling of Emotion in Spoken Language
por: Dutta, Soumya
Publicado: (2026)
por: Dutta, Soumya
Publicado: (2026)
Ejemplares similares
-
Decoding Vocal Articulations from Acoustic Latent Representations
por: Cámara, Mateo, et al.
Publicado: (2024) -
Sub-band Domain Multi-Hypothesis Acoustic Echo Canceler Based Acoustic Scene Analysis
por: Southwell, Benjamin J, et al.
Publicado: (2025) -
Scene-wide Acoustic Parameter Estimation
por: Falcon-Perez, Ricardo, et al.
Publicado: (2024) -
Leveraging Self-supervised Audio Representations for Data-Efficient Acoustic Scene Classification
por: Cai, Yiqiang, et al.
Publicado: (2024) -
Acoustic Scene Classification Using CNN-GRU Model Without Knowledge Distillation
por: Tan, Ee-Leng, et al.
Publicado: (2025)