Sentiment analysis in non-fixed length audios using a Fully Convolutional Neural Network
Fuente:
arXiv
Guardado en:
| Autores principales: | García-Ordás, María Teresa, Alaiz-Moretón, Héctor, Benítez-Andrades, José Alberto, García-Rodríguez, Isaías, García-Olalla, Oscar, Benavides, Carmen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multispecies bird sound recognition using a fully convolutional neural network
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
Determining the severity of Parkinson's disease in patients using a multi task neural network
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
Scaling up masked audio encoder learning for general audio classification
por: Dinkel, Heinrich, et al.
Publicado: (2024)
por: Dinkel, Heinrich, et al.
Publicado: (2024)
Towards audio language modeling -- an overview
por: Wu, Haibin, et al.
Publicado: (2024)
por: Wu, Haibin, et al.
Publicado: (2024)
Semi-intrusive audio evaluation: Casting non-intrusive assessment as a multi-modal text prediction task
por: Coldenhoff, Jozef, et al.
Publicado: (2024)
por: Coldenhoff, Jozef, et al.
Publicado: (2024)
Are audio DeepFake detection models polyglots?
por: Marek, Bartłomiej, et al.
Publicado: (2024)
por: Marek, Bartłomiej, et al.
Publicado: (2024)
Tweaking autoregressive methods for inpainting of gaps in audio signals
por: Mokrý, Ondřej, et al.
Publicado: (2024)
por: Mokrý, Ondřej, et al.
Publicado: (2024)
MBCodec:Thorough disentangle for high-fidelity audio compression
por: Zhang, Ruonan, et al.
Publicado: (2025)
por: Zhang, Ruonan, et al.
Publicado: (2025)
Real-time implementation of vibrato transfer as an audio effect
por: Hyrkas, Jeremy
Publicado: (2025)
por: Hyrkas, Jeremy
Publicado: (2025)
Speaker anonymization using neural audio codec language models
por: Panariello, Michele, et al.
Publicado: (2023)
por: Panariello, Michele, et al.
Publicado: (2023)
FxSearcher: gradient-free text-driven audio transformation
por: Ki, Hojoon, et al.
Publicado: (2025)
por: Ki, Hojoon, et al.
Publicado: (2025)
Regularized autoregressive modeling and its application to audio signal reconstruction
por: Mokrý, Ondřej, et al.
Publicado: (2024)
por: Mokrý, Ondřej, et al.
Publicado: (2024)
EDTC: enhance depth of text comprehension in automated audio captioning
por: Tan, Liwen, et al.
Publicado: (2024)
por: Tan, Liwen, et al.
Publicado: (2024)
AxLSTMs: learning self-supervised audio representations with xLSTMs
por: Yadav, Sarthak, et al.
Publicado: (2024)
por: Yadav, Sarthak, et al.
Publicado: (2024)
STASE: A spatialized text-to-audio synthesis engine for music generation
por: Chi, Tutti, et al.
Publicado: (2025)
por: Chi, Tutti, et al.
Publicado: (2025)
Deep learning based spatial aliasing reduction in beamforming for audio capture
por: Guzik, Mateusz, et al.
Publicado: (2025)
por: Guzik, Mateusz, et al.
Publicado: (2025)
Enhancement by postfiltering for speech and audio coding in ad-hoc sensor networks
por: Das, Sneha, et al.
Publicado: (2020)
por: Das, Sneha, et al.
Publicado: (2020)
Human-CLAP: Human-perception-based contrastive language-audio pretraining
por: Takano, Taisei, et al.
Publicado: (2025)
por: Takano, Taisei, et al.
Publicado: (2025)
DashengTokenizer: One layer is enough for unified audio understanding and generation
por: Dinkel, Heinrich, et al.
Publicado: (2026)
por: Dinkel, Heinrich, et al.
Publicado: (2026)
RELATE: Subjective evaluation dataset for automatic evaluation of relevance between text and audio
por: Kanamori, Yusuke, et al.
Publicado: (2025)
por: Kanamori, Yusuke, et al.
Publicado: (2025)
ICGAN: An implicit conditioning method for interpretable feature control of neural audio synthesis
por: Liu, Yunyi, et al.
Publicado: (2024)
por: Liu, Yunyi, et al.
Publicado: (2024)
A robust audio deepfake detection system via multi-view feature
por: Yang, Yujie, et al.
Publicado: (2024)
por: Yang, Yujie, et al.
Publicado: (2024)
Reconstructing the Charlie Parker Omnibook using an audio-to-score automatic transcription pipeline
por: Riley, Xavier, et al.
Publicado: (2024)
por: Riley, Xavier, et al.
Publicado: (2024)
Exploring trends in audio mixes and masters: Insights from a dataset analysis
por: Mourgela, Angeliki, et al.
Publicado: (2024)
por: Mourgela, Angeliki, et al.
Publicado: (2024)
Modeling strategies for speech enhancement in the latent space of a neural audio codec
por: Kammoun, Sofiene, et al.
Publicado: (2025)
por: Kammoun, Sofiene, et al.
Publicado: (2025)
ACAVCaps: Enabling large-scale training for fine-grained and diverse audio understanding
por: Niu, Yadong, et al.
Publicado: (2026)
por: Niu, Yadong, et al.
Publicado: (2026)
A tunable binaural audio telepresence system capable of balancing immersive and enhanced modes
por: Hsu, Yicheng, et al.
Publicado: (2024)
por: Hsu, Yicheng, et al.
Publicado: (2024)
WavJEPA: Semantic learning unlocks robust audio foundation models for raw waveforms
por: Yuksel, Goksenin, et al.
Publicado: (2025)
por: Yuksel, Goksenin, et al.
Publicado: (2025)
Codec-SUPERB @ SLT 2024: A lightweight benchmark for neural audio codec models
por: Wu, Haibin, et al.
Publicado: (2024)
por: Wu, Haibin, et al.
Publicado: (2024)
Efficient learning-based sound propagation for virtual and real-world audio processing applications
por: Ratnarajah, Anton Jeran
Publicado: (2024)
por: Ratnarajah, Anton Jeran
Publicado: (2024)
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
audio2chart: End to End Audio Transcription into playable Guitar Hero charts
por: Tripodi, Riccardo
Publicado: (2025)
por: Tripodi, Riccardo
Publicado: (2025)
Phase Repair for Time-Domain Convolutional Neural Networks in Music Super-Resolution
por: Zhang, Yenan, et al.
Publicado: (2023)
por: Zhang, Yenan, et al.
Publicado: (2023)
An audio-quality-based multi-strategy approach for target speaker extraction in the MISP 2023 Challenge
por: Han, Runduo, et al.
Publicado: (2024)
por: Han, Runduo, et al.
Publicado: (2024)
Self-supervised learning method using multiple sampling strategies for general-purpose audio representation
por: Kuroyanagi, Ibuki, et al.
Publicado: (2025)
por: Kuroyanagi, Ibuki, et al.
Publicado: (2025)
Blind estimation of audio effects using an auto-encoder approach and differentiable digital signal processing
por: Peladeau, Côme, et al.
Publicado: (2023)
por: Peladeau, Côme, et al.
Publicado: (2023)
AudioRepInceptionNeXt: A lightweight single-stream architecture for efficient audio recognition
por: Lau, Kin Wai, et al.
Publicado: (2024)
por: Lau, Kin Wai, et al.
Publicado: (2024)
AEROMamba: An efficient architecture for audio super-resolution using generative adversarial networks and state space models
por: Abreu, Wallace, et al.
Publicado: (2024)
por: Abreu, Wallace, et al.
Publicado: (2024)
An automatic mixing speech enhancement system for multi-track audio
por: Liu, Xiaojing, et al.
Publicado: (2024)
por: Liu, Xiaojing, et al.
Publicado: (2024)
Hybrid-Sep: Language-queried audio source separation via pre-trained Model Fusion and Adversarial Diffusion Training
por: Feng, Jianyuan, et al.
Publicado: (2025)
por: Feng, Jianyuan, et al.
Publicado: (2025)
Ejemplares similares
-
Multispecies bird sound recognition using a fully convolutional neural network
por: García-Ordás, María Teresa, et al.
Publicado: (2024) -
Determining the severity of Parkinson's disease in patients using a multi task neural network
por: García-Ordás, María Teresa, et al.
Publicado: (2024) -
Scaling up masked audio encoder learning for general audio classification
por: Dinkel, Heinrich, et al.
Publicado: (2024) -
Towards audio language modeling -- an overview
por: Wu, Haibin, et al.
Publicado: (2024) -
Semi-intrusive audio evaluation: Casting non-intrusive assessment as a multi-modal text prediction task
por: Coldenhoff, Jozef, et al.
Publicado: (2024)