audio dataset
Fuente:
Zenodo
Gespeichert in:
| 1. Verfasser: | handesha |
|---|---|
| Format: | Recurso digital |
| Veröffentlicht: |
Zenodo
2025
|
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Audio Dialogues: Dialogues dataset for audio and music understanding
von: Goel, Arushi, et al.
Veröffentlicht: (2024)
von: Goel, Arushi, et al.
Veröffentlicht: (2024)
Unsupervised outlier detection to improve bird audio dataset labels
von: Collins, Bruce
Veröffentlicht: (2025)
von: Collins, Bruce
Veröffentlicht: (2025)
An RFP dataset for Real, Fake, and Partially fake audio detection
von: AlAli, Abdulazeez, et al.
Veröffentlicht: (2024)
von: AlAli, Abdulazeez, et al.
Veröffentlicht: (2024)
Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning
von: Smeu, Stefan, et al.
Veröffentlicht: (2024)
von: Smeu, Stefan, et al.
Veröffentlicht: (2024)
RELATE: Subjective evaluation dataset for automatic evaluation of relevance between text and audio
von: Kanamori, Yusuke, et al.
Veröffentlicht: (2025)
von: Kanamori, Yusuke, et al.
Veröffentlicht: (2025)
A 1000-hour EEG-EMG-audio dataset of Japanese speech production
von: Sato, Motoshige, et al.
Veröffentlicht: (2026)
von: Sato, Motoshige, et al.
Veröffentlicht: (2026)
Exploring trends in audio mixes and masters: Insights from a dataset analysis
von: Mourgela, Angeliki, et al.
Veröffentlicht: (2024)
von: Mourgela, Angeliki, et al.
Veröffentlicht: (2024)
WikiMuTe: A web-sourced dataset of semantic descriptions for music audio
von: Weck, Benno, et al.
Veröffentlicht: (2023)
von: Weck, Benno, et al.
Veröffentlicht: (2023)
Chinese-LiPS: A Chinese audio-visual speech recognition dataset with Lip-reading and Presentation Slides
von: Zhao, Jinghua, et al.
Veröffentlicht: (2025)
von: Zhao, Jinghua, et al.
Veröffentlicht: (2025)
Scaling up masked audio encoder learning for general audio classification
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2024)
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2024)
Online incremental learning for audio classification using a pretrained audio model
von: Mulimani, Manjunath, et al.
Veröffentlicht: (2025)
von: Mulimani, Manjunath, et al.
Veröffentlicht: (2025)
Session six audio material
von: Abhari, Ahmad
Veröffentlicht: (2026)
von: Abhari, Ahmad
Veröffentlicht: (2026)
Introduzione all’audio forense
von: Cenceschi, Sonia, et al.
Veröffentlicht: (2026)
von: Cenceschi, Sonia, et al.
Veröffentlicht: (2026)
Spectrogram features for audio and speech analysis
von: McLoughlin, Ian, et al.
Veröffentlicht: (2026)
von: McLoughlin, Ian, et al.
Veröffentlicht: (2026)
Towards audio language modeling -- an overview
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
Stage-adaptive audio diffusion modeling
von: Zhang, Xuanhao, et al.
Veröffentlicht: (2026)
von: Zhang, Xuanhao, et al.
Veröffentlicht: (2026)
A SOUND APPROACH: Using Large Language Models to generate audio descriptions for egocentric text-audio retrieval
von: Oncescu, Andreea-Maria, et al.
Veröffentlicht: (2024)
von: Oncescu, Andreea-Maria, et al.
Veröffentlicht: (2024)
Interfacing with history: Curating with audio augmented objects
von: Cliffe, Laurence
Veröffentlicht: (2024)
von: Cliffe, Laurence
Veröffentlicht: (2024)
Joint sentiment analysis of lyrics and audio in music
von: Schaab, Lea, et al.
Veröffentlicht: (2024)
von: Schaab, Lea, et al.
Veröffentlicht: (2024)
Character-aware audio-visual subtitling in context
von: Huh, Jaesung, et al.
Veröffentlicht: (2024)
von: Huh, Jaesung, et al.
Veröffentlicht: (2024)
Are audio DeepFake detection models polyglots?
von: Marek, Bartłomiej, et al.
Veröffentlicht: (2024)
von: Marek, Bartłomiej, et al.
Veröffentlicht: (2024)
Versatile audio-visual learning for emotion recognition
von: Goncalves, Lucas, et al.
Veröffentlicht: (2023)
von: Goncalves, Lucas, et al.
Veröffentlicht: (2023)
Blind speaker identification for audio forensic purposes
von: Dora María Ballesteros Larrota
Veröffentlicht: (2017)
von: Dora María Ballesteros Larrota
Veröffentlicht: (2017)
animal2vec and MeerKAT: A self‐supervised transformer for rare‐event raw audio input and a large‐scale reference dataset for bioacoustics
von: Julian C. Schäfer‐Zimmermann, et al.
Veröffentlicht: (2025)
von: Julian C. Schäfer‐Zimmermann, et al.
Veröffentlicht: (2025)
animal2vec and MeerKAT: A self-supervised transformer for rare-event raw audio input and a large-scale reference dataset for bioacoustics
von: Schäfer-Zimmermann, Julian C., et al.
Veröffentlicht: (2024)
von: Schäfer-Zimmermann, Julian C., et al.
Veröffentlicht: (2024)
Multimodal recognition with deep learning: audio, image, and text
von: Gummula, Ravi, et al.
Veröffentlicht: (2025)
von: Gummula, Ravi, et al.
Veröffentlicht: (2025)
ADIFF: Explaining audio difference using natural language
von: Deshmukh, Soham, et al.
Veröffentlicht: (2025)
von: Deshmukh, Soham, et al.
Veröffentlicht: (2025)
ADIFF: Explaining audio difference using natural language
von: Deshmukh, Soham, et al.
Veröffentlicht: (2025)
von: Deshmukh, Soham, et al.
Veröffentlicht: (2025)
MBCodec:Thorough disentangle for high-fidelity audio compression
von: Zhang, Ruonan, et al.
Veröffentlicht: (2025)
von: Zhang, Ruonan, et al.
Veröffentlicht: (2025)
Transformation of audio embeddings into interpretable, concept-based representations
von: Zhang, Alice, et al.
Veröffentlicht: (2025)
von: Zhang, Alice, et al.
Veröffentlicht: (2025)
Cryfish: On deep audio analysis with Large Language Models
von: Mitrofanov, Anton, et al.
Veröffentlicht: (2025)
von: Mitrofanov, Anton, et al.
Veröffentlicht: (2025)
Recomposer: Event-roll-guided generative audio editing
von: Ellis, Daniel P. W., et al.
Veröffentlicht: (2025)
von: Ellis, Daniel P. W., et al.
Veröffentlicht: (2025)
Training chord recognition models on artificially generated audio
von: Majchrzak, Martyna, et al.
Veröffentlicht: (2025)
von: Majchrzak, Martyna, et al.
Veröffentlicht: (2025)
Mellow: a small audio language model for reasoning
von: Deshmukh, Soham, et al.
Veröffentlicht: (2025)
von: Deshmukh, Soham, et al.
Veröffentlicht: (2025)
Mixer Metaphors: audio interfaces for non-musical applications
von: McNamara, Tace, et al.
Veröffentlicht: (2025)
von: McNamara, Tace, et al.
Veröffentlicht: (2025)
Real-time implementation of vibrato transfer as an audio effect
von: Hyrkas, Jeremy
Veröffentlicht: (2025)
von: Hyrkas, Jeremy
Veröffentlicht: (2025)
Bird detection in audio: a survey and a challenge
von: Stowell, Dan, et al.
Veröffentlicht: (2016)
von: Stowell, Dan, et al.
Veröffentlicht: (2016)
WavLM model ensemble for audio deepfake detection
von: Combei, David, et al.
Veröffentlicht: (2024)
von: Combei, David, et al.
Veröffentlicht: (2024)
Tweaking autoregressive methods for inpainting of gaps in audio signals
von: Mokrý, Ondřej, et al.
Veröffentlicht: (2024)
von: Mokrý, Ondřej, et al.
Veröffentlicht: (2024)
Compositional nonlinear audio signal processing with Volterra series
von: Araujo-Simon, Jake
Veröffentlicht: (2023)
von: Araujo-Simon, Jake
Veröffentlicht: (2023)
Ähnliche Einträge
-
Audio Dialogues: Dialogues dataset for audio and music understanding
von: Goel, Arushi, et al.
Veröffentlicht: (2024) -
Unsupervised outlier detection to improve bird audio dataset labels
von: Collins, Bruce
Veröffentlicht: (2025) -
An RFP dataset for Real, Fake, and Partially fake audio detection
von: AlAli, Abdulazeez, et al.
Veröffentlicht: (2024) -
Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning
von: Smeu, Stefan, et al.
Veröffentlicht: (2024) -
RELATE: Subjective evaluation dataset for automatic evaluation of relevance between text and audio
von: Kanamori, Yusuke, et al.
Veröffentlicht: (2025)