Echoes: A semantically-aligned music deepfake detection dataset
Fuente:
arXiv
Salvato in:
| Autori principali: | Pascu, Octavian, Oneata, Dan, Cucu, Horia, Muller, Nicolas M. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Easy, Interpretable, Effective: openSMILE for voice deepfake detection
di: Pascu, Octavian, et al.
Pubblicazione: (2024)
di: Pascu, Octavian, et al.
Pubblicazione: (2024)
Towards generalisable and calibrated synthetic speech detection with self-supervised representations
di: Pascu, Octavian, et al.
Pubblicazione: (2023)
di: Pascu, Octavian, et al.
Pubblicazione: (2023)
WavLM model ensemble for audio deepfake detection
di: Combei, David, et al.
Pubblicazione: (2024)
di: Combei, David, et al.
Pubblicazione: (2024)
Unmasking real-world audio deepfakes: A data-centric approach
di: Combei, David, et al.
Pubblicazione: (2025)
di: Combei, David, et al.
Pubblicazione: (2025)
Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning
di: Smeu, Stefan, et al.
Pubblicazione: (2024)
di: Smeu, Stefan, et al.
Pubblicazione: (2024)
Where are we in audio deepfake detection? A systematic analysis over generative and detection models
di: Li, Xiang, et al.
Pubblicazione: (2024)
di: Li, Xiang, et al.
Pubblicazione: (2024)
TADA: Training-free Attribution and Out-of-Domain Detection of Audio Deepfakes
di: Stan, Adriana, et al.
Pubblicazione: (2025)
di: Stan, Adriana, et al.
Pubblicazione: (2025)
Forensic deepfake audio detection using segmental speech features
di: Yang, Tianle, et al.
Pubblicazione: (2025)
di: Yang, Tianle, et al.
Pubblicazione: (2025)
Understanding the strengths and weaknesses of SSL models for audio deepfake model attribution
di: Pîrlogeanu, Gabriel, et al.
Pubblicazione: (2026)
di: Pîrlogeanu, Gabriel, et al.
Pubblicazione: (2026)
AS-70: A Mandarin stuttered speech dataset for automatic speech recognition and stuttering event detection
di: Gong, Rong, et al.
Pubblicazione: (2024)
di: Gong, Rong, et al.
Pubblicazione: (2024)
Open Source State-Of-the-Art Solution for Romanian Speech Recognition
di: Pirlogeanu, Gabriel, et al.
Pubblicazione: (2025)
di: Pirlogeanu, Gabriel, et al.
Pubblicazione: (2025)
A robust audio deepfake detection system via multi-view feature
di: Yang, Yujie, et al.
Pubblicazione: (2024)
di: Yang, Yujie, et al.
Pubblicazione: (2024)
Detecting music deepfakes is easy but actually hard
di: Afchar, Darius, et al.
Pubblicazione: (2024)
di: Afchar, Darius, et al.
Pubblicazione: (2024)
A correlation-permutation approach for speech-music encoders model merging
di: Ritter-Gutierrez, Fabian, et al.
Pubblicazione: (2025)
di: Ritter-Gutierrez, Fabian, et al.
Pubblicazione: (2025)
Translating speech with just images
di: Oneata, Dan, et al.
Pubblicazione: (2024)
di: Oneata, Dan, et al.
Pubblicazione: (2024)
UniTTS: An end-to-end TTS system without decoupling of acoustic and semantic information
di: Wang, Rui, et al.
Pubblicazione: (2025)
di: Wang, Rui, et al.
Pubblicazione: (2025)
Deep learning for music generation. Four approaches and their comparative evaluation
di: Paroiu, Razvan, et al.
Pubblicazione: (2025)
di: Paroiu, Razvan, et al.
Pubblicazione: (2025)
CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment
di: Liu, Hanwen, et al.
Pubblicazione: (2026)
di: Liu, Hanwen, et al.
Pubblicazione: (2026)
CTC-aligned Audio-Text Embedding for Streaming Open-vocabulary Keyword Spotting
di: Jin, Sichen, et al.
Pubblicazione: (2024)
di: Jin, Sichen, et al.
Pubblicazione: (2024)
Harder or Different? Understanding Generalization of Audio Deepfake Detection
di: Müller, Nicolas M., et al.
Pubblicazione: (2024)
di: Müller, Nicolas M., et al.
Pubblicazione: (2024)
A New Approach to Voice Authenticity
di: Müller, Nicolas M., et al.
Pubblicazione: (2024)
di: Müller, Nicolas M., et al.
Pubblicazione: (2024)
Introducing voice timbre attribute detection
di: He, Jinghao, et al.
Pubblicazione: (2025)
di: He, Jinghao, et al.
Pubblicazione: (2025)
ECOSoundSet: a finely annotated dataset for the automated acoustic identification of Orthoptera and Cicadidae in North, Central and temperate Western Europe
di: Funosas, David, et al.
Pubblicazione: (2025)
di: Funosas, David, et al.
Pubblicazione: (2025)
EZhouNet:A framework based on graph neural network and anchor interval for the respiratory sound event detection
di: Chu, Yun, et al.
Pubblicazione: (2025)
di: Chu, Yun, et al.
Pubblicazione: (2025)
Replay Attacks Against Audio Deepfake Detection
di: Müller, Nicolas, et al.
Pubblicazione: (2025)
di: Müller, Nicolas, et al.
Pubblicazione: (2025)
A Toolchain for Comprehensive Audio/Video Analysis Using Deep Learning Based Multimodal Approach (A use case of riot or violent context detection)
di: Pham, Lam, et al.
Pubblicazione: (2024)
di: Pham, Lam, et al.
Pubblicazione: (2024)
LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging
di: Singh, Shubhr, et al.
Pubblicazione: (2025)
di: Singh, Shubhr, et al.
Pubblicazione: (2025)
Echoes of Ideology: Toward an Audio Analysis Pipeline to Unveil Character Traits in Historical Nazi Propaganda Films
di: Ruth, Nicolas, et al.
Pubblicazione: (2026)
di: Ruth, Nicolas, et al.
Pubblicazione: (2026)
Fast-VGAN: Lightweight Voice Conversion with Explicit Control of F0 and Duration Parameters
di: Abrassart, Mathilde, et al.
Pubblicazione: (2025)
di: Abrassart, Mathilde, et al.
Pubblicazione: (2025)
Generalizable speech deepfake detection via meta-learned LoRA
di: Laakkonen, Janne, et al.
Pubblicazione: (2025)
di: Laakkonen, Janne, et al.
Pubblicazione: (2025)
Sustaining model performance for covid-19 detection from dynamic audio data: Development and evaluation of a comprehensive drift-adaptive framework
di: Ganitidis, Theofanis, et al.
Pubblicazione: (2024)
di: Ganitidis, Theofanis, et al.
Pubblicazione: (2024)
MidiTok Visualizer: a tool for visualization and analysis of tokenized MIDI symbolic music
di: Wiszenko, Michał, et al.
Pubblicazione: (2024)
di: Wiszenko, Michał, et al.
Pubblicazione: (2024)
Lina-Speech: Gated Linear Attention and Initial-State Tuning for Multi-Sample Prompting Text-To-Speech Synthesis
di: Lemerle, Théodor, et al.
Pubblicazione: (2024)
di: Lemerle, Théodor, et al.
Pubblicazione: (2024)
Disambiguation of Chinese Polyphones in an End-to-End Framework with Semantic Features Extracted by Pre-trained BERT
di: Dai, Dongyang, et al.
Pubblicazione: (2025)
di: Dai, Dongyang, et al.
Pubblicazione: (2025)
CORD: Bridging the Audio-Text Reasoning Gap via Weighted On-policy Cross-modal Distillation
di: Hu, Jing, et al.
Pubblicazione: (2026)
di: Hu, Jing, et al.
Pubblicazione: (2026)
MoE Adapter for Large Audio Language Models: Sparsity, Disentanglement, and Gradient-Conflict-Free
di: Lei, Yishu, et al.
Pubblicazione: (2026)
di: Lei, Yishu, et al.
Pubblicazione: (2026)
Linear RNNs for autoregressive generation of long music samples
di: Szewczyk, Konrad, et al.
Pubblicazione: (2025)
di: Szewczyk, Konrad, et al.
Pubblicazione: (2025)
In-depth analysis of music structure as a text network
di: Tsai, Ping-Rui, et al.
Pubblicazione: (2023)
di: Tsai, Ping-Rui, et al.
Pubblicazione: (2023)
Symbotunes: unified hub for symbolic music generative models
di: Skierś, Paweł, et al.
Pubblicazione: (2024)
di: Skierś, Paweł, et al.
Pubblicazione: (2024)
Patient-Level Multimodal Question Answering from Multi-Site Auscultation Recordings
di: Wu, Fan, et al.
Pubblicazione: (2026)
di: Wu, Fan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Easy, Interpretable, Effective: openSMILE for voice deepfake detection
di: Pascu, Octavian, et al.
Pubblicazione: (2024) -
Towards generalisable and calibrated synthetic speech detection with self-supervised representations
di: Pascu, Octavian, et al.
Pubblicazione: (2023) -
WavLM model ensemble for audio deepfake detection
di: Combei, David, et al.
Pubblicazione: (2024) -
Unmasking real-world audio deepfakes: A data-centric approach
di: Combei, David, et al.
Pubblicazione: (2025) -
Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning
di: Smeu, Stefan, et al.
Pubblicazione: (2024)