Faked Speech Detection with Zero Prior Knowledge
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ajmi, Sahar Al, Hayat, Khizar, Obaidi, Alaa M. Al, Kumar, Naresh, Najmuldeen, Munaf, Magnier, Baptiste |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Identifying Chemicals Through Dimensionality Reduction
von: Anand, Emile, et al.
Veröffentlicht: (2022)
von: Anand, Emile, et al.
Veröffentlicht: (2022)
Parameter optimization comparison in QAOA using Stochastic Hill Climbing with Random Re-starts and Local Search with entangled and non-entangled mixing operators
von: Sarmina, Brian García, et al.
Veröffentlicht: (2024)
von: Sarmina, Brian García, et al.
Veröffentlicht: (2024)
Coarse-to-Fine Proposal Refinement Framework for Audio Temporal Forgery Detection and Localization
von: Wu, Junyan, et al.
Veröffentlicht: (2024)
von: Wu, Junyan, et al.
Veröffentlicht: (2024)
EvoPruneDeepTL: An Evolutionary Pruning Model for Transfer Learning based Deep Neural Networks
von: Poyatos, Javier, et al.
Veröffentlicht: (2022)
von: Poyatos, Javier, et al.
Veröffentlicht: (2022)
Rethinking Masking Strategies for Masked Prediction-based Audio Self-supervised Learning
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2026)
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2026)
Alternate Loss Functions for Classification and Robust Regression Can Improve the Accuracy of Artificial Neural Networks
von: Noel, Mathew Mithra, et al.
Veröffentlicht: (2023)
von: Noel, Mathew Mithra, et al.
Veröffentlicht: (2023)
RFOX (Rotated-Field Oscillatory eXchange) quantum algorithm: Towards Parameter-Free Quantum Optimizers
von: Sarmina, Brian García, et al.
Veröffentlicht: (2026)
von: Sarmina, Brian García, et al.
Veröffentlicht: (2026)
Exploring Entanglement and Parameter Sensitivity in QAOA through Quantum Fisher Information
von: Sarmina, Brian García, et al.
Veröffentlicht: (2025)
von: Sarmina, Brian García, et al.
Veröffentlicht: (2025)
A Unified Model For Voice and Accent Conversion In Speech and Singing using Self-Supervised Learning and Feature Extraction
von: Cheripally, Sowmya
Veröffentlicht: (2024)
von: Cheripally, Sowmya
Veröffentlicht: (2024)
Racism in the Machine: Visualization Ethics in Digital Humanities Projects
von: Hepworth, K. J., et al.
Veröffentlicht: (2024)
von: Hepworth, K. J., et al.
Veröffentlicht: (2024)
FakeSound: Deepfake General Audio Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
Prevailing Research Areas for Music AI in the Era of Foundation Models
von: Wei, Megan, et al.
Veröffentlicht: (2024)
von: Wei, Megan, et al.
Veröffentlicht: (2024)
FakeSound2: A Benchmark for Explainable and Generalizable Deepfake Sound Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2025)
von: Xie, Zeyu, et al.
Veröffentlicht: (2025)
Implementing Online Reinforcement Learning with Clustering Neural Networks
von: Smith, James E.
Veröffentlicht: (2024)
von: Smith, James E.
Veröffentlicht: (2024)
Text2midi-InferAlign: Improving Symbolic Music Generation with Inference-Time Alignment
von: Roy, Abhinaba, et al.
Veröffentlicht: (2025)
von: Roy, Abhinaba, et al.
Veröffentlicht: (2025)
STAR: Speech-to-Audio Generation via Representation Learning
von: Xie, Zeyu, et al.
Veröffentlicht: (2025)
von: Xie, Zeyu, et al.
Veröffentlicht: (2025)
FabuLight-ASD: Unveiling Speech Activity via Body Language
von: Carneiro, Hugo, et al.
Veröffentlicht: (2024)
von: Carneiro, Hugo, et al.
Veröffentlicht: (2024)
AiGAS-dEVL-RC: An Adaptive Growing Neural Gas Model for Recurrently Drifting Unsupervised Data Streams
von: Arostegi, Maria, et al.
Veröffentlicht: (2025)
von: Arostegi, Maria, et al.
Veröffentlicht: (2025)
Phase-Coded Memory and Morphological Resonance: A Next-Generation Retrieval-Augmented Generator Architecture
von: Saklakov, Denis V.
Veröffentlicht: (2025)
von: Saklakov, Denis V.
Veröffentlicht: (2025)
Evo* 2025 -- Late-Breaking Abstracts Volume
von: Mora, A. M., et al.
Veröffentlicht: (2025)
von: Mora, A. M., et al.
Veröffentlicht: (2025)
Representation Loss Minimization with Randomized Selection Strategy for Efficient Environmental Fake Audio Detection
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
Modeling Local Search Metaheuristics Using Markov Decision Processes
von: Ruiz-Torrubiano, Rubén
Veröffentlicht: (2024)
von: Ruiz-Torrubiano, Rubén
Veröffentlicht: (2024)
Physics-Driven AI Correction in Laser Absorption Sensing Quantification
von: Kang, Ruiyuan, et al.
Veröffentlicht: (2024)
von: Kang, Ruiyuan, et al.
Veröffentlicht: (2024)
Reservoir Computing with Evolved Critical Neural Cellular Automata
von: Pontes-Filho, Sidney, et al.
Veröffentlicht: (2025)
von: Pontes-Filho, Sidney, et al.
Veröffentlicht: (2025)
Loss shaping enhances exact gradient learning with Eventprop in spiking neural networks
von: Nowotny, Thomas, et al.
Veröffentlicht: (2022)
von: Nowotny, Thomas, et al.
Veröffentlicht: (2022)
ChordSync: Conformer-Based Alignment of Chord Annotations to Music Audio
von: Poltronieri, Andrea, et al.
Veröffentlicht: (2024)
von: Poltronieri, Andrea, et al.
Veröffentlicht: (2024)
Multi-View Multi-Task Modeling with Speech Foundation Models for Speech Forensic Tasks
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
Investigating Prosodic Signatures via Speech Pre-Trained Models for Audio Deepfake Source Attribution
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
SeQuiFi: Mitigating Catastrophic Forgetting in Speech Emotion Recognition with Sequential Class-Finetuning
von: Jain, Sarthak, et al.
Veröffentlicht: (2024)
von: Jain, Sarthak, et al.
Veröffentlicht: (2024)
M2D-CLAP: Exploring General-purpose Audio-Language Representations Beyond CLAP
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2025)
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2025)
PicoAudio: Enabling Precise Timestamp and Frequency Controllability of Audio Events in Text-to-audio Generation
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
AudioTime: A Temporally-aligned Audio-text Benchmark Dataset
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
CAST-TTS: A Simple Cross-Attention Framework for Unified Timbre Control in TTS
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
PicoAudio2: Temporal Controllable Text-to-Audio Generation with Natural Language Description
von: Zheng, Zihao, et al.
Veröffentlicht: (2025)
von: Zheng, Zihao, et al.
Veröffentlicht: (2025)
PolyGlotFake: A Novel Multilingual and Multimodal DeepFake Dataset
von: Hou, Yang, et al.
Veröffentlicht: (2024)
von: Hou, Yang, et al.
Veröffentlicht: (2024)
Water-Based Metaheuristics: How Water Dynamics Can Help Us to Solve NP-Hard Problems
von: Rubio, Fernando, et al.
Veröffentlicht: (2024)
von: Rubio, Fernando, et al.
Veröffentlicht: (2024)
AI-based Drone Assisted Human Rescue in Disaster Environments: Challenges and Opportunities
von: Papyan, Narek, et al.
Veröffentlicht: (2024)
von: Papyan, Narek, et al.
Veröffentlicht: (2024)
Evo* 2023 -- Late-Breaking Abstracts Volume
von: Mora, A. M., et al.
Veröffentlicht: (2024)
von: Mora, A. M., et al.
Veröffentlicht: (2024)
Composition of Relational Features with an Application to Explaining Black-Box Predictors
von: Srinivasan, Ashwin, et al.
Veröffentlicht: (2022)
von: Srinivasan, Ashwin, et al.
Veröffentlicht: (2022)
Beyond Speech and More: Investigating the Emergent Ability of Speech Foundation Models for Classifying Physiological Time-Series Signals
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Identifying Chemicals Through Dimensionality Reduction
von: Anand, Emile, et al.
Veröffentlicht: (2022) -
Parameter optimization comparison in QAOA using Stochastic Hill Climbing with Random Re-starts and Local Search with entangled and non-entangled mixing operators
von: Sarmina, Brian García, et al.
Veröffentlicht: (2024) -
Coarse-to-Fine Proposal Refinement Framework for Audio Temporal Forgery Detection and Localization
von: Wu, Junyan, et al.
Veröffentlicht: (2024) -
EvoPruneDeepTL: An Evolutionary Pruning Model for Transfer Learning based Deep Neural Networks
von: Poyatos, Javier, et al.
Veröffentlicht: (2022) -
Rethinking Masking Strategies for Masked Prediction-based Audio Self-supervised Learning
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2026)