Human Perception of Audio Deepfakes
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Müller, Nicolas M., Pizzi, Karla, Williams, Jennifer |
|---|---|
| Format: | Preprint |
| Publié: |
2021
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Replay Attacks Against Audio Deepfake Detection
par: Müller, Nicolas, et autres
Publié: (2025)
par: Müller, Nicolas, et autres
Publié: (2025)
LSTM-CNN Network for Audio Signature Analysis in Noisy Environments
par: Damacharla, Praveen, et autres
Publié: (2023)
par: Damacharla, Praveen, et autres
Publié: (2023)
Between the AI and Me: Analysing Listeners' Perspectives on AI- and Human-Composed Progressive Metal Music
par: Sarmento, Pedro, et autres
Publié: (2024)
par: Sarmento, Pedro, et autres
Publié: (2024)
Learning Relationships Between Separate Audio Tracks for Creative Applications
par: Bujard, Balthazar, et autres
Publié: (2025)
par: Bujard, Balthazar, et autres
Publié: (2025)
Step-Audio-EditX Technical Report
par: Yan, Chao, et autres
Publié: (2025)
par: Yan, Chao, et autres
Publié: (2025)
Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings
par: Li, Yinan, et autres
Publié: (2026)
par: Li, Yinan, et autres
Publié: (2026)
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
par: Huang, Ailin, et autres
Publié: (2025)
par: Huang, Ailin, et autres
Publié: (2025)
How Private is Low-Frequency Speech Audio in the Wild? An Analysis of Verbal Intelligibility by Humans and Machines
par: Liu, Ailin, et autres
Publié: (2024)
par: Liu, Ailin, et autres
Publié: (2024)
DeformTune: A Deformable XAI Music Prototype for Non-Musicians
par: Xu, Ziqing, et autres
Publié: (2025)
par: Xu, Ziqing, et autres
Publié: (2025)
CabinSep: IR-Augmented Mask-Based MVDR for Real-Time In-Car Speech Separation with Distributed Heterogeneous Arrays
par: Han, Runduo, et autres
Publié: (2025)
par: Han, Runduo, et autres
Publié: (2025)
SynthScribe: Deep Multimodal Tools for Synthesizer Sound Retrieval and Exploration
par: Brade, Stephen, et autres
Publié: (2023)
par: Brade, Stephen, et autres
Publié: (2023)
Exploring Situated Stabilities of a Rhythm Generation System through Variational Cross-Examination
par: Kotowski, Błażej, et autres
Publié: (2025)
par: Kotowski, Błażej, et autres
Publié: (2025)
Revisiting Your Memory: Reconstruction of Affect-Contextualized Memory via EEG-guided Audiovisual Generation
par: Kwon, Joonwoo, et autres
Publié: (2024)
par: Kwon, Joonwoo, et autres
Publié: (2024)
VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching
par: Guo, Yiwei, et autres
Publié: (2023)
par: Guo, Yiwei, et autres
Publié: (2023)
EvolveCaptions: Empowering DHH Users Through Real-Time Collaborative Captioning
par: Wu, Liang-Yuan, et autres
Publié: (2025)
par: Wu, Liang-Yuan, et autres
Publié: (2025)
Reimagining Dance: Real-time Music Co-creation between Dancers and AI
par: Vechtomova, Olga, et autres
Publié: (2025)
par: Vechtomova, Olga, et autres
Publié: (2025)
A Cross-Modal Approach to Silent Speech with LLM-Enhanced Recognition
par: Benster, Tyler, et autres
Publié: (2024)
par: Benster, Tyler, et autres
Publié: (2024)
MCP2OSC: Parametric Control by Natural Language
par: Fan, Yuan-Yi
Publié: (2025)
par: Fan, Yuan-Yi
Publié: (2025)
STAA-Net: A Sparse and Transferable Adversarial Attack for Speech Emotion Recognition
par: Chang, Yi, et autres
Publié: (2024)
par: Chang, Yi, et autres
Publié: (2024)
GMM-ResNext: Combining Generative and Discriminative Models for Speaker Verification
par: Yan, Hui, et autres
Publié: (2024)
par: Yan, Hui, et autres
Publié: (2024)
A Theory-Based Explainable Deep Learning Architecture for Music Emotion
par: Fong, Hortense, et autres
Publié: (2024)
par: Fong, Hortense, et autres
Publié: (2024)
Interactive Melody Generation System for Enhancing the Creativity of Musicians
par: Hirawata, So, et autres
Publié: (2024)
par: Hirawata, So, et autres
Publié: (2024)
Tidal MerzA: Combining affective modelling and autonomous code generation through Reinforcement Learning
par: Wilson, Elizabeth, et autres
Publié: (2024)
par: Wilson, Elizabeth, et autres
Publié: (2024)
Tipping Points, Pulse Elasticity and Tonal Tension: An Empirical Study on What Generates Tipping Points
par: Naik, Canishk, et autres
Publié: (2024)
par: Naik, Canishk, et autres
Publié: (2024)
Seeing Beyond Sound: Visualization and Abstraction in Audio Data Representation
par: Blum'e, Ashlae
Publié: (2025)
par: Blum'e, Ashlae
Publié: (2025)
SACM: SEEG-Audio Contrastive Matching for Chinese Speech Decoding
par: Wang, Hongbin, et autres
Publié: (2025)
par: Wang, Hongbin, et autres
Publié: (2025)
Cross-Lingual Speech Emotion Recognition: Humans vs. Self-Supervised Models
par: Han, Zhichen, et autres
Publié: (2024)
par: Han, Zhichen, et autres
Publié: (2024)
More-than-Human Storytelling: Designing Longitudinal Narrative Engagements with Generative AI
par: Fabre, Émilie, et autres
Publié: (2025)
par: Fabre, Émilie, et autres
Publié: (2025)
FeatureSense: Protecting Speaker Attributes in Always-On Audio Sensing System
par: Chhaglani, Bhawana, et autres
Publié: (2025)
par: Chhaglani, Bhawana, et autres
Publié: (2025)
A Mapping Strategy for Interacting with Latent Audio Synthesis Using Artistic Materials
par: Zheng, Shuoyang, et autres
Publié: (2024)
par: Zheng, Shuoyang, et autres
Publié: (2024)
Sound Judgment: Properties of Consequential Sounds Affecting Human-Perception of Robots
par: Allen, Aimee, et autres
Publié: (2025)
par: Allen, Aimee, et autres
Publié: (2025)
DESAMO: A Device for Elder-Friendly Smart Homes Powered by Embedded LLM with Audio Modality
par: Choi, Youngwon, et autres
Publié: (2025)
par: Choi, Youngwon, et autres
Publié: (2025)
Harder or Different? Understanding Generalization of Audio Deepfake Detection
par: Müller, Nicolas M., et autres
Publié: (2024)
par: Müller, Nicolas M., et autres
Publié: (2024)
Music Generation using Human-In-The-Loop Reinforcement Learning
par: Justus, Aju Ani
Publié: (2025)
par: Justus, Aju Ani
Publié: (2025)
Composers' Evaluations of an AI Music Tool: Insights for Human-Centred Design
par: Row, Eleanor, et autres
Publié: (2024)
par: Row, Eleanor, et autres
Publié: (2024)
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese
par: Wang, Xihuai, et autres
Publié: (2025)
par: Wang, Xihuai, et autres
Publié: (2025)
MetaBGM: Dynamic Soundtrack Transformation For Continuous Multi-Scene Experiences With Ambient Awareness And Personalization
par: Liu, Haoxuan, et autres
Publié: (2024)
par: Liu, Haoxuan, et autres
Publié: (2024)
Proceedings of The second international workshop on eXplainable AI for the Arts (XAIxArts)
par: Bryan-Kinns, Nick, et autres
Publié: (2024)
par: Bryan-Kinns, Nick, et autres
Publié: (2024)
SounDiT: Geo-Contextual Soundscape-to-Landscape Generation
par: Wang, Junbo, et autres
Publié: (2025)
par: Wang, Junbo, et autres
Publié: (2025)
Freetalker: Controllable Speech and Text-Driven Gesture Generation Based on Diffusion Models for Enhanced Speaker Naturalness
par: Yang, Sicheng, et autres
Publié: (2024)
par: Yang, Sicheng, et autres
Publié: (2024)
Documents similaires
-
Replay Attacks Against Audio Deepfake Detection
par: Müller, Nicolas, et autres
Publié: (2025) -
LSTM-CNN Network for Audio Signature Analysis in Noisy Environments
par: Damacharla, Praveen, et autres
Publié: (2023) -
Between the AI and Me: Analysing Listeners' Perspectives on AI- and Human-Composed Progressive Metal Music
par: Sarmento, Pedro, et autres
Publié: (2024) -
Learning Relationships Between Separate Audio Tracks for Creative Applications
par: Bujard, Balthazar, et autres
Publié: (2025) -
Step-Audio-EditX Technical Report
par: Yan, Chao, et autres
Publié: (2025)