I Hear, Therefore I Trust: A Socio-Technical Investigation of Humans as Synthetic Speech Detectors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Erscoi, Lelia, Kinnunen, Tomi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Cross-Modal Approach to Silent Speech with LLM-Enhanced Recognition
von: Benster, Tyler, et al.
Veröffentlicht: (2024)
von: Benster, Tyler, et al.
Veröffentlicht: (2024)
STAA-Net: A Sparse and Transferable Adversarial Attack for Speech Emotion Recognition
von: Chang, Yi, et al.
Veröffentlicht: (2024)
von: Chang, Yi, et al.
Veröffentlicht: (2024)
Beamforming-LLM: What, Where and When Did I Miss?
von: Choudhari, Vishal
Veröffentlicht: (2025)
von: Choudhari, Vishal
Veröffentlicht: (2025)
VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching
von: Guo, Yiwei, et al.
Veröffentlicht: (2023)
von: Guo, Yiwei, et al.
Veröffentlicht: (2023)
Human Perception of Audio Deepfakes
von: Müller, Nicolas M., et al.
Veröffentlicht: (2021)
von: Müller, Nicolas M., et al.
Veröffentlicht: (2021)
Cross-Lingual Speech Emotion Recognition: Humans vs. Self-Supervised Models
von: Han, Zhichen, et al.
Veröffentlicht: (2024)
von: Han, Zhichen, et al.
Veröffentlicht: (2024)
CabinSep: IR-Augmented Mask-Based MVDR for Real-Time In-Car Speech Separation with Distributed Heterogeneous Arrays
von: Han, Runduo, et al.
Veröffentlicht: (2025)
von: Han, Runduo, et al.
Veröffentlicht: (2025)
Between the AI and Me: Analysing Listeners' Perspectives on AI- and Human-Composed Progressive Metal Music
von: Sarmento, Pedro, et al.
Veröffentlicht: (2024)
von: Sarmento, Pedro, et al.
Veröffentlicht: (2024)
Step-Audio-EditX Technical Report
von: Yan, Chao, et al.
Veröffentlicht: (2025)
von: Yan, Chao, et al.
Veröffentlicht: (2025)
InsightPulse: An IoT-based System for User Experience Interview Analysis
von: Lyu, Dian, et al.
Veröffentlicht: (2024)
von: Lyu, Dian, et al.
Veröffentlicht: (2024)
Toward a Realistic Encoding Model of Auditory Affective Understanding in the Brain
von: Pan, Guandong, et al.
Veröffentlicht: (2025)
von: Pan, Guandong, et al.
Veröffentlicht: (2025)
PersonaCite: VoC-Grounded Interviewable Agentic Synthetic AI Personas for Verifiable User and Design Research
von: Truss, Mario
Veröffentlicht: (2026)
von: Truss, Mario
Veröffentlicht: (2026)
Zero-Shot KWS for Children's Speech using Layer-Wise Features from SSL Models
von: Kutum, Subham, et al.
Veröffentlicht: (2025)
von: Kutum, Subham, et al.
Veröffentlicht: (2025)
Freetalker: Controllable Speech and Text-Driven Gesture Generation Based on Diffusion Models for Enhanced Speaker Naturalness
von: Yang, Sicheng, et al.
Veröffentlicht: (2024)
von: Yang, Sicheng, et al.
Veröffentlicht: (2024)
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
von: Huang, Ailin, et al.
Veröffentlicht: (2025)
von: Huang, Ailin, et al.
Veröffentlicht: (2025)
DeformTune: A Deformable XAI Music Prototype for Non-Musicians
von: Xu, Ziqing, et al.
Veröffentlicht: (2025)
von: Xu, Ziqing, et al.
Veröffentlicht: (2025)
A Theory-Based Explainable Deep Learning Architecture for Music Emotion
von: Fong, Hortense, et al.
Veröffentlicht: (2024)
von: Fong, Hortense, et al.
Veröffentlicht: (2024)
Tidal MerzA: Combining affective modelling and autonomous code generation through Reinforcement Learning
von: Wilson, Elizabeth, et al.
Veröffentlicht: (2024)
von: Wilson, Elizabeth, et al.
Veröffentlicht: (2024)
A Near-Real-Time Processing Ego Speech Filtering Pipeline Designed for Speech Interruption During Human-Robot Interaction
von: Li, Yue, et al.
Veröffentlicht: (2024)
von: Li, Yue, et al.
Veröffentlicht: (2024)
Timbre-Aware LLM-based Direct Speech-to-Speech Translation Extendable to Multiple Language Pairs
von: Arya, Lalaram, et al.
Veröffentlicht: (2026)
von: Arya, Lalaram, et al.
Veröffentlicht: (2026)
Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models
von: Dietrich, Juergen
Veröffentlicht: (2026)
von: Dietrich, Juergen
Veröffentlicht: (2026)
Meta-Learning Approaches for Improving Detection of Unseen Speech Deepfakes
von: Kukanov, Ivan, et al.
Veröffentlicht: (2024)
von: Kukanov, Ivan, et al.
Veröffentlicht: (2024)
SynthScribe: Deep Multimodal Tools for Synthesizer Sound Retrieval and Exploration
von: Brade, Stephen, et al.
Veröffentlicht: (2023)
von: Brade, Stephen, et al.
Veröffentlicht: (2023)
Exploring Situated Stabilities of a Rhythm Generation System through Variational Cross-Examination
von: Kotowski, Błażej, et al.
Veröffentlicht: (2025)
von: Kotowski, Błażej, et al.
Veröffentlicht: (2025)
Revisiting Your Memory: Reconstruction of Affect-Contextualized Memory via EEG-guided Audiovisual Generation
von: Kwon, Joonwoo, et al.
Veröffentlicht: (2024)
von: Kwon, Joonwoo, et al.
Veröffentlicht: (2024)
EvolveCaptions: Empowering DHH Users Through Real-Time Collaborative Captioning
von: Wu, Liang-Yuan, et al.
Veröffentlicht: (2025)
von: Wu, Liang-Yuan, et al.
Veröffentlicht: (2025)
Reimagining Dance: Real-time Music Co-creation between Dancers and AI
von: Vechtomova, Olga, et al.
Veröffentlicht: (2025)
von: Vechtomova, Olga, et al.
Veröffentlicht: (2025)
LSTM-CNN Network for Audio Signature Analysis in Noisy Environments
von: Damacharla, Praveen, et al.
Veröffentlicht: (2023)
von: Damacharla, Praveen, et al.
Veröffentlicht: (2023)
MCP2OSC: Parametric Control by Natural Language
von: Fan, Yuan-Yi
Veröffentlicht: (2025)
von: Fan, Yuan-Yi
Veröffentlicht: (2025)
GMM-ResNext: Combining Generative and Discriminative Models for Speaker Verification
von: Yan, Hui, et al.
Veröffentlicht: (2024)
von: Yan, Hui, et al.
Veröffentlicht: (2024)
Interactive Melody Generation System for Enhancing the Creativity of Musicians
von: Hirawata, So, et al.
Veröffentlicht: (2024)
von: Hirawata, So, et al.
Veröffentlicht: (2024)
Tipping Points, Pulse Elasticity and Tonal Tension: An Empirical Study on What Generates Tipping Points
von: Naik, Canishk, et al.
Veröffentlicht: (2024)
von: Naik, Canishk, et al.
Veröffentlicht: (2024)
Cluster-to-Predict Affect Contours from Speech
von: Kuşçu, Gökhan, et al.
Veröffentlicht: (2024)
von: Kuşçu, Gökhan, et al.
Veröffentlicht: (2024)
More-than-Human Storytelling: Designing Longitudinal Narrative Engagements with Generative AI
von: Fabre, Émilie, et al.
Veröffentlicht: (2025)
von: Fabre, Émilie, et al.
Veröffentlicht: (2025)
USpeech: Ultrasound-Enhanced Speech with Minimal Human Effort via Cross-Modal Synthesis
von: Yu, Luca Jiang-Tao, et al.
Veröffentlicht: (2024)
von: Yu, Luca Jiang-Tao, et al.
Veröffentlicht: (2024)
Adaptation and Optimization of Automatic Speech Recognition (ASR) for the Maritime Domain in the Field of VHF Communication
von: Nakilcioglu, Emin Cagatay, et al.
Veröffentlicht: (2023)
von: Nakilcioglu, Emin Cagatay, et al.
Veröffentlicht: (2023)
Layer-Wise Analysis of Self-Supervised Representations for Age and Gender Classification in Children's Speech
von: Sinha, Abhijit, et al.
Veröffentlicht: (2025)
von: Sinha, Abhijit, et al.
Veröffentlicht: (2025)
Emotion-Disentangled Embedding Alignment for Noise-Robust and Cross-Corpus Speech Emotion Recognition
von: Tiwari, Upasana, et al.
Veröffentlicht: (2025)
von: Tiwari, Upasana, et al.
Veröffentlicht: (2025)
Toward using Speech to Sense Student Emotion in Remote Learning Environments
von: Vyas, Sargam, et al.
Veröffentlicht: (2026)
von: Vyas, Sargam, et al.
Veröffentlicht: (2026)
Speech-driven Personalized Gesture Synthetics: Harnessing Automatic Fuzzy Feature Inference
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Cross-Modal Approach to Silent Speech with LLM-Enhanced Recognition
von: Benster, Tyler, et al.
Veröffentlicht: (2024) -
STAA-Net: A Sparse and Transferable Adversarial Attack for Speech Emotion Recognition
von: Chang, Yi, et al.
Veröffentlicht: (2024) -
Beamforming-LLM: What, Where and When Did I Miss?
von: Choudhari, Vishal
Veröffentlicht: (2025) -
VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching
von: Guo, Yiwei, et al.
Veröffentlicht: (2023) -
Human Perception of Audio Deepfakes
von: Müller, Nicolas M., et al.
Veröffentlicht: (2021)