Using Voice and Biofeedback to Predict User Engagement during Product Feedback Interviews
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ferrari, Alessio, Huichapa, Thaide, Spoletini, Paola, Novielli, Nicole, Fucci, Davide, Girardi, Daniela |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2021
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
STAR: Speech-to-Audio Generation via Representation Learning
von: Xie, Zeyu, et al.
Veröffentlicht: (2025)
von: Xie, Zeyu, et al.
Veröffentlicht: (2025)
FakeSound2: A Benchmark for Explainable and Generalizable Deepfake Sound Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2025)
von: Xie, Zeyu, et al.
Veröffentlicht: (2025)
PicoAudio: Enabling Precise Timestamp and Frequency Controllability of Audio Events in Text-to-audio Generation
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
AudioTime: A Temporally-aligned Audio-text Benchmark Dataset
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
Sound Safeguarding for Acoustic Measurement Using Any Sounds: Tools and Applications
von: Kawahara, Hideki, et al.
Veröffentlicht: (2025)
von: Kawahara, Hideki, et al.
Veröffentlicht: (2025)
CAST-TTS: A Simple Cross-Attention Framework for Unified Timbre Control in TTS
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
PicoAudio2: Temporal Controllable Text-to-Audio Generation with Natural Language Description
von: Zheng, Zihao, et al.
Veröffentlicht: (2025)
von: Zheng, Zihao, et al.
Veröffentlicht: (2025)
FakeSound: Deepfake General Audio Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
von: Xie, Zeyu, et al.
Veröffentlicht: (2024)
Self-supervised Pretraining for Robust Personalized Voice Activity Detection in Adverse Conditions
von: Bovbjerg, Holger Severin, et al.
Veröffentlicht: (2023)
von: Bovbjerg, Holger Severin, et al.
Veröffentlicht: (2023)
Noise-Robust Target-Speaker Voice Activity Detection Through Self-Supervised Pretraining
von: Bovbjerg, Holger Severin, et al.
Veröffentlicht: (2025)
von: Bovbjerg, Holger Severin, et al.
Veröffentlicht: (2025)
Are Music Foundation Models Better at Singing Voice Deepfake Detection? Far-Better Fuse them with Speech Foundation Models
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
Investigating Prosodic Signatures via Speech Pre-Trained Models for Audio Deepfake Source Attribution
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
Strong Alone, Stronger Together: Synergizing Modality-Binding Foundation Models with Optimal Transport for Non-Verbal Emotion Recognition
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
Multi-View Multi-Task Modeling with Speech Foundation Models for Speech Forensic Tasks
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
Avengers Assemble: Amalgamation of Non-Semantic Features for Depression Detection
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
SeQuiFi: Mitigating Catastrophic Forgetting in Speech Emotion Recognition with Sequential Class-Finetuning
von: Jain, Sarthak, et al.
Veröffentlicht: (2024)
von: Jain, Sarthak, et al.
Veröffentlicht: (2024)
Representation Loss Minimization with Randomized Selection Strategy for Efficient Environmental Fake Audio Detection
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
MENASpeechBank: A Reference Voice Bank with Persona-Conditioned Multi-Turn Conversations for AudioLLMs
von: Ali, Zien Sheikh, et al.
Veröffentlicht: (2026)
von: Ali, Zien Sheikh, et al.
Veröffentlicht: (2026)
Sink or SWIM: Tackling Real-Time ASR at Scale
von: Bruzzone, Federico, et al.
Veröffentlicht: (2026)
von: Bruzzone, Federico, et al.
Veröffentlicht: (2026)
TorchFX: A modern approach to Audio DSP with PyTorch and GPU acceleration
von: Spanio, Matteo, et al.
Veröffentlicht: (2025)
von: Spanio, Matteo, et al.
Veröffentlicht: (2025)
A Unified Model For Voice and Accent Conversion In Speech and Singing using Self-Supervised Learning and Feature Extraction
von: Cheripally, Sowmya
Veröffentlicht: (2024)
von: Cheripally, Sowmya
Veröffentlicht: (2024)
SARA: Stress Test Reasoning in Audio Deepfake Detection
von: Nguyen, Binh, et al.
Veröffentlicht: (2026)
von: Nguyen, Binh, et al.
Veröffentlicht: (2026)
Proposal of protocols for speech materials acquisition and presentation assisted by tools based on structured test signals
von: Kawahara, Hideki, et al.
Veröffentlicht: (2024)
von: Kawahara, Hideki, et al.
Veröffentlicht: (2024)
Rethinking Masking Strategies for Masked Prediction-based Audio Self-supervised Learning
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2026)
von: Niizumi, Daisuke, et al.
Veröffentlicht: (2026)
Beyond Hearing: Learning Task-Agnostic ExG Representations from Earphones via Physiology-Informed Tokenization
von: Yoon, Hyungjun, et al.
Veröffentlicht: (2025)
von: Yoon, Hyungjun, et al.
Veröffentlicht: (2025)
A Voice-based Triage for Type 2 Diabetes using a Conversational Virtual Assistant in the Home Environment
von: Summoogum, Kelvin, et al.
Veröffentlicht: (2024)
von: Summoogum, Kelvin, et al.
Veröffentlicht: (2024)
ML-ASPA: A Contemplation of Machine Learning-based Acoustic Signal Processing Analysis for Sounds, & Strains Emerging Technology
von: Ali, Ratul, et al.
Veröffentlicht: (2023)
von: Ali, Ratul, et al.
Veröffentlicht: (2023)
Window Size Versus Accuracy Experiments in Voice Activity Detectors
von: McKinnon, Max, et al.
Veröffentlicht: (2026)
von: McKinnon, Max, et al.
Veröffentlicht: (2026)
Comparison of parameters of vowel sounds of russian and english languages
von: Fedoseev, V. I., et al.
Veröffentlicht: (2024)
von: Fedoseev, V. I., et al.
Veröffentlicht: (2024)
Noise-Robust Keyword Spotting through Self-supervised Pretraining
von: Mørk, Jacob, et al.
Veröffentlicht: (2024)
von: Mørk, Jacob, et al.
Veröffentlicht: (2024)
Learning Robust Spatial Representations from Binaural Audio through Feature Distillation
von: Bovbjerg, Holger Severin, et al.
Veröffentlicht: (2025)
von: Bovbjerg, Holger Severin, et al.
Veröffentlicht: (2025)
Monaural Multi-Speaker Speech Separation Using Efficient Transformer Model
von: Rijal, S., et al.
Veröffentlicht: (2023)
von: Rijal, S., et al.
Veröffentlicht: (2023)
Detecting Check-Worthy Claims in Political Debates, Speeches, and Interviews Using Audio Data
von: Ivanov, Petar, et al.
Veröffentlicht: (2023)
von: Ivanov, Petar, et al.
Veröffentlicht: (2023)
Toward Low-Latency End-to-End Voice Agents for Telecommunications Using Streaming ASR, Quantized LLMs, and Real-Time TTS
von: Ethiraj, Vignesh, et al.
Veröffentlicht: (2025)
von: Ethiraj, Vignesh, et al.
Veröffentlicht: (2025)
Multi-Task Instruction Tuning via Data Scheduling for Low-Resource Arabic AudioLLMs
von: Bhatti, Hunzalah Hassan, et al.
Veröffentlicht: (2026)
von: Bhatti, Hunzalah Hassan, et al.
Veröffentlicht: (2026)
AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks
von: Maben, Leander Melroy, et al.
Veröffentlicht: (2025)
von: Maben, Leander Melroy, et al.
Veröffentlicht: (2025)
Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English
von: Zhang, Haoyang, et al.
Veröffentlicht: (2025)
von: Zhang, Haoyang, et al.
Veröffentlicht: (2025)
SeamlessEdit: Background Noise Aware Zero-Shot Speech Editing with in-Context Enhancement
von: Chen, Kuan-Yu, et al.
Veröffentlicht: (2025)
von: Chen, Kuan-Yu, et al.
Veröffentlicht: (2025)
Large Vocabulary Spontaneous Speech Recognition for Tigrigna
von: Kahsu, Ataklti, et al.
Veröffentlicht: (2023)
von: Kahsu, Ataklti, et al.
Veröffentlicht: (2023)
Text2midi-InferAlign: Improving Symbolic Music Generation with Inference-Time Alignment
von: Roy, Abhinaba, et al.
Veröffentlicht: (2025)
von: Roy, Abhinaba, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
STAR: Speech-to-Audio Generation via Representation Learning
von: Xie, Zeyu, et al.
Veröffentlicht: (2025) -
FakeSound2: A Benchmark for Explainable and Generalizable Deepfake Sound Detection
von: Xie, Zeyu, et al.
Veröffentlicht: (2025) -
PicoAudio: Enabling Precise Timestamp and Frequency Controllability of Audio Events in Text-to-audio Generation
von: Xie, Zeyu, et al.
Veröffentlicht: (2024) -
AudioTime: A Temporally-aligned Audio-text Benchmark Dataset
von: Xie, Zeyu, et al.
Veröffentlicht: (2024) -
Sound Safeguarding for Acoustic Measurement Using Any Sounds: Tools and Applications
von: Kawahara, Hideki, et al.
Veröffentlicht: (2025)