DOO-RE: A dataset of ambient sensors in a meeting room for activity recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Hyunju, Kim, Geon, Lee, Taehoon, Kim, Kisoo, Lee, Dongman |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DESAMO: A Device for Elder-Friendly Smart Homes Powered by Embedded LLM with Audio Modality
von: Choi, Youngwon, et al.
Veröffentlicht: (2025)
von: Choi, Youngwon, et al.
Veröffentlicht: (2025)
Evaluating Spatialized Auditory Cues for Rapid Attention Capture in XR
von: Kim, Yoonsang, et al.
Veröffentlicht: (2026)
von: Kim, Yoonsang, et al.
Veröffentlicht: (2026)
The effect of self-motion and room familiarity on sound source localization in virtual environments
von: Isserstedt, Niklas, et al.
Veröffentlicht: (2024)
von: Isserstedt, Niklas, et al.
Veröffentlicht: (2024)
SCDiar: a streaming diarization system based on speaker change detection and speech recognition
von: Zheng, Naijun, et al.
Veröffentlicht: (2025)
von: Zheng, Naijun, et al.
Veröffentlicht: (2025)
Quality Audio Prototyping: a prototype system for unified sound retrieval and procedural generation
von: Garcia, Nelly, et al.
Veröffentlicht: (2026)
von: Garcia, Nelly, et al.
Veröffentlicht: (2026)
Artificial Neural Networks to Recognize Speakers Division from Continuous Bengali Speech
von: Ali, Hasmot, et al.
Veröffentlicht: (2024)
von: Ali, Hasmot, et al.
Veröffentlicht: (2024)
Harnessing Smartwatch Microphone Sensors for Cough Detection and Classification
von: Jaiswal, Pranay, et al.
Veröffentlicht: (2024)
von: Jaiswal, Pranay, et al.
Veröffentlicht: (2024)
SonicSieve: Bringing Directional Speech Extraction to Smartphones Using Acoustic Microstructures
von: Yuan, Kuang, et al.
Veröffentlicht: (2025)
von: Yuan, Kuang, et al.
Veröffentlicht: (2025)
Improving AI-generated music with user-guided training
von: Singh, Vishwa Mohan, et al.
Veröffentlicht: (2025)
von: Singh, Vishwa Mohan, et al.
Veröffentlicht: (2025)
Human Feedback Driven Dynamic Speech Emotion Recognition
von: Fedorov, Ilya, et al.
Veröffentlicht: (2025)
von: Fedorov, Ilya, et al.
Veröffentlicht: (2025)
Detecting COPD Through Speech Analysis: A Dataset of Danish Speech and Machine Learning Approach
von: Sankey-Olsen, Cuno, et al.
Veröffentlicht: (2025)
von: Sankey-Olsen, Cuno, et al.
Veröffentlicht: (2025)
Optimizing Multilingual Text-To-Speech with Accents & Emotions
von: Pawar, Pranav, et al.
Veröffentlicht: (2025)
von: Pawar, Pranav, et al.
Veröffentlicht: (2025)
Voice Passing : a Non-Binary Voice Gender Prediction System for evaluating Transgender voice transition
von: Doukhan, David, et al.
Veröffentlicht: (2024)
von: Doukhan, David, et al.
Veröffentlicht: (2024)
NeoLightning: A Modern Reimagination of Gesture-Based Sound Design
von: Kim, Yonghyun, et al.
Veröffentlicht: (2025)
von: Kim, Yonghyun, et al.
Veröffentlicht: (2025)
Audio2Face-3D: Audio-driven Realistic Facial Animation For Digital Avatars
von: NVIDIA, et al.
Veröffentlicht: (2025)
von: NVIDIA, et al.
Veröffentlicht: (2025)
AI TrackMate: Finally, Someone Who Will Give Your Music More Than Just "Sounds Great!"
von: Jiang, Yi-Lin, et al.
Veröffentlicht: (2024)
von: Jiang, Yi-Lin, et al.
Veröffentlicht: (2024)
Workflow-Based Evaluation of Music Generation Systems
von: Dadman, Shayan, et al.
Veröffentlicht: (2025)
von: Dadman, Shayan, et al.
Veröffentlicht: (2025)
Revisiting Your Memory: Reconstruction of Affect-Contextualized Memory via EEG-guided Audiovisual Generation
von: Kwon, Joonwoo, et al.
Veröffentlicht: (2024)
von: Kwon, Joonwoo, et al.
Veröffentlicht: (2024)
Active noise cancellation on open-ear smart glasses
von: Yuan, Kuang, et al.
Veröffentlicht: (2026)
von: Yuan, Kuang, et al.
Veröffentlicht: (2026)
VoXtream: Full-Stream Text-to-Speech with Extremely Low Latency
von: Torgashov, Nikita, et al.
Veröffentlicht: (2025)
von: Torgashov, Nikita, et al.
Veröffentlicht: (2025)
Abjad-Kids: An Arabic Speech Classification Dataset for Primary Education
von: Snoubara, Abdul Aziz, et al.
Veröffentlicht: (2026)
von: Snoubara, Abdul Aziz, et al.
Veröffentlicht: (2026)
Spontaneous Informal Speech Dataset for Punctuation Restoration
von: Liu, Xing Yi, et al.
Veröffentlicht: (2024)
von: Liu, Xing Yi, et al.
Veröffentlicht: (2024)
AIx Speed: Playback Speed Optimization Using Listening Comprehension of Speech Recognition Models
von: Kawamura, Kazuki, et al.
Veröffentlicht: (2024)
von: Kawamura, Kazuki, et al.
Veröffentlicht: (2024)
Loop Copilot: Conducting AI Ensembles for Music Generation and Iterative Editing
von: Zhang, Yixiao, et al.
Veröffentlicht: (2023)
von: Zhang, Yixiao, et al.
Veröffentlicht: (2023)
A conversational gesture synthesis system based on emotions and semantics
von: Hoang-Minh, Thanh
Veröffentlicht: (2025)
von: Hoang-Minh, Thanh
Veröffentlicht: (2025)
LLAMAPIE: Proactive In-Ear Conversation Assistants
von: Chen, Tuochao, et al.
Veröffentlicht: (2025)
von: Chen, Tuochao, et al.
Veröffentlicht: (2025)
VoXtream2: Full-stream TTS with dynamic speaking rate control
von: Torgashov, Nikita, et al.
Veröffentlicht: (2026)
von: Torgashov, Nikita, et al.
Veröffentlicht: (2026)
Literary and Colloquial Tamil Dialect Identification
von: Nanmalar, M., et al.
Veröffentlicht: (2024)
von: Nanmalar, M., et al.
Veröffentlicht: (2024)
Efficient Ensemble for Multimodal Punctuation Restoration using Time-Delay Neural Network
von: Liu, Xing Yi, et al.
Veröffentlicht: (2023)
von: Liu, Xing Yi, et al.
Veröffentlicht: (2023)
Investigating the Effects of Large-Scale Pseudo-Stereo Data and Different Speech Foundation Model on Dialogue Generative Spoken Language Model
von: Fu, Yu-Kuan, et al.
Veröffentlicht: (2024)
von: Fu, Yu-Kuan, et al.
Veröffentlicht: (2024)
Tailors: New Music Timbre Visualizer to Entertain Music Through Imagery
von: Lee, ChungHa
Veröffentlicht: (2024)
von: Lee, ChungHa
Veröffentlicht: (2024)
Adaptation and Optimization of Automatic Speech Recognition (ASR) for the Maritime Domain in the Field of VHF Communication
von: Nakilcioglu, Emin Cagatay, et al.
Veröffentlicht: (2023)
von: Nakilcioglu, Emin Cagatay, et al.
Veröffentlicht: (2023)
Towards Decoding Brain Activity During Passive Listening of Speech
von: Fodor, Milán András, et al.
Veröffentlicht: (2024)
von: Fodor, Milán András, et al.
Veröffentlicht: (2024)
Layer-Wise Analysis of Self-Supervised Representations for Age and Gender Classification in Children's Speech
von: Sinha, Abhijit, et al.
Veröffentlicht: (2025)
von: Sinha, Abhijit, et al.
Veröffentlicht: (2025)
Emotion-Disentangled Embedding Alignment for Noise-Robust and Cross-Corpus Speech Emotion Recognition
von: Tiwari, Upasana, et al.
Veröffentlicht: (2025)
von: Tiwari, Upasana, et al.
Veröffentlicht: (2025)
Composers' Evaluations of an AI Music Tool: Insights for Human-Centred Design
von: Row, Eleanor, et al.
Veröffentlicht: (2024)
von: Row, Eleanor, et al.
Veröffentlicht: (2024)
Homogeneous Speaker Features for On-the-Fly Dysarthric and Elderly Speaker Adaptation
von: Geng, Mengzhe, et al.
Veröffentlicht: (2024)
von: Geng, Mengzhe, et al.
Veröffentlicht: (2024)
Music Generation using Human-In-The-Loop Reinforcement Learning
von: Justus, Aju Ani
Veröffentlicht: (2025)
von: Justus, Aju Ani
Veröffentlicht: (2025)
Learning Relationships Between Separate Audio Tracks for Creative Applications
von: Bujard, Balthazar, et al.
Veröffentlicht: (2025)
von: Bujard, Balthazar, et al.
Veröffentlicht: (2025)
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
von: Khanday, Owais Mujtaba, et al.
Veröffentlicht: (2025)
von: Khanday, Owais Mujtaba, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DESAMO: A Device for Elder-Friendly Smart Homes Powered by Embedded LLM with Audio Modality
von: Choi, Youngwon, et al.
Veröffentlicht: (2025) -
Evaluating Spatialized Auditory Cues for Rapid Attention Capture in XR
von: Kim, Yoonsang, et al.
Veröffentlicht: (2026) -
The effect of self-motion and room familiarity on sound source localization in virtual environments
von: Isserstedt, Niklas, et al.
Veröffentlicht: (2024) -
SCDiar: a streaming diarization system based on speaker change detection and speech recognition
von: Zheng, Naijun, et al.
Veröffentlicht: (2025) -
Quality Audio Prototyping: a prototype system for unified sound retrieval and procedural generation
von: Garcia, Nelly, et al.
Veröffentlicht: (2026)