Definition-independent Formalization of Soundscapes: Towards a Formal Methodology
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jedrusiak, Mikel D., Harweg, Thomas, Haselhoff, Timo, Lawrence, Bryce T., Moebus, Susanne, Weichert, Frank |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PSM: Learning Probabilistic Embeddings for Multi-scale Zero-Shot Soundscape Mapping
von: Khanal, Subash, et al.
Veröffentlicht: (2024)
von: Khanal, Subash, et al.
Veröffentlicht: (2024)
Seeing Soundscapes: Audio-Visual Generation and Separation from Soundscapes Using Audio-Visual Separator
von: Kang, Minjae, et al.
Veröffentlicht: (2025)
von: Kang, Minjae, et al.
Veröffentlicht: (2025)
Self-Supervised Audio-Visual Soundscape Stylization
von: Li, Tingle, et al.
Veröffentlicht: (2024)
von: Li, Tingle, et al.
Veröffentlicht: (2024)
Feature Selection via Graph Topology Inference for Soundscape Emotion Recognition
von: Rey, Samuel, et al.
Veröffentlicht: (2025)
von: Rey, Samuel, et al.
Veröffentlicht: (2025)
Autonomous Soundscape Augmentation with Multimodal Fusion of Visual and Participant-linked Inputs
von: Ooi, Kenneth, et al.
Veröffentlicht: (2023)
von: Ooi, Kenneth, et al.
Veröffentlicht: (2023)
Robust Bioacoustic Detection via Richly Labelled Synthetic Soundscape Augmentation
von: Soltero, Kaspar, et al.
Veröffentlicht: (2025)
von: Soltero, Kaspar, et al.
Veröffentlicht: (2025)
Generating Diverse Audio-Visual 360 Soundscapes for Sound Event Localization and Detection
von: Roman, Adrian S., et al.
Veröffentlicht: (2025)
von: Roman, Adrian S., et al.
Veröffentlicht: (2025)
ARAUS: A Large-Scale Dataset and Baseline Models of Affective Responses to Augmented Urban Soundscapes
von: Ooi, Kenneth, et al.
Veröffentlicht: (2022)
von: Ooi, Kenneth, et al.
Veröffentlicht: (2022)
Automating Urban Soundscape Enhancements with AI: In-situ Assessment of Quality and Restorativeness in Traffic-Exposed Residential Areas
von: Lam, Bhan, et al.
Veröffentlicht: (2024)
von: Lam, Bhan, et al.
Veröffentlicht: (2024)
Soundscape Captioning using Sound Affective Quality Network and Large Language Model
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
Fine-grained Soundscape Control for Augmented Hearing
von: Oh, Seunghyun, et al.
Veröffentlicht: (2026)
von: Oh, Seunghyun, et al.
Veröffentlicht: (2026)
Sound Tagging in Infant-centric Home Soundscapes
von: Khan, Mohammad Nur Hossain, et al.
Veröffentlicht: (2024)
von: Khan, Mohammad Nur Hossain, et al.
Veröffentlicht: (2024)
Misophonia Trigger Sound Detection on Synthetic Soundscapes Using a Hybrid Model with a Frozen Pre-Trained CNN and a Time-Series Module
von: Sashida, Kurumi, et al.
Veröffentlicht: (2026)
von: Sashida, Kurumi, et al.
Veröffentlicht: (2026)
Towards Video to Piano Music Generation with Chain-of-Perform Support Benchmarks
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
Omni-AVSR: Towards Unified Multimodal Speech Recognition with Large Language Models
von: Cappellazzo, Umberto, et al.
Veröffentlicht: (2025)
von: Cappellazzo, Umberto, et al.
Veröffentlicht: (2025)
Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation
von: Sun, Peiwen, et al.
Veröffentlicht: (2024)
von: Sun, Peiwen, et al.
Veröffentlicht: (2024)
SSAVSV: Towards Unified Model for Self-Supervised Audio-Visual Speaker Verification
von: Rajasekhar, Gnana Praveen, et al.
Veröffentlicht: (2025)
von: Rajasekhar, Gnana Praveen, et al.
Veröffentlicht: (2025)
Towards Reliable Audio Deepfake Attribution and Model Recognition: A Multi-Level Autoencoder-Based Framework
von: Di Pierno, Andrea, et al.
Veröffentlicht: (2025)
von: Di Pierno, Andrea, et al.
Veröffentlicht: (2025)
VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction
von: Fu, Chaoyou, et al.
Veröffentlicht: (2025)
von: Fu, Chaoyou, et al.
Veröffentlicht: (2025)
Generating Moving 3D Soundscapes with Latent Diffusion Models
von: Templin, Christian, et al.
Veröffentlicht: (2025)
von: Templin, Christian, et al.
Veröffentlicht: (2025)
Do We Need EMA for Diffusion-Based Speech Enhancement? Toward a Magnitude-Preserving Network Architecture
von: Richter, Julius, et al.
Veröffentlicht: (2025)
von: Richter, Julius, et al.
Veröffentlicht: (2025)
DeepAudio-V1:Towards Multi-Modal Multi-Stage End-to-End Video to Speech and Audio Generation
von: Zhang, Haomin, et al.
Veröffentlicht: (2025)
von: Zhang, Haomin, et al.
Veröffentlicht: (2025)
UWAV: Uncertainty-weighted Weakly-supervised Audio-Visual Video Parsing
von: Lai, Yung-Hsuan, et al.
Veröffentlicht: (2025)
von: Lai, Yung-Hsuan, et al.
Veröffentlicht: (2025)
mWhisper-Flamingo for Multilingual Audio-Visual Noise-Robust Speech Recognition
von: Rouditchenko, Andrew, et al.
Veröffentlicht: (2025)
von: Rouditchenko, Andrew, et al.
Veröffentlicht: (2025)
Listen, Chat, and Remix: Text-Guided Soundscape Remixing for Enhanced Auditory Experience
von: Jiang, Xilin, et al.
Veröffentlicht: (2024)
von: Jiang, Xilin, et al.
Veröffentlicht: (2024)
Whisper-Flamingo: Integrating Visual Features into Whisper for Audio-Visual Speech Recognition and Translation
von: Rouditchenko, Andrew, et al.
Veröffentlicht: (2024)
von: Rouditchenko, Andrew, et al.
Veröffentlicht: (2024)
MDSGen: Fast and Efficient Masked Diffusion Temporal-Aware Transformers for Open-Domain Sound Generation
von: Pham, Trung X., et al.
Veröffentlicht: (2024)
von: Pham, Trung X., et al.
Veröffentlicht: (2024)
Towards a generalized monaural and binaural auditory model for psychoacoustics and speech intelligibility
von: Biberger, Thomas, et al.
Veröffentlicht: (2021)
von: Biberger, Thomas, et al.
Veröffentlicht: (2021)
Towards Accurate Lip-to-Speech Synthesis in-the-Wild
von: Hegde, Sindhu, et al.
Veröffentlicht: (2024)
von: Hegde, Sindhu, et al.
Veröffentlicht: (2024)
Improving Acoustic Scene Classification with City Features
von: Cai, Yiqiang, et al.
Veröffentlicht: (2025)
von: Cai, Yiqiang, et al.
Veröffentlicht: (2025)
ASiT: Local-Global Audio Spectrogram vIsion Transformer for Event Classification
von: Atito, Sara, et al.
Veröffentlicht: (2022)
von: Atito, Sara, et al.
Veröffentlicht: (2022)
Emotional Vietnamese Speech-Based Depression Diagnosis Using Dynamic Attention Mechanism
von: D., Quang-Anh N., et al.
Veröffentlicht: (2024)
von: D., Quang-Anh N., et al.
Veröffentlicht: (2024)
Gotta Hear Them All: Towards Sound Source Aware Audio Generation
von: Guo, Wei, et al.
Veröffentlicht: (2024)
von: Guo, Wei, et al.
Veröffentlicht: (2024)
On Feature Learning for Titi Monkey Activity Detection
von: Ravuri, Aditya, et al.
Veröffentlicht: (2024)
von: Ravuri, Aditya, et al.
Veröffentlicht: (2024)
Spatial Scaper: A Library to Simulate and Augment Soundscapes for Sound Event Localization and Detection in Realistic Rooms
von: Roman, Iran R., et al.
Veröffentlicht: (2024)
von: Roman, Iran R., et al.
Veröffentlicht: (2024)
ReverbFX: A Dataset of Room Impulse Responses Derived from Reverb Effect Plugins for Singing Voice Dereverberation
von: Richter, Julius, et al.
Veröffentlicht: (2025)
von: Richter, Julius, et al.
Veröffentlicht: (2025)
SoundWeaver: Semantic Warm-Starting for Text-to-Audio Diffusion Serving
von: Barik, Ayush, et al.
Veröffentlicht: (2026)
von: Barik, Ayush, et al.
Veröffentlicht: (2026)
Few-shot Acoustic Synthesis with Multimodal Flow Matching
von: Brunetto, Amandine
Veröffentlicht: (2026)
von: Brunetto, Amandine
Veröffentlicht: (2026)
pycnet-audio: A Python package to support bioacoustics data processing
von: Ruff, Zachary J., et al.
Veröffentlicht: (2025)
von: Ruff, Zachary J., et al.
Veröffentlicht: (2025)
V2SFlow: Video-to-Speech Generation with Speech Decomposition and Rectified Flow
von: Choi, Jeongsoo, et al.
Veröffentlicht: (2024)
von: Choi, Jeongsoo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
PSM: Learning Probabilistic Embeddings for Multi-scale Zero-Shot Soundscape Mapping
von: Khanal, Subash, et al.
Veröffentlicht: (2024) -
Seeing Soundscapes: Audio-Visual Generation and Separation from Soundscapes Using Audio-Visual Separator
von: Kang, Minjae, et al.
Veröffentlicht: (2025) -
Self-Supervised Audio-Visual Soundscape Stylization
von: Li, Tingle, et al.
Veröffentlicht: (2024) -
Feature Selection via Graph Topology Inference for Soundscape Emotion Recognition
von: Rey, Samuel, et al.
Veröffentlicht: (2025) -
Autonomous Soundscape Augmentation with Multimodal Fusion of Visual and Participant-linked Inputs
von: Ooi, Kenneth, et al.
Veröffentlicht: (2023)