Soundify: Matching Sound Effects to Video
Fuente:
arXiv
Salvato in:
| Autori principali: | Lin, David Chuan-En, Germanidis, Anastasis, Valenzuela, Cristóbal, Shi, Yining, Martelaro, Nikolas |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2021
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
NeoLightning: A Modern Reimagination of Gesture-Based Sound Design
di: Kim, Yonghyun, et al.
Pubblicazione: (2025)
di: Kim, Yonghyun, et al.
Pubblicazione: (2025)
VidTune: Creating Video Soundtracks with Generative Music and Contextual Thumbnails
di: Huh, Mina, et al.
Pubblicazione: (2026)
di: Huh, Mina, et al.
Pubblicazione: (2026)
MR-DAW: Towards Collaborative Digital Audio Workstations in Mixed Reality
di: Hopkins, Torin, et al.
Pubblicazione: (2026)
di: Hopkins, Torin, et al.
Pubblicazione: (2026)
Real-Time Word-Level Temporal Segmentation in Streaming Speech Recognition
di: Nishida, Naoto, et al.
Pubblicazione: (2025)
di: Nishida, Naoto, et al.
Pubblicazione: (2025)
Capturing Cancer as Music: Cancer Mechanisms Expressed through Musification
di: Hnatyshyn, Rostyslav, et al.
Pubblicazione: (2024)
di: Hnatyshyn, Rostyslav, et al.
Pubblicazione: (2024)
Assessing the Viability of Wave Field Synthesis in VR-Based Cognitive Research
di: Kahl, Benjamin
Pubblicazione: (2025)
di: Kahl, Benjamin
Pubblicazione: (2025)
Creating Aesthetic Sonifications on the Web with SIREN
di: Peng, Tristan, et al.
Pubblicazione: (2024)
di: Peng, Tristan, et al.
Pubblicazione: (2024)
Robust Dual-Modal Speech Keyword Spotting for XR Headsets
di: Cai, Zhuojiang, et al.
Pubblicazione: (2024)
di: Cai, Zhuojiang, et al.
Pubblicazione: (2024)
AVE Speech: A Comprehensive Multi-Modal Dataset for Speech Recognition Integrating Audio, Visual, and Electromyographic Signals
di: Zhou, Dongliang, et al.
Pubblicazione: (2025)
di: Zhou, Dongliang, et al.
Pubblicazione: (2025)
AI TrackMate: Finally, Someone Who Will Give Your Music More Than Just "Sounds Great!"
di: Jiang, Yi-Lin, et al.
Pubblicazione: (2024)
di: Jiang, Yi-Lin, et al.
Pubblicazione: (2024)
Flowers Revisited: A Preliminary Replication of Flowers et al. 1997
di: Enge, Kajetan, et al.
Pubblicazione: (2024)
di: Enge, Kajetan, et al.
Pubblicazione: (2024)
Towards Reliable Large Audio Language Model
di: Ma, Ziyang, et al.
Pubblicazione: (2025)
di: Ma, Ziyang, et al.
Pubblicazione: (2025)
Acoustic Wave Modeling Using 2D FDTD: Applications in Unreal Engine For Dynamic Sound Rendering
di: Samsurya, Bilkent
Pubblicazione: (2025)
di: Samsurya, Bilkent
Pubblicazione: (2025)
MetaBGM: Dynamic Soundtrack Transformation For Continuous Multi-Scene Experiences With Ambient Awareness And Personalization
di: Liu, Haoxuan, et al.
Pubblicazione: (2024)
di: Liu, Haoxuan, et al.
Pubblicazione: (2024)
Proceedings of The second international workshop on eXplainable AI for the Arts (XAIxArts)
di: Bryan-Kinns, Nick, et al.
Pubblicazione: (2024)
di: Bryan-Kinns, Nick, et al.
Pubblicazione: (2024)
A Multi-Agent AI Framework for Immersive Audiobook Production through Spatial Audio and Neural Narration
di: Selvamani, Shaja Arul, et al.
Pubblicazione: (2025)
di: Selvamani, Shaja Arul, et al.
Pubblicazione: (2025)
Workflow-Based Evaluation of Music Generation Systems
di: Dadman, Shayan, et al.
Pubblicazione: (2025)
di: Dadman, Shayan, et al.
Pubblicazione: (2025)
Freetalker: Controllable Speech and Text-Driven Gesture Generation Based on Diffusion Models for Enhanced Speaker Naturalness
di: Yang, Sicheng, et al.
Pubblicazione: (2024)
di: Yang, Sicheng, et al.
Pubblicazione: (2024)
G-STAR: End-to-End Global Speaker-Tracking Attributed Recognition
di: Peng, Jing, et al.
Pubblicazione: (2026)
di: Peng, Jing, et al.
Pubblicazione: (2026)
DiM-Gestor: Co-Speech Gesture Generation with Adaptive Layer Normalization Mamba-2
di: Zhang, Fan, et al.
Pubblicazione: (2024)
di: Zhang, Fan, et al.
Pubblicazione: (2024)
MIST: Multimodal Interactive Speech-based Tool-calling Conversational Assistants for Smart Homes
di: Chen, Maximillian, et al.
Pubblicazione: (2026)
di: Chen, Maximillian, et al.
Pubblicazione: (2026)
HiCMAE: Hierarchical Contrastive Masked Autoencoder for Self-Supervised Audio-Visual Emotion Recognition
di: Sun, Licai, et al.
Pubblicazione: (2024)
di: Sun, Licai, et al.
Pubblicazione: (2024)
SoundShift: Exploring Sound Manipulations for Accessible Mixed-Reality Awareness
di: Chang, Ruei-Che, et al.
Pubblicazione: (2024)
di: Chang, Ruei-Che, et al.
Pubblicazione: (2024)
Seeing Beyond Sound: Visualization and Abstraction in Audio Data Representation
di: Blum'e, Ashlae
Pubblicazione: (2025)
di: Blum'e, Ashlae
Pubblicazione: (2025)
Revival: Collaborative Artistic Creation through Human-AI Interactions in Musical Creativity
di: Lee, Keon Ju M., et al.
Pubblicazione: (2025)
di: Lee, Keon Ju M., et al.
Pubblicazione: (2025)
SACM: SEEG-Audio Contrastive Matching for Chinese Speech Decoding
di: Wang, Hongbin, et al.
Pubblicazione: (2025)
di: Wang, Hongbin, et al.
Pubblicazione: (2025)
Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings
di: Li, Yinan, et al.
Pubblicazione: (2026)
di: Li, Yinan, et al.
Pubblicazione: (2026)
UltrasonicSpheres: Localized, Multi-Channel Sound Spheres Using Off-the-Shelf Speakers and Earables
di: Küttner, Michael, et al.
Pubblicazione: (2025)
di: Küttner, Michael, et al.
Pubblicazione: (2025)
Sound Judgment: Properties of Consequential Sounds Affecting Human-Perception of Robots
di: Allen, Aimee, et al.
Pubblicazione: (2025)
di: Allen, Aimee, et al.
Pubblicazione: (2025)
Sound-Based Recognition of Touch Gestures and Emotions for Enhanced Human-Robot Interaction
di: Hou, Yuanbo, et al.
Pubblicazione: (2024)
di: Hou, Yuanbo, et al.
Pubblicazione: (2024)
Robotic Blended Sonification: Consequential Robot Sound as Creative Material for Human-Robot Interaction
di: Johansen, Stine S., et al.
Pubblicazione: (2024)
di: Johansen, Stine S., et al.
Pubblicazione: (2024)
YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls
di: Chen, Zihao, et al.
Pubblicazione: (2024)
di: Chen, Zihao, et al.
Pubblicazione: (2024)
Directional Source Separation for Robust Speech Recognition on Smart Glasses
di: Feng, Tiantian, et al.
Pubblicazione: (2023)
di: Feng, Tiantian, et al.
Pubblicazione: (2023)
SoundingActions: Learning How Actions Sound from Narrated Egocentric Videos
di: Chen, Changan, et al.
Pubblicazione: (2024)
di: Chen, Changan, et al.
Pubblicazione: (2024)
SoundMind: RL-Incentivized Logic Reasoning for Audio-Language Models
di: Diao, Xingjian, et al.
Pubblicazione: (2025)
di: Diao, Xingjian, et al.
Pubblicazione: (2025)
Video-Guided Foley Sound Generation with Multimodal Controls
di: Chen, Ziyang, et al.
Pubblicazione: (2024)
di: Chen, Ziyang, et al.
Pubblicazione: (2024)
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2025)
A Mapping Strategy for Interacting with Latent Audio Synthesis Using Artistic Materials
di: Zheng, Shuoyang, et al.
Pubblicazione: (2024)
di: Zheng, Shuoyang, et al.
Pubblicazione: (2024)
Enhancing DMI Interactions by Integrating Haptic Feedback for Intricate Vibrato Technique
di: Piao, Ziyue, et al.
Pubblicazione: (2024)
di: Piao, Ziyue, et al.
Pubblicazione: (2024)
Documenti analoghi
-
NeoLightning: A Modern Reimagination of Gesture-Based Sound Design
di: Kim, Yonghyun, et al.
Pubblicazione: (2025) -
VidTune: Creating Video Soundtracks with Generative Music and Contextual Thumbnails
di: Huh, Mina, et al.
Pubblicazione: (2026) -
MR-DAW: Towards Collaborative Digital Audio Workstations in Mixed Reality
di: Hopkins, Torin, et al.
Pubblicazione: (2026) -
Real-Time Word-Level Temporal Segmentation in Streaming Speech Recognition
di: Nishida, Naoto, et al.
Pubblicazione: (2025) -
Capturing Cancer as Music: Cancer Mechanisms Expressed through Musification
di: Hnatyshyn, Rostyslav, et al.
Pubblicazione: (2024)