Salvato in:
| Autori principali: | Ritu, Jarin, Van Dine, Alexandra, Peeples, Joshua |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2504.14843 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Structural and Statistical Audio Texture Knowledge Distillation for Acoustic Classification
di: Ritu, Jarin, et al.
Pubblicazione: (2025)
di: Ritu, Jarin, et al.
Pubblicazione: (2025)
Neural Edge Histogram Descriptors for Underwater Acoustic Target Recognition
di: Agashe, Atharva, et al.
Pubblicazione: (2025)
di: Agashe, Atharva, et al.
Pubblicazione: (2025)
Cross-Domain Knowledge Transfer for Underwater Acoustic Classification Using Pre-trained Models
di: Mohammadi, Amirmohammad, et al.
Pubblicazione: (2024)
di: Mohammadi, Amirmohammad, et al.
Pubblicazione: (2024)
Investigation of Time-Frequency Feature Combinations with Histogram Layer Time Delay Neural Networks
di: Mohammadi, Amirmohammad, et al.
Pubblicazione: (2024)
di: Mohammadi, Amirmohammad, et al.
Pubblicazione: (2024)
Quantitative Analysis of Proxy Tasks for Anomalous Sound Detection
di: Shin, Seunghyeon, et al.
Pubblicazione: (2026)
di: Shin, Seunghyeon, et al.
Pubblicazione: (2026)
A Statistics-Driven Differentiable Approach for Sound Texture Synthesis and Analysis
di: Gutiérrez, Esteban, et al.
Pubblicazione: (2025)
di: Gutiérrez, Esteban, et al.
Pubblicazione: (2025)
Towards a Quantitative Analysis of Coarticulation with a Phoneme-to-Articulatory Model
di: Fan, Chaofei, et al.
Pubblicazione: (2024)
di: Fan, Chaofei, et al.
Pubblicazione: (2024)
Raw Audio Classification with Cosine Convolutional Neural Network (CosCovNN)
di: Haque, Kazi Nazmul, et al.
Pubblicazione: (2024)
di: Haque, Kazi Nazmul, et al.
Pubblicazione: (2024)
HiRIS: an Airborne Sonar Sensor with a 1024 Channel Microphone Array for In-Air Acoustic Imaging
di: Laurijssen, Dennis, et al.
Pubblicazione: (2024)
di: Laurijssen, Dennis, et al.
Pubblicazione: (2024)
Multiple Mobile Target Detection and Tracking in Active Sonar Array Using a Track-Before-Detect Approach
di: Abu, Avi, et al.
Pubblicazione: (2024)
di: Abu, Avi, et al.
Pubblicazione: (2024)
Subspace Track-before-Detect for Passive Multi-Target Tracking with Unknown Emitted Signals
di: Ito, Nobutaka, et al.
Pubblicazione: (2026)
di: Ito, Nobutaka, et al.
Pubblicazione: (2026)
METEOR: Melody-aware Texture-controllable Symbolic Orchestral Music Generation via Transformer VAE
di: Le, Dinh-Viet-Toan, et al.
Pubblicazione: (2024)
di: Le, Dinh-Viet-Toan, et al.
Pubblicazione: (2024)
SEABAD: A Tropical Bird Activity Detection Dataset for Passive Acoustic Monitoring
di: Zabidi, Muhammad Mun'im Ahmad, et al.
Pubblicazione: (2026)
di: Zabidi, Muhammad Mun'im Ahmad, et al.
Pubblicazione: (2026)
Fine-Grained Quantitative Emotion Editing for Speech Generation
di: Inoue, Sho, et al.
Pubblicazione: (2024)
di: Inoue, Sho, et al.
Pubblicazione: (2024)
Spectrographic Portamento Gradient Analysis: A Quantitative Method for Historical Cello Recordings with Application to Beethoven's Piano and Cello Sonatas, 1930--2012
di: Sole, Ignasi
Pubblicazione: (2026)
di: Sole, Ignasi
Pubblicazione: (2026)
Multitask Learning with Capsule Networks for Speech-to-Intent Applications
di: Poncelet, Jakob, et al.
Pubblicazione: (2020)
di: Poncelet, Jakob, et al.
Pubblicazione: (2020)
Disentangled-Transformer: An Explainable End-to-End Automatic Speech Recognition Model with Speech Content-Context Separation
di: Wang, Pu, et al.
Pubblicazione: (2024)
di: Wang, Pu, et al.
Pubblicazione: (2024)
Towards Controllable Audio Texture Morphing
di: Gupta, Chitralekha, et al.
Pubblicazione: (2023)
di: Gupta, Chitralekha, et al.
Pubblicazione: (2023)
Unsupervised Online Continual Learning for Automatic Speech Recognition
di: Eeckt, Steven Vander, et al.
Pubblicazione: (2024)
di: Eeckt, Steven Vander, et al.
Pubblicazione: (2024)
Using Adapters to Overcome Catastrophic Forgetting in End-to-End Automatic Speech Recognition
di: Eeckt, Steven Vander, et al.
Pubblicazione: (2022)
di: Eeckt, Steven Vander, et al.
Pubblicazione: (2022)
Inverse-Hessian Regularization for Continual Learning in ASR
di: Eeckt, Steven Vander, et al.
Pubblicazione: (2026)
di: Eeckt, Steven Vander, et al.
Pubblicazione: (2026)
Passive acoustic non-line-of-sight localization without a relay surface
di: Sommer, Tal I., et al.
Pubblicazione: (2025)
di: Sommer, Tal I., et al.
Pubblicazione: (2025)
Audio Texture Manipulation by Exemplar-Based Analogy
di: Cheng, Kan Jen, et al.
Pubblicazione: (2025)
di: Cheng, Kan Jen, et al.
Pubblicazione: (2025)
Learning Control of Neural Sound Effects Synthesis from Physically Inspired Models
di: Zong, Yisu, et al.
Pubblicazione: (2025)
di: Zong, Yisu, et al.
Pubblicazione: (2025)
Assessing Data Replication in Symbolic Music via Adapted Structural Similarity Index Measure
di: Ji, Shulei, et al.
Pubblicazione: (2025)
di: Ji, Shulei, et al.
Pubblicazione: (2025)
Measuring Prosody Diversity in Zero-Shot TTS: A New Metric, Benchmark, and Exploration
di: Yang, Yifan, et al.
Pubblicazione: (2025)
di: Yang, Yifan, et al.
Pubblicazione: (2025)
Improving Neural Pitch Estimation with SWIPE Kernels
di: Marttila, David, et al.
Pubblicazione: (2025)
di: Marttila, David, et al.
Pubblicazione: (2025)
Learning Vocal-Tract Area and Radiation with a Physics-Informed Webster Model
di: Lu, Minhui, et al.
Pubblicazione: (2026)
di: Lu, Minhui, et al.
Pubblicazione: (2026)
Four Decades of Digital Waveguides
di: de Paula, Pablo Tablas, et al.
Pubblicazione: (2026)
di: de Paula, Pablo Tablas, et al.
Pubblicazione: (2026)
On feature representations for marmoset vocal communication analysis
di: Sarkar, Eklavya, et al.
Pubblicazione: (2025)
di: Sarkar, Eklavya, et al.
Pubblicazione: (2025)
Measuring Audio's Impact on Correctness: Audio-Contribution-Aware Post-Training of Large Audio Language Models
di: He, Haolin, et al.
Pubblicazione: (2025)
di: He, Haolin, et al.
Pubblicazione: (2025)
Example-Based Framework for Perceptually Guided Audio Texture Generation
di: Kamath, Purnima, et al.
Pubblicazione: (2023)
di: Kamath, Purnima, et al.
Pubblicazione: (2023)
Assessing speech quality metrics for evaluation of neural audio codecs under clean speech conditions
di: Mack, Wolfgang, et al.
Pubblicazione: (2025)
di: Mack, Wolfgang, et al.
Pubblicazione: (2025)
Objective Measurements of Voice Quality
di: Dhamyal, Hira, et al.
Pubblicazione: (2024)
di: Dhamyal, Hira, et al.
Pubblicazione: (2024)
Leveraging Broadcast Media Subtitle Transcripts for Automatic Speech Recognition and Subtitling
di: Poncelet, Jakob, et al.
Pubblicazione: (2025)
di: Poncelet, Jakob, et al.
Pubblicazione: (2025)
Unsupervised Accent Adaptation Through Masked Language Model Correction Of Discrete Self-Supervised Speech Units
di: Poncelet, Jakob, et al.
Pubblicazione: (2023)
di: Poncelet, Jakob, et al.
Pubblicazione: (2023)
Comparison of Self-Supervised Speech Pre-Training Methods on Flemish Dutch
di: Poncelet, Jakob, et al.
Pubblicazione: (2021)
di: Poncelet, Jakob, et al.
Pubblicazione: (2021)
Differentiable Black-box and Gray-box Modeling of Nonlinear Audio Effects
di: Comunità, Marco, et al.
Pubblicazione: (2025)
di: Comunità, Marco, et al.
Pubblicazione: (2025)
Weakly Supervised Phonological Features for Pathological Speech Analysis
di: Thienpondt, Jenthe, et al.
Pubblicazione: (2025)
di: Thienpondt, Jenthe, et al.
Pubblicazione: (2025)
XANE Background Acoustic Embeddings: Ablation and Clustering Analysis
di: Sharma, Dushyant, et al.
Pubblicazione: (2024)
di: Sharma, Dushyant, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Structural and Statistical Audio Texture Knowledge Distillation for Acoustic Classification
di: Ritu, Jarin, et al.
Pubblicazione: (2025) -
Neural Edge Histogram Descriptors for Underwater Acoustic Target Recognition
di: Agashe, Atharva, et al.
Pubblicazione: (2025) -
Cross-Domain Knowledge Transfer for Underwater Acoustic Classification Using Pre-trained Models
di: Mohammadi, Amirmohammad, et al.
Pubblicazione: (2024) -
Investigation of Time-Frequency Feature Combinations with Histogram Layer Time Delay Neural Networks
di: Mohammadi, Amirmohammad, et al.
Pubblicazione: (2024) -
Quantitative Analysis of Proxy Tasks for Anomalous Sound Detection
di: Shin, Seunghyeon, et al.
Pubblicazione: (2026)