Saved in:
| Main Authors: | Ritu, Jarin, Van Dine, Alexandra, Peeples, Joshua |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2504.14843 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Structural and Statistical Audio Texture Knowledge Distillation for Acoustic Classification
by: Ritu, Jarin, et al.
Published: (2025)
by: Ritu, Jarin, et al.
Published: (2025)
Neural Edge Histogram Descriptors for Underwater Acoustic Target Recognition
by: Agashe, Atharva, et al.
Published: (2025)
by: Agashe, Atharva, et al.
Published: (2025)
Cross-Domain Knowledge Transfer for Underwater Acoustic Classification Using Pre-trained Models
by: Mohammadi, Amirmohammad, et al.
Published: (2024)
by: Mohammadi, Amirmohammad, et al.
Published: (2024)
Investigation of Time-Frequency Feature Combinations with Histogram Layer Time Delay Neural Networks
by: Mohammadi, Amirmohammad, et al.
Published: (2024)
by: Mohammadi, Amirmohammad, et al.
Published: (2024)
Quantitative Analysis of Proxy Tasks for Anomalous Sound Detection
by: Shin, Seunghyeon, et al.
Published: (2026)
by: Shin, Seunghyeon, et al.
Published: (2026)
A Statistics-Driven Differentiable Approach for Sound Texture Synthesis and Analysis
by: Gutiérrez, Esteban, et al.
Published: (2025)
by: Gutiérrez, Esteban, et al.
Published: (2025)
Towards a Quantitative Analysis of Coarticulation with a Phoneme-to-Articulatory Model
by: Fan, Chaofei, et al.
Published: (2024)
by: Fan, Chaofei, et al.
Published: (2024)
Raw Audio Classification with Cosine Convolutional Neural Network (CosCovNN)
by: Haque, Kazi Nazmul, et al.
Published: (2024)
by: Haque, Kazi Nazmul, et al.
Published: (2024)
HiRIS: an Airborne Sonar Sensor with a 1024 Channel Microphone Array for In-Air Acoustic Imaging
by: Laurijssen, Dennis, et al.
Published: (2024)
by: Laurijssen, Dennis, et al.
Published: (2024)
Multiple Mobile Target Detection and Tracking in Active Sonar Array Using a Track-Before-Detect Approach
by: Abu, Avi, et al.
Published: (2024)
by: Abu, Avi, et al.
Published: (2024)
Subspace Track-before-Detect for Passive Multi-Target Tracking with Unknown Emitted Signals
by: Ito, Nobutaka, et al.
Published: (2026)
by: Ito, Nobutaka, et al.
Published: (2026)
METEOR: Melody-aware Texture-controllable Symbolic Orchestral Music Generation via Transformer VAE
by: Le, Dinh-Viet-Toan, et al.
Published: (2024)
by: Le, Dinh-Viet-Toan, et al.
Published: (2024)
SEABAD: A Tropical Bird Activity Detection Dataset for Passive Acoustic Monitoring
by: Zabidi, Muhammad Mun'im Ahmad, et al.
Published: (2026)
by: Zabidi, Muhammad Mun'im Ahmad, et al.
Published: (2026)
Fine-Grained Quantitative Emotion Editing for Speech Generation
by: Inoue, Sho, et al.
Published: (2024)
by: Inoue, Sho, et al.
Published: (2024)
Spectrographic Portamento Gradient Analysis: A Quantitative Method for Historical Cello Recordings with Application to Beethoven's Piano and Cello Sonatas, 1930--2012
by: Sole, Ignasi
Published: (2026)
by: Sole, Ignasi
Published: (2026)
Multitask Learning with Capsule Networks for Speech-to-Intent Applications
by: Poncelet, Jakob, et al.
Published: (2020)
by: Poncelet, Jakob, et al.
Published: (2020)
Disentangled-Transformer: An Explainable End-to-End Automatic Speech Recognition Model with Speech Content-Context Separation
by: Wang, Pu, et al.
Published: (2024)
by: Wang, Pu, et al.
Published: (2024)
Towards Controllable Audio Texture Morphing
by: Gupta, Chitralekha, et al.
Published: (2023)
by: Gupta, Chitralekha, et al.
Published: (2023)
Unsupervised Online Continual Learning for Automatic Speech Recognition
by: Eeckt, Steven Vander, et al.
Published: (2024)
by: Eeckt, Steven Vander, et al.
Published: (2024)
Using Adapters to Overcome Catastrophic Forgetting in End-to-End Automatic Speech Recognition
by: Eeckt, Steven Vander, et al.
Published: (2022)
by: Eeckt, Steven Vander, et al.
Published: (2022)
Inverse-Hessian Regularization for Continual Learning in ASR
by: Eeckt, Steven Vander, et al.
Published: (2026)
by: Eeckt, Steven Vander, et al.
Published: (2026)
Passive acoustic non-line-of-sight localization without a relay surface
by: Sommer, Tal I., et al.
Published: (2025)
by: Sommer, Tal I., et al.
Published: (2025)
Audio Texture Manipulation by Exemplar-Based Analogy
by: Cheng, Kan Jen, et al.
Published: (2025)
by: Cheng, Kan Jen, et al.
Published: (2025)
Learning Control of Neural Sound Effects Synthesis from Physically Inspired Models
by: Zong, Yisu, et al.
Published: (2025)
by: Zong, Yisu, et al.
Published: (2025)
Assessing Data Replication in Symbolic Music via Adapted Structural Similarity Index Measure
by: Ji, Shulei, et al.
Published: (2025)
by: Ji, Shulei, et al.
Published: (2025)
Measuring Prosody Diversity in Zero-Shot TTS: A New Metric, Benchmark, and Exploration
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
Improving Neural Pitch Estimation with SWIPE Kernels
by: Marttila, David, et al.
Published: (2025)
by: Marttila, David, et al.
Published: (2025)
Learning Vocal-Tract Area and Radiation with a Physics-Informed Webster Model
by: Lu, Minhui, et al.
Published: (2026)
by: Lu, Minhui, et al.
Published: (2026)
Four Decades of Digital Waveguides
by: de Paula, Pablo Tablas, et al.
Published: (2026)
by: de Paula, Pablo Tablas, et al.
Published: (2026)
On feature representations for marmoset vocal communication analysis
by: Sarkar, Eklavya, et al.
Published: (2025)
by: Sarkar, Eklavya, et al.
Published: (2025)
Measuring Audio's Impact on Correctness: Audio-Contribution-Aware Post-Training of Large Audio Language Models
by: He, Haolin, et al.
Published: (2025)
by: He, Haolin, et al.
Published: (2025)
Example-Based Framework for Perceptually Guided Audio Texture Generation
by: Kamath, Purnima, et al.
Published: (2023)
by: Kamath, Purnima, et al.
Published: (2023)
Assessing speech quality metrics for evaluation of neural audio codecs under clean speech conditions
by: Mack, Wolfgang, et al.
Published: (2025)
by: Mack, Wolfgang, et al.
Published: (2025)
Objective Measurements of Voice Quality
by: Dhamyal, Hira, et al.
Published: (2024)
by: Dhamyal, Hira, et al.
Published: (2024)
Leveraging Broadcast Media Subtitle Transcripts for Automatic Speech Recognition and Subtitling
by: Poncelet, Jakob, et al.
Published: (2025)
by: Poncelet, Jakob, et al.
Published: (2025)
Unsupervised Accent Adaptation Through Masked Language Model Correction Of Discrete Self-Supervised Speech Units
by: Poncelet, Jakob, et al.
Published: (2023)
by: Poncelet, Jakob, et al.
Published: (2023)
Comparison of Self-Supervised Speech Pre-Training Methods on Flemish Dutch
by: Poncelet, Jakob, et al.
Published: (2021)
by: Poncelet, Jakob, et al.
Published: (2021)
Differentiable Black-box and Gray-box Modeling of Nonlinear Audio Effects
by: Comunità, Marco, et al.
Published: (2025)
by: Comunità, Marco, et al.
Published: (2025)
Weakly Supervised Phonological Features for Pathological Speech Analysis
by: Thienpondt, Jenthe, et al.
Published: (2025)
by: Thienpondt, Jenthe, et al.
Published: (2025)
XANE Background Acoustic Embeddings: Ablation and Clustering Analysis
by: Sharma, Dushyant, et al.
Published: (2024)
by: Sharma, Dushyant, et al.
Published: (2024)
Similar Items
-
Structural and Statistical Audio Texture Knowledge Distillation for Acoustic Classification
by: Ritu, Jarin, et al.
Published: (2025) -
Neural Edge Histogram Descriptors for Underwater Acoustic Target Recognition
by: Agashe, Atharva, et al.
Published: (2025) -
Cross-Domain Knowledge Transfer for Underwater Acoustic Classification Using Pre-trained Models
by: Mohammadi, Amirmohammad, et al.
Published: (2024) -
Investigation of Time-Frequency Feature Combinations with Histogram Layer Time Delay Neural Networks
by: Mohammadi, Amirmohammad, et al.
Published: (2024) -
Quantitative Analysis of Proxy Tasks for Anomalous Sound Detection
by: Shin, Seunghyeon, et al.
Published: (2026)