LISTEN: Lightweight Industrial Sound-representable Transformer for Edge Notification
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Han, Changheon, Kang, Yun Seok, Sim, Yuseop, Park, Hyung Wook, Jun, Martin Byung-Guk |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
IMPACT: Industrial Machine Perception via Acoustic Cognitive Transformer
par: Han, Changheon, et autres
Publié: (2025)
par: Han, Changheon, et autres
Publié: (2025)
Adversarial Domain Adaptation for Metal Cutting Sound Detection: Leveraging Abundant Lab Data for Scarce Industry Data
par: Mostafiz, Mir Imtiaz, et autres
Publié: (2024)
par: Mostafiz, Mir Imtiaz, et autres
Publié: (2024)
Track Role Prediction of Single-Instrumental Sequences
par: Han, Changheon, et autres
Publié: (2024)
par: Han, Changheon, et autres
Publié: (2024)
CNN-based Robust Sound Source Localization with SRP-PHAT for the Extreme Edge
par: Yin, Jun, et autres
Publié: (2025)
par: Yin, Jun, et autres
Publié: (2025)
Reverse Attention for Lightweight Speech Enhancement on Edge Devices
par: Ojha, Shuubham, et autres
Publié: (2025)
par: Ojha, Shuubham, et autres
Publié: (2025)
JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection
par: Nam, Hyeonuk, et autres
Publié: (2025)
par: Nam, Hyeonuk, et autres
Publié: (2025)
Listen through the Sound: Generative Speech Restoration Leveraging Acoustic Context Representation
par: Chung, Soo-Whan, et autres
Publié: (2025)
par: Chung, Soo-Whan, et autres
Publié: (2025)
Towards Understanding of Frequency Dependence on Sound Event Detection
par: Nam, Hyeonuk, et autres
Publié: (2025)
par: Nam, Hyeonuk, et autres
Publié: (2025)
Statistical Beamformer Exploiting Non-stationarity and Sparsity with Spatially Constrained ICA for Robust Speech Recognition
par: Shin, Ui-Hyeop, et autres
Publié: (2023)
par: Shin, Ui-Hyeop, et autres
Publié: (2023)
Asymmetric Encoder-Decoder Based on Time-Frequency Correlation for Speech Separation
par: Shin, Ui-Hyeop, et autres
Publié: (2026)
par: Shin, Ui-Hyeop, et autres
Publié: (2026)
HILCodec: High-Fidelity and Lightweight Neural Audio Codec
par: Ahn, Sunghwan, et autres
Publié: (2024)
par: Ahn, Sunghwan, et autres
Publié: (2024)
Study of Lightweight Transformer Architectures for Single-Channel Speech Enhancement
par: Zhao, Haixin, et autres
Publié: (2025)
par: Zhao, Haixin, et autres
Publié: (2025)
Vision Transformer Segmentation for Visual Bird Sound Denoising
par: Kumar, Sahil, et autres
Publié: (2024)
par: Kumar, Sahil, et autres
Publié: (2024)
Learning to Solve Inverse Problems for Perceptual Sound Matching
par: Han, Han, et autres
Publié: (2023)
par: Han, Han, et autres
Publié: (2023)
Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection
par: Nam, Hyeonuk, et autres
Publié: (2025)
par: Nam, Hyeonuk, et autres
Publié: (2025)
Effective Pre-Training of Audio Transformers for Sound Event Detection
par: Schmid, Florian, et autres
Publié: (2024)
par: Schmid, Florian, et autres
Publié: (2024)
NeXt-TDNN: Modernizing Multi-Scale Temporal Convolution Backbone for Speaker Verification
par: Heo, Hyun-Jun, et autres
Publié: (2023)
par: Heo, Hyun-Jun, et autres
Publié: (2023)
TF-CorrNet: Leveraging Spatial Correlation for Continuous Speech Separation
par: Shin, Ui-Hyeop, et autres
Publié: (2025)
par: Shin, Ui-Hyeop, et autres
Publié: (2025)
Pushing the Limit of Sound Event Detection with Multi-Dilated Frequency Dynamic Convolution
par: Nam, Hyeonuk, et autres
Publié: (2024)
par: Nam, Hyeonuk, et autres
Publié: (2024)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
par: Chen, Yuanjian, et autres
Publié: (2025)
par: Chen, Yuanjian, et autres
Publié: (2025)
ELF: Encoding Speaker-Specific Latent Speech Feature for Speech Synthesis
par: Kong, Jungil, et autres
Publié: (2023)
par: Kong, Jungil, et autres
Publié: (2023)
Ensemble Confidence Calibration for Sound Event Detection in Open-environment
par: Chen, Yuanjian, et autres
Publié: (2025)
par: Chen, Yuanjian, et autres
Publié: (2025)
SoloAudio: Target Sound Extraction with Language-oriented Audio Diffusion Transformer
par: Wang, Helin, et autres
Publié: (2024)
par: Wang, Helin, et autres
Publié: (2024)
SoundLoCD: An Efficient Conditional Discrete Contrastive Latent Diffusion Model for Text-to-Sound Generation
par: Niu, Xinlei, et autres
Publié: (2024)
par: Niu, Xinlei, et autres
Publié: (2024)
Enhancing Stereo Sound Event Detection with BiMamba and Pretrained PSELDnet
par: Gao, Wenmiao, et autres
Publié: (2025)
par: Gao, Wenmiao, et autres
Publié: (2025)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
par: Lee, Gyeong-Tae, et autres
Publié: (2025)
par: Lee, Gyeong-Tae, et autres
Publié: (2025)
PSELDNets: Pre-trained Neural Networks on a Large-scale Synthetic Dataset for Sound Event Localization and Detection
par: Hu, Jinbo, et autres
Publié: (2024)
par: Hu, Jinbo, et autres
Publié: (2024)
Inverse Nonlinearity Compensation of Hyperelastic Deformation in Dielectric Elastomer for Acoustic Actuation
par: Lee, Jin Woo, et autres
Publié: (2024)
par: Lee, Jin Woo, et autres
Publié: (2024)
Evaluating CNN with Stacked Feature Representations and Audio Spectrogram Transformer Models for Sound Classification
par: Dehaghania, Parinaz Binandeh, et autres
Publié: (2026)
par: Dehaghania, Parinaz Binandeh, et autres
Publié: (2026)
Improving Audio Spectrogram Transformers for Sound Event Detection Through Multi-Stage Training
par: Schmid, Florian, et autres
Publié: (2024)
par: Schmid, Florian, et autres
Publié: (2024)
DroFiT: A Lightweight Band-fused Frequency Attention Toward Real-time UAV Speech Enhancement
par: Lee, Jeongmin, et autres
Publié: (2025)
par: Lee, Jeongmin, et autres
Publié: (2025)
Re-Parameterization of Lightweight Transformer for On-Device Speech Emotion Recognition
par: Zhang, Zixing, et autres
Publié: (2024)
par: Zhang, Zixing, et autres
Publié: (2024)
Diversifying and Expanding Frequency-Adaptive Convolution Kernels for Sound Event Detection
par: Nam, Hyeonuk, et autres
Publié: (2024)
par: Nam, Hyeonuk, et autres
Publié: (2024)
Hierarchical Pooling Structure for Weakly Labeled Sound Event Detection
par: He, Ke-Xin, et autres
Publié: (2019)
par: He, Ke-Xin, et autres
Publié: (2019)
JELLY: Joint Emotion Recognition and Context Reasoning with LLMs for Conversational Speech Synthesis
par: Cha, Jun-Hyeok, et autres
Publié: (2025)
par: Cha, Jun-Hyeok, et autres
Publié: (2025)
The Sound Demixing Challenge 2023 $\unicode{x2013}$ Music Demixing Track
par: Fabbro, Giorgio, et autres
Publié: (2023)
par: Fabbro, Giorgio, et autres
Publié: (2023)
Leveraging Sound Source Trajectories for Universal Sound Separation
par: Wu, Donghang, et autres
Publié: (2024)
par: Wu, Donghang, et autres
Publié: (2024)
Sound Zone Control Robust To Sound Speed Change
par: Bhattacharjee, Sankha Subhra, et autres
Publié: (2024)
par: Bhattacharjee, Sankha Subhra, et autres
Publié: (2024)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
par: Yin, Han, et autres
Publié: (2024)
par: Yin, Han, et autres
Publié: (2024)
ENACT-Heart -- ENsemble-based Assessment Using CNN and Transformer on Heart Sounds
par: Han, Jiho, et autres
Publié: (2025)
par: Han, Jiho, et autres
Publié: (2025)
Documents similaires
-
IMPACT: Industrial Machine Perception via Acoustic Cognitive Transformer
par: Han, Changheon, et autres
Publié: (2025) -
Adversarial Domain Adaptation for Metal Cutting Sound Detection: Leveraging Abundant Lab Data for Scarce Industry Data
par: Mostafiz, Mir Imtiaz, et autres
Publié: (2024) -
Track Role Prediction of Single-Instrumental Sequences
par: Han, Changheon, et autres
Publié: (2024) -
CNN-based Robust Sound Source Localization with SRP-PHAT for the Extreme Edge
par: Yin, Jun, et autres
Publié: (2025) -
Reverse Attention for Lightweight Speech Enhancement on Edge Devices
par: Ojha, Shuubham, et autres
Publié: (2025)