LIWhiz: A Non-Intrusive Lyric Intelligibility Prediction System for the Cadenza Challenge
Fuente:
arXiv
Guardado en:
| Autores principales: | Shekar, Ram C. M. C., López-Espejo, Iván |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Non-Intrusive Intelligibility Prediction for Hearing Aids: Recent Advances, Trends, and Challenges
por: Zezario, Ryandhimas E.
Publicado: (2025)
por: Zezario, Ryandhimas E.
Publicado: (2025)
Feature Importance across Domains for Improving Non-Intrusive Speech Intelligibility Prediction in Hearing Aids
por: Zezario, Ryandhimas E., et al.
Publicado: (2025)
por: Zezario, Ryandhimas E., et al.
Publicado: (2025)
Non-Intrusive Binaural Speech Intelligibility Prediction Using Mamba for Hearing-Impaired Listeners
por: Yamamoto, Katsuhiko, et al.
Publicado: (2025)
por: Yamamoto, Katsuhiko, et al.
Publicado: (2025)
Frame-Aligned Fusion of Canary and WavLM for Non-Intrusive Intelligibility Prediction of Hearing-Aid-Processed Speech
por: Nakazawa, Kazushi
Publicado: (2026)
por: Nakazawa, Kazushi
Publicado: (2026)
A Study on Zero-Shot Non-Intrusive Speech Intelligibility for Hearing Aids Using Large Language Models
por: Zezario, Ryandhimas E., et al.
Publicado: (2025)
por: Zezario, Ryandhimas E., et al.
Publicado: (2025)
Non-Intrusive Speech Intelligibility Prediction for Hearing Aids using Whisper and Metadata
por: Zezario, Ryandhimas E., et al.
Publicado: (2023)
por: Zezario, Ryandhimas E., et al.
Publicado: (2023)
Leveraging Multiple Speech Enhancers for Non-Intrusive Intelligibility Prediction for Hearing-Impaired Listeners
por: Cao, Boxuan, et al.
Publicado: (2025)
por: Cao, Boxuan, et al.
Publicado: (2025)
ASASVIcomtech: The Vicomtech-UGR Speech Deepfake Detection and SASV Systems for the ASVspoof5 Challenge
por: Martín-Doñas, Juan M., et al.
Publicado: (2024)
por: Martín-Doñas, Juan M., et al.
Publicado: (2024)
Exploiting Music Source Separation for Automatic Lyrics Transcription with Whisper
por: Syed, Jaza, et al.
Publicado: (2025)
por: Syed, Jaza, et al.
Publicado: (2025)
Evaluating Speech Enhancement Systems Through Listening Effort
por: Gelderblom, Femke B., et al.
Publicado: (2024)
por: Gelderblom, Femke B., et al.
Publicado: (2024)
Non-Intrusive Speech Intelligibility Prediction for Hearing-Impaired Users using Intermediate ASR Features and Human Memory Models
por: Mogridge, Rhiannon, et al.
Publicado: (2024)
por: Mogridge, Rhiannon, et al.
Publicado: (2024)
Multimodal Lyrics-Rhythm Matching
por: Liao, Callie C., et al.
Publicado: (2023)
por: Liao, Callie C., et al.
Publicado: (2023)
Enhancing Lyrics Transcription on Music Mixtures with Consistency Loss
por: Huang, Jiawen, et al.
Publicado: (2025)
por: Huang, Jiawen, et al.
Publicado: (2025)
WhiSQA: Non-Intrusive Speech Quality Prediction Using Whisper Encoder Features
por: Close, George, et al.
Publicado: (2025)
por: Close, George, et al.
Publicado: (2025)
The first Cadenza challenges: using machine learning competitions to improve music for listeners with a hearing loss
por: Dabike, Gerardo Roa, et al.
Publicado: (2024)
por: Dabike, Gerardo Roa, et al.
Publicado: (2024)
Embedding-Based Intrusive Evaluation Metrics for Musical Source Separation Using MERT Representations
por: Bereuter, Paul A., et al.
Publicado: (2026)
por: Bereuter, Paul A., et al.
Publicado: (2026)
Music Enhancement with Deep Filters: A Technical Report for The ICASSP 2024 Cadenza Challenge
por: Shao, Keren, et al.
Publicado: (2024)
por: Shao, Keren, et al.
Publicado: (2024)
Relationships between Keywords and Strong Beats in Lyrical Music
por: Liao, Callie C., et al.
Publicado: (2024)
por: Liao, Callie C., et al.
Publicado: (2024)
LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT
por: Zhuo, Le, et al.
Publicado: (2023)
por: Zhuo, Le, et al.
Publicado: (2023)
YingMusic-Singer-Plus: Controllable Singing Voice Synthesis with Flexible Lyric Manipulation and Annotation-free Melody Guidance
por: Hao, Chunbo, et al.
Publicado: (2026)
por: Hao, Chunbo, et al.
Publicado: (2026)
Multi-Task Pseudo-Label Learning for Non-Intrusive Speech Quality Assessment Model
por: Zezario, Ryandhimas E., et al.
Publicado: (2023)
por: Zezario, Ryandhimas E., et al.
Publicado: (2023)
EventTrojan: Manipulating Non-Intrusive Speech Quality Assessment via Imperceptible Events
por: Ren, Ying, et al.
Publicado: (2023)
por: Ren, Ying, et al.
Publicado: (2023)
A Computational Analysis of Lyric Similarity Perception
por: Kim, Haven, et al.
Publicado: (2024)
por: Kim, Haven, et al.
Publicado: (2024)
MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence
por: Kumar, Sonal, et al.
Publicado: (2025)
por: Kumar, Sonal, et al.
Publicado: (2025)
REFFLY: Melody-Constrained Lyrics Editing Model
por: Zhao, Songyan, et al.
Publicado: (2024)
por: Zhao, Songyan, et al.
Publicado: (2024)
Deep Learning-based Non-Intrusive Multi-Objective Speech Assessment Model with Cross-Domain Features
por: Zezario, Ryandhimas E., et al.
Publicado: (2021)
por: Zezario, Ryandhimas E., et al.
Publicado: (2021)
The T12 System for AudioMOS Challenge 2025: Audio Aesthetics Score Prediction System Using KAN- and VERSA-based Models
por: Yamamoto, Katsuhiko, et al.
Publicado: (2025)
por: Yamamoto, Katsuhiko, et al.
Publicado: (2025)
A Real-Time Lyrics Alignment System Using Chroma And Phonetic Features For Classical Vocal Performance
por: Park, Jiyun, et al.
Publicado: (2024)
por: Park, Jiyun, et al.
Publicado: (2024)
SongCreator: Lyrics-based Universal Song Generation
por: Lei, Shun, et al.
Publicado: (2024)
por: Lei, Shun, et al.
Publicado: (2024)
XMUspeech Systems for the ASVspoof 5 Challenge
por: Li, Wangjie, et al.
Publicado: (2025)
por: Li, Wangjie, et al.
Publicado: (2025)
BUT Systems and Analyses for the ASVspoof 5 Challenge
por: Rohdin, Johan, et al.
Publicado: (2024)
por: Rohdin, Johan, et al.
Publicado: (2024)
The VoiceMOS Challenge 2024: Beyond Speech Quality Prediction
por: Huang, Wen-Chin, et al.
Publicado: (2024)
por: Huang, Wen-Chin, et al.
Publicado: (2024)
Joint Learning of Wording and Formatting for Singable Melody-to-Lyric Generation
por: Ou, Longshen, et al.
Publicado: (2023)
por: Ou, Longshen, et al.
Publicado: (2023)
The USTC-NERCSLIP Systems for The ICMC-ASR Challenge
por: Wu, Minghui, et al.
Publicado: (2024)
por: Wu, Minghui, et al.
Publicado: (2024)
STCON System for the CHiME-8 Challenge
por: Mitrofanov, Anton, et al.
Publicado: (2024)
por: Mitrofanov, Anton, et al.
Publicado: (2024)
Towards Building an End-to-End Multilingual Automatic Lyrics Transcription Model
por: Huang, Jiawen, et al.
Publicado: (2024)
por: Huang, Jiawen, et al.
Publicado: (2024)
GESI: Gammachirp Envelope Similarity Index for Predicting Intelligibility of Simulated Hearing Loss Sounds
por: Yamamoto, Ayako, et al.
Publicado: (2023)
por: Yamamoto, Ayako, et al.
Publicado: (2023)
CUHK-EE Systems for the vTAD Challenge at NCMMSC 2025
por: Chiu, Aemon Yat Fei, et al.
Publicado: (2025)
por: Chiu, Aemon Yat Fei, et al.
Publicado: (2025)
The SJTU X-LANCE Lab System for MSR Challenge 2025
por: Zhu, Jinxuan, et al.
Publicado: (2026)
por: Zhu, Jinxuan, et al.
Publicado: (2026)
The USTC-NERCSLIP Systems for the CHiME-8 MMCSG Challenge
por: Jiang, Ya, et al.
Publicado: (2024)
por: Jiang, Ya, et al.
Publicado: (2024)
Ejemplares similares
-
Non-Intrusive Intelligibility Prediction for Hearing Aids: Recent Advances, Trends, and Challenges
por: Zezario, Ryandhimas E.
Publicado: (2025) -
Feature Importance across Domains for Improving Non-Intrusive Speech Intelligibility Prediction in Hearing Aids
por: Zezario, Ryandhimas E., et al.
Publicado: (2025) -
Non-Intrusive Binaural Speech Intelligibility Prediction Using Mamba for Hearing-Impaired Listeners
por: Yamamoto, Katsuhiko, et al.
Publicado: (2025) -
Frame-Aligned Fusion of Canary and WavLM for Non-Intrusive Intelligibility Prediction of Hearing-Aid-Processed Speech
por: Nakazawa, Kazushi
Publicado: (2026) -
A Study on Zero-Shot Non-Intrusive Speech Intelligibility for Hearing Aids Using Large Language Models
por: Zezario, Ryandhimas E., et al.
Publicado: (2025)