Learning to Discover: A Generalized Framework for Raga Identification without Forgetting
Fuente:
arXiv
Saved in:
| Main Authors: | Singh, Parampreet, Kumar, Somya, Nitawe, Chaitanya Shailendra, Arora, Vipul |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Explainable Deep Learning Analysis for Raga Identification in Indian Art Music
by: Singh, Parampreet, et al.
Published: (2024)
by: Singh, Parampreet, et al.
Published: (2024)
Identification and Clustering of Unseen Ragas in Indian Art Music
by: Singh, Parampreet, et al.
Published: (2024)
by: Singh, Parampreet, et al.
Published: (2024)
Automatic Detection and Analysis of Singing Mistakes for Music Pedagogy
by: Kumar, Sumit, et al.
Published: (2026)
by: Kumar, Sumit, et al.
Published: (2026)
Recognizing Ornaments in Vocal Indian Art Music with Active Annotation
by: Kumar, Sumit, et al.
Published: (2025)
by: Kumar, Sumit, et al.
Published: (2025)
Learning from Limited Labels: Transductive Graph Label Propagation for Indian Music Analysis
by: Singh, Parampreet, et al.
Published: (2026)
by: Singh, Parampreet, et al.
Published: (2026)
Improving Active Learning for Melody Estimation by Disentangling Uncertainties
by: Jaiswal, Aayush, et al.
Published: (2025)
by: Jaiswal, Aayush, et al.
Published: (2025)
TeLeS: Temporal Lexeme Similarity Score to Estimate Confidence in End-to-End ASR
by: Ravi, Nagarathna, et al.
Published: (2024)
by: Ravi, Nagarathna, et al.
Published: (2024)
$T\bar{a}laGen:$ A System for Automatic $T\bar{a}la$ Identification and Generation
by: Kodag, Rahul Bapusaheb, et al.
Published: (2024)
by: Kodag, Rahul Bapusaheb, et al.
Published: (2024)
Carnatic Raga Identification System using Rigorous Time-Delay Neural Network
by: Natesan, Sanjay, et al.
Published: (2024)
by: Natesan, Sanjay, et al.
Published: (2024)
Discovering and Steering Interpretable Concepts in Large Generative Music Models
by: Singh, Nikhil, et al.
Published: (2025)
by: Singh, Nikhil, et al.
Published: (2025)
Weakly Supervised Tabla Stroke Transcription via TI-SDRM: A Rhythm-Aware Lattice Rescoring Framework
by: Kodag, Rahul Bapusaheb, et al.
Published: (2026)
by: Kodag, Rahul Bapusaheb, et al.
Published: (2026)
BEST-STD2.0: Balanced and Efficient Speech Tokenizer for Spoken Term Detection
by: Singh, Anup, et al.
Published: (2025)
by: Singh, Anup, et al.
Published: (2025)
AudioNet: Supervised Deep Hashing for Retrieval of Similar Audio Events
by: Dutta, Sagar, et al.
Published: (2025)
by: Dutta, Sagar, et al.
Published: (2025)
Bayesian Parameter-Efficient Fine-Tuning for Overcoming Catastrophic Forgetting
by: Chen, Haolin, et al.
Published: (2024)
by: Chen, Haolin, et al.
Published: (2024)
Attention-Based Audio Embeddings for Query-by-Example
by: Singh, Anup, et al.
Published: (2022)
by: Singh, Anup, et al.
Published: (2022)
H-QuEST: Accelerating Query-by-Example Spoken Term Detection with Hierarchical Indexing
by: Singh, Akanksha, et al.
Published: (2025)
by: Singh, Akanksha, et al.
Published: (2025)
Uncertainty Quantification in Melody Estimation using Histogram Representation
by: Saxena, Kavya Ranjan, et al.
Published: (2025)
by: Saxena, Kavya Ranjan, et al.
Published: (2025)
Meta-learning-based percussion transcription and $t\bar{a}la$ identification from low-resource audio
by: Kodag, Rahul Bapusaheb, et al.
Published: (2025)
by: Kodag, Rahul Bapusaheb, et al.
Published: (2025)
SyncNet: correlating objective for time delay estimation in audio signals
by: Raina, Akshay, et al.
Published: (2022)
by: Raina, Akshay, et al.
Published: (2022)
Interactive singing melody extraction based on active adaptation
by: Saxena, Kavya Ranjan, et al.
Published: (2024)
by: Saxena, Kavya Ranjan, et al.
Published: (2024)
DeCoR: Defy Knowledge Forgetting by Predicting Earlier Audio Codes
by: Jiang, Xilin, et al.
Published: (2023)
by: Jiang, Xilin, et al.
Published: (2023)
Plug-in Losses for Evidential Deep Learning: A Simplified Framework for Uncertainty Estimation that Includes the Softmax Classifier
by: Hayta, Berk, et al.
Published: (2026)
by: Hayta, Berk, et al.
Published: (2026)
Literary and Colloquial Dialect Identification for Tamil using Acoustic Features
by: Nanmalar, M., et al.
Published: (2024)
by: Nanmalar, M., et al.
Published: (2024)
Exploring speech style spaces with language models: Emotional TTS without emotion labels
by: Chandra, Shreeram Suresh, et al.
Published: (2024)
by: Chandra, Shreeram Suresh, et al.
Published: (2024)
SpeechOp: Inference-Time Task Composition for Generative Speech Processing
by: Lovelace, Justin, et al.
Published: (2025)
by: Lovelace, Justin, et al.
Published: (2025)
EnvId: A Metric Learning Approach for Forensic Few-Shot Identification of Unseen Environments
by: Moussa, Denise, et al.
Published: (2024)
by: Moussa, Denise, et al.
Published: (2024)
BenSParX: A Robust Explainable Machine Learning Framework for Parkinson's Disease Detection from Bengali Conversational Speech
by: Hossain, Riad, et al.
Published: (2025)
by: Hossain, Riad, et al.
Published: (2025)
Contrastive Learning from Synthetic Audio Doppelgängers
by: Cherep, Manuel, et al.
Published: (2024)
by: Cherep, Manuel, et al.
Published: (2024)
A Differentiable Alignment Framework for Sequence-to-Sequence Modeling via Optimal Transport
by: Kaloga, Yacouba, et al.
Published: (2025)
by: Kaloga, Yacouba, et al.
Published: (2025)
Benchmarking Automatic Speech Recognition coupled LLM Modules for Medical Diagnostics
by: Kumar, Kabir
Published: (2025)
by: Kumar, Kabir
Published: (2025)
Creative Text-to-Audio Generation via Synthesizer Programming
by: Cherep, Manuel, et al.
Published: (2024)
by: Cherep, Manuel, et al.
Published: (2024)
Identifying birdsong syllables without labelled data
by: Teng, Mélisande, et al.
Published: (2025)
by: Teng, Mélisande, et al.
Published: (2025)
Scaling NVIDIA's Multi-speaker Multi-lingual TTS Systems with Zero-Shot TTS to Indic Languages
by: Arora, Akshit, et al.
Published: (2024)
by: Arora, Akshit, et al.
Published: (2024)
Learn and Don't Forget: Adding a New Language to ASR Foundation Models
by: Qian, Mengjie, et al.
Published: (2024)
by: Qian, Mengjie, et al.
Published: (2024)
SA-SSL-MOS: Self-supervised Learning MOS Prediction with Spectral Augmentation for Generalized Multi-Rate Speech Assessment
by: Cao, Fengyuan, et al.
Published: (2026)
by: Cao, Fengyuan, et al.
Published: (2026)
Beat this! Accurate beat tracking without DBN postprocessing
by: Foscarin, Francesco, et al.
Published: (2024)
by: Foscarin, Francesco, et al.
Published: (2024)
BEST-STD: Bidirectional Mamba-Enhanced Speech Tokenization for Spoken Term Detection
by: Singh, Anup, et al.
Published: (2024)
by: Singh, Anup, et al.
Published: (2024)
ASTRA: Aligning Speech and Text Representations for Asr without Sampling
by: Gaur, Neeraj, et al.
Published: (2024)
by: Gaur, Neeraj, et al.
Published: (2024)
A Novel Score-CAM based Denoiser for Spectrographic Signature Extraction without Ground Truth
by: Elias, Noel
Published: (2024)
by: Elias, Noel
Published: (2024)
Speech Understanding on Tiny Devices with A Learning Cache
by: Benazir, Afsara, et al.
Published: (2023)
by: Benazir, Afsara, et al.
Published: (2023)
Similar Items
-
Explainable Deep Learning Analysis for Raga Identification in Indian Art Music
by: Singh, Parampreet, et al.
Published: (2024) -
Identification and Clustering of Unseen Ragas in Indian Art Music
by: Singh, Parampreet, et al.
Published: (2024) -
Automatic Detection and Analysis of Singing Mistakes for Music Pedagogy
by: Kumar, Sumit, et al.
Published: (2026) -
Recognizing Ornaments in Vocal Indian Art Music with Active Annotation
by: Kumar, Sumit, et al.
Published: (2025) -
Learning from Limited Labels: Transductive Graph Label Propagation for Indian Music Analysis
by: Singh, Parampreet, et al.
Published: (2026)