$T\bar{a}laGen:$ A System for Automatic $T\bar{a}la$ Identification and Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kodag, Rahul Bapusaheb, Jindal, Himanshu, Arora, Vipul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Meta-learning-based percussion transcription and $t\bar{a}la$ identification from low-resource audio
von: Kodag, Rahul Bapusaheb, et al.
Veröffentlicht: (2025)
von: Kodag, Rahul Bapusaheb, et al.
Veröffentlicht: (2025)
Weakly Supervised Tabla Stroke Transcription via TI-SDRM: A Rhythm-Aware Lattice Rescoring Framework
von: Kodag, Rahul Bapusaheb, et al.
Veröffentlicht: (2026)
von: Kodag, Rahul Bapusaheb, et al.
Veröffentlicht: (2026)
Identification and Clustering of Unseen Ragas in Indian Art Music
von: Singh, Parampreet, et al.
Veröffentlicht: (2024)
von: Singh, Parampreet, et al.
Veröffentlicht: (2024)
Explainable Deep Learning Analysis for Raga Identification in Indian Art Music
von: Singh, Parampreet, et al.
Veröffentlicht: (2024)
von: Singh, Parampreet, et al.
Veröffentlicht: (2024)
AudioNet: Supervised Deep Hashing for Retrieval of Similar Audio Events
von: Dutta, Sagar, et al.
Veröffentlicht: (2025)
von: Dutta, Sagar, et al.
Veröffentlicht: (2025)
Learning to Discover: A Generalized Framework for Raga Identification without Forgetting
von: Singh, Parampreet, et al.
Veröffentlicht: (2026)
von: Singh, Parampreet, et al.
Veröffentlicht: (2026)
Uncertainty Quantification in Melody Estimation using Histogram Representation
von: Saxena, Kavya Ranjan, et al.
Veröffentlicht: (2025)
von: Saxena, Kavya Ranjan, et al.
Veröffentlicht: (2025)
SyncNet: correlating objective for time delay estimation in audio signals
von: Raina, Akshay, et al.
Veröffentlicht: (2022)
von: Raina, Akshay, et al.
Veröffentlicht: (2022)
Automatic Detection and Analysis of Singing Mistakes for Music Pedagogy
von: Kumar, Sumit, et al.
Veröffentlicht: (2026)
von: Kumar, Sumit, et al.
Veröffentlicht: (2026)
BEST-STD2.0: Balanced and Efficient Speech Tokenizer for Spoken Term Detection
von: Singh, Anup, et al.
Veröffentlicht: (2025)
von: Singh, Anup, et al.
Veröffentlicht: (2025)
Improving Active Learning for Melody Estimation by Disentangling Uncertainties
von: Jaiswal, Aayush, et al.
Veröffentlicht: (2025)
von: Jaiswal, Aayush, et al.
Veröffentlicht: (2025)
Interactive singing melody extraction based on active adaptation
von: Saxena, Kavya Ranjan, et al.
Veröffentlicht: (2024)
von: Saxena, Kavya Ranjan, et al.
Veröffentlicht: (2024)
H-QuEST: Accelerating Query-by-Example Spoken Term Detection with Hierarchical Indexing
von: Singh, Akanksha, et al.
Veröffentlicht: (2025)
von: Singh, Akanksha, et al.
Veröffentlicht: (2025)
Attention-Based Audio Embeddings for Query-by-Example
von: Singh, Anup, et al.
Veröffentlicht: (2022)
von: Singh, Anup, et al.
Veröffentlicht: (2022)
TeLeS: Temporal Lexeme Similarity Score to Estimate Confidence in End-to-End ASR
von: Ravi, Nagarathna, et al.
Veröffentlicht: (2024)
von: Ravi, Nagarathna, et al.
Veröffentlicht: (2024)
Learning from Limited Labels: Transductive Graph Label Propagation for Indian Music Analysis
von: Singh, Parampreet, et al.
Veröffentlicht: (2026)
von: Singh, Parampreet, et al.
Veröffentlicht: (2026)
BEST-STD: Bidirectional Mamba-Enhanced Speech Tokenization for Spoken Term Detection
von: Singh, Anup, et al.
Veröffentlicht: (2024)
von: Singh, Anup, et al.
Veröffentlicht: (2024)
Recognizing Ornaments in Vocal Indian Art Music with Active Annotation
von: Kumar, Sumit, et al.
Veröffentlicht: (2025)
von: Kumar, Sumit, et al.
Veröffentlicht: (2025)
PiCoGen: Generate Piano Covers with a Two-stage Approach
von: Tan, Chih-Pin, et al.
Veröffentlicht: (2024)
von: Tan, Chih-Pin, et al.
Veröffentlicht: (2024)
A Generalized Weighted Overlap-Add (WOLA) Filter Bank for Improved Subband System Identification
von: Sharma, Mohit, et al.
Veröffentlicht: (2025)
von: Sharma, Mohit, et al.
Veröffentlicht: (2025)
Automatic Music Mixing using a Generative Model of Effect Embeddings
von: Moliner, Eloi, et al.
Veröffentlicht: (2025)
von: Moliner, Eloi, et al.
Veröffentlicht: (2025)
Automatic Live Music Song Identification Using Multi-level Deep Sequence Similarity Learning
von: Hakala, Aapo, et al.
Veröffentlicht: (2025)
von: Hakala, Aapo, et al.
Veröffentlicht: (2025)
Cochleagram-based Noise Adapted Speaker Identification System for Distorted Speech
von: Ahmed, Sabbir, et al.
Veröffentlicht: (2025)
von: Ahmed, Sabbir, et al.
Veröffentlicht: (2025)
Full-Duplex-Bench-v2: A Multi-Turn Evaluation Framework for Duplex Dialogue Systems with an Automated Examiner
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2025)
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2025)
The RoyalFlush Automatic Speech Diarization and Recognition System for In-Car Multi-Channel Automatic Speech Recognition Challenge
von: Tian, Jingguang, et al.
Veröffentlicht: (2024)
von: Tian, Jingguang, et al.
Veröffentlicht: (2024)
Pitfalls and Limits in Automatic Dementia Assessment
von: Braun, Franziska, et al.
Veröffentlicht: (2025)
von: Braun, Franziska, et al.
Veröffentlicht: (2025)
Non-Intrusive Automatic Speech Recognition Refinement: A Survey
von: Peyghan, Mohammad Reza, et al.
Veröffentlicht: (2025)
von: Peyghan, Mohammad Reza, et al.
Veröffentlicht: (2025)
Augmenting Polish Automatic Speech Recognition System With Synthetic Data
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling
von: Yao, Jixun, et al.
Veröffentlicht: (2025)
von: Yao, Jixun, et al.
Veröffentlicht: (2025)
DNN based HRIRs Identification with a Continuously Rotating Speaker Array
von: Ko, Byeong-Yun, et al.
Veröffentlicht: (2025)
von: Ko, Byeong-Yun, et al.
Veröffentlicht: (2025)
Advanced Signal Analysis in Detecting Replay Attacks for Automatic Speaker Verification Systems
von: Kuang, Lee Shih
Veröffentlicht: (2024)
von: Kuang, Lee Shih
Veröffentlicht: (2024)
CAFA: a Controllable Automatic Foley Artist
von: Benita, Roi, et al.
Veröffentlicht: (2025)
von: Benita, Roi, et al.
Veröffentlicht: (2025)
Unsupervised Online Continual Learning for Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2024)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2024)
Trainable Adaptive Score Normalization for Automatic Speaker Verification
von: Choi, Jeong-Hwan, et al.
Veröffentlicht: (2025)
von: Choi, Jeong-Hwan, et al.
Veröffentlicht: (2025)
Using Songs to Improve Kazakh Automatic Speech Recognition
von: Yeshpanov, Rustem
Veröffentlicht: (2026)
von: Yeshpanov, Rustem
Veröffentlicht: (2026)
Rhythm Features for Speaker Identification
von: Mehlman, Nick, et al.
Veröffentlicht: (2025)
von: Mehlman, Nick, et al.
Veröffentlicht: (2025)
Sequence-to-Sequence Neural Diarization with Automatic Speaker Detection and Representation
von: Cheng, Ming, et al.
Veröffentlicht: (2024)
von: Cheng, Ming, et al.
Veröffentlicht: (2024)
Fairness of Automatic Speech Recognition in Cleft Lip and Palate Speech
von: Bhattacharjee, Susmita, et al.
Veröffentlicht: (2025)
von: Bhattacharjee, Susmita, et al.
Veröffentlicht: (2025)
MusicEval: A Generative Music Dataset with Expert Ratings for Automatic Text-to-Music Evaluation
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
The DKU System for Multi-Speaker Automatic Speech Recognition in MLC-SLM Challenge
von: Lin, Yuke, et al.
Veröffentlicht: (2025)
von: Lin, Yuke, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Meta-learning-based percussion transcription and $t\bar{a}la$ identification from low-resource audio
von: Kodag, Rahul Bapusaheb, et al.
Veröffentlicht: (2025) -
Weakly Supervised Tabla Stroke Transcription via TI-SDRM: A Rhythm-Aware Lattice Rescoring Framework
von: Kodag, Rahul Bapusaheb, et al.
Veröffentlicht: (2026) -
Identification and Clustering of Unseen Ragas in Indian Art Music
von: Singh, Parampreet, et al.
Veröffentlicht: (2024) -
Explainable Deep Learning Analysis for Raga Identification in Indian Art Music
von: Singh, Parampreet, et al.
Veröffentlicht: (2024) -
AudioNet: Supervised Deep Hashing for Retrieval of Similar Audio Events
von: Dutta, Sagar, et al.
Veröffentlicht: (2025)