PhonologyBench: Evaluating Phonological Skills of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Suvarna, Ashima, Khandelwal, Harshita, Peng, Nanyun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
[b]=[d]-[t]+[p]: Self-supervised Speech Models Discover Phonological Vector Arithmetic
von: Choi, Kwanghee, et al.
Veröffentlicht: (2026)
von: Choi, Kwanghee, et al.
Veröffentlicht: (2026)
Spoken Language Intelligence of Large Language Models for Language Learning
von: Peng, Linkai, et al.
Veröffentlicht: (2023)
von: Peng, Linkai, et al.
Veröffentlicht: (2023)
Towards Signal Processing In Large Language Models
von: Verma, Prateek, et al.
Veröffentlicht: (2024)
von: Verma, Prateek, et al.
Veröffentlicht: (2024)
Adaptive Large Language Models By Layerwise Attention Shortcuts
von: Verma, Prateek, et al.
Veröffentlicht: (2024)
von: Verma, Prateek, et al.
Veröffentlicht: (2024)
Large Language Models' Internal Perception of Symbolic Music
von: Shin, Andrew, et al.
Veröffentlicht: (2025)
von: Shin, Andrew, et al.
Veröffentlicht: (2025)
Simultaneous Interpretation Corpus Construction by Large Language Models in Distant Language Pair
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
KoALa-Bench: Evaluating Large Audio Language Models on Korean Speech Understanding and Faithfulness
von: Kim, Jinyoung, et al.
Veröffentlicht: (2026)
von: Kim, Jinyoung, et al.
Veröffentlicht: (2026)
Large Language Models are Efficient Learners of Noise-Robust Speech Recognition
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
Whisper-GPT: A Hybrid Representation Audio Large Language Model
von: Verma, Prateek
Veröffentlicht: (2024)
von: Verma, Prateek
Veröffentlicht: (2024)
Turn-taking and Backchannel Prediction with Acoustic and Large Language Model Fusion
von: Wang, Jinhan, et al.
Veröffentlicht: (2024)
von: Wang, Jinhan, et al.
Veröffentlicht: (2024)
GenTranslate: Large Language Models are Generative Multilingual Speech and Machine Translators
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
SEAL: Speaker Error Correction using Acoustic-conditioned Large Language Models
von: Kumar, Anurag, et al.
Veröffentlicht: (2025)
von: Kumar, Anurag, et al.
Veröffentlicht: (2025)
C3LLM: Conditional Multimodal Content Generation Using Large Language Models
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
Generative Speech Recognition Error Correction with Large Language Models and Task-Activating Prompting
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2023)
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2023)
Probing for Phonology in Self-Supervised Speech Representations: A Case Study on Accent Perception
von: Venkateswaran, Nitin, et al.
Veröffentlicht: (2025)
von: Venkateswaran, Nitin, et al.
Veröffentlicht: (2025)
Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models
von: Kuan, Chun-Yi, et al.
Veröffentlicht: (2026)
von: Kuan, Chun-Yi, et al.
Veröffentlicht: (2026)
XMAD-Bench: Cross-Domain Multilingual Audio Deepfake Benchmark
von: Ciobanu, Ioan-Paul, et al.
Veröffentlicht: (2025)
von: Ciobanu, Ioan-Paul, et al.
Veröffentlicht: (2025)
Phonology-Guided Speech-to-Speech Translation for African Languages
von: Ochieng, Peter, et al.
Veröffentlicht: (2024)
von: Ochieng, Peter, et al.
Veröffentlicht: (2024)
Listen Again and Choose the Right Answer: A New Paradigm for Automatic Speech Recognition with Large Language Models
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
AQUA-Bench: Beyond Finding Answers to Knowing When There Are None in Audio Question Answering
von: Kuan, Chun-Yi, et al.
Veröffentlicht: (2026)
von: Kuan, Chun-Yi, et al.
Veröffentlicht: (2026)
Large Language Model Based Generative Error Correction: A Challenge and Baselines for Speech Recognition, Speaker Tagging, and Emotion Recognition
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2024)
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2024)
What Do Language Models Hear? Probing for Auditory Representations in Language Models
von: Ngo, Jerry, et al.
Veröffentlicht: (2024)
von: Ngo, Jerry, et al.
Veröffentlicht: (2024)
AIR-Bench: Benchmarking Large Audio-Language Models via Generative Comprehension
von: Yang, Qian, et al.
Veröffentlicht: (2024)
von: Yang, Qian, et al.
Veröffentlicht: (2024)
Wavelet GPT: Wavelet Inspired Large Language Models
von: Verma, Prateek
Veröffentlicht: (2024)
von: Verma, Prateek
Veröffentlicht: (2024)
Advancing Speech Understanding in Speech-Aware Language Models with GRPO
von: Elmakies, Avishai, et al.
Veröffentlicht: (2025)
von: Elmakies, Avishai, et al.
Veröffentlicht: (2025)
MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
von: Li, Yizhi, et al.
Veröffentlicht: (2023)
von: Li, Yizhi, et al.
Veröffentlicht: (2023)
Jailbreak-AudioBench: In-Depth Evaluation and Analysis of Jailbreak Threats for Large Audio Language Models
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
A Variational Framework for Improving Naturalness in Generative Spoken Language Models
von: Chen, Li-Wei, et al.
Veröffentlicht: (2025)
von: Chen, Li-Wei, et al.
Veröffentlicht: (2025)
Audio-Reasoner: Improving Reasoning Capability in Large Audio Language Models
von: Xie, Zhifei, et al.
Veröffentlicht: (2025)
von: Xie, Zhifei, et al.
Veröffentlicht: (2025)
Imagine to Hear: Auditory Knowledge Generation can be an Effective Assistant for Language Models
von: Yoo, Suho, et al.
Veröffentlicht: (2025)
von: Yoo, Suho, et al.
Veröffentlicht: (2025)
Unsupervised Speech Segmentation: A General Approach Using Speech Language Models
von: Elmakies, Avishai, et al.
Veröffentlicht: (2025)
von: Elmakies, Avishai, et al.
Veröffentlicht: (2025)
Slamming: Training a Speech Language Model on One GPU in a Day
von: Maimon, Gallil, et al.
Veröffentlicht: (2025)
von: Maimon, Gallil, et al.
Veröffentlicht: (2025)
Decoding Poultry Vocalizations -- Natural Language Processing and Transformer Models for Semantic and Emotional Analysis
von: Manikandan, Venkatraman, et al.
Veröffentlicht: (2024)
von: Manikandan, Venkatraman, et al.
Veröffentlicht: (2024)
Joint Fine-tuning and Conversion of Pretrained Speech and Language Models towards Linear Complexity
von: He, Mutian, et al.
Veröffentlicht: (2024)
von: He, Mutian, et al.
Veröffentlicht: (2024)
Analyzing Mitigation Strategies for Catastrophic Forgetting in End-to-End Training of Spoken Language Models
von: Hsiao, Chi-Yuan, et al.
Veröffentlicht: (2025)
von: Hsiao, Chi-Yuan, et al.
Veröffentlicht: (2025)
RALL-E: Robust Codec Language Modeling with Chain-of-Thought Prompting for Text-to-Speech Synthesis
von: Xin, Detai, et al.
Veröffentlicht: (2024)
von: Xin, Detai, et al.
Veröffentlicht: (2024)
Towards Controllable Speech Synthesis in the Era of Large Language Models: A Systematic Survey
von: Xie, Tianxin, et al.
Veröffentlicht: (2024)
von: Xie, Tianxin, et al.
Veröffentlicht: (2024)
NGPU-LM: GPU-Accelerated N-Gram Language Model for Context-Biasing in Greedy ASR Decoding
von: Bataev, Vladimir, et al.
Veröffentlicht: (2025)
von: Bataev, Vladimir, et al.
Veröffentlicht: (2025)
LLM Gesticulator: Leveraging Large Language Models for Scalable and Controllable Co-Speech Gesture Synthesis
von: Pang, Haozhou, et al.
Veröffentlicht: (2024)
von: Pang, Haozhou, et al.
Veröffentlicht: (2024)
Benchmarking and Confidence Evaluation of LALMs For Temporal Reasoning
von: Bhattacharya, Debarpan, et al.
Veröffentlicht: (2025)
von: Bhattacharya, Debarpan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
[b]=[d]-[t]+[p]: Self-supervised Speech Models Discover Phonological Vector Arithmetic
von: Choi, Kwanghee, et al.
Veröffentlicht: (2026) -
Spoken Language Intelligence of Large Language Models for Language Learning
von: Peng, Linkai, et al.
Veröffentlicht: (2023) -
Towards Signal Processing In Large Language Models
von: Verma, Prateek, et al.
Veröffentlicht: (2024) -
Adaptive Large Language Models By Layerwise Attention Shortcuts
von: Verma, Prateek, et al.
Veröffentlicht: (2024) -
Large Language Models' Internal Perception of Symbolic Music
von: Shin, Andrew, et al.
Veröffentlicht: (2025)