Hearing Between the Lines: Unlocking the Reasoning Power of LLMs for Speech Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Chandra, Arjun, Miller, Kevin, Ravichandran, Venkatesh, Papayiannis, Constantinos, Saligrama, Venkatesh |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BabyVLM: Data-Efficient Pretraining of VLMs Inspired by Infant Learning
by: Wang, Shengao, et al.
Published: (2025)
by: Wang, Shengao, et al.
Published: (2025)
Read Quietly, Think Aloud: Decoupling Comprehension and Reasoning in LLMs
by: Wang, Yuanxin, et al.
Published: (2025)
by: Wang, Yuanxin, et al.
Published: (2025)
Deep Companion Learning: Enhancing Generalization Through Historical Consistency
by: Zhu, Ruizhao, et al.
Published: (2024)
by: Zhu, Ruizhao, et al.
Published: (2024)
Repeated-Token Counting Reveals a Dissociation Between Representations and Outputs
by: Venkatesh, Sohan
Published: (2026)
by: Venkatesh, Sohan
Published: (2026)
Scaling Up Temporal Domain Generalization via Temporal Experts Averaging
by: Liu, Aoming, et al.
Published: (2025)
by: Liu, Aoming, et al.
Published: (2025)
Multi-Stage Multi-Modal Pre-Training for Automatic Speech Recognition
by: Jain, Yash, et al.
Published: (2024)
by: Jain, Yash, et al.
Published: (2024)
SPARC: Score Prompting and Adaptive Fusion for Zero-Shot Multi-Label Recognition in Vision-Language Models
by: Miller, Kevin, et al.
Published: (2025)
by: Miller, Kevin, et al.
Published: (2025)
Negative Before Positive: Asymmetric Valence Processing in Large Language Models
by: Venkatesh, Sohan
Published: (2026)
by: Venkatesh, Sohan
Published: (2026)
Architecture, Not Scale: Circuit Localization in Large Language Models
by: Venkatesh, Sohan
Published: (2026)
by: Venkatesh, Sohan
Published: (2026)
SCRAMBLe : Enhancing Multimodal LLM Compositionality with Synthetic Preference Data
by: Mishra, Samarth, et al.
Published: (2025)
by: Mishra, Samarth, et al.
Published: (2025)
Hear Both Sides: Efficient Multi-Agent Debate via Diversity-Aware Message Retention
by: Nguyen, Manh, et al.
Published: (2026)
by: Nguyen, Manh, et al.
Published: (2026)
Symmetry Reveals Layerwise Dynamics: How Transformers Perform In-Context Classification
by: Lutz, Patrick, et al.
Published: (2026)
by: Lutz, Patrick, et al.
Published: (2026)
Constrained Linear Thompson Sampling
by: Gangrade, Aditya, et al.
Published: (2025)
by: Gangrade, Aditya, et al.
Published: (2025)
Multi-Document Financial Question Answering using LLMs
by: Shah, Shalin, et al.
Published: (2024)
by: Shah, Shalin, et al.
Published: (2024)
Investigating the Shortcomings of LLMs in Step-by-Step Legal Reasoning
by: Mishra, Venkatesh, et al.
Published: (2025)
by: Mishra, Venkatesh, et al.
Published: (2025)
Cross-Lingual Auto Evaluation for Assessing Multilingual LLMs
by: Doddapaneni, Sumanth, et al.
Published: (2024)
by: Doddapaneni, Sumanth, et al.
Published: (2024)
TextBandit: Evaluating Probabilistic Reasoning in LLMs Through Language-Only Decision Tasks
by: Lim, Jimin, et al.
Published: (2025)
by: Lim, Jimin, et al.
Published: (2025)
Hearing to Translate: The Effectiveness of Speech Modality Integration into LLMs
by: Papi, Sara, et al.
Published: (2025)
by: Papi, Sara, et al.
Published: (2025)
Evaluating o1-Like LLMs: Unlocking Reasoning for Translation through Comprehensive Analysis
by: Chen, Andong, et al.
Published: (2025)
by: Chen, Andong, et al.
Published: (2025)
Can Small Language Models Handle Context-Summarized Multi-Turn Customer-Service QA? A Synthetic Data-Driven Comparative Evaluation
by: Cooray, Lakshan, et al.
Published: (2026)
by: Cooray, Lakshan, et al.
Published: (2026)
Mapping Hymns and Organizing Concepts in the Rigveda: Quantitatively Connecting the Vedic Suktas
by: Bollineni, Venkatesh, et al.
Published: (2025)
by: Bollineni, Venkatesh, et al.
Published: (2025)
The Thin Line Between Comprehension and Persuasion in LLMs
by: de Wynter, Adrian, et al.
Published: (2025)
by: de Wynter, Adrian, et al.
Published: (2025)
From Amateur to Master: Infusing Knowledge into LLMs via Automated Curriculum Learning
by: Neema, Nishit, et al.
Published: (2025)
by: Neema, Nishit, et al.
Published: (2025)
The Reasoning Bottleneck in Graph-RAG: Structured Prompting and Context Compression for Multi-Hop QA
by: Zarrinkia, Yasaman, et al.
Published: (2026)
by: Zarrinkia, Yasaman, et al.
Published: (2026)
The Ultimate Guide to Fine-Tuning LLMs from Basics to Breakthroughs: An Exhaustive Review of Technologies, Research, Best Practices, Applied Research Challenges and Opportunities
by: Parthasarathy, Venkatesh Balavadhani, et al.
Published: (2024)
by: Parthasarathy, Venkatesh Balavadhani, et al.
Published: (2024)
Speech LLMs are Contextual Reasoning Transcribers
by: Deng, Keqi, et al.
Published: (2026)
by: Deng, Keqi, et al.
Published: (2026)
The Future of MLLM Prompting is Adaptive: A Comprehensive Experimental Evaluation of Prompt Engineering Methods for Robust Multimodal Performance
by: Mohanty, Anwesha, et al.
Published: (2025)
by: Mohanty, Anwesha, et al.
Published: (2025)
Investigating and Addressing Hallucinations of LLMs in Tasks Involving Negation
by: Varshney, Neeraj, et al.
Published: (2024)
by: Varshney, Neeraj, et al.
Published: (2024)
Thinking to Recall: How Reasoning Unlocks Parametric Knowledge in LLMs
by: Gekhman, Zorik, et al.
Published: (2026)
by: Gekhman, Zorik, et al.
Published: (2026)
Beyond Reasoning: Reinforcement Learning Unlocks Parametric Knowledge in LLMs
by: Yang, Wanli, et al.
Published: (2026)
by: Yang, Wanli, et al.
Published: (2026)
Large Language Models are Algorithmically Blind
by: Venkatesh, Sohan, et al.
Published: (2026)
by: Venkatesh, Sohan, et al.
Published: (2026)
Safe Linear Bandits over Unknown Polytopes
by: Gangrade, Aditya, et al.
Published: (2022)
by: Gangrade, Aditya, et al.
Published: (2022)
RE-GrievanceAssist: Enhancing Customer Experience through ML-Powered Complaint Management
by: C, Venkatesh, et al.
Published: (2024)
by: C, Venkatesh, et al.
Published: (2024)
SRAG: RAG with Structured Data Improves Vector Retrieval
by: Shah, Shalin, et al.
Published: (2026)
by: Shah, Shalin, et al.
Published: (2026)
SimulS2S-LLM: Unlocking Simultaneous Inference of Speech LLMs for Speech-to-Speech Translation
by: Deng, Keqi, et al.
Published: (2025)
by: Deng, Keqi, et al.
Published: (2025)
Hypothesis Generation for Materials Discovery and Design Using Goal-Driven and Constraint-Guided LLM Agents
by: Kumbhar, Shrinidhi, et al.
Published: (2025)
by: Kumbhar, Shrinidhi, et al.
Published: (2025)
HearSay Benchmark: Do Audio LLMs Leak What They Hear?
by: Wang, Jin, et al.
Published: (2026)
by: Wang, Jin, et al.
Published: (2026)
FAMA: Failure-Aware Meta-Agentic Framework for Open-Source LLMs in Interactive Tool Use Environments
by: Saeidi, Amir, et al.
Published: (2026)
by: Saeidi, Amir, et al.
Published: (2026)
Say Anything but This: When Tokenizer Betrays Reasoning in LLMs
by: Ayoobi, Navid, et al.
Published: (2026)
by: Ayoobi, Navid, et al.
Published: (2026)
Merge-based syntax is mediated by distinct neurocognitive mechanisms: A clustering analysis of comprehension abilities in 84,000 individuals with language deficits across nine languages
by: Murphy, Elliot, et al.
Published: (2025)
by: Murphy, Elliot, et al.
Published: (2025)
Similar Items
-
BabyVLM: Data-Efficient Pretraining of VLMs Inspired by Infant Learning
by: Wang, Shengao, et al.
Published: (2025) -
Read Quietly, Think Aloud: Decoupling Comprehension and Reasoning in LLMs
by: Wang, Yuanxin, et al.
Published: (2025) -
Deep Companion Learning: Enhancing Generalization Through Historical Consistency
by: Zhu, Ruizhao, et al.
Published: (2024) -
Repeated-Token Counting Reveals a Dissociation Between Representations and Outputs
by: Venkatesh, Sohan
Published: (2026) -
Scaling Up Temporal Domain Generalization via Temporal Experts Averaging
by: Liu, Aoming, et al.
Published: (2025)