Hearing Between the Lines: Unlocking the Reasoning Power of LLMs for Speech Evaluation
Fuente:
arXiv
Guardado en:
| Autores principales: | Chandra, Arjun, Miller, Kevin, Ravichandran, Venkatesh, Papayiannis, Constantinos, Saligrama, Venkatesh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
BabyVLM: Data-Efficient Pretraining of VLMs Inspired by Infant Learning
por: Wang, Shengao, et al.
Publicado: (2025)
por: Wang, Shengao, et al.
Publicado: (2025)
Read Quietly, Think Aloud: Decoupling Comprehension and Reasoning in LLMs
por: Wang, Yuanxin, et al.
Publicado: (2025)
por: Wang, Yuanxin, et al.
Publicado: (2025)
Deep Companion Learning: Enhancing Generalization Through Historical Consistency
por: Zhu, Ruizhao, et al.
Publicado: (2024)
por: Zhu, Ruizhao, et al.
Publicado: (2024)
Repeated-Token Counting Reveals a Dissociation Between Representations and Outputs
por: Venkatesh, Sohan
Publicado: (2026)
por: Venkatesh, Sohan
Publicado: (2026)
Scaling Up Temporal Domain Generalization via Temporal Experts Averaging
por: Liu, Aoming, et al.
Publicado: (2025)
por: Liu, Aoming, et al.
Publicado: (2025)
Multi-Stage Multi-Modal Pre-Training for Automatic Speech Recognition
por: Jain, Yash, et al.
Publicado: (2024)
por: Jain, Yash, et al.
Publicado: (2024)
SPARC: Score Prompting and Adaptive Fusion for Zero-Shot Multi-Label Recognition in Vision-Language Models
por: Miller, Kevin, et al.
Publicado: (2025)
por: Miller, Kevin, et al.
Publicado: (2025)
Negative Before Positive: Asymmetric Valence Processing in Large Language Models
por: Venkatesh, Sohan
Publicado: (2026)
por: Venkatesh, Sohan
Publicado: (2026)
Architecture, Not Scale: Circuit Localization in Large Language Models
por: Venkatesh, Sohan
Publicado: (2026)
por: Venkatesh, Sohan
Publicado: (2026)
SCRAMBLe : Enhancing Multimodal LLM Compositionality with Synthetic Preference Data
por: Mishra, Samarth, et al.
Publicado: (2025)
por: Mishra, Samarth, et al.
Publicado: (2025)
Hear Both Sides: Efficient Multi-Agent Debate via Diversity-Aware Message Retention
por: Nguyen, Manh, et al.
Publicado: (2026)
por: Nguyen, Manh, et al.
Publicado: (2026)
Symmetry Reveals Layerwise Dynamics: How Transformers Perform In-Context Classification
por: Lutz, Patrick, et al.
Publicado: (2026)
por: Lutz, Patrick, et al.
Publicado: (2026)
Constrained Linear Thompson Sampling
por: Gangrade, Aditya, et al.
Publicado: (2025)
por: Gangrade, Aditya, et al.
Publicado: (2025)
Multi-Document Financial Question Answering using LLMs
por: Shah, Shalin, et al.
Publicado: (2024)
por: Shah, Shalin, et al.
Publicado: (2024)
Investigating the Shortcomings of LLMs in Step-by-Step Legal Reasoning
por: Mishra, Venkatesh, et al.
Publicado: (2025)
por: Mishra, Venkatesh, et al.
Publicado: (2025)
Cross-Lingual Auto Evaluation for Assessing Multilingual LLMs
por: Doddapaneni, Sumanth, et al.
Publicado: (2024)
por: Doddapaneni, Sumanth, et al.
Publicado: (2024)
TextBandit: Evaluating Probabilistic Reasoning in LLMs Through Language-Only Decision Tasks
por: Lim, Jimin, et al.
Publicado: (2025)
por: Lim, Jimin, et al.
Publicado: (2025)
Hearing to Translate: The Effectiveness of Speech Modality Integration into LLMs
por: Papi, Sara, et al.
Publicado: (2025)
por: Papi, Sara, et al.
Publicado: (2025)
Evaluating o1-Like LLMs: Unlocking Reasoning for Translation through Comprehensive Analysis
por: Chen, Andong, et al.
Publicado: (2025)
por: Chen, Andong, et al.
Publicado: (2025)
Can Small Language Models Handle Context-Summarized Multi-Turn Customer-Service QA? A Synthetic Data-Driven Comparative Evaluation
por: Cooray, Lakshan, et al.
Publicado: (2026)
por: Cooray, Lakshan, et al.
Publicado: (2026)
Mapping Hymns and Organizing Concepts in the Rigveda: Quantitatively Connecting the Vedic Suktas
por: Bollineni, Venkatesh, et al.
Publicado: (2025)
por: Bollineni, Venkatesh, et al.
Publicado: (2025)
The Thin Line Between Comprehension and Persuasion in LLMs
por: de Wynter, Adrian, et al.
Publicado: (2025)
por: de Wynter, Adrian, et al.
Publicado: (2025)
From Amateur to Master: Infusing Knowledge into LLMs via Automated Curriculum Learning
por: Neema, Nishit, et al.
Publicado: (2025)
por: Neema, Nishit, et al.
Publicado: (2025)
The Reasoning Bottleneck in Graph-RAG: Structured Prompting and Context Compression for Multi-Hop QA
por: Zarrinkia, Yasaman, et al.
Publicado: (2026)
por: Zarrinkia, Yasaman, et al.
Publicado: (2026)
The Ultimate Guide to Fine-Tuning LLMs from Basics to Breakthroughs: An Exhaustive Review of Technologies, Research, Best Practices, Applied Research Challenges and Opportunities
por: Parthasarathy, Venkatesh Balavadhani, et al.
Publicado: (2024)
por: Parthasarathy, Venkatesh Balavadhani, et al.
Publicado: (2024)
Speech LLMs are Contextual Reasoning Transcribers
por: Deng, Keqi, et al.
Publicado: (2026)
por: Deng, Keqi, et al.
Publicado: (2026)
The Future of MLLM Prompting is Adaptive: A Comprehensive Experimental Evaluation of Prompt Engineering Methods for Robust Multimodal Performance
por: Mohanty, Anwesha, et al.
Publicado: (2025)
por: Mohanty, Anwesha, et al.
Publicado: (2025)
Investigating and Addressing Hallucinations of LLMs in Tasks Involving Negation
por: Varshney, Neeraj, et al.
Publicado: (2024)
por: Varshney, Neeraj, et al.
Publicado: (2024)
Thinking to Recall: How Reasoning Unlocks Parametric Knowledge in LLMs
por: Gekhman, Zorik, et al.
Publicado: (2026)
por: Gekhman, Zorik, et al.
Publicado: (2026)
Beyond Reasoning: Reinforcement Learning Unlocks Parametric Knowledge in LLMs
por: Yang, Wanli, et al.
Publicado: (2026)
por: Yang, Wanli, et al.
Publicado: (2026)
Large Language Models are Algorithmically Blind
por: Venkatesh, Sohan, et al.
Publicado: (2026)
por: Venkatesh, Sohan, et al.
Publicado: (2026)
Safe Linear Bandits over Unknown Polytopes
por: Gangrade, Aditya, et al.
Publicado: (2022)
por: Gangrade, Aditya, et al.
Publicado: (2022)
RE-GrievanceAssist: Enhancing Customer Experience through ML-Powered Complaint Management
por: C, Venkatesh, et al.
Publicado: (2024)
por: C, Venkatesh, et al.
Publicado: (2024)
SRAG: RAG with Structured Data Improves Vector Retrieval
por: Shah, Shalin, et al.
Publicado: (2026)
por: Shah, Shalin, et al.
Publicado: (2026)
SimulS2S-LLM: Unlocking Simultaneous Inference of Speech LLMs for Speech-to-Speech Translation
por: Deng, Keqi, et al.
Publicado: (2025)
por: Deng, Keqi, et al.
Publicado: (2025)
Hypothesis Generation for Materials Discovery and Design Using Goal-Driven and Constraint-Guided LLM Agents
por: Kumbhar, Shrinidhi, et al.
Publicado: (2025)
por: Kumbhar, Shrinidhi, et al.
Publicado: (2025)
HearSay Benchmark: Do Audio LLMs Leak What They Hear?
por: Wang, Jin, et al.
Publicado: (2026)
por: Wang, Jin, et al.
Publicado: (2026)
FAMA: Failure-Aware Meta-Agentic Framework for Open-Source LLMs in Interactive Tool Use Environments
por: Saeidi, Amir, et al.
Publicado: (2026)
por: Saeidi, Amir, et al.
Publicado: (2026)
Say Anything but This: When Tokenizer Betrays Reasoning in LLMs
por: Ayoobi, Navid, et al.
Publicado: (2026)
por: Ayoobi, Navid, et al.
Publicado: (2026)
Merge-based syntax is mediated by distinct neurocognitive mechanisms: A clustering analysis of comprehension abilities in 84,000 individuals with language deficits across nine languages
por: Murphy, Elliot, et al.
Publicado: (2025)
por: Murphy, Elliot, et al.
Publicado: (2025)
Ejemplares similares
-
BabyVLM: Data-Efficient Pretraining of VLMs Inspired by Infant Learning
por: Wang, Shengao, et al.
Publicado: (2025) -
Read Quietly, Think Aloud: Decoupling Comprehension and Reasoning in LLMs
por: Wang, Yuanxin, et al.
Publicado: (2025) -
Deep Companion Learning: Enhancing Generalization Through Historical Consistency
por: Zhu, Ruizhao, et al.
Publicado: (2024) -
Repeated-Token Counting Reveals a Dissociation Between Representations and Outputs
por: Venkatesh, Sohan
Publicado: (2026) -
Scaling Up Temporal Domain Generalization via Temporal Experts Averaging
por: Liu, Aoming, et al.
Publicado: (2025)