AugSumm: towards generalizable speech summarization using synthetic labels from large language model
Fuente:
arXiv
Salvato in:
| Autori principali: | Jung, Jee-weon, Sharma, Roshan, Chen, William, Raj, Bhiksha, Watanabe, Shinji |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Can large audio language models understand child stuttering speech? speech summarization, and source separation
di: Okocha, Chibuzor, et al.
Pubblicazione: (2025)
di: Okocha, Chibuzor, et al.
Pubblicazione: (2025)
Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text continuation tasks
di: Maiti, Soumi, et al.
Pubblicazione: (2023)
di: Maiti, Soumi, et al.
Pubblicazione: (2023)
UniverSLU: Universal Spoken Language Understanding for Diverse Tasks with Natural Language Instructions
di: Arora, Siddhant, et al.
Pubblicazione: (2023)
di: Arora, Siddhant, et al.
Pubblicazione: (2023)
Factual consistency evaluation of summarization in the Era of large language models
di: Luo, Zheheng, et al.
Pubblicazione: (2024)
di: Luo, Zheheng, et al.
Pubblicazione: (2024)
The current status of large language models in summarizing radiology report impressions
di: Hu, Danqing, et al.
Pubblicazione: (2024)
di: Hu, Danqing, et al.
Pubblicazione: (2024)
On the Evaluation of Speech Foundation Models for Spoken Language Understanding
di: Arora, Siddhant, et al.
Pubblicazione: (2024)
di: Arora, Siddhant, et al.
Pubblicazione: (2024)
Improving Design of Input Condition Invariant Speech Enhancement
di: Zhang, Wangyou, et al.
Pubblicazione: (2024)
di: Zhang, Wangyou, et al.
Pubblicazione: (2024)
A dataset and benchmark for hospital course summarization with adapted large language models
di: Aali, Asad, et al.
Pubblicazione: (2024)
di: Aali, Asad, et al.
Pubblicazione: (2024)
Chain-of-Thought Training for Open E2E Spoken Dialogue Systems
di: Arora, Siddhant, et al.
Pubblicazione: (2025)
di: Arora, Siddhant, et al.
Pubblicazione: (2025)
Evaluating and Improving Continual Learning in Spoken Language Understanding
di: Yang, Muqiao, et al.
Pubblicazione: (2024)
di: Yang, Muqiao, et al.
Pubblicazione: (2024)
Closing the gap between open-source and commercial large language models for medical evidence summarization
di: Zhang, Gongbo, et al.
Pubblicazione: (2024)
di: Zhang, Gongbo, et al.
Pubblicazione: (2024)
Language-agnostic, automated assessment of listeners' speech recall using large language models
di: Herrmann, Björn
Pubblicazione: (2025)
di: Herrmann, Björn
Pubblicazione: (2025)
Simple synthetic data reduces sycophancy in large language models
di: Wei, Jerry, et al.
Pubblicazione: (2023)
di: Wei, Jerry, et al.
Pubblicazione: (2023)
TMT: Tri-Modal Translation between Speech, Image, and Text by Processing Different Modalities as Different Languages
di: Kim, Minsu, et al.
Pubblicazione: (2024)
di: Kim, Minsu, et al.
Pubblicazione: (2024)
Speech vs. Transcript: Does It Matter for Human Annotators in Speech Summarization?
di: Sharma, Roshan, et al.
Pubblicazione: (2024)
di: Sharma, Roshan, et al.
Pubblicazione: (2024)
Transferable speech-to-text large language model alignment module
di: Wu, Boyong, et al.
Pubblicazione: (2024)
di: Wu, Boyong, et al.
Pubblicazione: (2024)
OWSM v3.1: Better and Faster Open Whisper-Style Speech Models based on E-Branchformer
di: Peng, Yifan, et al.
Pubblicazione: (2024)
di: Peng, Yifan, et al.
Pubblicazione: (2024)
Zero-shot generation of synthetic neurosurgical data with large language models
di: Barr, Austin A., et al.
Pubblicazione: (2025)
di: Barr, Austin A., et al.
Pubblicazione: (2025)
Natural language guidance of high-fidelity text-to-speech with synthetic annotations
di: Lyth, Dan, et al.
Pubblicazione: (2024)
di: Lyth, Dan, et al.
Pubblicazione: (2024)
Beyond Silence: Bias Analysis through Loss and Asymmetric Approach in Audio Anti-Spoofing
di: Shim, Hye-jin, et al.
Pubblicazione: (2024)
di: Shim, Hye-jin, et al.
Pubblicazione: (2024)
On the steerability of large language models toward data-driven personas
di: Li, Junyi, et al.
Pubblicazione: (2023)
di: Li, Junyi, et al.
Pubblicazione: (2023)
Leveraging language models for summarizing mental state examinations: A comprehensive evaluation and dataset release
di: Sahu, Nilesh Kumar, et al.
Pubblicazione: (2024)
di: Sahu, Nilesh Kumar, et al.
Pubblicazione: (2024)
Beyond Performance Plateaus: A Comprehensive Study on Scalability in Speech Enhancement
di: Zhang, Wangyou, et al.
Pubblicazione: (2024)
di: Zhang, Wangyou, et al.
Pubblicazione: (2024)
FENICE: Factuality Evaluation of summarization based on Natural language Inference and Claim Extraction
di: Scirè, Alessandro, et al.
Pubblicazione: (2024)
di: Scirè, Alessandro, et al.
Pubblicazione: (2024)
Revisiting Acoustic Features for Robust ASR
di: Shah, Muhammad A., et al.
Pubblicazione: (2024)
di: Shah, Muhammad A., et al.
Pubblicazione: (2024)
Do large language models resemble humans in language use?
di: Cai, Zhenguang G., et al.
Pubblicazione: (2023)
di: Cai, Zhenguang G., et al.
Pubblicazione: (2023)
DELULU: Discriminative Embedding Learning Using Latent Units for Speaker-Aware Self-Trained Speech Foundational Model
di: Baali, Massa, et al.
Pubblicazione: (2025)
di: Baali, Massa, et al.
Pubblicazione: (2025)
QFS-Composer: Query-focused summarization pipeline for less resourced languages
di: Đuranović, Vuk, et al.
Pubblicazione: (2026)
di: Đuranović, Vuk, et al.
Pubblicazione: (2026)
Enhancing reasoning accuracy in large language models during inference time
di: Sharma, Vinay, et al.
Pubblicazione: (2026)
di: Sharma, Vinay, et al.
Pubblicazione: (2026)
Generative adversarial networks vs large language models: a comparative study on synthetic tabular data generation
di: Barr, Austin A., et al.
Pubblicazione: (2025)
di: Barr, Austin A., et al.
Pubblicazione: (2025)
Correcting misinformation on social media with a large language model
di: Zhou, Xinyi, et al.
Pubblicazione: (2024)
di: Zhou, Xinyi, et al.
Pubblicazione: (2024)
On the Robust Approximation of ASR Metrics
di: Waheed, Abdul, et al.
Pubblicazione: (2025)
di: Waheed, Abdul, et al.
Pubblicazione: (2025)
WhisperRT -- Turning Whisper into a Causal Streaming Model
di: Krichli, Tomer, et al.
Pubblicazione: (2025)
di: Krichli, Tomer, et al.
Pubblicazione: (2025)
What and When to Learn: CURriculum Ranking Loss for Large-Scale Speaker Verification
di: Baali, Massa, et al.
Pubblicazione: (2026)
di: Baali, Massa, et al.
Pubblicazione: (2026)
Measuring short-form factuality in large language models
di: Wei, Jason, et al.
Pubblicazione: (2024)
di: Wei, Jason, et al.
Pubblicazione: (2024)
The potential -- and the pitfalls -- of using pre-trained language models as cognitive science theories
di: Shah, Raj Sanjay, et al.
Pubblicazione: (2025)
di: Shah, Raj Sanjay, et al.
Pubblicazione: (2025)
Dissociating language and thought in large language models
di: Mahowald, Kyle, et al.
Pubblicazione: (2023)
di: Mahowald, Kyle, et al.
Pubblicazione: (2023)
Positive effects of speech and language therapy group interventions in primary progressive aphasia: A systematic review
di: Miyuki Watanabe, et al.
Pubblicazione: (2024)
di: Miyuki Watanabe, et al.
Pubblicazione: (2024)
Explainable Depression Detection using Masked Hard Instance Mining
di: Prakrankamanant, Patawee, et al.
Pubblicazione: (2025)
di: Prakrankamanant, Patawee, et al.
Pubblicazione: (2025)
AdvSumm: Adversarial Training for Bias Mitigation in Text Summarization
di: Gupta, Mukur, et al.
Pubblicazione: (2025)
di: Gupta, Mukur, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Can large audio language models understand child stuttering speech? speech summarization, and source separation
di: Okocha, Chibuzor, et al.
Pubblicazione: (2025) -
Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text continuation tasks
di: Maiti, Soumi, et al.
Pubblicazione: (2023) -
UniverSLU: Universal Spoken Language Understanding for Diverse Tasks with Natural Language Instructions
di: Arora, Siddhant, et al.
Pubblicazione: (2023) -
Factual consistency evaluation of summarization in the Era of large language models
di: Luo, Zheheng, et al.
Pubblicazione: (2024) -
The current status of large language models in summarizing radiology report impressions
di: Hu, Danqing, et al.
Pubblicazione: (2024)