HITSZ's End-To-End Speech Translation Systems Combining Sequence-to-Sequence Auto Speech Recognition Model and Indic Large Language Model for IWSLT 2025 in Indic Track
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wei, Xuchen, Wu, Yangxin, Zhang, Yaoyin, Liu, Henglyu, Chen, Kehai, Bai, Xuefeng, Zhang, Min |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive Inner Speech-Text Alignment for LLM-based Speech Translation
von: Liu, Henglyu, et al.
Veröffentlicht: (2025)
von: Liu, Henglyu, et al.
Veröffentlicht: (2025)
XIFBench: Evaluating Large Language Models on Multilingual Instruction Following
von: Li, Zhenyu, et al.
Veröffentlicht: (2025)
von: Li, Zhenyu, et al.
Veröffentlicht: (2025)
CMU's IWSLT 2025 Simultaneous Speech Translation System
von: Ouyang, Siqi, et al.
Veröffentlicht: (2025)
von: Ouyang, Siqi, et al.
Veröffentlicht: (2025)
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey
von: Qu, Bingzheng, et al.
Veröffentlicht: (2026)
von: Qu, Bingzheng, et al.
Veröffentlicht: (2026)
LLaST: Improved End-to-end Speech Translation System Leveraged by Large Language Models
von: Chen, Xi, et al.
Veröffentlicht: (2024)
von: Chen, Xi, et al.
Veröffentlicht: (2024)
Exploring the Translation Mechanism of Large Language Models
von: Zhang, Hongbin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongbin, et al.
Veröffentlicht: (2025)
IndicMMLU-Pro: Benchmarking Indic Large Language Models on Multi-Task Language Understanding
von: KJ, Sankalp, et al.
Veröffentlicht: (2025)
von: KJ, Sankalp, et al.
Veröffentlicht: (2025)
Abusive Speech Detection in Indic Languages Using Acoustic Features
von: Spiesberger, Anika A., et al.
Veröffentlicht: (2024)
von: Spiesberger, Anika A., et al.
Veröffentlicht: (2024)
Indic-CodecFake meets SATYAM: Towards Detecting Neural Audio Codec Synthesized Speech Deepfakes in Indic Languages
von: Girish, et al.
Veröffentlicht: (2026)
von: Girish, et al.
Veröffentlicht: (2026)
Statistical Machine Translation for Indic Languages
von: Das, Sudhansu Bala, et al.
Veröffentlicht: (2023)
von: Das, Sudhansu Bala, et al.
Veröffentlicht: (2023)
Exploring an Inter-Pausal Unit (IPU) based Approach for Indic End-to-End TTS Systems
von: Prakash, Anusha, et al.
Veröffentlicht: (2024)
von: Prakash, Anusha, et al.
Veröffentlicht: (2024)
Paying More Attention to Source Context: Mitigating Unfaithful Translations from Large Language Model
von: Zhang, Hongbin, et al.
Veröffentlicht: (2024)
von: Zhang, Hongbin, et al.
Veröffentlicht: (2024)
Integrating Pre-Trained Speech and Language Models for End-to-End Speech Recognition
von: Hono, Yukiya, et al.
Veröffentlicht: (2023)
von: Hono, Yukiya, et al.
Veröffentlicht: (2023)
Simultaneous Translation with Offline Speech and LLM Models in CUNI Submission to IWSLT 2025
von: Macháček, Dominik, et al.
Veröffentlicht: (2025)
von: Macháček, Dominik, et al.
Veröffentlicht: (2025)
IndicSTR12: A Dataset for Indic Scene Text Recognition
von: Lunia, Harsh, et al.
Veröffentlicht: (2024)
von: Lunia, Harsh, et al.
Veröffentlicht: (2024)
A Non-autoregressive Generation Framework for End-to-End Simultaneous Speech-to-Speech Translation
von: Ma, Zhengrui, et al.
Veröffentlicht: (2024)
von: Ma, Zhengrui, et al.
Veröffentlicht: (2024)
CMU's IWSLT 2024 Simultaneous Speech Translation System
von: Xu, Xi, et al.
Veröffentlicht: (2024)
von: Xu, Xi, et al.
Veröffentlicht: (2024)
NAIST Simultaneous Speech Translation System for IWSLT 2024
von: Ko, Yuka, et al.
Veröffentlicht: (2024)
von: Ko, Yuka, et al.
Veröffentlicht: (2024)
GMU Systems for the IWSLT 2025 Low-Resource Speech Translation Shared Task
von: Meng, Chutong, et al.
Veröffentlicht: (2025)
von: Meng, Chutong, et al.
Veröffentlicht: (2025)
Safer in Translation? Presupposition Robustness in Indic Languages
von: Palnitkar, Aadi, et al.
Veröffentlicht: (2025)
von: Palnitkar, Aadi, et al.
Veröffentlicht: (2025)
Improving Multilingual Neural Machine Translation System for Indic Languages
von: Das, Sudhansu Bala, et al.
Veröffentlicht: (2022)
von: Das, Sudhansu Bala, et al.
Veröffentlicht: (2022)
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
von: Luu, Nam, et al.
Veröffentlicht: (2025)
von: Luu, Nam, et al.
Veröffentlicht: (2025)
Representation Purification for End-to-End Speech Translation
von: Zhang, Chengwei, et al.
Veröffentlicht: (2024)
von: Zhang, Chengwei, et al.
Veröffentlicht: (2024)
End-to-End Speech Recognition with Pre-trained Masked Language Model
von: Higuchi, Yosuke, et al.
Veröffentlicht: (2024)
von: Higuchi, Yosuke, et al.
Veröffentlicht: (2024)
KIT's Offline Speech Translation and Instruction Following Submission for IWSLT 2025
von: Koneru, Sai, et al.
Veröffentlicht: (2025)
von: Koneru, Sai, et al.
Veröffentlicht: (2025)
IndicVoices: Towards building an Inclusive Multilingual Speech Dataset for Indian Languages
von: Javed, Tahir, et al.
Veröffentlicht: (2024)
von: Javed, Tahir, et al.
Veröffentlicht: (2024)
KIT's Low-resource Speech Translation Systems for IWSLT2025: System Enhancement with Synthetic Data and Model Regularization
von: Li, Zhaolin, et al.
Veröffentlicht: (2025)
von: Li, Zhaolin, et al.
Veröffentlicht: (2025)
Blending LLMs into Cascaded Speech Translation: KIT's Offline Speech Translation System for IWSLT 2024
von: Koneru, Sai, et al.
Veröffentlicht: (2024)
von: Koneru, Sai, et al.
Veröffentlicht: (2024)
BeaverTalk: Oregon State University's IWSLT 2025 Simultaneous Speech Translation System
von: Raffel, Matthew, et al.
Veröffentlicht: (2025)
von: Raffel, Matthew, et al.
Veröffentlicht: (2025)
IndicParam: Benchmark to evaluate LLMs on low-resource Indic Languages
von: Maheshwari, Ayush, et al.
Veröffentlicht: (2025)
von: Maheshwari, Ayush, et al.
Veröffentlicht: (2025)
An End-to-End Speech Summarization Using Large Language Model
von: Shang, Hengchao, et al.
Veröffentlicht: (2024)
von: Shang, Hengchao, et al.
Veröffentlicht: (2024)
Vividh-ASR: A Complexity-Tiered Benchmark and Optimization Dynamics for Robust Indic Speech Recognition
von: Juvekar, Kush, et al.
Veröffentlicht: (2026)
von: Juvekar, Kush, et al.
Veröffentlicht: (2026)
MLLP-VRAIN UPV system for the IWSLT 2025 Simultaneous Speech Translation Translation task
von: Iranzo-Sánchez, Jorge, et al.
Veröffentlicht: (2025)
von: Iranzo-Sánchez, Jorge, et al.
Veröffentlicht: (2025)
Agentic Tool Use in Large Language Models
von: Hu, Jinchao, et al.
Veröffentlicht: (2026)
von: Hu, Jinchao, et al.
Veröffentlicht: (2026)
PSP: An Interpretable Per-Dimension Accent Benchmark for Indic Text-to-Speech
von: Menta, Venkata Pushpak Teja
Veröffentlicht: (2026)
von: Menta, Venkata Pushpak Teja
Veröffentlicht: (2026)
Harnessing the Zero-Shot Power of Instruction-Tuned Large Language Model in End-to-End Speech Recognition
von: Higuchi, Yosuke, et al.
Veröffentlicht: (2023)
von: Higuchi, Yosuke, et al.
Veröffentlicht: (2023)
IndicSentEval: How Effectively do Multilingual Transformer Models encode Linguistic Properties for Indic Languages?
von: Aravapalli, Akhilesh, et al.
Veröffentlicht: (2024)
von: Aravapalli, Akhilesh, et al.
Veröffentlicht: (2024)
Instituto de Telecomunicações at IWSLT 2025: Aligning Small-Scale Speech and Language Models for Speech-to-Text Learning
von: Attanasio, Giuseppe, et al.
Veröffentlicht: (2025)
von: Attanasio, Giuseppe, et al.
Veröffentlicht: (2025)
Towards Visually-Guided Movie Subtitle Translation for Indic Languages
von: Chintada, Tarun, et al.
Veröffentlicht: (2026)
von: Chintada, Tarun, et al.
Veröffentlicht: (2026)
Graph-Assisted Culturally Adaptable Idiomatic Translation for Indic Languages
von: Singh, Pratik Rakesh, et al.
Veröffentlicht: (2025)
von: Singh, Pratik Rakesh, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Adaptive Inner Speech-Text Alignment for LLM-based Speech Translation
von: Liu, Henglyu, et al.
Veröffentlicht: (2025) -
XIFBench: Evaluating Large Language Models on Multilingual Instruction Following
von: Li, Zhenyu, et al.
Veröffentlicht: (2025) -
CMU's IWSLT 2025 Simultaneous Speech Translation System
von: Ouyang, Siqi, et al.
Veröffentlicht: (2025) -
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey
von: Qu, Bingzheng, et al.
Veröffentlicht: (2026) -
LLaST: Improved End-to-end Speech Translation System Leveraged by Large Language Models
von: Chen, Xi, et al.
Veröffentlicht: (2024)