CMU's IWSLT 2024 Simultaneous Speech Translation System
Fuente:
arXiv
Guardado en:
| Autores principales: | Xu, Xi, Ouyang, Siqi, Yan, Brian, Fernandes, Patrick, Chen, William, Li, Lei, Neubig, Graham, Watanabe, Shinji |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CMU's IWSLT 2025 Simultaneous Speech Translation System
por: Ouyang, Siqi, et al.
Publicado: (2025)
por: Ouyang, Siqi, et al.
Publicado: (2025)
InfiniSST: Simultaneous Translation of Unbounded Speech with Large Language Model
por: Ouyang, Siqi, et al.
Publicado: (2025)
por: Ouyang, Siqi, et al.
Publicado: (2025)
FASST: Fast LLM-based Simultaneous Speech Translation
por: Ouyang, Siqi, et al.
Publicado: (2024)
por: Ouyang, Siqi, et al.
Publicado: (2024)
CA*: Addressing Evaluation Pitfalls in Computation-Aware Latency for Simultaneous Speech Translation
por: Xu, Xi, et al.
Publicado: (2024)
por: Xu, Xi, et al.
Publicado: (2024)
SimulSeamless: FBK at IWSLT 2024 Simultaneous Speech Translation
por: Papi, Sara, et al.
Publicado: (2024)
por: Papi, Sara, et al.
Publicado: (2024)
Blending LLMs into Cascaded Speech Translation: KIT's Offline Speech Translation System for IWSLT 2024
por: Koneru, Sai, et al.
Publicado: (2024)
por: Koneru, Sai, et al.
Publicado: (2024)
KIT's Offline Speech Translation and Instruction Following Submission for IWSLT 2025
por: Koneru, Sai, et al.
Publicado: (2025)
por: Koneru, Sai, et al.
Publicado: (2025)
NAIST Simultaneous Speech Translation System for IWSLT 2024
por: Ko, Yuka, et al.
Publicado: (2024)
por: Ko, Yuka, et al.
Publicado: (2024)
KIT's Low-resource Speech Translation Systems for IWSLT2025: System Enhancement with Synthetic Data and Model Regularization
por: Li, Zhaolin, et al.
Publicado: (2025)
por: Li, Zhaolin, et al.
Publicado: (2025)
OWLS: Scaling Laws for Multilingual Speech Recognition and Translation Models
por: Chen, William, et al.
Publicado: (2025)
por: Chen, William, et al.
Publicado: (2025)
Task Arithmetic for Language Expansion in Speech Translation
por: Cheng, Yao-Fei, et al.
Publicado: (2024)
por: Cheng, Yao-Fei, et al.
Publicado: (2024)
Instituto de Telecomunicações at IWSLT 2025: Aligning Small-Scale Speech and Language Models for Speech-to-Text Learning
por: Attanasio, Giuseppe, et al.
Publicado: (2025)
por: Attanasio, Giuseppe, et al.
Publicado: (2025)
Effective Strategies for Asynchronous Software Engineering Agents
por: Geng, Jiayi, et al.
Publicado: (2026)
por: Geng, Jiayi, et al.
Publicado: (2026)
Better Instruction-Following Through Minimum Bayes Risk
por: Wu, Ian, et al.
Publicado: (2024)
por: Wu, Ian, et al.
Publicado: (2024)
BeaverTalk: Oregon State University's IWSLT 2025 Simultaneous Speech Translation System
por: Raffel, Matthew, et al.
Publicado: (2025)
por: Raffel, Matthew, et al.
Publicado: (2025)
StressTransfer: Stress-Aware Speech-to-Speech Translation with Emphasis Preservation
por: Chen, Xi, et al.
Publicado: (2025)
por: Chen, Xi, et al.
Publicado: (2025)
Hierarchical Policy Optimization for Simultaneous Translation of Unbounded Speech
por: Ouyang, Siqi, et al.
Publicado: (2026)
por: Ouyang, Siqi, et al.
Publicado: (2026)
End-to-End Evaluation for Low-Latency Simultaneous Speech Translation
por: Huber, Christian, et al.
Publicado: (2023)
por: Huber, Christian, et al.
Publicado: (2023)
RASST: Fast Cross-modal Retrieval-Augmented Simultaneous Speech Translation
por: Luo, Jiaxuan, et al.
Publicado: (2026)
por: Luo, Jiaxuan, et al.
Publicado: (2026)
Recent Advances in End-to-End Simultaneous Speech Translation
por: Liu, Xiaoqian, et al.
Publicado: (2024)
por: Liu, Xiaoqian, et al.
Publicado: (2024)
MLLP-VRAIN UPV system for the IWSLT 2025 Simultaneous Speech Translation Translation task
por: Iranzo-Sánchez, Jorge, et al.
Publicado: (2025)
por: Iranzo-Sánchez, Jorge, et al.
Publicado: (2025)
Multimodal Sentiment Analysis on CMU-MOSEI Dataset using Transformer-based Models
por: Gajjar, Jugal, et al.
Publicado: (2025)
por: Gajjar, Jugal, et al.
Publicado: (2025)
Simultaneous Translation with Offline Speech and LLM Models in CUNI Submission to IWSLT 2025
por: Macháček, Dominik, et al.
Publicado: (2025)
por: Macháček, Dominik, et al.
Publicado: (2025)
TeamCMU at Touché: Adversarial Co-Evolution for Advertisement Integration and Detection in Conversational Search
por: Kim, To Eun, et al.
Publicado: (2025)
por: Kim, To Eun, et al.
Publicado: (2025)
StreamSpeech: Simultaneous Speech-to-Speech Translation with Multi-task Learning
por: Zhang, Shaolei, et al.
Publicado: (2024)
por: Zhang, Shaolei, et al.
Publicado: (2024)
M-Prometheus: A Suite of Open Multilingual LLM Judges
por: Pombal, José, et al.
Publicado: (2025)
por: Pombal, José, et al.
Publicado: (2025)
How "Real" is Your Real-Time Simultaneous Speech-to-Text Translation System?
por: Papi, Sara, et al.
Publicado: (2024)
por: Papi, Sara, et al.
Publicado: (2024)
Better Late Than Never: Meta-Evaluation of Latency Metrics for Simultaneous Speech-to-Text Translation
por: Polák, Peter, et al.
Publicado: (2025)
por: Polák, Peter, et al.
Publicado: (2025)
R-BI: Regularized Batched Inputs enhance Incremental Decoding Framework for Low-Latency Simultaneous Speech Translation
por: Guo, Jiaxin, et al.
Publicado: (2024)
por: Guo, Jiaxin, et al.
Publicado: (2024)
Can Large Language Models be Trusted for Evaluation? Scalable Meta-Evaluation of LLMs as Evaluators via Agent Debate
por: Chern, Steffi, et al.
Publicado: (2024)
por: Chern, Steffi, et al.
Publicado: (2024)
Training Versatile Coding Agents in Synthetic Environments
por: Zhu, Yiqi, et al.
Publicado: (2025)
por: Zhu, Yiqi, et al.
Publicado: (2025)
SimulU: Training-free Policy for Long-form Simultaneous Speech-to-Speech Translation
por: Djanibekov, Amirbek, et al.
Publicado: (2026)
por: Djanibekov, Amirbek, et al.
Publicado: (2026)
AugSumm: towards generalizable speech summarization using synthetic labels from large language model
por: Jung, Jee-weon, et al.
Publicado: (2024)
por: Jung, Jee-weon, et al.
Publicado: (2024)
Glancing Future for Simultaneous Machine Translation
por: Guo, Shoutao, et al.
Publicado: (2023)
por: Guo, Shoutao, et al.
Publicado: (2023)
A Non-autoregressive Generation Framework for End-to-End Simultaneous Speech-to-Speech Translation
por: Ma, Zhengrui, et al.
Publicado: (2024)
por: Ma, Zhengrui, et al.
Publicado: (2024)
SimulPL: Aligning Human Preferences in Simultaneous Machine Translation
por: Yu, Donglei, et al.
Publicado: (2025)
por: Yu, Donglei, et al.
Publicado: (2025)
DOA: Training-Free Decoder-Only Attention Policy for Long-Form Simultaneous Translation with SpeechLLMs
por: Papi, Sara, et al.
Publicado: (2026)
por: Papi, Sara, et al.
Publicado: (2026)
Alignment for Honesty
por: Yang, Yuqing, et al.
Publicado: (2023)
por: Yang, Yuqing, et al.
Publicado: (2023)
SELF-GUIDE: Better Task-Specific Instruction Following via Self-Synthetic Finetuning
por: Zhao, Chenyang, et al.
Publicado: (2024)
por: Zhao, Chenyang, et al.
Publicado: (2024)
What Are Tools Anyway? A Survey from the Language Model Perspective
por: Wang, Zhiruo, et al.
Publicado: (2024)
por: Wang, Zhiruo, et al.
Publicado: (2024)
Ejemplares similares
-
CMU's IWSLT 2025 Simultaneous Speech Translation System
por: Ouyang, Siqi, et al.
Publicado: (2025) -
InfiniSST: Simultaneous Translation of Unbounded Speech with Large Language Model
por: Ouyang, Siqi, et al.
Publicado: (2025) -
FASST: Fast LLM-based Simultaneous Speech Translation
por: Ouyang, Siqi, et al.
Publicado: (2024) -
CA*: Addressing Evaluation Pitfalls in Computation-Aware Latency for Simultaneous Speech Translation
por: Xu, Xi, et al.
Publicado: (2024) -
SimulSeamless: FBK at IWSLT 2024 Simultaneous Speech Translation
por: Papi, Sara, et al.
Publicado: (2024)