Contrastive Feedback Mechanism for Simultaneous Speech Translation
Fuente:
arXiv
Guardado en:
| Autores principales: | Tan, Haotian, Sakti, Sakriani |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SimulSense: Sense-Driven Interpreting for Efficient Simultaneous Speech Translation
por: Tan, Haotian, et al.
Publicado: (2025)
por: Tan, Haotian, et al.
Publicado: (2025)
SC-SOT: Conditioning the Decoder on Diarized Speaker Information for End-to-End Overlapped Speech Recognition
por: Hirano, Yuta, et al.
Publicado: (2025)
por: Hirano, Yuta, et al.
Publicado: (2025)
NAIST Simultaneous Speech Translation System for IWSLT 2024
por: Ko, Yuka, et al.
Publicado: (2024)
por: Ko, Yuka, et al.
Publicado: (2024)
On the Problem of Text-To-Speech Model Selection for Synthetic Data Generation in Automatic Speech Recognition
por: Rossenbach, Nick, et al.
Publicado: (2024)
por: Rossenbach, Nick, et al.
Publicado: (2024)
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
por: Adila, Aulia, et al.
Publicado: (2024)
por: Adila, Aulia, et al.
Publicado: (2024)
Indonesian-English Code-Switching Speech Synthesizer Utilizing Multilingual STEN-TTS and Bert LID
por: Handoyo, Ahmad Alfani, et al.
Publicado: (2024)
por: Handoyo, Ahmad Alfani, et al.
Publicado: (2024)
Continual Learning in Machine Speech Chain Using Gradient Episodic Memory
por: Tyndall, Geoffrey, et al.
Publicado: (2024)
por: Tyndall, Geoffrey, et al.
Publicado: (2024)
High-Fidelity Simultaneous Speech-To-Speech Translation
por: Labiausse, Tom, et al.
Publicado: (2025)
por: Labiausse, Tom, et al.
Publicado: (2025)
Hierarchical Policy Optimization for Simultaneous Translation of Unbounded Speech
por: Ouyang, Siqi, et al.
Publicado: (2026)
por: Ouyang, Siqi, et al.
Publicado: (2026)
CMU's IWSLT 2025 Simultaneous Speech Translation System
por: Ouyang, Siqi, et al.
Publicado: (2025)
por: Ouyang, Siqi, et al.
Publicado: (2025)
Toward Natural Emotional Text-To-Speech System with Fine-Grained Non-Verbal Expression Control
por: Zhou, Wangzixi, et al.
Publicado: (2026)
por: Zhou, Wangzixi, et al.
Publicado: (2026)
Simultaneous Speech-to-Speech Translation Without Aligned Data
por: Labiausse, Tom, et al.
Publicado: (2026)
por: Labiausse, Tom, et al.
Publicado: (2026)
Continuous Rating as Reliable Human Evaluation of Simultaneous Speech Translation
por: Javorský, Dávid, et al.
Publicado: (2022)
por: Javorský, Dávid, et al.
Publicado: (2022)
Efficient and Adaptive Simultaneous Speech Translation with Fully Unidirectional Architecture
por: Fu, Biao, et al.
Publicado: (2025)
por: Fu, Biao, et al.
Publicado: (2025)
MLLP-VRAIN UPV system for the IWSLT 2025 Simultaneous Speech Translation Translation task
por: Iranzo-Sánchez, Jorge, et al.
Publicado: (2025)
por: Iranzo-Sánchez, Jorge, et al.
Publicado: (2025)
FASST: Fast LLM-based Simultaneous Speech Translation
por: Ouyang, Siqi, et al.
Publicado: (2024)
por: Ouyang, Siqi, et al.
Publicado: (2024)
CMU's IWSLT 2024 Simultaneous Speech Translation System
por: Xu, Xi, et al.
Publicado: (2024)
por: Xu, Xi, et al.
Publicado: (2024)
Exploring the Correlation between Human and Machine Evaluation of Simultaneous Speech Translation
por: Wang, Xiaoman, et al.
Publicado: (2024)
por: Wang, Xiaoman, et al.
Publicado: (2024)
RASST: Fast Cross-modal Retrieval-Augmented Simultaneous Speech Translation
por: Luo, Jiaxuan, et al.
Publicado: (2026)
por: Luo, Jiaxuan, et al.
Publicado: (2026)
DPO-Tuned Large Language Models for Segmentation in Simultaneous Speech Translation
por: Yang, Zeyu, et al.
Publicado: (2025)
por: Yang, Zeyu, et al.
Publicado: (2025)
SASST: Leveraging Syntax-Aware Chunking and LLMs for Simultaneous Speech Translation
por: Yang, Zeyu, et al.
Publicado: (2025)
por: Yang, Zeyu, et al.
Publicado: (2025)
SimulTron: On-Device Simultaneous Speech to Speech Translation
por: Agranovich, Alex, et al.
Publicado: (2024)
por: Agranovich, Alex, et al.
Publicado: (2024)
End-to-End Evaluation for Low-Latency Simultaneous Speech Translation
por: Huber, Christian, et al.
Publicado: (2023)
por: Huber, Christian, et al.
Publicado: (2023)
Simultaneous Translation with Offline Speech and LLM Models in CUNI Submission to IWSLT 2025
por: Macháček, Dominik, et al.
Publicado: (2025)
por: Macháček, Dominik, et al.
Publicado: (2025)
StreamSpeech: Simultaneous Speech-to-Speech Translation with Multi-task Learning
por: Zhang, Shaolei, et al.
Publicado: (2024)
por: Zhang, Shaolei, et al.
Publicado: (2024)
SimulS2S-LLM: Unlocking Simultaneous Inference of Speech LLMs for Speech-to-Speech Translation
por: Deng, Keqi, et al.
Publicado: (2025)
por: Deng, Keqi, et al.
Publicado: (2025)
InfiniSST: Simultaneous Translation of Unbounded Speech with Large Language Model
por: Ouyang, Siqi, et al.
Publicado: (2025)
por: Ouyang, Siqi, et al.
Publicado: (2025)
PROST-LLM: Progressively Enhancing the Speech-to-Speech Translation Capability in LLMs
por: Xu, Jing, et al.
Publicado: (2026)
por: Xu, Jing, et al.
Publicado: (2026)
Recent Advances in End-to-End Simultaneous Speech Translation
por: Liu, Xiaoqian, et al.
Publicado: (2024)
por: Liu, Xiaoqian, et al.
Publicado: (2024)
CA*: Addressing Evaluation Pitfalls in Computation-Aware Latency for Simultaneous Speech Translation
por: Xu, Xi, et al.
Publicado: (2024)
por: Xu, Xi, et al.
Publicado: (2024)
On the Hallucination in Simultaneous Machine Translation
por: Zhong, Meizhi, et al.
Publicado: (2024)
por: Zhong, Meizhi, et al.
Publicado: (2024)
SimulU: Training-free Policy for Long-form Simultaneous Speech-to-Speech Translation
por: Djanibekov, Amirbek, et al.
Publicado: (2026)
por: Djanibekov, Amirbek, et al.
Publicado: (2026)
Joint Training And Decoding for Multilingual End-to-End Simultaneous Speech Translation
por: Huang, Wuwei, et al.
Publicado: (2025)
por: Huang, Wuwei, et al.
Publicado: (2025)
Label-Synchronous Neural Transducer for E2E Simultaneous Speech Translation
por: Deng, Keqi, et al.
Publicado: (2024)
por: Deng, Keqi, et al.
Publicado: (2024)
BeaverTalk: Oregon State University's IWSLT 2025 Simultaneous Speech Translation System
por: Raffel, Matthew, et al.
Publicado: (2025)
por: Raffel, Matthew, et al.
Publicado: (2025)
SimulSeamless: FBK at IWSLT 2024 Simultaneous Speech Translation
por: Papi, Sara, et al.
Publicado: (2024)
por: Papi, Sara, et al.
Publicado: (2024)
A Non-autoregressive Generation Framework for End-to-End Simultaneous Speech-to-Speech Translation
por: Ma, Zhengrui, et al.
Publicado: (2024)
por: Ma, Zhengrui, et al.
Publicado: (2024)
A Modular-based Strategy for Mitigating Gradient Conflicts in Simultaneous Speech Translation
por: Liu, Xiaoqian, et al.
Publicado: (2024)
por: Liu, Xiaoqian, et al.
Publicado: (2024)
Better Late Than Never: Meta-Evaluation of Latency Metrics for Simultaneous Speech-to-Text Translation
por: Polák, Peter, et al.
Publicado: (2025)
por: Polák, Peter, et al.
Publicado: (2025)
REINA: Regularized Entropy Information-Based Loss for Efficient Simultaneous Speech Translation
por: Hirschkind, Nameer, et al.
Publicado: (2025)
por: Hirschkind, Nameer, et al.
Publicado: (2025)
Ejemplares similares
-
SimulSense: Sense-Driven Interpreting for Efficient Simultaneous Speech Translation
por: Tan, Haotian, et al.
Publicado: (2025) -
SC-SOT: Conditioning the Decoder on Diarized Speaker Information for End-to-End Overlapped Speech Recognition
por: Hirano, Yuta, et al.
Publicado: (2025) -
NAIST Simultaneous Speech Translation System for IWSLT 2024
por: Ko, Yuka, et al.
Publicado: (2024) -
On the Problem of Text-To-Speech Model Selection for Synthetic Data Generation in Automatic Speech Recognition
por: Rossenbach, Nick, et al.
Publicado: (2024) -
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
por: Adila, Aulia, et al.
Publicado: (2024)