Pushing the Limits of Zero-shot End-to-End Speech Translation
Fuente:
arXiv
Salvato in:
| Autori principali: | Tsiamas, Ioannis, Gállego, Gerard I., Fonollosa, José A. R., Costa-jussà, Marta R. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Improving Language and Modality Transfer in Translation by Character-level Modeling
di: Tsiamas, Ioannis, et al.
Pubblicazione: (2025)
di: Tsiamas, Ioannis, et al.
Pubblicazione: (2025)
Unveiling the Role of Pretraining in Direct Speech Translation
di: Alastruey, Belen, et al.
Pubblicazione: (2024)
di: Alastruey, Belen, et al.
Pubblicazione: (2024)
SpeechAlign: a Framework for Speech Translation Alignment Evaluation
di: Alastruey, Belen, et al.
Pubblicazione: (2023)
di: Alastruey, Belen, et al.
Pubblicazione: (2023)
Speech is More Than Words: Do Speech-to-Text Translation Systems Leverage Prosody?
di: Tsiamas, Ioannis, et al.
Pubblicazione: (2024)
di: Tsiamas, Ioannis, et al.
Pubblicazione: (2024)
A Case Study on Filtering for End-to-End Speech Translation
di: Alam, Md Mahfuz Ibn, et al.
Pubblicazione: (2024)
di: Alam, Md Mahfuz Ibn, et al.
Pubblicazione: (2024)
Representation Purification for End-to-End Speech Translation
di: Zhang, Chengwei, et al.
Pubblicazione: (2024)
di: Zhang, Chengwei, et al.
Pubblicazione: (2024)
On the Similarity of Circuits across Languages: a Case Study on the Subject-verb Agreement Task
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
End-to-End Evaluation for Low-Latency Simultaneous Speech Translation
di: Huber, Christian, et al.
Pubblicazione: (2023)
di: Huber, Christian, et al.
Pubblicazione: (2023)
End-to-End Speech-to-Text Translation: A Survey
di: Sethiya, Nivedita, et al.
Pubblicazione: (2023)
di: Sethiya, Nivedita, et al.
Pubblicazione: (2023)
Gender-specific Machine Translation with Large Language Models
di: Sánchez, Eduardo, et al.
Pubblicazione: (2023)
di: Sánchez, Eduardo, et al.
Pubblicazione: (2023)
Revisiting Direct Speech-to-Text Translation with Speech LLMs: Better Scaling than CoT Prompting?
di: Pareras, Oriol, et al.
Pubblicazione: (2025)
di: Pareras, Oriol, et al.
Pubblicazione: (2025)
Recent Advances in End-to-End Simultaneous Speech Translation
di: Liu, Xiaoqian, et al.
Pubblicazione: (2024)
di: Liu, Xiaoqian, et al.
Pubblicazione: (2024)
Translate, then Detect: Leveraging Machine Translation for Cross-Lingual Toxicity Classification
di: Bell, Samuel J., et al.
Pubblicazione: (2025)
di: Bell, Samuel J., et al.
Pubblicazione: (2025)
Joint Training And Decoding for Multilingual End-to-End Simultaneous Speech Translation
di: Huang, Wuwei, et al.
Pubblicazione: (2025)
di: Huang, Wuwei, et al.
Pubblicazione: (2025)
Listening or Reading? Evaluating Speech Awareness in Chain-of-Thought Speech-to-Text Translation
di: Romero-Díaz, Jacobo, et al.
Pubblicazione: (2025)
di: Romero-Díaz, Jacobo, et al.
Pubblicazione: (2025)
DTW-Align: Bridging the Modality Gap in End-to-End Speech Translation with Dynamic Time Warping Alignment
di: Issam, Abderrahmane, et al.
Pubblicazione: (2025)
di: Issam, Abderrahmane, et al.
Pubblicazione: (2025)
A Non-autoregressive Generation Framework for End-to-End Simultaneous Speech-to-Speech Translation
di: Ma, Zhengrui, et al.
Pubblicazione: (2024)
di: Ma, Zhengrui, et al.
Pubblicazione: (2024)
Leveraging Synthetic Audio Data for End-to-End Low-Resource Speech Translation
di: Moslem, Yasmin
Pubblicazione: (2024)
di: Moslem, Yasmin
Pubblicazione: (2024)
End-to-End Intracortical Speech Decoding from Neural Activity
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2026)
di: Khanday, Owais Mujtaba, et al.
Pubblicazione: (2026)
End-to-End Speech Translation for Low-Resource Languages Using Weakly Labeled Data
di: Pothula, Aishwarya, et al.
Pubblicazione: (2025)
di: Pothula, Aishwarya, et al.
Pubblicazione: (2025)
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
di: Luu, Nam, et al.
Pubblicazione: (2025)
di: Luu, Nam, et al.
Pubblicazione: (2025)
WildSpeech-Bench: Benchmarking End-to-End SpeechLLMs in the Wild
di: Zhang, Linhao, et al.
Pubblicazione: (2025)
di: Zhang, Linhao, et al.
Pubblicazione: (2025)
When End-to-End is Overkill: Rethinking Cascaded Speech-to-Text Translation
di: Min, Anna, et al.
Pubblicazione: (2025)
di: Min, Anna, et al.
Pubblicazione: (2025)
Soft Language Identification for Language-Agnostic Many-to-One End-to-End Speech Translation
di: Wang, Peidong, et al.
Pubblicazione: (2024)
di: Wang, Peidong, et al.
Pubblicazione: (2024)
Long-Form End-to-End Speech Translation via Latent Alignment Segmentation
di: Polák, Peter, et al.
Pubblicazione: (2023)
di: Polák, Peter, et al.
Pubblicazione: (2023)
AdaST: Dynamically Adapting Encoder States in the Decoder for End-to-End Speech-to-Text Translation
di: Huang, Wuwei, et al.
Pubblicazione: (2025)
di: Huang, Wuwei, et al.
Pubblicazione: (2025)
StutterZero and StutterFormer: End-to-End Speech Conversion for Stuttering Transcription and Correction
di: Xu, Qianheng
Pubblicazione: (2025)
di: Xu, Qianheng
Pubblicazione: (2025)
MuTox: Universal MUltilingual Audio-based TOXicity Dataset and Zero-shot Detector
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
Cross-modality Data Augmentation for End-to-End Sign Language Translation
di: Ye, Jinhui, et al.
Pubblicazione: (2023)
di: Ye, Jinhui, et al.
Pubblicazione: (2023)
A Primer on the Inner Workings of Transformer-based Language Models
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
HITSZ's End-To-End Speech Translation Systems Combining Sequence-to-Sequence Auto Speech Recognition Model and Indic Large Language Model for IWSLT 2025 in Indic Track
di: Wei, Xuchen, et al.
Pubblicazione: (2025)
di: Wei, Xuchen, et al.
Pubblicazione: (2025)
Harnessing the Zero-Shot Power of Instruction-Tuned Large Language Model in End-to-End Speech Recognition
di: Higuchi, Yosuke, et al.
Pubblicazione: (2023)
di: Higuchi, Yosuke, et al.
Pubblicazione: (2023)
End-to-End Training for Back-Translation with Categorical Reparameterization Trick
di: Heo, DongNyeong, et al.
Pubblicazione: (2022)
di: Heo, DongNyeong, et al.
Pubblicazione: (2022)
BOUQuET: dataset, Benchmark and Open initiative for Universal Quality Evaluation in Translation
di: The Omnilingual MT Team, et al.
Pubblicazione: (2025)
di: The Omnilingual MT Team, et al.
Pubblicazione: (2025)
Growing Trees on Sounds: Assessing Strategies for End-to-End Dependency Parsing of Speech
di: Pupier, Adrien, et al.
Pubblicazione: (2024)
di: Pupier, Adrien, et al.
Pubblicazione: (2024)
LELA: An End-to-end LLM-based Entity Linking Framework with Zero-shot Domain Adaptation
di: Haffoudhi, Samy, et al.
Pubblicazione: (2026)
di: Haffoudhi, Samy, et al.
Pubblicazione: (2026)
LLaST: Improved End-to-end Speech Translation System Leveraged by Large Language Models
di: Chen, Xi, et al.
Pubblicazione: (2024)
di: Chen, Xi, et al.
Pubblicazione: (2024)
Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation
di: Lan, Zhibin, et al.
Pubblicazione: (2024)
di: Lan, Zhibin, et al.
Pubblicazione: (2024)
Enhancing Speech-to-Speech Dialogue Modeling with End-to-End Retrieval-Augmented Generation
di: Feng, Pengchao, et al.
Pubblicazione: (2025)
di: Feng, Pengchao, et al.
Pubblicazione: (2025)
With Good MT There is No Need For End-to-End: A Case for Translate-then-Summarize Cross-lingual Summarization
di: Varab, Daniel, et al.
Pubblicazione: (2024)
di: Varab, Daniel, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Improving Language and Modality Transfer in Translation by Character-level Modeling
di: Tsiamas, Ioannis, et al.
Pubblicazione: (2025) -
Unveiling the Role of Pretraining in Direct Speech Translation
di: Alastruey, Belen, et al.
Pubblicazione: (2024) -
SpeechAlign: a Framework for Speech Translation Alignment Evaluation
di: Alastruey, Belen, et al.
Pubblicazione: (2023) -
Speech is More Than Words: Do Speech-to-Text Translation Systems Leverage Prosody?
di: Tsiamas, Ioannis, et al.
Pubblicazione: (2024) -
A Case Study on Filtering for End-to-End Speech Translation
di: Alam, Md Mahfuz Ibn, et al.
Pubblicazione: (2024)