Decoupled Vocabulary Learning Enables Zero-Shot Translation from Unseen Languages
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mullov, Carlos, Pham, Ngoc-Quan, Waibel, Alexander |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Weight Factorization and Centralization for Continual Learning in Speech Recognition
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
Adapting Language Balance in Code-Switching Speech
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
Zero-Shot Strategies for Length-Controllable Summarization
von: Retkowski, Fabian, et al.
Veröffentlicht: (2024)
von: Retkowski, Fabian, et al.
Veröffentlicht: (2024)
Towards continually learning new languages
von: Pham, Ngoc-Quan, et al.
Veröffentlicht: (2022)
von: Pham, Ngoc-Quan, et al.
Veröffentlicht: (2022)
Cocktail-Party Audio-Visual Speech Recognition
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2025)
Accent conversion using discrete units with parallel data synthesized from controllable accented TTS
von: Nguyen, Tuan Nam, et al.
Veröffentlicht: (2024)
von: Nguyen, Tuan Nam, et al.
Veröffentlicht: (2024)
Predictive Speech Recognition and End-of-Utterance Detection Towards Spoken Dialog Systems
von: Zink, Oswald, et al.
Veröffentlicht: (2024)
von: Zink, Oswald, et al.
Veröffentlicht: (2024)
Bayesian Low-Rank Factorization for Robust Model Adaptation
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
End-to-End Evaluation for Low-Latency Simultaneous Speech Translation
von: Huber, Christian, et al.
Veröffentlicht: (2023)
von: Huber, Christian, et al.
Veröffentlicht: (2023)
Blending LLMs into Cascaded Speech Translation: KIT's Offline Speech Translation System for IWSLT 2024
von: Koneru, Sai, et al.
Veröffentlicht: (2024)
von: Koneru, Sai, et al.
Veröffentlicht: (2024)
Streaming Non-Autoregressive Model for Accent Conversion and Pronunciation Improvement
von: Nguyen, Tuan-Nam, et al.
Veröffentlicht: (2025)
von: Nguyen, Tuan-Nam, et al.
Veröffentlicht: (2025)
PIER: A Novel Metric for Evaluating What Matters in Code-Switching
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025)
A Cocktail-Party Benchmark: Multi-Modal dataset and Comparative Evaluation Results
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2025)
KIT's Low-resource Speech Translation Systems for IWSLT2025: System Enhancement with Synthetic Data and Model Regularization
von: Li, Zhaolin, et al.
Veröffentlicht: (2025)
von: Li, Zhaolin, et al.
Veröffentlicht: (2025)
Continuously Learning New Words in Automatic Speech Recognition
von: Huber, Christian, et al.
Veröffentlicht: (2024)
von: Huber, Christian, et al.
Veröffentlicht: (2024)
Towards Zero-Shot, Controllable Dialog Planning with LLMs
von: Väth, Dirk, et al.
Veröffentlicht: (2024)
von: Väth, Dirk, et al.
Veröffentlicht: (2024)
CharSpan: Utilizing Lexical Similarity to Enable Zero-Shot Machine Translation for Extremely Low-resource Languages
von: Maurya, Kaushal Kumar, et al.
Veröffentlicht: (2023)
von: Maurya, Kaushal Kumar, et al.
Veröffentlicht: (2023)
Nsanku: Evaluating Zero-Shot Translation Performance of LLMs for Ghanaian Languages
von: Moore, Stephen E., et al.
Veröffentlicht: (2026)
von: Moore, Stephen E., et al.
Veröffentlicht: (2026)
From Text Segmentation to Smart Chaptering: A Novel Benchmark for Structuring Video Transcriptions
von: Retkowski, Fabian, et al.
Veröffentlicht: (2024)
von: Retkowski, Fabian, et al.
Veröffentlicht: (2024)
Paragraph Segmentation Revisited: Towards a Standard Task for Structuring Speech
von: Retkowski, Fabian, et al.
Veröffentlicht: (2025)
von: Retkowski, Fabian, et al.
Veröffentlicht: (2025)
Spectral Prompt Tuning:Unveiling Unseen Classes for Zero-Shot Semantic Segmentation
von: Xu, Wenhao, et al.
Veröffentlicht: (2023)
von: Xu, Wenhao, et al.
Veröffentlicht: (2023)
Machine Translation Models are Zero-Shot Detectors of Translation Direction
von: Wastl, Michelle, et al.
Veröffentlicht: (2024)
von: Wastl, Michelle, et al.
Veröffentlicht: (2024)
Towards Zero-Shot Multimodal Machine Translation
von: Futeral, Matthieu, et al.
Veröffentlicht: (2024)
von: Futeral, Matthieu, et al.
Veröffentlicht: (2024)
A Zero-Shot Open-Vocabulary Pipeline for Dialogue Understanding
von: Safa, Abdulfattah, et al.
Veröffentlicht: (2024)
von: Safa, Abdulfattah, et al.
Veröffentlicht: (2024)
Zero-Shot Hierarchical Classification on the Common Procurement Vocabulary Taxonomy
von: Moiraghi, Federico, et al.
Veröffentlicht: (2024)
von: Moiraghi, Federico, et al.
Veröffentlicht: (2024)
Lombard Speech Synthesis for Any Voice with Controllable Style Embeddings
von: Akti, Seymanur, et al.
Veröffentlicht: (2026)
von: Akti, Seymanur, et al.
Veröffentlicht: (2026)
Languages Transferred Within the Encoder: On Representation Transfer in Zero-Shot Multilingual Translation
von: Qu, Zhi, et al.
Veröffentlicht: (2024)
von: Qu, Zhi, et al.
Veröffentlicht: (2024)
LLMs Are Zero-Shot Context-Aware Simultaneous Translators
von: Koshkin, Roman, et al.
Veröffentlicht: (2024)
von: Koshkin, Roman, et al.
Veröffentlicht: (2024)
Context Biasing for Pronunciation-Orthography Mismatch in Automatic Speech Recognition
von: Huber, Christian, et al.
Veröffentlicht: (2025)
von: Huber, Christian, et al.
Veröffentlicht: (2025)
Zero-Shot Detection of LLM-Generated Text using Token Cohesiveness
von: Ma, Shixuan, et al.
Veröffentlicht: (2024)
von: Ma, Shixuan, et al.
Veröffentlicht: (2024)
LCS: A Language Converter Strategy for Zero-Shot Neural Machine Translation
von: Sun, Zengkui, et al.
Veröffentlicht: (2024)
von: Sun, Zengkui, et al.
Veröffentlicht: (2024)
Understanding and Mitigating the Uncertainty in Zero-Shot Translation
von: Wang, Wenxuan, et al.
Veröffentlicht: (2022)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2022)
Handling Numeric Expressions in Automatic Speech Recognition
von: Huber, Christian, et al.
Veröffentlicht: (2024)
von: Huber, Christian, et al.
Veröffentlicht: (2024)
KIT's Offline Speech Translation and Instruction Following Submission for IWSLT 2025
von: Koneru, Sai, et al.
Veröffentlicht: (2025)
von: Koneru, Sai, et al.
Veröffentlicht: (2025)
Structure-aware Prompt Adaptation from Seen to Unseen for Open-Vocabulary Compositional Zero-Shot Learning
von: Duan, Yihang, et al.
Veröffentlicht: (2026)
von: Duan, Yihang, et al.
Veröffentlicht: (2026)
Leveraging Sentence-oriented Augmentation and Transformer-Based Architecture for Vietnamese-Bahnaric Translation
von: Nguyen, Tan Sang, et al.
Veröffentlicht: (2026)
von: Nguyen, Tan Sang, et al.
Veröffentlicht: (2026)
Tuning LLMs with Contrastive Alignment Instructions for Machine Translation in Unseen, Low-resource Languages
von: Mao, Zhuoyuan, et al.
Veröffentlicht: (2024)
von: Mao, Zhuoyuan, et al.
Veröffentlicht: (2024)
The AI Co-Ethnographer: How Far Can Automation Take Qualitative Research?
von: Retkowski, Fabian, et al.
Veröffentlicht: (2025)
von: Retkowski, Fabian, et al.
Veröffentlicht: (2025)
Who Watches the Watchmen? Humans Disagree With Translation Metrics on Unseen Domains
von: Schmidt, Finn, et al.
Veröffentlicht: (2026)
von: Schmidt, Finn, et al.
Veröffentlicht: (2026)
MSA-ASR: Efficient Multilingual Speaker Attribution with frozen ASR Models
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2024)
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Weight Factorization and Centralization for Continual Learning in Speech Recognition
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025) -
Adapting Language Balance in Code-Switching Speech
von: Ugan, Enes Yavuz, et al.
Veröffentlicht: (2025) -
Zero-Shot Strategies for Length-Controllable Summarization
von: Retkowski, Fabian, et al.
Veröffentlicht: (2024) -
Towards continually learning new languages
von: Pham, Ngoc-Quan, et al.
Veröffentlicht: (2022) -
Cocktail-Party Audio-Visual Speech Recognition
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2025)