Token-level Ensembling of Models with Different Vocabularies
Fuente:
arXiv
Salvato in:
| Autori principali: | Wicks, Rachel, Ravisankar, Kartik, Yang, Xinchen, Koehn, Philipp, Post, Matt |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Recovering document annotations for sentence-level bitext
di: Wicks, Rachel, et al.
Pubblicazione: (2024)
di: Wicks, Rachel, et al.
Pubblicazione: (2024)
Escaping the sentence-level paradigm in machine translation
di: Post, Matt, et al.
Pubblicazione: (2023)
di: Post, Matt, et al.
Pubblicazione: (2023)
Can you map it to English? The Role of Cross-Lingual Alignment in Multilingual Performance of LLMs
di: Ravisankar, Kartik, et al.
Pubblicazione: (2025)
di: Ravisankar, Kartik, et al.
Pubblicazione: (2025)
Bridging the Gap between Different Vocabularies for LLM Ensemble
di: Xu, Yangyifan, et al.
Pubblicazione: (2024)
di: Xu, Yangyifan, et al.
Pubblicazione: (2024)
Speech Vecalign: an Embedding-based Method for Aligning Parallel Speech Documents
di: Meng, Chutong, et al.
Pubblicazione: (2025)
di: Meng, Chutong, et al.
Pubblicazione: (2025)
Text Style Transfer with Parameter-efficient LLM Finetuning and Round-trip Translation
di: Liu, Ruoxi, et al.
Pubblicazione: (2026)
di: Liu, Ruoxi, et al.
Pubblicazione: (2026)
Learn and Unlearn: Addressing Misinformation in Multilingual LLMs
di: Lu, Taiming, et al.
Pubblicazione: (2024)
di: Lu, Taiming, et al.
Pubblicazione: (2024)
Pointer-Generator Networks for Low-Resource Machine Translation: Don't Copy That!
di: Bafna, Niyati, et al.
Pubblicazione: (2024)
di: Bafna, Niyati, et al.
Pubblicazione: (2024)
PEAR: Pairwise Evaluation for Automatic Relative Scoring in Machine Translation
di: Proietti, Lorenzo, et al.
Pubblicazione: (2026)
di: Proietti, Lorenzo, et al.
Pubblicazione: (2026)
Error Norm Truncation: Robust Training in the Presence of Data Noise for Text Generation Models
di: Li, Tianjian, et al.
Pubblicazione: (2023)
di: Li, Tianjian, et al.
Pubblicazione: (2023)
Steering Large Language Models with Register Analysis for Arbitrary Style Transfer
di: Yang, Xinchen, et al.
Pubblicazione: (2025)
di: Yang, Xinchen, et al.
Pubblicazione: (2025)
SLIDE: Reference-free Evaluation for Machine Translation using a Sliding Document Window
di: Raunak, Vikas, et al.
Pubblicazione: (2023)
di: Raunak, Vikas, et al.
Pubblicazione: (2023)
Exploring Tokenization Strategies and Vocabulary Sizes for Enhanced Arabic Language Models
di: Alrefaie, Mohamed Taher, et al.
Pubblicazione: (2024)
di: Alrefaie, Mohamed Taher, et al.
Pubblicazione: (2024)
TokAlign++: Advancing Vocabulary Adaptation via Better Token Alignment
di: Li, Chong, et al.
Pubblicazione: (2026)
di: Li, Chong, et al.
Pubblicazione: (2026)
Navigating the Metrics Maze: Reconciling Score Magnitudes and Accuracies
di: Kocmi, Tom, et al.
Pubblicazione: (2024)
di: Kocmi, Tom, et al.
Pubblicazione: (2024)
Adaptive BPE Tokenization for Enhanced Vocabulary Adaptation in Finetuning Pretrained Language Models
di: Balde, Gunjan, et al.
Pubblicazione: (2024)
di: Balde, Gunjan, et al.
Pubblicazione: (2024)
DiffNorm: Self-Supervised Normalization for Non-autoregressive Speech-to-speech Translation
di: Tan, Weiting, et al.
Pubblicazione: (2024)
di: Tan, Weiting, et al.
Pubblicazione: (2024)
TokAlign: Efficient Vocabulary Adaptation via Token Alignment
di: Li, Chong, et al.
Pubblicazione: (2025)
di: Li, Chong, et al.
Pubblicazione: (2025)
Parallel Tokenizers: Rethinking Vocabulary Design for Cross-Lingual Transfer
di: Kautsar, Muhammad Dehan Al, et al.
Pubblicazione: (2025)
di: Kautsar, Muhammad Dehan Al, et al.
Pubblicazione: (2025)
When to Ensemble: Identifying Token-Level Points for Stable and Fast LLM Ensembling
di: Yun, Heecheol, et al.
Pubblicazione: (2025)
di: Yun, Heecheol, et al.
Pubblicazione: (2025)
HiMATE: A Hierarchical Multi-Agent Framework for Machine Translation Evaluation
di: Zhang, Shijie, et al.
Pubblicazione: (2025)
di: Zhang, Shijie, et al.
Pubblicazione: (2025)
BPE Gets Picky: Efficient Vocabulary Refinement During Tokenizer Training
di: Chizhov, Pavel, et al.
Pubblicazione: (2024)
di: Chizhov, Pavel, et al.
Pubblicazione: (2024)
Over-Tokenized Transformer: Vocabulary is Generally Worth Scaling
di: Huang, Hongzhi, et al.
Pubblicazione: (2025)
di: Huang, Hongzhi, et al.
Pubblicazione: (2025)
Neuron-Level Emotion Control in Speech-Generative Large Audio-Language Models
di: Zhao, Xiutian, et al.
Pubblicazione: (2026)
di: Zhao, Xiutian, et al.
Pubblicazione: (2026)
Interchangeable Token Embeddings for Extendable Vocabulary and Alpha-Equivalence
di: Işık, İlker, et al.
Pubblicazione: (2024)
di: Işık, İlker, et al.
Pubblicazione: (2024)
When the Same Coefficients Reach Different Places: Asymmetric Realizability in Transplanting Tokenizers across Large Language Models
di: Liu, Xiaoze, et al.
Pubblicazione: (2025)
di: Liu, Xiaoze, et al.
Pubblicazione: (2025)
Small Vocabularies, Big Gains: Pretraining and Tokenization in Time Series Models
di: Roger, Alexis, et al.
Pubblicazione: (2025)
di: Roger, Alexis, et al.
Pubblicazione: (2025)
X-ALMA: Plug & Play Modules and Adaptive Rejection for Quality Translation at Scale
di: Xu, Haoran, et al.
Pubblicazione: (2024)
di: Xu, Haoran, et al.
Pubblicazione: (2024)
Vocabulary-level Memory Efficiency for Language Model Fine-tuning
di: Williams, Miles, et al.
Pubblicazione: (2023)
di: Williams, Miles, et al.
Pubblicazione: (2023)
PyMarian: Fast Neural Machine Translation and Evaluation in Python
di: Gowda, Thamme, et al.
Pubblicazione: (2024)
di: Gowda, Thamme, et al.
Pubblicazione: (2024)
Optimizing LLMs for Italian: Reducing Token Fertility and Enhancing Efficiency Through Vocabulary Adaptation
di: Moroni, Luca, et al.
Pubblicazione: (2025)
di: Moroni, Luca, et al.
Pubblicazione: (2025)
Cut Your Losses in Large-Vocabulary Language Models
di: Wijmans, Erik, et al.
Pubblicazione: (2024)
di: Wijmans, Erik, et al.
Pubblicazione: (2024)
The Tokenization Bottleneck: How Vocabulary Extension Improves Chemistry Representation Learning in Pretrained Language Models
di: Kalamkar, Prathamesh, et al.
Pubblicazione: (2025)
di: Kalamkar, Prathamesh, et al.
Pubblicazione: (2025)
Exploring the Effect of Segmentation and Vocabulary Size on Speech Tokenization for Speech Language Models
di: Kando, Shunsuke, et al.
Pubblicazione: (2025)
di: Kando, Shunsuke, et al.
Pubblicazione: (2025)
The First to Know: How Token Distributions Reveal Hidden Knowledge in Large Vision-Language Models?
di: Zhao, Qinyu, et al.
Pubblicazione: (2024)
di: Zhao, Qinyu, et al.
Pubblicazione: (2024)
Rethinking Tokenization: Crafting Better Tokenizers for Large Language Models
di: Yang, Jinbiao
Pubblicazione: (2024)
di: Yang, Jinbiao
Pubblicazione: (2024)
Iterative Auto-Annotation for Scientific Named Entity Recognition Using BERT-Based Models
di: Gupta, Kartik
Pubblicazione: (2025)
di: Gupta, Kartik
Pubblicazione: (2025)
CTC-GMM: CTC guided modality matching for fast and accurate streaming speech translation
di: Zhao, Rui, et al.
Pubblicazione: (2024)
di: Zhao, Rui, et al.
Pubblicazione: (2024)
Overcoming Vocabulary Constraints with Pixel-level Fallback
di: Lotz, Jonas F., et al.
Pubblicazione: (2025)
di: Lotz, Jonas F., et al.
Pubblicazione: (2025)
DualToken: Towards Unifying Visual Understanding and Generation with Dual Visual Vocabularies
di: Song, Wei, et al.
Pubblicazione: (2025)
di: Song, Wei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Recovering document annotations for sentence-level bitext
di: Wicks, Rachel, et al.
Pubblicazione: (2024) -
Escaping the sentence-level paradigm in machine translation
di: Post, Matt, et al.
Pubblicazione: (2023) -
Can you map it to English? The Role of Cross-Lingual Alignment in Multilingual Performance of LLMs
di: Ravisankar, Kartik, et al.
Pubblicazione: (2025) -
Bridging the Gap between Different Vocabularies for LLM Ensemble
di: Xu, Yangyifan, et al.
Pubblicazione: (2024) -
Speech Vecalign: an Embedding-based Method for Aligning Parallel Speech Documents
di: Meng, Chutong, et al.
Pubblicazione: (2025)