Zero-Shot Tokenizer Transfer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Minixhofer, Benjamin, Ponti, Edoardo Maria, Vulić, Ivan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Universal Cross-Tokenizer Distillation via Approximate Likelihood Matching
von: Minixhofer, Benjamin, et al.
Veröffentlicht: (2025)
von: Minixhofer, Benjamin, et al.
Veröffentlicht: (2025)
Retrofitting Large Language Models with Dynamic Tokenization
von: Feher, Darius, et al.
Veröffentlicht: (2024)
von: Feher, Darius, et al.
Veröffentlicht: (2024)
Emergent Communication Pretraining for Few-Shot Machine Translation
von: Li, Yaoyiran, et al.
Veröffentlicht: (2020)
von: Li, Yaoyiran, et al.
Veröffentlicht: (2020)
Segment Any Text: A Universal Approach for Robust, Efficient and Adaptable Sentence Segmentation
von: Frohmann, Markus, et al.
Veröffentlicht: (2024)
von: Frohmann, Markus, et al.
Veröffentlicht: (2024)
Scaling Sparse Fine-Tuning to Large Language Models
von: Ansell, Alan, et al.
Veröffentlicht: (2024)
von: Ansell, Alan, et al.
Veröffentlicht: (2024)
Analyzing and Adapting Large Language Models for Few-Shot Multilingual NLU: Are We There Yet?
von: Razumovskaia, Evgeniia, et al.
Veröffentlicht: (2024)
von: Razumovskaia, Evgeniia, et al.
Veröffentlicht: (2024)
Bolmo: Byteifying the Next Generation of Language Models
von: Minixhofer, Benjamin, et al.
Veröffentlicht: (2025)
von: Minixhofer, Benjamin, et al.
Veröffentlicht: (2025)
Cross-Lingual and Cross-Cultural Variation in Image Descriptions
von: Berger, Uri, et al.
Veröffentlicht: (2024)
von: Berger, Uri, et al.
Veröffentlicht: (2024)
Language Fusion for Parameter-Efficient Cross-lingual Transfer
von: Borchert, Philipp, et al.
Veröffentlicht: (2025)
von: Borchert, Philipp, et al.
Veröffentlicht: (2025)
FUN with Fisher: Improving Generalization of Adapter-Based Cross-lingual Transfer with Scheduled Unfreezing
von: Liu, Chen Cecilia, et al.
Veröffentlicht: (2023)
von: Liu, Chen Cecilia, et al.
Veröffentlicht: (2023)
Navigating the Alignment-Calibration Trade-off: A Pareto-Superior Frontier via Model Merging
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
DARE: Diverse Visual Question Answering with Robustness Evaluation
von: Sterz, Hannah, et al.
Veröffentlicht: (2024)
von: Sterz, Hannah, et al.
Veröffentlicht: (2024)
Quantifying Language Disparities in Multilingual Large Language Models
von: Hu, Songbo, et al.
Veröffentlicht: (2025)
von: Hu, Songbo, et al.
Veröffentlicht: (2025)
A Grounded Typology of Word Classes
von: Haley, Coleman, et al.
Veröffentlicht: (2024)
von: Haley, Coleman, et al.
Veröffentlicht: (2024)
Specialising and Analysing Instruction-Tuned and Byte-Level Language Models for Organic Reaction Prediction
von: Pang, Jiayun, et al.
Veröffentlicht: (2024)
von: Pang, Jiayun, et al.
Veröffentlicht: (2024)
Generalization Measures for Zero-Shot Cross-Lingual Transfer
von: Bassi, Saksham, et al.
Veröffentlicht: (2024)
von: Bassi, Saksham, et al.
Veröffentlicht: (2024)
Zero-Shot Detection of LLM-Generated Text using Token Cohesiveness
von: Ma, Shixuan, et al.
Veröffentlicht: (2024)
von: Ma, Shixuan, et al.
Veröffentlicht: (2024)
Is Information Density Uniform when Utterances are Grounded on Perception and Discourse?
von: Gay, Matteo, et al.
Veröffentlicht: (2026)
von: Gay, Matteo, et al.
Veröffentlicht: (2026)
Languages Transferred Within the Encoder: On Representation Transfer in Zero-Shot Multilingual Translation
von: Qu, Zhi, et al.
Veröffentlicht: (2024)
von: Qu, Zhi, et al.
Veröffentlicht: (2024)
Tokenization Matters: Improving Zero-Shot NER for Indic Languages
von: Pattnayak, Priyaranjan, et al.
Veröffentlicht: (2025)
von: Pattnayak, Priyaranjan, et al.
Veröffentlicht: (2025)
Beyond the Next Token: Towards Prompt-Robust Zero-Shot Classification via Efficient Multi-Token Prediction
von: Qian, Junlang, et al.
Veröffentlicht: (2025)
von: Qian, Junlang, et al.
Veröffentlicht: (2025)
Post-hoc Reward Calibration: A Case Study on Length Bias
von: Huang, Zeyu, et al.
Veröffentlicht: (2024)
von: Huang, Zeyu, et al.
Veröffentlicht: (2024)
Donors and Recipients: On Asymmetric Transfer Across Tasks and Languages with Parameter-Efficient Fine-Tuning
von: Dymkiewicz, Kajetan, et al.
Veröffentlicht: (2025)
von: Dymkiewicz, Kajetan, et al.
Veröffentlicht: (2025)
Fine-tuning Large Language Models with Sequential Instructions
von: Hu, Hanxu, et al.
Veröffentlicht: (2024)
von: Hu, Hanxu, et al.
Veröffentlicht: (2024)
SQATIN: Supervised Instruction Tuning Meets Question Answering for Improved Dialogue NLU
von: Razumovskaia, Evgeniia, et al.
Veröffentlicht: (2023)
von: Razumovskaia, Evgeniia, et al.
Veröffentlicht: (2023)
Polyglot Teachers: Evaluating Language Models for Multilingual Synthetic Data Generation
von: Miranda, Lester James V., et al.
Veröffentlicht: (2026)
von: Miranda, Lester James V., et al.
Veröffentlicht: (2026)
Probing the Emergence of Cross-lingual Alignment during LLM Training
von: Wang, Hetong, et al.
Veröffentlicht: (2024)
von: Wang, Hetong, et al.
Veröffentlicht: (2024)
Layer Swapping for Zero-Shot Cross-Lingual Transfer in Large Language Models
von: Bandarkar, Lucas, et al.
Veröffentlicht: (2024)
von: Bandarkar, Lucas, et al.
Veröffentlicht: (2024)
Model-Based Ranking of Source Languages for Zero-Shot Cross-Lingual Transfer
von: Ebrahimi, Abteen, et al.
Veröffentlicht: (2025)
von: Ebrahimi, Abteen, et al.
Veröffentlicht: (2025)
Self-Distillation for Model Stacking Unlocks Cross-Lingual NLU in 200+ Languages
von: Schmidt, Fabian David, et al.
Veröffentlicht: (2024)
von: Schmidt, Fabian David, et al.
Veröffentlicht: (2024)
ReCoVeR the Target Language: Language Steering without Sacrificing Task Performance
von: Sterz, Hannah, et al.
Veröffentlicht: (2025)
von: Sterz, Hannah, et al.
Veröffentlicht: (2025)
TTSDS2: Resources and Benchmark for Evaluating Human-Quality Text to Speech Systems
von: Minixhofer, Christoph, et al.
Veröffentlicht: (2025)
von: Minixhofer, Christoph, et al.
Veröffentlicht: (2025)
Inference-Time Hyper-Scaling with KV Cache Compression
von: Łańcucki, Adrian, et al.
Veröffentlicht: (2025)
von: Łańcucki, Adrian, et al.
Veröffentlicht: (2025)
Reuse Your Rewards: Reward Model Transfer for Zero-Shot Cross-Lingual Alignment
von: Wu, Zhaofeng, et al.
Veröffentlicht: (2024)
von: Wu, Zhaofeng, et al.
Veröffentlicht: (2024)
Parallel-SFT: Improving Zero-Shot Cross-Programming-Language Transfer for Code RL
von: Wu, Zhaofeng, et al.
Veröffentlicht: (2026)
von: Wu, Zhaofeng, et al.
Veröffentlicht: (2026)
Just Pass Twice: Efficient Token Classification with LLMs for Zero-Shot NER
von: Ewais, Ahmed, et al.
Veröffentlicht: (2026)
von: Ewais, Ahmed, et al.
Veröffentlicht: (2026)
Dynamic Memory Compression: Retrofitting LLMs for Accelerated Inference
von: Nawrot, Piotr, et al.
Veröffentlicht: (2024)
von: Nawrot, Piotr, et al.
Veröffentlicht: (2024)
Modular Deep Learning
von: Pfeiffer, Jonas, et al.
Veröffentlicht: (2023)
von: Pfeiffer, Jonas, et al.
Veröffentlicht: (2023)
Beyond Oversmoothing: Evaluating DDPM and MSE for Scalable Speech Synthesis in ASR
von: Minixhofer, Christoph, et al.
Veröffentlicht: (2024)
von: Minixhofer, Christoph, et al.
Veröffentlicht: (2024)
Self-Augmented In-Context Learning for Unsupervised Word Translation
von: Li, Yaoyiran, et al.
Veröffentlicht: (2024)
von: Li, Yaoyiran, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Universal Cross-Tokenizer Distillation via Approximate Likelihood Matching
von: Minixhofer, Benjamin, et al.
Veröffentlicht: (2025) -
Retrofitting Large Language Models with Dynamic Tokenization
von: Feher, Darius, et al.
Veröffentlicht: (2024) -
Emergent Communication Pretraining for Few-Shot Machine Translation
von: Li, Yaoyiran, et al.
Veröffentlicht: (2020) -
Segment Any Text: A Universal Approach for Robust, Efficient and Adaptable Sentence Segmentation
von: Frohmann, Markus, et al.
Veröffentlicht: (2024) -
Scaling Sparse Fine-Tuning to Large Language Models
von: Ansell, Alan, et al.
Veröffentlicht: (2024)