Overcoming Vocabulary Constraints with Pixel-level Fallback
Fuente:
arXiv
Salvato in:
| Autori principali: | Lotz, Jonas F., Setiawan, Hendra, Peitz, Stephan, Kementchedjhieva, Yova |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond Text Compression: Evaluating Tokenizers Across Scales
di: Lotz, Jonas F., et al.
Pubblicazione: (2025)
di: Lotz, Jonas F., et al.
Pubblicazione: (2025)
The Devil is in the EOS: Sequence Training for Detailed Image Captioning
di: Mohamed, Abdelrahman, et al.
Pubblicazione: (2025)
di: Mohamed, Abdelrahman, et al.
Pubblicazione: (2025)
Accurate Knowledge Distillation with n-best Reranking
di: Setiawan, Hendra
Pubblicazione: (2023)
di: Setiawan, Hendra
Pubblicazione: (2023)
CONCAP: Seeing Beyond English with Concepts Retrieval-Augmented Captioning
di: Ibrahim, George, et al.
Pubblicazione: (2025)
di: Ibrahim, George, et al.
Pubblicazione: (2025)
Answerability in Retrieval-Augmented Open-Domain Question Answering
di: Abdumalikov, Rustam, et al.
Pubblicazione: (2024)
di: Abdumalikov, Rustam, et al.
Pubblicazione: (2024)
SPECS: Specificity-Enhanced CLIP-Score for Long Image Caption Evaluation
di: Chen, Xiaofu, et al.
Pubblicazione: (2025)
di: Chen, Xiaofu, et al.
Pubblicazione: (2025)
Multimodal Large Language Models to Support Real-World Fact-Checking
di: Geng, Jiahui, et al.
Pubblicazione: (2024)
di: Geng, Jiahui, et al.
Pubblicazione: (2024)
A Discriminative Latent-Variable Model for Bilingual Lexicon Induction
di: Ruder, Sebastian, et al.
Pubblicazione: (2018)
di: Ruder, Sebastian, et al.
Pubblicazione: (2018)
Do Vision and Language Models Share Concepts? A Vector Space Alignment Study
di: Li, Jiaang, et al.
Pubblicazione: (2023)
di: Li, Jiaang, et al.
Pubblicazione: (2023)
MuLan: A Study of Fact Mutability in Language Models
di: Fierro, Constanza, et al.
Pubblicazione: (2024)
di: Fierro, Constanza, et al.
Pubblicazione: (2024)
A Simple Data Augmentation Strategy for Text-in-Image Scientific VQA
di: Shoer, Belal, et al.
Pubblicazione: (2025)
di: Shoer, Belal, et al.
Pubblicazione: (2025)
EFSA: Episodic Few-Shot Adaptation for Text-to-Image Retrieval
di: Huzaifa, Muhammad, et al.
Pubblicazione: (2024)
di: Huzaifa, Muhammad, et al.
Pubblicazione: (2024)
LLMs Can Compensate for Deficiencies in Visual Representations
di: Takishita, Sho, et al.
Pubblicazione: (2025)
di: Takishita, Sho, et al.
Pubblicazione: (2025)
Multilingual Pretraining for Pixel Language Models
di: Kesen, Ilker, et al.
Pubblicazione: (2025)
di: Kesen, Ilker, et al.
Pubblicazione: (2025)
LinguDistill: Recovering Linguistic Ability in Vision-Language Models via Selective Cross-Modal Distillation
di: Irawan, Patrick Amadeus, et al.
Pubblicazione: (2026)
di: Irawan, Patrick Amadeus, et al.
Pubblicazione: (2026)
Long Story Short: Disentangling Compositionality and Long-Caption Understanding in Contrastive VLMs
di: Salazar, Israfel, et al.
Pubblicazione: (2025)
di: Salazar, Israfel, et al.
Pubblicazione: (2025)
VLMs Need Words: Vision Language Models Ignore Visual Detail In Favor of Semantic Anchors
di: Shahgir, Haz Sameen, et al.
Pubblicazione: (2026)
di: Shahgir, Haz Sameen, et al.
Pubblicazione: (2026)
Overcoming Vocabulary Mismatch: Vocabulary-agnostic Teacher Guided Language Modeling
di: Shin, Haebin, et al.
Pubblicazione: (2025)
di: Shin, Haebin, et al.
Pubblicazione: (2025)
Safer Policy Compliance with Dynamic Epistemic Fallback
di: Imperial, Joseph Marvin, et al.
Pubblicazione: (2026)
di: Imperial, Joseph Marvin, et al.
Pubblicazione: (2026)
JEEM: Vision-Language Understanding in Four Arabic Dialects
di: Kadaoui, Karima, et al.
Pubblicazione: (2025)
di: Kadaoui, Karima, et al.
Pubblicazione: (2025)
The Role of Data Curation in Image Captioning
di: Li, Wenyan, et al.
Pubblicazione: (2023)
di: Li, Wenyan, et al.
Pubblicazione: (2023)
From Loops to Oops: Fallback Behaviors of Language Models Under Uncertainty
di: Ivgi, Maor, et al.
Pubblicazione: (2024)
di: Ivgi, Maor, et al.
Pubblicazione: (2024)
DualEdit: Mitigating Safety Fallback in LLM Backdoor Editing via Affirmation-Refusal Regulation
di: Jiang, Houcheng, et al.
Pubblicazione: (2025)
di: Jiang, Houcheng, et al.
Pubblicazione: (2025)
From Fallback to Frontline: When Can LLMs be Superior Annotators of Human Perspectives?
di: Amin, Hasan, et al.
Pubblicazione: (2026)
di: Amin, Hasan, et al.
Pubblicazione: (2026)
Token-level Ensembling of Models with Different Vocabularies
di: Wicks, Rachel, et al.
Pubblicazione: (2025)
di: Wicks, Rachel, et al.
Pubblicazione: (2025)
Building Consumer Loyalty: Understanding E-Satisfaction in Fast-Fashion Purchases on E-Commerce Platform
di: Wijaya, Hendra, et al.
Pubblicazione: (2025)
di: Wijaya, Hendra, et al.
Pubblicazione: (2025)
Vocabulary-level Memory Efficiency for Language Model Fine-tuning
di: Williams, Miles, et al.
Pubblicazione: (2023)
di: Williams, Miles, et al.
Pubblicazione: (2023)
Noise is an Efficient Learner for Zero-Shot Vision-Language Models
di: Imam, Raza, et al.
Pubblicazione: (2025)
di: Imam, Raza, et al.
Pubblicazione: (2025)
Trivial Vocabulary Bans Improve LLM Reasoning More Than Deep Linguistic Constraints
di: Jehu-Appiah, Rodney
Pubblicazione: (2026)
di: Jehu-Appiah, Rodney
Pubblicazione: (2026)
Generation with Dynamic Vocabulary
di: Liu, Yanting, et al.
Pubblicazione: (2024)
di: Liu, Yanting, et al.
Pubblicazione: (2024)
Gold Panning in Vocabulary: An Adaptive Method for Vocabulary Expansion of Domain-Specific LLMs
di: Liu, Chengyuan, et al.
Pubblicazione: (2024)
di: Liu, Chengyuan, et al.
Pubblicazione: (2024)
The Polish Vocabulary Size Test: A Novel Adaptive Test for Receptive Vocabulary Assessment
di: Fokin, Danil, et al.
Pubblicazione: (2025)
di: Fokin, Danil, et al.
Pubblicazione: (2025)
NusaBERT: Teaching IndoBERT to be Multilingual and Multicultural
di: Wongso, Wilson, et al.
Pubblicazione: (2024)
di: Wongso, Wilson, et al.
Pubblicazione: (2024)
Vocabulary shapes cross-lingual variation of word-order learnability in language models
di: Martins, Jonas Mayer, et al.
Pubblicazione: (2026)
di: Martins, Jonas Mayer, et al.
Pubblicazione: (2026)
Scaling Laws with Vocabulary: Larger Models Deserve Larger Vocabularies
di: Tao, Chaofan, et al.
Pubblicazione: (2024)
di: Tao, Chaofan, et al.
Pubblicazione: (2024)
Speculative Decoding with a Speculative Vocabulary
di: Williams, Miles, et al.
Pubblicazione: (2026)
di: Williams, Miles, et al.
Pubblicazione: (2026)
EVOKE: Emotion Vocabulary Of Korean and English
di: Jung, Yoonwon, et al.
Pubblicazione: (2026)
di: Jung, Yoonwon, et al.
Pubblicazione: (2026)
ASVRI-Legal: Fine-Tuning LLMs with Retrieval Augmented Generation for Enhanced Legal Regulation
di: Octadion, One, et al.
Pubblicazione: (2025)
di: Octadion, One, et al.
Pubblicazione: (2025)
PixelWorld: How Far Are We from Perceiving Everything as Pixels?
di: Lyu, Zhiheng, et al.
Pubblicazione: (2025)
di: Lyu, Zhiheng, et al.
Pubblicazione: (2025)
LLM Distillation for Efficient Few-Shot Multiple Choice Question Answering
di: Sutanto, Patrick, et al.
Pubblicazione: (2024)
di: Sutanto, Patrick, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Beyond Text Compression: Evaluating Tokenizers Across Scales
di: Lotz, Jonas F., et al.
Pubblicazione: (2025) -
The Devil is in the EOS: Sequence Training for Detailed Image Captioning
di: Mohamed, Abdelrahman, et al.
Pubblicazione: (2025) -
Accurate Knowledge Distillation with n-best Reranking
di: Setiawan, Hendra
Pubblicazione: (2023) -
CONCAP: Seeing Beyond English with Concepts Retrieval-Augmented Captioning
di: Ibrahim, George, et al.
Pubblicazione: (2025) -
Answerability in Retrieval-Augmented Open-Domain Question Answering
di: Abdumalikov, Rustam, et al.
Pubblicazione: (2024)