Generation with Dynamic Vocabulary
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Yanting, Ji, Tao, Sun, Changzhi, Wu, Yuanbin, Wang, Xiaoling |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DVAGen: Dynamic Vocabulary Augmented Generation
by: Du, Wei, et al.
Published: (2025)
by: Du, Wei, et al.
Published: (2025)
Stabilizing LLM Supervised Fine-Tuning via Explicit Distributional Control
by: Wang, Xinyu, et al.
Published: (2026)
by: Wang, Xinyu, et al.
Published: (2026)
Protein Design with Dynamic Protein Vocabulary
by: Liu, Nuowei, et al.
Published: (2025)
by: Liu, Nuowei, et al.
Published: (2025)
Logic-Regularized Verifier Elicits Reasoning from LLMs
by: Wang, Xinyu, et al.
Published: (2026)
by: Wang, Xinyu, et al.
Published: (2026)
Length Generalization of Causal Transformers without Position Encoding
by: Wang, Jie, et al.
Published: (2024)
by: Wang, Jie, et al.
Published: (2024)
CERD: A Comprehensive Chinese Rhetoric Dataset for Rhetorical Understanding and Generation in Essays
by: Liu, Nuowei, et al.
Published: (2024)
by: Liu, Nuowei, et al.
Published: (2024)
Investigating and Mitigating Object Hallucinations in Pretrained Vision-Language (CLIP) Models
by: Liu, Yufang, et al.
Published: (2024)
by: Liu, Yufang, et al.
Published: (2024)
TASO: Task-Aligned Sparse Optimization for Parameter-Efficient Model Adaptation
by: Miao, Daiye, et al.
Published: (2025)
by: Miao, Daiye, et al.
Published: (2025)
AntLM: Bridging Causal and Masked Language Models
by: Yu, Xinru, et al.
Published: (2024)
by: Yu, Xinru, et al.
Published: (2024)
TCMBench: A Comprehensive Benchmark for Evaluating Large Language Models in Traditional Chinese Medicine
by: Yue, Wenjing, et al.
Published: (2024)
by: Yue, Wenjing, et al.
Published: (2024)
Gold Panning in Vocabulary: An Adaptive Method for Vocabulary Expansion of Domain-Specific LLMs
by: Liu, Chengyuan, et al.
Published: (2024)
by: Liu, Chengyuan, et al.
Published: (2024)
Predicting Contextual Informativeness for Vocabulary Learning using Deep Learning
by: Wu, Tao, et al.
Published: (2026)
by: Wu, Tao, et al.
Published: (2026)
Text2MDT: Extracting Medical Decision Trees from Medical Texts
by: Zhu, Wei, et al.
Published: (2024)
by: Zhu, Wei, et al.
Published: (2024)
Overcoming Vocabulary Mismatch: Vocabulary-agnostic Teacher Guided Language Modeling
by: Shin, Haebin, et al.
Published: (2025)
by: Shin, Haebin, et al.
Published: (2025)
Beyond Heuristics: A Decision-Theoretic Framework for Agent Memory Management
by: Sun, Changzhi, et al.
Published: (2025)
by: Sun, Changzhi, et al.
Published: (2025)
Scaling Laws with Vocabulary: Larger Models Deserve Larger Vocabularies
by: Tao, Chaofan, et al.
Published: (2024)
by: Tao, Chaofan, et al.
Published: (2024)
PDFBench: A Benchmark for De novo Protein Design from Function
by: Kuang, Jiahao, et al.
Published: (2025)
by: Kuang, Jiahao, et al.
Published: (2025)
Over-Tokenized Transformer: Vocabulary is Generally Worth Scaling
by: Huang, Hongzhi, et al.
Published: (2025)
by: Huang, Hongzhi, et al.
Published: (2025)
Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-based LLMs
by: Ji, Tao, et al.
Published: (2025)
by: Ji, Tao, et al.
Published: (2025)
ISSR: Iterative Selection with Self-Review for Vocabulary Test Distractor Generation
by: Liu, Yu-Cheng, et al.
Published: (2025)
by: Liu, Yu-Cheng, et al.
Published: (2025)
WordCraft: Scaffolding the Keyword Method for L2 Vocabulary Learning with Multimodal LLMs
by: Shao, Yuheng, et al.
Published: (2026)
by: Shao, Yuheng, et al.
Published: (2026)
The CoNLL-2013 Shared Task on Grammatical Error Correction
by: Ng, Hwee Tou, et al.
Published: (2025)
by: Ng, Hwee Tou, et al.
Published: (2025)
On Support Samples of Next Word Prediction
by: Li, Yuqian, et al.
Published: (2025)
by: Li, Yuqian, et al.
Published: (2025)
ANALOGYKB: Unlocking Analogical Reasoning of Language Models with A Million-scale Knowledge Base
by: Yuan, Siyu, et al.
Published: (2023)
by: Yuan, Siyu, et al.
Published: (2023)
DualToken: Towards Unifying Visual Understanding and Generation with Dual Visual Vocabularies
by: Song, Wei, et al.
Published: (2025)
by: Song, Wei, et al.
Published: (2025)
The Polish Vocabulary Size Test: A Novel Adaptive Test for Receptive Vocabulary Assessment
by: Fokin, Danil, et al.
Published: (2025)
by: Fokin, Danil, et al.
Published: (2025)
Dynamic Vocabulary Pruning in Early-Exit LLMs
by: Vincenti, Jort, et al.
Published: (2024)
by: Vincenti, Jort, et al.
Published: (2024)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary
by: Sudo, Yui, et al.
Published: (2024)
by: Sudo, Yui, et al.
Published: (2024)
LEVOS: Leveraging Vocabulary Overlap with Sanskrit to Generate Technical Lexicons in Indian Languages
by: J, Karthika N, et al.
Published: (2024)
by: J, Karthika N, et al.
Published: (2024)
Second Language (Arabic) Acquisition of LLMs via Progressive Vocabulary Expansion
by: Zhu, Jianqing, et al.
Published: (2024)
by: Zhu, Jianqing, et al.
Published: (2024)
Automated Generation of Multiple-Choice Cloze Questions for Assessing English Vocabulary Using GPT-turbo 3.5
by: Wang, Qiao, et al.
Published: (2024)
by: Wang, Qiao, et al.
Published: (2024)
Grounded Token Initialization for New Vocabulary in LMs for Generative Recommendation
by: Chen, Daiwei, et al.
Published: (2026)
by: Chen, Daiwei, et al.
Published: (2026)
Analysing The Impact of Sequence Composition on Language Model Pre-Training
by: Zhao, Yu, et al.
Published: (2024)
by: Zhao, Yu, et al.
Published: (2024)
UniC-RAG: Universal Knowledge Corruption Attacks to Retrieval-Augmented Generation
by: Geng, Runpeng, et al.
Published: (2025)
by: Geng, Runpeng, et al.
Published: (2025)
Boosting Large Language Models with Continual Learning for Aspect-based Sentiment Analysis
by: Ding, Xuanwen, et al.
Published: (2024)
by: Ding, Xuanwen, et al.
Published: (2024)
Open Vocabulary Panoptic Segmentation With Retrieval Augmentation
by: Sadeq, Nafis, et al.
Published: (2026)
by: Sadeq, Nafis, et al.
Published: (2026)
Fine-Grained Detoxification via Instance-Level Prefixes for Large Language Models
by: Yi, Xin, et al.
Published: (2024)
by: Yi, Xin, et al.
Published: (2024)
Beyond Film Subtitles: Is YouTube the Best Approximation of Spoken Vocabulary?
by: Nohejl, Adam, et al.
Published: (2024)
by: Nohejl, Adam, et al.
Published: (2024)
Speculative Decoding with a Speculative Vocabulary
by: Williams, Miles, et al.
Published: (2026)
by: Williams, Miles, et al.
Published: (2026)
EVOKE: Emotion Vocabulary Of Korean and English
by: Jung, Yoonwon, et al.
Published: (2026)
by: Jung, Yoonwon, et al.
Published: (2026)
Similar Items
-
DVAGen: Dynamic Vocabulary Augmented Generation
by: Du, Wei, et al.
Published: (2025) -
Stabilizing LLM Supervised Fine-Tuning via Explicit Distributional Control
by: Wang, Xinyu, et al.
Published: (2026) -
Protein Design with Dynamic Protein Vocabulary
by: Liu, Nuowei, et al.
Published: (2025) -
Logic-Regularized Verifier Elicits Reasoning from LLMs
by: Wang, Xinyu, et al.
Published: (2026) -
Length Generalization of Causal Transformers without Position Encoding
by: Wang, Jie, et al.
Published: (2024)