Correlation and Navigation in the Vocabulary Key Representation Space of Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Peng, Letian, An, Chenyang, Shang, Jingbo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Linear Correlation in LM's Compositional Generalization and Hallucination
by: Peng, Letian, et al.
Published: (2025)
by: Peng, Letian, et al.
Published: (2025)
Incubating Text Classifiers Following User Instruction with Nothing but LLM
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
Quantifying and Optimizing Global Faithfulness in Persona-driven Role-playing
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
Codifying Character Logic in Role-Playing
by: Peng, Letian, et al.
Published: (2025)
by: Peng, Letian, et al.
Published: (2025)
UltraGen: Extremely Fine-grained Controllable Generation via Attribute Reconstruction and Global Preference Optimization
by: Yun, Longfei, et al.
Published: (2025)
by: Yun, Longfei, et al.
Published: (2025)
Controllable Data Augmentation for Few-Shot Text Mining with Chain-of-Thought Attribute Manipulation
by: Peng, Letian, et al.
Published: (2023)
by: Peng, Letian, et al.
Published: (2023)
The Price of Format: Diversity Collapse in LLMs
by: Yun, Longfei, et al.
Published: (2025)
by: Yun, Longfei, et al.
Published: (2025)
Cuckoo: An IE Free Rider Hatched by Massive Nutrition in LLM's Nest
by: Peng, Letian, et al.
Published: (2025)
by: Peng, Letian, et al.
Published: (2025)
Codified Finite-state Machines for Role-playing
by: Peng, Letian, et al.
Published: (2026)
by: Peng, Letian, et al.
Published: (2026)
MetaIE: Distilling a Meta Model from LLM for All Kinds of Information Extraction Tasks
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
Text Grafting: Near-Distribution Weak Supervision for Minority Classes in Text Classification
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
Deriving Character Logic from Storyline as Codified Decision Trees
by: Peng, Letian, et al.
Published: (2026)
by: Peng, Letian, et al.
Published: (2026)
Codified Foreshadowing-Payoff Text Generation
by: Yun, Longfei, et al.
Published: (2026)
by: Yun, Longfei, et al.
Published: (2026)
Dual-Space Knowledge Distillation with Key-Query Matching for Large Language Models with Vocabulary Mismatch
by: Tsiapali, Stella Eva, et al.
Published: (2026)
by: Tsiapali, Stella Eva, et al.
Published: (2026)
Watermarks for Language Models via Probabilistic Automata
by: Wang, Yangkun, et al.
Published: (2025)
by: Wang, Yangkun, et al.
Published: (2025)
Smaller Language Models are capable of selecting Instruction-Tuning Training Data for Larger Language Models
by: Mekala, Dheeraj, et al.
Published: (2024)
by: Mekala, Dheeraj, et al.
Published: (2024)
Answer is All You Need: Instruction-following Text Embedding via Answering the Question
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
BOOKMARKS: Efficient Active Storyline Memory for Role-playing
by: Peng, Letian, et al.
Published: (2026)
by: Peng, Letian, et al.
Published: (2026)
Model-diff: A Tool for Comparative Study of Language Models in the Input Space
by: Liu, Weitang, et al.
Published: (2024)
by: Liu, Weitang, et al.
Published: (2024)
Bidirectional LMs are Better Knowledge Memorizers? A Benchmark for Real-world Knowledge Injection
by: Zhang, Yuwei, et al.
Published: (2025)
by: Zhang, Yuwei, et al.
Published: (2025)
Navigating the Concept Space of Language Models
by: Marcílio-Jr, Wilson E., et al.
Published: (2026)
by: Marcílio-Jr, Wilson E., et al.
Published: (2026)
Simulating Organized Group Behavior: New Framework, Benchmark, and Analysis
by: Zou, Xinkai, et al.
Published: (2026)
by: Zou, Xinkai, et al.
Published: (2026)
Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
by: Zhong, Li, et al.
Published: (2024)
by: Zhong, Li, et al.
Published: (2024)
Backward Lens: Projecting Language Model Gradients into the Vocabulary Space
by: Katz, Shahar, et al.
Published: (2024)
by: Katz, Shahar, et al.
Published: (2024)
Vector-ICL: In-context Learning with Continuous Vector Representations
by: Zhuang, Yufan, et al.
Published: (2024)
by: Zhuang, Yufan, et al.
Published: (2024)
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
by: Wang, Zhaowei, et al.
Published: (2024)
by: Wang, Zhaowei, et al.
Published: (2024)
Navigating Brain Language Representations: A Comparative Analysis of Neural Language Models and Psychologically Plausible Models
by: Zhang, Yunhao, et al.
Published: (2024)
by: Zhang, Yunhao, et al.
Published: (2024)
Entangled Relations: Leveraging NLI and Meta-analysis to Enhance Biomedical Relation Extraction
by: Hogan, William, et al.
Published: (2024)
by: Hogan, William, et al.
Published: (2024)
Overcoming Vocabulary Mismatch: Vocabulary-agnostic Teacher Guided Language Modeling
by: Shin, Haebin, et al.
Published: (2025)
by: Shin, Haebin, et al.
Published: (2025)
Disentangling MLP Neuron Weights in Vocabulary Space
by: Avrahamy, Asaf, et al.
Published: (2026)
by: Avrahamy, Asaf, et al.
Published: (2026)
The Tokenization Bottleneck: How Vocabulary Extension Improves Chemistry Representation Learning in Pretrained Language Models
by: Kalamkar, Prathamesh, et al.
Published: (2025)
by: Kalamkar, Prathamesh, et al.
Published: (2025)
Blind Spot Navigation in Large Language Model Reasoning with Thought Space Explorer
by: Zhang, Jinghan, et al.
Published: (2024)
by: Zhang, Jinghan, et al.
Published: (2024)
Large Language Models for Time Series: A Survey
by: Zhang, Xiyuan, et al.
Published: (2024)
by: Zhang, Xiyuan, et al.
Published: (2024)
Finish First, Perfect Later: Test-Time Token-Level Cross-Validation for Diffusion Large Language Models
by: Tian, Runchu, et al.
Published: (2025)
by: Tian, Runchu, et al.
Published: (2025)
Large Concept Models: Language Modeling in a Sentence Representation Space
by: LCM team, et al.
Published: (2024)
by: LCM team, et al.
Published: (2024)
REVS: Unlearning Sensitive Information in Language Models via Rank Editing in the Vocabulary Space
by: Ashuach, Tomer, et al.
Published: (2024)
by: Ashuach, Tomer, et al.
Published: (2024)
Learning Distribution-Wise Control in Representation Space for Language Models
by: Deng, Chunyuan, et al.
Published: (2025)
by: Deng, Chunyuan, et al.
Published: (2025)
Large Vocabulary Size Improves Large Language Models
by: Takase, Sho, et al.
Published: (2024)
by: Takase, Sho, et al.
Published: (2024)
Evaluating the Smooth Control of Attribute Intensity in Text Generation with LLMs
by: Zhou, Shang, et al.
Published: (2024)
by: Zhou, Shang, et al.
Published: (2024)
Vocabulary-level Memory Efficiency for Language Model Fine-tuning
by: Williams, Miles, et al.
Published: (2023)
by: Williams, Miles, et al.
Published: (2023)
Similar Items
-
Linear Correlation in LM's Compositional Generalization and Hallucination
by: Peng, Letian, et al.
Published: (2025) -
Incubating Text Classifiers Following User Instruction with Nothing but LLM
by: Peng, Letian, et al.
Published: (2024) -
Quantifying and Optimizing Global Faithfulness in Persona-driven Role-playing
by: Peng, Letian, et al.
Published: (2024) -
Codifying Character Logic in Role-Playing
by: Peng, Letian, et al.
Published: (2025) -
UltraGen: Extremely Fine-grained Controllable Generation via Attribute Reconstruction and Global Preference Optimization
by: Yun, Longfei, et al.
Published: (2025)