Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Christine, Jurafsky, Dan, Shani, Chen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond Tokens: Concept-Level Training Objectives for LLMs
von: Iyer, Laya, et al.
Veröffentlicht: (2026)
von: Iyer, Laya, et al.
Veröffentlicht: (2026)
The Roots of Performance Disparity in Multilingual Language Models: Intrinsic Modeling Difficulty or Design Choices?
von: Shani, Chen, et al.
Veröffentlicht: (2026)
von: Shani, Chen, et al.
Veröffentlicht: (2026)
Rethinking Word Similarity: Semantic Similarity through Classification Confusion
von: Zhou, Kaitlyn, et al.
Veröffentlicht: (2025)
von: Zhou, Kaitlyn, et al.
Veröffentlicht: (2025)
From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning
von: Shani, Chen, et al.
Veröffentlicht: (2025)
von: Shani, Chen, et al.
Veröffentlicht: (2025)
Cooking Up Creativity: Enhancing LLM Creativity through Structured Recombination
von: Mizrahi, Moran, et al.
Veröffentlicht: (2025)
von: Mizrahi, Moran, et al.
Veröffentlicht: (2025)
STARS: Synchronous Token Alignment for Robust Supervision in Large Language Models
von: Quamar, Mohammad Atif, et al.
Veröffentlicht: (2025)
von: Quamar, Mohammad Atif, et al.
Veröffentlicht: (2025)
False Friends Are Not Foes: Investigating Vocabulary Overlap in Multilingual Language Models
von: Kallini, Julie, et al.
Veröffentlicht: (2025)
von: Kallini, Julie, et al.
Veröffentlicht: (2025)
Transcribe, Translate, or Transliterate: An Investigation of Intermediate Representations in Spoken Language Models
von: Ògúnrèmí, Tolúlopé, et al.
Veröffentlicht: (2025)
von: Ògúnrèmí, Tolúlopé, et al.
Veröffentlicht: (2025)
Persona-judge: Personalized Alignment of Large Language Models via Token-level Self-judgment
von: Zhang, Xiaotian, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaotian, et al.
Veröffentlicht: (2025)
A Benchmark for Learning to Translate a New Language from One Grammar Book
von: Tanzer, Garrett, et al.
Veröffentlicht: (2023)
von: Tanzer, Garrett, et al.
Veröffentlicht: (2023)
A layer-wise analysis of Mandarin and English suprasegmentals in SSL speech models
von: de la Fuente, Antón, et al.
Veröffentlicht: (2024)
von: de la Fuente, Antón, et al.
Veröffentlicht: (2024)
Grounding Gaps in Language Model Generations
von: Shaikh, Omar, et al.
Veröffentlicht: (2023)
von: Shaikh, Omar, et al.
Veröffentlicht: (2023)
SumTablets: A Transliteration Dataset of Sumerian Tablets
von: Simmons, Cole, et al.
Veröffentlicht: (2026)
von: Simmons, Cole, et al.
Veröffentlicht: (2026)
HumT DumT: Measuring and controlling human-like language in LLMs
von: Cheng, Myra, et al.
Veröffentlicht: (2025)
von: Cheng, Myra, et al.
Veröffentlicht: (2025)
Categorize Early, Integrate Late: Divergent Processing Strategies in Automatic Speech Recognition
von: Roll, Nathan, et al.
Veröffentlicht: (2026)
von: Roll, Nathan, et al.
Veröffentlicht: (2026)
Othering and low status framing of immigrant cuisines in US restaurant reviews and large language models
von: Luo, Yiwei, et al.
Veröffentlicht: (2023)
von: Luo, Yiwei, et al.
Veröffentlicht: (2023)
Adapting Pretrained Language Models for Citation Classification via Self-Supervised Contrastive Learning
von: Li, Tong, et al.
Veröffentlicht: (2025)
von: Li, Tong, et al.
Veröffentlicht: (2025)
SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
Data Checklist: On Unit-Testing Datasets with Usable Information
von: Zhang, Heidi C., et al.
Veröffentlicht: (2024)
von: Zhang, Heidi C., et al.
Veröffentlicht: (2024)
Humans overrely on overconfident language models, across languages
von: Rathi, Neil, et al.
Veröffentlicht: (2025)
von: Rathi, Neil, et al.
Veröffentlicht: (2025)
Breaking Token Into Concepts: Exploring Extreme Compression in Token Representation Via Compositional Shared Semantics
von: R V, Kavin, et al.
Veröffentlicht: (2025)
von: R V, Kavin, et al.
Veröffentlicht: (2025)
Safety-Tuned LLaMAs: Lessons From Improving the Safety of Large Language Models that Follow Instructions
von: Bianchi, Federico, et al.
Veröffentlicht: (2023)
von: Bianchi, Federico, et al.
Veröffentlicht: (2023)
Accommodation and Epistemic Vigilance: A Pragmatic Account of Why LLMs Fail to Challenge Harmful Beliefs
von: Cheng, Myra, et al.
Veröffentlicht: (2026)
von: Cheng, Myra, et al.
Veröffentlicht: (2026)
A Comprehensive Analysis of Tokenization and Self-Supervised Learning in End-to-End Automatic Speech Recognition applied on French Language
von: Bañeras-Roux, Thibault, et al.
Veröffentlicht: (2026)
von: Bañeras-Roux, Thibault, et al.
Veröffentlicht: (2026)
AlignDistil: Token-Level Language Model Alignment as Adaptive Policy Distillation
von: Zhang, Songming, et al.
Veröffentlicht: (2025)
von: Zhang, Songming, et al.
Veröffentlicht: (2025)
Dynamic Cheatsheet: Test-Time Learning with Adaptive Memory
von: Suzgun, Mirac, et al.
Veröffentlicht: (2025)
von: Suzgun, Mirac, et al.
Veröffentlicht: (2025)
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
Evaluating Morphological Alignment of Tokenizers in 70 Languages
von: Arnett, Catherine, et al.
Veröffentlicht: (2025)
von: Arnett, Catherine, et al.
Veröffentlicht: (2025)
Belief in the Machine: Investigating Epistemological Blind Spots of Language Models
von: Suzgun, Mirac, et al.
Veröffentlicht: (2024)
von: Suzgun, Mirac, et al.
Veröffentlicht: (2024)
Probabilistic Token Alignment for Large Language Model Fusion
von: Zeng, Runjia, et al.
Veröffentlicht: (2025)
von: Zeng, Runjia, et al.
Veröffentlicht: (2025)
CausalGym: Benchmarking causal interpretability methods on linguistic tasks
von: Arora, Aryaman, et al.
Veröffentlicht: (2024)
von: Arora, Aryaman, et al.
Veröffentlicht: (2024)
SeMe: Training-Free Language Model Merging via Semantic Alignment
von: Gu, Jian, et al.
Veröffentlicht: (2025)
von: Gu, Jian, et al.
Veröffentlicht: (2025)
Self-Supervised Visual Preference Alignment
von: Zhu, Ke, et al.
Veröffentlicht: (2024)
von: Zhu, Ke, et al.
Veröffentlicht: (2024)
SemToken: Semantic-Aware Tokenization for Efficient Long-Context Language Modeling
von: Liu, Dong, et al.
Veröffentlicht: (2025)
von: Liu, Dong, et al.
Veröffentlicht: (2025)
AnthroScore: A Computational Linguistic Measure of Anthropomorphism
von: Cheng, Myra, et al.
Veröffentlicht: (2024)
von: Cheng, Myra, et al.
Veröffentlicht: (2024)
Self-supervised Analogical Learning using Language Models
von: Zhou, Ben, et al.
Veröffentlicht: (2025)
von: Zhou, Ben, et al.
Veröffentlicht: (2025)
HIGHT: Hierarchical Graph Tokenization for Molecule-Language Alignment
von: Chen, Yongqiang, et al.
Veröffentlicht: (2024)
von: Chen, Yongqiang, et al.
Veröffentlicht: (2024)
Self-Supervised Alignment with Mutual Information: Learning to Follow Principles without Preference Labels
von: Fränken, Jan-Philipp, et al.
Veröffentlicht: (2024)
von: Fränken, Jan-Philipp, et al.
Veröffentlicht: (2024)
Enhancing Large Language Models for Mobility Analytics with Semantic Location Tokenization
von: Chen, Yile, et al.
Veröffentlicht: (2025)
von: Chen, Yile, et al.
Veröffentlicht: (2025)
Semantic Convergence: Harmonizing Recommender Systems via Two-Stage Alignment and Behavioral Semantic Tokenization
von: Li, Guanghan, et al.
Veröffentlicht: (2024)
von: Li, Guanghan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Beyond Tokens: Concept-Level Training Objectives for LLMs
von: Iyer, Laya, et al.
Veröffentlicht: (2026) -
The Roots of Performance Disparity in Multilingual Language Models: Intrinsic Modeling Difficulty or Design Choices?
von: Shani, Chen, et al.
Veröffentlicht: (2026) -
Rethinking Word Similarity: Semantic Similarity through Classification Confusion
von: Zhou, Kaitlyn, et al.
Veröffentlicht: (2025) -
From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning
von: Shani, Chen, et al.
Veröffentlicht: (2025) -
Cooking Up Creativity: Enhancing LLM Creativity through Structured Recombination
von: Mizrahi, Moran, et al.
Veröffentlicht: (2025)