Learning Mutually Informed Representations for Characters and Subwords
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yilin, Hu, Xinyi, Gormley, Matthew R. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Autoregressive Training with Dynamic Oracles
by: Yang, Jianing, et al.
Published: (2024)
by: Yang, Jianing, et al.
Published: (2024)
Enhancing Sindhi Word Segmentation using Subword Representation Learning and Position-aware Self-attention
by: Ali, Wazir, et al.
Published: (2020)
by: Ali, Wazir, et al.
Published: (2020)
T-FREE: Subword Tokenizer-Free Generative LLMs via Sparse Representations for Memory-Efficient Embeddings
by: Deiseroth, Björn, et al.
Published: (2024)
by: Deiseroth, Björn, et al.
Published: (2024)
Understanding Subword Compositionality of Large Language Models
by: Peng, Qiwei, et al.
Published: (2025)
by: Peng, Qiwei, et al.
Published: (2025)
On the Effect of (Near) Duplicate Subwords in Language Modelling
by: Schäfer, Anton, et al.
Published: (2024)
by: Schäfer, Anton, et al.
Published: (2024)
Pretraining Language Models with Subword Regularization: An Empirical Study of BPE Dropout in Low-Resource NLP
by: Visser, Ruan, et al.
Published: (2026)
by: Visser, Ruan, et al.
Published: (2026)
TokDrift: When LLM Speaks in Subwords but Code Speaks in Grammar
by: Li, Yinxi, et al.
Published: (2025)
by: Li, Yinxi, et al.
Published: (2025)
Representation Learning with Conditional Information Flow Maximization
by: Hu, Dou, et al.
Published: (2024)
by: Hu, Dou, et al.
Published: (2024)
Learning Task Representations from In-Context Learning
by: Saglam, Baturay, et al.
Published: (2025)
by: Saglam, Baturay, et al.
Published: (2025)
Significance-Gain Pair Encoding for LLMs: A Statistical Alternative to Frequency-Based Subword Merging
by: Nouri, Azam
Published: (2026)
by: Nouri, Azam
Published: (2026)
Token Alignment via Character Matching for Subword Completion
by: Athiwaratkun, Ben, et al.
Published: (2024)
by: Athiwaratkun, Ben, et al.
Published: (2024)
STENCIL: Submodular Mutual Information Based Weak Supervision for Cold-Start Active Learning
by: Beck, Nathan, et al.
Published: (2024)
by: Beck, Nathan, et al.
Published: (2024)
Efficient RLVR Training via Weighted Mutual Information Data Selection
by: Zhou, Xinyu, et al.
Published: (2026)
by: Zhou, Xinyu, et al.
Published: (2026)
Single Character Perturbations Break LLM Alignment
by: Lin, Leon, et al.
Published: (2024)
by: Lin, Leon, et al.
Published: (2024)
Learning Harmonized Representations for Speculative Sampling
by: Zhang, Lefan, et al.
Published: (2024)
by: Zhang, Lefan, et al.
Published: (2024)
Intrinsic Mutual Information as a Modulator for Preference Optimization
by: Liao, Peng, et al.
Published: (2026)
by: Liao, Peng, et al.
Published: (2026)
GCC-Spam: Spam Detection via GAN, Contrastive Learning, and Character Similarity Networks
by: Wang, Zhijie, et al.
Published: (2025)
by: Wang, Zhijie, et al.
Published: (2025)
Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models
by: Liu, Xiaoze, et al.
Published: (2026)
by: Liu, Xiaoze, et al.
Published: (2026)
Persona Vectors: Monitoring and Controlling Character Traits in Language Models
by: Chen, Runjin, et al.
Published: (2025)
by: Chen, Runjin, et al.
Published: (2025)
CharED: Character-wise Ensemble Decoding for Large Language Models
by: Gu, Kevin, et al.
Published: (2024)
by: Gu, Kevin, et al.
Published: (2024)
Exploration of Summarization by Generative Language Models for Automated Scoring of Long Essays
by: Hua, Haowei, et al.
Published: (2025)
by: Hua, Haowei, et al.
Published: (2025)
Inducing Meaningful Units from Character Sequences with Dynamic Capacity Slot Attention
by: Behjati, Melika, et al.
Published: (2021)
by: Behjati, Melika, et al.
Published: (2021)
Structured Style-Rewrite with Chain-of-Thought Planning for Low-Resource Character Dialogue
by: Zhu, Chanhui
Published: (2026)
by: Zhu, Chanhui
Published: (2026)
Many of Your DPOs are Secretly One: Attempting Unification Through Mutual Information
by: Tutnov, Rasul, et al.
Published: (2025)
by: Tutnov, Rasul, et al.
Published: (2025)
Disentangled Representation Learning with Large Language Models for Text-Attributed Graphs
by: Qin, Yijian, et al.
Published: (2023)
by: Qin, Yijian, et al.
Published: (2023)
OUTLINEFORGE: Hierarchical Reinforcement Learning with Explicit States for Scientific Writing
by: Bao, Yilin, et al.
Published: (2026)
by: Bao, Yilin, et al.
Published: (2026)
An Information-theoretic Multi-task Representation Learning Framework for Natural Language Understanding
by: Hu, Dou, et al.
Published: (2025)
by: Hu, Dou, et al.
Published: (2025)
Utilizing and Calibrating Hindsight Process Rewards via Reinforcement with Mutual Information Self-Evaluation
by: Yao, Jiashu, et al.
Published: (2026)
by: Yao, Jiashu, et al.
Published: (2026)
Patent Representation Learning via Self-supervision
by: Zuo, You, et al.
Published: (2025)
by: Zuo, You, et al.
Published: (2025)
On the Relationship Between the Choice of Representation and In-Context Learning
by: Marinescu, Ioana, et al.
Published: (2025)
by: Marinescu, Ioana, et al.
Published: (2025)
Training a Bilingual Language Model by Mapping Tokens onto a Shared Character Space
by: Rom, Aviad, et al.
Published: (2024)
by: Rom, Aviad, et al.
Published: (2024)
Mutual Reinforcement of LLM Dialogue Synthesis and Summarization Capabilities for Few-Shot Dialogue Summarization
by: Lu, Yen-Ju, et al.
Published: (2025)
by: Lu, Yen-Ju, et al.
Published: (2025)
Improving Academic Skills Assessment with NLP and Ensemble Learning
by: Huang, Xinyi, et al.
Published: (2024)
by: Huang, Xinyi, et al.
Published: (2024)
Behavior Structformer: Learning Players Representations with Structured Tokenization
by: Smirnov, Oleg, et al.
Published: (2024)
by: Smirnov, Oleg, et al.
Published: (2024)
MLP-KAN: Unifying Deep Representation and Function Learning
by: He, Yunhong, et al.
Published: (2024)
by: He, Yunhong, et al.
Published: (2024)
Large Language Models are Superpositions of All Characters: Attaining Arbitrary Role-play via Self-Alignment
by: Lu, Keming, et al.
Published: (2024)
by: Lu, Keming, et al.
Published: (2024)
MindMap: Knowledge Graph Prompting Sparks Graph of Thoughts in Large Language Models
by: Wen, Yilin, et al.
Published: (2023)
by: Wen, Yilin, et al.
Published: (2023)
Can Character-based Language Models Improve Downstream Task Performance in Low-Resource and Noisy Language Scenarios?
by: Riabi, Arij, et al.
Published: (2021)
by: Riabi, Arij, et al.
Published: (2021)
FLoE: Fisher-Based Layer Selection for Efficient Sparse Adaptation of Low-Rank Experts
by: Wang, Xinyi, et al.
Published: (2025)
by: Wang, Xinyi, et al.
Published: (2025)
Multimodal Contrastive Representation Learning in Augmented Biomedical Knowledge Graphs
by: Dang, Tien, et al.
Published: (2025)
by: Dang, Tien, et al.
Published: (2025)
Similar Items
-
Improving Autoregressive Training with Dynamic Oracles
by: Yang, Jianing, et al.
Published: (2024) -
Enhancing Sindhi Word Segmentation using Subword Representation Learning and Position-aware Self-attention
by: Ali, Wazir, et al.
Published: (2020) -
T-FREE: Subword Tokenizer-Free Generative LLMs via Sparse Representations for Memory-Efficient Embeddings
by: Deiseroth, Björn, et al.
Published: (2024) -
Understanding Subword Compositionality of Large Language Models
by: Peng, Qiwei, et al.
Published: (2025) -
On the Effect of (Near) Duplicate Subwords in Language Modelling
by: Schäfer, Anton, et al.
Published: (2024)