MEXMA: Token-level objectives improve sentence representations
Fuente:
arXiv
Saved in:
| Main Authors: | Janeiro, João Maria, Piwowarski, Benjamin, Gallinari, Patrick, Barrault, Loïc |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Lossless Token Pruning in Late-Interaction Retrieval Models
by: Zong, Yuxuan, et al.
Published: (2025)
by: Zong, Yuxuan, et al.
Published: (2025)
Structural Deep Encoding for Table Question Answering
by: Mouravieff, Raphaël, et al.
Published: (2025)
by: Mouravieff, Raphaël, et al.
Published: (2025)
On the Representations of Entities in Auto-regressive Large Language Models
by: Morand, Victor, et al.
Published: (2025)
by: Morand, Victor, et al.
Published: (2025)
Hyperbolic sentence representations for solving Textual Entailment
by: Petrovski, Igor
Published: (2024)
by: Petrovski, Igor
Published: (2024)
ToMMeR -- Efficient Entity Mention Detection from Large Language Models
by: Morand, Victor, et al.
Published: (2025)
by: Morand, Victor, et al.
Published: (2025)
Probing Language Models on Their Knowledge Source
by: Tighidet, Zineddine, et al.
Published: (2024)
by: Tighidet, Zineddine, et al.
Published: (2024)
Interference Matrix: Quantifying Cross-Lingual Interference in Transformer Encoders
by: Alastruey, Belen, et al.
Published: (2025)
by: Alastruey, Belen, et al.
Published: (2025)
Give it Space! Explicit Disentangling of Positional and Semantic Representations in Encoders
by: Lequeu, Pierre-Antoine, et al.
Published: (2026)
by: Lequeu, Pierre-Antoine, et al.
Published: (2026)
Generating bilingual example sentences with large language models as lexicography assistants
by: Merx, Raphael, et al.
Published: (2024)
by: Merx, Raphael, et al.
Published: (2024)
Jailbreak Instruction-Tuned LLMs via end-of-sentence MLP Re-weighting
by: Luo, Yifan, et al.
Published: (2024)
by: Luo, Yifan, et al.
Published: (2024)
EzSQL: An SQL intermediate representation for improving SQL-to-text Generation
by: Bhardwaj, Meher, et al.
Published: (2024)
by: Bhardwaj, Meher, et al.
Published: (2024)
Rubrics to Tokens: Bridging Response-level Rubrics and Token-level Rewards in Instruction Following Tasks
by: Xu, Tianze, et al.
Published: (2026)
by: Xu, Tianze, et al.
Published: (2026)
DefSent+: Improving sentence embeddings of language models by projecting definition sentences into a quasi-isotropic or isotropic vector space of unlimited dictionary entries
by: Liu, Xiaodong
Published: (2024)
by: Liu, Xiaodong
Published: (2024)
LLMs are Not Just Next Token Predictors
by: Downes, Stephen M., et al.
Published: (2024)
by: Downes, Stephen M., et al.
Published: (2024)
A new approach for fine-tuning sentence transformers for intent classification and out-of-scope detection tasks
by: Zhang, Tianyi, et al.
Published: (2024)
by: Zhang, Tianyi, et al.
Published: (2024)
Context Copying Modulation: The Role of Entropy Neurons in Managing Parametric and Contextual Knowledge Conflicts
by: Tighidet, Zineddine, et al.
Published: (2025)
by: Tighidet, Zineddine, et al.
Published: (2025)
How does fine-tuning improve sensorimotor representations in large language models?
by: Wu, Minghua, et al.
Published: (2026)
by: Wu, Minghua, et al.
Published: (2026)
Token-level Direct Preference Optimization
by: Zeng, Yongcheng, et al.
Published: (2024)
by: Zeng, Yongcheng, et al.
Published: (2024)
SentenceVAE: Enable Next-sentence Prediction for Large Language Models with Faster Speed, Higher Accuracy and Longer Context
by: An, Hongjun, et al.
Published: (2024)
by: An, Hongjun, et al.
Published: (2024)
ARC-Encoder: learning compressed text representations for large language models
by: Pilchen, Hippolyte, et al.
Published: (2025)
by: Pilchen, Hippolyte, et al.
Published: (2025)
Token-Level Uncertainty-Aware Objective for Language Model Post-Training
by: Liu, Tingkai, et al.
Published: (2025)
by: Liu, Tingkai, et al.
Published: (2025)
Empowering Character-level Text Infilling by Eliminating Sub-Tokens
by: Ren, Houxing, et al.
Published: (2024)
by: Ren, Houxing, et al.
Published: (2024)
Reinforcement Learning with Token-level Feedback for Controllable Text Generation
by: Li, Wendi, et al.
Published: (2024)
by: Li, Wendi, et al.
Published: (2024)
Explainable Token-level Noise Filtering for LLM Fine-tuning Datasets
by: Yang, Yuchen, et al.
Published: (2026)
by: Yang, Yuchen, et al.
Published: (2026)
The Token Tax: Systematic Bias in Multilingual Tokenization
by: Lundin, Jessica M., et al.
Published: (2025)
by: Lundin, Jessica M., et al.
Published: (2025)
Full-ECE: A Metric For Token-level Calibration on Large Language Models
by: Liu, Han, et al.
Published: (2024)
by: Liu, Han, et al.
Published: (2024)
State over Tokens: Characterizing the Role of Reasoning Tokens
by: Levy, Mosh, et al.
Published: (2025)
by: Levy, Mosh, et al.
Published: (2025)
Are language models aware of the road not taken? Token-level uncertainty and hidden state dynamics
by: Zur, Amir, et al.
Published: (2025)
by: Zur, Amir, et al.
Published: (2025)
Persona-judge: Personalized Alignment of Large Language Models via Token-level Self-judgment
by: Zhang, Xiaotian, et al.
Published: (2025)
by: Zhang, Xiaotian, et al.
Published: (2025)
Learned Hallucination Detection in Black-Box LLMs using Token-level Entropy Production Rate
by: Moslonka, Charles, et al.
Published: (2025)
by: Moslonka, Charles, et al.
Published: (2025)
Reverse Probing: Supervised Token-level Uncertainty Quantification for Large Language Models in Clinical Text
by: Xiao, Bushi, et al.
Published: (2026)
by: Xiao, Bushi, et al.
Published: (2026)
TiTok: Transfer Token-level Knowledge via Contrastive Excess to Transplant LoRA
by: Jung, Chanjoo, et al.
Published: (2025)
by: Jung, Chanjoo, et al.
Published: (2025)
Tokenization Matters! Degrading Large Language Models through Challenging Their Tokenization
by: Wang, Dixuan, et al.
Published: (2024)
by: Wang, Dixuan, et al.
Published: (2024)
Can LLMs interpret figurative language as humans do?: surface-level vs representational similarity
by: Bollepally, Samhita, et al.
Published: (2026)
by: Bollepally, Samhita, et al.
Published: (2026)
SemToken: Semantic-Aware Tokenization for Efficient Long-Context Language Modeling
by: Liu, Dong, et al.
Published: (2025)
by: Liu, Dong, et al.
Published: (2025)
One Token Is Enough: Improving Diffusion Language Models with a Sink Token
by: Zhang, Zihou, et al.
Published: (2026)
by: Zhang, Zihou, et al.
Published: (2026)
TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching
by: Nguyen, Truong, et al.
Published: (2026)
by: Nguyen, Truong, et al.
Published: (2026)
Learning from Sufficient Rationales: Analysing the Relationship Between Explanation Faithfulness and Token-level Regularisation Strategies
by: Kamp, Jonathan, et al.
Published: (2025)
by: Kamp, Jonathan, et al.
Published: (2025)
TokenSeek: Memory Efficient Fine Tuning via Instance-Aware Token Ditching
by: Zeng, Runjia, et al.
Published: (2026)
by: Zeng, Runjia, et al.
Published: (2026)
Token-Guard: Towards Token-Level Hallucination Control via Self-Checking Decoding
by: Zhu, Yifan, et al.
Published: (2026)
by: Zhu, Yifan, et al.
Published: (2026)
Similar Items
-
Towards Lossless Token Pruning in Late-Interaction Retrieval Models
by: Zong, Yuxuan, et al.
Published: (2025) -
Structural Deep Encoding for Table Question Answering
by: Mouravieff, Raphaël, et al.
Published: (2025) -
On the Representations of Entities in Auto-regressive Large Language Models
by: Morand, Victor, et al.
Published: (2025) -
Hyperbolic sentence representations for solving Textual Entailment
by: Petrovski, Igor
Published: (2024) -
ToMMeR -- Efficient Entity Mention Detection from Large Language Models
by: Morand, Victor, et al.
Published: (2025)