Token Prepending: A Training-Free Approach for Eliciting Better Sentence Embeddings from LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Fu, Yuchen, Cheng, Zifeng, Jiang, Zhiwei, Wang, Zhonghui, Yin, Yafeng, Li, Zhengliang, Gu, Qing |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Contrastive Prompting Enhances Sentence Embeddings in LLMs through Inference-Time Steering
por: Cheng, Zifeng, et al.
Publicado: (2025)
por: Cheng, Zifeng, et al.
Publicado: (2025)
Steering When Necessary: Flexible Steering Large Language Models with Backtracking
por: Cheng, Zifeng, et al.
Publicado: (2025)
por: Cheng, Zifeng, et al.
Publicado: (2025)
GenEOL: Harnessing the Generative Power of LLMs for Training-Free Sentence Embeddings
por: Thirukovalluru, Raghuveer, et al.
Publicado: (2024)
por: Thirukovalluru, Raghuveer, et al.
Publicado: (2024)
RegionMarker: A Region-Triggered Semantic Watermarking Framework for Embedding-as-a-Service Copyright Protection
por: Yang, Shufan, et al.
Publicado: (2025)
por: Yang, Shufan, et al.
Publicado: (2025)
Multi-Prompting Decoder Helps Better Language Understanding
por: Cheng, Zifeng, et al.
Publicado: (2024)
por: Cheng, Zifeng, et al.
Publicado: (2024)
Advanced Sign Language Video Generation with Compressed and Quantized Multi-Condition Tokenization
por: Wang, Cong, et al.
Publicado: (2025)
por: Wang, Cong, et al.
Publicado: (2025)
Eliciting Better Multilingual Structured Reasoning from LLMs through Code
por: Li, Bryan, et al.
Publicado: (2024)
por: Li, Bryan, et al.
Publicado: (2024)
Computer Environments Elicit General Agentic Intelligence in LLMs
por: Cheng, Daixuan, et al.
Publicado: (2026)
por: Cheng, Daixuan, et al.
Publicado: (2026)
Verbal Process Supervision Elicits Better Coding Agents
por: Chen, Hao-Yuan, et al.
Publicado: (2025)
por: Chen, Hao-Yuan, et al.
Publicado: (2025)
Memory Tokens: Large Language Models Can Generate Reversible Sentence Embeddings
por: Sastre, Ignacio, et al.
Publicado: (2025)
por: Sastre, Ignacio, et al.
Publicado: (2025)
Executable Code Actions Elicit Better LLM Agents
por: Wang, Xingyao, et al.
Publicado: (2024)
por: Wang, Xingyao, et al.
Publicado: (2024)
LuxEmbedder: A Cross-Lingual Approach to Enhanced Luxembourgish Sentence Embeddings
por: Philippy, Fred, et al.
Publicado: (2024)
por: Philippy, Fred, et al.
Publicado: (2024)
Logic-Regularized Verifier Elicits Reasoning from LLMs
por: Wang, Xinyu, et al.
Publicado: (2026)
por: Wang, Xinyu, et al.
Publicado: (2026)
AgentGroupChat: An Interactive Group Chat Simulacra For Better Eliciting Emergent Behavior
por: Gu, Zhouhong, et al.
Publicado: (2024)
por: Gu, Zhouhong, et al.
Publicado: (2024)
Latent Reasoning via Sentence Embedding Prediction
por: Hwang, Hyeonbin, et al.
Publicado: (2025)
por: Hwang, Hyeonbin, et al.
Publicado: (2025)
Training LLMs Beyond Next Token Prediction -- Filling the Mutual Information Gap
por: Yang, Chun-Hao, et al.
Publicado: (2025)
por: Yang, Chun-Hao, et al.
Publicado: (2025)
T-FREE: Subword Tokenizer-Free Generative LLMs via Sparse Representations for Memory-Efficient Embeddings
por: Deiseroth, Björn, et al.
Publicado: (2024)
por: Deiseroth, Björn, et al.
Publicado: (2024)
AlphaToken: Decoupling Adaptation and Stability for Path-Aware Response Token Valuation in LLM Post-Training
por: Qing, Liu, et al.
Publicado: (2026)
por: Qing, Liu, et al.
Publicado: (2026)
RSQ: Learning from Important Tokens Leads to Better Quantized LLMs
por: Sung, Yi-Lin, et al.
Publicado: (2025)
por: Sung, Yi-Lin, et al.
Publicado: (2025)
Is Factuality Enhancement a Free Lunch For LLMs? Better Factuality Can Lead to Worse Context-Faithfulness
por: Bi, Baolong, et al.
Publicado: (2024)
por: Bi, Baolong, et al.
Publicado: (2024)
LeDex: Training LLMs to Better Self-Debug and Explain Code
por: Jiang, Nan, et al.
Publicado: (2024)
por: Jiang, Nan, et al.
Publicado: (2024)
Static Word Embeddings for Sentence Semantic Representation
por: Wada, Takashi, et al.
Publicado: (2025)
por: Wada, Takashi, et al.
Publicado: (2025)
Predicting Oscar-Nominated Screenplays with Sentence Embeddings
por: Gross, Francis
Publicado: (2025)
por: Gross, Francis
Publicado: (2025)
A Training-Free Length Extrapolation Approach for LLMs: Greedy Attention Logit Interpolation (GALI)
por: Li, Yan, et al.
Publicado: (2025)
por: Li, Yan, et al.
Publicado: (2025)
Template-assisted Contrastive Learning of Task-oriented Dialogue Sentence Embeddings
por: Oh, Minsik, et al.
Publicado: (2023)
por: Oh, Minsik, et al.
Publicado: (2023)
TNCSE: Tensor's Norm Constraints for Unsupervised Contrastive Learning of Sentence Embeddings
por: Zong, Tianyu, et al.
Publicado: (2025)
por: Zong, Tianyu, et al.
Publicado: (2025)
Generating Hierarchical JSON Representations of Scientific Sentences Using LLMs
por: Nimmagadda, Satya Sri Rajiteswari, et al.
Publicado: (2026)
por: Nimmagadda, Satya Sri Rajiteswari, et al.
Publicado: (2026)
Unsupervised Candidate Ranking for Lexical Substitution via Holistic Sentence Semantics
por: Hu, Zhongyang, et al.
Publicado: (2025)
por: Hu, Zhongyang, et al.
Publicado: (2025)
Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs
por: Han, Pengrui, et al.
Publicado: (2026)
por: Han, Pengrui, et al.
Publicado: (2026)
Logit Arithmetic Elicits Long Reasoning Capabilities Without Training
por: Zhang, Yunxiang, et al.
Publicado: (2025)
por: Zhang, Yunxiang, et al.
Publicado: (2025)
SBERT studies Meaning Representations: Decomposing Sentence Embeddings into Explainable Semantic Features
por: Opitz, Juri, et al.
Publicado: (2022)
por: Opitz, Juri, et al.
Publicado: (2022)
HNCSE: Advancing Sentence Embeddings via Hybrid Contrastive Learning with Hard Negatives
por: Liu, Wenxiao, et al.
Publicado: (2024)
por: Liu, Wenxiao, et al.
Publicado: (2024)
Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation
por: Liu, Jingyu, et al.
Publicado: (2025)
por: Liu, Jingyu, et al.
Publicado: (2025)
Discrete Tokenization for Multimodal LLMs: A Comprehensive Survey
por: Li, Jindong, et al.
Publicado: (2025)
por: Li, Jindong, et al.
Publicado: (2025)
Causal2Vec: Improving Decoder-only LLMs as Embedding Models through a Contextual Token
por: Lin, Ailiang, et al.
Publicado: (2025)
por: Lin, Ailiang, et al.
Publicado: (2025)
Efficient Switchable Safety Control in LLMs via Magic-Token-Guided Co-Training
por: Si, Jianfeng, et al.
Publicado: (2025)
por: Si, Jianfeng, et al.
Publicado: (2025)
Putting People in LLMs' Shoes: Generating Better Answers via Question Rewriter
por: Chen, Junhao, et al.
Publicado: (2024)
por: Chen, Junhao, et al.
Publicado: (2024)
Can LLMs Learn by Teaching for Better Reasoning? A Preliminary Study
por: Ning, Xuefei, et al.
Publicado: (2024)
por: Ning, Xuefei, et al.
Publicado: (2024)
Training-Free Tokenizer Transplantation via Orthogonal Matching Pursuit
por: Goddard, Charles, et al.
Publicado: (2025)
por: Goddard, Charles, et al.
Publicado: (2025)
Better Embeddings with Coupled Adam
por: Stollenwerk, Felix, et al.
Publicado: (2025)
por: Stollenwerk, Felix, et al.
Publicado: (2025)
Ejemplares similares
-
Contrastive Prompting Enhances Sentence Embeddings in LLMs through Inference-Time Steering
por: Cheng, Zifeng, et al.
Publicado: (2025) -
Steering When Necessary: Flexible Steering Large Language Models with Backtracking
por: Cheng, Zifeng, et al.
Publicado: (2025) -
GenEOL: Harnessing the Generative Power of LLMs for Training-Free Sentence Embeddings
por: Thirukovalluru, Raghuveer, et al.
Publicado: (2024) -
RegionMarker: A Region-Triggered Semantic Watermarking Framework for Embedding-as-a-Service Copyright Protection
por: Yang, Shufan, et al.
Publicado: (2025) -
Multi-Prompting Decoder Helps Better Language Understanding
por: Cheng, Zifeng, et al.
Publicado: (2024)