Token Prepending: A Training-Free Approach for Eliciting Better Sentence Embeddings from LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Fu, Yuchen, Cheng, Zifeng, Jiang, Zhiwei, Wang, Zhonghui, Yin, Yafeng, Li, Zhengliang, Gu, Qing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Contrastive Prompting Enhances Sentence Embeddings in LLMs through Inference-Time Steering
di: Cheng, Zifeng, et al.
Pubblicazione: (2025)
di: Cheng, Zifeng, et al.
Pubblicazione: (2025)
Steering When Necessary: Flexible Steering Large Language Models with Backtracking
di: Cheng, Zifeng, et al.
Pubblicazione: (2025)
di: Cheng, Zifeng, et al.
Pubblicazione: (2025)
GenEOL: Harnessing the Generative Power of LLMs for Training-Free Sentence Embeddings
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024)
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024)
RegionMarker: A Region-Triggered Semantic Watermarking Framework for Embedding-as-a-Service Copyright Protection
di: Yang, Shufan, et al.
Pubblicazione: (2025)
di: Yang, Shufan, et al.
Pubblicazione: (2025)
Multi-Prompting Decoder Helps Better Language Understanding
di: Cheng, Zifeng, et al.
Pubblicazione: (2024)
di: Cheng, Zifeng, et al.
Pubblicazione: (2024)
Advanced Sign Language Video Generation with Compressed and Quantized Multi-Condition Tokenization
di: Wang, Cong, et al.
Pubblicazione: (2025)
di: Wang, Cong, et al.
Pubblicazione: (2025)
Eliciting Better Multilingual Structured Reasoning from LLMs through Code
di: Li, Bryan, et al.
Pubblicazione: (2024)
di: Li, Bryan, et al.
Pubblicazione: (2024)
Computer Environments Elicit General Agentic Intelligence in LLMs
di: Cheng, Daixuan, et al.
Pubblicazione: (2026)
di: Cheng, Daixuan, et al.
Pubblicazione: (2026)
Verbal Process Supervision Elicits Better Coding Agents
di: Chen, Hao-Yuan, et al.
Pubblicazione: (2025)
di: Chen, Hao-Yuan, et al.
Pubblicazione: (2025)
Memory Tokens: Large Language Models Can Generate Reversible Sentence Embeddings
di: Sastre, Ignacio, et al.
Pubblicazione: (2025)
di: Sastre, Ignacio, et al.
Pubblicazione: (2025)
Executable Code Actions Elicit Better LLM Agents
di: Wang, Xingyao, et al.
Pubblicazione: (2024)
di: Wang, Xingyao, et al.
Pubblicazione: (2024)
LuxEmbedder: A Cross-Lingual Approach to Enhanced Luxembourgish Sentence Embeddings
di: Philippy, Fred, et al.
Pubblicazione: (2024)
di: Philippy, Fred, et al.
Pubblicazione: (2024)
Logic-Regularized Verifier Elicits Reasoning from LLMs
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
di: Wang, Xinyu, et al.
Pubblicazione: (2026)
AgentGroupChat: An Interactive Group Chat Simulacra For Better Eliciting Emergent Behavior
di: Gu, Zhouhong, et al.
Pubblicazione: (2024)
di: Gu, Zhouhong, et al.
Pubblicazione: (2024)
Latent Reasoning via Sentence Embedding Prediction
di: Hwang, Hyeonbin, et al.
Pubblicazione: (2025)
di: Hwang, Hyeonbin, et al.
Pubblicazione: (2025)
Training LLMs Beyond Next Token Prediction -- Filling the Mutual Information Gap
di: Yang, Chun-Hao, et al.
Pubblicazione: (2025)
di: Yang, Chun-Hao, et al.
Pubblicazione: (2025)
T-FREE: Subword Tokenizer-Free Generative LLMs via Sparse Representations for Memory-Efficient Embeddings
di: Deiseroth, Björn, et al.
Pubblicazione: (2024)
di: Deiseroth, Björn, et al.
Pubblicazione: (2024)
AlphaToken: Decoupling Adaptation and Stability for Path-Aware Response Token Valuation in LLM Post-Training
di: Qing, Liu, et al.
Pubblicazione: (2026)
di: Qing, Liu, et al.
Pubblicazione: (2026)
RSQ: Learning from Important Tokens Leads to Better Quantized LLMs
di: Sung, Yi-Lin, et al.
Pubblicazione: (2025)
di: Sung, Yi-Lin, et al.
Pubblicazione: (2025)
Is Factuality Enhancement a Free Lunch For LLMs? Better Factuality Can Lead to Worse Context-Faithfulness
di: Bi, Baolong, et al.
Pubblicazione: (2024)
di: Bi, Baolong, et al.
Pubblicazione: (2024)
LeDex: Training LLMs to Better Self-Debug and Explain Code
di: Jiang, Nan, et al.
Pubblicazione: (2024)
di: Jiang, Nan, et al.
Pubblicazione: (2024)
Static Word Embeddings for Sentence Semantic Representation
di: Wada, Takashi, et al.
Pubblicazione: (2025)
di: Wada, Takashi, et al.
Pubblicazione: (2025)
Predicting Oscar-Nominated Screenplays with Sentence Embeddings
di: Gross, Francis
Pubblicazione: (2025)
di: Gross, Francis
Pubblicazione: (2025)
A Training-Free Length Extrapolation Approach for LLMs: Greedy Attention Logit Interpolation (GALI)
di: Li, Yan, et al.
Pubblicazione: (2025)
di: Li, Yan, et al.
Pubblicazione: (2025)
Template-assisted Contrastive Learning of Task-oriented Dialogue Sentence Embeddings
di: Oh, Minsik, et al.
Pubblicazione: (2023)
di: Oh, Minsik, et al.
Pubblicazione: (2023)
TNCSE: Tensor's Norm Constraints for Unsupervised Contrastive Learning of Sentence Embeddings
di: Zong, Tianyu, et al.
Pubblicazione: (2025)
di: Zong, Tianyu, et al.
Pubblicazione: (2025)
Generating Hierarchical JSON Representations of Scientific Sentences Using LLMs
di: Nimmagadda, Satya Sri Rajiteswari, et al.
Pubblicazione: (2026)
di: Nimmagadda, Satya Sri Rajiteswari, et al.
Pubblicazione: (2026)
Unsupervised Candidate Ranking for Lexical Substitution via Holistic Sentence Semantics
di: Hu, Zhongyang, et al.
Pubblicazione: (2025)
di: Hu, Zhongyang, et al.
Pubblicazione: (2025)
Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs
di: Han, Pengrui, et al.
Pubblicazione: (2026)
di: Han, Pengrui, et al.
Pubblicazione: (2026)
Logit Arithmetic Elicits Long Reasoning Capabilities Without Training
di: Zhang, Yunxiang, et al.
Pubblicazione: (2025)
di: Zhang, Yunxiang, et al.
Pubblicazione: (2025)
SBERT studies Meaning Representations: Decomposing Sentence Embeddings into Explainable Semantic Features
di: Opitz, Juri, et al.
Pubblicazione: (2022)
di: Opitz, Juri, et al.
Pubblicazione: (2022)
HNCSE: Advancing Sentence Embeddings via Hybrid Contrastive Learning with Hard Negatives
di: Liu, Wenxiao, et al.
Pubblicazione: (2024)
di: Liu, Wenxiao, et al.
Pubblicazione: (2024)
Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation
di: Liu, Jingyu, et al.
Pubblicazione: (2025)
di: Liu, Jingyu, et al.
Pubblicazione: (2025)
Discrete Tokenization for Multimodal LLMs: A Comprehensive Survey
di: Li, Jindong, et al.
Pubblicazione: (2025)
di: Li, Jindong, et al.
Pubblicazione: (2025)
Causal2Vec: Improving Decoder-only LLMs as Embedding Models through a Contextual Token
di: Lin, Ailiang, et al.
Pubblicazione: (2025)
di: Lin, Ailiang, et al.
Pubblicazione: (2025)
Efficient Switchable Safety Control in LLMs via Magic-Token-Guided Co-Training
di: Si, Jianfeng, et al.
Pubblicazione: (2025)
di: Si, Jianfeng, et al.
Pubblicazione: (2025)
Putting People in LLMs' Shoes: Generating Better Answers via Question Rewriter
di: Chen, Junhao, et al.
Pubblicazione: (2024)
di: Chen, Junhao, et al.
Pubblicazione: (2024)
Can LLMs Learn by Teaching for Better Reasoning? A Preliminary Study
di: Ning, Xuefei, et al.
Pubblicazione: (2024)
di: Ning, Xuefei, et al.
Pubblicazione: (2024)
Training-Free Tokenizer Transplantation via Orthogonal Matching Pursuit
di: Goddard, Charles, et al.
Pubblicazione: (2025)
di: Goddard, Charles, et al.
Pubblicazione: (2025)
Better Embeddings with Coupled Adam
di: Stollenwerk, Felix, et al.
Pubblicazione: (2025)
di: Stollenwerk, Felix, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Contrastive Prompting Enhances Sentence Embeddings in LLMs through Inference-Time Steering
di: Cheng, Zifeng, et al.
Pubblicazione: (2025) -
Steering When Necessary: Flexible Steering Large Language Models with Backtracking
di: Cheng, Zifeng, et al.
Pubblicazione: (2025) -
GenEOL: Harnessing the Generative Power of LLMs for Training-Free Sentence Embeddings
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024) -
RegionMarker: A Region-Triggered Semantic Watermarking Framework for Embedding-as-a-Service Copyright Protection
di: Yang, Shufan, et al.
Pubblicazione: (2025) -
Multi-Prompting Decoder Helps Better Language Understanding
di: Cheng, Zifeng, et al.
Pubblicazione: (2024)