Extensible Embedding: A Flexible Multipler For LLM's Context Length
Fuente:
arXiv
Saved in:
| Main Authors: | Shao, Ninglu, Xiao, Shitao, Liu, Zheng, Zhang, Peitian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Flexibly Scaling Large Language Models Contexts Through Extensible Tokenization
by: Shao, Ninglu, et al.
Published: (2024)
by: Shao, Ninglu, et al.
Published: (2024)
Compressing Lengthy Context With UltraGist
by: Zhang, Peitian, et al.
Published: (2024)
by: Zhang, Peitian, et al.
Published: (2024)
Long Context Compression with Activation Beacon
by: Zhang, Peitian, et al.
Published: (2024)
by: Zhang, Peitian, et al.
Published: (2024)
Extending Llama-3's Context Ten-Fold Overnight
by: Zhang, Peitian, et al.
Published: (2024)
by: Zhang, Peitian, et al.
Published: (2024)
Lighter And Better: Towards Flexible Context Adaptation For Retrieval Augmented Generation
by: Liu, Zheng, et al.
Published: (2024)
by: Liu, Zheng, et al.
Published: (2024)
M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation
by: Chen, Jianlv, et al.
Published: (2024)
by: Chen, Jianlv, et al.
Published: (2024)
C-Pack: Packed Resources For General Chinese Embeddings
by: Xiao, Shitao, et al.
Published: (2023)
by: Xiao, Shitao, et al.
Published: (2023)
BGE Landmark Embedding: A Chunking-Free Embedding Method For Retrieval Augmented Long-Context Large Language Models
by: Luo, Kun, et al.
Published: (2024)
by: Luo, Kun, et al.
Published: (2024)
Matryoshka Re-Ranker: A Flexible Re-Ranking Architecture With Configurable Depth and Width
by: Liu, Zheng, et al.
Published: (2025)
by: Liu, Zheng, et al.
Published: (2025)
Understanding Privacy Risks of Embeddings Induced by Large Language Models
by: Zhu, Zhihao, et al.
Published: (2024)
by: Zhu, Zhihao, et al.
Published: (2024)
Does RAG Really Perform Bad For Long-Context Processing?
by: Luo, Kun, et al.
Published: (2025)
by: Luo, Kun, et al.
Published: (2025)
Boosting Long-Context Management via Query-Guided Activation Refilling
by: Qian, Hongjin, et al.
Published: (2024)
by: Qian, Hongjin, et al.
Published: (2024)
Are Long-LLMs A Necessity For Long-Context Tasks?
by: Qian, Hongjin, et al.
Published: (2024)
by: Qian, Hongjin, et al.
Published: (2024)
Llama2Vec: Unsupervised Adaptation of Large Language Models for Dense Retrieval
by: Liu, Zheng, et al.
Published: (2023)
by: Liu, Zheng, et al.
Published: (2023)
VISTA: Visualized Text Embedding For Universal Multi-Modal Retrieval
by: Zhou, Junjie, et al.
Published: (2024)
by: Zhou, Junjie, et al.
Published: (2024)
MemoRAG: Boosting Long Context Processing with Global Memory-Enhanced Retrieval Augmentation
by: Qian, Hongjin, et al.
Published: (2024)
by: Qian, Hongjin, et al.
Published: (2024)
MedOrch: Medical Diagnosis with Tool-Augmented Reasoning Agents for Flexible Extensibility
by: He, Yexiao, et al.
Published: (2025)
by: He, Yexiao, et al.
Published: (2025)
Beyond Length: Quantifying Long-Range Information for Long-Context LLM Pretraining Data
by: Deng, Haoran, et al.
Published: (2025)
by: Deng, Haoran, et al.
Published: (2025)
Squeezed Attention: Accelerating Long Context Length LLM Inference
by: Hooper, Coleman, et al.
Published: (2024)
by: Hooper, Coleman, et al.
Published: (2024)
Large Language Models as Foundations for Next-Gen Dense Retrieval: A Comprehensive Empirical Assessment
by: Luo, Kun, et al.
Published: (2024)
by: Luo, Kun, et al.
Published: (2024)
Megalodon: Efficient LLM Pretraining and Inference with Unlimited Context Length
by: Ma, Xuezhe, et al.
Published: (2024)
by: Ma, Xuezhe, et al.
Published: (2024)
GREAT: Guiding Query Generation with a Trie for Recommending Related Search about Video at Kuaishou
by: Shao, Ninglu, et al.
Published: (2025)
by: Shao, Ninglu, et al.
Published: (2025)
ByteSized32Refactored: Towards an Extensible Interactive Text Games Corpus for LLM World Modeling and Evaluation
by: Wang, Haonan, et al.
Published: (2025)
by: Wang, Haonan, et al.
Published: (2025)
EasySteer: A Unified Framework for High-Performance and Extensible LLM Steering
by: Xu, Haolei, et al.
Published: (2025)
by: Xu, Haolei, et al.
Published: (2025)
Easy Dataset: A Unified and Extensible Framework for Synthesizing LLM Fine-Tuning Data from Unstructured Documents
by: Miao, Ziyang, et al.
Published: (2025)
by: Miao, Ziyang, et al.
Published: (2025)
InftyThink: Breaking the Length Limits of Long-Context Reasoning in Large Language Models
by: Yan, Yuchen, et al.
Published: (2025)
by: Yan, Yuchen, et al.
Published: (2025)
Growing Pains: Extensible and Efficient LLM Benchmarking Via Fixed Parameter Calibration
by: Habba, Eliya, et al.
Published: (2026)
by: Habba, Eliya, et al.
Published: (2026)
Length-Induced Embedding Collapse in PLM-based Models
by: Zhou, Yuqi, et al.
Published: (2024)
by: Zhou, Yuqi, et al.
Published: (2024)
Base of RoPE Bounds Context Length
by: Men, Xin, et al.
Published: (2024)
by: Men, Xin, et al.
Published: (2024)
Context Length Alone Hurts LLM Performance Despite Perfect Retrieval
by: Du, Yufeng, et al.
Published: (2025)
by: Du, Yufeng, et al.
Published: (2025)
SpikeLLM: Scaling up Spiking Neural Network to Large Language Models via Saliency-based Spiking
by: Xing, Xingrun, et al.
Published: (2024)
by: Xing, Xingrun, et al.
Published: (2024)
Context-aware Biases for Length Extrapolation
by: Veisi, Ali, et al.
Published: (2025)
by: Veisi, Ali, et al.
Published: (2025)
In-Context Learning (and Unlearning) of Length Biases
by: Schoch, Stephanie, et al.
Published: (2025)
by: Schoch, Stephanie, et al.
Published: (2025)
ParallelComp: Parallel Long-Context Compressor for Length Extrapolation
by: Xiong, Jing, et al.
Published: (2025)
by: Xiong, Jing, et al.
Published: (2025)
Bootstrap Your Own Context Length
by: Wang, Liang, et al.
Published: (2024)
by: Wang, Liang, et al.
Published: (2024)
Context Discipline and Performance Correlation: Analyzing LLM Performance and Quality Degradation Under Varying Context Lengths
by: Ponnusamy, Ahilan Ayyachamy Nadar, et al.
Published: (2025)
by: Ponnusamy, Ahilan Ayyachamy Nadar, et al.
Published: (2025)
Making Text Embedders Few-Shot Learners
by: Li, Chaofan, et al.
Published: (2024)
by: Li, Chaofan, et al.
Published: (2024)
LLM Ensemble for RAG: Role of Context Length in Zero-Shot Question Answering for BioASQ Challenge
by: Galat, Dima, et al.
Published: (2025)
by: Galat, Dima, et al.
Published: (2025)
Fast and Extensible Hybrid Embeddings with Micros
by: Bocirnea, Sean, et al.
Published: (2025)
by: Bocirnea, Sean, et al.
Published: (2025)
Building Task Bots with Self-learning for Enhanced Adaptability, Extensibility, and Factuality
by: Zhang, Xiaoying
Published: (2025)
by: Zhang, Xiaoying
Published: (2025)
Similar Items
-
Flexibly Scaling Large Language Models Contexts Through Extensible Tokenization
by: Shao, Ninglu, et al.
Published: (2024) -
Compressing Lengthy Context With UltraGist
by: Zhang, Peitian, et al.
Published: (2024) -
Long Context Compression with Activation Beacon
by: Zhang, Peitian, et al.
Published: (2024) -
Extending Llama-3's Context Ten-Fold Overnight
by: Zhang, Peitian, et al.
Published: (2024) -
Lighter And Better: Towards Flexible Context Adaptation For Retrieval Augmented Generation
by: Liu, Zheng, et al.
Published: (2024)