Learning to Compress: Unlocking the Potential of Large Language Models for Text Representation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yeqin, Zhao, Yizheng, Hu, Chen, Jiao, Binxing, Jiang, Daxin, Miao, Ruihang, Nguyen, Cam-Tu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Retrospex: Language Agent Meets Offline Reinforcement Learning Critic
by: Xiang, Yufei, et al.
Published: (2025)
by: Xiang, Yufei, et al.
Published: (2025)
LLM-based Embeddings: Attention Values Encode Sentence Semantics Better Than Hidden States
by: Zhang, Yeqin, et al.
Published: (2026)
by: Zhang, Yeqin, et al.
Published: (2026)
Mitigating the Impact of False Negatives in Dense Retrieval with Contrastive Confidence Regularization
by: Wang, Shiqi, et al.
Published: (2023)
by: Wang, Shiqi, et al.
Published: (2023)
A Survey on Retrieval-Augmented Text Generation for Large Language Models
by: Huang, Yizheng, et al.
Published: (2024)
by: Huang, Yizheng, et al.
Published: (2024)
Unlocking Anticipatory Text Generation: A Constrained Approach for Large Language Models Decoding
by: Tu, Lifu, et al.
Published: (2023)
by: Tu, Lifu, et al.
Published: (2023)
Unlocking the Potential of Model Merging for Low-Resource Languages
by: Tao, Mingxu, et al.
Published: (2024)
by: Tao, Mingxu, et al.
Published: (2024)
XGrammar: Flexible and Efficient Structured Generation Engine for Large Language Models
by: Dong, Yixin, et al.
Published: (2024)
by: Dong, Yixin, et al.
Published: (2024)
Unlocking the Potential: Benchmarking Large Language Models in Water Engineering and Research
by: Xu, Boyan, et al.
Published: (2024)
by: Xu, Boyan, et al.
Published: (2024)
Text Embeddings by Weakly-Supervised Contrastive Pre-training
by: Wang, Liang, et al.
Published: (2022)
by: Wang, Liang, et al.
Published: (2022)
Mixed Preference Optimization: Reinforcement Learning with Data Selection and Better Reference Model
by: Gou, Qi, et al.
Published: (2024)
by: Gou, Qi, et al.
Published: (2024)
WizardCoder: Empowering Code Large Language Models with Evol-Instruct
by: Luo, Ziyang, et al.
Published: (2023)
by: Luo, Ziyang, et al.
Published: (2023)
Reward Difference Optimization For Sample Reweighting In Offline RLHF
by: Wang, Shiqi, et al.
Published: (2024)
by: Wang, Shiqi, et al.
Published: (2024)
Unlocking the Potential of Large Language Models for Clinical Text Anonymization: A Comparative Study
by: Pissarra, David, et al.
Published: (2024)
by: Pissarra, David, et al.
Published: (2024)
Unlocking the Power of Large Language Models for Entity Alignment
by: Jiang, Xuhui, et al.
Published: (2024)
by: Jiang, Xuhui, et al.
Published: (2024)
CMT: A Memory Compression Method for Continual Knowledge Learning of Large Language Models
by: Li, Dongfang, et al.
Published: (2024)
by: Li, Dongfang, et al.
Published: (2024)
EmbedLLM: Learning Compact Representations of Large Language Models
by: Zhuang, Richard, et al.
Published: (2024)
by: Zhuang, Richard, et al.
Published: (2024)
Large Language Model Enhanced Knowledge Representation Learning: A Survey
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
All Languages Matter: On the Multilingual Safety of Large Language Models
by: Wang, Wenxuan, et al.
Published: (2023)
by: Wang, Wenxuan, et al.
Published: (2023)
Unlocking the Potential of Diffusion Language Models through Template Infilling
by: Lee, Junhoo, et al.
Published: (2025)
by: Lee, Junhoo, et al.
Published: (2025)
Unveiling the Misuse Potential of Base Large Language Models via In-Context Learning
by: Wang, Xiao, et al.
Published: (2024)
by: Wang, Xiao, et al.
Published: (2024)
Random Policy Valuation is Enough for LLM Reasoning with Verifiable Rewards
by: He, Haoran, et al.
Published: (2025)
by: He, Haoran, et al.
Published: (2025)
Improving Attributed Text Generation of Large Language Models via Preference Learning
by: Li, Dongfang, et al.
Published: (2024)
by: Li, Dongfang, et al.
Published: (2024)
Adam: Dense Retrieval Distillation with Adaptive Dark Examples
by: Tao, Chongyang, et al.
Published: (2022)
by: Tao, Chongyang, et al.
Published: (2022)
Adaptive Ability Decomposing for Unlocking Large Reasoning Model Effective Reinforcement Learning
by: Chen, Zhipeng, et al.
Published: (2026)
by: Chen, Zhipeng, et al.
Published: (2026)
Dynamic Compressing Prompts for Efficient Inference of Large Language Models
by: Hu, Jinwu, et al.
Published: (2025)
by: Hu, Jinwu, et al.
Published: (2025)
Streamlining Redundant Layers to Compress Large Language Models
by: Chen, Xiaodong, et al.
Published: (2024)
by: Chen, Xiaodong, et al.
Published: (2024)
MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining
by: Xiaomi, LLM-Core, et al.
Published: (2025)
by: Xiaomi, LLM-Core, et al.
Published: (2025)
Q-Mirror: Unlocking the Multi-Modal Potential of Scientific Text-Only QA Pairs
by: Wang, Junying, et al.
Published: (2025)
by: Wang, Junying, et al.
Published: (2025)
Automated Triaging and Transfer Learning of Incident Learning Safety Reports Using Large Language Representational Models
by: Beidler, Peter, et al.
Published: (2025)
by: Beidler, Peter, et al.
Published: (2025)
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks
by: Nguyen, Hieu Minh "Jord"
Published: (2025)
by: Nguyen, Hieu Minh "Jord"
Published: (2025)
Large Language Models for Detection of Life-Threatening Texts
by: Nguyen, Thanh Thi, et al.
Published: (2025)
by: Nguyen, Thanh Thi, et al.
Published: (2025)
Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models
by: Li, Yanwei, et al.
Published: (2024)
by: Li, Yanwei, et al.
Published: (2024)
Information Capacity: Evaluating the Efficiency of Large Language Models via Text Compression
by: Yuan, Cheng, et al.
Published: (2025)
by: Yuan, Cheng, et al.
Published: (2025)
Revisiting Entropy Regularization: Adaptive Coefficient Unlocks Its Potential for LLM Reinforcement Learning
by: Zhang, Xiaoyun, et al.
Published: (2025)
by: Zhang, Xiaoyun, et al.
Published: (2025)
In-context Autoencoder for Context Compression in a Large Language Model
by: Ge, Tao, et al.
Published: (2023)
by: Ge, Tao, et al.
Published: (2023)
Probing the Decision Boundaries of In-context Learning in Large Language Models
by: Zhao, Siyan, et al.
Published: (2024)
by: Zhao, Siyan, et al.
Published: (2024)
TOOLCAD: Exploring Tool-Using Large Language Models in Text-to-CAD Generation with Reinforcement Learning
by: Gong, Yifei, et al.
Published: (2026)
by: Gong, Yifei, et al.
Published: (2026)
Unlocking Multimodal Mathematical Reasoning via Process Reward Model
by: Luo, Ruilin, et al.
Published: (2025)
by: Luo, Ruilin, et al.
Published: (2025)
Unlocking the Working Memory of Large Language Models for Latent Reasoning
by: Aichberger, Lukas, et al.
Published: (2026)
by: Aichberger, Lukas, et al.
Published: (2026)
Unlocking Reasoning Capability on Machine Translation in Large Language Models
by: Rajaee, Sara, et al.
Published: (2026)
by: Rajaee, Sara, et al.
Published: (2026)
Similar Items
-
Retrospex: Language Agent Meets Offline Reinforcement Learning Critic
by: Xiang, Yufei, et al.
Published: (2025) -
LLM-based Embeddings: Attention Values Encode Sentence Semantics Better Than Hidden States
by: Zhang, Yeqin, et al.
Published: (2026) -
Mitigating the Impact of False Negatives in Dense Retrieval with Contrastive Confidence Regularization
by: Wang, Shiqi, et al.
Published: (2023) -
A Survey on Retrieval-Augmented Text Generation for Large Language Models
by: Huang, Yizheng, et al.
Published: (2024) -
Unlocking Anticipatory Text Generation: A Constrained Approach for Large Language Models Decoding
by: Tu, Lifu, et al.
Published: (2023)