Entropy Law: The Story Behind Data Compression and LLM Performance
Fuente:
arXiv
Saved in:
| Main Authors: | Yin, Mingjia, Wu, Chuhan, Wang, Yufei, Wang, Hao, Guo, Wei, Wang, Yasheng, Liu, Yong, Tang, Ruiming, Lian, Defu, Chen, Enhong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Performance Law of Large Language Models
by: Wu, Chuhan, et al.
Published: (2024)
by: Wu, Chuhan, et al.
Published: (2024)
Understanding the planning of LLM agents: A survey
by: Huang, Xu, et al.
Published: (2024)
by: Huang, Xu, et al.
Published: (2024)
Optimizing Sequential Recommendation Models with Scaling Laws and Approximate Entropy
by: Shen, Tingjia, et al.
Published: (2024)
by: Shen, Tingjia, et al.
Published: (2024)
Position: The Real Barrier to LLM Agent Usability is Agentic ROI
by: Liu, Weiwen, et al.
Published: (2025)
by: Liu, Weiwen, et al.
Published: (2025)
ToolACE-DEV: Self-Improving Tool Learning via Decomposition and EVolution
by: Huang, Xu, et al.
Published: (2025)
by: Huang, Xu, et al.
Published: (2025)
Instruction-Tuning Data Synthesis from Scratch via Web Reconstruction
by: Jiang, Yuxin, et al.
Published: (2025)
by: Jiang, Yuxin, et al.
Published: (2025)
Crowd Comparative Reasoning: Unlocking Comprehensive Evaluations for LLM-as-a-Judge
by: Zhang, Qiyuan, et al.
Published: (2025)
by: Zhang, Qiyuan, et al.
Published: (2025)
WESE: Weak Exploration to Strong Exploitation for LLM Agents
by: Huang, Xu, et al.
Published: (2024)
by: Huang, Xu, et al.
Published: (2024)
RevisEval: Improving LLM-as-a-Judge via Response-Adapted References
by: Zhang, Qiyuan, et al.
Published: (2024)
by: Zhang, Qiyuan, et al.
Published: (2024)
LLM Cache Bandit Revisited: Addressing Query Heterogeneity for Cost-Effective LLM Inference
by: Yang, Hantao, et al.
Published: (2025)
by: Yang, Hantao, et al.
Published: (2025)
Advancing and Benchmarking Personalized Tool Invocation for LLMs
by: Huang, Xu, et al.
Published: (2025)
by: Huang, Xu, et al.
Published: (2025)
IE as Cache: Information Extraction Enhanced Agentic Reasoning
by: Lv, Hang, et al.
Published: (2026)
by: Lv, Hang, et al.
Published: (2026)
Bridging and Modeling Correlations in Pairwise Data for Direct Preference Optimization
by: Jiang, Yuxin, et al.
Published: (2024)
by: Jiang, Yuxin, et al.
Published: (2024)
ToolACE: Winning the Points of LLM Function Calling
by: Liu, Weiwen, et al.
Published: (2024)
by: Liu, Weiwen, et al.
Published: (2024)
Dataset Regeneration for Sequential Recommendation
by: Yin, Mingjia, et al.
Published: (2024)
by: Yin, Mingjia, et al.
Published: (2024)
CoSteer: Collaborative Decoding-Time Personalization via Local Delta Steering
by: Lv, Hang, et al.
Published: (2025)
by: Lv, Hang, et al.
Published: (2025)
MDAP: A Multi-view Disentangled and Adaptive Preference Learning Framework for Cross-Domain Recommendation
by: Tong, Junxiong, et al.
Published: (2024)
by: Tong, Junxiong, et al.
Published: (2024)
SpecSteer: Synergizing Local Context and Global Reasoning for Efficient Personalized Generation
by: Lv, Hang, et al.
Published: (2026)
by: Lv, Hang, et al.
Published: (2026)
A Survey on Multi-Turn Interaction Capabilities of Large Language Models
by: Zhang, Chen, et al.
Published: (2025)
by: Zhang, Chen, et al.
Published: (2025)
FuXi-β: Towards a Lightweight and Fast Large-Scale Generative Recommendation Model
by: Ye, Yufei, et al.
Published: (2025)
by: Ye, Yufei, et al.
Published: (2025)
Learning Partially Aligned Item Representation for Cross-Domain Sequential Recommendation
by: Yin, Mingjia, et al.
Published: (2024)
by: Yin, Mingjia, et al.
Published: (2024)
Generative Large Recommendation Models: Emerging Trends in LLMs for Recommendation
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
Learning to Substitute Components for Compositional Generalization
by: Li, Zhaoyi, et al.
Published: (2025)
by: Li, Zhaoyi, et al.
Published: (2025)
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
by: Dong, Kuicai, et al.
Published: (2025)
by: Dong, Kuicai, et al.
Published: (2025)
RAPID: Efficient Retrieval-Augmented Long Text Generation with Writing Planning and Information Discovery
by: Gu, Hongchao, et al.
Published: (2025)
by: Gu, Hongchao, et al.
Published: (2025)
RethinkMCTS: Refining Erroneous Thoughts in Monte Carlo Tree Search for Code Generation
by: Li, Qingyao, et al.
Published: (2024)
by: Li, Qingyao, et al.
Published: (2024)
From Feature Interaction to Feature Generation: A Generative Paradigm of CTR Prediction Models
by: Yin, Mingjia, et al.
Published: (2025)
by: Yin, Mingjia, et al.
Published: (2025)
HumanLLM: Towards Personalized Understanding and Simulation of Human Nature
by: Lei, Yuxuan, et al.
Published: (2026)
by: Lei, Yuxuan, et al.
Published: (2026)
A Universal Framework for Compressing Embeddings in CTR Prediction
by: Wang, Kefan, et al.
Published: (2025)
by: Wang, Kefan, et al.
Published: (2025)
Adaptive Tool Use in Large Language Models with Meta-Cognition Trigger
by: Li, Wenjun, et al.
Published: (2025)
by: Li, Wenjun, et al.
Published: (2025)
Learning to Edit: Aligning LLMs with Knowledge Editing
by: Jiang, Yuxin, et al.
Published: (2024)
by: Jiang, Yuxin, et al.
Published: (2024)
CoIR: A Comprehensive Benchmark for Code Information Retrieval Models
by: Li, Xiangyang, et al.
Published: (2024)
by: Li, Xiangyang, et al.
Published: (2024)
Learning from Emptiness: De-biasing Listwise Rerankers with Content-Agnostic Probability Calibration
by: Lv, Hang, et al.
Published: (2026)
by: Lv, Hang, et al.
Published: (2026)
FuXi-$α$: Scaling Recommendation Model with Feature Interaction Enhanced Transformer
by: Ye, Yufei, et al.
Published: (2025)
by: Ye, Yufei, et al.
Published: (2025)
Understanding DNNs in Feature Interaction Models: A Dimensional Collapse Perspective
by: Wang, Jiancheng, et al.
Published: (2026)
by: Wang, Jiancheng, et al.
Published: (2026)
Enhancing CTR Prediction with De-correlated Expert Networks
by: Wang, Jiancheng, et al.
Published: (2025)
by: Wang, Jiancheng, et al.
Published: (2025)
A Unified Framework for Adaptive Representation Enhancement and Inversed Learning in Cross-Domain Recommendation
by: Zhang, Luankang, et al.
Published: (2024)
by: Zhang, Luankang, et al.
Published: (2024)
LogitsCoder: Towards Efficient Chain-of-Thought Path Search via Logits Preference Decoding for Code Generation
by: Chen, Jizheng, et al.
Published: (2026)
by: Chen, Jizheng, et al.
Published: (2026)
Can Recommender Systems Teach Themselves? A Recursive Self-Improving Framework with Fidelity Control
by: Zhang, Luankang, et al.
Published: (2026)
by: Zhang, Luankang, et al.
Published: (2026)
No One Left Behind: How to Exploit the Incomplete and Skewed Multi-Label Data for Conversion Rate Prediction
by: Jia, Qinglin, et al.
Published: (2025)
by: Jia, Qinglin, et al.
Published: (2025)
Similar Items
-
Performance Law of Large Language Models
by: Wu, Chuhan, et al.
Published: (2024) -
Understanding the planning of LLM agents: A survey
by: Huang, Xu, et al.
Published: (2024) -
Optimizing Sequential Recommendation Models with Scaling Laws and Approximate Entropy
by: Shen, Tingjia, et al.
Published: (2024) -
Position: The Real Barrier to LLM Agent Usability is Agentic ROI
by: Liu, Weiwen, et al.
Published: (2025) -
ToolACE-DEV: Self-Improving Tool Learning via Decomposition and EVolution
by: Huang, Xu, et al.
Published: (2025)