Entropy-UID: A Method for Optimizing Information Density
Fuente:
arXiv
Guardado en:
| Autor principal: | Shou, Xinpeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Revisiting the UID Hypothesis in LLM Reasoning Traces
por: Gwak, Minju, et al.
Publicado: (2025)
por: Gwak, Minju, et al.
Publicado: (2025)
TokenTiming: A Dynamic Alignment Method for Universal Speculative Decoding Model Pairs
por: Xiao, Sibo, et al.
Publicado: (2025)
por: Xiao, Sibo, et al.
Publicado: (2025)
PLHF: Prompt Optimization with Few-Shot Human Feedback
por: Yang, Chun-Pai, et al.
Publicado: (2025)
por: Yang, Chun-Pai, et al.
Publicado: (2025)
Decoupling Strategy and Execution in Task-Focused Dialogue via Goal-Oriented Preference Optimization
por: Xu, Jingyi, et al.
Publicado: (2026)
por: Xu, Jingyi, et al.
Publicado: (2026)
Mechanistic Decoding of Cognitive Constructs in Large Language Models
por: Shou, Yitong, et al.
Publicado: (2026)
por: Shou, Yitong, et al.
Publicado: (2026)
VEPO: Variable Entropy Policy Optimization for Low-Resource Language Foundation Models
por: Liu, Chonghan, et al.
Publicado: (2026)
por: Liu, Chonghan, et al.
Publicado: (2026)
Entropy Controllable Direct Preference Optimization
por: Omura, Motoki, et al.
Publicado: (2024)
por: Omura, Motoki, et al.
Publicado: (2024)
Revisiting the Uniform Information Density Hypothesis in LLM Reasoning
por: Gwak, Minju, et al.
Publicado: (2025)
por: Gwak, Minju, et al.
Publicado: (2025)
Think Before Refusal : Triggering Safety Reflection in LLMs to Mitigate False Refusal Behavior
por: Si, Shengyun, et al.
Publicado: (2025)
por: Si, Shengyun, et al.
Publicado: (2025)
Is It Thinking or Cheating? Detecting Implicit Reward Hacking by Measuring Reasoning Effort
por: Wang, Xinpeng, et al.
Publicado: (2025)
por: Wang, Xinpeng, et al.
Publicado: (2025)
DELIA: Diversity-Enhanced Learning for Instruction Adaptation in Large Language Models
por: Zeng, Yuanhao, et al.
Publicado: (2024)
por: Zeng, Yuanhao, et al.
Publicado: (2024)
Look at the Text: Instruction-Tuned Language Models are More Robust Multiple Choice Selectors than You Think
por: Wang, Xinpeng, et al.
Publicado: (2024)
por: Wang, Xinpeng, et al.
Publicado: (2024)
InfoFlow: Reinforcing Search Agent Via Reward Density Optimization
por: Luo, Kun, et al.
Publicado: (2025)
por: Luo, Kun, et al.
Publicado: (2025)
Human-Inspired Learning for Large Language Models via Obvious Record and Maximum-Entropy Method Discovery
por: Su, Hong
Publicado: (2025)
por: Su, Hong
Publicado: (2025)
InfoDensity: Rewarding Information-Dense Traces for Efficient Reasoning
por: Wei, Chengwei, et al.
Publicado: (2026)
por: Wei, Chengwei, et al.
Publicado: (2026)
Entropy-Tree: Tree-Based Decoding with Entropy-Guided Exploration
por: Wei, Longxuan, et al.
Publicado: (2026)
por: Wei, Longxuan, et al.
Publicado: (2026)
Density-Guided Response Optimization: Community-Grounded Alignment via Implicit Acceptance Signals
por: Gerard, Patrick, et al.
Publicado: (2026)
por: Gerard, Patrick, et al.
Publicado: (2026)
o-MEGA: Optimized Methods for Explanation Generation and Analysis
por: Kriš, Ľuboš, et al.
Publicado: (2025)
por: Kriš, Ľuboš, et al.
Publicado: (2025)
Compound AI Systems Optimization: A Survey of Methods, Challenges, and Future Directions
por: Lee, Yu-Ang, et al.
Publicado: (2025)
por: Lee, Yu-Ang, et al.
Publicado: (2025)
The Role of Entropy in Visual Grounding: Analysis and Optimization
por: Li, Shuo, et al.
Publicado: (2025)
por: Li, Shuo, et al.
Publicado: (2025)
Training LLMs Beyond Next Token Prediction -- Filling the Mutual Information Gap
por: Yang, Chun-Hao, et al.
Publicado: (2025)
por: Yang, Chun-Hao, et al.
Publicado: (2025)
Semantic Density Effect (SDE): Maximizing Information Per Token Improves LLM Accuracy
por: Ahmed, Amr
Publicado: (2026)
por: Ahmed, Amr
Publicado: (2026)
Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key
por: Wang, Tianle, et al.
Publicado: (2026)
por: Wang, Tianle, et al.
Publicado: (2026)
Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict
por: Chen, Yihang, et al.
Publicado: (2026)
por: Chen, Yihang, et al.
Publicado: (2026)
TECP: Token-Entropy Conformal Prediction for LLMs
por: Xu, Beining, et al.
Publicado: (2025)
por: Xu, Beining, et al.
Publicado: (2025)
Entropy-aware Masking for Masked Language Modeling
por: Srinivasagan, Gokul, et al.
Publicado: (2026)
por: Srinivasagan, Gokul, et al.
Publicado: (2026)
DER-GCN: Dialogue and Event Relation-Aware Graph Convolutional Neural Network for Multimodal Dialogue Emotion Recognition
por: Ai, Wei, et al.
Publicado: (2023)
por: Ai, Wei, et al.
Publicado: (2023)
Agentic Entropy-Balanced Policy Optimization
por: Dong, Guanting, et al.
Publicado: (2025)
por: Dong, Guanting, et al.
Publicado: (2025)
GTPO: Stabilizing Group Relative Policy Optimization via Gradient and Entropy Control
por: Simoni, Marco, et al.
Publicado: (2025)
por: Simoni, Marco, et al.
Publicado: (2025)
IGOT: Information Gain Optimized Tokenizer on Domain Adaptive Pretraining
por: Feng, Dawei, et al.
Publicado: (2024)
por: Feng, Dawei, et al.
Publicado: (2024)
Entropy-based Exploration Conduction for Multi-step Reasoning
por: Zhang, Jinghan, et al.
Publicado: (2025)
por: Zhang, Jinghan, et al.
Publicado: (2025)
Entropy-Gated Branching for Efficient Test-Time Reasoning
por: Li, Xianzhi, et al.
Publicado: (2025)
por: Li, Xianzhi, et al.
Publicado: (2025)
Evaluation of the Automated Labeling Method for Taxonomic Nomenclature Through Prompt-Optimized Large Language Model
por: Inoshita, Keito, et al.
Publicado: (2025)
por: Inoshita, Keito, et al.
Publicado: (2025)
PPO-BR: Dual-Signal Entropy-Reward Adaptation for Trust Region Policy Optimization
por: Rahman, Ben
Publicado: (2025)
por: Rahman, Ben
Publicado: (2025)
Entropy-Aware Speculative Decoding Toward Improved LLM Reasoning
por: Su, Tiancheng, et al.
Publicado: (2025)
por: Su, Tiancheng, et al.
Publicado: (2025)
Entropy-Based Block Pruning for Efficient Large Language Models
por: Yang, Liangwei, et al.
Publicado: (2025)
por: Yang, Liangwei, et al.
Publicado: (2025)
ETR: Entropy Trend Reward for Efficient Chain-of-Thought Reasoning
por: Xiong, Xuan, et al.
Publicado: (2026)
por: Xiong, Xuan, et al.
Publicado: (2026)
SE-GCL: An Event-Based Simple and Effective Graph Contrastive Learning for Text Representation
por: Meng, Tao, et al.
Publicado: (2024)
por: Meng, Tao, et al.
Publicado: (2024)
Relative Density Ratio Optimization for Stable and Statistically Consistent Model Alignment
por: Takahashi, Hiroshi, et al.
Publicado: (2026)
por: Takahashi, Hiroshi, et al.
Publicado: (2026)
Judge Q: Trainable Queries for Optimized Information Retention in KV Cache Eviction
por: Liu, Yijun, et al.
Publicado: (2025)
por: Liu, Yijun, et al.
Publicado: (2025)
Ejemplares similares
-
Revisiting the UID Hypothesis in LLM Reasoning Traces
por: Gwak, Minju, et al.
Publicado: (2025) -
TokenTiming: A Dynamic Alignment Method for Universal Speculative Decoding Model Pairs
por: Xiao, Sibo, et al.
Publicado: (2025) -
PLHF: Prompt Optimization with Few-Shot Human Feedback
por: Yang, Chun-Pai, et al.
Publicado: (2025) -
Decoupling Strategy and Execution in Task-Focused Dialogue via Goal-Oriented Preference Optimization
por: Xu, Jingyi, et al.
Publicado: (2026) -
Mechanistic Decoding of Cognitive Constructs in Large Language Models
por: Shou, Yitong, et al.
Publicado: (2026)