Entropy-Aligned Decoding of LMs for Better Writing and Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ahmed, Kareem, Singh, Sameer |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MUStReason: A Benchmark for Diagnosing Pragmatic Reasoning in Video-LMs for Multimodal Sarcasm Detection
von: Saha, Anisha, et al.
Veröffentlicht: (2025)
von: Saha, Anisha, et al.
Veröffentlicht: (2025)
Sample Smart, Not Hard: Correctness-First Decoding for Better Reasoning in LLMs
von: Li, Xueyan, et al.
Veröffentlicht: (2025)
von: Li, Xueyan, et al.
Veröffentlicht: (2025)
Nudging: Inference-time Alignment of LLMs via Guided Decoding
von: Fei, Yu, et al.
Veröffentlicht: (2024)
von: Fei, Yu, et al.
Veröffentlicht: (2024)
Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
von: Damani, Mehul, et al.
Veröffentlicht: (2025)
von: Damani, Mehul, et al.
Veröffentlicht: (2025)
Compositional preference models for aligning LMs
von: Go, Dongyoung, et al.
Veröffentlicht: (2023)
von: Go, Dongyoung, et al.
Veröffentlicht: (2023)
When Thoughts Meet Facts: Reusable Reasoning for Long-Context LMs
von: Jeong, Soyeong, et al.
Veröffentlicht: (2025)
von: Jeong, Soyeong, et al.
Veröffentlicht: (2025)
Enabling Approximate Joint Sampling in Diffusion LMs
von: Bansal, Parikshit, et al.
Veröffentlicht: (2025)
von: Bansal, Parikshit, et al.
Veröffentlicht: (2025)
Training Bilingual LMs with Data Constraints in the Targeted Language
von: Seto, Skyler, et al.
Veröffentlicht: (2024)
von: Seto, Skyler, et al.
Veröffentlicht: (2024)
Grammar-Aligned Decoding
von: Park, Kanghee, et al.
Veröffentlicht: (2024)
von: Park, Kanghee, et al.
Veröffentlicht: (2024)
Aligning LLMs by Predicting Preferences from User Writing Samples
von: Aroca-Ouellette, Stéphane, et al.
Veröffentlicht: (2025)
von: Aroca-Ouellette, Stéphane, et al.
Veröffentlicht: (2025)
Diffusion LMs Can Approximate Optimal Infilling Lengths Implicitly
von: Liu, Hengchang, et al.
Veröffentlicht: (2026)
von: Liu, Hengchang, et al.
Veröffentlicht: (2026)
Learning How to Ask: Querying LMs with Mixtures of Soft Prompts
von: Qin, Guanghui, et al.
Veröffentlicht: (2021)
von: Qin, Guanghui, et al.
Veröffentlicht: (2021)
Dodo: Dynamic Contextual Compression for Decoder-only LMs
von: Qin, Guanghui, et al.
Veröffentlicht: (2023)
von: Qin, Guanghui, et al.
Veröffentlicht: (2023)
Reflect-RL: Two-Player Online RL Fine-Tuning for LMs
von: Zhou, Runlong, et al.
Veröffentlicht: (2024)
von: Zhou, Runlong, et al.
Veröffentlicht: (2024)
A Little Help Goes a Long Way: Efficient LLM Training by Leveraging Small LMs
von: Rawat, Ankit Singh, et al.
Veröffentlicht: (2024)
von: Rawat, Ankit Singh, et al.
Veröffentlicht: (2024)
ALIEN: Aligned Entropy Head for Improving Uncertainty Estimation of LLMs
von: Zabolotnyi, Artem, et al.
Veröffentlicht: (2025)
von: Zabolotnyi, Artem, et al.
Veröffentlicht: (2025)
Decoder-Hybrid-Decoder Architecture for Efficient Reasoning with Long Generation
von: Ren, Liliang, et al.
Veröffentlicht: (2025)
von: Ren, Liliang, et al.
Veröffentlicht: (2025)
Scaling Speculative Decoding with Lookahead Reasoning
von: Fu, Yichao, et al.
Veröffentlicht: (2025)
von: Fu, Yichao, et al.
Veröffentlicht: (2025)
DeCAL Tokenwise Compression
von: Panwar, Sameer
Veröffentlicht: (2025)
von: Panwar, Sameer
Veröffentlicht: (2025)
Sentence-level Reward Model can Generalize Better for Aligning LLM from Human Preference
von: Qiu, Wenjie, et al.
Veröffentlicht: (2025)
von: Qiu, Wenjie, et al.
Veröffentlicht: (2025)
Controllable Generation via Locally Constrained Resampling
von: Ahmed, Kareem, et al.
Veröffentlicht: (2024)
von: Ahmed, Kareem, et al.
Veröffentlicht: (2024)
Grounded Token Initialization for New Vocabulary in LMs for Generative Recommendation
von: Chen, Daiwei, et al.
Veröffentlicht: (2026)
von: Chen, Daiwei, et al.
Veröffentlicht: (2026)
SoftLMs: Efficient Adaptive Low-Rank Approximation of Language Models using Soft-Thresholding Mechanism
von: Bhatnagar, Priyansh, et al.
Veröffentlicht: (2024)
von: Bhatnagar, Priyansh, et al.
Veröffentlicht: (2024)
POSS: Position Specialist Generates Better Draft for Speculative Decoding
von: Huang, Langlin, et al.
Veröffentlicht: (2025)
von: Huang, Langlin, et al.
Veröffentlicht: (2025)
Better Think Thrice: Learning to Reason Causally with Double Counterfactual Consistency
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
Distribution-Aligned Sequence Distillation for Superior Long-CoT Reasoning
von: Yan, Shaotian, et al.
Veröffentlicht: (2026)
von: Yan, Shaotian, et al.
Veröffentlicht: (2026)
Addressing the Ecological Fallacy in Larger LMs with Human Context
von: Soni, Nikita, et al.
Veröffentlicht: (2026)
von: Soni, Nikita, et al.
Veröffentlicht: (2026)
Causal Reasoning Favors Encoders: On The Limits of Decoder-Only Models
von: Roy, Amartya, et al.
Veröffentlicht: (2025)
von: Roy, Amartya, et al.
Veröffentlicht: (2025)
Better LLM Reasoning via Dual-Play
von: Zhang, Zhengxin, et al.
Veröffentlicht: (2025)
von: Zhang, Zhengxin, et al.
Veröffentlicht: (2025)
TIDE: Textual Identity Detection for Evaluating and Augmenting Classification and Language Models
von: Klu, Emmanuel, et al.
Veröffentlicht: (2023)
von: Klu, Emmanuel, et al.
Veröffentlicht: (2023)
Towards Trustworthy Multimodal Moderation via Policy-Aligned Reasoning and Hierarchical Labeling
von: Li, Anqi, et al.
Veröffentlicht: (2025)
von: Li, Anqi, et al.
Veröffentlicht: (2025)
Entropy Adaptive Decoding: Dynamic Model Switching for Efficient Inference
von: Simonds, Toby
Veröffentlicht: (2025)
von: Simonds, Toby
Veröffentlicht: (2025)
Distillation Contrastive Decoding: Improving LLMs Reasoning with Contrastive Decoding and Distillation
von: Phan, Phuc, et al.
Veröffentlicht: (2024)
von: Phan, Phuc, et al.
Veröffentlicht: (2024)
ShifaMind: A Multiplicative Concept Bottleneck for Interpretable ICD-10 Coding
von: Syed, Mohammed Sameer, et al.
Veröffentlicht: (2026)
von: Syed, Mohammed Sameer, et al.
Veröffentlicht: (2026)
MindAlign: Bridging EEG, Vision, and Language for Zero-Shot Visual Decoding
von: Chen, Zexuan, et al.
Veröffentlicht: (2026)
von: Chen, Zexuan, et al.
Veröffentlicht: (2026)
Where is the signal in tokenization space?
von: Geh, Renato Lui, et al.
Veröffentlicht: (2024)
von: Geh, Renato Lui, et al.
Veröffentlicht: (2024)
SceneAlign: Aligning Multimodal Reasoning to Scene Graphs in Complex Visual Scenes
von: Wang, Chuhan, et al.
Veröffentlicht: (2026)
von: Wang, Chuhan, et al.
Veröffentlicht: (2026)
Parallel Token Prediction for Language Models
von: Draxler, Felix, et al.
Veröffentlicht: (2025)
von: Draxler, Felix, et al.
Veröffentlicht: (2025)
A Pseudo-Semantic Loss for Autoregressive Models with Logical Constraints
von: Ahmed, Kareem, et al.
Veröffentlicht: (2023)
von: Ahmed, Kareem, et al.
Veröffentlicht: (2023)
Restoring Exploration after Post-Training: Latent Exploration Decoding for Large Reasoning Models
von: Tan, Wenhui, et al.
Veröffentlicht: (2026)
von: Tan, Wenhui, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MUStReason: A Benchmark for Diagnosing Pragmatic Reasoning in Video-LMs for Multimodal Sarcasm Detection
von: Saha, Anisha, et al.
Veröffentlicht: (2025) -
Sample Smart, Not Hard: Correctness-First Decoding for Better Reasoning in LLMs
von: Li, Xueyan, et al.
Veröffentlicht: (2025) -
Nudging: Inference-time Alignment of LLMs via Guided Decoding
von: Fei, Yu, et al.
Veröffentlicht: (2024) -
Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
von: Damani, Mehul, et al.
Veröffentlicht: (2025) -
Compositional preference models for aligning LMs
von: Go, Dongyoung, et al.
Veröffentlicht: (2023)