From Hallucination to Structure Snowballing: The Alignment Tax of Constrained Decoding in LLM Reflection
Fuente:
arXiv
Guardado en:
| Autor principal: | Zhou, Hongxu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
por: Zhong, Weihong, et al.
Publicado: (2024)
por: Zhong, Weihong, et al.
Publicado: (2024)
The Hallucination Tax of Reinforcement Finetuning
por: Song, Linxin, et al.
Publicado: (2025)
por: Song, Linxin, et al.
Publicado: (2025)
Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning
por: Gan, Zeyu, et al.
Publicado: (2025)
por: Gan, Zeyu, et al.
Publicado: (2025)
The UNDO Flip-Flop: A Controlled Probe for Reversible Semantic State Management in State Space Model
por: Zhou, Hongxu
Publicado: (2026)
por: Zhou, Hongxu
Publicado: (2026)
Joint Effects of Argumentation Theory, Audio Modality and Data Enrichment on LLM-Based Fallacy Classification
por: Zhou, Hongxu, et al.
Publicado: (2025)
por: Zhou, Hongxu, et al.
Publicado: (2025)
MM-Snowball: Evaluating and Mitigating Hallucination Snowballing in Multimodal Multi-Turn Dialogue
por: Jiang, Yue, et al.
Publicado: (2026)
por: Jiang, Yue, et al.
Publicado: (2026)
Online Merging Optimizers for Boosting Rewards and Mitigating Tax in Alignment
por: Lu, Keming, et al.
Publicado: (2024)
por: Lu, Keming, et al.
Publicado: (2024)
Transfer Q Star: Principled Decoding for LLM Alignment
por: Chakraborty, Souradip, et al.
Publicado: (2024)
por: Chakraborty, Souradip, et al.
Publicado: (2024)
Flexible and Efficient Grammar-Constrained Decoding
por: Park, Kanghee, et al.
Publicado: (2025)
por: Park, Kanghee, et al.
Publicado: (2025)
Copy-as-Decode: Grammar-Constrained Parallel Prefill for LLM Editing
por: Liu, Ziyang
Publicado: (2026)
por: Liu, Ziyang
Publicado: (2026)
Collab: Controlled Decoding using Mixture of Agents for LLM Alignment
por: Chakraborty, Souradip, et al.
Publicado: (2025)
por: Chakraborty, Souradip, et al.
Publicado: (2025)
Reducing Tool Hallucination via Reliability Alignment
por: Xu, Hongshen, et al.
Publicado: (2024)
por: Xu, Hongshen, et al.
Publicado: (2024)
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
por: Wang, Chao, et al.
Publicado: (2025)
por: Wang, Chao, et al.
Publicado: (2025)
MAD: Multi-Alignment MEG-to-Text Decoding
por: Yang, Yiqian, et al.
Publicado: (2024)
por: Yang, Yiqian, et al.
Publicado: (2024)
Safety Alignment as Continual Learning: Mitigating the Alignment Tax via Orthogonal Gradient Projection
por: Sun, Guanglong, et al.
Publicado: (2026)
por: Sun, Guanglong, et al.
Publicado: (2026)
LLM Hallucination Detection: HSAD
por: Li, JinXin, et al.
Publicado: (2025)
por: Li, JinXin, et al.
Publicado: (2025)
Shakespearean Sparks: The Dance of Hallucination and Creativity in LLMs' Decoding Layers
por: He, Zicong, et al.
Publicado: (2025)
por: He, Zicong, et al.
Publicado: (2025)
Draft-Conditioned Constrained Decoding for Structured Generation in LLMs
por: Reddy, Avinash, et al.
Publicado: (2026)
por: Reddy, Avinash, et al.
Publicado: (2026)
Safety Alignment via Constrained Knowledge Unlearning
por: Shi, Zesheng, et al.
Publicado: (2025)
por: Shi, Zesheng, et al.
Publicado: (2025)
TaxPraBen: A Scalable Benchmark for Structured Evaluation of LLMs in Chinese Real-World Tax Practice
por: Hu, Gang, et al.
Publicado: (2026)
por: Hu, Gang, et al.
Publicado: (2026)
Alignment-Augmented Speculative Decoding with Alignment Sampling and Conditional Verification
por: Wang, Jikai, et al.
Publicado: (2025)
por: Wang, Jikai, et al.
Publicado: (2025)
Banishing LLM Hallucinations Requires Rethinking Generalization
por: Li, Johnny, et al.
Publicado: (2024)
por: Li, Johnny, et al.
Publicado: (2024)
Constrained Decoding with Speculative Lookaheads
por: Nakshatri, Nishanth, et al.
Publicado: (2024)
por: Nakshatri, Nishanth, et al.
Publicado: (2024)
GRAD: Graph-Retrieved Adaptive Decoding for Hallucination Mitigation
por: Nguyen, Manh, et al.
Publicado: (2025)
por: Nguyen, Manh, et al.
Publicado: (2025)
Grammar-Constrained Decoding for Structured NLP Tasks without Finetuning
por: Geng, Saibo, et al.
Publicado: (2023)
por: Geng, Saibo, et al.
Publicado: (2023)
Schema Key Wording as an Instruction Channel in Structured Generation under Constrained Decoding
por: Le, Yifan
Publicado: (2026)
por: Le, Yifan
Publicado: (2026)
Reinforcement Learning with Semantic Rewards Enables Low-Resource Language Expansion without Alignment Tax
por: Su, Zeli, et al.
Publicado: (2026)
por: Su, Zeli, et al.
Publicado: (2026)
MARCH: Multi-Agent Reinforced Self-Check for LLM Hallucination
por: Li, Zhuo, et al.
Publicado: (2026)
por: Li, Zhuo, et al.
Publicado: (2026)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
por: Zhang, Xiaoying, et al.
Publicado: (2024)
por: Zhang, Xiaoying, et al.
Publicado: (2024)
Multi-Drafter Speculative Decoding with Alignment Feedback
por: Kim, Taehyeon, et al.
Publicado: (2026)
por: Kim, Taehyeon, et al.
Publicado: (2026)
The Format Tax
por: Lee, Ivan Yee, et al.
Publicado: (2026)
por: Lee, Ivan Yee, et al.
Publicado: (2026)
Detecting Hallucinations in Authentic LLM-Human Interactions
por: Ren, Yujie, et al.
Publicado: (2025)
por: Ren, Yujie, et al.
Publicado: (2025)
HICD: Hallucination-Inducing via Attention Dispersion for Contrastive Decoding to Mitigate Hallucinations in Large Language Models
por: Jiang, Xinyan, et al.
Publicado: (2025)
por: Jiang, Xinyan, et al.
Publicado: (2025)
On-Policy Self-Alignment with Fine-grained Knowledge Feedback for Hallucination Mitigation
por: Wen, Xueru, et al.
Publicado: (2024)
por: Wen, Xueru, et al.
Publicado: (2024)
Disperse-Then-Merge: Pushing the Limits of Instruction Tuning via Alignment Tax Reduction
por: Fu, Tingchen, et al.
Publicado: (2024)
por: Fu, Tingchen, et al.
Publicado: (2024)
Mitigating Multimodal Hallucinations via Gradient-based Self-Reflection
por: Wang, Shan, et al.
Publicado: (2025)
por: Wang, Shan, et al.
Publicado: (2025)
GRIFFIN: Effective Token Alignment for Faster Speculative Decoding
por: Hu, Shijing, et al.
Publicado: (2025)
por: Hu, Shijing, et al.
Publicado: (2025)
FLAP: Flow-Adhering Planning with Constrained Decoding in LLMs
por: Roy, Shamik, et al.
Publicado: (2024)
por: Roy, Shamik, et al.
Publicado: (2024)
Decoding Alignment: A Critical Survey of LLM Development Initiatives through Value-setting and Data-centric Lens
por: Chalkidis, Ilias
Publicado: (2025)
por: Chalkidis, Ilias
Publicado: (2025)
Tool-Call Dependency Structure is Linearly Decodable in LLM Agent Residual Streams
por: Sun, Tianda, et al.
Publicado: (2026)
por: Sun, Tianda, et al.
Publicado: (2026)
Ejemplares similares
-
Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
por: Zhong, Weihong, et al.
Publicado: (2024) -
The Hallucination Tax of Reinforcement Finetuning
por: Song, Linxin, et al.
Publicado: (2025) -
Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning
por: Gan, Zeyu, et al.
Publicado: (2025) -
The UNDO Flip-Flop: A Controlled Probe for Reversible Semantic State Management in State Space Model
por: Zhou, Hongxu
Publicado: (2026) -
Joint Effects of Argumentation Theory, Audio Modality and Data Enrichment on LLM-Based Fallacy Classification
por: Zhou, Hongxu, et al.
Publicado: (2025)