Knowing but Not Correcting: Routine Task Requests Suppress Factual Correction in LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Zixuan, Lin, Hao, Chen, Zizhe, Tian, Yizhou, Yang, Garry, Wang, Depeng, Guo, Ya, Zhu, Huijia, Cheng, James |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hista and Numca: Estimate State Value Effectively for LLM Reinforcement Learning
von: Chen, Zizhe, et al.
Veröffentlicht: (2026)
von: Chen, Zizhe, et al.
Veröffentlicht: (2026)
EchoingPixels: Cross-Modal Adaptive Token Reduction for Efficient Audio-Visual LLMs
von: Gong, Chao, et al.
Veröffentlicht: (2025)
von: Gong, Chao, et al.
Veröffentlicht: (2025)
AVID: A Benchmark for Omni-Modal Audio-Visual Inconsistency Understanding via Agent-Driven Construction
von: Chen, Zixuan, et al.
Veröffentlicht: (2026)
von: Chen, Zixuan, et al.
Veröffentlicht: (2026)
CriticBench: Benchmarking LLMs for Critique-Correct Reasoning
von: Lin, Zicheng, et al.
Veröffentlicht: (2024)
von: Lin, Zicheng, et al.
Veröffentlicht: (2024)
RW-TTT: Batched Serving for Request-Owned Test-Time Training State
von: Yang, Jian, et al.
Veröffentlicht: (2026)
von: Yang, Jian, et al.
Veröffentlicht: (2026)
Low-Rank Correction for Quantized LLMs
von: Scetbon, Meyer, et al.
Veröffentlicht: (2024)
von: Scetbon, Meyer, et al.
Veröffentlicht: (2024)
KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality
von: Ren, Baochang, et al.
Veröffentlicht: (2025)
von: Ren, Baochang, et al.
Veröffentlicht: (2025)
Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression
von: Mao, Yixiu, et al.
Veröffentlicht: (2024)
von: Mao, Yixiu, et al.
Veröffentlicht: (2024)
RAC: Efficient LLM Factuality Correction with Retrieval Augmentation
von: Li, Changmao, et al.
Veröffentlicht: (2024)
von: Li, Changmao, et al.
Veröffentlicht: (2024)
Reinforcement Learning from Denoising Feedback
von: He, Qi, et al.
Veröffentlicht: (2026)
von: He, Qi, et al.
Veröffentlicht: (2026)
KnowCoder: Coding Structured Knowledge into LLMs for Universal Information Extraction
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
Improving Expressive Power of Spectral Graph Neural Networks with Eigenvalue Correction
von: Lu, Kangkang, et al.
Veröffentlicht: (2024)
von: Lu, Kangkang, et al.
Veröffentlicht: (2024)
Can LLMs Refuse Questions They Do Not Know? Measuring Knowledge-Aware Refusal in Factual Tasks
von: Pan, Wenbo, et al.
Veröffentlicht: (2025)
von: Pan, Wenbo, et al.
Veröffentlicht: (2025)
Mechanistic Interpretability of Code Correctness in LLMs via Sparse Autoencoders
von: Tahimic, Kriz, et al.
Veröffentlicht: (2025)
von: Tahimic, Kriz, et al.
Veröffentlicht: (2025)
Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in Coding Tasks
von: Li, Yuangang, et al.
Veröffentlicht: (2026)
von: Li, Yuangang, et al.
Veröffentlicht: (2026)
Efficient Multi-Task Reinforcement Learning via Task-Specific Action Correction
von: Feng, Jinyuan, et al.
Veröffentlicht: (2024)
von: Feng, Jinyuan, et al.
Veröffentlicht: (2024)
Reactive Model Correction: Mitigating Harm to Task-Relevant Features via Conditional Bias Suppression
von: Bareeva, Dilyara, et al.
Veröffentlicht: (2024)
von: Bareeva, Dilyara, et al.
Veröffentlicht: (2024)
Towards a Holistic Evaluation of LLMs on Factual Knowledge Recall
von: Yuan, Jiaqing, et al.
Veröffentlicht: (2024)
von: Yuan, Jiaqing, et al.
Veröffentlicht: (2024)
Teaching LLMs for Step-Level Automatic Math Correction via Reinforcement Learning
von: Li, Junsong, et al.
Veröffentlicht: (2025)
von: Li, Junsong, et al.
Veröffentlicht: (2025)
ORGEval: Graph-Theoretic Evaluation of LLMs in Optimization Modeling
von: Wang, Zhuohan, et al.
Veröffentlicht: (2025)
von: Wang, Zhuohan, et al.
Veröffentlicht: (2025)
Is Conformal Factuality for RAG-based LLMs Robust? Novel Metrics and Systematic Insights
von: Chen, Yi, et al.
Veröffentlicht: (2026)
von: Chen, Yi, et al.
Veröffentlicht: (2026)
Learning to Correct for QA Reasoning with Black-box LLMs
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
von: Kim, Jaehyung, et al.
Veröffentlicht: (2024)
MathlibPR: Pull Request Merge-Readiness Benchmark for Formal Mathematical Libraries
von: Xie, Zixuan, et al.
Veröffentlicht: (2026)
von: Xie, Zixuan, et al.
Veröffentlicht: (2026)
Know When You're Wrong: Aligning Confidence with Correctness for LLM Error Detection
von: Xiaohu, Xie, et al.
Veröffentlicht: (2026)
von: Xiaohu, Xie, et al.
Veröffentlicht: (2026)
Correcting Suppressed Log-Probabilities in Language Models with Post-Transformer Adapters
von: Sanchez, Bryan
Veröffentlicht: (2026)
von: Sanchez, Bryan
Veröffentlicht: (2026)
Persuasion Tokens for Editing Factual Knowledge in LLMs
von: Youssef, Paul, et al.
Veröffentlicht: (2026)
von: Youssef, Paul, et al.
Veröffentlicht: (2026)
On LLMs' Internal Representation of Code Correctness
von: Ribeiro, Francisco, et al.
Veröffentlicht: (2025)
von: Ribeiro, Francisco, et al.
Veröffentlicht: (2025)
KnowHalu: Hallucination Detection via Multi-Form Knowledge Based Factual Checking
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
Attention Head Entropy of LLMs Predicts Answer Correctness
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2026)
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2026)
Evaluating the Correctness of Inference Patterns Used by LLMs for Judgment
von: Chen, Lu, et al.
Veröffentlicht: (2024)
von: Chen, Lu, et al.
Veröffentlicht: (2024)
Hide in Plain Sight: Clean-Label Backdoor for Auditing Membership Inference
von: Chen, Depeng, et al.
Veröffentlicht: (2024)
von: Chen, Depeng, et al.
Veröffentlicht: (2024)
MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models
von: Yang, Garry, et al.
Veröffentlicht: (2025)
von: Yang, Garry, et al.
Veröffentlicht: (2025)
Causality-Inspired Safe Residual Correction for Multivariate Time Series
von: Xie, Jianxiang, et al.
Veröffentlicht: (2025)
von: Xie, Jianxiang, et al.
Veröffentlicht: (2025)
When Models Know When They Do Not Know: Calibration, Cascading, and Cleaning
von: Hao, Chenjie, et al.
Veröffentlicht: (2026)
von: Hao, Chenjie, et al.
Veröffentlicht: (2026)
Factual Confidence of LLMs: on Reliability and Robustness of Current Estimators
von: Mahaut, Matéo, et al.
Veröffentlicht: (2024)
von: Mahaut, Matéo, et al.
Veröffentlicht: (2024)
Adaptive Requesting in Decentralized Edge Networks via Non-Stationary Bandits
von: Zhuang, Yi, et al.
Veröffentlicht: (2026)
von: Zhuang, Yi, et al.
Veröffentlicht: (2026)
Partial Domain Adaptation via Importance Sampling-based Shift Correction
von: Guo, Cheng-Jun, et al.
Veröffentlicht: (2025)
von: Guo, Cheng-Jun, et al.
Veröffentlicht: (2025)
Sirius: Contextual Sparsity with Correction for Efficient LLMs
von: Zhou, Yang, et al.
Veröffentlicht: (2024)
von: Zhou, Yang, et al.
Veröffentlicht: (2024)
CLFEC: A New Task for Unified Linguistic and Factual Error Correction in paragraph-level Chinese Professional Writing
von: Kai, Jian, et al.
Veröffentlicht: (2026)
von: Kai, Jian, et al.
Veröffentlicht: (2026)
Unsupervised Evaluation of Code LLMs with Round-Trip Correctness
von: Allamanis, Miltiadis, et al.
Veröffentlicht: (2024)
von: Allamanis, Miltiadis, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Hista and Numca: Estimate State Value Effectively for LLM Reinforcement Learning
von: Chen, Zizhe, et al.
Veröffentlicht: (2026) -
EchoingPixels: Cross-Modal Adaptive Token Reduction for Efficient Audio-Visual LLMs
von: Gong, Chao, et al.
Veröffentlicht: (2025) -
AVID: A Benchmark for Omni-Modal Audio-Visual Inconsistency Understanding via Agent-Driven Construction
von: Chen, Zixuan, et al.
Veröffentlicht: (2026) -
CriticBench: Benchmarking LLMs for Critique-Correct Reasoning
von: Lin, Zicheng, et al.
Veröffentlicht: (2024) -
RW-TTT: Batched Serving for Request-Owned Test-Time Training State
von: Yang, Jian, et al.
Veröffentlicht: (2026)