KARL: Mitigating Hallucinations in LLMs via Knowledge-Boundary-Aware Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Gao, Cheng, Huang, Cheng, Luo, Kangyang, Qiao, Ziqing, Si, Shuzheng, Chen, Huimin, Xiao, Chaojun, Sun, Maosong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
H-Neurons: On the Existence, Impact, and Origin of Hallucination-Associated Neurons in LLMs
por: Gao, Cheng, et al.
Publicado: (2025)
por: Gao, Cheng, et al.
Publicado: (2025)
Aligning Large Language Models to Follow Instructions and Hallucinate Less via Effective Data Filtering
por: Si, Shuzheng, et al.
Publicado: (2025)
por: Si, Shuzheng, et al.
Publicado: (2025)
KARL: Knowledge-Aware Reasoning and Reinforcement Learning for Knowledge-Intensive Visual Grounding
por: Ma, Xinyu, et al.
Publicado: (2025)
por: Ma, Xinyu, et al.
Publicado: (2025)
InFi-Check: Interpretable and Fine-Grained Fact-Checking of LLMs
por: Bai, Yuzhuo, et al.
Publicado: (2026)
por: Bai, Yuzhuo, et al.
Publicado: (2026)
Teaching Large Language Models to Maintain Contextual Faithfulness via Synthetic Tasks and Reinforcement Learning
por: Si, Shuzheng, et al.
Publicado: (2025)
por: Si, Shuzheng, et al.
Publicado: (2025)
FaithLens: Detecting and Explaining Faithfulness Hallucination
por: Si, Shuzheng, et al.
Publicado: (2025)
por: Si, Shuzheng, et al.
Publicado: (2025)
A Goal Without a Plan Is Just a Wish: Efficient and Effective Global Planner Training for Long-Horizon Agent Tasks
por: Si, Shuzheng, et al.
Publicado: (2025)
por: Si, Shuzheng, et al.
Publicado: (2025)
GLTW: Joint Improved Graph Transformer and LLM via Three-Word Language for Knowledge Graph Completion
por: Luo, Kangyang, et al.
Publicado: (2025)
por: Luo, Kangyang, et al.
Publicado: (2025)
ImCoref-CeS: An Improved Lightweight Pipeline for Coreference Resolution with LLM-based Checker-Splitter Refinement
por: Luo, Kangyang, et al.
Publicado: (2025)
por: Luo, Kangyang, et al.
Publicado: (2025)
GATEAU: Selecting Influential Samples for Long Context Alignment
por: Si, Shuzheng, et al.
Publicado: (2024)
por: Si, Shuzheng, et al.
Publicado: (2024)
From Unaligned to Aligned: Scaling Multilingual LLMs with Multi-Way Parallel Corpora
por: Shen, Yingli, et al.
Publicado: (2025)
por: Shen, Yingli, et al.
Publicado: (2025)
Enhancing Legal Case Retrieval via Scaling High-quality Synthetic Query-Candidate Pairs
por: Gao, Cheng, et al.
Publicado: (2024)
por: Gao, Cheng, et al.
Publicado: (2024)
KARL: Knowledge-Aware Retrieval and Representations aid Retention and Learning in Students
por: Shu, Matthew, et al.
Publicado: (2024)
por: Shu, Matthew, et al.
Publicado: (2024)
Med-HEAL: Analyzing and Mitigating Hallucinations in Medical LLMs with Hallucination-Aware In-Context Learning
por: Liao, Yiming, et al.
Publicado: (2026)
por: Liao, Yiming, et al.
Publicado: (2026)
FactNet: A Billion-Scale Knowledge Graph for Multilingual Factual Grounding
por: Shen, Yingli, et al.
Publicado: (2026)
por: Shen, Yingli, et al.
Publicado: (2026)
Rowen: Adaptive Retrieval-Augmented Generation for Hallucination Mitigation in LLMs
por: Ding, Hanxing, et al.
Publicado: (2024)
por: Ding, Hanxing, et al.
Publicado: (2024)
Looking Beyond Text: Reducing Language bias in Large Vision-Language Models via Multimodal Dual-Attention and Soft-Image Guidance
por: Zhao, Haozhe, et al.
Publicado: (2024)
por: Zhao, Haozhe, et al.
Publicado: (2024)
MEIC-DT: Memory-Efficient Incremental Clustering for Long-Text Coreference Resolution with Dual-Threshold Constraints
por: Luo, Kangyang, et al.
Publicado: (2025)
por: Luo, Kangyang, et al.
Publicado: (2025)
Mitigating Hallucination on Hallucination in RAG via Ensemble Voting
por: Xie, Zequn, et al.
Publicado: (2026)
por: Xie, Zequn, et al.
Publicado: (2026)
DCAD-2000: A Multilingual Dataset across 2000+ Languages with Data Cleaning as Anomaly Detection
por: Shen, Yingli, et al.
Publicado: (2025)
por: Shen, Yingli, et al.
Publicado: (2025)
InfLLM: Training-Free Long-Context Extrapolation for LLMs with an Efficient Context Memory
por: Xiao, Chaojun, et al.
Publicado: (2024)
por: Xiao, Chaojun, et al.
Publicado: (2024)
RhinoInsight: Improving Deep Research through Control Mechanisms for Model Behavior and Context
por: Lei, Yu, et al.
Publicado: (2025)
por: Lei, Yu, et al.
Publicado: (2025)
Reducing Hallucinations in LLMs via Factuality-Aware Preference Learning
por: Chaduvula, Sindhuja, et al.
Publicado: (2026)
por: Chaduvula, Sindhuja, et al.
Publicado: (2026)
Beyond Reasoning: Reinforcement Learning Unlocks Parametric Knowledge in LLMs
por: Yang, Wanli, et al.
Publicado: (2026)
por: Yang, Wanli, et al.
Publicado: (2026)
When Personalization Misleads: Understanding and Mitigating Hallucinations in Personalized LLMs
por: Sun, Zhongxiang, et al.
Publicado: (2026)
por: Sun, Zhongxiang, et al.
Publicado: (2026)
Think More, Hallucinate Less: Mitigating Hallucinations via Dual Process of Fast and Slow Thinking
por: Cheng, Xiaoxue, et al.
Publicado: (2025)
por: Cheng, Xiaoxue, et al.
Publicado: (2025)
Mitigating Object and Action Hallucinations in Multimodal LLMs via Self-Augmented Contrastive Alignment
por: Chang, Kai-Po, et al.
Publicado: (2025)
por: Chang, Kai-Po, et al.
Publicado: (2025)
Embedding Domain Knowledge for Large Language Models via Reinforcement Learning from Augmented Generation
por: Nie, Chaojun, et al.
Publicado: (2025)
por: Nie, Chaojun, et al.
Publicado: (2025)
Densing Law of LLMs
por: Xiao, Chaojun, et al.
Publicado: (2024)
por: Xiao, Chaojun, et al.
Publicado: (2024)
Delta -- Contrastive Decoding Mitigates Text Hallucinations in Large Language Models
por: Huang, Cheng Peng, et al.
Publicado: (2025)
por: Huang, Cheng Peng, et al.
Publicado: (2025)
Improving the Robustness of Distantly-Supervised Named Entity Recognition via Uncertainty-Aware Teacher Learning and Student-Student Collaborative Learning
por: Si, Shuzheng, et al.
Publicado: (2023)
por: Si, Shuzheng, et al.
Publicado: (2023)
Understanding New-Knowledge-Induced Factual Hallucinations in LLMs: Analysis and Interpretation
por: Dang, Renfei, et al.
Publicado: (2025)
por: Dang, Renfei, et al.
Publicado: (2025)
DRAG: Distilling RAG for SLMs from LLMs to Transfer Knowledge and Mitigate Hallucination via Evidence and Graph-based Distillation
por: Chen, Jennifer, et al.
Publicado: (2025)
por: Chen, Jennifer, et al.
Publicado: (2025)
From Context to EDUs: Faithful and Structured Context Compression via Elementary Discourse Unit Decomposition
por: Zhou, Yiqing, et al.
Publicado: (2025)
por: Zhou, Yiqing, et al.
Publicado: (2025)
Search-Based LLMs for Code Optimization
por: Gao, Shuzheng, et al.
Publicado: (2024)
por: Gao, Shuzheng, et al.
Publicado: (2024)
Bridging External and Parametric Knowledge: Mitigating Hallucination of LLMs with Shared-Private Semantic Synergy in Dual-Stream Knowledge
por: Sui, Yi, et al.
Publicado: (2025)
por: Sui, Yi, et al.
Publicado: (2025)
KARL: Knowledge Agents via Reinforcement Learning
por: Chang, Jonathan D., et al.
Publicado: (2026)
por: Chang, Jonathan D., et al.
Publicado: (2026)
KARE-RAG: Knowledge-Aware Refinement and Enhancement for RAG
por: Li, Yongjian, et al.
Publicado: (2025)
por: Li, Yongjian, et al.
Publicado: (2025)
Mitigating Language-Level Performance Disparity in mPLMs via Teacher Language Selection and Cross-lingual Self-Distillation
por: Zhao, Haozhe, et al.
Publicado: (2024)
por: Zhao, Haozhe, et al.
Publicado: (2024)
Dynamic Attention-Guided Context Decoding for Mitigating Context Faithfulness Hallucinations in Large Language Models
por: Huang, Yanwen, et al.
Publicado: (2025)
por: Huang, Yanwen, et al.
Publicado: (2025)
Ejemplares similares
-
H-Neurons: On the Existence, Impact, and Origin of Hallucination-Associated Neurons in LLMs
por: Gao, Cheng, et al.
Publicado: (2025) -
Aligning Large Language Models to Follow Instructions and Hallucinate Less via Effective Data Filtering
por: Si, Shuzheng, et al.
Publicado: (2025) -
KARL: Knowledge-Aware Reasoning and Reinforcement Learning for Knowledge-Intensive Visual Grounding
por: Ma, Xinyu, et al.
Publicado: (2025) -
InFi-Check: Interpretable and Fine-Grained Fact-Checking of LLMs
por: Bai, Yuzhuo, et al.
Publicado: (2026) -
Teaching Large Language Models to Maintain Contextual Faithfulness via Synthetic Tasks and Reinforcement Learning
por: Si, Shuzheng, et al.
Publicado: (2025)