Enregistré dans:
| Auteurs principaux: | Jaaouine, Imane, King, Ross D. |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2512.00931 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Adaptive Activation Cancellation for Hallucination Mitigation in Large Language Models
par: Yocam, Eric, et autres
Publié: (2026)
par: Yocam, Eric, et autres
Publié: (2026)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
par: Wu, Shuai, et autres
Publié: (2026)
par: Wu, Shuai, et autres
Publié: (2026)
Weakly Supervised Distillation of Hallucination Signals into Transformer Representations
par: Salehmohamed, Shoaib Sadiq, et autres
Publié: (2026)
par: Salehmohamed, Shoaib Sadiq, et autres
Publié: (2026)
EvoIdeator: Evolving Scientific Ideas through Checklist-Grounded Reinforcement Learning
par: Sauter, Andreas, et autres
Publié: (2026)
par: Sauter, Andreas, et autres
Publié: (2026)
Beyond Hallucinations: A Composite Score for Measuring Reliability in Open-Source Large Language Models
par: Salla, Rohit Kumar, et autres
Publié: (2025)
par: Salla, Rohit Kumar, et autres
Publié: (2025)
Incentives or Ontology? A Structural Rebuttal to OpenAI's Hallucination Thesis
par: Ackermann, Richard, et autres
Publié: (2025)
par: Ackermann, Richard, et autres
Publié: (2025)
Adapting While Learning: Grounding LLMs for Scientific Problems with Intelligent Tool Usage Adaptation
par: Lyu, Bohan, et autres
Publié: (2024)
par: Lyu, Bohan, et autres
Publié: (2024)
Improving Commonsense Bias Classification by Mitigating the Influence of Demographic Terms
par: Lee, JinKyu, et autres
Publié: (2024)
par: Lee, JinKyu, et autres
Publié: (2024)
In-Context Fixation: When Demonstrated Labels Override Semantics in Few-Shot Classification
par: Liu, Ming
Publié: (2026)
par: Liu, Ming
Publié: (2026)
Synthius-Mem: Brain-Inspired Hallucination-Resistant Persona Memory Achieving 94.4% Memory Accuracy and 99.6% Adversarial Robustness on LoCoMo
par: Gadzhiev, Artem, et autres
Publié: (2026)
par: Gadzhiev, Artem, et autres
Publié: (2026)
Graph Your Way to Inspiration: Integrating Co-Author Graphs with Retrieval-Augmented Generation for Large Language Model Based Scientific Idea Generation
par: Xie, Pengzhen, et autres
Publié: (2025)
par: Xie, Pengzhen, et autres
Publié: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
par: Fadli, Samih
Publié: (2025)
par: Fadli, Samih
Publié: (2025)
Accelerating Complex Disease Treatment through Network Medicine and GenAI: A Case Study on Drug Repurposing for Breast Cancer
par: Hamed, Ahmed Abdeen, et autres
Publié: (2024)
par: Hamed, Ahmed Abdeen, et autres
Publié: (2024)
ImmigrationQA: A Source-Grounded Dataset and Small-Model Adaptation for U.S. Immigration Law
par: Shportun, Nazarii
Publié: (2026)
par: Shportun, Nazarii
Publié: (2026)
A Hierarchical Error Framework for Reliable Automated Coding in Communication Research: Applications to Health and Political Communication
par: Zhao, Zhilong, et autres
Publié: (2025)
par: Zhao, Zhilong, et autres
Publié: (2025)
Distilling Self-Consistency into Verbal Confidence: A Pre-Registered Negative Result and Post-Hoc Rescue on Gemma 3 4B
par: Cacioli, Jon-Paul
Publié: (2026)
par: Cacioli, Jon-Paul
Publié: (2026)
Re-Mask and Redirect: Exploiting Denoising Irreversibility in Diffusion Language Models
par: Singh, Arth
Publié: (2026)
par: Singh, Arth
Publié: (2026)
Let's Think Dot by Dot: Hidden Computation in Transformer Language Models
par: Pfau, Jacob, et autres
Publié: (2024)
par: Pfau, Jacob, et autres
Publié: (2024)
CRAFT: Clustered Regression for Adaptive Filtering of Training data
par: Panda, Parthasarathi, et autres
Publié: (2026)
par: Panda, Parthasarathi, et autres
Publié: (2026)
EMA Is Not All You Need: Mapping the Boundary Between Structure and Content in Recurrent Context
par: Singh, Arth
Publié: (2026)
par: Singh, Arth
Publié: (2026)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
par: Liu, Zhongxin, et autres
Publié: (2025)
par: Liu, Zhongxin, et autres
Publié: (2025)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
par: Wen, Yuqiao, et autres
Publié: (2024)
par: Wen, Yuqiao, et autres
Publié: (2024)
Conscious Gaze: Adaptive Attention Mechanisms for Hallucination Mitigation in Vision-Language Models
par: Bu, Weijue, et autres
Publié: (2025)
par: Bu, Weijue, et autres
Publié: (2025)
KAConvText: Novel Approach to Burmese Sentence Classification using Kolmogorov-Arnold Convolution
par: Thu, Ye Kyaw, et autres
Publié: (2025)
par: Thu, Ye Kyaw, et autres
Publié: (2025)
When Persuasion Overrides Truth in Multi-Agent LLM Debates: Introducing a Confidence-Weighted Persuasion Override Rate (CW-POR)
par: Agarwal, Mahak, et autres
Publié: (2025)
par: Agarwal, Mahak, et autres
Publié: (2025)
Bridging the Reasoning Gap: Small LLMs Can Plan with Generalised Strategies
par: Borro, Andrey, et autres
Publié: (2025)
par: Borro, Andrey, et autres
Publié: (2025)
Resource for Error Analysis in Text Simplification: New Taxonomy and Test Collection
par: Vendeville, Benjamin, et autres
Publié: (2025)
par: Vendeville, Benjamin, et autres
Publié: (2025)
Assessing Large Language Models on Islamic Legal Reasoning: Evidence from Inheritance Law Evaluation
par: Bouchekif, Abdessalam, et autres
Publié: (2025)
par: Bouchekif, Abdessalam, et autres
Publié: (2025)
Can AI Read Between The Lines? Benchmarking LLMs On Financial Nuance
par: Kubica, Dominick, et autres
Publié: (2025)
par: Kubica, Dominick, et autres
Publié: (2025)
SECURA: Sigmoid-Enhanced CUR Decomposition with Uninterrupted Retention and Low-Rank Adaptation in Large Language Models
par: Zhang, Yuxuan
Publié: (2025)
par: Zhang, Yuxuan
Publié: (2025)
Exemplar Retrieval Without Overhypothesis Induction: Limits of Distributional Sequence Learning in Early Word Learning
par: Cacioli, Jon-Paul
Publié: (2026)
par: Cacioli, Jon-Paul
Publié: (2026)
Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models
par: Wang, Yating, et autres
Publié: (2026)
par: Wang, Yating, et autres
Publié: (2026)
Align and Shine: Building High-Quality Sentence-Aligned Corpora for Multilingual Text Simplification
par: Hilasaca, Kenji, et autres
Publié: (2026)
par: Hilasaca, Kenji, et autres
Publié: (2026)
Intention Collapse: Intention-Level Metrics for Reasoning in Language Models
par: Vera, Patricio
Publié: (2026)
par: Vera, Patricio
Publié: (2026)
Whether, Not Which: Mechanistic Interpretability Reveals Dissociable Affect Reception and Emotion Categorization in LLMs
par: Keeman, Michael
Publié: (2026)
par: Keeman, Michael
Publié: (2026)
Why Models Know But Don't Say: Chain-of-Thought Faithfulness Divergence Between Thinking Tokens and Answers in Open-Weight Reasoning Models
par: Young, Richard J.
Publié: (2026)
par: Young, Richard J.
Publié: (2026)
The Pragmatic Persona: Discovering LLM Persona through Bridging Inference
par: Yang, Jisoo, et autres
Publié: (2026)
par: Yang, Jisoo, et autres
Publié: (2026)
Training Language Models to Win Debates with Self-Play Improves Judge Accuracy
par: Arnesen, Samuel, et autres
Publié: (2024)
par: Arnesen, Samuel, et autres
Publié: (2024)
RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences
par: Zhou, Yangyang, et autres
Publié: (2026)
par: Zhou, Yangyang, et autres
Publié: (2026)
UrduBench: An Urdu Reasoning Benchmark using Contextually Ensembled Translations with Human-in-the-Loop
par: Shafique, Muhammad Ali, et autres
Publié: (2026)
par: Shafique, Muhammad Ali, et autres
Publié: (2026)
Documents similaires
-
Adaptive Activation Cancellation for Hallucination Mitigation in Large Language Models
par: Yocam, Eric, et autres
Publié: (2026) -
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
par: Wu, Shuai, et autres
Publié: (2026) -
Weakly Supervised Distillation of Hallucination Signals into Transformer Representations
par: Salehmohamed, Shoaib Sadiq, et autres
Publié: (2026) -
EvoIdeator: Evolving Scientific Ideas through Checklist-Grounded Reinforcement Learning
par: Sauter, Andreas, et autres
Publié: (2026) -
Beyond Hallucinations: A Composite Score for Measuring Reliability in Open-Source Large Language Models
par: Salla, Rohit Kumar, et autres
Publié: (2025)