Mitigating Hallucinations in Zero-Shot Scientific Summarisation: A Pilot Study
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jaaouine, Imane, King, Ross D. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive Activation Cancellation for Hallucination Mitigation in Large Language Models
von: Yocam, Eric, et al.
Veröffentlicht: (2026)
von: Yocam, Eric, et al.
Veröffentlicht: (2026)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
Weakly Supervised Distillation of Hallucination Signals into Transformer Representations
von: Salehmohamed, Shoaib Sadiq, et al.
Veröffentlicht: (2026)
von: Salehmohamed, Shoaib Sadiq, et al.
Veröffentlicht: (2026)
EvoIdeator: Evolving Scientific Ideas through Checklist-Grounded Reinforcement Learning
von: Sauter, Andreas, et al.
Veröffentlicht: (2026)
von: Sauter, Andreas, et al.
Veröffentlicht: (2026)
Beyond Hallucinations: A Composite Score for Measuring Reliability in Open-Source Large Language Models
von: Salla, Rohit Kumar, et al.
Veröffentlicht: (2025)
von: Salla, Rohit Kumar, et al.
Veröffentlicht: (2025)
Incentives or Ontology? A Structural Rebuttal to OpenAI's Hallucination Thesis
von: Ackermann, Richard, et al.
Veröffentlicht: (2025)
von: Ackermann, Richard, et al.
Veröffentlicht: (2025)
Adapting While Learning: Grounding LLMs for Scientific Problems with Intelligent Tool Usage Adaptation
von: Lyu, Bohan, et al.
Veröffentlicht: (2024)
von: Lyu, Bohan, et al.
Veröffentlicht: (2024)
Improving Commonsense Bias Classification by Mitigating the Influence of Demographic Terms
von: Lee, JinKyu, et al.
Veröffentlicht: (2024)
von: Lee, JinKyu, et al.
Veröffentlicht: (2024)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
von: Liu, Zhongxin, et al.
Veröffentlicht: (2025)
von: Liu, Zhongxin, et al.
Veröffentlicht: (2025)
Re-Mask and Redirect: Exploiting Denoising Irreversibility in Diffusion Language Models
von: Singh, Arth
Veröffentlicht: (2026)
von: Singh, Arth
Veröffentlicht: (2026)
Let's Think Dot by Dot: Hidden Computation in Transformer Language Models
von: Pfau, Jacob, et al.
Veröffentlicht: (2024)
von: Pfau, Jacob, et al.
Veröffentlicht: (2024)
CRAFT: Clustered Regression for Adaptive Filtering of Training data
von: Panda, Parthasarathi, et al.
Veröffentlicht: (2026)
von: Panda, Parthasarathi, et al.
Veröffentlicht: (2026)
EMA Is Not All You Need: Mapping the Boundary Between Structure and Content in Recurrent Context
von: Singh, Arth
Veröffentlicht: (2026)
von: Singh, Arth
Veröffentlicht: (2026)
In-Context Fixation: When Demonstrated Labels Override Semantics in Few-Shot Classification
von: Liu, Ming
Veröffentlicht: (2026)
von: Liu, Ming
Veröffentlicht: (2026)
Synthius-Mem: Brain-Inspired Hallucination-Resistant Persona Memory Achieving 94.4% Memory Accuracy and 99.6% Adversarial Robustness on LoCoMo
von: Gadzhiev, Artem, et al.
Veröffentlicht: (2026)
von: Gadzhiev, Artem, et al.
Veröffentlicht: (2026)
ImmigrationQA: A Source-Grounded Dataset and Small-Model Adaptation for U.S. Immigration Law
von: Shportun, Nazarii
Veröffentlicht: (2026)
von: Shportun, Nazarii
Veröffentlicht: (2026)
A Hierarchical Error Framework for Reliable Automated Coding in Communication Research: Applications to Health and Political Communication
von: Zhao, Zhilong, et al.
Veröffentlicht: (2025)
von: Zhao, Zhilong, et al.
Veröffentlicht: (2025)
Distilling Self-Consistency into Verbal Confidence: A Pre-Registered Negative Result and Post-Hoc Rescue on Gemma 3 4B
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
Accelerating Complex Disease Treatment through Network Medicine and GenAI: A Case Study on Drug Repurposing for Breast Cancer
von: Hamed, Ahmed Abdeen, et al.
Veröffentlicht: (2024)
von: Hamed, Ahmed Abdeen, et al.
Veröffentlicht: (2024)
Graph Your Way to Inspiration: Integrating Co-Author Graphs with Retrieval-Augmented Generation for Large Language Model Based Scientific Idea Generation
von: Xie, Pengzhen, et al.
Veröffentlicht: (2025)
von: Xie, Pengzhen, et al.
Veröffentlicht: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
KAConvText: Novel Approach to Burmese Sentence Classification using Kolmogorov-Arnold Convolution
von: Thu, Ye Kyaw, et al.
Veröffentlicht: (2025)
von: Thu, Ye Kyaw, et al.
Veröffentlicht: (2025)
When Persuasion Overrides Truth in Multi-Agent LLM Debates: Introducing a Confidence-Weighted Persuasion Override Rate (CW-POR)
von: Agarwal, Mahak, et al.
Veröffentlicht: (2025)
von: Agarwal, Mahak, et al.
Veröffentlicht: (2025)
Bridging the Reasoning Gap: Small LLMs Can Plan with Generalised Strategies
von: Borro, Andrey, et al.
Veröffentlicht: (2025)
von: Borro, Andrey, et al.
Veröffentlicht: (2025)
Resource for Error Analysis in Text Simplification: New Taxonomy and Test Collection
von: Vendeville, Benjamin, et al.
Veröffentlicht: (2025)
von: Vendeville, Benjamin, et al.
Veröffentlicht: (2025)
Assessing Large Language Models on Islamic Legal Reasoning: Evidence from Inheritance Law Evaluation
von: Bouchekif, Abdessalam, et al.
Veröffentlicht: (2025)
von: Bouchekif, Abdessalam, et al.
Veröffentlicht: (2025)
Can AI Read Between The Lines? Benchmarking LLMs On Financial Nuance
von: Kubica, Dominick, et al.
Veröffentlicht: (2025)
von: Kubica, Dominick, et al.
Veröffentlicht: (2025)
SECURA: Sigmoid-Enhanced CUR Decomposition with Uninterrupted Retention and Low-Rank Adaptation in Large Language Models
von: Zhang, Yuxuan
Veröffentlicht: (2025)
von: Zhang, Yuxuan
Veröffentlicht: (2025)
Exemplar Retrieval Without Overhypothesis Induction: Limits of Distributional Sequence Learning in Early Word Learning
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models
von: Wang, Yating, et al.
Veröffentlicht: (2026)
von: Wang, Yating, et al.
Veröffentlicht: (2026)
Align and Shine: Building High-Quality Sentence-Aligned Corpora for Multilingual Text Simplification
von: Hilasaca, Kenji, et al.
Veröffentlicht: (2026)
von: Hilasaca, Kenji, et al.
Veröffentlicht: (2026)
Intention Collapse: Intention-Level Metrics for Reasoning in Language Models
von: Vera, Patricio
Veröffentlicht: (2026)
von: Vera, Patricio
Veröffentlicht: (2026)
Whether, Not Which: Mechanistic Interpretability Reveals Dissociable Affect Reception and Emotion Categorization in LLMs
von: Keeman, Michael
Veröffentlicht: (2026)
von: Keeman, Michael
Veröffentlicht: (2026)
Why Models Know But Don't Say: Chain-of-Thought Faithfulness Divergence Between Thinking Tokens and Answers in Open-Weight Reasoning Models
von: Young, Richard J.
Veröffentlicht: (2026)
von: Young, Richard J.
Veröffentlicht: (2026)
The Pragmatic Persona: Discovering LLM Persona through Bridging Inference
von: Yang, Jisoo, et al.
Veröffentlicht: (2026)
von: Yang, Jisoo, et al.
Veröffentlicht: (2026)
Training Language Models to Win Debates with Self-Play Improves Judge Accuracy
von: Arnesen, Samuel, et al.
Veröffentlicht: (2024)
von: Arnesen, Samuel, et al.
Veröffentlicht: (2024)
RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences
von: Zhou, Yangyang, et al.
Veröffentlicht: (2026)
von: Zhou, Yangyang, et al.
Veröffentlicht: (2026)
UrduBench: An Urdu Reasoning Benchmark using Contextually Ensembled Translations with Human-in-the-Loop
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2026)
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2026)
Eyla: Toward an Identity-Anchored LLM Architecture with Integrated Biological Priors -- Vision, Implementation Attempt, and Lessons from AI-Assisted Development
von: Aditto, Arif
Veröffentlicht: (2026)
von: Aditto, Arif
Veröffentlicht: (2026)
Truth as a Compression Artifact in Language Model Training
von: Krestnikov, Konstantin
Veröffentlicht: (2026)
von: Krestnikov, Konstantin
Veröffentlicht: (2026)
Ähnliche Einträge
-
Adaptive Activation Cancellation for Hallucination Mitigation in Large Language Models
von: Yocam, Eric, et al.
Veröffentlicht: (2026) -
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
von: Wu, Shuai, et al.
Veröffentlicht: (2026) -
Weakly Supervised Distillation of Hallucination Signals into Transformer Representations
von: Salehmohamed, Shoaib Sadiq, et al.
Veröffentlicht: (2026) -
EvoIdeator: Evolving Scientific Ideas through Checklist-Grounded Reinforcement Learning
von: Sauter, Andreas, et al.
Veröffentlicht: (2026) -
Beyond Hallucinations: A Composite Score for Measuring Reliability in Open-Source Large Language Models
von: Salla, Rohit Kumar, et al.
Veröffentlicht: (2025)