Osiris: A Lightweight Open-Source Hallucination Detection System
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Shan, Alex, Bauer, John, Manning, Christopher D. |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Semgrex and Ssurgeon, Searching and Manipulating Dependency Graphs
par: Bauer, John, et autres
Publié: (2024)
par: Bauer, John, et autres
Publié: (2024)
Do "English" Named Entity Recognizers Work Well on Global Englishes?
par: Shan, Alexander, et autres
Publié: (2024)
par: Shan, Alexander, et autres
Publié: (2024)
A New Pair of GloVes
par: Carlson, Riley, et autres
Publié: (2025)
par: Carlson, Riley, et autres
Publié: (2025)
Drop Dropout on Single-Epoch Language Model Pretraining
par: Liu, Houjun, et autres
Publié: (2025)
par: Liu, Houjun, et autres
Publié: (2025)
Lynx: An Open Source Hallucination Evaluation Model
par: Ravi, Selvan Sunitha, et autres
Publié: (2024)
par: Ravi, Selvan Sunitha, et autres
Publié: (2024)
The Case for Repeatable, Open, and Expert-Grounded Hallucination Benchmarks in Large Language Models
par: Norman, Justin D., et autres
Publié: (2025)
par: Norman, Justin D., et autres
Publié: (2025)
Hallucination-Free? Assessing the Reliability of Leading AI Legal Research Tools
par: Magesh, Varun, et autres
Publié: (2024)
par: Magesh, Varun, et autres
Publié: (2024)
Stronger Baselines for Retrieval-Augmented Generation with Long-Context Language Models
par: Laitenberger, Alex, et autres
Publié: (2025)
par: Laitenberger, Alex, et autres
Publié: (2025)
Marcel: A Lightweight and Open-Source Conversational Agent for University Student Support
par: Trienes, Jan, et autres
Publié: (2025)
par: Trienes, Jan, et autres
Publié: (2025)
Humans and transformer LMs: Abstraction drives language learning
par: Jian, Jasper, et autres
Publié: (2026)
par: Jian, Jasper, et autres
Publié: (2026)
Source or It Didn't Happen: A Multi-Agent Framework for Citation Hallucination Detection
par: Li, Mingzhe, et autres
Publié: (2026)
par: Li, Mingzhe, et autres
Publié: (2026)
ReCOGS: How Incidental Details of a Logical Form Overshadow an Evaluation of Semantic Interpretation
par: Wu, Zhengxuan, et autres
Publié: (2023)
par: Wu, Zhengxuan, et autres
Publié: (2023)
The Hallucinations Leaderboard -- An Open Effort to Measure Hallucinations in Large Language Models
par: Hong, Giwon, et autres
Publié: (2024)
par: Hong, Giwon, et autres
Publié: (2024)
ShED-HD: A Shannon Entropy Distribution Framework for Lightweight Hallucination Detection on Edge Devices
par: Vathul, Aneesh, et autres
Publié: (2025)
par: Vathul, Aneesh, et autres
Publié: (2025)
Instruction Following without Instruction Tuning
par: Hewitt, John, et autres
Publié: (2024)
par: Hewitt, John, et autres
Publié: (2024)
Sneaking Syntax into Transformer Language Models with Tree Regularization
par: Nandi, Ananjan, et autres
Publié: (2024)
par: Nandi, Ananjan, et autres
Publié: (2024)
HalluDetect: Detecting, Mitigating, and Benchmarking Hallucinations in Conversational Systems in the Legal Domain
par: Anaokar, Spandan, et autres
Publié: (2025)
par: Anaokar, Spandan, et autres
Publié: (2025)
LUMINA: Detecting Hallucinations in RAG System with Context-Knowledge Signals
par: Yeh, Samuel, et autres
Publié: (2025)
par: Yeh, Samuel, et autres
Publié: (2025)
MetaRAG: Metamorphic Testing for Hallucination Detection in RAG Systems
par: Sok, Channdeth, et autres
Publié: (2025)
par: Sok, Channdeth, et autres
Publié: (2025)
Hallucination Detection and Hallucination Mitigation: An Investigation
par: Luo, Junliang, et autres
Publié: (2024)
par: Luo, Junliang, et autres
Publié: (2024)
Hallucinations and Key Information Extraction in Medical Texts: A Comprehensive Assessment of Open-Source Large Language Models
par: Das, Anindya Bijoy, et autres
Publié: (2025)
par: Das, Anindya Bijoy, et autres
Publié: (2025)
Hallucination Detection via Activations of Open-Weight Proxy Analyzers
par: Singh, Akshita, et autres
Publié: (2026)
par: Singh, Akshita, et autres
Publié: (2026)
HalluCiteChecker: A Lightweight Toolkit for Hallucinated Citation Detection and Verification in the Era of AI Scientists
par: Sakai, Yusuke, et autres
Publié: (2026)
par: Sakai, Yusuke, et autres
Publié: (2026)
HARP: Hallucination Detection via Reasoning Subspace Projection
par: Hu, Junjie, et autres
Publié: (2025)
par: Hu, Junjie, et autres
Publié: (2025)
LLM Hallucination Detection: HSAD
par: Li, JinXin, et autres
Publié: (2025)
par: Li, JinXin, et autres
Publié: (2025)
Learning Auxiliary Tasks Improves Reference-Free Hallucination Detection in Open-Domain Long-Form Generation
par: Qin, Chengwei, et autres
Publié: (2025)
par: Qin, Chengwei, et autres
Publié: (2025)
LLM Hallucination Detection: A Fast Fourier Transform Method Based on Hidden Layer Temporal Signals
par: Li, Jinxin, et autres
Publié: (2025)
par: Li, Jinxin, et autres
Publié: (2025)
Improved Representation Steering for Language Models
par: Wu, Zhengxuan, et autres
Publié: (2025)
par: Wu, Zhengxuan, et autres
Publié: (2025)
MQuAKE: Assessing Knowledge Editing in Language Models via Multi-Hop Questions
par: Zhong, Zexuan, et autres
Publié: (2023)
par: Zhong, Zexuan, et autres
Publié: (2023)
Transcribe, Translate, or Transliterate: An Investigation of Intermediate Representations in Spoken Language Models
par: Ògúnrèmí, Tolúlopé, et autres
Publié: (2025)
par: Ògúnrèmí, Tolúlopé, et autres
Publié: (2025)
Mechanisms vs. Outcomes: Probing for Syntax Fails to Explain Performance on Targeted Syntactic Evaluations
par: Agarwal, Ananth, et autres
Publié: (2025)
par: Agarwal, Ananth, et autres
Publié: (2025)
NNetNav: Unsupervised Learning of Browser Agents Through Environment Interaction in the Wild
par: Murty, Shikhar, et autres
Publié: (2024)
par: Murty, Shikhar, et autres
Publié: (2024)
Few-Shot Optimized Framework for Hallucination Detection in Resource-Limited NLP Systems
par: Hikal, Baraa, et autres
Publié: (2025)
par: Hikal, Baraa, et autres
Publié: (2025)
Towards Long Context Hallucination Detection
par: Liu, Siyi, et autres
Publié: (2025)
par: Liu, Siyi, et autres
Publié: (2025)
Can LLMs Detect Their Own Hallucinations?
par: Kadotani, Sora, et autres
Publié: (2025)
par: Kadotani, Sora, et autres
Publié: (2025)
EduCoder: An Open-Source Annotation System for Education Transcript Data
par: Ashraf, Saad, et autres
Publié: (2025)
par: Ashraf, Saad, et autres
Publié: (2025)
Hallucination is Inevitable for LLMs with the Open World Assumption
par: Xu, Bowen
Publié: (2025)
par: Xu, Bowen
Publié: (2025)
Model Editing with Canonical Examples
par: Hewitt, John, et autres
Publié: (2024)
par: Hewitt, John, et autres
Publié: (2024)
Detecting Hallucinations in Authentic LLM-Human Interactions
par: Ren, Yujie, et autres
Publié: (2025)
par: Ren, Yujie, et autres
Publié: (2025)
CLATTER: Comprehensive Entailment Reasoning for Hallucination Detection
par: Eliav, Ron, et autres
Publié: (2025)
par: Eliav, Ron, et autres
Publié: (2025)
Documents similaires
-
Semgrex and Ssurgeon, Searching and Manipulating Dependency Graphs
par: Bauer, John, et autres
Publié: (2024) -
Do "English" Named Entity Recognizers Work Well on Global Englishes?
par: Shan, Alexander, et autres
Publié: (2024) -
A New Pair of GloVes
par: Carlson, Riley, et autres
Publié: (2025) -
Drop Dropout on Single-Epoch Language Model Pretraining
par: Liu, Houjun, et autres
Publié: (2025) -
Lynx: An Open Source Hallucination Evaluation Model
par: Ravi, Selvan Sunitha, et autres
Publié: (2024)