Hallucination-Resistant, Domain-Specific Research Assistant with Self-Evaluation and Vector-Grounded Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Bhavsar, Vivek, Ereifej, Joseph, Gurusami, Aravanan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLMs to Support a Domain Specific Knowledge Assistant
by: Lovin, Maria-Flavia
Published: (2025)
by: Lovin, Maria-Flavia
Published: (2025)
Domain-Specific Improvement on Psychotherapy Chatbot Using Assistant
by: Kang, Cheng, et al.
Published: (2024)
by: Kang, Cheng, et al.
Published: (2024)
How Many Parameters Does it Take to Change a Light Bulb? Evaluating Performance in Self-Play of Conversational Games as a Function of Model Characteristics
by: Bhavsar, Nidhir, et al.
Published: (2024)
by: Bhavsar, Nidhir, et al.
Published: (2024)
ProgRAG: Hallucination-Resistant Progressive Retrieval and Reasoning over Knowledge Graphs
by: Park, Minbae, et al.
Published: (2025)
by: Park, Minbae, et al.
Published: (2025)
GraphRAG-Bench: Challenging Domain-Specific Reasoning for Evaluating Graph Retrieval-Augmented Generation
by: Xiao, Yilin, et al.
Published: (2025)
by: Xiao, Yilin, et al.
Published: (2025)
Human-AI Collaborative Taxonomy Construction: A Case Study in Profession-Specific Writing Assistants
by: Lee, Minhwa, et al.
Published: (2024)
by: Lee, Minhwa, et al.
Published: (2024)
InteGround: On the Evaluation of Verification and Retrieval Planning in Integrative Grounding
by: Jiayang, Cheng, et al.
Published: (2025)
by: Jiayang, Cheng, et al.
Published: (2025)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
by: Zhang, Xiaoying, et al.
Published: (2024)
by: Zhang, Xiaoying, et al.
Published: (2024)
Domain-Specific Retrieval-Augmented Generation Using Vector Stores, Knowledge Graphs, and Tensor Factorization
by: Barron, Ryan C., et al.
Published: (2024)
by: Barron, Ryan C., et al.
Published: (2024)
Evaluating ChatGPT on Nuclear Domain-Specific Data
by: Anwar, Muhammad, et al.
Published: (2024)
by: Anwar, Muhammad, et al.
Published: (2024)
GRAMMAR: Grounded and Modular Methodology for Assessment of Closed-Domain Retrieval-Augmented Language Model
by: Li, Xinzhe, et al.
Published: (2024)
by: Li, Xinzhe, et al.
Published: (2024)
Vectoring Languages
by: Chen, Joseph
Published: (2024)
by: Chen, Joseph
Published: (2024)
RAFT: Adapting Language Model to Domain Specific RAG
by: Zhang, Tianjun, et al.
Published: (2024)
by: Zhang, Tianjun, et al.
Published: (2024)
Enhancing Domain-Specific Retrieval-Augmented Generation: Synthetic Data Generation and Evaluation using Reasoning Models
by: Jadon, Aryan, et al.
Published: (2025)
by: Jadon, Aryan, et al.
Published: (2025)
RAGalyst: Automated Human-Aligned Agentic Evaluation for Domain-Specific RAG
by: Gao, Joshua, et al.
Published: (2025)
by: Gao, Joshua, et al.
Published: (2025)
Rankers, Judges, and Assistants: Towards Understanding the Interplay of LLMs in Information Retrieval Evaluation
by: Balog, Krisztian, et al.
Published: (2025)
by: Balog, Krisztian, et al.
Published: (2025)
ChatSOS: Vector Database Augmented Generative Question Answering Assistant in Safety Engineering
by: Tang, Haiyang, et al.
Published: (2024)
by: Tang, Haiyang, et al.
Published: (2024)
From Guidelines to Guarantees: A Graph-Based Evaluation Harness for Domain-Specific Evaluation of LLMs
by: Lundin, Jessica M., et al.
Published: (2025)
by: Lundin, Jessica M., et al.
Published: (2025)
VeriTrail: Closed-Domain Hallucination Detection with Traceability
by: Metropolitansky, Dasha, et al.
Published: (2025)
by: Metropolitansky, Dasha, et al.
Published: (2025)
GAIus: Combining Genai with Legal Clauses Retrieval for Knowledge-based Assistant
by: Matak, Michał, et al.
Published: (2025)
by: Matak, Michał, et al.
Published: (2025)
A Functionality-Grounded Benchmark for Evaluating Web Agents in E-commerce Domains
by: Zhang, Xianren, et al.
Published: (2025)
by: Zhang, Xianren, et al.
Published: (2025)
Sequence-Level Certainty Reduces Hallucination In Knowledge-Grounded Dialogue Generation
by: Wan, Yixin, et al.
Published: (2023)
by: Wan, Yixin, et al.
Published: (2023)
Fusion-Eval: Integrating Assistant Evaluators with LLMs
by: Shu, Lei, et al.
Published: (2023)
by: Shu, Lei, et al.
Published: (2023)
DO-RAG: A Domain-Specific QA Framework Using Knowledge Graph-Enhanced Retrieval-Augmented Generation
by: Opoku, David Osei, et al.
Published: (2025)
by: Opoku, David Osei, et al.
Published: (2025)
Mitigating Hallucination in Abstractive Summarization with Domain-Conditional Mutual Information
by: Chae, Kyubyung, et al.
Published: (2024)
by: Chae, Kyubyung, et al.
Published: (2024)
DeCoRe: Decoding by Contrasting Retrieval Heads to Mitigate Hallucinations
by: Gema, Aryo Pradipta, et al.
Published: (2024)
by: Gema, Aryo Pradipta, et al.
Published: (2024)
RepIt: Steering Language Models with Concept-Specific Refusal Vectors
by: Siu, Vincent, et al.
Published: (2025)
by: Siu, Vincent, et al.
Published: (2025)
DeepWriter: A Fact-Grounded Multimodal Writing Assistant Based On Offline Knowledge Base
by: Mao, Song, et al.
Published: (2025)
by: Mao, Song, et al.
Published: (2025)
A FAIR and Free Prompt-based Research Assistant
by: Shamsabadi, Mahsa, et al.
Published: (2024)
by: Shamsabadi, Mahsa, et al.
Published: (2024)
SLearnLLM: A Self-Learning Framework for Efficient Domain-Specific Adaptation of Large Language Models
by: Liu, Xiang, et al.
Published: (2025)
by: Liu, Xiang, et al.
Published: (2025)
Think$^{2}$: Grounded Metacognitive Reasoning in Large Language Models
by: Elenjical, Abraham Paul, et al.
Published: (2026)
by: Elenjical, Abraham Paul, et al.
Published: (2026)
Where Fake Citations Are Made: Tracing Field-Level Hallucination to Specific Neurons in LLMs
by: Chen, Yuefei, et al.
Published: (2026)
by: Chen, Yuefei, et al.
Published: (2026)
Evaluating LLM Reasoning in the Operations Research Domain with ORQA
by: Mostajabdaveh, Mahdi, et al.
Published: (2024)
by: Mostajabdaveh, Mahdi, et al.
Published: (2024)
Cross-Domain Content Generation with Domain-Specific Small Language Models
by: Maloo, Ankit, et al.
Published: (2024)
by: Maloo, Ankit, et al.
Published: (2024)
Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation
by: Mündler, Niels, et al.
Published: (2023)
by: Mündler, Niels, et al.
Published: (2023)
Lynx: An Open Source Hallucination Evaluation Model
by: Ravi, Selvan Sunitha, et al.
Published: (2024)
by: Ravi, Selvan Sunitha, et al.
Published: (2024)
Don't Let It Hallucinate: Premise Verification via Retrieval-Augmented Logical Reasoning
by: Qin, Yuehan, et al.
Published: (2025)
by: Qin, Yuehan, et al.
Published: (2025)
Micro-Macro Retrieval: Reducing Long-Form Hallucination in Large Language Models
by: Feng, Yujie, et al.
Published: (2026)
by: Feng, Yujie, et al.
Published: (2026)
Do Hallucination Neurons Generalize? Evidence from Cross-Domain Transfer in LLMs
by: Vaddi, Snehit, et al.
Published: (2026)
by: Vaddi, Snehit, et al.
Published: (2026)
BioRAGent: A Retrieval-Augmented Generation System for Showcasing Generative Query Expansion and Domain-Specific Search for Scientific Q&A
by: Ateia, Samy, et al.
Published: (2024)
by: Ateia, Samy, et al.
Published: (2024)
Similar Items
-
LLMs to Support a Domain Specific Knowledge Assistant
by: Lovin, Maria-Flavia
Published: (2025) -
Domain-Specific Improvement on Psychotherapy Chatbot Using Assistant
by: Kang, Cheng, et al.
Published: (2024) -
How Many Parameters Does it Take to Change a Light Bulb? Evaluating Performance in Self-Play of Conversational Games as a Function of Model Characteristics
by: Bhavsar, Nidhir, et al.
Published: (2024) -
ProgRAG: Hallucination-Resistant Progressive Retrieval and Reasoning over Knowledge Graphs
by: Park, Minbae, et al.
Published: (2025) -
GraphRAG-Bench: Challenging Domain-Specific Reasoning for Evaluating Graph Retrieval-Augmented Generation
by: Xiao, Yilin, et al.
Published: (2025)