Retromorphic Testing with Hierarchical Verification for Hallucination Detection in RAG
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yu, Boxi, Zhang, Yuzhong, Lin, Liting, Briand, Lionel, Muñoz, Emir |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FACTUM: Mechanistic Detection of Citation Hallucination in Long-Form RAG
von: Dassen, Maxime, et al.
Veröffentlicht: (2026)
von: Dassen, Maxime, et al.
Veröffentlicht: (2026)
Biomedical systems biology workflow orchestration and execution with PoSyMed
von: Süwer, Simon, et al.
Veröffentlicht: (2026)
von: Süwer, Simon, et al.
Veröffentlicht: (2026)
An Industrial-Scale Retrieval-Augmented Generation Framework for Requirements Engineering: Empirical Evaluation with Automotive Manufacturing Data
von: Khalid, Muhammad, et al.
Veröffentlicht: (2026)
von: Khalid, Muhammad, et al.
Veröffentlicht: (2026)
SkillForge: Forging Domain-Specific, Self-Evolving Agent Skills in Cloud Technical Support
von: Liu, Xingyan, et al.
Veröffentlicht: (2026)
von: Liu, Xingyan, et al.
Veröffentlicht: (2026)
When LLM meets Fuzzy-TOPSIS for Personnel Selection through Automated Profile Analysis
von: Hoque, Shahria, et al.
Veröffentlicht: (2026)
von: Hoque, Shahria, et al.
Veröffentlicht: (2026)
Contextually Aware E-Commerce Product Question Answering using RAG
von: Tangarajan, Praveen, et al.
Veröffentlicht: (2025)
von: Tangarajan, Praveen, et al.
Veröffentlicht: (2025)
BridgeRAG: Training-Free Bridge-Conditioned Retrieval for Multi-Hop Question Answering
von: Bacellar, Andre
Veröffentlicht: (2026)
von: Bacellar, Andre
Veröffentlicht: (2026)
Benchmarking Google Embeddings 2 against Open-Source Models for Multilingual Dense Retrieval and RAG Systems
von: Cirillo, Stefano, et al.
Veröffentlicht: (2026)
von: Cirillo, Stefano, et al.
Veröffentlicht: (2026)
SLA Management in Reconfigurable Multi-Agent RAG: A Systems Approach to Question Answering
von: Iannelli, Michael, et al.
Veröffentlicht: (2024)
von: Iannelli, Michael, et al.
Veröffentlicht: (2024)
LLMLogAnalyzer: A Clustering-Based Log Analysis Chatbot using Large Language Models
von: Cai, Peng, et al.
Veröffentlicht: (2025)
von: Cai, Peng, et al.
Veröffentlicht: (2025)
The Case for Intent-Based Query Rewriting
von: Nicolai, Gianna Lisa, et al.
Veröffentlicht: (2025)
von: Nicolai, Gianna Lisa, et al.
Veröffentlicht: (2025)
A Grounded Memory System For Smart Personal Assistants
von: Ocker, Felix, et al.
Veröffentlicht: (2025)
von: Ocker, Felix, et al.
Veröffentlicht: (2025)
Codebase-Memory: Tree-Sitter-Based Knowledge Graphs for LLM Code Exploration via MCP
von: Vogel, Martin, et al.
Veröffentlicht: (2026)
von: Vogel, Martin, et al.
Veröffentlicht: (2026)
Less LLM, More Documents: Searching for Improved RAG
von: Ning, Jingjie, et al.
Veröffentlicht: (2025)
von: Ning, Jingjie, et al.
Veröffentlicht: (2025)
Evaluating the Effectiveness of Large Language Models in Automated News Article Summarization
von: Houamegni, Lionel Richy Panlap, et al.
Veröffentlicht: (2025)
von: Houamegni, Lionel Richy Panlap, et al.
Veröffentlicht: (2025)
Memory Architectures for Multi-Turn Text-to-SQL: A Benchmark and Empirical Study
von: Tummalapenta, Ravi Kumar, et al.
Veröffentlicht: (2026)
von: Tummalapenta, Ravi Kumar, et al.
Veröffentlicht: (2026)
FinGround: Detecting and Grounding Financial Hallucinations via Atomic Claim Verification
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
MLDocRAG: Multimodal Long-Context Document Retrieval Augmented Generation
von: Zhang, Yongyue, et al.
Veröffentlicht: (2026)
von: Zhang, Yongyue, et al.
Veröffentlicht: (2026)
Enhancing Long-term RAG Chatbots with Psychological Models of Memory Importance and Forgetting
von: Sumida, Ryuichi, et al.
Veröffentlicht: (2024)
von: Sumida, Ryuichi, et al.
Veröffentlicht: (2024)
ToolForge: A Data Synthesis Pipeline for Multi-Hop Search without Real-World APIs
von: Chen, Hao, et al.
Veröffentlicht: (2025)
von: Chen, Hao, et al.
Veröffentlicht: (2025)
Mitigating Hallucinations in Large Language Models via Self-Refinement-Enhanced Knowledge Retrieval
von: Niu, Mengjia, et al.
Veröffentlicht: (2024)
von: Niu, Mengjia, et al.
Veröffentlicht: (2024)
The Reasoning Bottleneck in Graph-RAG: Structured Prompting and Context Compression for Multi-Hop QA
von: Zarrinkia, Yasaman, et al.
Veröffentlicht: (2026)
von: Zarrinkia, Yasaman, et al.
Veröffentlicht: (2026)
Flippi: End To End GenAI Assistant for E-Commerce
von: Rajasekar, Anand A., et al.
Veröffentlicht: (2025)
von: Rajasekar, Anand A., et al.
Veröffentlicht: (2025)
Fine-Grained Emotion Recognition via In-Context Learning
von: Ren, Zhaochun, et al.
Veröffentlicht: (2025)
von: Ren, Zhaochun, et al.
Veröffentlicht: (2025)
Adaptive ToR: Complexity-Aware Tree-Based Retrieval for Pareto-Optimal Multi-Intent NLU
von: Yoo, Hee-Kyong, et al.
Veröffentlicht: (2026)
von: Yoo, Hee-Kyong, et al.
Veröffentlicht: (2026)
MasterSet: A Large-Scale Benchmark for Must-Cite Citation Recommendation in the AI/ML Literature
von: Ratul, Md Toyaha Rahman, et al.
Veröffentlicht: (2026)
von: Ratul, Md Toyaha Rahman, et al.
Veröffentlicht: (2026)
NewsScope: Schema-Grounded Cross-Domain News Claim Extraction with Open Models
von: Pandya, Nidhi
Veröffentlicht: (2025)
von: Pandya, Nidhi
Veröffentlicht: (2025)
Evaluation of Chunking Strategies for Effective Text Embedding in Low-Resource Language on Agricultural Documents
von: Chhoun, Sovandara, et al.
Veröffentlicht: (2026)
von: Chhoun, Sovandara, et al.
Veröffentlicht: (2026)
Stage-Audit: Auditable Source-Frontier Discovery for Cross-Wiki Tables
von: Shen, Chen
Veröffentlicht: (2026)
von: Shen, Chen
Veröffentlicht: (2026)
Train Once, Use Flexibly: A Modular Framework for Multi-Aspect Neural News Recommendation
von: Iana, Andreea, et al.
Veröffentlicht: (2023)
von: Iana, Andreea, et al.
Veröffentlicht: (2023)
Towards Adaptive Context Management for Intelligent Conversational Question Answering
von: Perera, Manoj Madushanka, et al.
Veröffentlicht: (2025)
von: Perera, Manoj Madushanka, et al.
Veröffentlicht: (2025)
Session Context Embedding for Intent Understanding in Product Search
von: Mehrdad, Navid, et al.
Veröffentlicht: (2024)
von: Mehrdad, Navid, et al.
Veröffentlicht: (2024)
Scaling Multilingual Semantic Search in Uber Eats Delivery
von: Ling, Bo, et al.
Veröffentlicht: (2026)
von: Ling, Bo, et al.
Veröffentlicht: (2026)
When Stored Evidence Stops Being Usable: Scale-Conditioned Evaluation of Agent Memory
von: Shao, Jiaqi, et al.
Veröffentlicht: (2026)
von: Shao, Jiaqi, et al.
Veröffentlicht: (2026)
Does UMBRELA Work on Other LLMs?
von: Farzi, Naghmeh, et al.
Veröffentlicht: (2025)
von: Farzi, Naghmeh, et al.
Veröffentlicht: (2025)
Optimizing Retrieval-Augmented Generation for Electrical Engineering: A Case Study on ABB Circuit Breakers
von: Alawadhi, Salahuddin, et al.
Veröffentlicht: (2025)
von: Alawadhi, Salahuddin, et al.
Veröffentlicht: (2025)
2024 Google Scholar Research Interest Ranking for Top 3260 Computer Science Authors
von: Rasane, Atharva
Veröffentlicht: (2024)
von: Rasane, Atharva
Veröffentlicht: (2024)
Local Hybrid Retrieval-Augmented Document QA
von: Astrino, Paolo
Veröffentlicht: (2025)
von: Astrino, Paolo
Veröffentlicht: (2025)
QUARK: Robust Retrieval under Non-Faithful Queries via Query-Anchored Aggregation
von: Lyu, Rita Qiuran, et al.
Veröffentlicht: (2026)
von: Lyu, Rita Qiuran, et al.
Veröffentlicht: (2026)
Improving and Evaluating Open Deep Research Agents
von: Allabadi, Doaa, et al.
Veröffentlicht: (2025)
von: Allabadi, Doaa, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FACTUM: Mechanistic Detection of Citation Hallucination in Long-Form RAG
von: Dassen, Maxime, et al.
Veröffentlicht: (2026) -
Biomedical systems biology workflow orchestration and execution with PoSyMed
von: Süwer, Simon, et al.
Veröffentlicht: (2026) -
An Industrial-Scale Retrieval-Augmented Generation Framework for Requirements Engineering: Empirical Evaluation with Automotive Manufacturing Data
von: Khalid, Muhammad, et al.
Veröffentlicht: (2026) -
SkillForge: Forging Domain-Specific, Self-Evolving Agent Skills in Cloud Technical Support
von: Liu, Xingyan, et al.
Veröffentlicht: (2026) -
When LLM meets Fuzzy-TOPSIS for Personnel Selection through Automated Profile Analysis
von: Hoque, Shahria, et al.
Veröffentlicht: (2026)