HalluMat: Detecting Hallucinations in LLM-Generated Materials Science Content Through Multi-Stage Verification
Fuente:
arXiv
Saved in:
| Main Authors: | Vangala, Bhanu Prakash, Mahmud, Sajid, Neupane, Pawan, Selvaraj, Joel, Cheng, Jianlin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Stage Verification-Centric Framework for Mitigating Hallucination in Multi-Modal RAG
by: Chen, Baiyu, et al.
Published: (2025)
by: Chen, Baiyu, et al.
Published: (2025)
Efficient Multi-Model Orchestration for Self-Hosted Large Language Models
by: Vangala, Bhanu Prakash, et al.
Published: (2025)
by: Vangala, Bhanu Prakash, et al.
Published: (2025)
MatSciRE: Leveraging Pointer Networks to Automate Entity and Relation Extraction for Material Science Knowledge-base Construction
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
Towards Unification of Hallucination Detection and Fact Verification for Large Language Models
by: Su, Weihang, et al.
Published: (2025)
by: Su, Weihang, et al.
Published: (2025)
MatPROV: A Provenance Graph Dataset of Material Synthesis Extracted from Scientific Literature
by: Tsuruta, Hirofumi, et al.
Published: (2025)
by: Tsuruta, Hirofumi, et al.
Published: (2025)
A Pointer Network-based Approach for Joint Extraction and Detection of Multi-Label Multi-Class Intents
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
Medico: Towards Hallucination Detection and Correction with Multi-source Evidence Fusion
by: Zhao, Xinping, et al.
Published: (2024)
by: Zhao, Xinping, et al.
Published: (2024)
FVA-RAG: Falsification-Verification Alignment for Mitigating Sycophantic Hallucinations
by: Ravishankara, Mayank
Published: (2025)
by: Ravishankara, Mayank
Published: (2025)
Improving AlphaFold2 ‐ and AlphaFold3 ‐Based Protein Complex Structure Prediction With MULTICOM4 in CASP16
by: Jian Liu, et al.
Published: (2025)
by: Jian Liu, et al.
Published: (2025)
PSBench: a large-scale benchmark for estimating the accuracy of protein complex structural models
by: Neupane, Pawan, et al.
Published: (2025)
by: Neupane, Pawan, et al.
Published: (2025)
ReSearch: A Multi-Stage Machine Learning Framework for Earth Science Data Discovery
by: Sun, Youran, et al.
Published: (2026)
by: Sun, Youran, et al.
Published: (2026)
TriMat: Context-aware Recommendation by Tri-Matrix Factorization
by: Wang, Hao
Published: (2025)
by: Wang, Hao
Published: (2025)
Full Stage Learning to Rank: A Unified Framework for Multi-Stage Systems
by: Zheng, Kai, et al.
Published: (2024)
by: Zheng, Kai, et al.
Published: (2024)
Relevance Matters: A Multi-Task and Multi-Stage Large Language Model Approach for E-commerce Query Rewriting
by: Dai, Aijun, et al.
Published: (2026)
by: Dai, Aijun, et al.
Published: (2026)
HalluLens: LLM Hallucination Benchmark
by: Bang, Yejin, et al.
Published: (2025)
by: Bang, Yejin, et al.
Published: (2025)
Single-Turn LLM Reformulation Powered Multi-Stage Hybrid Re-Ranking for Tip-of-the-Tongue Known-Item Retrieval
by: Mukhopadhyay, Debayan, et al.
Published: (2026)
by: Mukhopadhyay, Debayan, et al.
Published: (2026)
AI-Generated Code Is Not Reproducible (Yet): An Empirical Study of Dependency Gaps in LLM-Based Coding Agents
by: Vangala, Bhanu Prakash, et al.
Published: (2025)
by: Vangala, Bhanu Prakash, et al.
Published: (2025)
Hallucination Detection and Evaluation of Large Language Model
by: Zhang, Chenggong, et al.
Published: (2025)
by: Zhang, Chenggong, et al.
Published: (2025)
The Discovery Gap: How Product Hunt Startups Vanish in LLM Organic Discovery Queries
by: Sharma, Amit Prakash
Published: (2026)
by: Sharma, Amit Prakash
Published: (2026)
ZeroMat: Solving Cold-start Problem of Recommender System with No Input Data
by: Wang, Hao
Published: (2021)
by: Wang, Hao
Published: (2021)
Alleviating LLM-based Generative Retrieval Hallucination in Alipay Search
by: Shen, Yedan, et al.
Published: (2025)
by: Shen, Yedan, et al.
Published: (2025)
Legommenders: A Comprehensive Content-Based Recommendation Library with LLM Support
by: Liu, Qijiong, et al.
Published: (2024)
by: Liu, Qijiong, et al.
Published: (2024)
Recall-Augmented Ranking: Enhancing Click-Through Rate Prediction Accuracy with Cross-Stage Data
by: Huang, Junjie, et al.
Published: (2024)
by: Huang, Junjie, et al.
Published: (2024)
Enhancing Semantic Interoperability Across Materials Science With HIVE4MAT
by: Greenberg, Jane, et al.
Published: (2024)
by: Greenberg, Jane, et al.
Published: (2024)
From Relevance to Utility: Evidence Retrieval with Feedback for Fact Verification
by: Zhang, Hengran, et al.
Published: (2023)
by: Zhang, Hengran, et al.
Published: (2023)
ColBERT-serve: Efficient Multi-Stage Memory-Mapped Scoring
by: Huang, Kaili, et al.
Published: (2025)
by: Huang, Kaili, et al.
Published: (2025)
Hierarchical Multi-field Representations for Two-Stage E-commerce Retrieval
by: Freymuth, Niklas, et al.
Published: (2025)
by: Freymuth, Niklas, et al.
Published: (2025)
A Hybrid Architecture for Multi-Stage Claim Document Understanding: Combining Vision-Language Models and Machine Learning for Real-Time Processing
by: Cheng, Lilu, et al.
Published: (2026)
by: Cheng, Lilu, et al.
Published: (2026)
Illusions of Relevance: Arbitrary Content Injection Attacks Deceive Retrievers, Rerankers, and LLM Judges
by: Tamber, Manveer Singh, et al.
Published: (2025)
by: Tamber, Manveer Singh, et al.
Published: (2025)
Hyper-RAG: Combating LLM Hallucinations using Hypergraph-Driven Retrieval-Augmented Generation
by: Feng, Yifan, et al.
Published: (2025)
by: Feng, Yifan, et al.
Published: (2025)
G-RAG: Knowledge Expansion in Material Science
by: Mostafa, Radeen, et al.
Published: (2024)
by: Mostafa, Radeen, et al.
Published: (2024)
Combating Biomedical Misinformation through Multi-modal Claim Detection and Evidence-based Verification
by: Barone, Mariano, et al.
Published: (2025)
by: Barone, Mariano, et al.
Published: (2025)
Improving Multi-modal Recommender Systems by Denoising and Aligning Multi-modal Content and User Feedback
by: Xv, Guipeng, et al.
Published: (2024)
by: Xv, Guipeng, et al.
Published: (2024)
Multi-Stage Field Extraction of Financial Documents with OCR and Compact Vision-Language Models
by: Jin, Yichao, et al.
Published: (2025)
by: Jin, Yichao, et al.
Published: (2025)
MultiRAG: A Knowledge-guided Framework for Mitigating Hallucination in Multi-source Retrieval Augmented Generation
by: Wu, Wenlong, et al.
Published: (2025)
by: Wu, Wenlong, et al.
Published: (2025)
Revisiting Human-vs-LLM judgments using the TREC Podcast Track
by: Mansour, Watheq, et al.
Published: (2026)
by: Mansour, Watheq, et al.
Published: (2026)
Tracing Content Requirements in Financial Documents using Multi-granularity Text Analysis
by: Li, Xiaochen, et al.
Published: (2021)
by: Li, Xiaochen, et al.
Published: (2021)
VeriCite: Towards Reliable Citations in Retrieval-Augmented Generation via Rigorous Verification
by: Qian, Haosheng, et al.
Published: (2025)
by: Qian, Haosheng, et al.
Published: (2025)
KadiAssistant: A conversational AI Agent for information retrieval in Kadi4Mat
by: Cierpka, Adrian, et al.
Published: (2026)
by: Cierpka, Adrian, et al.
Published: (2026)
FinGround: Detecting and Grounding Financial Hallucinations via Atomic Claim Verification
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
Similar Items
-
Multi-Stage Verification-Centric Framework for Mitigating Hallucination in Multi-Modal RAG
by: Chen, Baiyu, et al.
Published: (2025) -
Efficient Multi-Model Orchestration for Self-Hosted Large Language Models
by: Vangala, Bhanu Prakash, et al.
Published: (2025) -
MatSciRE: Leveraging Pointer Networks to Automate Entity and Relation Extraction for Material Science Knowledge-base Construction
by: Mullick, Ankan, et al.
Published: (2024) -
Towards Unification of Hallucination Detection and Fact Verification for Large Language Models
by: Su, Weihang, et al.
Published: (2025) -
MatPROV: A Provenance Graph Dataset of Material Synthesis Extracted from Scientific Literature
by: Tsuruta, Hirofumi, et al.
Published: (2025)