BibTeX Citation Hallucinations in Scientific Publishing Agents: Evaluation and Mitigation
Fuente:
arXiv
Saved in:
| Main Authors: | Rao, Delip, Callison-Burch, Chris |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CiteAssist: A System for Automated Preprint Citation and BibTeX Generation
by: Kaesberg, Lars Benedikt, et al.
Published: (2024)
by: Kaesberg, Lars Benedikt, et al.
Published: (2024)
WithdrarXiv: A Large-Scale Dataset for Retraction Study
by: Rao, Delip, et al.
Published: (2024)
by: Rao, Delip, et al.
Published: (2024)
Detecting and Correcting Reference Hallucinations in Commercial LLMs and Deep Research Agents
by: Rao, Delip, et al.
Published: (2026)
by: Rao, Delip, et al.
Published: (2026)
Autorubric: Unifying Rubric-based LLM Evaluation
by: Rao, Delip, et al.
Published: (2026)
by: Rao, Delip, et al.
Published: (2026)
Agreement Metrics for LLM-as-Judge Evaluation: What to Report and Why
by: Rao, Delip, et al.
Published: (2026)
by: Rao, Delip, et al.
Published: (2026)
What Do Claim Verification Datasets Actually Test? A Reasoning Trace Analysis
by: Rao, Delip, et al.
Published: (2026)
by: Rao, Delip, et al.
Published: (2026)
NSF-SciFy: Mining the NSF Awards Database for Scientific Claims
by: Rao, Delip, et al.
Published: (2025)
by: Rao, Delip, et al.
Published: (2025)
Citation Grounding: Detecting and Reducing LLM Citation Hallucinations via Legal Citation Graphs
by: Ovcharov, Volodymyr
Published: (2026)
by: Ovcharov, Volodymyr
Published: (2026)
SciRAG: Adaptive, Citation-Aware, and Outline-Guided Retrieval and Synthesis for Scientific Literature
by: Ding, Hang, et al.
Published: (2025)
by: Ding, Hang, et al.
Published: (2025)
HalluCitation Matters: Revealing the Impact of Hallucinated References with 300 Hallucinated Papers in ACL Conferences
by: Sakai, Yusuke, et al.
Published: (2026)
by: Sakai, Yusuke, et al.
Published: (2026)
Sentiment Analysis of Citations in Scientific Articles Using ChatGPT: Identifying Potential Biases and Conflicts of Interest
by: Hariri, Walid
Published: (2024)
by: Hariri, Walid
Published: (2024)
ThinknCheck: Grounded Claim Verification with Compact, Reasoning-Driven, and Interpretable Models
by: Rao, Delip, et al.
Published: (2026)
by: Rao, Delip, et al.
Published: (2026)
SCIRGC: Multi-Granularity Citation Recommendation and Citation Sentence Preference Alignment
by: Li, Xiangyu, et al.
Published: (2025)
by: Li, Xiangyu, et al.
Published: (2025)
HalluCiteChecker: A Lightweight Toolkit for Hallucinated Citation Detection and Verification in the Era of AI Scientists
by: Sakai, Yusuke, et al.
Published: (2026)
by: Sakai, Yusuke, et al.
Published: (2026)
BibAgent: An Agentic Framework for Traceable Miscitation Detection in Scientific Literature
by: Li, Peiran, et al.
Published: (2026)
by: Li, Peiran, et al.
Published: (2026)
Semantically Orthogonal Framework for Citation Classification: Disentangling Intent and Content
by: Duan, Changxu, et al.
Published: (2026)
by: Duan, Changxu, et al.
Published: (2026)
Improving Citation Text Generation: Overcoming Limitations in Length Control
by: Mandal, Biswadip, et al.
Published: (2024)
by: Mandal, Biswadip, et al.
Published: (2024)
Citation Amnesia: On The Recency Bias of NLP and Other Academic Fields
by: Wahle, Jan Philip, et al.
Published: (2024)
by: Wahle, Jan Philip, et al.
Published: (2024)
Overview of SCIDOCA 2025 Shared Task on Citation Prediction, Discovery, and Placement
by: Dao, An, et al.
Published: (2025)
by: Dao, An, et al.
Published: (2025)
Digging Up Citations: FOSSIL, a Dataset and Workflow for Reference Extraction in Law and the Humanities
by: Foppiano, Luca, et al.
Published: (2026)
by: Foppiano, Luca, et al.
Published: (2026)
CC30k: A Citation Contexts Dataset for Reproducibility-Oriented Sentiment Analysis
by: Obadage, Rochana R., et al.
Published: (2025)
by: Obadage, Rochana R., et al.
Published: (2025)
Hidden Division of Labor in Scientific Teams Revealed Through 1.6 Million LaTeX Files
by: Pei, Jiaxin, et al.
Published: (2025)
by: Pei, Jiaxin, et al.
Published: (2025)
When Verification Fails: How Compositionally Infeasible Claims Escape Rejection
by: Liu, Muxin, et al.
Published: (2026)
by: Liu, Muxin, et al.
Published: (2026)
SemanticCite: Citation Verification with AI-Powered Full-Text Analysis and Evidence-Based Reasoning
by: Haan, Sebastian
Published: (2025)
by: Haan, Sebastian
Published: (2025)
Can LLMs Predict Citation Intent? An Experimental Analysis of In-context Learning and Fine-tuning on Open LLMs
by: Koloveas, Paris, et al.
Published: (2025)
by: Koloveas, Paris, et al.
Published: (2025)
Measuring and Analyzing Subjective Uncertainty in Scientific Communications
by: Sourati, Jamshid, et al.
Published: (2025)
by: Sourati, Jamshid, et al.
Published: (2025)
Is there really a Citation Age Bias in NLP?
by: Nguyen, Hoa, et al.
Published: (2024)
by: Nguyen, Hoa, et al.
Published: (2024)
Rethinking Review Citations: Impact on Scientific Integrity
by: Aguilar-Ruiz, Jesus S.
Published: (2025)
by: Aguilar-Ruiz, Jesus S.
Published: (2025)
AI-Reporter: A Path to a New Genre of Scientific Communication
by: Graßhoff, Gerd
Published: (2025)
by: Graßhoff, Gerd
Published: (2025)
Citation Recommendation based on Argumentative Zoning of User Queries
by: Ma, Shutian, et al.
Published: (2025)
by: Ma, Shutian, et al.
Published: (2025)
SCI-IDEA: Context-Aware Scientific Ideation Using Token and Sentence Embeddings
by: Keya, Farhana, et al.
Published: (2025)
by: Keya, Farhana, et al.
Published: (2025)
LLM4SR: A Survey on Large Language Models for Scientific Research
by: Luo, Ziming, et al.
Published: (2025)
by: Luo, Ziming, et al.
Published: (2025)
HLM-Cite: Hybrid Language Model Workflow for Text-based Scientific Citation Prediction
by: Hao, Qianyue, et al.
Published: (2024)
by: Hao, Qianyue, et al.
Published: (2024)
Automatic Detection of Research Values from Scientific Abstracts Across Computer Science Subfields
by: Jiang, Hang, et al.
Published: (2025)
by: Jiang, Hang, et al.
Published: (2025)
A Multi-lingual Dataset of Classified Paragraphs from Open Access Scientific Publications
by: Jeangirard, Eric
Published: (2025)
by: Jeangirard, Eric
Published: (2025)
Citation Parsing and Analysis with Language Models
by: Sarin, Parth, et al.
Published: (2025)
by: Sarin, Parth, et al.
Published: (2025)
Multi-Disciplinary Dataset Discovery from Citation-Verified Literature Contexts
by: Tan, Zhiyin, et al.
Published: (2026)
by: Tan, Zhiyin, et al.
Published: (2026)
CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era
by: Shi, Kaiwen, et al.
Published: (2026)
by: Shi, Kaiwen, et al.
Published: (2026)
CiteCheck: Retrieval-Grounded Detection of LLM Citation Hallucinations in Scientific Text
by: Khajavi, Khashayar, et al.
Published: (2026)
by: Khajavi, Khashayar, et al.
Published: (2026)
FLAWS: A Benchmark for Error Identification and Localization in Scientific Papers
by: Xi, Sarina, et al.
Published: (2025)
by: Xi, Sarina, et al.
Published: (2025)
Similar Items
-
CiteAssist: A System for Automated Preprint Citation and BibTeX Generation
by: Kaesberg, Lars Benedikt, et al.
Published: (2024) -
WithdrarXiv: A Large-Scale Dataset for Retraction Study
by: Rao, Delip, et al.
Published: (2024) -
Detecting and Correcting Reference Hallucinations in Commercial LLMs and Deep Research Agents
by: Rao, Delip, et al.
Published: (2026) -
Autorubric: Unifying Rubric-based LLM Evaluation
by: Rao, Delip, et al.
Published: (2026) -
Agreement Metrics for LLM-as-Judge Evaluation: What to Report and Why
by: Rao, Delip, et al.
Published: (2026)