Hallucination-Free? Assessing the Reliability of Leading AI Legal Research Tools
Fuente:
arXiv
Saved in:
| Main Authors: | Magesh, Varun, Surani, Faiz, Dahl, Matthew, Suzgun, Mirac, Manning, Christopher D., Ho, Daniel E. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models
by: Dahl, Matthew, et al.
Published: (2024)
by: Dahl, Matthew, et al.
Published: (2024)
AI for Scaling Legal Reform: Mapping and Redacting Racial Covenants in Santa Clara County
by: Surani, Faiz, et al.
Published: (2025)
by: Surani, Faiz, et al.
Published: (2025)
Belief in the Machine: Investigating Epistemological Blind Spots of Language Models
by: Suzgun, Mirac, et al.
Published: (2024)
by: Suzgun, Mirac, et al.
Published: (2024)
Bye-bye, Bluebook? Automating Legal Procedure with Large Language Models
by: Dahl, Matthew
Published: (2025)
by: Dahl, Matthew
Published: (2025)
Meta-Prompting: Enhancing Language Models with Task-Agnostic Scaffolding
by: Suzgun, Mirac, et al.
Published: (2024)
by: Suzgun, Mirac, et al.
Published: (2024)
Do Language Models Know When They're Hallucinating References?
by: Agrawal, Ayush, et al.
Published: (2023)
by: Agrawal, Ayush, et al.
Published: (2023)
Aalap: AI Assistant for Legal & Paralegal Functions in India
by: Tiwari, Aman, et al.
Published: (2024)
by: Tiwari, Aman, et al.
Published: (2024)
Evaluating Commercial AI Chatbots as News Intermediaries
by: Suzgun, Mirac, et al.
Published: (2026)
by: Suzgun, Mirac, et al.
Published: (2026)
Gaps or Hallucinations? Gazing into Machine-Generated Legal Analysis for Fine-grained Text Evaluations
by: Hou, Abe Bohan, et al.
Published: (2024)
by: Hou, Abe Bohan, et al.
Published: (2024)
The Cambridge Law Corpus: A Dataset for Legal AI Research
by: Östling, Andreas, et al.
Published: (2023)
by: Östling, Andreas, et al.
Published: (2023)
Place Matters: Comparing LLM Hallucination Rates for Place-Based Legal Queries
by: Curran, Damian, et al.
Published: (2025)
by: Curran, Damian, et al.
Published: (2025)
A Benchmark for Learning to Translate a New Language from One Grammar Book
by: Tanzer, Garrett, et al.
Published: (2023)
by: Tanzer, Garrett, et al.
Published: (2023)
The Dark Side of ChatGPT: Legal and Ethical Challenges from Stochastic Parrots and Hallucination
by: Li, Zihao
Published: (2023)
by: Li, Zihao
Published: (2023)
ArabLegalEval: A Multitask Benchmark for Assessing Arabic Legal Knowledge in Large Language Models
by: Hijazi, Faris, et al.
Published: (2024)
by: Hijazi, Faris, et al.
Published: (2024)
Dynamic Cheatsheet: Test-Time Learning with Adaptive Memory
by: Suzgun, Mirac, et al.
Published: (2025)
by: Suzgun, Mirac, et al.
Published: (2025)
CheckIfExist: Detecting Citation Hallucinations in the Era of AI-Generated Content
by: Abbonato, Diletta
Published: (2026)
by: Abbonato, Diletta
Published: (2026)
Bridging Legal Interpretation and Formal Logic: Faithfulness, Assumption, and the Future of AI Legal Reasoning
by: Wang, Olivia Peiyu, et al.
Published: (2026)
by: Wang, Olivia Peiyu, et al.
Published: (2026)
Cost-of-Pass: An Economic Framework for Evaluating Language Models
by: Erol, Mehmet Hamza, et al.
Published: (2025)
by: Erol, Mehmet Hamza, et al.
Published: (2025)
Eskwai for Students: Generative AI Assistant for Legal Education in Ghana
by: Boateng, George, et al.
Published: (2026)
by: Boateng, George, et al.
Published: (2026)
Software Engineering Methods For AI-Driven Deductive Legal Reasoning
by: Padhye, Rohan
Published: (2024)
by: Padhye, Rohan
Published: (2024)
Safety-Tuned LLaMAs: Lessons From Improving the Safety of Large Language Models that Follow Instructions
by: Bianchi, Federico, et al.
Published: (2023)
by: Bianchi, Federico, et al.
Published: (2023)
Careless Whisper: Speech-to-Text Hallucination Harms
by: Koenecke, Allison, et al.
Published: (2024)
by: Koenecke, Allison, et al.
Published: (2024)
Giving AI Personalities Leads to More Human-Like Reasoning
by: Nighojkar, Animesh, et al.
Published: (2025)
by: Nighojkar, Animesh, et al.
Published: (2025)
Are Models Trained on Indian Legal Data Fair?
by: Girhepuje, Sahil, et al.
Published: (2023)
by: Girhepuje, Sahil, et al.
Published: (2023)
Towards Grammatical Tagging for the Legal Language of Cybersecurity
by: Castiglione, Gianpietro, et al.
Published: (2023)
by: Castiglione, Gianpietro, et al.
Published: (2023)
Mining Legal Arguments to Study Judicial Formalism
by: Koref, Tomáš, et al.
Published: (2025)
by: Koref, Tomáš, et al.
Published: (2025)
LLMs Provide Unstable Answers to Legal Questions
by: Blair-Stanek, Andrew, et al.
Published: (2025)
by: Blair-Stanek, Andrew, et al.
Published: (2025)
Legally Binding but Unfair? Towards Assessing Fairness of Privacy Policies
by: Freiberger, Vincent, et al.
Published: (2024)
by: Freiberger, Vincent, et al.
Published: (2024)
Osiris: A Lightweight Open-Source Hallucination Detection System
by: Shan, Alex, et al.
Published: (2025)
by: Shan, Alex, et al.
Published: (2025)
Caveat Lector: Large Language Models in Legal Practice
by: Mik, Eliza
Published: (2024)
by: Mik, Eliza
Published: (2024)
Exploring Possibilities of AI-Powered Legal Assistance in Bangladesh through Large Language Modeling
by: Wasi, Azmine Toushik, et al.
Published: (2024)
by: Wasi, Azmine Toushik, et al.
Published: (2024)
Legal Fact Prediction: The Missing Piece in Legal Judgment Prediction
by: Liu, Junkai, et al.
Published: (2024)
by: Liu, Junkai, et al.
Published: (2024)
Reducing Tool Hallucination via Reliability Alignment
by: Xu, Hongshen, et al.
Published: (2024)
by: Xu, Hongshen, et al.
Published: (2024)
SemCAFE: When Named Entities make the Difference Assessing Web Source Reliability through Entity-level Analytics
by: Shahi, Gautam Kishore, et al.
Published: (2025)
by: Shahi, Gautam Kishore, et al.
Published: (2025)
A Reasoning-Focused Legal Retrieval Benchmark
by: Zheng, Lucia, et al.
Published: (2025)
by: Zheng, Lucia, et al.
Published: (2025)
CLERC: A Dataset for Legal Case Retrieval and Retrieval-Augmented Analysis Generation
by: Hou, Abe Bohan, et al.
Published: (2024)
by: Hou, Abe Bohan, et al.
Published: (2024)
SecureForge: Finding and Preventing Vulnerabilities in LLM-Generated Code via Prompt Optimization
by: Liu, Houjun, et al.
Published: (2026)
by: Liu, Houjun, et al.
Published: (2026)
Medical Hallucinations in Foundation Models and Their Impact on Healthcare
by: Kim, Yubin, et al.
Published: (2025)
by: Kim, Yubin, et al.
Published: (2025)
Evaluating LLM-Generated Legal Explanations for Regulatory Compliance in Social Media Influencer Marketing
by: Gui, Haoyang, et al.
Published: (2025)
by: Gui, Haoyang, et al.
Published: (2025)
Legal Minds, Algorithmic Decisions: How LLMs Apply Constitutional Principles in Complex Scenarios
by: Bignotti, Camilla, et al.
Published: (2024)
by: Bignotti, Camilla, et al.
Published: (2024)
Similar Items
-
Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models
by: Dahl, Matthew, et al.
Published: (2024) -
AI for Scaling Legal Reform: Mapping and Redacting Racial Covenants in Santa Clara County
by: Surani, Faiz, et al.
Published: (2025) -
Belief in the Machine: Investigating Epistemological Blind Spots of Language Models
by: Suzgun, Mirac, et al.
Published: (2024) -
Bye-bye, Bluebook? Automating Legal Procedure with Large Language Models
by: Dahl, Matthew
Published: (2025) -
Meta-Prompting: Enhancing Language Models with Task-Agnostic Scaffolding
by: Suzgun, Mirac, et al.
Published: (2024)