Gaps or Hallucinations? Gazing into Machine-Generated Legal Analysis for Fine-grained Text Evaluations
Fuente:
arXiv
Saved in:
| Main Authors: | Hou, Abe Bohan, Jurayj, William, Holzenberger, Nils, Blair-Stanek, Andrew, Van Durme, Benjamin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Language Models and Logic Programs for Trustworthy Tax Reasoning
by: Jurayj, William, et al.
Published: (2025)
by: Jurayj, William, et al.
Published: (2025)
CLERC: A Dataset for Legal Case Retrieval and Retrieval-Augmented Analysis Generation
by: Hou, Abe Bohan, et al.
Published: (2024)
by: Hou, Abe Bohan, et al.
Published: (2024)
BLT: Can Large Language Models Handle Basic Legal Text?
by: Blair-Stanek, Andrew, et al.
Published: (2023)
by: Blair-Stanek, Andrew, et al.
Published: (2023)
LLMs Provide Unstable Answers to Legal Questions
by: Blair-Stanek, Andrew, et al.
Published: (2025)
by: Blair-Stanek, Andrew, et al.
Published: (2025)
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
by: Blair-Stanek, Andrew, et al.
Published: (2023)
by: Blair-Stanek, Andrew, et al.
Published: (2023)
Can LLMs Identify Tax Abuse?
by: Blair-Stanek, Andrew, et al.
Published: (2025)
by: Blair-Stanek, Andrew, et al.
Published: (2025)
DeonticBench: A Benchmark for Reasoning over Rules
by: Dou, Guangyao, et al.
Published: (2026)
by: Dou, Guangyao, et al.
Published: (2026)
Is That Your Final Answer? Test-Time Scaling Improves Selective Question Answering
by: Jurayj, William, et al.
Published: (2025)
by: Jurayj, William, et al.
Published: (2025)
k-SemStamp: A Clustering-Based Semantic Watermark for Detection of Machine-Generated Text
by: Hou, Abe Bohan, et al.
Published: (2024)
by: Hou, Abe Bohan, et al.
Published: (2024)
Can AI expose tax loopholes? Towards a new generation of legal policy assistants
by: Fratrič, Peter, et al.
Published: (2025)
by: Fratrič, Peter, et al.
Published: (2025)
Reframing Tax Law Entailment as Analogical Reasoning
by: Zou, Xinrui, et al.
Published: (2024)
by: Zou, Xinrui, et al.
Published: (2024)
Crystal: Characterizing Relative Impact of Scholarly Publications
by: Collison, Hannah, et al.
Published: (2026)
by: Collison, Hannah, et al.
Published: (2026)
Process Supervision of Confidence Margin for Calibrated LLM Reasoning
by: Wang, Liaoyaqi, et al.
Published: (2026)
by: Wang, Liaoyaqi, et al.
Published: (2026)
Weird Generalization is Weirdly Brittle
by: Wanner, Miriam, et al.
Published: (2026)
by: Wanner, Miriam, et al.
Published: (2026)
Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models
by: Dahl, Matthew, et al.
Published: (2024)
by: Dahl, Matthew, et al.
Published: (2024)
Hallucination-Free? Assessing the Reliability of Leading AI Legal Research Tools
by: Magesh, Varun, et al.
Published: (2024)
by: Magesh, Varun, et al.
Published: (2024)
Many-Tier Instruction Hierarchy in LLM Agents
by: Zhang, Jingyu, et al.
Published: (2026)
by: Zhang, Jingyu, et al.
Published: (2026)
Careless Whisper: Speech-to-Text Hallucination Harms
by: Koenecke, Allison, et al.
Published: (2024)
by: Koenecke, Allison, et al.
Published: (2024)
Why Avoid Generative Legal AI Systems? Hallucination, Overreliance, and their Impact on Explainability
by: Varkonyi, Gizem Gültekin
Published: (2026)
by: Varkonyi, Gizem Gültekin
Published: (2026)
Unified Multimodal Uncertain Inference
by: Zhang, Dengjia, et al.
Published: (2026)
by: Zhang, Dengjia, et al.
Published: (2026)
Place Matters: Comparing LLM Hallucination Rates for Place-Based Legal Queries
by: Curran, Damian, et al.
Published: (2025)
by: Curran, Damian, et al.
Published: (2025)
Bridging Research Gaps Between Academic Research and Legal Investigations of Algorithmic Discrimination
by: Chien, Colleen V., et al.
Published: (2025)
by: Chien, Colleen V., et al.
Published: (2025)
Always Tell Me The Odds: Fine-grained Conditional Probability Estimation
by: Wang, Liaoyaqi, et al.
Published: (2025)
by: Wang, Liaoyaqi, et al.
Published: (2025)
The Dark Side of ChatGPT: Legal and Ethical Challenges from Stochastic Parrots and Hallucination
by: Li, Zihao
Published: (2023)
by: Li, Zihao
Published: (2023)
Facets of Disparate Impact: Evaluating Legally Consistent Bias in Machine Learning
by: Briscoe, Jarren, et al.
Published: (2025)
by: Briscoe, Jarren, et al.
Published: (2025)
Mind The Gap: How The Technical Mechanism Of Agentic AI Outpace Global Legal Frameworks
by: Osmond, Marcel, et al.
Published: (2026)
by: Osmond, Marcel, et al.
Published: (2026)
Training for Technology: Adoption and Productive Use of Generative AI in Legal Analysis
by: Chen, Benjamin M., et al.
Published: (2026)
by: Chen, Benjamin M., et al.
Published: (2026)
From Argumentation to Deliberation: Perspectivized Stance Vectors for Fine-grained (Dis)agreement Analysis
by: Plenz, Moritz, et al.
Published: (2025)
by: Plenz, Moritz, et al.
Published: (2025)
A Multimodal, Multilingual, and Multidimensional Pipeline for Fine-grained Crowdsourcing Earthquake Damage Evaluation
by: Ma, Zihui, et al.
Published: (2025)
by: Ma, Zihui, et al.
Published: (2025)
Fine-grained Classification of A Million Life Trajectories from Wikipedia
by: Liu, Zhaoyang, et al.
Published: (2026)
by: Liu, Zhaoyang, et al.
Published: (2026)
Transfeminist AI Governance
by: Attard-Frost, Blair
Published: (2025)
by: Attard-Frost, Blair
Published: (2025)
Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurement
by: Kersting, Nicholas S., et al.
Published: (2026)
by: Kersting, Nicholas S., et al.
Published: (2026)
SemStamp: A Semantic Watermark with Paraphrastic Robustness for Text Generation
by: Hou, Abe Bohan, et al.
Published: (2023)
by: Hou, Abe Bohan, et al.
Published: (2023)
Artificial Intelligence and Legal Analysis: Implications for Legal Education and the Profession
by: Peoples, Lee
Published: (2025)
by: Peoples, Lee
Published: (2025)
Machines in the Crowd? Measuring the Footprint of Machine-Generated Text on Reddit
by: La Cava, Lucio, et al.
Published: (2025)
by: La Cava, Lucio, et al.
Published: (2025)
Authorship Attribution in Multilingual Machine-Generated Texts
by: La Cava, Lucio, et al.
Published: (2025)
by: La Cava, Lucio, et al.
Published: (2025)
Towards Unsupervised Question Answering System with Multi-level Summarization for Legal Text
by: Prabhu, M Manvith, et al.
Published: (2024)
by: Prabhu, M Manvith, et al.
Published: (2024)
Attention is All You Want: Machinic Gaze and the Anthropocene
by: Magee, Liam, et al.
Published: (2024)
by: Magee, Liam, et al.
Published: (2024)
Evaluating LLM-Generated Legal Explanations for Regulatory Compliance in Social Media Influencer Marketing
by: Gui, Haoyang, et al.
Published: (2025)
by: Gui, Haoyang, et al.
Published: (2025)
AI-Powered Legal Intelligence System Architecture: A Comprehensive Framework for Automated Legal Consultation and Analysis
by: Kalaycioglu, Sean, et al.
Published: (2025)
by: Kalaycioglu, Sean, et al.
Published: (2025)
Similar Items
-
Language Models and Logic Programs for Trustworthy Tax Reasoning
by: Jurayj, William, et al.
Published: (2025) -
CLERC: A Dataset for Legal Case Retrieval and Retrieval-Augmented Analysis Generation
by: Hou, Abe Bohan, et al.
Published: (2024) -
BLT: Can Large Language Models Handle Basic Legal Text?
by: Blair-Stanek, Andrew, et al.
Published: (2023) -
LLMs Provide Unstable Answers to Legal Questions
by: Blair-Stanek, Andrew, et al.
Published: (2025) -
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
by: Blair-Stanek, Andrew, et al.
Published: (2023)