A Reasoning-Focused Legal Retrieval Benchmark
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Lucia, Guha, Neel, Arifov, Javokhir, Zhang, Sarah, Skreta, Michal, Manning, Christopher D., Henderson, Peter, Ho, Daniel E. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AI for Scaling Legal Reform: Mapping and Redacting Racial Covenants in Santa Clara County
by: Surani, Faiz, et al.
Published: (2025)
by: Surani, Faiz, et al.
Published: (2025)
LawInstruct: A Resource for Studying Language Model Adaptation to the Legal Domain
by: Niklaus, Joel, et al.
Published: (2024)
by: Niklaus, Joel, et al.
Published: (2024)
Hallucination-Free? Assessing the Reliability of Leading AI Legal Research Tools
by: Magesh, Varun, et al.
Published: (2024)
by: Magesh, Varun, et al.
Published: (2024)
Legal-DC: Benchmarking Retrieval-Augmented Generation for Legal Documents
by: Li, Yaocong, et al.
Published: (2026)
by: Li, Yaocong, et al.
Published: (2026)
Benchmarking Legal RAG: The Promise and Limits of AI Statutory Surveys
by: Afane, Mohamed, et al.
Published: (2026)
by: Afane, Mohamed, et al.
Published: (2026)
OpenExempt: A Diagnostic Benchmark for Legal Reasoning and a Framework for Creating Custom Benchmarks on Demand
by: Servantez, Sergio, et al.
Published: (2026)
by: Servantez, Sergio, et al.
Published: (2026)
Statistical Uncertainty in Word Embeddings: GloVe-V
by: Vallebueno, Andrea, et al.
Published: (2024)
by: Vallebueno, Andrea, et al.
Published: (2024)
Stronger Baselines for Retrieval-Augmented Generation with Long-Context Language Models
by: Laitenberger, Alex, et al.
Published: (2025)
by: Laitenberger, Alex, et al.
Published: (2025)
GAIus: Combining Genai with Legal Clauses Retrieval for Knowledge-based Assistant
by: Matak, Michał, et al.
Published: (2025)
by: Matak, Michał, et al.
Published: (2025)
Korean Canonical Legal Benchmark: Toward Knowledge-Independent Evaluation of LLMs' Legal Reasoning Capabilities
by: Oh, Hongseok, et al.
Published: (2025)
by: Oh, Hongseok, et al.
Published: (2025)
Query-Focused Retrieval Heads Improve Long-Context Reasoning and Re-ranking
by: Zhang, Wuwei, et al.
Published: (2025)
by: Zhang, Wuwei, et al.
Published: (2025)
Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models
by: Dahl, Matthew, et al.
Published: (2024)
by: Dahl, Matthew, et al.
Published: (2024)
One Law, Many Languages: Benchmarking Multilingual Legal Reasoning for Judicial Support
by: Stern, Ronja, et al.
Published: (2023)
by: Stern, Ronja, et al.
Published: (2023)
AR-BENCH: Benchmarking Legal Reasoning with Judgment Error Detection, Classification and Correction
by: Li, Yifei, et al.
Published: (2026)
by: Li, Yifei, et al.
Published: (2026)
Humans and transformer LMs: Abstraction drives language learning
by: Jian, Jasper, et al.
Published: (2026)
by: Jian, Jasper, et al.
Published: (2026)
LegalGraphRAG: Multi-Agent Graph Retrieval-Augmented Generation for Reliable Legal Reasoning
by: Chen, Zerui, et al.
Published: (2026)
by: Chen, Zerui, et al.
Published: (2026)
ReCOGS: How Incidental Details of a Logical Form Overshadow an Evaluation of Semantic Interpretation
by: Wu, Zhengxuan, et al.
Published: (2023)
by: Wu, Zhengxuan, et al.
Published: (2023)
ALARB: An Arabic Legal Argument Reasoning Benchmark
by: Shairah, Harethah Abu, et al.
Published: (2025)
by: Shairah, Harethah Abu, et al.
Published: (2025)
JUÁ -- A Benchmark for Information Retrieval in Brazilian Legal Text Collections
by: Pereira, Jayr, et al.
Published: (2026)
by: Pereira, Jayr, et al.
Published: (2026)
CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval
by: Putta, Akshith Reddy, et al.
Published: (2026)
by: Putta, Akshith Reddy, et al.
Published: (2026)
Impacts of Continued Legal Pre-Training and IFT on LLMs' Latent Representations of Human-Defined Legal Concepts
by: Ho, Shaun
Published: (2024)
by: Ho, Shaun
Published: (2024)
LegalRikai: Open Benchmark -- Benchmark for Complex Japanese Corporate Legal Tasks
by: Fujita, Shogo, et al.
Published: (2025)
by: Fujita, Shogo, et al.
Published: (2025)
ACORD: An Expert-Annotated Retrieval Dataset for Legal Contract Drafting
by: Wang, Steven H., et al.
Published: (2025)
by: Wang, Steven H., et al.
Published: (2025)
RAR-b: Reasoning as Retrieval Benchmark
by: Xiao, Chenghao, et al.
Published: (2024)
by: Xiao, Chenghao, et al.
Published: (2024)
PIRB: A Comprehensive Benchmark of Polish Dense and Hybrid Text Retrieval Methods
by: Dadas, Sławomir, et al.
Published: (2024)
by: Dadas, Sławomir, et al.
Published: (2024)
RAPTOR: Recursive Abstractive Processing for Tree-Organized Retrieval
by: Sarthi, Parth, et al.
Published: (2024)
by: Sarthi, Parth, et al.
Published: (2024)
GreekBarBench: A Challenging Benchmark for Free-Text Legal Reasoning and Citations
by: Chlapanis, Odysseas S., et al.
Published: (2025)
by: Chlapanis, Odysseas S., et al.
Published: (2025)
MAWARITH: A Dataset and Benchmark for Legal Inheritance Reasoning with LLMs
by: Bouchekif, Abdessalam, et al.
Published: (2026)
by: Bouchekif, Abdessalam, et al.
Published: (2026)
AI-Assisted Moot Courts: Simulating Justice-Specific Questioning in Oral Arguments
by: Zhang, Kylie, et al.
Published: (2026)
by: Zhang, Kylie, et al.
Published: (2026)
MASLegalBench: Benchmarking Multi-Agent Systems in Deductive Legal Reasoning
by: Jing, Huihao, et al.
Published: (2025)
by: Jing, Huihao, et al.
Published: (2025)
Exploiting LLMs' Reasoning Capability to Infer Implicit Concepts in Legal Information Retrieval
by: Nguyen, Hai-Long, et al.
Published: (2024)
by: Nguyen, Hai-Long, et al.
Published: (2024)
Benchmarking the Legal Reasoning of LLMs in Arabic Islamic Inheritance Cases
by: AlDahoul, Nouar, et al.
Published: (2025)
by: AlDahoul, Nouar, et al.
Published: (2025)
LegalOne: A Family of Foundation Models for Reliable Legal Reasoning
by: Li, Haitao, et al.
Published: (2026)
by: Li, Haitao, et al.
Published: (2026)
Osiris: A Lightweight Open-Source Hallucination Detection System
by: Shan, Alex, et al.
Published: (2025)
by: Shan, Alex, et al.
Published: (2025)
Sneaking Syntax into Transformer Language Models with Tree Regularization
by: Nandi, Ananjan, et al.
Published: (2024)
by: Nandi, Ananjan, et al.
Published: (2024)
LegalSearchLM: Rethinking Legal Case Retrieval as Legal Elements Generation
by: Kim, Chaeeun, et al.
Published: (2025)
by: Kim, Chaeeun, et al.
Published: (2025)
LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation
by: Li, Haitao, et al.
Published: (2025)
by: Li, Haitao, et al.
Published: (2025)
IL-TUR: Benchmark for Indian Legal Text Understanding and Reasoning
by: Joshi, Abhinav, et al.
Published: (2024)
by: Joshi, Abhinav, et al.
Published: (2024)
BriefMe: A Legal NLP Benchmark for Assisting with Legal Briefs
by: Woo, Jesse, et al.
Published: (2025)
by: Woo, Jesse, et al.
Published: (2025)
Precise Legal Sentence Boundary Detection for Retrieval at Scale: NUPunkt and CharBoundary
by: Bommarito, Michael J, et al.
Published: (2025)
by: Bommarito, Michael J, et al.
Published: (2025)
Similar Items
-
AI for Scaling Legal Reform: Mapping and Redacting Racial Covenants in Santa Clara County
by: Surani, Faiz, et al.
Published: (2025) -
LawInstruct: A Resource for Studying Language Model Adaptation to the Legal Domain
by: Niklaus, Joel, et al.
Published: (2024) -
Hallucination-Free? Assessing the Reliability of Leading AI Legal Research Tools
by: Magesh, Varun, et al.
Published: (2024) -
Legal-DC: Benchmarking Retrieval-Augmented Generation for Legal Documents
by: Li, Yaocong, et al.
Published: (2026) -
Benchmarking Legal RAG: The Promise and Limits of AI Statutory Surveys
by: Afane, Mohamed, et al.
Published: (2026)