BriefMe: A Legal NLP Benchmark for Assisting with Legal Briefs
Fuente:
arXiv
Saved in:
| Main Authors: | Woo, Jesse, Chaleshtori, Fateme Hashemi, Marasović, Ana, Marino, Kenneth |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On Evaluating Explanation Utility for Human-AI Decision Making in NLP
by: Chaleshtori, Fateme Hashemi, et al.
Published: (2024)
by: Chaleshtori, Fateme Hashemi, et al.
Published: (2024)
Measuring Chain of Thought Faithfulness by Unlearning Reasoning Steps
by: Tutek, Martin, et al.
Published: (2025)
by: Tutek, Martin, et al.
Published: (2025)
Teaching People LLM's Errors and Getting it Right
by: Stringham, Nathan, et al.
Published: (2025)
by: Stringham, Nathan, et al.
Published: (2025)
Indian Legal NLP Benchmarks : A Survey
by: Kalamkar, Prathamesh, et al.
Published: (2021)
by: Kalamkar, Prathamesh, et al.
Published: (2021)
Whispers of Doubt Amidst Echoes of Triumph in NLP Robustness
by: Gupta, Ashim, et al.
Published: (2023)
by: Gupta, Ashim, et al.
Published: (2023)
Legal Briefs
Published: (2025)
Published: (2025)
Legal Briefs
Published: (2025)
Published: (2025)
Legal Briefs
Published: (2025)
Published: (2025)
Legal Briefs
Published: (2025)
Published: (2025)
Legal Briefs
Published: (2025)
Published: (2025)
Legal Briefs
Published: (2025)
Published: (2025)
Legal Briefs
Published: (2025)
Published: (2025)
Legal Briefs
Published: (2026)
Published: (2026)
Legal Briefs
Published: (2024)
Published: (2024)
Legal Briefs
Published: (2024)
Published: (2024)
Legal Briefs
Published: (2024)
Published: (2024)
Legal Briefs
Published: (2024)
Published: (2024)
Legal Briefs
Published: (2024)
Published: (2024)
Legal Briefs
Published: (2024)
Published: (2024)
LegalRikai: Open Benchmark -- Benchmark for Complex Japanese Corporate Legal Tasks
by: Fujita, Shogo, et al.
Published: (2025)
by: Fujita, Shogo, et al.
Published: (2025)
Legal-DC: Benchmarking Retrieval-Augmented Generation for Legal Documents
by: Li, Yaocong, et al.
Published: (2026)
by: Li, Yaocong, et al.
Published: (2026)
A Reasoning-Focused Legal Retrieval Benchmark
by: Zheng, Lucia, et al.
Published: (2025)
by: Zheng, Lucia, et al.
Published: (2025)
Korean Canonical Legal Benchmark: Toward Knowledge-Independent Evaluation of LLMs' Legal Reasoning Capabilities
by: Oh, Hongseok, et al.
Published: (2025)
by: Oh, Hongseok, et al.
Published: (2025)
LegalEval-Q: A New Benchmark for The Quality Evaluation of LLM-Generated Legal Text
by: yunhan, Li, et al.
Published: (2025)
by: yunhan, Li, et al.
Published: (2025)
LegalBench.PT: A Benchmark for Portuguese Law
by: Canaverde, Beatriz, et al.
Published: (2025)
by: Canaverde, Beatriz, et al.
Published: (2025)
Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth
by: Gur-Arieh, Yoav, et al.
Published: (2026)
by: Gur-Arieh, Yoav, et al.
Published: (2026)
ArabLegalEval: A Multitask Benchmark for Assessing Arabic Legal Knowledge in Large Language Models
by: Hijazi, Faris, et al.
Published: (2024)
by: Hijazi, Faris, et al.
Published: (2024)
LexTime: A Benchmark for Temporal Ordering of Legal Events
by: Barale, Claire, et al.
Published: (2025)
by: Barale, Claire, et al.
Published: (2025)
Towards Supporting Legal Argumentation with NLP: Is More Data Really All You Need?
by: Santosh, T. Y. S. S, et al.
Published: (2024)
by: Santosh, T. Y. S. S, et al.
Published: (2024)
A Small Claims Court for the NLP: Judging Legal Text Classification Strategies With Small Datasets
by: Noguti, Mariana Yukari, et al.
Published: (2024)
by: Noguti, Mariana Yukari, et al.
Published: (2024)
A Brief History of Named Entity Recognition
by: Munnangi, Monica
Published: (2024)
by: Munnangi, Monica
Published: (2024)
What Has Been Lost with Synthetic Evaluation?
by: Gill, Alexander, et al.
Published: (2025)
by: Gill, Alexander, et al.
Published: (2025)
LegalOne: A Family of Foundation Models for Reliable Legal Reasoning
by: Li, Haitao, et al.
Published: (2026)
by: Li, Haitao, et al.
Published: (2026)
LegalSearchLM: Rethinking Legal Case Retrieval as Legal Elements Generation
by: Kim, Chaeeun, et al.
Published: (2025)
by: Kim, Chaeeun, et al.
Published: (2025)
The Massive Legal Embedding Benchmark (MLEB)
by: Butler, Umar, et al.
Published: (2025)
by: Butler, Umar, et al.
Published: (2025)
LegalViz: Legal Text Visualization by Text To Diagram Generation
by: Onami, Eri, et al.
Published: (2025)
by: Onami, Eri, et al.
Published: (2025)
A Brief Summary of Explanatory Virtues
by: Zukerman, Ingrid
Published: (2024)
by: Zukerman, Ingrid
Published: (2024)
LAiW: A Chinese Legal Large Language Models Benchmark
by: Dai, Yongfu, et al.
Published: (2023)
by: Dai, Yongfu, et al.
Published: (2023)
Benchmarking Legal RAG: The Promise and Limits of AI Statutory Surveys
by: Afane, Mohamed, et al.
Published: (2026)
by: Afane, Mohamed, et al.
Published: (2026)
AppellateGen: A Benchmark for Appellate Legal Judgment Generation
by: Yang, Hongkun, et al.
Published: (2026)
by: Yang, Hongkun, et al.
Published: (2026)
Similar Items
-
On Evaluating Explanation Utility for Human-AI Decision Making in NLP
by: Chaleshtori, Fateme Hashemi, et al.
Published: (2024) -
Measuring Chain of Thought Faithfulness by Unlearning Reasoning Steps
by: Tutek, Martin, et al.
Published: (2025) -
Teaching People LLM's Errors and Getting it Right
by: Stringham, Nathan, et al.
Published: (2025) -
Indian Legal NLP Benchmarks : A Survey
by: Kalamkar, Prathamesh, et al.
Published: (2021) -
Whispers of Doubt Amidst Echoes of Triumph in NLP Robustness
by: Gupta, Ashim, et al.
Published: (2023)