LegalBench.PT: A Benchmark for Portuguese Law
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Canaverde, Beatriz, Pires, Telmo Pessoa, Ribeiro, Leonor Melo, Martins, André F. T. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
von: Teixeira, Tiago, et al.
Veröffentlicht: (2026)
von: Teixeira, Tiago, et al.
Veröffentlicht: (2026)
SEQUOR: A Multi-Turn Benchmark for Realistic Constraint Following
von: Canaverde, Beatriz, et al.
Veröffentlicht: (2026)
von: Canaverde, Beatriz, et al.
Veröffentlicht: (2026)
SaulLM-7B: A pioneering Large Language Model for Law
von: Colombo, Pierre, et al.
Veröffentlicht: (2024)
von: Colombo, Pierre, et al.
Veröffentlicht: (2024)
MedPT: A Massive Medical Question Answering Dataset for Brazilian-Portuguese Speakers
von: Färber, Fernanda Bufon, et al.
Veröffentlicht: (2025)
von: Färber, Fernanda Bufon, et al.
Veröffentlicht: (2025)
Advancing Neural Encoding of Portuguese with Transformer Albertina PT-*
von: Rodrigues, João, et al.
Veröffentlicht: (2023)
von: Rodrigues, João, et al.
Veröffentlicht: (2023)
SaulLM-54B & SaulLM-141B: Scaling Up Domain Adaptation for the Legal Domain
von: Colombo, Pierre, et al.
Veröffentlicht: (2024)
von: Colombo, Pierre, et al.
Veröffentlicht: (2024)
ClaimPT: A Portuguese Dataset of Annotated Claims in News Articles
von: Campos, Ricardo, et al.
Veröffentlicht: (2026)
von: Campos, Ricardo, et al.
Veröffentlicht: (2026)
CALRK-Bench: Evaluating Context-Aware Legal Reasoning in Korean Law
von: Jung, JiHyeok, et al.
Veröffentlicht: (2026)
von: Jung, JiHyeok, et al.
Veröffentlicht: (2026)
Advancing Generative AI for Portuguese with Open Decoder Gervásio PT*
von: Santos, Rodrigo, et al.
Veröffentlicht: (2024)
von: Santos, Rodrigo, et al.
Veröffentlicht: (2024)
Open Sentence Embeddings for Portuguese with the Serafim PT* encoders family
von: Gomes, Luís, et al.
Veröffentlicht: (2024)
von: Gomes, Luís, et al.
Veröffentlicht: (2024)
ALEXSIS-PT: A New Resource for Portuguese Lexical Simplification
von: North, Kai, et al.
Veröffentlicht: (2022)
von: North, Kai, et al.
Veröffentlicht: (2022)
ACE-2005-PT: Corpus for Event Extraction in Portuguese
von: Cunha, Luís Filipe, et al.
Veröffentlicht: (2024)
von: Cunha, Luís Filipe, et al.
Veröffentlicht: (2024)
GreekBarBench: A Challenging Benchmark for Free-Text Legal Reasoning and Citations
von: Chlapanis, Odysseas S., et al.
Veröffentlicht: (2025)
von: Chlapanis, Odysseas S., et al.
Veröffentlicht: (2025)
Fostering the Ecosystem of Open Neural Encoders for Portuguese with Albertina PT* Family
von: Santos, Rodrigo, et al.
Veröffentlicht: (2024)
von: Santos, Rodrigo, et al.
Veröffentlicht: (2024)
MASLegalBench: Benchmarking Multi-Agent Systems in Deductive Legal Reasoning
von: Jing, Huihao, et al.
Veröffentlicht: (2025)
von: Jing, Huihao, et al.
Veröffentlicht: (2025)
CommitBench: A Benchmark for Commit Message Generation
von: Schall, Maximilian, et al.
Veröffentlicht: (2024)
von: Schall, Maximilian, et al.
Veröffentlicht: (2024)
CAPITU: A Benchmark for Evaluating Instruction-Following in Brazilian Portuguese with Literary Context
von: Bonás, Giovana Kerche, et al.
Veröffentlicht: (2026)
von: Bonás, Giovana Kerche, et al.
Veröffentlicht: (2026)
Magis-Bench: Evaluating LLMs on Magistrate-Level Legal Tasks
von: Pires, Ramon, et al.
Veröffentlicht: (2026)
von: Pires, Ramon, et al.
Veröffentlicht: (2026)
CLARIN-PT-LDB: An Open LLM Leaderboard for Portuguese to assess Language, Culture and Civility
von: Silva, João, et al.
Veröffentlicht: (2026)
von: Silva, João, et al.
Veröffentlicht: (2026)
IslamicLegalBench: Evaluating LLMs Knowledge and Reasoning of Islamic Law Across 1,200 Years of Islamic Pluralist Legal Traditions
von: Elmahjub, Ezieddin, et al.
Veröffentlicht: (2026)
von: Elmahjub, Ezieddin, et al.
Veröffentlicht: (2026)
BenchBench: Benchmarking Automated Benchmark Generation
von: Zheng, Yandan, et al.
Veröffentlicht: (2026)
von: Zheng, Yandan, et al.
Veröffentlicht: (2026)
PLawBench: A Rubric-Based Benchmark for Evaluating LLMs in Real-World Legal Practice
von: Shi, Yuzhen, et al.
Veröffentlicht: (2026)
von: Shi, Yuzhen, et al.
Veröffentlicht: (2026)
PILOT-Bench: A Benchmark for Legal Reasoning in the Patent Domain with IRAC-Aligned Classification Tasks
von: Jang, Yehoon, et al.
Veröffentlicht: (2026)
von: Jang, Yehoon, et al.
Veröffentlicht: (2026)
VLegal-Bench: Cognitively Grounded Benchmark for Vietnamese Legal Reasoning of Large Language Models
von: Dong, Nguyen Tien, et al.
Veröffentlicht: (2025)
von: Dong, Nguyen Tien, et al.
Veröffentlicht: (2025)
LegalBench-BR: A Benchmark for Evaluating Large Language Models on Brazilian Legal Decision Classification
von: Neto, Pedro Barbosa de Carvalho
Veröffentlicht: (2026)
von: Neto, Pedro Barbosa de Carvalho
Veröffentlicht: (2026)
Reranking Laws for Language Generation: A Communication-Theoretic Perspective
von: Farinhas, António, et al.
Veröffentlicht: (2024)
von: Farinhas, António, et al.
Veröffentlicht: (2024)
SwiLTra-Bench: The Swiss Legal Translation Benchmark
von: Niklaus, Joel, et al.
Veröffentlicht: (2025)
von: Niklaus, Joel, et al.
Veröffentlicht: (2025)
LegalAgentBench: Evaluating LLM Agents in Legal Domain
von: Li, Haitao, et al.
Veröffentlicht: (2024)
von: Li, Haitao, et al.
Veröffentlicht: (2024)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
von: Ovcharov, Volodymyr
Veröffentlicht: (2026)
von: Ovcharov, Volodymyr
Veröffentlicht: (2026)
LegalRikai: Open Benchmark -- Benchmark for Complex Japanese Corporate Legal Tasks
von: Fujita, Shogo, et al.
Veröffentlicht: (2025)
von: Fujita, Shogo, et al.
Veröffentlicht: (2025)
AnnoCaseLaw: A Richly-Annotated Dataset For Benchmarking Explainable Legal Judgment Prediction
von: Sesodia, Magnus, et al.
Veröffentlicht: (2025)
von: Sesodia, Magnus, et al.
Veröffentlicht: (2025)
LegalCiteBench: Evaluating Citation Reliability in Legal Language Models
von: Chen, Sijia, et al.
Veröffentlicht: (2026)
von: Chen, Sijia, et al.
Veröffentlicht: (2026)
BriefMe: A Legal NLP Benchmark for Assisting with Legal Briefs
von: Woo, Jesse, et al.
Veröffentlicht: (2025)
von: Woo, Jesse, et al.
Veröffentlicht: (2025)
BenGER: Benchmarking LLM Systems on Subsumption-Based Legal Reasoning in German Law
von: Nagl, Sebastian, et al.
Veröffentlicht: (2026)
von: Nagl, Sebastian, et al.
Veröffentlicht: (2026)
Do These LLM Benchmarks Agree? Fixing Benchmark Evaluation with BenchBench
von: Perlitz, Yotam, et al.
Veröffentlicht: (2024)
von: Perlitz, Yotam, et al.
Veröffentlicht: (2024)
Indexing Portuguese NLP Resources with PT-Pump-Up
von: Almeida, Rúben, et al.
Veröffentlicht: (2024)
von: Almeida, Rúben, et al.
Veröffentlicht: (2024)
LEXam: Benchmarking Legal Reasoning on 340 Law Exams
von: Fan, Yu, et al.
Veröffentlicht: (2025)
von: Fan, Yu, et al.
Veröffentlicht: (2025)
Legal-DC: Benchmarking Retrieval-Augmented Generation for Legal Documents
von: Li, Yaocong, et al.
Veröffentlicht: (2026)
von: Li, Yaocong, et al.
Veröffentlicht: (2026)
AMALIA Technical Report: A Fully Open Source Large Language Model for European Portuguese
von: Simplício, Afonso, et al.
Veröffentlicht: (2026)
von: Simplício, Afonso, et al.
Veröffentlicht: (2026)
A Reasoning-Focused Legal Retrieval Benchmark
von: Zheng, Lucia, et al.
Veröffentlicht: (2025)
von: Zheng, Lucia, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
von: Teixeira, Tiago, et al.
Veröffentlicht: (2026) -
SEQUOR: A Multi-Turn Benchmark for Realistic Constraint Following
von: Canaverde, Beatriz, et al.
Veröffentlicht: (2026) -
SaulLM-7B: A pioneering Large Language Model for Law
von: Colombo, Pierre, et al.
Veröffentlicht: (2024) -
MedPT: A Massive Medical Question Answering Dataset for Brazilian-Portuguese Speakers
von: Färber, Fernanda Bufon, et al.
Veröffentlicht: (2025) -
Advancing Neural Encoding of Portuguese with Transformer Albertina PT-*
von: Rodrigues, João, et al.
Veröffentlicht: (2023)