LegalBench-BR: A Benchmark for Evaluating Large Language Models on Brazilian Legal Decision Classification
Fuente:
arXiv
Salvato in:
| Autore principale: | Neto, Pedro Barbosa de Carvalho |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Using Instruction-Tuned Large Language Models to Identify Indicators of Vulnerability in Police Incident Narratives
di: Relins, Sam, et al.
Pubblicazione: (2024)
di: Relins, Sam, et al.
Pubblicazione: (2024)
Data Processing for the OpenGPT-X Model Family
di: Brandizzi, Nicolo', et al.
Pubblicazione: (2024)
di: Brandizzi, Nicolo', et al.
Pubblicazione: (2024)
Detection of Personal Data in Structured Datasets Using a Large Language Model
di: Ntwali, Albert Agisha, et al.
Pubblicazione: (2025)
di: Ntwali, Albert Agisha, et al.
Pubblicazione: (2025)
Argument Quality Annotation and Gender Bias Detection in Financial Communication through Large Language Models
di: Alhamzeh, Alaa, et al.
Pubblicazione: (2025)
di: Alhamzeh, Alaa, et al.
Pubblicazione: (2025)
LLM-based Extraction of Contradictions from Patents
di: Trapp, Stefan, et al.
Pubblicazione: (2024)
di: Trapp, Stefan, et al.
Pubblicazione: (2024)
Supervised Semantic Differential for Cross-Cultural Concept Analysis: A Case Study of Human Affect
di: Sikora, Jan, et al.
Pubblicazione: (2026)
di: Sikora, Jan, et al.
Pubblicazione: (2026)
IndiaFinBench: An Evaluation Benchmark for Large Language Model Performance on Indian Financial Regulatory Text
di: Pall, Rajveer Singh
Pubblicazione: (2026)
di: Pall, Rajveer Singh
Pubblicazione: (2026)
Mining Large Language Models for Low-Resource Language Data: Comparing Elicitation Strategies for Hausa and Fongbe
di: Adjovi, Mahounan Pericles, et al.
Pubblicazione: (2026)
di: Adjovi, Mahounan Pericles, et al.
Pubblicazione: (2026)
CIG: Measuring Conversational Information Gain in Deliberative Dialogues with Semantic Memory Dynamics
di: Chen, Ming-Bin, et al.
Pubblicazione: (2026)
di: Chen, Ming-Bin, et al.
Pubblicazione: (2026)
Advancing Uto-Aztecan Language Technologies: A Case Study on the Endangered Comanche Language
di: C, Jesus Alvarez, et al.
Pubblicazione: (2025)
di: C, Jesus Alvarez, et al.
Pubblicazione: (2025)
When F1 Fails: Granularity-Aware Evaluation for Dialogue Topic Segmentation
di: Coen, Michael H.
Pubblicazione: (2025)
di: Coen, Michael H.
Pubblicazione: (2025)
AI-Friendly LaTeX: Using LaTeX Code as a Knowledge Source for Retrieval-Augmented Generation
di: Verhoeff, Tom
Pubblicazione: (2026)
di: Verhoeff, Tom
Pubblicazione: (2026)
TMT: A Simple Way to Translate Topic Models Using Dictionaries
di: Engl, Felix, et al.
Pubblicazione: (2025)
di: Engl, Felix, et al.
Pubblicazione: (2025)
A Survey of Text and Speech Resources for Hausa and Fongbe: Availability, Quality, and Gaps for NLP Development
di: Adjovi, Mahounan Pericles, et al.
Pubblicazione: (2026)
di: Adjovi, Mahounan Pericles, et al.
Pubblicazione: (2026)
LLMs as Architects and Critics for Multi-Source Opinion Summarization
di: Attri, Anuj, et al.
Pubblicazione: (2025)
di: Attri, Anuj, et al.
Pubblicazione: (2025)
Why We Feel What We Feel: Joint Detection of Emotions and Their Opinion Triggers in E-commerce
di: Attri, Arnav, et al.
Pubblicazione: (2025)
di: Attri, Arnav, et al.
Pubblicazione: (2025)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
di: Ovcharov, Volodymyr
Pubblicazione: (2026)
di: Ovcharov, Volodymyr
Pubblicazione: (2026)
PerSoMed: A Large-Scale Balanced Dataset for Persian Social Media Text Classification
di: Chehreh, Isun, et al.
Pubblicazione: (2026)
di: Chehreh, Isun, et al.
Pubblicazione: (2026)
When Does Data Augmentation Help? Evaluating LLM and Back-Translation Methods for Hausa and Fongbe NLP
di: Adjovi, Mahounan Pericles, et al.
Pubblicazione: (2026)
di: Adjovi, Mahounan Pericles, et al.
Pubblicazione: (2026)
Permanent Data Encoding (PDE): A Visual Language for Semantic Compression and Knowledge Preservation in 3-Character Units
di: Tsuyuki, Yoshiharu, et al.
Pubblicazione: (2025)
di: Tsuyuki, Yoshiharu, et al.
Pubblicazione: (2025)
Legal RAG Bench: an end-to-end benchmark for legal RAG
di: Butler, Abdur-Rahman, et al.
Pubblicazione: (2026)
di: Butler, Abdur-Rahman, et al.
Pubblicazione: (2026)
"This Suits You the Best": Query Focused Comparative Explainable Summarization
di: Attri, Arnav, et al.
Pubblicazione: (2025)
di: Attri, Arnav, et al.
Pubblicazione: (2025)
EPIC-EuroParl-UdS: Information-Theoretic Perspectives on Translation and Interpreting
di: Kunilovskaya, Maria, et al.
Pubblicazione: (2026)
di: Kunilovskaya, Maria, et al.
Pubblicazione: (2026)
Translationese as a Rational Response to Translation Task Difficulty
di: Kunilovskaya, Maria
Pubblicazione: (2026)
di: Kunilovskaya, Maria
Pubblicazione: (2026)
Tokenization Standards for Linguistic Integrity: Turkish as a Benchmark
di: Bayram, M. Ali, et al.
Pubblicazione: (2025)
di: Bayram, M. Ali, et al.
Pubblicazione: (2025)
Cognitive Linguistic Identity Fusion Score (CLIFS): A Scalable Cognition-Informed Approach to Quantifying Identity Fusion from Text
di: Wright, Devin R., et al.
Pubblicazione: (2025)
di: Wright, Devin R., et al.
Pubblicazione: (2025)
The Syntactic Acceptability Dataset (Preview): A Resource for Machine Learning and Linguistic Analysis of English
di: Juzek, Tom S
Pubblicazione: (2025)
di: Juzek, Tom S
Pubblicazione: (2025)
OpenGloss: A Synthetic Encyclopedic Dictionary and Semantic Knowledge Graph
di: Bommarito II, Michael J.
Pubblicazione: (2025)
di: Bommarito II, Michael J.
Pubblicazione: (2025)
Tokens with Meaning: A Hybrid Tokenization Approach for Turkish
di: Bayram, M. Ali, et al.
Pubblicazione: (2025)
di: Bayram, M. Ali, et al.
Pubblicazione: (2025)
LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
di: Pradhan, Anu, et al.
Pubblicazione: (2025)
di: Pradhan, Anu, et al.
Pubblicazione: (2025)
RenoBench: A Citation Parsing Benchmark
di: Sarin, Parth, et al.
Pubblicazione: (2026)
di: Sarin, Parth, et al.
Pubblicazione: (2026)
Low-Resource English-Tigrinya MT: Leveraging Multilingual Models, Custom Tokenizers, and Clean Evaluation Benchmarks
di: Teklehaymanot, Hailay Kidu, et al.
Pubblicazione: (2025)
di: Teklehaymanot, Hailay Kidu, et al.
Pubblicazione: (2025)
Large Language Models for Patent Classification: Strengths, Trade-offs, and the Long Tail Effect
di: Emer, Lorenzo, et al.
Pubblicazione: (2026)
di: Emer, Lorenzo, et al.
Pubblicazione: (2026)
MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
di: Li, Yingyun, et al.
Pubblicazione: (2026)
di: Li, Yingyun, et al.
Pubblicazione: (2026)
Leveraging LLMs to Create Content Corpora for Niche Domains
di: Zhang, Franklin, et al.
Pubblicazione: (2025)
di: Zhang, Franklin, et al.
Pubblicazione: (2025)
HiPS: Hierarchical PDF Segmentation of Textbooks
di: Wehnert, Sabine, et al.
Pubblicazione: (2025)
di: Wehnert, Sabine, et al.
Pubblicazione: (2025)
An efficient domain-independent approach for supervised keyphrase extraction and ranking
di: Ramaswamy, Sriraghavendra
Pubblicazione: (2024)
di: Ramaswamy, Sriraghavendra
Pubblicazione: (2024)
Reading Between the Waves: Robust Topic Segmentation Using Inter-Sentence Audio Features
di: Freisinger, Steffen, et al.
Pubblicazione: (2026)
di: Freisinger, Steffen, et al.
Pubblicazione: (2026)
Enhancing Scientific Literature Chatbots with Retrieval-Augmented Generation: A Performance Evaluation of Vector and Graph-Based Systems
di: Ghanadian, Hamideh, et al.
Pubblicazione: (2026)
di: Ghanadian, Hamideh, et al.
Pubblicazione: (2026)
How Large Language Models Are Changing MOOC Essay Answers: A Comparison of Pre- and Post-LLM Responses
di: Leppänen, Leo, et al.
Pubblicazione: (2025)
di: Leppänen, Leo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Using Instruction-Tuned Large Language Models to Identify Indicators of Vulnerability in Police Incident Narratives
di: Relins, Sam, et al.
Pubblicazione: (2024) -
Data Processing for the OpenGPT-X Model Family
di: Brandizzi, Nicolo', et al.
Pubblicazione: (2024) -
Detection of Personal Data in Structured Datasets Using a Large Language Model
di: Ntwali, Albert Agisha, et al.
Pubblicazione: (2025) -
Argument Quality Annotation and Gender Bias Detection in Financial Communication through Large Language Models
di: Alhamzeh, Alaa, et al.
Pubblicazione: (2025) -
LLM-based Extraction of Contradictions from Patents
di: Trapp, Stefan, et al.
Pubblicazione: (2024)