JUÁ -- A Benchmark for Information Retrieval in Brazilian Legal Text Collections
Fuente:
arXiv
Saved in:
| Main Authors: | Pereira, Jayr, Fernandes, Leandro, de Brito, Erick, Lotufo, Roberto, Bonifacio, Luiz |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Domain-Adaptive Dense Retrieval for Brazilian Legal Search
by: Pereira, Jayr, et al.
Published: (2026)
by: Pereira, Jayr, et al.
Published: (2026)
Quati: A Brazilian Portuguese Information Retrieval Dataset from Native Speakers
by: Bueno, Mirelle, et al.
Published: (2024)
by: Bueno, Mirelle, et al.
Published: (2024)
JurisTCU: A Brazilian Portuguese Information Retrieval Dataset with Query Relevance Judgments
by: Fernandes, Leandro Carísio, et al.
Published: (2025)
by: Fernandes, Leandro Carísio, et al.
Published: (2025)
Exploring Selective Retrieval-Augmentation for Long-Tail Legal Text Classification
by: Mao, Boheng
Published: (2025)
by: Mao, Boheng
Published: (2025)
ptt5-v2: A Closer Look at Continued Pretraining of T5 Models for the Portuguese Language
by: Piau, Marcos, et al.
Published: (2024)
by: Piau, Marcos, et al.
Published: (2024)
Retail-GPT: leveraging Retrieval Augmented Generation (RAG) for building E-commerce Chat Assistants
by: de Freitas, Bruno Amaral Teixeira, et al.
Published: (2024)
by: de Freitas, Bruno Amaral Teixeira, et al.
Published: (2024)
LegalSearchLM: Rethinking Legal Case Retrieval as Legal Elements Generation
by: Kim, Chaeeun, et al.
Published: (2025)
by: Kim, Chaeeun, et al.
Published: (2025)
LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation
by: Li, Haitao, et al.
Published: (2025)
by: Li, Haitao, et al.
Published: (2025)
Assessing the Performance Gap Between Lexical and Semantic Models for Information Retrieval With Formulaic Legal Language
by: Mori, Larissa, et al.
Published: (2025)
by: Mori, Larissa, et al.
Published: (2025)
CAPTAIN at COLIEE 2023: Efficient Methods for Legal Information Retrieval and Entailment Tasks
by: Nguyen, Chau, et al.
Published: (2024)
by: Nguyen, Chau, et al.
Published: (2024)
LegalRAG: A Hybrid RAG System for Multilingual Legal Information Retrieval
by: Kabir, Muhammad Rafsan, et al.
Published: (2025)
by: Kabir, Muhammad Rafsan, et al.
Published: (2025)
Logic Rules as Explanations for Legal Case Retrieval
by: Sun, Zhongxiang, et al.
Published: (2024)
by: Sun, Zhongxiang, et al.
Published: (2024)
Building Russian Benchmark for Evaluation of Information Retrieval Models
by: Kovalev, Grigory, et al.
Published: (2025)
by: Kovalev, Grigory, et al.
Published: (2025)
AIR-Bench: Automated Heterogeneous Information Retrieval Benchmark
by: Chen, Jianlyu, et al.
Published: (2024)
by: Chen, Jianlyu, et al.
Published: (2024)
Adaptive Two-Phase Finetuning LLMs for Japanese Legal Text Retrieval
by: Trung, Quang Hoang, et al.
Published: (2024)
by: Trung, Quang Hoang, et al.
Published: (2024)
CoIR: A Comprehensive Benchmark for Code Information Retrieval Models
by: Li, Xiangyang, et al.
Published: (2024)
by: Li, Xiangyang, et al.
Published: (2024)
ExaRanker-Open: Synthetic Explanation for IR using Open-Source LLMs
by: Ferraretto, Fernando, et al.
Published: (2024)
by: Ferraretto, Fernando, et al.
Published: (2024)
PosIR: Position-Aware Heterogeneous Information Retrieval Benchmark
by: Zeng, Ziyang, et al.
Published: (2026)
by: Zeng, Ziyang, et al.
Published: (2026)
Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integration
by: Dai, Sunhao, et al.
Published: (2024)
by: Dai, Sunhao, et al.
Published: (2024)
"Knowing When You Don't Know": A Multilingual Relevance Assessment Dataset for Robust Retrieval-Augmented Generation
by: Thakur, Nandan, et al.
Published: (2023)
by: Thakur, Nandan, et al.
Published: (2023)
Wikipedia-based Datasets in Russian Information Retrieval Benchmark RusBEIR
by: Kovalev, Grigory, et al.
Published: (2025)
by: Kovalev, Grigory, et al.
Published: (2025)
IFIR: A Comprehensive Benchmark for Evaluating Instruction-Following in Expert-Domain Information Retrieval
by: Song, Tingyu, et al.
Published: (2025)
by: Song, Tingyu, et al.
Published: (2025)
Optimizing Legal Document Retrieval in Vietnamese with Semi-Hard Negative Mining
by: Le, Van-Hoang, et al.
Published: (2025)
by: Le, Van-Hoang, et al.
Published: (2025)
InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation
by: Xi, Yunjia, et al.
Published: (2025)
by: Xi, Yunjia, et al.
Published: (2025)
Know When to Fuse: Investigating Non-English Hybrid Retrieval in the Legal Domain
by: Louis, Antoine, et al.
Published: (2024)
by: Louis, Antoine, et al.
Published: (2024)
Learning Interpretable Legal Case Retrieval via Knowledge-Guided Case Reformulation
by: Deng, Chenlong, et al.
Published: (2024)
by: Deng, Chenlong, et al.
Published: (2024)
Benchmarking Information Retrieval Models on Complex Retrieval Tasks
by: Killingback, Julian, et al.
Published: (2025)
by: Killingback, Julian, et al.
Published: (2025)
STELLA: Self-Reflective Terminology-Aware Framework for Building an Aerospace Information Retrieval Benchmark
by: Kim, Bongmin
Published: (2026)
by: Kim, Bongmin
Published: (2026)
Anveshana: A New Benchmark Dataset for Cross-Lingual Information Retrieval On English Queries and Sanskrit Documents
by: Jagadeeshan, Manoj Balaji, et al.
Published: (2025)
by: Jagadeeshan, Manoj Balaji, et al.
Published: (2025)
How Vital is the Jurisprudential Relevance: Law Article Intervened Legal Case Retrieval and Matching
by: Xu, Nuo, et al.
Published: (2025)
by: Xu, Nuo, et al.
Published: (2025)
MIRB: Mathematical Information Retrieval Benchmark
by: Ju, Haocheng, et al.
Published: (2025)
by: Ju, Haocheng, et al.
Published: (2025)
RAR-b: Reasoning as Retrieval Benchmark
by: Xiao, Chenghao, et al.
Published: (2024)
by: Xiao, Chenghao, et al.
Published: (2024)
DAPR: A Benchmark on Document-Aware Passage Retrieval
by: Wang, Kexin, et al.
Published: (2023)
by: Wang, Kexin, et al.
Published: (2023)
BERGEN: A Benchmarking Library for Retrieval-Augmented Generation
by: Rau, David, et al.
Published: (2024)
by: Rau, David, et al.
Published: (2024)
LegalMALR:Multi-Agent Query Understanding and LLM-Based Reranking for Chinese Statute Retrieval
by: Li, Yunhan, et al.
Published: (2026)
by: Li, Yunhan, et al.
Published: (2026)
Check-Eval: A Checklist-based Approach for Evaluating Text Quality
by: Pereira, Jayr, et al.
Published: (2024)
by: Pereira, Jayr, et al.
Published: (2024)
FollowTable: A Benchmark for Instruction-Following Table Retrieval
by: Jin, Rihui, et al.
Published: (2026)
by: Jin, Rihui, et al.
Published: (2026)
PJB: A Reasoning-Aware Benchmark for Person-Job Retrieval
by: Wang, Guangzhi, et al.
Published: (2026)
by: Wang, Guangzhi, et al.
Published: (2026)
Hindi-BEIR : A Large Scale Retrieval Benchmark in Hindi
by: Acharya, Arkadeep, et al.
Published: (2024)
by: Acharya, Arkadeep, et al.
Published: (2024)
RAPID: Retrieval-Augmented Parallel Inference Drafting for Text-Based Video Event Retrieval
by: Nguyen, Long, et al.
Published: (2025)
by: Nguyen, Long, et al.
Published: (2025)
Similar Items
-
Domain-Adaptive Dense Retrieval for Brazilian Legal Search
by: Pereira, Jayr, et al.
Published: (2026) -
Quati: A Brazilian Portuguese Information Retrieval Dataset from Native Speakers
by: Bueno, Mirelle, et al.
Published: (2024) -
JurisTCU: A Brazilian Portuguese Information Retrieval Dataset with Query Relevance Judgments
by: Fernandes, Leandro Carísio, et al.
Published: (2025) -
Exploring Selective Retrieval-Augmentation for Long-Tail Legal Text Classification
by: Mao, Boheng
Published: (2025) -
ptt5-v2: A Closer Look at Continued Pretraining of T5 Models for the Portuguese Language
by: Piau, Marcos, et al.
Published: (2024)