Fine-grained Claim-level RAG Benchmark for Law
Fuente:
arXiv
Saved in:
| Main Authors: | Das, Souvick, Abualhaija, Sallam, Bianculli, Domenico |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Scaling: Predicting Patent Approval with Domain-specific Fine-grained Claim Dependency Graph
by: Gao, Xiaochen Kev, et al.
Published: (2024)
by: Gao, Xiaochen Kev, et al.
Published: (2024)
Decoding Large-Language Models: A Systematic Overview of Socio-Technical Impacts, Constraints, and Emerging Questions
by: Kaya, Zeyneb N., et al.
Published: (2024)
by: Kaya, Zeyneb N., et al.
Published: (2024)
Entropic Claim Resolution: Uncertainty-Driven Evidence Selection for RAG
by: Di Gioia, Davide
Published: (2026)
by: Di Gioia, Davide
Published: (2026)
FT-RAG: A Fine-grained Retrieval-Augmented Generation Framework for Complex Table Reasoning
by: Guo, Zebin, et al.
Published: (2026)
by: Guo, Zebin, et al.
Published: (2026)
RPC-Bench: A Fine-grained Benchmark for Research Paper Comprehension
by: Chen, Yelin, et al.
Published: (2026)
by: Chen, Yelin, et al.
Published: (2026)
Legal Requirements Analysis
by: Abualhaija, Sallam, et al.
Published: (2023)
by: Abualhaija, Sallam, et al.
Published: (2023)
ClaimIQ at CheckThat! 2025: Comparing Prompted and Fine-Tuned Language Models for Verifying Numerical Claims
by: Anik, Anirban Saha, et al.
Published: (2025)
by: Anik, Anirban Saha, et al.
Published: (2025)
HalluVerse25: Fine-grained Multilingual Benchmark Dataset for LLM Hallucinations
by: Abdaljalil, Samir, et al.
Published: (2025)
by: Abdaljalil, Samir, et al.
Published: (2025)
A Claim Decomposition Benchmark for Long-form Answer Verification
by: Zhang, Zhihao, et al.
Published: (2024)
by: Zhang, Zhihao, et al.
Published: (2024)
Evaluating LLM-based Approaches to Legal Citation Prediction: Domain-specific Pre-training, Fine-tuning, or RAG? A Benchmark and an Australian Law Case Study
by: Han, Jiuzhou, et al.
Published: (2024)
by: Han, Jiuzhou, et al.
Published: (2024)
Benchmarking LLM Faithfulness in RAG with Evolving Leaderboards
by: Tamber, Manveer Singh, et al.
Published: (2025)
by: Tamber, Manveer Singh, et al.
Published: (2025)
FineSurE: Fine-grained Summarization Evaluation using LLMs
by: Song, Hwanjun, et al.
Published: (2024)
by: Song, Hwanjun, et al.
Published: (2024)
Assessment of RAG and Fine-Tuning for Industrial Question-Answering-Applications
by: Sturm, Jakob, et al.
Published: (2026)
by: Sturm, Jakob, et al.
Published: (2026)
FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding
by: Faure, Gueter Josmy, et al.
Published: (2026)
by: Faure, Gueter Josmy, et al.
Published: (2026)
Fine-tuning with RAG for Improving LLM Learning of New Skills
by: Ibrahim, Humaid, et al.
Published: (2025)
by: Ibrahim, Humaid, et al.
Published: (2025)
PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models
by: Song, Mingyang, et al.
Published: (2025)
by: Song, Mingyang, et al.
Published: (2025)
MIKE: A New Benchmark for Fine-grained Multimodal Entity Knowledge Editing
by: Li, Jiaqi, et al.
Published: (2024)
by: Li, Jiaqi, et al.
Published: (2024)
Generating Leakage-Free Benchmarks for Robust RAG Evaluation
by: Liu, Jiayi, et al.
Published: (2026)
by: Liu, Jiayi, et al.
Published: (2026)
Courtroom-Style Multi-Agent Debate with Progressive RAG and Role-Switching for Controversial Claim Verification
by: Chowdhury, Masnun Nuha, et al.
Published: (2026)
by: Chowdhury, Masnun Nuha, et al.
Published: (2026)
Fine-grained Narrative Classification in Biased News Articles
by: Afroz, Zeba, et al.
Published: (2025)
by: Afroz, Zeba, et al.
Published: (2025)
GenerationPrograms: Fine-grained Attribution with Executable Programs
by: Wan, David, et al.
Published: (2025)
by: Wan, David, et al.
Published: (2025)
Fine-tuning BERT with Bidirectional LSTM for Fine-grained Movie Reviews Sentiment Analysis
by: Nkhata, Gibson, et al.
Published: (2025)
by: Nkhata, Gibson, et al.
Published: (2025)
Can AI Validate Science? Benchmarking LLMs for Accurate Scientific Claim $\rightarrow$ Evidence Reasoning
by: Javaji, Shashidhar Reddy, et al.
Published: (2025)
by: Javaji, Shashidhar Reddy, et al.
Published: (2025)
GaRAGe: A Benchmark with Grounding Annotations for RAG Evaluation
by: Sorodoc, Ionut-Teodor, et al.
Published: (2025)
by: Sorodoc, Ionut-Teodor, et al.
Published: (2025)
FinS-Pilot: A Benchmark for Online Financial RAG System
by: Wang, Feng, et al.
Published: (2025)
by: Wang, Feng, et al.
Published: (2025)
UNH at CheckThat! 2025: Fine-tuning Vs Prompting in Claim Extraction
by: Wilder, Joe, et al.
Published: (2025)
by: Wilder, Joe, et al.
Published: (2025)
A System for Comprehensive Assessment of RAG Frameworks
by: Rengo, Mattia, et al.
Published: (2025)
by: Rengo, Mattia, et al.
Published: (2025)
To Memorize or to Retrieve: Scaling Laws for RAG-Considerate Pretraining
by: Singh, Karan, et al.
Published: (2026)
by: Singh, Karan, et al.
Published: (2026)
LMUnit: Fine-grained Evaluation with Natural Language Unit Tests
by: Saad-Falcon, Jon, et al.
Published: (2024)
by: Saad-Falcon, Jon, et al.
Published: (2024)
Repurposing Synthetic Data for Fine-grained Search Agent Supervision
by: Zhao, Yida, et al.
Published: (2025)
by: Zhao, Yida, et al.
Published: (2025)
QAPyramid: Fine-grained Evaluation of Content Selection for Text Summarization
by: Zhang, Shiyue, et al.
Published: (2024)
by: Zhang, Shiyue, et al.
Published: (2024)
ClaimGen-CN: A Large-scale Chinese Dataset for Legal Claim Generation
by: Zhou, Siying, et al.
Published: (2025)
by: Zhou, Siying, et al.
Published: (2025)
A New Pipeline For Generating Instruction Dataset via RAG and Self Fine-Tuning
by: Song, Chih-Wei, et al.
Published: (2024)
by: Song, Chih-Wei, et al.
Published: (2024)
Should We Fine-Tune or RAG? Evaluating Different Techniques to Adapt LLMs for Dialogue
by: Alghisi, Simone, et al.
Published: (2024)
by: Alghisi, Simone, et al.
Published: (2024)
FLRC: Fine-grained Low-Rank Compressor for Efficient LLM Inference
by: Lu, Yu-Chen, et al.
Published: (2025)
by: Lu, Yu-Chen, et al.
Published: (2025)
RAGChecker: A Fine-grained Framework for Diagnosing Retrieval-Augmented Generation
by: Ru, Dongyu, et al.
Published: (2024)
by: Ru, Dongyu, et al.
Published: (2024)
FG-RAG: Enhancing Query-Focused Summarization with Context-Aware Fine-Grained Graph RAG
by: Hong, Yubin, et al.
Published: (2025)
by: Hong, Yubin, et al.
Published: (2025)
ClaimBrush: A Novel Framework for Automated Patent Claim Refinement Based on Large Language Models
by: Kawano, Seiya, et al.
Published: (2024)
by: Kawano, Seiya, et al.
Published: (2024)
Fine-Tuning MedGemma for Clinical Captioning to Enhance Multimodal RAG over Malaysia CPGs
by: Zun, Lee Qi, et al.
Published: (2025)
by: Zun, Lee Qi, et al.
Published: (2025)
HiFACTMix: A Code-Mixed Benchmark and Graph-Aware Model for EvidenceBased Political Claim Verification in Hinglish
by: Thakur, Rakesh, et al.
Published: (2025)
by: Thakur, Rakesh, et al.
Published: (2025)
Similar Items
-
Beyond Scaling: Predicting Patent Approval with Domain-specific Fine-grained Claim Dependency Graph
by: Gao, Xiaochen Kev, et al.
Published: (2024) -
Decoding Large-Language Models: A Systematic Overview of Socio-Technical Impacts, Constraints, and Emerging Questions
by: Kaya, Zeyneb N., et al.
Published: (2024) -
Entropic Claim Resolution: Uncertainty-Driven Evidence Selection for RAG
by: Di Gioia, Davide
Published: (2026) -
FT-RAG: A Fine-grained Retrieval-Augmented Generation Framework for Complex Table Reasoning
by: Guo, Zebin, et al.
Published: (2026) -
RPC-Bench: A Fine-grained Benchmark for Research Paper Comprehension
by: Chen, Yelin, et al.
Published: (2026)