GreekBarBench: A Challenging Benchmark for Free-Text Legal Reasoning and Citations
Fuente:
arXiv
Saved in:
| Main Authors: | Chlapanis, Odysseas S., Galanis, Dimitrios, Aletras, Nikolaos, Androutsopoulos, Ion |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LAR-ECHR: A New Legal Argument Reasoning Task and Dataset for Cases of the European Court of Human Rights
by: Chlapanis, Odysseas S., et al.
Published: (2024)
by: Chlapanis, Odysseas S., et al.
Published: (2024)
Archimedes-AUEB at SemEval-2024 Task 5: LLM explains Civil Procedure
by: Chlapanis, Odysseas S., et al.
Published: (2024)
by: Chlapanis, Odysseas S., et al.
Published: (2024)
AUEB-Archimedes at RIRAG-2025: Is obligation concatenation really all you need?
by: Chasandras, Ioannis, et al.
Published: (2024)
by: Chasandras, Ioannis, et al.
Published: (2024)
Local Explanations and Self-Explanations for Assessing Faithfulness in black-box LLMs
by: Fragkathoulas, Christos, et al.
Published: (2024)
by: Fragkathoulas, Christos, et al.
Published: (2024)
The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans
by: Chlapanis, Odysseas S., et al.
Published: (2026)
by: Chlapanis, Odysseas S., et al.
Published: (2026)
Vocabulary-level Memory Efficiency for Language Model Fine-tuning
by: Williams, Miles, et al.
Published: (2023)
by: Williams, Miles, et al.
Published: (2023)
On the Impact of Calibration Data in Post-training Quantization and Pruning
by: Williams, Miles, et al.
Published: (2023)
by: Williams, Miles, et al.
Published: (2023)
Comparing Data Augmentation Methods for End-to-End Task-Oriented Dialog Systems
by: Vlachos, Christos, et al.
Published: (2024)
by: Vlachos, Christos, et al.
Published: (2024)
MCiteBench: A Multimodal Benchmark for Generating Text with Citations
by: Hu, Caiyu, et al.
Published: (2025)
by: Hu, Caiyu, et al.
Published: (2025)
How Can We Effectively Expand the Vocabulary of LLMs with 0.01GB of Target Language Text?
by: Yamaguchi, Atsuki, et al.
Published: (2024)
by: Yamaguchi, Atsuki, et al.
Published: (2024)
GR-NLP-TOOLKIT: An Open-Source NLP Toolkit for Modern Greek
by: Loukas, Lefteris, et al.
Published: (2024)
by: Loukas, Lefteris, et al.
Published: (2024)
Comparing Explanation Faithfulness between Multilingual and Monolingual Fine-tuned Language Models
by: Zhao, Zhixue, et al.
Published: (2024)
by: Zhao, Zhixue, et al.
Published: (2024)
Designing Synthetic Discussion Generation Systems: A Case Study for Online Facilitation
by: Tsirmpas, Dimitris, et al.
Published: (2025)
by: Tsirmpas, Dimitris, et al.
Published: (2025)
Should I try multiple optimizers when fine-tuning pre-trained Transformers for NLP tasks? Should I tune their hyperparameters?
by: Gkouti, Nefeli, et al.
Published: (2024)
by: Gkouti, Nefeli, et al.
Published: (2024)
LegalCiteBench: Evaluating Citation Reliability in Legal Language Models
by: Chen, Sijia, et al.
Published: (2026)
by: Chen, Sijia, et al.
Published: (2026)
Incorporating Attribution Importance for Improving Faithfulness Metrics
by: Zhao, Zhixue, et al.
Published: (2023)
by: Zhao, Zhixue, et al.
Published: (2023)
Building Open-Retrieval Conversational Question Answering Systems by Generating Synthetic Data and Decontextualizing User Questions
by: Vlachos, Christos, et al.
Published: (2025)
by: Vlachos, Christos, et al.
Published: (2025)
Fine-Tuning on Noisy Instructions: Effects on Generalization and Performance
by: Alajrami, Ahmed, et al.
Published: (2025)
by: Alajrami, Ahmed, et al.
Published: (2025)
Progressive Depth Up-scaling via Optimal Transport
by: Cao, Mingzi, et al.
Published: (2025)
by: Cao, Mingzi, et al.
Published: (2025)
Self-calibration for Language Model Quantization and Pruning
by: Williams, Miles, et al.
Published: (2024)
by: Williams, Miles, et al.
Published: (2024)
Enhancing Linguistic Competence of Language Models through Pre-training with Language Learning Tasks
by: Yamaguchi, Atsuki, et al.
Published: (2026)
by: Yamaguchi, Atsuki, et al.
Published: (2026)
How Private are Language Models in Abstractive Summarization?
by: Hughes, Anthony, et al.
Published: (2024)
by: Hughes, Anthony, et al.
Published: (2024)
Enhancing Logical Reasoning in Language Models via Symbolically-Guided Monte Carlo Process Supervision
by: Tan, Xingwei, et al.
Published: (2025)
by: Tan, Xingwei, et al.
Published: (2025)
Improving Multimodal Classification of Social Media Posts by Leveraging Image-Text Auxiliary Tasks
by: Villegas, Danae Sánchez, et al.
Published: (2023)
by: Villegas, Danae Sánchez, et al.
Published: (2023)
A Data-Driven Guided Decoding Mechanism for Diagnostic Captioning
by: Kaliosis, Panagiotis, et al.
Published: (2024)
by: Kaliosis, Panagiotis, et al.
Published: (2024)
An Empirical Study on Cross-lingual Vocabulary Adaptation for Efficient Language Model Inference
by: Yamaguchi, Atsuki, et al.
Published: (2024)
by: Yamaguchi, Atsuki, et al.
Published: (2024)
MASLegalBench: Benchmarking Multi-Agent Systems in Deductive Legal Reasoning
by: Jing, Huihao, et al.
Published: (2025)
by: Jing, Huihao, et al.
Published: (2025)
LegalBench.PT: A Benchmark for Portuguese Law
by: Canaverde, Beatriz, et al.
Published: (2025)
by: Canaverde, Beatriz, et al.
Published: (2025)
Compressing Language Models for Specialized Domains
by: Williams, Miles, et al.
Published: (2025)
by: Williams, Miles, et al.
Published: (2025)
Examining the Limitations of Computational Rumor Detection Models Trained on Static Datasets
by: Mu, Yida, et al.
Published: (2023)
by: Mu, Yida, et al.
Published: (2023)
Enhancing Data Quality through Simple De-duplication: Navigating Responsible Computational Social Science Research
by: Mu, Yida, et al.
Published: (2024)
by: Mu, Yida, et al.
Published: (2024)
Where does output diversity collapse in post-training?
by: Karouzos, Constantinos, et al.
Published: (2026)
by: Karouzos, Constantinos, et al.
Published: (2026)
An Empirical Study on Preference Tuning Generalization and Diversity Under Domain Shift
by: Karouzos, Constantinos, et al.
Published: (2026)
by: Karouzos, Constantinos, et al.
Published: (2026)
Deconstructing Attention: Investigating Design Principles for Effective Language Modeling
by: Xue, Huiyin, et al.
Published: (2025)
by: Xue, Huiyin, et al.
Published: (2025)
FarFetched: Entity-centric Reasoning and Claim Validation for the Greek Language based on Textually Represented Environments
by: Papadopoulos, Dimitris, et al.
Published: (2024)
by: Papadopoulos, Dimitris, et al.
Published: (2024)
Maistros: A Greek Large Language Model Adapted Through Knowledge Distillation From Large Reasoning Models
by: Giarelis, Nikolaos, et al.
Published: (2026)
by: Giarelis, Nikolaos, et al.
Published: (2026)
PILOT-Bench: A Benchmark for Legal Reasoning in the Patent Domain with IRAC-Aligned Classification Tasks
by: Jang, Yehoon, et al.
Published: (2026)
by: Jang, Yehoon, et al.
Published: (2026)
VLegal-Bench: Cognitively Grounded Benchmark for Vietnamese Legal Reasoning of Large Language Models
by: Dong, Nguyen Tien, et al.
Published: (2025)
by: Dong, Nguyen Tien, et al.
Published: (2025)
Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models
by: Villegas, Danae Sánchez, et al.
Published: (2026)
by: Villegas, Danae Sánchez, et al.
Published: (2026)
RenoBench: A Citation Parsing Benchmark
by: Sarin, Parth, et al.
Published: (2026)
by: Sarin, Parth, et al.
Published: (2026)
Similar Items
-
LAR-ECHR: A New Legal Argument Reasoning Task and Dataset for Cases of the European Court of Human Rights
by: Chlapanis, Odysseas S., et al.
Published: (2024) -
Archimedes-AUEB at SemEval-2024 Task 5: LLM explains Civil Procedure
by: Chlapanis, Odysseas S., et al.
Published: (2024) -
AUEB-Archimedes at RIRAG-2025: Is obligation concatenation really all you need?
by: Chasandras, Ioannis, et al.
Published: (2024) -
Local Explanations and Self-Explanations for Assessing Faithfulness in black-box LLMs
by: Fragkathoulas, Christos, et al.
Published: (2024) -
The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans
by: Chlapanis, Odysseas S., et al.
Published: (2026)