The Judge Variable: Challenging Judge-Agnostic Legal Judgment Prediction
Fuente:
arXiv
Saved in:
| Main Author: | Zambrano, Guillaume |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploiting LLM-as-a-Judge Disposition on Free Text Legal QA via Prompt Optimization
by: Elganayni, Mohamed Hesham, et al.
Published: (2026)
by: Elganayni, Mohamed Hesham, et al.
Published: (2026)
Practical Design and Benchmarking of Generative AI Applications for Surgical Billing and Coding
by: Rollman, John C., et al.
Published: (2025)
by: Rollman, John C., et al.
Published: (2025)
BLT: Can Large Language Models Handle Basic Legal Text?
by: Blair-Stanek, Andrew, et al.
Published: (2023)
by: Blair-Stanek, Andrew, et al.
Published: (2023)
Continuous Predictive Modeling of Clinical Notes and ICD Codes in Patient Health Records
by: Caralt, Mireia Hernandez, et al.
Published: (2024)
by: Caralt, Mireia Hernandez, et al.
Published: (2024)
Natural Language Processing for the Legal Domain: A Survey of Tasks, Datasets, Models, and Challenges
by: Ariai, Farid, et al.
Published: (2024)
by: Ariai, Farid, et al.
Published: (2024)
Legal RAG Bench: an end-to-end benchmark for legal RAG
by: Butler, Abdur-Rahman, et al.
Published: (2026)
by: Butler, Abdur-Rahman, et al.
Published: (2026)
Temporal Relation Extraction in Clinical Texts: A Span-based Graph Transformer Approach
by: Chaturvedi, Rochana, et al.
Published: (2025)
by: Chaturvedi, Rochana, et al.
Published: (2025)
Counterfactual Causal Inference in Natural Language with Large Language Models
by: Gendron, Gaël, et al.
Published: (2024)
by: Gendron, Gaël, et al.
Published: (2024)
An Unsupervised Natural Language Processing Pipeline for Assessing Referral Appropriateness
by: Torri, Vittorio, et al.
Published: (2025)
by: Torri, Vittorio, et al.
Published: (2025)
Comparative Analysis of LoRA-Adapted Embedding Models for Clinical Cardiology Text Representation
by: Young, Richard J., et al.
Published: (2025)
by: Young, Richard J., et al.
Published: (2025)
Reasoning Over the Glyphs: Evaluation of LLM's Decipherment of Rare Scripts
by: Shih, Yu-Fei, et al.
Published: (2025)
by: Shih, Yu-Fei, et al.
Published: (2025)
Disease Entity Recognition and Normalization is Improved with Large Language Model Derived Synthetic Normalized Mentions
by: Sasse, Kuleen, et al.
Published: (2024)
by: Sasse, Kuleen, et al.
Published: (2024)
Towards Leveraging Large Language Models for Automated Medical Q&A Evaluation
by: Krolik, Jack, et al.
Published: (2024)
by: Krolik, Jack, et al.
Published: (2024)
A Llama walks into the 'Bar': Efficient Supervised Fine-Tuning for Legal Reasoning in the Multi-state Bar Exam
by: Fernandes, Rean, et al.
Published: (2025)
by: Fernandes, Rean, et al.
Published: (2025)
Prompting from the bench: Large-scale pretraining is not sufficient to prepare LLMs for ordinary meaning analysis
by: Purushothama, Abhishek, et al.
Published: (2025)
by: Purushothama, Abhishek, et al.
Published: (2025)
Neither Valid nor Reliable? Investigating the Use of LLMs as Judges
by: Chehbouni, Khaoula, et al.
Published: (2025)
by: Chehbouni, Khaoula, et al.
Published: (2025)
RPRA: Predicting an LLM-Judge for Efficient but Performant Inference
by: Ashley, Dylan R., et al.
Published: (2026)
by: Ashley, Dylan R., et al.
Published: (2026)
Repetition Without Exclusivity: Scale Sensitivity of Referential Mechanisms in Child-Scale Language Models
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Improving the Capabilities of Large Language Model Based Marketing Analytics Copilots With Semantic Search And Fine-Tuning
by: Gao, Yilin, et al.
Published: (2024)
by: Gao, Yilin, et al.
Published: (2024)
Temporal Concept Drift in Legal Judgment Prediction: Neural Baselines Across Three Epochs of Ukrainian Court Decisions
by: Ovcharov, Volodymyr
Published: (2026)
by: Ovcharov, Volodymyr
Published: (2026)
AMEL: Accumulated Message Effects on LLM Judgments
by: Temkit, Sid-Ali
Published: (2026)
by: Temkit, Sid-Ali
Published: (2026)
LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
by: Pradhan, Anu, et al.
Published: (2025)
by: Pradhan, Anu, et al.
Published: (2025)
Suicide Risk Assessment Using Multimodal Speech Features: A Study on the SW1 Challenge Dataset
by: Marie, Ambre, et al.
Published: (2025)
by: Marie, Ambre, et al.
Published: (2025)
GenAI Content Detection Task 3: Cross-Domain Machine-Generated Text Detection Challenge
by: Dugan, Liam, et al.
Published: (2025)
by: Dugan, Liam, et al.
Published: (2025)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
by: Oketunji, Abiodun Finbarrs
Published: (2023)
by: Oketunji, Abiodun Finbarrs
Published: (2023)
A-VERT: Agnostic Verification with Embedding Ranking Targets
by: Aguirre, Nicolás, et al.
Published: (2025)
by: Aguirre, Nicolás, et al.
Published: (2025)
Language Models and Retrieval Augmented Generation for Automated Structured Data Extraction from Diagnostic Reports
by: Jabal, Mohamed Sobhi, et al.
Published: (2024)
by: Jabal, Mohamed Sobhi, et al.
Published: (2024)
Lon-ea at SemEval-2023 Task 11: A Comparison of Activation Functions for Soft and Hard Label Prediction
by: Hosseini, Peyman, et al.
Published: (2023)
by: Hosseini, Peyman, et al.
Published: (2023)
Future Token Prediction -- Causal Language Modelling with Per-Token Semantic State Vector for Multi-Token Prediction
by: Walker, Nicholas
Published: (2024)
by: Walker, Nicholas
Published: (2024)
Opinion Mining on Offshore Wind Energy for Environmental Engineering
by: Bittencourt, Isabele, et al.
Published: (2024)
by: Bittencourt, Isabele, et al.
Published: (2024)
LegalCheck: Retrieval- and Context-Augmented Generation for Drafting Municipal Legal Advice Letters
by: van der Meer, Virgill, et al.
Published: (2026)
by: van der Meer, Virgill, et al.
Published: (2026)
Named Entity Recognition and Classification on Historical Documents: A Survey
by: Ehrmann, Maud, et al.
Published: (2021)
by: Ehrmann, Maud, et al.
Published: (2021)
The Foundational Capabilities of Large Language Models in Predicting Postoperative Risks Using Clinical Notes
by: Alba, Charles, et al.
Published: (2024)
by: Alba, Charles, et al.
Published: (2024)
The Need for Guardrails with Large Language Models in Medical Safety-Critical Settings: An Artificial Intelligence Application in the Pharmacovigilance Ecosystem
by: Hakim, Joe B, et al.
Published: (2024)
by: Hakim, Joe B, et al.
Published: (2024)
Interpretable phenotyping of Heart Failure patients with Dutch discharge letters
by: Torri, Vittorio, et al.
Published: (2025)
by: Torri, Vittorio, et al.
Published: (2025)
Advancing Italian Biomedical Information Extraction with Transformers-based Models: Methodological Insights and Multicenter Practical Application
by: Crema, Claudio, et al.
Published: (2023)
by: Crema, Claudio, et al.
Published: (2023)
Transparency-First Medical Language Models: Datasheets, Model Cards, and End-to-End Data Provenance for Clinical NLP
by: Imanov, Olaf Yunus Laitinen, et al.
Published: (2026)
by: Imanov, Olaf Yunus Laitinen, et al.
Published: (2026)
DAEDRA: A language model for predicting outcomes in passive pharmacovigilance reporting
by: von Csefalvay, Chris
Published: (2024)
by: von Csefalvay, Chris
Published: (2024)
Annotation Entropy Predicts Per-Example Learning Dynamics in LoRA Fine-Tuning
by: Steele, Brady
Published: (2026)
by: Steele, Brady
Published: (2026)
Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility
by: Szilvasy, Gergely, et al.
Published: (2026)
by: Szilvasy, Gergely, et al.
Published: (2026)
Similar Items
-
Exploiting LLM-as-a-Judge Disposition on Free Text Legal QA via Prompt Optimization
by: Elganayni, Mohamed Hesham, et al.
Published: (2026) -
Practical Design and Benchmarking of Generative AI Applications for Surgical Billing and Coding
by: Rollman, John C., et al.
Published: (2025) -
BLT: Can Large Language Models Handle Basic Legal Text?
by: Blair-Stanek, Andrew, et al.
Published: (2023) -
Continuous Predictive Modeling of Clinical Notes and ICD Codes in Patient Health Records
by: Caralt, Mireia Hernandez, et al.
Published: (2024) -
Natural Language Processing for the Legal Domain: A Survey of Tasks, Datasets, Models, and Challenges
by: Ariai, Farid, et al.
Published: (2024)