The Judge Variable: Challenging Judge-Agnostic Legal Judgment Prediction
Fuente:
arXiv
Salvato in:
| Autore principale: | Zambrano, Guillaume |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Exploiting LLM-as-a-Judge Disposition on Free Text Legal QA via Prompt Optimization
di: Elganayni, Mohamed Hesham, et al.
Pubblicazione: (2026)
di: Elganayni, Mohamed Hesham, et al.
Pubblicazione: (2026)
Practical Design and Benchmarking of Generative AI Applications for Surgical Billing and Coding
di: Rollman, John C., et al.
Pubblicazione: (2025)
di: Rollman, John C., et al.
Pubblicazione: (2025)
BLT: Can Large Language Models Handle Basic Legal Text?
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2023)
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2023)
Continuous Predictive Modeling of Clinical Notes and ICD Codes in Patient Health Records
di: Caralt, Mireia Hernandez, et al.
Pubblicazione: (2024)
di: Caralt, Mireia Hernandez, et al.
Pubblicazione: (2024)
Natural Language Processing for the Legal Domain: A Survey of Tasks, Datasets, Models, and Challenges
di: Ariai, Farid, et al.
Pubblicazione: (2024)
di: Ariai, Farid, et al.
Pubblicazione: (2024)
Legal RAG Bench: an end-to-end benchmark for legal RAG
di: Butler, Abdur-Rahman, et al.
Pubblicazione: (2026)
di: Butler, Abdur-Rahman, et al.
Pubblicazione: (2026)
Temporal Relation Extraction in Clinical Texts: A Span-based Graph Transformer Approach
di: Chaturvedi, Rochana, et al.
Pubblicazione: (2025)
di: Chaturvedi, Rochana, et al.
Pubblicazione: (2025)
Counterfactual Causal Inference in Natural Language with Large Language Models
di: Gendron, Gaël, et al.
Pubblicazione: (2024)
di: Gendron, Gaël, et al.
Pubblicazione: (2024)
An Unsupervised Natural Language Processing Pipeline for Assessing Referral Appropriateness
di: Torri, Vittorio, et al.
Pubblicazione: (2025)
di: Torri, Vittorio, et al.
Pubblicazione: (2025)
Comparative Analysis of LoRA-Adapted Embedding Models for Clinical Cardiology Text Representation
di: Young, Richard J., et al.
Pubblicazione: (2025)
di: Young, Richard J., et al.
Pubblicazione: (2025)
Reasoning Over the Glyphs: Evaluation of LLM's Decipherment of Rare Scripts
di: Shih, Yu-Fei, et al.
Pubblicazione: (2025)
di: Shih, Yu-Fei, et al.
Pubblicazione: (2025)
Disease Entity Recognition and Normalization is Improved with Large Language Model Derived Synthetic Normalized Mentions
di: Sasse, Kuleen, et al.
Pubblicazione: (2024)
di: Sasse, Kuleen, et al.
Pubblicazione: (2024)
Towards Leveraging Large Language Models for Automated Medical Q&A Evaluation
di: Krolik, Jack, et al.
Pubblicazione: (2024)
di: Krolik, Jack, et al.
Pubblicazione: (2024)
A Llama walks into the 'Bar': Efficient Supervised Fine-Tuning for Legal Reasoning in the Multi-state Bar Exam
di: Fernandes, Rean, et al.
Pubblicazione: (2025)
di: Fernandes, Rean, et al.
Pubblicazione: (2025)
Prompting from the bench: Large-scale pretraining is not sufficient to prepare LLMs for ordinary meaning analysis
di: Purushothama, Abhishek, et al.
Pubblicazione: (2025)
di: Purushothama, Abhishek, et al.
Pubblicazione: (2025)
Neither Valid nor Reliable? Investigating the Use of LLMs as Judges
di: Chehbouni, Khaoula, et al.
Pubblicazione: (2025)
di: Chehbouni, Khaoula, et al.
Pubblicazione: (2025)
RPRA: Predicting an LLM-Judge for Efficient but Performant Inference
di: Ashley, Dylan R., et al.
Pubblicazione: (2026)
di: Ashley, Dylan R., et al.
Pubblicazione: (2026)
Repetition Without Exclusivity: Scale Sensitivity of Referential Mechanisms in Child-Scale Language Models
di: Cacioli, Jon-Paul
Pubblicazione: (2026)
di: Cacioli, Jon-Paul
Pubblicazione: (2026)
Improving the Capabilities of Large Language Model Based Marketing Analytics Copilots With Semantic Search And Fine-Tuning
di: Gao, Yilin, et al.
Pubblicazione: (2024)
di: Gao, Yilin, et al.
Pubblicazione: (2024)
Temporal Concept Drift in Legal Judgment Prediction: Neural Baselines Across Three Epochs of Ukrainian Court Decisions
di: Ovcharov, Volodymyr
Pubblicazione: (2026)
di: Ovcharov, Volodymyr
Pubblicazione: (2026)
AMEL: Accumulated Message Effects on LLM Judgments
di: Temkit, Sid-Ali
Pubblicazione: (2026)
di: Temkit, Sid-Ali
Pubblicazione: (2026)
LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
di: Pradhan, Anu, et al.
Pubblicazione: (2025)
di: Pradhan, Anu, et al.
Pubblicazione: (2025)
Suicide Risk Assessment Using Multimodal Speech Features: A Study on the SW1 Challenge Dataset
di: Marie, Ambre, et al.
Pubblicazione: (2025)
di: Marie, Ambre, et al.
Pubblicazione: (2025)
GenAI Content Detection Task 3: Cross-Domain Machine-Generated Text Detection Challenge
di: Dugan, Liam, et al.
Pubblicazione: (2025)
di: Dugan, Liam, et al.
Pubblicazione: (2025)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
A-VERT: Agnostic Verification with Embedding Ranking Targets
di: Aguirre, Nicolás, et al.
Pubblicazione: (2025)
di: Aguirre, Nicolás, et al.
Pubblicazione: (2025)
Language Models and Retrieval Augmented Generation for Automated Structured Data Extraction from Diagnostic Reports
di: Jabal, Mohamed Sobhi, et al.
Pubblicazione: (2024)
di: Jabal, Mohamed Sobhi, et al.
Pubblicazione: (2024)
Lon-ea at SemEval-2023 Task 11: A Comparison of Activation Functions for Soft and Hard Label Prediction
di: Hosseini, Peyman, et al.
Pubblicazione: (2023)
di: Hosseini, Peyman, et al.
Pubblicazione: (2023)
Future Token Prediction -- Causal Language Modelling with Per-Token Semantic State Vector for Multi-Token Prediction
di: Walker, Nicholas
Pubblicazione: (2024)
di: Walker, Nicholas
Pubblicazione: (2024)
Opinion Mining on Offshore Wind Energy for Environmental Engineering
di: Bittencourt, Isabele, et al.
Pubblicazione: (2024)
di: Bittencourt, Isabele, et al.
Pubblicazione: (2024)
LegalCheck: Retrieval- and Context-Augmented Generation for Drafting Municipal Legal Advice Letters
di: van der Meer, Virgill, et al.
Pubblicazione: (2026)
di: van der Meer, Virgill, et al.
Pubblicazione: (2026)
Named Entity Recognition and Classification on Historical Documents: A Survey
di: Ehrmann, Maud, et al.
Pubblicazione: (2021)
di: Ehrmann, Maud, et al.
Pubblicazione: (2021)
The Foundational Capabilities of Large Language Models in Predicting Postoperative Risks Using Clinical Notes
di: Alba, Charles, et al.
Pubblicazione: (2024)
di: Alba, Charles, et al.
Pubblicazione: (2024)
The Need for Guardrails with Large Language Models in Medical Safety-Critical Settings: An Artificial Intelligence Application in the Pharmacovigilance Ecosystem
di: Hakim, Joe B, et al.
Pubblicazione: (2024)
di: Hakim, Joe B, et al.
Pubblicazione: (2024)
Interpretable phenotyping of Heart Failure patients with Dutch discharge letters
di: Torri, Vittorio, et al.
Pubblicazione: (2025)
di: Torri, Vittorio, et al.
Pubblicazione: (2025)
Advancing Italian Biomedical Information Extraction with Transformers-based Models: Methodological Insights and Multicenter Practical Application
di: Crema, Claudio, et al.
Pubblicazione: (2023)
di: Crema, Claudio, et al.
Pubblicazione: (2023)
Transparency-First Medical Language Models: Datasheets, Model Cards, and End-to-End Data Provenance for Clinical NLP
di: Imanov, Olaf Yunus Laitinen, et al.
Pubblicazione: (2026)
di: Imanov, Olaf Yunus Laitinen, et al.
Pubblicazione: (2026)
DAEDRA: A language model for predicting outcomes in passive pharmacovigilance reporting
di: von Csefalvay, Chris
Pubblicazione: (2024)
di: von Csefalvay, Chris
Pubblicazione: (2024)
Annotation Entropy Predicts Per-Example Learning Dynamics in LoRA Fine-Tuning
di: Steele, Brady
Pubblicazione: (2026)
di: Steele, Brady
Pubblicazione: (2026)
Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility
di: Szilvasy, Gergely, et al.
Pubblicazione: (2026)
di: Szilvasy, Gergely, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Exploiting LLM-as-a-Judge Disposition on Free Text Legal QA via Prompt Optimization
di: Elganayni, Mohamed Hesham, et al.
Pubblicazione: (2026) -
Practical Design and Benchmarking of Generative AI Applications for Surgical Billing and Coding
di: Rollman, John C., et al.
Pubblicazione: (2025) -
BLT: Can Large Language Models Handle Basic Legal Text?
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2023) -
Continuous Predictive Modeling of Clinical Notes and ICD Codes in Patient Health Records
di: Caralt, Mireia Hernandez, et al.
Pubblicazione: (2024) -
Natural Language Processing for the Legal Domain: A Survey of Tasks, Datasets, Models, and Challenges
di: Ariai, Farid, et al.
Pubblicazione: (2024)