BLT: Can Large Language Models Handle Basic Legal Text?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Blair-Stanek, Andrew, Holzenberger, Nils, Van Durme, Benjamin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can LLMs Identify Tax Abuse?
von: Blair-Stanek, Andrew, et al.
Veröffentlicht: (2025)
von: Blair-Stanek, Andrew, et al.
Veröffentlicht: (2025)
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
von: Blair-Stanek, Andrew, et al.
Veröffentlicht: (2023)
von: Blair-Stanek, Andrew, et al.
Veröffentlicht: (2023)
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
von: Hashemi, Helia, et al.
Veröffentlicht: (2024)
von: Hashemi, Helia, et al.
Veröffentlicht: (2024)
Multi-Hierarchical Feature Detection for Large Language Model Generated Text
von: Zhang, Luyan, et al.
Veröffentlicht: (2025)
von: Zhang, Luyan, et al.
Veröffentlicht: (2025)
LegalCheck: Retrieval- and Context-Augmented Generation for Drafting Municipal Legal Advice Letters
von: van der Meer, Virgill, et al.
Veröffentlicht: (2026)
von: van der Meer, Virgill, et al.
Veröffentlicht: (2026)
Towards Conditioning Clinical Text Generation for User Control
von: Koraş, Osman Alperen, et al.
Veröffentlicht: (2025)
von: Koraş, Osman Alperen, et al.
Veröffentlicht: (2025)
Meta-Evaluation of Translation Evaluation Methods: a systematic up-to-date overview
von: Han, Lifeng, et al.
Veröffentlicht: (2016)
von: Han, Lifeng, et al.
Veröffentlicht: (2016)
HalluScan: A Systematic Benchmark for Detecting and Mitigating Hallucinations in Instruction-Following LLMs
von: Cherif, Ahmed
Veröffentlicht: (2026)
von: Cherif, Ahmed
Veröffentlicht: (2026)
An Industrial-Scale Insurance LLM Achieving Verifiable Domain Mastery and Hallucination Control without Competence Trade-offs
von: Zhu, Qian, et al.
Veröffentlicht: (2026)
von: Zhu, Qian, et al.
Veröffentlicht: (2026)
Improving the Capabilities of Large Language Model Based Marketing Analytics Copilots With Semantic Search And Fine-Tuning
von: Gao, Yilin, et al.
Veröffentlicht: (2024)
von: Gao, Yilin, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models on Historical Health Crisis Knowledge in Resource-Limited Settings: A Hybrid Multi-Metric Study
von: Hasan, Mohammed Rakibul
Veröffentlicht: (2026)
von: Hasan, Mohammed Rakibul
Veröffentlicht: (2026)
The Need for Guardrails with Large Language Models in Medical Safety-Critical Settings: An Artificial Intelligence Application in the Pharmacovigilance Ecosystem
von: Hakim, Joe B, et al.
Veröffentlicht: (2024)
von: Hakim, Joe B, et al.
Veröffentlicht: (2024)
LLMs Simulate Big Five Personality Traits: Further Evidence
von: Sorokovikova, Aleksandra, et al.
Veröffentlicht: (2024)
von: Sorokovikova, Aleksandra, et al.
Veröffentlicht: (2024)
Scaling In, Not Up? Testing Thick Citation Context Analysis with GPT-5 and Fragile Prompts
von: Simons, Arno
Veröffentlicht: (2026)
von: Simons, Arno
Veröffentlicht: (2026)
OEMA: Ontology-Enhanced Multi-Agent Collaboration Framework for Zero-Shot Clinical Named Entity Recognition
von: Tao, Xinli, et al.
Veröffentlicht: (2025)
von: Tao, Xinli, et al.
Veröffentlicht: (2025)
Practical Design and Benchmarking of Generative AI Applications for Surgical Billing and Coding
von: Rollman, John C., et al.
Veröffentlicht: (2025)
von: Rollman, John C., et al.
Veröffentlicht: (2025)
Leveraging Large Language Models to Extract and Translate Medical Information in Doctors' Notes for Health Records and Diagnostic Billing Codes
von: Hartnett, Peter, et al.
Veröffentlicht: (2026)
von: Hartnett, Peter, et al.
Veröffentlicht: (2026)
Temporal Relation Extraction in Clinical Texts: A Span-based Graph Transformer Approach
von: Chaturvedi, Rochana, et al.
Veröffentlicht: (2025)
von: Chaturvedi, Rochana, et al.
Veröffentlicht: (2025)
PubMed Reasoner: Dynamic Reasoning-based Retrieval for Evidence-Grounded Biomedical Question Answering
von: Zhang, Yiqing, et al.
Veröffentlicht: (2026)
von: Zhang, Yiqing, et al.
Veröffentlicht: (2026)
LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization
von: Nguyen, Huyen, et al.
Veröffentlicht: (2026)
von: Nguyen, Huyen, et al.
Veröffentlicht: (2026)
Scaling Laws for State Dynamics in Large Language Models
von: Li, Jacob X, et al.
Veröffentlicht: (2025)
von: Li, Jacob X, et al.
Veröffentlicht: (2025)
Large Language Models for History, Philosophy, and Sociology of Science: Interpretive Uses, Methodological Challenges, and Critical Perspectives
von: Simons, Arno, et al.
Veröffentlicht: (2025)
von: Simons, Arno, et al.
Veröffentlicht: (2025)
Cognitive bias in LLM reasoning compromises interpretation of clinical oncology notes
von: Kenaston, Matthew W., et al.
Veröffentlicht: (2025)
von: Kenaston, Matthew W., et al.
Veröffentlicht: (2025)
AskSport: Web Application for Sports Question-Answering
von: Onofre, Enzo B, et al.
Veröffentlicht: (2025)
von: Onofre, Enzo B, et al.
Veröffentlicht: (2025)
How to Evaluate Medical AI
von: Kopanichuk, Ilia, et al.
Veröffentlicht: (2025)
von: Kopanichuk, Ilia, et al.
Veröffentlicht: (2025)
Comparing the Performance of LLMs in RAG-based Question-Answering: A Case Study in Computer Science Literature
von: Dayarathne, Ranul, et al.
Veröffentlicht: (2025)
von: Dayarathne, Ranul, et al.
Veröffentlicht: (2025)
Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs
von: Rodriguez, David, et al.
Veröffentlicht: (2025)
von: Rodriguez, David, et al.
Veröffentlicht: (2025)
ReTreVal: Reasoning Tree with Validation -- A Hybrid Framework for Enhanced LLM Multi-Step Reasoning
von: HS, Abhishek, et al.
Veröffentlicht: (2026)
von: HS, Abhishek, et al.
Veröffentlicht: (2026)
Challenges and Opportunities of NLP for HR Applications: A Discussion Paper
von: Leidner, Jochen L., et al.
Veröffentlicht: (2024)
von: Leidner, Jochen L., et al.
Veröffentlicht: (2024)
Comparative Analysis of AI Agent Architectures for Entity Relationship Classification
von: Berijanian, Maryam, et al.
Veröffentlicht: (2025)
von: Berijanian, Maryam, et al.
Veröffentlicht: (2025)
Teaching a Language Model to Speak the Language of Tools
von: Emanuilov, Simeon
Veröffentlicht: (2025)
von: Emanuilov, Simeon
Veröffentlicht: (2025)
A Llama walks into the 'Bar': Efficient Supervised Fine-Tuning for Legal Reasoning in the Multi-state Bar Exam
von: Fernandes, Rean, et al.
Veröffentlicht: (2025)
von: Fernandes, Rean, et al.
Veröffentlicht: (2025)
PerkwE_COQA: Enhanced Persian Conversational Question Answering by combining contextual keyword extraction with Large Language Models
von: Moradbeiki, Pardis, et al.
Veröffentlicht: (2024)
von: Moradbeiki, Pardis, et al.
Veröffentlicht: (2024)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
ChemPro: A Progressive Chemistry Benchmark for Large Language Models
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2026)
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2026)
Open-TI: Open Traffic Intelligence with Augmented Language Model
von: Da, Longchao, et al.
Veröffentlicht: (2023)
von: Da, Longchao, et al.
Veröffentlicht: (2023)
Exploiting LLM-as-a-Judge Disposition on Free Text Legal QA via Prompt Optimization
von: Elganayni, Mohamed Hesham, et al.
Veröffentlicht: (2026)
von: Elganayni, Mohamed Hesham, et al.
Veröffentlicht: (2026)
Robustness of Large Language Models to Perturbations in Text
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
From RAG to QA-RAG: Integrating Generative AI for Pharmaceutical Regulatory Compliance Process
von: Kim, Jaewoong, et al.
Veröffentlicht: (2024)
von: Kim, Jaewoong, et al.
Veröffentlicht: (2024)
AutoTRIZ: Automating Engineering Innovation with TRIZ and Large Language Models
von: Jiang, Shuo, et al.
Veröffentlicht: (2024)
von: Jiang, Shuo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Can LLMs Identify Tax Abuse?
von: Blair-Stanek, Andrew, et al.
Veröffentlicht: (2025) -
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
von: Blair-Stanek, Andrew, et al.
Veröffentlicht: (2023) -
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
von: Hashemi, Helia, et al.
Veröffentlicht: (2024) -
Multi-Hierarchical Feature Detection for Large Language Model Generated Text
von: Zhang, Luyan, et al.
Veröffentlicht: (2025) -
LegalCheck: Retrieval- and Context-Augmented Generation for Drafting Municipal Legal Advice Letters
von: van der Meer, Virgill, et al.
Veröffentlicht: (2026)