Saved in:
| Main Authors: | Buonocore, Tommaso Mario, Rancati, Simone, Parimbelli, Enea |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2407.06011 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reshaping Free-Text Radiology Notes Into Structured Reports With Generative Transformers
by: Bergomi, Laura, et al.
Published: (2024)
by: Bergomi, Laura, et al.
Published: (2024)
Advancing Italian Biomedical Information Extraction with Transformers-based Models: Methodological Insights and Multicenter Practical Application
by: Crema, Claudio, et al.
Published: (2023)
by: Crema, Claudio, et al.
Published: (2023)
RAR: Setting Knowledge Tripwires for Retrieval Augmented Rejection
by: Buonocore, Tommaso Mario, et al.
Published: (2025)
by: Buonocore, Tommaso Mario, et al.
Published: (2025)
BALI: Enhancing Biomedical Language Representations through Knowledge Graph and Language Model Alignment
by: Sakhovskiy, Andrey, et al.
Published: (2025)
by: Sakhovskiy, Andrey, et al.
Published: (2025)
BLT: Can Large Language Models Handle Basic Legal Text?
by: Blair-Stanek, Andrew, et al.
Published: (2023)
by: Blair-Stanek, Andrew, et al.
Published: (2023)
A Method for the Architecture of a Medical Vertical Large Language Model Based on Deepseek R1
by: Zhang, Mingda, et al.
Published: (2025)
by: Zhang, Mingda, et al.
Published: (2025)
Evaluating Large Language Models for IUCN Red List Species Information
by: Uryu, Shinya
Published: (2025)
by: Uryu, Shinya
Published: (2025)
ELMTEX: Fine-Tuning Large Language Models for Structured Clinical Information Extraction. A Case Study on Clinical Reports
by: Guluzade, Aynur, et al.
Published: (2025)
by: Guluzade, Aynur, et al.
Published: (2025)
AcuityBench: Evaluating Clinical Acuity Identification and Uncertainty Alignment
by: Linzmayer, Robin, et al.
Published: (2026)
by: Linzmayer, Robin, et al.
Published: (2026)
CuraView: A Multi-Agent Framework for Medical Hallucination Detection with GraphRAG-Enhanced Knowledge Verification
by: Ye, Severin, et al.
Published: (2026)
by: Ye, Severin, et al.
Published: (2026)
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters
by: Shah, Aaryan, et al.
Published: (2026)
by: Shah, Aaryan, et al.
Published: (2026)
On Fusing ChatGPT and Ensemble Learning in Discon-tinuous Named Entity Recognition in Health Corpora
by: Chen, Tzu-Chieh, et al.
Published: (2024)
by: Chen, Tzu-Chieh, et al.
Published: (2024)
MedPI: Evaluating AI Systems in Medical Patient-facing Interactions
by: V., Diego Fajardo, et al.
Published: (2025)
by: V., Diego Fajardo, et al.
Published: (2025)
Evaluating the Challenges of LLMs in Real-world Medical Follow-up: A Comparative Study and An Optimized Framework
by: Liu, Jinyan, et al.
Published: (2025)
by: Liu, Jinyan, et al.
Published: (2025)
Expertise Is What We Want
by: Ashworth, Alan, et al.
Published: (2025)
by: Ashworth, Alan, et al.
Published: (2025)
Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes
by: Liu, Ming
Published: (2026)
by: Liu, Ming
Published: (2026)
Exploiting LLM-as-a-Judge Disposition on Free Text Legal QA via Prompt Optimization
by: Elganayni, Mohamed Hesham, et al.
Published: (2026)
by: Elganayni, Mohamed Hesham, et al.
Published: (2026)
AI-MASLD Metabolic Dysfunction and Information Steatosis of Large Language Models in Unstructured Clinical Narratives
by: Shen, Yuan, et al.
Published: (2025)
by: Shen, Yuan, et al.
Published: (2025)
Temporal Relation Extraction in Clinical Texts: A Span-based Graph Transformer Approach
by: Chaturvedi, Rochana, et al.
Published: (2025)
by: Chaturvedi, Rochana, et al.
Published: (2025)
OEMA: Ontology-Enhanced Multi-Agent Collaboration Framework for Zero-Shot Clinical Named Entity Recognition
by: Tao, Xinli, et al.
Published: (2025)
by: Tao, Xinli, et al.
Published: (2025)
Shallow Robustness, Deep Vulnerabilities: Multi-Turn Evaluation of Medical LLMs
by: Manczak, Blazej, et al.
Published: (2025)
by: Manczak, Blazej, et al.
Published: (2025)
BabyReasoningBench: Generating Developmentally-Inspired Reasoning Tasks for Evaluating Baby Language Models
by: Dhole, Kaustubh D.
Published: (2026)
by: Dhole, Kaustubh D.
Published: (2026)
From RAG to QA-RAG: Integrating Generative AI for Pharmaceutical Regulatory Compliance Process
by: Kim, Jaewoong, et al.
Published: (2024)
by: Kim, Jaewoong, et al.
Published: (2024)
ARGUS: Seeing the Influence of Narrative Features on Persuasion in Argumentative Texts
by: Nabhani, Sara, et al.
Published: (2026)
by: Nabhani, Sara, et al.
Published: (2026)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
by: Peters, Sydney, et al.
Published: (2025)
by: Peters, Sydney, et al.
Published: (2025)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
by: Oketunji, Abiodun Finbarrs
Published: (2023)
by: Oketunji, Abiodun Finbarrs
Published: (2023)
The Good, the Bad, and the Hulk-like GPT: Analyzing Emotional Decisions of Large Language Models in Cooperation and Bargaining Games
by: Mozikov, Mikhail, et al.
Published: (2024)
by: Mozikov, Mikhail, et al.
Published: (2024)
CPEMH: An Agentic Framework for Prompt-Driven Behavior Evaluation and Assurance in Foundation-Model Systems for Mental Health Screening
by: Lorenzoni, Giuliano, et al.
Published: (2026)
by: Lorenzoni, Giuliano, et al.
Published: (2026)
Categorical Perception in Large Language Model Hidden States: Structural Warping at Digit-Count Boundaries
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
End-to-End Evaluation and Governance of an EHR-Embedded AI Agent for Clinicians
by: Shah, Aaryan, et al.
Published: (2026)
by: Shah, Aaryan, et al.
Published: (2026)
Interpretability without actionability: mechanistic methods cannot correct language model errors despite near-perfect internal representations
by: Basu, Sanjay, et al.
Published: (2026)
by: Basu, Sanjay, et al.
Published: (2026)
Classifiers of Data Sharing Statements in Clinical Trial Records
by: Mamaghani, Saber Jelodari, et al.
Published: (2025)
by: Mamaghani, Saber Jelodari, et al.
Published: (2025)
Model selection meets clinical semantics: Optimizing ICD-10-CM prediction via LLM-as-Judge evaluation, redundancy-aware sampling, and section-aware fine-tuning
by: Dai, Hong-Jie, et al.
Published: (2025)
by: Dai, Hong-Jie, et al.
Published: (2025)
Robustness of Large Language Models to Perturbations in Text
by: Singh, Ayush, et al.
Published: (2024)
by: Singh, Ayush, et al.
Published: (2024)
NL2LOGIC: AST-Guided Translation of Natural Language into First-Order Logic with Large Language Models
by: Putra, Rizky Ramadhana, et al.
Published: (2026)
by: Putra, Rizky Ramadhana, et al.
Published: (2026)
Sparse Autoencoders Map Brain-LLM Alignment onto Cortical Semantic Topography
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
Cognitive bias in LLM reasoning compromises interpretation of clinical oncology notes
by: Kenaston, Matthew W., et al.
Published: (2025)
by: Kenaston, Matthew W., et al.
Published: (2025)
PerkwE_COQA: Enhanced Persian Conversational Question Answering by combining contextual keyword extraction with Large Language Models
by: Moradbeiki, Pardis, et al.
Published: (2024)
by: Moradbeiki, Pardis, et al.
Published: (2024)
Enigme: Generative Text Puzzles for Evaluating Reasoning in Language Models
by: Hawkins, John
Published: (2025)
by: Hawkins, John
Published: (2025)
Reasoning over Uncertain Text by Generative Large Language Models
by: Nafar, Aliakbar, et al.
Published: (2024)
by: Nafar, Aliakbar, et al.
Published: (2024)
Similar Items
-
Reshaping Free-Text Radiology Notes Into Structured Reports With Generative Transformers
by: Bergomi, Laura, et al.
Published: (2024) -
Advancing Italian Biomedical Information Extraction with Transformers-based Models: Methodological Insights and Multicenter Practical Application
by: Crema, Claudio, et al.
Published: (2023) -
RAR: Setting Knowledge Tripwires for Retrieval Augmented Rejection
by: Buonocore, Tommaso Mario, et al.
Published: (2025) -
BALI: Enhancing Biomedical Language Representations through Knowledge Graph and Language Model Alignment
by: Sakhovskiy, Andrey, et al.
Published: (2025) -
BLT: Can Large Language Models Handle Basic Legal Text?
by: Blair-Stanek, Andrew, et al.
Published: (2023)