Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs
Fuente:
arXiv
Saved in:
| Main Authors: | Rodriguez, David, Seymour, William, Del Alamo, Jose M., Such, Jose |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
How to Evaluate Medical AI
by: Kopanichuk, Ilia, et al.
Published: (2025)
by: Kopanichuk, Ilia, et al.
Published: (2025)
Evaluating Large Language Models on Historical Health Crisis Knowledge in Resource-Limited Settings: A Hybrid Multi-Metric Study
by: Hasan, Mohammed Rakibul
Published: (2026)
by: Hasan, Mohammed Rakibul
Published: (2026)
AskSport: Web Application for Sports Question-Answering
by: Onofre, Enzo B, et al.
Published: (2025)
by: Onofre, Enzo B, et al.
Published: (2025)
Comparing the Performance of LLMs in RAG-based Question-Answering: A Case Study in Computer Science Literature
by: Dayarathne, Ranul, et al.
Published: (2025)
by: Dayarathne, Ranul, et al.
Published: (2025)
ReTreVal: Reasoning Tree with Validation -- A Hybrid Framework for Enhanced LLM Multi-Step Reasoning
by: HS, Abhishek, et al.
Published: (2026)
by: HS, Abhishek, et al.
Published: (2026)
Challenges and Opportunities of NLP for HR Applications: A Discussion Paper
by: Leidner, Jochen L., et al.
Published: (2024)
by: Leidner, Jochen L., et al.
Published: (2024)
Comparative Analysis of AI Agent Architectures for Entity Relationship Classification
by: Berijanian, Maryam, et al.
Published: (2025)
by: Berijanian, Maryam, et al.
Published: (2025)
VERITAS-NLI : Validation and Extraction of Reliable Information Through Automated Scraping and Natural Language Inference
by: Shah, Arjun, et al.
Published: (2024)
by: Shah, Arjun, et al.
Published: (2024)
BLT: Can Large Language Models Handle Basic Legal Text?
by: Blair-Stanek, Andrew, et al.
Published: (2023)
by: Blair-Stanek, Andrew, et al.
Published: (2023)
Comprehensive Evaluation and Insights into the Use of Large Language Models in the Automation of Behavior-Driven Development Acceptance Test Formulation
by: Karpurapu, Shanthi, et al.
Published: (2024)
by: Karpurapu, Shanthi, et al.
Published: (2024)
Introducing Brain-like Concepts to Embodied Hand-crafted Dialog Management System
by: Joublin, Frank, et al.
Published: (2024)
by: Joublin, Frank, et al.
Published: (2024)
EvoIdeator: Evolving Scientific Ideas through Checklist-Grounded Reinforcement Learning
by: Sauter, Andreas, et al.
Published: (2026)
by: Sauter, Andreas, et al.
Published: (2026)
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
by: Hashemi, Helia, et al.
Published: (2024)
by: Hashemi, Helia, et al.
Published: (2024)
Meta-Evaluation of Translation Evaluation Methods: a systematic up-to-date overview
by: Han, Lifeng, et al.
Published: (2016)
by: Han, Lifeng, et al.
Published: (2016)
Towards Conditioning Clinical Text Generation for User Control
by: Koraş, Osman Alperen, et al.
Published: (2025)
by: Koraş, Osman Alperen, et al.
Published: (2025)
From RAG to QA-RAG: Integrating Generative AI for Pharmaceutical Regulatory Compliance Process
by: Kim, Jaewoong, et al.
Published: (2024)
by: Kim, Jaewoong, et al.
Published: (2024)
AutoTRIZ: Automating Engineering Innovation with TRIZ and Large Language Models
by: Jiang, Shuo, et al.
Published: (2024)
by: Jiang, Shuo, et al.
Published: (2024)
Lossless Prompt Compression via Dictionary-Encoding and In-Context Learning: Enabling Cost-Effective LLM Analysis of Repetitive Data
by: de Campos, Andresa Rodrigues, et al.
Published: (2026)
by: de Campos, Andresa Rodrigues, et al.
Published: (2026)
Automating the Deep Space Network Data Systems; A Case Study in Adaptive Anomaly Detection through Agentic AI
by: Chou, Evan J., et al.
Published: (2025)
by: Chou, Evan J., et al.
Published: (2025)
Open-TI: Open Traffic Intelligence with Augmented Language Model
by: Da, Longchao, et al.
Published: (2023)
by: Da, Longchao, et al.
Published: (2023)
Deciphering Digital Detectives: Understanding LLM Behaviors and Capabilities in Multi-Agent Mystery Games
by: Wu, Dekun, et al.
Published: (2023)
by: Wu, Dekun, et al.
Published: (2023)
FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation
by: Hildebrand, Samuel, et al.
Published: (2025)
by: Hildebrand, Samuel, et al.
Published: (2025)
Mitigating Structural Noise in Low-Resource S2TT: An Optimized Cascaded Nepali-English Pipeline with Punctuation Restoration
by: Chongbang, Tangsang, et al.
Published: (2026)
by: Chongbang, Tangsang, et al.
Published: (2026)
Exploring the Structure of AI-Induced Language Change in Scientific English
by: Galpin, Riley, et al.
Published: (2025)
by: Galpin, Riley, et al.
Published: (2025)
A Llama walks into the 'Bar': Efficient Supervised Fine-Tuning for Legal Reasoning in the Multi-state Bar Exam
by: Fernandes, Rean, et al.
Published: (2025)
by: Fernandes, Rean, et al.
Published: (2025)
LLMs Simulate Big Five Personality Traits: Further Evidence
by: Sorokovikova, Aleksandra, et al.
Published: (2024)
by: Sorokovikova, Aleksandra, et al.
Published: (2024)
Scaling In, Not Up? Testing Thick Citation Context Analysis with GPT-5 and Fragile Prompts
by: Simons, Arno
Published: (2026)
by: Simons, Arno
Published: (2026)
Teaching a Language Model to Speak the Language of Tools
by: Emanuilov, Simeon
Published: (2025)
by: Emanuilov, Simeon
Published: (2025)
OEMA: Ontology-Enhanced Multi-Agent Collaboration Framework for Zero-Shot Clinical Named Entity Recognition
by: Tao, Xinli, et al.
Published: (2025)
by: Tao, Xinli, et al.
Published: (2025)
AI-Powered Detection of Inappropriate Language in Medical School Curricula
by: Salavati, Chiman, et al.
Published: (2025)
by: Salavati, Chiman, et al.
Published: (2025)
LegalCheck: Retrieval- and Context-Augmented Generation for Drafting Municipal Legal Advice Letters
by: van der Meer, Virgill, et al.
Published: (2026)
by: van der Meer, Virgill, et al.
Published: (2026)
Project Riley: Multimodal Multi-Agent LLM Collaboration with Emotional Reasoning and Voting
by: Ortigoso, Ana Rita, et al.
Published: (2025)
by: Ortigoso, Ana Rita, et al.
Published: (2025)
Scaling Laws for State Dynamics in Large Language Models
by: Li, Jacob X, et al.
Published: (2025)
by: Li, Jacob X, et al.
Published: (2025)
HalluScan: A Systematic Benchmark for Detecting and Mitigating Hallucinations in Instruction-Following LLMs
by: Cherif, Ahmed
Published: (2026)
by: Cherif, Ahmed
Published: (2026)
Multi-Hierarchical Feature Detection for Large Language Model Generated Text
by: Zhang, Luyan, et al.
Published: (2025)
by: Zhang, Luyan, et al.
Published: (2025)
An Industrial-Scale Insurance LLM Achieving Verifiable Domain Mastery and Hallucination Control without Competence Trade-offs
by: Zhu, Qian, et al.
Published: (2026)
by: Zhu, Qian, et al.
Published: (2026)
The Need for Guardrails with Large Language Models in Medical Safety-Critical Settings: An Artificial Intelligence Application in the Pharmacovigilance Ecosystem
by: Hakim, Joe B, et al.
Published: (2024)
by: Hakim, Joe B, et al.
Published: (2024)
A Graph-based RAG for Energy Efficiency Question Answering
by: Campi, Riccardo, et al.
Published: (2025)
by: Campi, Riccardo, et al.
Published: (2025)
Social and Ethical Risks Posed by General-Purpose LLMs for Settling Newcomers in Canada
by: Nejadgholi, Isar, et al.
Published: (2024)
by: Nejadgholi, Isar, et al.
Published: (2024)
PestMA: LLM-based Multi-Agent System for Informed Pest Management
by: Shi, Hongrui, et al.
Published: (2025)
by: Shi, Hongrui, et al.
Published: (2025)
Similar Items
-
How to Evaluate Medical AI
by: Kopanichuk, Ilia, et al.
Published: (2025) -
Evaluating Large Language Models on Historical Health Crisis Knowledge in Resource-Limited Settings: A Hybrid Multi-Metric Study
by: Hasan, Mohammed Rakibul
Published: (2026) -
AskSport: Web Application for Sports Question-Answering
by: Onofre, Enzo B, et al.
Published: (2025) -
Comparing the Performance of LLMs in RAG-based Question-Answering: A Case Study in Computer Science Literature
by: Dayarathne, Ranul, et al.
Published: (2025) -
ReTreVal: Reasoning Tree with Validation -- A Hybrid Framework for Enhanced LLM Multi-Step Reasoning
by: HS, Abhishek, et al.
Published: (2026)