CPEMH: An Agentic Framework for Prompt-Driven Behavior Evaluation and Assurance in Foundation-Model Systems for Mental Health Screening
Fuente:
arXiv
Guardado en:
| Autores principales: | Lorenzoni, Giuliano, Portugal, Ivens, Alencar, Paulo, Cowan, Donald |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Do We Always Need Query-Level Workflows? Rethinking Agentic Workflow Generation for Multi-Agent Systems
por: Wang, Zixu, et al.
Publicado: (2026)
por: Wang, Zixu, et al.
Publicado: (2026)
Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews
por: Li, Bowen, et al.
Publicado: (2026)
por: Li, Bowen, et al.
Publicado: (2026)
Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework
por: Wang, Xiaohua, et al.
Publicado: (2026)
por: Wang, Xiaohua, et al.
Publicado: (2026)
ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting
por: Chang, Jiale, et al.
Publicado: (2026)
por: Chang, Jiale, et al.
Publicado: (2026)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
por: Wang, Yuchen, et al.
Publicado: (2026)
por: Wang, Yuchen, et al.
Publicado: (2026)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
por: Costa, Rimom
Publicado: (2025)
por: Costa, Rimom
Publicado: (2025)
Applying Cognitive Design Patterns to General LLM Agents
por: Wray, Robert E., et al.
Publicado: (2025)
por: Wray, Robert E., et al.
Publicado: (2025)
Good to Go: The LOOP Skill Engine That Hits 99% Success and Slashes Token Usage by 99% via One-Shot Recording and Deterministic Replay
por: Wang, Xiaohua, et al.
Publicado: (2026)
por: Wang, Xiaohua, et al.
Publicado: (2026)
Post Hoc Extraction of Pareto Fronts for Continuous Control
por: Thakar, Raghav, et al.
Publicado: (2026)
por: Thakar, Raghav, et al.
Publicado: (2026)
MEMTIER: Tiered Memory Architecture and Retrieval Bottleneck Analysis for Long-Running Autonomous AI Agents
por: Sidik, Bronislav, et al.
Publicado: (2026)
por: Sidik, Bronislav, et al.
Publicado: (2026)
FediLoRA: Practical Federated Fine-Tuning of Foundation Models Under Missing-Modality Constraints
por: Yang, Lishan, et al.
Publicado: (2025)
por: Yang, Lishan, et al.
Publicado: (2025)
Beyond Benchmark Islands: Toward Representative Trustworthiness Evaluation for Agentic AI
por: Qi, Jinhu, et al.
Publicado: (2026)
por: Qi, Jinhu, et al.
Publicado: (2026)
Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines
por: Ray, Aninda
Publicado: (2026)
por: Ray, Aninda
Publicado: (2026)
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
por: Annapureddy, Sasank
Publicado: (2026)
por: Annapureddy, Sasank
Publicado: (2026)
Latent Cache Flow: Model-to-Model Communication Without Text
por: Rossi, Maximillian, et al.
Publicado: (2026)
por: Rossi, Maximillian, et al.
Publicado: (2026)
Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models
por: Gokdemir, Ozan, et al.
Publicado: (2025)
por: Gokdemir, Ozan, et al.
Publicado: (2025)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
por: Wu, Shuai, et al.
Publicado: (2026)
por: Wu, Shuai, et al.
Publicado: (2026)
Contrastive Learning-Enhanced Large Language Models for Monolith-to-Microservice Decomposition
por: Sellami, Khaled, et al.
Publicado: (2025)
por: Sellami, Khaled, et al.
Publicado: (2025)
elsciRL: Integrating Language Solutions into Reinforcement Learning Problem Settings
por: Osborne, Philip, et al.
Publicado: (2025)
por: Osborne, Philip, et al.
Publicado: (2025)
CRAwDAD: Causal Reasoning Augmentation with Dual-Agent Debate
por: Vamosi, Finn G., et al.
Publicado: (2025)
por: Vamosi, Finn G., et al.
Publicado: (2025)
GSAR: Typed Grounding for Hallucination Detection and Recovery in Multi-Agent LLMs
por: Kamelhar, Federico A.
Publicado: (2026)
por: Kamelhar, Federico A.
Publicado: (2026)
Generative AI Toolkit -- a framework for increasing the quality of LLM-based applications over their whole life cycle
por: Kohl, Jens, et al.
Publicado: (2024)
por: Kohl, Jens, et al.
Publicado: (2024)
Eliciting Problem Specifications via Large Language Models
por: Wray, Robert E., et al.
Publicado: (2024)
por: Wray, Robert E., et al.
Publicado: (2024)
MMiC: Mitigating Modality Incompleteness in Clustered Federated Learning
por: Yang, Lishan, et al.
Publicado: (2025)
por: Yang, Lishan, et al.
Publicado: (2025)
Agent WARPP: Workflow Adherence via Runtime Parallel Personalization
por: Mazzolenis, Maria Emilia, et al.
Publicado: (2025)
por: Mazzolenis, Maria Emilia, et al.
Publicado: (2025)
Exploring Design of Multi-Agent LLM Dialogues for Research Ideation
por: Ueda, Keisuke, et al.
Publicado: (2025)
por: Ueda, Keisuke, et al.
Publicado: (2025)
A Scalable Communication Protocol for Networks of Large Language Models
por: Marro, Samuele, et al.
Publicado: (2024)
por: Marro, Samuele, et al.
Publicado: (2024)
A Prompt-Aware Structuring Framework for Reliable Reuse of AI-Generated Content in the Agentic Web
por: Egami, Shusaku, et al.
Publicado: (2026)
por: Egami, Shusaku, et al.
Publicado: (2026)
Agentic Automation of BT-RADS Scoring: End-to-End Multi-Agent System for Standardized Brain Tumor Follow-up Assessment
por: Jabal, Mohamed Sobhi, et al.
Publicado: (2026)
por: Jabal, Mohamed Sobhi, et al.
Publicado: (2026)
Controlling Long-Horizon Behavior in Language Model Agents with Explicit State Dynamics
por: Subaharan, Sukesh
Publicado: (2026)
por: Subaharan, Sukesh
Publicado: (2026)
Advancing Transformer Architecture in Long-Context Large Language Models: A Comprehensive Survey
por: Huang, Yunpeng, et al.
Publicado: (2023)
por: Huang, Yunpeng, et al.
Publicado: (2023)
Fuzzy, Symbolic, and Contextual: Enhancing LLM Instruction via Cognitive Scaffolding
por: Figueiredo, Vanessa
Publicado: (2025)
por: Figueiredo, Vanessa
Publicado: (2025)
Tool-RoCo: An Agent-as-Tool Self-organization Large Language Model Benchmark in Multi-robot Cooperation
por: Zhang, Ke, et al.
Publicado: (2025)
por: Zhang, Ke, et al.
Publicado: (2025)
Umwelt Engineering: Designing the Cognitive Worlds of Linguistic Agents
por: Jehu-Appiah, Rodney
Publicado: (2026)
por: Jehu-Appiah, Rodney
Publicado: (2026)
Towards Resource-Efficient Multimodal Intelligence: Learned Routing among Specialized Expert Models
por: Saini, Mayank, et al.
Publicado: (2025)
por: Saini, Mayank, et al.
Publicado: (2025)
RefiningGPT: Specialized language Models for Automated Refinery Unit-level Process Diagram Synthesis
por: Liu, Dongxiao, et al.
Publicado: (2026)
por: Liu, Dongxiao, et al.
Publicado: (2026)
Prompt Fencing: A Cryptographic Approach to Establishing Security Boundaries in Large Language Model Prompts
por: Peh, Steven
Publicado: (2025)
por: Peh, Steven
Publicado: (2025)
A Simple Architecture for Enterprise Large Language Model Applications based on Role based security and Clearance Levels using Retrieval-Augmented Generation or Mixture of Experts
por: Özgür, Atilla, et al.
Publicado: (2024)
por: Özgür, Atilla, et al.
Publicado: (2024)
ART: Adaptive Response Tuning Framework -- A Multi-Agent Tournament-Based Approach to LLM Response Optimization
por: Khan, Omer Jauhar
Publicado: (2025)
por: Khan, Omer Jauhar
Publicado: (2025)
Complementarity, Augmentation, or Substitutivity? The Impact of Generative Artificial Intelligence on the U.S. Federal Workforce
por: Resh, William G., et al.
Publicado: (2025)
por: Resh, William G., et al.
Publicado: (2025)
Ejemplares similares
-
Do We Always Need Query-Level Workflows? Rethinking Agentic Workflow Generation for Multi-Agent Systems
por: Wang, Zixu, et al.
Publicado: (2026) -
Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews
por: Li, Bowen, et al.
Publicado: (2026) -
Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework
por: Wang, Xiaohua, et al.
Publicado: (2026) -
ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting
por: Chang, Jiale, et al.
Publicado: (2026) -
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
por: Wang, Yuchen, et al.
Publicado: (2026)