Preventing Another Tessa: Modular Safety Middleware For Health-Adjacent AI Assistants
Fuente:
arXiv
Salvato in:
| Autori principali: | Reddy, Pavan, Reddy, Nithin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
EthicsMH: A Pilot Benchmark for Ethical Reasoning in Mental Health AI
di: Kasu, Sai Kartheek Reddy
Pubblicazione: (2025)
di: Kasu, Sai Kartheek Reddy
Pubblicazione: (2025)
Balancing Safety and Helpfulness in Healthcare AI Assistants through Iterative Preference Alignment
di: Nghiem, Huy, et al.
Pubblicazione: (2025)
di: Nghiem, Huy, et al.
Pubblicazione: (2025)
Development of Semantics-Based Distributed Middleware for Heterogeneous Data Integration and its Application for Drought
di: Akanbi, A
Pubblicazione: (2024)
di: Akanbi, A
Pubblicazione: (2024)
A Knowledge-Component-Based Methodology for Evaluating AI Assistants
di: Qi, Laryn, et al.
Pubblicazione: (2024)
di: Qi, Laryn, et al.
Pubblicazione: (2024)
AgentSociety: Incentivizing Agentic Social Intelligence
di: Kesari, Aditya Vema Reddy, et al.
Pubblicazione: (2026)
di: Kesari, Aditya Vema Reddy, et al.
Pubblicazione: (2026)
Patentformer: A demonstration of AI-assisted automated patent drafting
di: Mudhiganti, Sai Krishna Reddy, et al.
Pubblicazione: (2025)
di: Mudhiganti, Sai Krishna Reddy, et al.
Pubblicazione: (2025)
AI-driven Personalized Privacy Assistants: a Systematic Literature Review
di: Morel, Victor, et al.
Pubblicazione: (2025)
di: Morel, Victor, et al.
Pubblicazione: (2025)
Human or AI? Comparing Design Thinking Assessments by Teaching Assistants and Bots
di: Khan, Sumbul, et al.
Pubblicazione: (2025)
di: Khan, Sumbul, et al.
Pubblicazione: (2025)
Aalap: AI Assistant for Legal & Paralegal Functions in India
di: Tiwari, Aman, et al.
Pubblicazione: (2024)
di: Tiwari, Aman, et al.
Pubblicazione: (2024)
AI Safety is Stuck in Technical Terms -- A System Safety Response to the International AI Safety Report
di: Dobbe, Roel
Pubblicazione: (2025)
di: Dobbe, Roel
Pubblicazione: (2025)
The Impact of Artificial Intelligence on Traditional Art Forms: A Disruption or Enhancement
di: Marella, Viswa Chaitanya, et al.
Pubblicazione: (2025)
di: Marella, Viswa Chaitanya, et al.
Pubblicazione: (2025)
Digital Companionship: Overlapping Uses of AI Companions and AI Assistants
di: Manoli, Aikaterina, et al.
Pubblicazione: (2025)
di: Manoli, Aikaterina, et al.
Pubblicazione: (2025)
International Agreements on AI Safety: Review and Recommendations for a Conditional AI Safety Treaty
di: Scholefield, Rebecca, et al.
Pubblicazione: (2025)
di: Scholefield, Rebecca, et al.
Pubblicazione: (2025)
Framework, Standards, Applications and Best practices of Responsible AI : A Comprehensive Survey
di: Gadekallu, Thippa Reddy, et al.
Pubblicazione: (2025)
di: Gadekallu, Thippa Reddy, et al.
Pubblicazione: (2025)
Preventing AI Deepfake Abuse: An Islamic Ethics Framework
di: Uriawan, Wisnu, et al.
Pubblicazione: (2025)
di: Uriawan, Wisnu, et al.
Pubblicazione: (2025)
Safety Cases: A Scalable Approach to Frontier AI Safety
di: Hilton, Benjamin, et al.
Pubblicazione: (2025)
di: Hilton, Benjamin, et al.
Pubblicazione: (2025)
Safety Cases: How to Justify the Safety of Advanced AI Systems
di: Clymer, Joshua, et al.
Pubblicazione: (2024)
di: Clymer, Joshua, et al.
Pubblicazione: (2024)
Toward an African Agenda for AI Safety
di: Segun, Samuel T., et al.
Pubblicazione: (2025)
di: Segun, Samuel T., et al.
Pubblicazione: (2025)
Concrete Problems in AI Safety, Revisited
di: Raji, Inioluwa Deborah, et al.
Pubblicazione: (2023)
di: Raji, Inioluwa Deborah, et al.
Pubblicazione: (2023)
Examining Student Interactions with a Pedagogical AI-Assistant for Essay Writing and their Impact on Students Writing Quality
di: Febriantoro, Wicaksono, et al.
Pubblicazione: (2025)
di: Febriantoro, Wicaksono, et al.
Pubblicazione: (2025)
How effective are VLMs in assisting humans in inferring the quality of mental models from Multimodal short answers?
di: Sil, Pritam, et al.
Pubblicazione: (2026)
di: Sil, Pritam, et al.
Pubblicazione: (2026)
Emerging Practices in Frontier AI Safety Frameworks
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2025)
di: Buhl, Marie Davidsen, et al.
Pubblicazione: (2025)
AI Safety: Necessary, but insufficient and possibly problematic
di: P, Deepak
Pubblicazione: (2024)
di: P, Deepak
Pubblicazione: (2024)
Intelligent Approaches to Predictive Analytics in Occupational Health and Safety in India
di: Saxena, Ritwik Raj
Pubblicazione: (2024)
di: Saxena, Ritwik Raj
Pubblicazione: (2024)
Combining Cost-Constrained Runtime Monitors for AI Safety
di: Hua, Tim Tian, et al.
Pubblicazione: (2025)
di: Hua, Tim Tian, et al.
Pubblicazione: (2025)
Building Effective Safety Guardrails in AI Education Tools
di: Clark, Hannah-Beth, et al.
Pubblicazione: (2025)
di: Clark, Hannah-Beth, et al.
Pubblicazione: (2025)
What Is AI Safety? What Do We Want It to Be?
di: Harding, Jacqueline, et al.
Pubblicazione: (2025)
di: Harding, Jacqueline, et al.
Pubblicazione: (2025)
The Singapore Consensus on Global AI Safety Research Priorities
di: Bengio, Yoshua, et al.
Pubblicazione: (2025)
di: Bengio, Yoshua, et al.
Pubblicazione: (2025)
The Ghost in the Grammar: Methodological Anthropomorphism in AI Safety Evaluations
di: Costa, Mariana Lins
Pubblicazione: (2026)
di: Costa, Mariana Lins
Pubblicazione: (2026)
Upstream and Downstream AI Safety: Both on the Same River?
di: McDermid, John, et al.
Pubblicazione: (2024)
di: McDermid, John, et al.
Pubblicazione: (2024)
Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents
di: Li, Miles Q., et al.
Pubblicazione: (2026)
di: Li, Miles Q., et al.
Pubblicazione: (2026)
Agentic Microphysics: A Manifesto for Generative AI Safety
di: Pierucci, Federico, et al.
Pubblicazione: (2026)
di: Pierucci, Federico, et al.
Pubblicazione: (2026)
Probabilistic Analysis of Copyright Disputes and Generative AI Safety
di: Chiba-Okabe, Hiroaki
Pubblicazione: (2024)
di: Chiba-Okabe, Hiroaki
Pubblicazione: (2024)
Interoperability in AI Safety Governance: Ethics, Regulations, and Standards
di: Chin, Yik Chan, et al.
Pubblicazione: (2026)
di: Chin, Yik Chan, et al.
Pubblicazione: (2026)
Customer Service Representative's Perception of the AI Assistant in an Organization's Call Center
di: Qin, Kai, et al.
Pubblicazione: (2025)
di: Qin, Kai, et al.
Pubblicazione: (2025)
Towards Ethical Personal AI Applications: Practical Considerations for AI Assistants with Long-Term Memory
di: Lee, Eunhae
Pubblicazione: (2024)
di: Lee, Eunhae
Pubblicazione: (2024)
Could ChatGPT get an Engineering Degree? Evaluating Higher Education Vulnerability to AI Assistants
di: Borges, Beatriz, et al.
Pubblicazione: (2024)
di: Borges, Beatriz, et al.
Pubblicazione: (2024)
Responsible Evaluation of AI for Mental Health
di: Arnaout, Hiba, et al.
Pubblicazione: (2026)
di: Arnaout, Hiba, et al.
Pubblicazione: (2026)
The Elephant in the Room -- Why AI Safety Demands Diverse Teams
di: Rostcheck, David, et al.
Pubblicazione: (2024)
di: Rostcheck, David, et al.
Pubblicazione: (2024)
Bridging Today and the Future of Humanity: AI Safety in 2024 and Beyond
di: Han, Shanshan
Pubblicazione: (2024)
di: Han, Shanshan
Pubblicazione: (2024)
Documenti analoghi
-
EthicsMH: A Pilot Benchmark for Ethical Reasoning in Mental Health AI
di: Kasu, Sai Kartheek Reddy
Pubblicazione: (2025) -
Balancing Safety and Helpfulness in Healthcare AI Assistants through Iterative Preference Alignment
di: Nghiem, Huy, et al.
Pubblicazione: (2025) -
Development of Semantics-Based Distributed Middleware for Heterogeneous Data Integration and its Application for Drought
di: Akanbi, A
Pubblicazione: (2024) -
A Knowledge-Component-Based Methodology for Evaluating AI Assistants
di: Qi, Laryn, et al.
Pubblicazione: (2024) -
AgentSociety: Incentivizing Agentic Social Intelligence
di: Kesari, Aditya Vema Reddy, et al.
Pubblicazione: (2026)