SAIF: A Comprehensive Framework for Evaluating the Risks of Generative AI in the Public Sector
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Kyeongryul, Kim, Heehyeon, Whang, Joyce Jiyoung |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Unveiling the Threat of Fraud Gangs to Graph Neural Networks: Multi-Target Graph Injection Attacks Against GNN-Based Fraud Detectors
di: Choi, Jinhyeok, et al.
Pubblicazione: (2024)
di: Choi, Jinhyeok, et al.
Pubblicazione: (2024)
Dynamic Relation-Attentive Graph Neural Networks for Fraud Detection
di: Kim, Heehyeon, et al.
Pubblicazione: (2023)
di: Kim, Heehyeon, et al.
Pubblicazione: (2023)
Classroom AI: Large Language Models as Grade-Specific Teachers
di: Oh, Jio, et al.
Pubblicazione: (2026)
di: Oh, Jio, et al.
Pubblicazione: (2026)
SpoT-Mamba: Learning Long-Range Dependency on Spatio-Temporal Graphs with Selective State Spaces
di: Choi, Jinhyeok, et al.
Pubblicazione: (2024)
di: Choi, Jinhyeok, et al.
Pubblicazione: (2024)
Generative AI in Saudi Arabia: A National Survey of Adoption, Risks, and Public Perceptions
di: AlDakheel, Abdulaziz, et al.
Pubblicazione: (2026)
di: AlDakheel, Abdulaziz, et al.
Pubblicazione: (2026)
Why So Gullible? Enhancing the Robustness of Retrieval-Augmented Models against Counterfactual Noise
di: Hong, Giwon, et al.
Pubblicazione: (2023)
di: Hong, Giwon, et al.
Pubblicazione: (2023)
Evaluating Patient Safety Risks in Generative AI: Development and Validation of a FMECA Framework for Generated Clinical Content
di: Bednarczyk, Lydie, et al.
Pubblicazione: (2026)
di: Bednarczyk, Lydie, et al.
Pubblicazione: (2026)
Unifying Inductive, Cross-Domain, and Multimodal Learning for Robust and Generalizable Recommendation
di: Chung, Chanyoung, et al.
Pubblicazione: (2025)
di: Chung, Chanyoung, et al.
Pubblicazione: (2025)
Evaluation Framework for AI Systems in "the Wild"
di: Jabbour, Sarah, et al.
Pubblicazione: (2025)
di: Jabbour, Sarah, et al.
Pubblicazione: (2025)
AssurAI: Experience with Constructing Korean Socio-cultural Datasets to Discover Potential Risks of Generative AI
di: Lim, Chae-Gyun, et al.
Pubblicazione: (2025)
di: Lim, Chae-Gyun, et al.
Pubblicazione: (2025)
General Scales Unlock AI Evaluation with Explanatory and Predictive Power
di: Zhou, Lexin, et al.
Pubblicazione: (2025)
di: Zhou, Lexin, et al.
Pubblicazione: (2025)
A Study on the Framework for Evaluating the Ethics and Trustworthiness of Generative AI
di: Jeong, Cheonsu, et al.
Pubblicazione: (2025)
di: Jeong, Cheonsu, et al.
Pubblicazione: (2025)
AI-Assisted Systematization for Evaluating GenAI Systems
di: Agarwal, Dhruv, et al.
Pubblicazione: (2026)
di: Agarwal, Dhruv, et al.
Pubblicazione: (2026)
A Unified Framework to Quantify Cultural Intelligence of AI
di: Dev, Sunipa, et al.
Pubblicazione: (2026)
di: Dev, Sunipa, et al.
Pubblicazione: (2026)
Understanding and Mitigating Risks of Generative AI in Financial Services
di: Gehrmann, Sebastian, et al.
Pubblicazione: (2025)
di: Gehrmann, Sebastian, et al.
Pubblicazione: (2025)
EvalCards: A Framework for Standardized Evaluation Reporting
di: Dhar, Ruchira, et al.
Pubblicazione: (2025)
di: Dhar, Ruchira, et al.
Pubblicazione: (2025)
Interpreting Public Sentiment in Diplomacy Events: A Counterfactual Analysis Framework Using Large Language Models
di: Ouyang, Leyi
Pubblicazione: (2025)
di: Ouyang, Leyi
Pubblicazione: (2025)
From Complexity to Clarity: How AI Enhances Perceptions of Scientists and the Public's Understanding of Science
di: Markowitz, David M.
Pubblicazione: (2024)
di: Markowitz, David M.
Pubblicazione: (2024)
Trans-EnV: A Framework for Evaluating the Linguistic Robustness of LLMs Against English Varieties
di: Lee, Jiyoung, et al.
Pubblicazione: (2025)
di: Lee, Jiyoung, et al.
Pubblicazione: (2025)
SAIF: A Sparse Autoencoder Framework for Interpreting and Steering Instruction Following of Language Models
di: He, Zirui, et al.
Pubblicazione: (2025)
di: He, Zirui, et al.
Pubblicazione: (2025)
CLEAR: A Clinically-Grounded Tabular Framework for Radiology Report Evaluation
di: Jiang, Yuyang, et al.
Pubblicazione: (2025)
di: Jiang, Yuyang, et al.
Pubblicazione: (2025)
Frontier Lag: A Bibliometric Audit of Capability Misrepresentation in Academic AI Evaluation
di: Gringras, David, et al.
Pubblicazione: (2026)
di: Gringras, David, et al.
Pubblicazione: (2026)
Security Challenges in AI Agent Deployment: Insights from a Large Scale Public Competition
di: Zou, Andy, et al.
Pubblicazione: (2025)
di: Zou, Andy, et al.
Pubblicazione: (2025)
Why They Disagree: Decoding Differences in Opinions about AI Risk on the Lex Fridman Podcast
di: Truong, Nghi, et al.
Pubblicazione: (2025)
di: Truong, Nghi, et al.
Pubblicazione: (2025)
An analysis of AI Decision under Risk: Prospect theory emerges in Large Language Models
di: Payne, Kenneth
Pubblicazione: (2025)
di: Payne, Kenneth
Pubblicazione: (2025)
Comprehensive Framework for Evaluating Conversational AI Chatbots
di: Gupta, Shailja, et al.
Pubblicazione: (2025)
di: Gupta, Shailja, et al.
Pubblicazione: (2025)
Evaluation of AI Ethics Tools in Language Models: A Developers' Perspective Case Study
di: Silva, Jhessica, et al.
Pubblicazione: (2025)
di: Silva, Jhessica, et al.
Pubblicazione: (2025)
Mitigating Gambling-Like Risk-Taking Behaviors in Large Language Models: A Behavioral Economics Approach to AI Safety
di: Du, Y.
Pubblicazione: (2025)
di: Du, Y.
Pubblicazione: (2025)
Teaching at Scale: Leveraging AI to Evaluate and Elevate Engineering Education
di: Chamberland, Jean-Francois, et al.
Pubblicazione: (2025)
di: Chamberland, Jean-Francois, et al.
Pubblicazione: (2025)
Explainability and Certification of AI-Generated Educational Assessments
di: Yaacoub, Antoun, et al.
Pubblicazione: (2026)
di: Yaacoub, Antoun, et al.
Pubblicazione: (2026)
Cultural Compass: A Framework for Organizing Societal Norms to Detect Violations in Human-AI Conversations
di: Cheng, Myra, et al.
Pubblicazione: (2026)
di: Cheng, Myra, et al.
Pubblicazione: (2026)
Nunchi-Bench: Benchmarking Language Models on Cultural Reasoning with a Focus on Korean Superstition
di: Kim, Kyuhee, et al.
Pubblicazione: (2025)
di: Kim, Kyuhee, et al.
Pubblicazione: (2025)
Rethinking Test-Time Scaling for Medical AI: Model and Task-Aware Strategies for LLMs and VLMs
di: Oh, Gyutaek, et al.
Pubblicazione: (2025)
di: Oh, Gyutaek, et al.
Pubblicazione: (2025)
The Responsible Development of Automated Student Feedback with Generative AI
di: Lindsay, Euan D, et al.
Pubblicazione: (2023)
di: Lindsay, Euan D, et al.
Pubblicazione: (2023)
Measuring Human Contribution in AI-Assisted Content Generation
di: Xie, Yueqi, et al.
Pubblicazione: (2024)
di: Xie, Yueqi, et al.
Pubblicazione: (2024)
Must Read: A Comprehensive Survey of Computational Persuasion
di: Bozdag, Nimet Beyza, et al.
Pubblicazione: (2025)
di: Bozdag, Nimet Beyza, et al.
Pubblicazione: (2025)
Detecting AI-Generated Text in Educational Content: Leveraging Machine Learning and Explainable AI for Academic Integrity
di: Najjar, Ayat A., et al.
Pubblicazione: (2025)
di: Najjar, Ayat A., et al.
Pubblicazione: (2025)
Generative AI Purpose-built for Social and Mental Health: A Real-World Pilot
di: Hull, Thomas D., et al.
Pubblicazione: (2025)
di: Hull, Thomas D., et al.
Pubblicazione: (2025)
From Black-Box Confidence to Measurable Trust in Clinical AI: A Framework for Evidence, Supervision, and Staged Autonomy
di: Zabolotnii, Serhii, et al.
Pubblicazione: (2026)
di: Zabolotnii, Serhii, et al.
Pubblicazione: (2026)
Effect of Gender Fair Job Description on Generative AI Images
di: Böckling, Finn, et al.
Pubblicazione: (2025)
di: Böckling, Finn, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Unveiling the Threat of Fraud Gangs to Graph Neural Networks: Multi-Target Graph Injection Attacks Against GNN-Based Fraud Detectors
di: Choi, Jinhyeok, et al.
Pubblicazione: (2024) -
Dynamic Relation-Attentive Graph Neural Networks for Fraud Detection
di: Kim, Heehyeon, et al.
Pubblicazione: (2023) -
Classroom AI: Large Language Models as Grade-Specific Teachers
di: Oh, Jio, et al.
Pubblicazione: (2026) -
SpoT-Mamba: Learning Long-Range Dependency on Spatio-Temporal Graphs with Selective State Spaces
di: Choi, Jinhyeok, et al.
Pubblicazione: (2024) -
Generative AI in Saudi Arabia: A National Survey of Adoption, Risks, and Public Perceptions
di: AlDakheel, Abdulaziz, et al.
Pubblicazione: (2026)