How to Assess Trustworthy AI in Practice
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zicari, Roberto V., Amann, Julia, Bruneault, Frédérick, Coffee, Megan, Düdder, Boris, Hickman, Eleanore, Gallucci, Alessio, Gilbert, Thomas Krendl, Hagendorff, Thilo, van Halem, Irmhild, Hildt, Elisabeth, Holm, Sune, Kararigas, Georgios, Kringen, Pedro, Madai, Vince I., Mathez, Emilie Wiinblad, Tithi, Jesmin Jahan, Vetter, Dennis, Westerlund, Magnus, Wurth, Renee |
|---|---|
| Format: | Preprint |
| Publié: |
2022
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Co-design for Trustworthy AI: An Interpretable and Explainable Tool for Type 2 Diabetes Prediction Using Genomic Polygenic Risk Scores
par: Beuthan, Ralf, et autres
Publié: (2026)
par: Beuthan, Ralf, et autres
Publié: (2026)
Lessons Learned in Performing a Trustworthy AI and Fundamental Rights Assessment
par: Boonstra, Marjolein, et autres
Publié: (2024)
par: Boonstra, Marjolein, et autres
Publié: (2024)
Getting Ready for the EU AI Act in Healthcare. A call for Sustainable AI Development and Deployment
par: Brodersen, John Brandt, et autres
Publié: (2025)
par: Brodersen, John Brandt, et autres
Publié: (2025)
Mapping the Ethics of Generative AI: A Comprehensive Scoping Review
par: Hagendorff, Thilo
Publié: (2024)
par: Hagendorff, Thilo
Publié: (2024)
On the Inevitability of Left-Leaning Political Bias in Aligned Language Models
par: Hagendorff, Thilo
Publié: (2025)
par: Hagendorff, Thilo
Publié: (2025)
Deception Abilities Emerged in Large Language Models
par: Hagendorff, Thilo
Publié: (2023)
par: Hagendorff, Thilo
Publié: (2023)
Beyond Chains of Thought: Benchmarking Latent-Space Reasoning Abilities in Large Language Models
par: Hagendorff, Thilo, et autres
Publié: (2025)
par: Hagendorff, Thilo, et autres
Publié: (2025)
When Image Generation Goes Wrong: A Safety Analysis of Stable Diffusion Models
par: Schneider, Matthias, et autres
Publié: (2024)
par: Schneider, Matthias, et autres
Publié: (2024)
PRIDE -- Parameter-Efficient Reduction of Identity Discrimination for Equality in LLMs
par: Menke, Maluna, et autres
Publié: (2025)
par: Menke, Maluna, et autres
Publié: (2025)
Fairness Hacking: The Malicious Practice of Shrouding Unfairness in Algorithms
par: Meding, Kristof, et autres
Publié: (2023)
par: Meding, Kristof, et autres
Publié: (2023)
Efficient Parallel Multi-Hop Reasoning: A Scalable Approach for Knowledge Graph Analysis
par: Tithi, Jesmin Jahan, et autres
Publié: (2024)
par: Tithi, Jesmin Jahan, et autres
Publié: (2024)
Ridgeline: A 2D Roofline Model for Distributed Systems
par: Checconi, Fabio, et autres
Publié: (2022)
par: Checconi, Fabio, et autres
Publié: (2022)
Evaluation Awareness in Language Models Has Limited Effect on Behaviour
par: Knecht, Amelie, et autres
Publié: (2026)
par: Knecht, Amelie, et autres
Publié: (2026)
Emergently Misaligned Language Models Show Behavioral Self-Awareness That Shifts With Subsequent Realignment
par: Vaugrante, Laurène, et autres
Publié: (2026)
par: Vaugrante, Laurène, et autres
Publié: (2026)
Large Reasoning Models Are Autonomous Jailbreak Agents
par: Hagendorff, Thilo, et autres
Publié: (2025)
par: Hagendorff, Thilo, et autres
Publié: (2025)
A Looming Replication Crisis in Evaluating Behavior in Language Models? Evidence and Solutions
par: Vaugrante, Laurène, et autres
Publié: (2024)
par: Vaugrante, Laurène, et autres
Publié: (2024)
Scaling Intelligence: Designing Data Centers for Next-Gen Language Models
par: Tithi, Jesmin Jahan, et autres
Publié: (2025)
par: Tithi, Jesmin Jahan, et autres
Publié: (2025)
Generative AI and generative education
par: Thomas Krendl Gilbert
Publié: (2024)
par: Thomas Krendl Gilbert
Publié: (2024)
Levers for successful implementation of the EU Nature Restoration Law: preparing for systemic biodiversity litigation
par: Laura Hildt
Publié: (2025)
par: Laura Hildt
Publié: (2025)
APHANES ARVENSIS (ROSACEAE) EN EL EXTREMO NORDESTE DE LA ARGENTINA
par: Gustavo Hildt
Publié: (2014)
par: Gustavo Hildt
Publié: (2014)
DUNE: A Machine Learning Deep UNet++ based Ensemble Approach to Monthly, Seasonal and Annual Climate Forecasting
par: Shukla, Pratik, et autres
Publié: (2024)
par: Shukla, Pratik, et autres
Publié: (2024)
Compromising Honesty and Harmlessness in Language Models via Deception Attacks
par: Vaugrante, Laurène, et autres
Publié: (2025)
par: Vaugrante, Laurène, et autres
Publié: (2025)
Speciesism in AI: Evaluating Discrimination Against Animals in Large Language Models
par: Jotautaitė, Monika, et autres
Publié: (2025)
par: Jotautaitė, Monika, et autres
Publié: (2025)
"Dark Triad" Model Organisms of Misalignment: Narrow Fine-Tuning Mirrors Human Antisocial Behavior
par: Lulla, Roshni, et autres
Publié: (2026)
par: Lulla, Roshni, et autres
Publié: (2026)
Textbook Censorship and Intolerance in the Classroom.
par: Richardson, Eleanore H.
Publié: (1982)
par: Richardson, Eleanore H.
Publié: (1982)
A Study of Microfiche as an Alternative to the Reserve Room Function.
par: Ficke, Eleanore R.
Publié: (1974)
par: Ficke, Eleanore R.
Publié: (1974)
Reconcilable Differences: Comparative Analysis of EU and US Ethical AI Frameworks with Focus on Divergent Ethical Aspects
par: Cameron M Pierson, et autres
Publié: (2025)
par: Cameron M Pierson, et autres
Publié: (2025)
Verdi in Victorian London
par: Zicari, Massimo
Publié: (2017)
par: Zicari, Massimo
Publié: (2017)
The Voice of the Century
par: Zicari, Massimo
Publié: (2022)
par: Zicari, Massimo
Publié: (2022)
Topología, dominación y subjetividad. Las teorías del poder de Michael Foucault y de Norbert Elías en perspectiva comparada
par: Julián Zícari
Publié: (2018)
par: Julián Zícari
Publié: (2018)
Imagen tomográfica infrecuente en patología biliar
par: Marcelo Zícari
Publié: (2015)
par: Marcelo Zícari
Publié: (2015)
The Power of Lithium in South America
par: Julián Zicari
Publié: (2017)
par: Julián Zicari
Publié: (2017)
¿Igualdad natural, desigualdad artificial? Hobbes, el problema del igualitarismo y las ficciones del ‘como si’
par: Julián Zícari
Publié: (2017)
par: Julián Zícari
Publié: (2017)
Finanzas personales y ciclo de vida: un desafío actual
par: Adrián Zicari
Publié: (2008)
par: Adrián Zicari
Publié: (2008)
Responsabilidad social empresaria: del dicho al hecho. Poniéndole números a la responsabilidad social
par: Adrián Zicari
Publié: (2006)
par: Adrián Zicari
Publié: (2006)
A rare cause of abdominal tumor
par: Marcelo Zícari
Publié: (2012)
par: Marcelo Zícari
Publié: (2012)
Narrativa literaria e historia, algunos puntos de debate: la concepción metahistórica de Hayden White frente a las críticas de Chris Lorenz
par: Julián Zícari
Publié: (2015)
par: Julián Zícari
Publié: (2015)
Fondos responsables: una exploración de su viabilidad en el mercado de capitales argentino
par: Adrián Zicari
Publié: (2007)
par: Adrián Zicari
Publié: (2007)
ReLATE: Learning Efficient Sparse Encoding for High-Performance Tensor Decomposition
par: Helal, Ahmed E., et autres
Publié: (2025)
par: Helal, Ahmed E., et autres
Publié: (2025)
Deciding Termination of Simple Randomized Loops
par: Meyer, Éléanore, et autres
Publié: (2025)
par: Meyer, Éléanore, et autres
Publié: (2025)
Documents similaires
-
Co-design for Trustworthy AI: An Interpretable and Explainable Tool for Type 2 Diabetes Prediction Using Genomic Polygenic Risk Scores
par: Beuthan, Ralf, et autres
Publié: (2026) -
Lessons Learned in Performing a Trustworthy AI and Fundamental Rights Assessment
par: Boonstra, Marjolein, et autres
Publié: (2024) -
Getting Ready for the EU AI Act in Healthcare. A call for Sustainable AI Development and Deployment
par: Brodersen, John Brandt, et autres
Publié: (2025) -
Mapping the Ethics of Generative AI: A Comprehensive Scoping Review
par: Hagendorff, Thilo
Publié: (2024) -
On the Inevitability of Left-Leaning Political Bias in Aligned Language Models
par: Hagendorff, Thilo
Publié: (2025)