Towards Effective Discrimination Testing for Generative AI
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zollo, Thomas P., Rajaneesh, Nikita, Zemel, Richard, Gillis, Talia B., Black, Emily |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Test-Time Warmup for Multimodal Large Language Models
von: Rajaneesh, Nikita, et al.
Veröffentlicht: (2025)
von: Rajaneesh, Nikita, et al.
Veröffentlicht: (2025)
Adaptive Elicitation of Latent Information Using Natural Language
von: Wang, Jimmy, et al.
Veröffentlicht: (2025)
von: Wang, Jimmy, et al.
Veröffentlicht: (2025)
Prompt Risk Control: A Rigorous Framework for Responsible Deployment of Large Language Models
von: Zollo, Thomas P., et al.
Veröffentlicht: (2023)
von: Zollo, Thomas P., et al.
Veröffentlicht: (2023)
Improving Predictor Reliability with Selective Recalibration
von: Zollo, Thomas P., et al.
Veröffentlicht: (2024)
von: Zollo, Thomas P., et al.
Veröffentlicht: (2024)
Whom to Query for What: Adaptive Group Elicitation via Multi-Turn LLM Interactions
von: Ding, Ruomeng, et al.
Veröffentlicht: (2026)
von: Ding, Ruomeng, et al.
Veröffentlicht: (2026)
Towards Responsible Development of Generative AI for Education: An Evaluation-Driven Approach
von: Jurenka, Irina, et al.
Veröffentlicht: (2024)
von: Jurenka, Irina, et al.
Veröffentlicht: (2024)
Generative AI Meets Future Cities: Towards an Era of Autonomous Urban Intelligence
von: Wang, Dongjie, et al.
Veröffentlicht: (2023)
von: Wang, Dongjie, et al.
Veröffentlicht: (2023)
Towards Environmentally Equitable AI
von: Hajiesmaili, Mohammad, et al.
Veröffentlicht: (2024)
von: Hajiesmaili, Mohammad, et al.
Veröffentlicht: (2024)
Thousands of AI Authors on the Future of AI
von: Grace, Katja, et al.
Veröffentlicht: (2024)
von: Grace, Katja, et al.
Veröffentlicht: (2024)
A Longitudinal Measurement of Privacy Policy Evolution for Large Language Models
von: Tao, Zhen, et al.
Veröffentlicht: (2025)
von: Tao, Zhen, et al.
Veröffentlicht: (2025)
Towards Socially and Environmentally Responsible AI
von: Li, Pengfei, et al.
Veröffentlicht: (2024)
von: Li, Pengfei, et al.
Veröffentlicht: (2024)
What happens when generative AI models train recursively on each others' outputs?
von: Vu, Hung Anh, et al.
Veröffentlicht: (2025)
von: Vu, Hung Anh, et al.
Veröffentlicht: (2025)
Towards a Science of AI Agent Reliability
von: Rabanser, Stephan, et al.
Veröffentlicht: (2026)
von: Rabanser, Stephan, et al.
Veröffentlicht: (2026)
Confidence Calibration in Vision-Language-Action Models
von: Zollo, Thomas P, et al.
Veröffentlicht: (2025)
von: Zollo, Thomas P, et al.
Veröffentlicht: (2025)
Towards AI Transparency and Accountability: A Global Framework for Exchanging Information on AI Systems
von: Buckley, Warren, et al.
Veröffentlicht: (2023)
von: Buckley, Warren, et al.
Veröffentlicht: (2023)
Measuring What AI Systems Might Do: Towards A Measurement Science in AI
von: Voudouris, Konstantinos, et al.
Veröffentlicht: (2026)
von: Voudouris, Konstantinos, et al.
Veröffentlicht: (2026)
Testing autonomous vehicles and AI: perspectives and challenges from cybersecurity, transparency, robustness and fairness
von: Llorca, David Fernández, et al.
Veröffentlicht: (2024)
von: Llorca, David Fernández, et al.
Veröffentlicht: (2024)
Being Considerate as a Pathway Towards Pluralistic Alignment for Agentic AI
von: Alamdari, Parand A., et al.
Veröffentlicht: (2024)
von: Alamdari, Parand A., et al.
Veröffentlicht: (2024)
Generative AI Training and Copyright Law
von: Stober, Sebastian, et al.
Veröffentlicht: (2025)
von: Stober, Sebastian, et al.
Veröffentlicht: (2025)
How Effective Is Constitutional AI in Small LLMs? A Study on DeepSeek-R1 and Its Peers
von: Menke, Antonio-Gabriel Chacón, et al.
Veröffentlicht: (2025)
von: Menke, Antonio-Gabriel Chacón, et al.
Veröffentlicht: (2025)
Unsupervised Confidence Calibration for Reasoning LLMs from a Single Generation
von: Zollo, Thomas, et al.
Veröffentlicht: (2026)
von: Zollo, Thomas, et al.
Veröffentlicht: (2026)
AI and Generative AI Transforming Disaster Management: A Survey of Damage Assessment and Response Techniques
von: Raj, Aman, et al.
Veröffentlicht: (2025)
von: Raj, Aman, et al.
Veröffentlicht: (2025)
Case Studies of AI Policy Development in Africa
von: Diallo, Kadijatou, et al.
Veröffentlicht: (2024)
von: Diallo, Kadijatou, et al.
Veröffentlicht: (2024)
Towards Counterfactual Fairness-aware Domain Generalization in Changing Environments
von: Lin, Yujie, et al.
Veröffentlicht: (2023)
von: Lin, Yujie, et al.
Veröffentlicht: (2023)
Tell Me What To Learn: Generalizing Neural Memory to be Controllable in Natural Language
von: Bennett, Max S., et al.
Veröffentlicht: (2026)
von: Bennett, Max S., et al.
Veröffentlicht: (2026)
AssurAI: Experience with Constructing Korean Socio-cultural Datasets to Discover Potential Risks of Generative AI
von: Lim, Chae-Gyun, et al.
Veröffentlicht: (2025)
von: Lim, Chae-Gyun, et al.
Veröffentlicht: (2025)
Uncertain Boundaries: Multidisciplinary Approaches to Copyright Issues in Generative AI
von: Amon, Archer, et al.
Veröffentlicht: (2024)
von: Amon, Archer, et al.
Veröffentlicht: (2024)
The Case for ESM3 as a General-Purpose AI Model with Systemic Risk Under the EU AI Act
von: Qureshi, Taro, et al.
Veröffentlicht: (2026)
von: Qureshi, Taro, et al.
Veröffentlicht: (2026)
SecGenAI: Enhancing Security of Cloud-based Generative AI Applications within Australian Critical Technologies of National Interest
von: Haryanto, Christoforus Yoga, et al.
Veröffentlicht: (2024)
von: Haryanto, Christoforus Yoga, et al.
Veröffentlicht: (2024)
Impacts of Racial Bias in Historical Training Data for News AI
von: Bhargava, Rahul, et al.
Veröffentlicht: (2025)
von: Bhargava, Rahul, et al.
Veröffentlicht: (2025)
Understanding and Mitigating Risks of Generative AI in Financial Services
von: Gehrmann, Sebastian, et al.
Veröffentlicht: (2025)
von: Gehrmann, Sebastian, et al.
Veröffentlicht: (2025)
Open Shouldn't Mean Exempt: Open-Source Exceptionalism and Generative AI
von: Atkinson, David
Veröffentlicht: (2025)
von: Atkinson, David
Veröffentlicht: (2025)
Towards Urban General Intelligence: A Review and Outlook of Urban Foundation Models
von: Zhang, Weijia, et al.
Veröffentlicht: (2024)
von: Zhang, Weijia, et al.
Veröffentlicht: (2024)
On the Effectiveness and Generalization of Race Representations for Debiasing High-Stakes Decisions
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
Enhancing Team Diversity with Generative AI: A Novel Project Management Framework
von: Chan, Johnny, et al.
Veröffentlicht: (2025)
von: Chan, Johnny, et al.
Veröffentlicht: (2025)
Machine Unlearning Doesn't Do What You Think: Lessons for Generative AI Policy and Research
von: Cooper, A. Feder, et al.
Veröffentlicht: (2024)
von: Cooper, A. Feder, et al.
Veröffentlicht: (2024)
Risk Sources and Risk Management Measures in Support of Standards for General-Purpose AI Systems
von: Gipiškis, Rokas, et al.
Veröffentlicht: (2024)
von: Gipiškis, Rokas, et al.
Veröffentlicht: (2024)
Unlocking Fair Use in the Generative AI Supply Chain: A Systematized Literature Review
von: Mahuli, Amruta, et al.
Veröffentlicht: (2024)
von: Mahuli, Amruta, et al.
Veröffentlicht: (2024)
Towards AI-$45^{\circ}$ Law: A Roadmap to Trustworthy AGI
von: Yang, Chao, et al.
Veröffentlicht: (2024)
von: Yang, Chao, et al.
Veröffentlicht: (2024)
Biothreat Benchmark Generation Framework for Evaluating Frontier AI Models I: The Task-Query Architecture
von: Ackerman, Gary, et al.
Veröffentlicht: (2025)
von: Ackerman, Gary, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Test-Time Warmup for Multimodal Large Language Models
von: Rajaneesh, Nikita, et al.
Veröffentlicht: (2025) -
Adaptive Elicitation of Latent Information Using Natural Language
von: Wang, Jimmy, et al.
Veröffentlicht: (2025) -
Prompt Risk Control: A Rigorous Framework for Responsible Deployment of Large Language Models
von: Zollo, Thomas P., et al.
Veröffentlicht: (2023) -
Improving Predictor Reliability with Selective Recalibration
von: Zollo, Thomas P., et al.
Veröffentlicht: (2024) -
Whom to Query for What: Adaptive Group Elicitation via Multi-Turn LLM Interactions
von: Ding, Ruomeng, et al.
Veröffentlicht: (2026)