Who Evaluates AI's Social Impacts? Mapping Coverage and Gaps in First and Third Party Evaluations
Fuente:
arXiv
Salvato in:
| Autori principali: | Reuel, Anka, Ghosh, Avijit, Chim, Jenny, Tran, Andrew, Long, Yanan, Mickel, Jennifer, Gohar, Usman, Yadav, Srishti, Ammanamanchi, Pawan Sasanka, Allaham, Mowafak, Rahmani, Hossein A., Akhtar, Mubashara, Friedrich, Felix, Scholz, Robert, Riegler, Michael Alexander, Batzner, Jan, Habba, Eliya, Saxena, Arushi, Kornilova, Anastassia, Wei, Kevin, Soni, Prajna, Mathew, Yohan, Klyman, Kevin, Sania, Jeba, Sahoo, Subramanyam, Bruvik, Olivia Beyer, Sadeghi, Pouya, Goswami, Sujata, Wang, Angelina, Jernite, Yacine, Talat, Zeerak, Biderman, Stella, Kochenderfer, Mykel, Koyejo, Sanmi, Solaiman, Irene |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Evaluating the Capabilities of LLMs for Supporting Anticipatory Impact Assessment
di: Allaham, Mowafak, et al.
Pubblicazione: (2024)
di: Allaham, Mowafak, et al.
Pubblicazione: (2024)
When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation
di: Akhtar, Mubashara, et al.
Pubblicazione: (2026)
di: Akhtar, Mubashara, et al.
Pubblicazione: (2026)
Synthetic Sources?: Auditing Generative Search Engine Citations for Evidence of AI-Generated Sources
di: Allaham, Mowafak, et al.
Pubblicazione: (2026)
di: Allaham, Mowafak, et al.
Pubblicazione: (2026)
Towards Leveraging News Media to Support Impact Assessment of AI Technologies
di: Allaham, Mowafak, et al.
Pubblicazione: (2024)
di: Allaham, Mowafak, et al.
Pubblicazione: (2024)
Global Perspectives of AI Risks and Harms: Analyzing the Negative Impacts of AI Technologies as Prioritized by News Media
di: Allaham, Mowafak, et al.
Pubblicazione: (2025)
di: Allaham, Mowafak, et al.
Pubblicazione: (2025)
Informing AI Risk Assessment with News Media: Analyzing National and Political Variation in the Coverage of AI Risks
di: Allaham, Mowafak, et al.
Pubblicazione: (2025)
di: Allaham, Mowafak, et al.
Pubblicazione: (2025)
Emergence WebVoyager: Toward Consistent and Transparent Evaluation of (Web) Agents in The Wild
di: Akkil, Deepak, et al.
Pubblicazione: (2026)
di: Akkil, Deepak, et al.
Pubblicazione: (2026)
Enhancing LLMs for Governance with Human Oversight: Evaluating and Aligning LLMs on Expert Classification of Climate Misinformation for Detecting False or Misleading Claims about Climate Change
di: Allaham, Mowafak, et al.
Pubblicazione: (2025)
di: Allaham, Mowafak, et al.
Pubblicazione: (2025)
Audit Cards: Contextualizing AI Evaluations
di: Staufer, Leon, et al.
Pubblicazione: (2025)
di: Staufer, Leon, et al.
Pubblicazione: (2025)
Analyzing And Editing Inner Mechanisms Of Backdoored Language Models
di: Lamparth, Max, et al.
Pubblicazione: (2023)
di: Lamparth, Max, et al.
Pubblicazione: (2023)
Fairness in Reinforcement Learning: A Survey
di: Reuel, Anka, et al.
Pubblicazione: (2024)
di: Reuel, Anka, et al.
Pubblicazione: (2024)
ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments
di: Lior, Gili, et al.
Pubblicazione: (2025)
di: Lior, Gili, et al.
Pubblicazione: (2025)
Acceptable Use Policies for Foundation Models
di: Klyman, Kevin
Pubblicazione: (2024)
di: Klyman, Kevin
Pubblicazione: (2024)
Generative AI Needs Adaptive Governance
di: Reuel, Anka, et al.
Pubblicazione: (2024)
di: Reuel, Anka, et al.
Pubblicazione: (2024)
JSON Whisperer: Efficient JSON Editing with LLMs
di: Duanis, Sarel, et al.
Pubblicazione: (2025)
di: Duanis, Sarel, et al.
Pubblicazione: (2025)
Beyond Benchmarks: On The False Promise of AI Regulation
di: Stanovsky, Gabriel, et al.
Pubblicazione: (2025)
di: Stanovsky, Gabriel, et al.
Pubblicazione: (2025)
PromptSuite: A Task-Agnostic Framework for Multi-Prompt Generation
di: Habba, Eliya, et al.
Pubblicazione: (2025)
di: Habba, Eliya, et al.
Pubblicazione: (2025)
From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs
di: Itzhak, Itay, et al.
Pubblicazione: (2026)
di: Itzhak, Itay, et al.
Pubblicazione: (2026)
From Symptoms to Systems: An Expert-Guided Approach to Understanding Risks of Generative AI for Eating Disorders
di: Winecoff, Amy, et al.
Pubblicazione: (2025)
di: Winecoff, Amy, et al.
Pubblicazione: (2025)
SpecEval: Evaluating Model Adherence to Behavior Specifications
di: Ahmed, Ahmed, et al.
Pubblicazione: (2025)
di: Ahmed, Ahmed, et al.
Pubblicazione: (2025)
Recommendations and Reporting Checklist for Rigorous & Transparent Human Baselines in Model Evaluations
di: Wei, Kevin L., et al.
Pubblicazione: (2025)
di: Wei, Kevin L., et al.
Pubblicazione: (2025)
DOVE: A Large-Scale Multi-Dimensional Predictions Dataset Towards Meaningful LLM Evaluation
di: Habba, Eliya, et al.
Pubblicazione: (2025)
di: Habba, Eliya, et al.
Pubblicazione: (2025)
Position Paper: Technical Research and Talent is Needed for Effective AI Governance
di: Reuel, Anka, et al.
Pubblicazione: (2024)
di: Reuel, Anka, et al.
Pubblicazione: (2024)
ScheMatiQ: From Research Question to Structured Data through Interactive Schema Discovery
di: Levy, Shahar, et al.
Pubblicazione: (2026)
di: Levy, Shahar, et al.
Pubblicazione: (2026)
Measurement to Meaning: A Validity-Centered Framework for AI Evaluation
di: Salaudeen, Olawale, et al.
Pubblicazione: (2025)
di: Salaudeen, Olawale, et al.
Pubblicazione: (2025)
Ev2R: Evaluating Evidence Retrieval in Automated Fact-Checking
di: Akhtar, Mubashara, et al.
Pubblicazione: (2024)
di: Akhtar, Mubashara, et al.
Pubblicazione: (2024)
Welfare, Improvability, and Variance: A Principal-Agent Approach to Optimal Benchmark Item Aggregation
di: Haupt, Andreas, et al.
Pubblicazione: (2026)
di: Haupt, Andreas, et al.
Pubblicazione: (2026)
Do AI Companies Make Good on Voluntary Commitments to the White House?
di: Wang, Jennifer, et al.
Pubblicazione: (2025)
di: Wang, Jennifer, et al.
Pubblicazione: (2025)
Spectral statistics and energy-gap scaling in $k-$local spin Hamiltonians
di: Dowarah, Sasanka
Pubblicazione: (2025)
di: Dowarah, Sasanka
Pubblicazione: (2025)
Beyond Release: Access Considerations for Generative AI Systems
di: Solaiman, Irene, et al.
Pubblicazione: (2025)
di: Solaiman, Irene, et al.
Pubblicazione: (2025)
Evaluating the Social Impact of Generative AI Systems in Systems and Society
di: Solaiman, Irene, et al.
Pubblicazione: (2023)
di: Solaiman, Irene, et al.
Pubblicazione: (2023)
Two New Species of Elaphoglossum (Elaphoglossaceae) from Amazonas, Venezuela
di: Mickel, John T.
Pubblicazione: (1990)
di: Mickel, John T.
Pubblicazione: (1990)
Racial/Ethnic Categories in AI and Algorithmic Fairness: Why They Matter and What They Represent
di: Mickel, Jennifer
Pubblicazione: (2024)
di: Mickel, Jennifer
Pubblicazione: (2024)
More of the Same: Persistent Representational Harms Under Increased Representation
di: Mickel, Jennifer, et al.
Pubblicazione: (2025)
di: Mickel, Jennifer, et al.
Pubblicazione: (2025)
New Tools are Needed for Tracking Adherence to AI Model Behavioral Use Clauses
di: McDuff, Daniel, et al.
Pubblicazione: (2025)
di: McDuff, Daniel, et al.
Pubblicazione: (2025)
Recourse, Repair, Reparation, & Prevention: A Stakeholder Analysis of AI Supply Chains
di: Hopkins, Aspen K., et al.
Pubblicazione: (2025)
di: Hopkins, Aspen K., et al.
Pubblicazione: (2025)
Comparing Apples to Oranges: A Taxonomy for Navigating the Global Landscape of AI Regulation
di: Alanoca, Sacha, et al.
Pubblicazione: (2025)
di: Alanoca, Sacha, et al.
Pubblicazione: (2025)
Escalation Risks from Language Models in Military and Diplomatic Decision-Making
di: Rivera, Juan-Pablo, et al.
Pubblicazione: (2024)
di: Rivera, Juan-Pablo, et al.
Pubblicazione: (2024)
BetterBench: Assessing AI Benchmarks, Uncovering Issues, and Establishing Best Practices
di: Reuel, Anka, et al.
Pubblicazione: (2024)
di: Reuel, Anka, et al.
Pubblicazione: (2024)
Lessons from the Trenches on Reproducible Evaluation of Language Models
di: Biderman, Stella, et al.
Pubblicazione: (2024)
di: Biderman, Stella, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Evaluating the Capabilities of LLMs for Supporting Anticipatory Impact Assessment
di: Allaham, Mowafak, et al.
Pubblicazione: (2024) -
When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation
di: Akhtar, Mubashara, et al.
Pubblicazione: (2026) -
Synthetic Sources?: Auditing Generative Search Engine Citations for Evidence of AI-Generated Sources
di: Allaham, Mowafak, et al.
Pubblicazione: (2026) -
Towards Leveraging News Media to Support Impact Assessment of AI Technologies
di: Allaham, Mowafak, et al.
Pubblicazione: (2024) -
Global Perspectives of AI Risks and Harms: Analyzing the Negative Impacts of AI Technologies as Prioritized by News Media
di: Allaham, Mowafak, et al.
Pubblicazione: (2025)