The Reality of AI and Biorisk
Fuente:
arXiv
Guardado en:
| Autores principales: | Peppin, Aidan, Reuel, Anka, Casper, Stephen, Jones, Elliot, Strait, Andrew, Anwar, Usman, Agrawal, Anurag, Kapoor, Sayash, Koyejo, Sanmi, Pellat, Marie, Bommasani, Rishi, Frosst, Nick, Hooker, Sara |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Welfare, Improvability, and Variance: A Principal-Agent Approach to Optimal Benchmark Item Aggregation
por: Haupt, Andreas, et al.
Publicado: (2026)
por: Haupt, Andreas, et al.
Publicado: (2026)
Audit Cards: Contextualizing AI Evaluations
por: Staufer, Leon, et al.
Publicado: (2025)
por: Staufer, Leon, et al.
Publicado: (2025)
The 2024 Foundation Model Transparency Index
por: Bommasani, Rishi, et al.
Publicado: (2024)
por: Bommasani, Rishi, et al.
Publicado: (2024)
Analyzing And Editing Inner Mechanisms Of Backdoored Language Models
por: Lamparth, Max, et al.
Publicado: (2023)
por: Lamparth, Max, et al.
Publicado: (2023)
Fairness in Reinforcement Learning: A Survey
por: Reuel, Anka, et al.
Publicado: (2024)
por: Reuel, Anka, et al.
Publicado: (2024)
The Societal Impact of Foundation Models: Advancing Evidence-based AI Policy
por: Bommasani, Rishi
Publicado: (2025)
por: Bommasani, Rishi
Publicado: (2025)
NeurIPS should lead scientific consensus on AI policy
por: Bommasani, Rishi
Publicado: (2025)
por: Bommasani, Rishi
Publicado: (2025)
Foundation Model Transparency Reports
por: Bommasani, Rishi, et al.
Publicado: (2024)
por: Bommasani, Rishi, et al.
Publicado: (2024)
The 2025 Foundation Model Transparency Index
por: Wan, Alexander, et al.
Publicado: (2025)
por: Wan, Alexander, et al.
Publicado: (2025)
Generative AI Needs Adaptive Governance
por: Reuel, Anka, et al.
Publicado: (2024)
por: Reuel, Anka, et al.
Publicado: (2024)
Measurement to Meaning: A Validity-Centered Framework for AI Evaluation
por: Salaudeen, Olawale, et al.
Publicado: (2025)
por: Salaudeen, Olawale, et al.
Publicado: (2025)
Fantastic Bugs and Where to Find Them in AI Benchmarks
por: Truong, Sang, et al.
Publicado: (2025)
por: Truong, Sang, et al.
Publicado: (2025)
AI Cartography: Mapping the Latent Landscape of AI Benchmark Ecosystems
por: Hardy, Michael, et al.
Publicado: (2026)
por: Hardy, Michael, et al.
Publicado: (2026)
The Leaderboard Illusion
por: Singh, Shivalika, et al.
Publicado: (2025)
por: Singh, Shivalika, et al.
Publicado: (2025)
Trustworthy Social Bias Measurement
por: Bommasani, Rishi, et al.
Publicado: (2022)
por: Bommasani, Rishi, et al.
Publicado: (2022)
More than Marketing? On the Information Value of AI Benchmarks for Practitioners
por: Hardy, Amelia, et al.
Publicado: (2024)
por: Hardy, Amelia, et al.
Publicado: (2024)
Causally Inspired Regularization Enables Domain General Representations
por: Salaudeen, Olawale, et al.
Publicado: (2024)
por: Salaudeen, Olawale, et al.
Publicado: (2024)
Let's Measure Information Step-by-Step: AI-Based Evaluation Beyond Vibes
por: Robertson, Zachary, et al.
Publicado: (2025)
por: Robertson, Zachary, et al.
Publicado: (2025)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
por: Vo, Truong, et al.
Publicado: (2025)
por: Vo, Truong, et al.
Publicado: (2025)
The Limits of Inference Scaling Through Resampling
por: Stroebl, Benedikt, et al.
Publicado: (2024)
por: Stroebl, Benedikt, et al.
Publicado: (2024)
Build Agent Advocates, Not Platform Agents
por: Kapoor, Sayash, et al.
Publicado: (2025)
por: Kapoor, Sayash, et al.
Publicado: (2025)
Promises and pitfalls of artificial intelligence for legal applications
por: Kapoor, Sayash, et al.
Publicado: (2024)
por: Kapoor, Sayash, et al.
Publicado: (2024)
Toward an Evaluation Science for Generative AI Systems
por: Weidinger, Laura, et al.
Publicado: (2025)
por: Weidinger, Laura, et al.
Publicado: (2025)
A Framework for Objective-Driven Dynamical Stochastic Fields
por: Zhang, Yibo Jacky, et al.
Publicado: (2025)
por: Zhang, Yibo Jacky, et al.
Publicado: (2025)
Position Paper: Technical Research and Talent is Needed for Effective AI Governance
por: Reuel, Anka, et al.
Publicado: (2024)
por: Reuel, Anka, et al.
Publicado: (2024)
In-Context Learning of Energy Functions
por: Schaeffer, Rylan, et al.
Publicado: (2024)
por: Schaeffer, Rylan, et al.
Publicado: (2024)
Discovering Implicit Large Language Model Alignment Objectives
por: Chen, Edward, et al.
Publicado: (2026)
por: Chen, Edward, et al.
Publicado: (2026)
High-Dimensional Markov-switching Ordinary Differential Processes
por: Tsai, Katherine, et al.
Publicado: (2024)
por: Tsai, Katherine, et al.
Publicado: (2024)
Distributional Machine Unlearning via Selective Data Removal
por: Allouah, Youssef, et al.
Publicado: (2025)
por: Allouah, Youssef, et al.
Publicado: (2025)
SCENEBench: An Audio Understanding Benchmark Grounded in Assistive and Industrial Use Cases
por: Iyer, Laya, et al.
Publicado: (2026)
por: Iyer, Laya, et al.
Publicado: (2026)
HiFA: High-fidelity Text-to-3D Generation with Advanced Diffusion Guidance
por: Zhu, Junzhe, et al.
Publicado: (2023)
por: Zhu, Junzhe, et al.
Publicado: (2023)
Best Practices for Biorisk Evaluations on Open-Weight Bio-Foundation Models
por: Wei, Boyi, et al.
Publicado: (2025)
por: Wei, Boyi, et al.
Publicado: (2025)
Reasoning Models Don't Just Think Longer, They Move Differently
por: Gjølbye, Anders, et al.
Publicado: (2026)
por: Gjølbye, Anders, et al.
Publicado: (2026)
Is Backpropagation Optimal? When Synthetic Gradients Improve Sample Efficiency
por: Zhang, Yibo Jacky, et al.
Publicado: (2026)
por: Zhang, Yibo Jacky, et al.
Publicado: (2026)
The Inadequacy of Offline LLM Evaluations: A Need to Account for Personalization in Model Behavior
por: Wang, Angelina, et al.
Publicado: (2025)
por: Wang, Angelina, et al.
Publicado: (2025)
Principled Federated Domain Adaptation: Gradient Projection and Auto-Weighting
por: Jiang, Enyi, et al.
Publicado: (2023)
por: Jiang, Enyi, et al.
Publicado: (2023)
Open-World Evaluations for Measuring Frontier AI Capabilities
por: Kapoor, Sayash, et al.
Publicado: (2026)
por: Kapoor, Sayash, et al.
Publicado: (2026)
Do AI Companies Make Good on Voluntary Commitments to the White House?
por: Wang, Jennifer, et al.
Publicado: (2025)
por: Wang, Jennifer, et al.
Publicado: (2025)
Legal Alignment for Safe and Ethical AI
por: Kolt, Noam, et al.
Publicado: (2026)
por: Kolt, Noam, et al.
Publicado: (2026)
Why Do Safety Guardrails Degrade Across Languages?
por: Zhang, Max, et al.
Publicado: (2026)
por: Zhang, Max, et al.
Publicado: (2026)
Ejemplares similares
-
Welfare, Improvability, and Variance: A Principal-Agent Approach to Optimal Benchmark Item Aggregation
por: Haupt, Andreas, et al.
Publicado: (2026) -
Audit Cards: Contextualizing AI Evaluations
por: Staufer, Leon, et al.
Publicado: (2025) -
The 2024 Foundation Model Transparency Index
por: Bommasani, Rishi, et al.
Publicado: (2024) -
Analyzing And Editing Inner Mechanisms Of Backdoored Language Models
por: Lamparth, Max, et al.
Publicado: (2023) -
Fairness in Reinforcement Learning: A Survey
por: Reuel, Anka, et al.
Publicado: (2024)