Skip to content
Universidad del Mar SIBUMAR Descubridor Institucional UMAR
  • Inicio
  • Búsqueda avanzada
  • Explorar
  • Login
    • English
    • Deutsch
    • Español
    • Français
    • Italiano
Advanced
  • The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning
Cover Image

The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Li, Nathaniel, Pan, Alexander, Gopal, Anjali, Yue, Summer, Berrios, Daniel, Gatti, Alice, Li, Justin D., Dombrowski, Ann-Kathrin, Goel, Shashwat, Phan, Long, Mukobi, Gabriel, Helm-Burger, Nathan, Lababidi, Rassin, Justen, Lennart, Liu, Andrew B., Chen, Michael, Barrass, Isabelle, Zhang, Oliver, Zhu, Xiaoyuan, Tamirisa, Rishub, Bharathi, Bhrugu, Khoja, Adam, Zhao, Zhenqi, Herbert-Voss, Ariel, Breuer, Cort B., Marks, Samuel, Patel, Oam, Zou, Andy, Mazeika, Mantas, Wang, Zifan, Oswal, Palash, Lin, Weiran, Hunt, Adam A., Tienken-Harder, Justin, Shih, Kevin Y., Talley, Kemper, Guan, John, Kaplan, Russell, Steneker, Ian, Campbell, David, Jokubaitis, Brad, Levinson, Alex, Wang, Jean, Qian, William, Karmakar, Kallol Krishna, Basart, Steven, Fitz, Stephen, Levine, Mindy, Kumaraguru, Ponnurangam, Tupakula, Uday, Varadharajan, Vijay, Wang, Ruoyu, Shoshitaishvili, Yan, Ba, Jimmy, Esvelt, Kevin M., Wang, Alexandr, Hendrycks, Dan
Format: Preprint
Published: 2024
Subjects:
Machine Learning
Artificial Intelligence
Computation and Language
Computers and Society
Online Access:
Acceder al recurso
Tags: Add Tag
No Tags, Be the first to tag this record!
  • Cite this
  • Text this
  • Email this
  • Print
  • Export Record
    • Export to RefWorks
    • Export to EndNoteWeb
    • Export to EndNote
  • Save to List
  • Permanent link
  • Holdings
  • Description
  • Comments
  • Similar Items
  • Staff View
Description
Description not available.

Similar Items

  • Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs
    by: Mazeika, Mantas, et al.
    Published: (2025)
  • Tamper-Resistant Safeguards for Open-Weight LLMs
    by: Tamirisa, Rishub, et al.
    Published: (2024)
  • Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?
    by: Ren, Richard, et al.
    Published: (2024)
  • Foundation models may exhibit staged progression in novel CBRN threat disclosure
    by: Esvelt, Kevin M
    Published: (2025)
  • FedSelect: Customized Selection of Parameters for Fine-Tuning during Personalized Federated Learning
    by: Tamirisa, Rishub, et al.
    Published: (2023)
Universidad del Mar
Universidad del MarSistema Bibliotecario de la Universidad del MarDescubridor Institucional UMARImplementación y desarrollo: Mtro. Carlos Alonso Albores Pérez
InicioBúsqueda avanzadaExplorar
Visitas al Descubridor: 33,245© 2026 Universidad del Mar