Skip to content
Universidad del Mar SIBUMAR Descubridor Institucional UMAR
  • Inicio
  • Búsqueda avanzada
  • Explorar
  • Login
    • English
    • Deutsch
    • Español
    • Français
    • Italiano
Advanced
  • Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Cover Image

Foundational Challenges in Assuring Alignment and Safety of Large Language Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Anwar, Usman, Saparov, Abulhair, Rando, Javier, Paleka, Daniel, Turpin, Miles, Hase, Peter, Lubana, Ekdeep Singh, Jenner, Erik, Casper, Stephen, Sourbut, Oliver, Edelman, Benjamin L., Zhang, Zhaowei, Günther, Mario, Korinek, Anton, Hernandez-Orallo, Jose, Hammond, Lewis, Bigelow, Eric, Pan, Alexander, Langosco, Lauro, Korbak, Tomasz, Zhang, Heidi, Zhong, Ruiqi, hÉigeartaigh, Seán Ó, Recchia, Gabriel, Corsi, Giulio, Chan, Alan, Anderljung, Markus, Edwards, Lilian, Petrov, Aleksandar, de Witt, Christian Schroeder, Motwan, Sumeet Ramesh, Bengio, Yoshua, Chen, Danqi, Torr, Philip H. S., Albanie, Samuel, Maharaj, Tegan, Foerster, Jakob, Tramer, Florian, He, He, Kasirzadeh, Atoosa, Choi, Yejin, Krueger, David
Format: Preprint
Published: 2024
Subjects:
Machine Learning
Artificial Intelligence
Computation and Language
Computers and Society
Online Access:
Acceder al recurso
Tags: Add Tag
No Tags, Be the first to tag this record!
  • Cite this
  • Text this
  • Email this
  • Print
  • Export Record
    • Export to RefWorks
    • Export to EndNoteWeb
    • Export to EndNote
  • Save to List
  • Permanent link
  • Holdings
  • Description
  • Comments
  • Similar Items
  • Staff View

Internet

https://arxiv.org/abs/2404.09932

Similar Items

  • Measurement challenges in AI catastrophic risk governance and safety frameworks
    by: Kasirzadeh, Atoosa
    Published: (2024)
  • Two Types of AI Existential Risk: Decisive and Accumulative
    by: Kasirzadeh, Atoosa
    Published: (2024)
  • Transformers Can Learn Connectivity in Some Graphs but Not Others
    by: Roy, Amit, et al.
    Published: (2025)
  • Language Models Might Not Understand You: Evaluating Theory of Mind via Story Prompting
    by: Getachew, Nathaniel, et al.
    Published: (2025)
  • Do Language Models Follow Occam's Razor? An Evaluation of Parsimony in Inductive and Abductive Reasoning
    by: Sun, Yunxin, et al.
    Published: (2025)
Universidad del Mar
Universidad del MarSistema Bibliotecario de la Universidad del MarDescubridor Institucional UMARImplementación y desarrollo: Mtro. Carlos Alonso Albores Pérez
InicioBúsqueda avanzadaExplorar
Visitas al Descubridor: 33,245© 2026 Universidad del Mar