Small models, big threats: Characterizing safety challenges from low-compute AI models
Fuente:
arXiv
Guardado en:
| Autor principal: | Puri, Prateek |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Measurement challenges in AI catastrophic risk governance and safety frameworks
por: Kasirzadeh, Atoosa
Publicado: (2024)
por: Kasirzadeh, Atoosa
Publicado: (2024)
The recessionary pressures of generative AI: A threat to wellbeing
por: Occhipinti, Jo-An, et al.
Publicado: (2024)
por: Occhipinti, Jo-An, et al.
Publicado: (2024)
The threat of analytic flexibility in using large language models to simulate human data
por: Cummins, Jamie
Publicado: (2025)
por: Cummins, Jamie
Publicado: (2025)
Safety challenges of AI in medicine in the era of large language models
por: Wang, Xiaoye, et al.
Publicado: (2024)
por: Wang, Xiaoye, et al.
Publicado: (2024)
Irresponsible AI: big tech's influence on AI research and associated impacts
por: Hernandez-Garcia, Alex, et al.
Publicado: (2025)
por: Hernandez-Garcia, Alex, et al.
Publicado: (2025)
Characterizing and modeling harms from interactions with design patterns in AI interfaces
por: Ibrahim, Lujain, et al.
Publicado: (2024)
por: Ibrahim, Lujain, et al.
Publicado: (2024)
AI Emergency Preparedness: Examining the federal government's ability to detect and respond to AI-related national security threats
por: Wasil, Akash, et al.
Publicado: (2024)
por: Wasil, Akash, et al.
Publicado: (2024)
Dynamic safety cases for frontier AI
por: Cârlan, Carmen, et al.
Publicado: (2024)
por: Cârlan, Carmen, et al.
Publicado: (2024)
Deep opacity and AI: A threat to XAI and to privacy protection mechanisms
por: Müller, Vincent C.
Publicado: (2025)
por: Müller, Vincent C.
Publicado: (2025)
AI threats to national security can be countered through an incident regime
por: Ortega, Alejandro
Publicado: (2025)
por: Ortega, Alejandro
Publicado: (2025)
Assessing confidence in frontier AI safety cases
por: Barrett, Stephen, et al.
Publicado: (2025)
por: Barrett, Stephen, et al.
Publicado: (2025)
Foundation models may exhibit staged progression in novel CBRN threat disclosure
por: Esvelt, Kevin M
Publicado: (2025)
por: Esvelt, Kevin M
Publicado: (2025)
Digital Lifelong Learning in the Age of AI: Trends and Insights
por: Puri, Geeta, et al.
Publicado: (2026)
por: Puri, Geeta, et al.
Publicado: (2026)
AI as a Medical Ally: Evaluating ChatGPT's Usage and Impact in Indian Healthcare
por: Raina, Aryaman, et al.
Publicado: (2024)
por: Raina, Aryaman, et al.
Publicado: (2024)
Transforming disaster risk reduction with AI and big data: Legal and interdisciplinary perspectives
por: Chun, Kwok P, et al.
Publicado: (2024)
por: Chun, Kwok P, et al.
Publicado: (2024)
Is a model equivalent to its computer implementation?
por: Hiesmayr, Beatrix C., et al.
Publicado: (2024)
por: Hiesmayr, Beatrix C., et al.
Publicado: (2024)
Including frameworks of public health ethics in computational modelling of infectious disease interventions
por: Zarebski, Alexander E., et al.
Publicado: (2025)
por: Zarebski, Alexander E., et al.
Publicado: (2025)
Third-party compliance reviews for frontier AI safety frameworks
por: Homewood, Aidan, et al.
Publicado: (2025)
por: Homewood, Aidan, et al.
Publicado: (2025)
A systematic review of research on large language models for computer programming education
por: Zhu, Meina, et al.
Publicado: (2025)
por: Zhu, Meina, et al.
Publicado: (2025)
A risk model and analysis method for the psychological safety of human and autonomous vehicles interaction
por: Sirgabsou, Yandika, et al.
Publicado: (2024)
por: Sirgabsou, Yandika, et al.
Publicado: (2024)
The rising costs of training frontier AI models
por: Cottier, Ben, et al.
Publicado: (2024)
por: Cottier, Ben, et al.
Publicado: (2024)
Fairness in AI: challenges in bridging the gap between algorithms and law
por: Giannopoulos, Giorgos, et al.
Publicado: (2024)
por: Giannopoulos, Giorgos, et al.
Publicado: (2024)
Sleeper Social Bots: a new generation of AI disinformation bots are already a political threat
por: Doshi, Jaiv, et al.
Publicado: (2024)
por: Doshi, Jaiv, et al.
Publicado: (2024)
Large Language Models Enable Design of Personalized Nudges across Cultures
por: Maksimenko, Vladimir, et al.
Publicado: (2025)
por: Maksimenko, Vladimir, et al.
Publicado: (2025)
Defining bias in AI-systems: Biased models are fair models
por: Lindloff, Chiara, et al.
Publicado: (2025)
por: Lindloff, Chiara, et al.
Publicado: (2025)
The coordination gap in frontier AI safety policies
por: Mengesha, Isaak
Publicado: (2026)
por: Mengesha, Isaak
Publicado: (2026)
A survey on fairness of large language models in e-commerce: progress, application, and challenge
por: Ren, Qingyang, et al.
Publicado: (2024)
por: Ren, Qingyang, et al.
Publicado: (2024)
Affirmative safety: An approach to risk management for high-risk AI
por: Wasil, Akash R., et al.
Publicado: (2024)
por: Wasil, Akash R., et al.
Publicado: (2024)
A cross-regional review of AI safety regulations in the commercial aviation
por: Barr, Penny A., et al.
Publicado: (2025)
por: Barr, Penny A., et al.
Publicado: (2025)
How do digital threats change requirements for the software industry?
por: Halttunen, Veikko
Publicado: (2024)
por: Halttunen, Veikko
Publicado: (2024)
A computational model for gender asset gap management with a focus on gender disparity in land acquisition and land tenure security
por: Ogundare, Oluwatosin, et al.
Publicado: (2024)
por: Ogundare, Oluwatosin, et al.
Publicado: (2024)
AI-AI Bias: large language models favor communications generated by large language models
por: Laurito, Walter, et al.
Publicado: (2024)
por: Laurito, Walter, et al.
Publicado: (2024)
Generative AI has lowered the barriers to computational social sciences
por: Zhang, Yongjun
Publicado: (2023)
por: Zhang, Yongjun
Publicado: (2023)
A semantic embedding space based on large language models for modelling human beliefs
por: Lee, Byunghwee, et al.
Publicado: (2024)
por: Lee, Byunghwee, et al.
Publicado: (2024)
Mitigating biases in big mobility data: a case study of monitoring large-scale transit systems
por: Wang, Feilong, et al.
Publicado: (2024)
por: Wang, Feilong, et al.
Publicado: (2024)
Adoption of smartphones among older adults and the role of perceived threat of cyberattacks
por: Pucer, Patrik, et al.
Publicado: (2024)
por: Pucer, Patrik, et al.
Publicado: (2024)
The potential functions of an international institution for AI safety. Insights from adjacent policy areas and recent trends
por: De Castris, A. Leone, et al.
Publicado: (2024)
por: De Castris, A. Leone, et al.
Publicado: (2024)
Characterizing AI Fact-Checkers and Their Contributions on Community Notes
por: Gong, Yilin, et al.
Publicado: (2026)
por: Gong, Yilin, et al.
Publicado: (2026)
STAMP/STPA Informed Characterization of Factors Leading to Loss of Control in AI Systems
por: Barrett, Steve, et al.
Publicado: (2025)
por: Barrett, Steve, et al.
Publicado: (2025)
Comprehensive AI governance requires addressing non-model gains
por: Goemans, Arthur, et al.
Publicado: (2026)
por: Goemans, Arthur, et al.
Publicado: (2026)
Ejemplares similares
-
Measurement challenges in AI catastrophic risk governance and safety frameworks
por: Kasirzadeh, Atoosa
Publicado: (2024) -
The recessionary pressures of generative AI: A threat to wellbeing
por: Occhipinti, Jo-An, et al.
Publicado: (2024) -
The threat of analytic flexibility in using large language models to simulate human data
por: Cummins, Jamie
Publicado: (2025) -
Safety challenges of AI in medicine in the era of large language models
por: Wang, Xiaoye, et al.
Publicado: (2024) -
Irresponsible AI: big tech's influence on AI research and associated impacts
por: Hernandez-Garcia, Alex, et al.
Publicado: (2025)