Assessing Socio-Cultural Alignment and Technical Safety of Sovereign LLMs

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Chae, Kyubyung, Kim, Gihoon, Lee, Gyuseong, Kim, Taesup, Lee, Jaejin, Kim, Heejin
Formato: Preprint
Publicado: 2025
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866909849853362176
author Chae, Kyubyung
Kim, Gihoon
Lee, Gyuseong
Kim, Taesup
Lee, Jaejin
Kim, Heejin
author_facet Chae, Kyubyung
Kim, Gihoon
Lee, Gyuseong
Kim, Taesup
Lee, Jaejin
Kim, Heejin
contents Recent trends in LLMs development clearly show growing interest in the use and application of sovereign LLMs. The global debate over sovereign LLMs highlights the need for governments to develop their LLMs, tailored to their unique socio-cultural and historical contexts. However, there remains a shortage of frameworks and datasets to verify two critical questions: (1) how well these models align with users' socio-cultural backgrounds, and (2) whether they maintain safety and technical robustness without exposing users to potential harms and risks. To address this gap, we construct a new dataset and introduce an analytic framework for extracting and evaluating the socio-cultural elements of sovereign LLMs, alongside assessments of their technical robustness. Our experimental results demonstrate that while sovereign LLMs play a meaningful role in supporting low-resource languages, they do not always meet the popular claim that these models serve their target users well. We also show that pursuing this untested claim may lead to underestimating critical quality attributes such as safety. Our study suggests that advancing sovereign LLMs requires a more extensive evaluation that incorporates a broader range of well-grounded and practical criteria.
format Preprint
id arxiv_https___arxiv_org_abs_2510_14565
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Assessing Socio-Cultural Alignment and Technical Safety of Sovereign LLMs
Chae, Kyubyung
Kim, Gihoon
Lee, Gyuseong
Kim, Taesup
Lee, Jaejin
Kim, Heejin
Computation and Language
Recent trends in LLMs development clearly show growing interest in the use and application of sovereign LLMs. The global debate over sovereign LLMs highlights the need for governments to develop their LLMs, tailored to their unique socio-cultural and historical contexts. However, there remains a shortage of frameworks and datasets to verify two critical questions: (1) how well these models align with users' socio-cultural backgrounds, and (2) whether they maintain safety and technical robustness without exposing users to potential harms and risks. To address this gap, we construct a new dataset and introduce an analytic framework for extracting and evaluating the socio-cultural elements of sovereign LLMs, alongside assessments of their technical robustness. Our experimental results demonstrate that while sovereign LLMs play a meaningful role in supporting low-resource languages, they do not always meet the popular claim that these models serve their target users well. We also show that pursuing this untested claim may lead to underestimating critical quality attributes such as safety. Our study suggests that advancing sovereign LLMs requires a more extensive evaluation that incorporates a broader range of well-grounded and practical criteria.
title Assessing Socio-Cultural Alignment and Technical Safety of Sovereign LLMs
topic Computation and Language
url https://arxiv.org/abs/2510.14565