LLMs as Strategic Actors: Behavioral Alignment, Risk Calibration, and Argumentation Framing in Geopolitical Simulations
Fuente:
arXiv
Salvato in:
| Autori principali: | Solopova, Veronika, Skorik, Viktoria, Tereshchenko, Maksym, Haidun, Alina, Vykhopen, Ostap |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Online Density-Based Clustering for Real-Time Narrative Evolution Monitorin
di: Vykhopen, Ostap, et al.
Pubblicazione: (2026)
di: Vykhopen, Ostap, et al.
Pubblicazione: (2026)
Beyond Text-to-SQL: Autonomous Research-Driven Database Exploration with DAR
di: Vykhopen, Ostap, et al.
Pubblicazione: (2025)
di: Vykhopen, Ostap, et al.
Pubblicazione: (2025)
From Trust to Truth: Actionable policies for the use of AI in fact-checking in Germany and Ukraine
di: Solopova, Veronika
Pubblicazione: (2025)
di: Solopova, Veronika
Pubblicazione: (2025)
Implicit Behavioral Alignment of Language Agents in High-Stakes Crowd Simulations
di: Wang, Yunzhe, et al.
Pubblicazione: (2025)
di: Wang, Yunzhe, et al.
Pubblicazione: (2025)
Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction
di: Li, Ming, et al.
Pubblicazione: (2025)
di: Li, Ming, et al.
Pubblicazione: (2025)
Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment
di: Li, Bobo, et al.
Pubblicazione: (2026)
di: Li, Bobo, et al.
Pubblicazione: (2026)
Fluent but Foreign: Even Regional LLMs Lack Cultural Alignment
di: Agarwal, Dhruv, et al.
Pubblicazione: (2025)
di: Agarwal, Dhruv, et al.
Pubblicazione: (2025)
Are LLMs (Really) Ideological? An IRT-based Analysis and Alignment Tool for Perceived Socio-Economic Bias in LLMs
di: Wachter, Jasmin, et al.
Pubblicazione: (2025)
di: Wachter, Jasmin, et al.
Pubblicazione: (2025)
Break the Checkbox: Challenging Closed-Style Evaluations of Cultural Alignment in LLMs
di: Kabir, Mohsinul, et al.
Pubblicazione: (2025)
di: Kabir, Mohsinul, et al.
Pubblicazione: (2025)
PICACO: Pluralistic In-Context Value Alignment of LLMs via Total Correlation Optimization
di: Jiang, Han, et al.
Pubblicazione: (2025)
di: Jiang, Han, et al.
Pubblicazione: (2025)
Multilingual != Multicultural: Evaluating Gaps Between Multilingual Capabilities and Cultural Alignment in LLMs
di: Rystrøm, Jonathan, et al.
Pubblicazione: (2025)
di: Rystrøm, Jonathan, et al.
Pubblicazione: (2025)
Fair Play in the Newsroom: Actor-Based Filtering Gender Discrimination in Text Corpora
di: Urchs, Stefanie, et al.
Pubblicazione: (2025)
di: Urchs, Stefanie, et al.
Pubblicazione: (2025)
HugAgent: Benchmarking LLMs for Simulation of Individualized Human Reasoning
di: Li, Chance Jiajie, et al.
Pubblicazione: (2025)
di: Li, Chance Jiajie, et al.
Pubblicazione: (2025)
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
di: Sakhawat, Adib, et al.
Pubblicazione: (2026)
di: Sakhawat, Adib, et al.
Pubblicazione: (2026)
Self-Alignment of Large Language Models via Monopolylogue-based Social Scene Simulation
di: Pang, Xianghe, et al.
Pubblicazione: (2024)
di: Pang, Xianghe, et al.
Pubblicazione: (2024)
Mitigating Gambling-Like Risk-Taking Behaviors in Large Language Models: A Behavioral Economics Approach to AI Safety
di: Du, Y.
Pubblicazione: (2025)
di: Du, Y.
Pubblicazione: (2025)
Valence-Arousal Subspace in LLMs: Circular Emotion Geometry and Multi-Behavioral Control
di: Sun, Lihao, et al.
Pubblicazione: (2026)
di: Sun, Lihao, et al.
Pubblicazione: (2026)
Fine-tuning with Hierarchical Prompting for Robust Propaganda Classification Across Annotation Schemas
di: Stähelin, Lukas, et al.
Pubblicazione: (2026)
di: Stähelin, Lukas, et al.
Pubblicazione: (2026)
Validity Arguments For Constructed Response Scoring Using Generative Artificial Intelligence Applications
di: Casabianca, Jodi M., et al.
Pubblicazione: (2025)
di: Casabianca, Jodi M., et al.
Pubblicazione: (2025)
LLMs as Debate Partners: Utilizing Genetic Algorithms and Adversarial Search for Adaptive Arguments
di: Aryan, Prakash
Pubblicazione: (2024)
di: Aryan, Prakash
Pubblicazione: (2024)
Fine-Grained Behavior Simulation with Role-Playing Large Language Model on Social Media
di: Li, Kun, et al.
Pubblicazione: (2024)
di: Li, Kun, et al.
Pubblicazione: (2024)
From Argumentation to Deliberation: Perspectivized Stance Vectors for Fine-grained (Dis)agreement Analysis
di: Plenz, Moritz, et al.
Pubblicazione: (2025)
di: Plenz, Moritz, et al.
Pubblicazione: (2025)
ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety
di: Lee, Michael S., et al.
Pubblicazione: (2026)
di: Lee, Michael S., et al.
Pubblicazione: (2026)
Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection
di: Hu, Beizhe, et al.
Pubblicazione: (2023)
di: Hu, Beizhe, et al.
Pubblicazione: (2023)
Human vs. Machine: Behavioral Differences Between Expert Humans and Language Models in Wargame Simulations
di: Lamparth, Max, et al.
Pubblicazione: (2024)
di: Lamparth, Max, et al.
Pubblicazione: (2024)
EigenBench: A Comparative Behavioral Measure of Value Alignment
di: Chang, Jonathn, et al.
Pubblicazione: (2025)
di: Chang, Jonathn, et al.
Pubblicazione: (2025)
Wikipedia in the Era of LLMs: Evolution and Risks
di: Huang, Siming, et al.
Pubblicazione: (2025)
di: Huang, Siming, et al.
Pubblicazione: (2025)
Mechanistic Interpretability of Socio-Political Frames in Language Models
di: Asghari, Hadi, et al.
Pubblicazione: (2025)
di: Asghari, Hadi, et al.
Pubblicazione: (2025)
Societal Alignment Frameworks Can Improve LLM Alignment
di: Stańczak, Karolina, et al.
Pubblicazione: (2025)
di: Stańczak, Karolina, et al.
Pubblicazione: (2025)
SYN-DIGITS: A Synthetic Control Framework for Calibrated Digital Twin Simulation
di: Fan, Grace Jiarui, et al.
Pubblicazione: (2026)
di: Fan, Grace Jiarui, et al.
Pubblicazione: (2026)
CAMO: An Agentic Framework for Automated Causal Discovery from Micro Behaviors to Macro Emergence in LLM Agent Simulations
di: Yu, Xiangning, et al.
Pubblicazione: (2026)
di: Yu, Xiangning, et al.
Pubblicazione: (2026)
Scopes of Alignment
di: Varshney, Kush R., et al.
Pubblicazione: (2025)
di: Varshney, Kush R., et al.
Pubblicazione: (2025)
How English Print Media Frames Human-Elephant Conflicts in India
di: Punith, Bonala Sai, et al.
Pubblicazione: (2026)
di: Punith, Bonala Sai, et al.
Pubblicazione: (2026)
Uncovering Latent Arguments in Social Media Messaging by Employing LLMs-in-the-Loop Strategy
di: Islam, Tunazzina, et al.
Pubblicazione: (2024)
di: Islam, Tunazzina, et al.
Pubblicazione: (2024)
When Style Breaks Safety: Defending LLMs Against Superficial Style Alignment
di: Xiao, Yuxin, et al.
Pubblicazione: (2025)
di: Xiao, Yuxin, et al.
Pubblicazione: (2025)
FrameRef: A Framing Dataset and Simulation Testbed for Modeling Bounded Rational Information Health
di: De Lima, Victor, et al.
Pubblicazione: (2026)
di: De Lima, Victor, et al.
Pubblicazione: (2026)
The Personality Illusion: Revealing Dissociation Between Self-Reports & Behavior in LLMs
di: Han, Pengrui, et al.
Pubblicazione: (2025)
di: Han, Pengrui, et al.
Pubblicazione: (2025)
Incorporating LLMs for Large-Scale Urban Complex Mobility Simulation
di: Song, Yu-Lun, et al.
Pubblicazione: (2025)
di: Song, Yu-Lun, et al.
Pubblicazione: (2025)
Improving Alignment and Robustness with Circuit Breakers
di: Zou, Andy, et al.
Pubblicazione: (2024)
di: Zou, Andy, et al.
Pubblicazione: (2024)
Bias in the Tails: How Name-conditioned Evaluative Framing in Resume Summaries Destabilizes LLM-based Hiring
di: Nghiem, Huy, et al.
Pubblicazione: (2026)
di: Nghiem, Huy, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Online Density-Based Clustering for Real-Time Narrative Evolution Monitorin
di: Vykhopen, Ostap, et al.
Pubblicazione: (2026) -
Beyond Text-to-SQL: Autonomous Research-Driven Database Exploration with DAR
di: Vykhopen, Ostap, et al.
Pubblicazione: (2025) -
From Trust to Truth: Actionable policies for the use of AI in fact-checking in Germany and Ukraine
di: Solopova, Veronika
Pubblicazione: (2025) -
Implicit Behavioral Alignment of Language Agents in High-Stakes Crowd Simulations
di: Wang, Yunzhe, et al.
Pubblicazione: (2025) -
Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction
di: Li, Ming, et al.
Pubblicazione: (2025)