Red Lines and Grey Zones in the Fog of War: Benchmarking Legal Risk, Moral Harm, and Regional Bias in Large Language Model Military Decision-Making
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Drinkall, Toby |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Escalation Risks from Language Models in Military and Diplomatic Decision-Making
von: Rivera, Juan-Pablo, et al.
Veröffentlicht: (2024)
von: Rivera, Juan-Pablo, et al.
Veröffentlicht: (2024)
Measuring Free-Form Decision-Making Inconsistency of Language Models in Military Crisis Simulations
von: Shrivastava, Aryan, et al.
Veröffentlicht: (2024)
von: Shrivastava, Aryan, et al.
Veröffentlicht: (2024)
When AI Navigates the Fog of War
von: Li, Ming, et al.
Veröffentlicht: (2026)
von: Li, Ming, et al.
Veröffentlicht: (2026)
ArabLegalEval: A Multitask Benchmark for Assessing Arabic Legal Knowledge in Large Language Models
von: Hijazi, Faris, et al.
Veröffentlicht: (2024)
von: Hijazi, Faris, et al.
Veröffentlicht: (2024)
AccessEval: Benchmarking Disability Bias in Large Language Models
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
von: Panda, Srikant, et al.
Veröffentlicht: (2025)
Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models
von: Dahl, Matthew, et al.
Veröffentlicht: (2024)
von: Dahl, Matthew, et al.
Veröffentlicht: (2024)
War and Peace (WarAgent): Large Language Model-based Multi-Agent Simulation of World Wars
von: Hua, Wenyue, et al.
Veröffentlicht: (2023)
von: Hua, Wenyue, et al.
Veröffentlicht: (2023)
JobFair: A Framework for Benchmarking Gender Hiring Bias in Large Language Models
von: Wang, Ze, et al.
Veröffentlicht: (2024)
von: Wang, Ze, et al.
Veröffentlicht: (2024)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
PluriHarms: Benchmarking the Full Spectrum of Human Judgments on AI Harm
von: Li, Jing-Jing, et al.
Veröffentlicht: (2026)
von: Li, Jing-Jing, et al.
Veröffentlicht: (2026)
An analysis of AI Decision under Risk: Prospect theory emerges in Large Language Models
von: Payne, Kenneth
Veröffentlicht: (2025)
von: Payne, Kenneth
Veröffentlicht: (2025)
Cross-Language Bias Examination in Large Language Models
von: Liang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Liang, Yuxuan, et al.
Veröffentlicht: (2025)
Quantifying Risk Propensities of Large Language Models: Ethical Focus and Bias Detection through Role-Play
von: Zeng, Yifan, et al.
Veröffentlicht: (2024)
von: Zeng, Yifan, et al.
Veröffentlicht: (2024)
From Individuals to Interactions: Benchmarking Gender Bias in Multimodal Large Language Models from the Lens of Social Relationship
von: Xu, Yue, et al.
Veröffentlicht: (2025)
von: Xu, Yue, et al.
Veröffentlicht: (2025)
Unequal Opportunities: Examining the Bias in Geographical Recommendations by Large Language Models
von: Dudy, Shiran, et al.
Veröffentlicht: (2025)
von: Dudy, Shiran, et al.
Veröffentlicht: (2025)
Bye-bye, Bluebook? Automating Legal Procedure with Large Language Models
von: Dahl, Matthew
Veröffentlicht: (2025)
von: Dahl, Matthew
Veröffentlicht: (2025)
WARBENCH: A Comprehensive Benchmark for Evaluating LLMs in Military Decision-Making
von: Li, Zongjie, et al.
Veröffentlicht: (2026)
von: Li, Zongjie, et al.
Veröffentlicht: (2026)
Gender Bias in Machine Translation and The Era of Large Language Models
von: Vanmassenhove, Eva
Veröffentlicht: (2024)
von: Vanmassenhove, Eva
Veröffentlicht: (2024)
Evaluation of Large Language Models in Legal Applications: Challenges, Methods, and Future Directions
von: Hu, Yiran, et al.
Veröffentlicht: (2026)
von: Hu, Yiran, et al.
Veröffentlicht: (2026)
Moral Susceptibility and Robustness under Persona Role-Play in Large Language Models
von: Costa, Davi Bastos, et al.
Veröffentlicht: (2025)
von: Costa, Davi Bastos, et al.
Veröffentlicht: (2025)
Large Language Models' Complicit Responses to Illicit Instructions across Socio-Legal Contexts
von: Wang, Xing, et al.
Veröffentlicht: (2025)
von: Wang, Xing, et al.
Veröffentlicht: (2025)
Exploring Possibilities of AI-Powered Legal Assistance in Bangladesh through Large Language Modeling
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
Addressing Moral Uncertainty using Large Language Models for Ethical Decision-Making
von: Dubey, Rohit K., et al.
Veröffentlicht: (2025)
von: Dubey, Rohit K., et al.
Veröffentlicht: (2025)
Born With a Silver Spoon? Investigating Socioeconomic Bias in Large Language Models
von: Singh, Smriti, et al.
Veröffentlicht: (2024)
von: Singh, Smriti, et al.
Veröffentlicht: (2024)
Invisible Filters: Cultural Bias in Hiring Evaluations Using Large Language Models
von: Rao, Pooja S. B., et al.
Veröffentlicht: (2025)
von: Rao, Pooja S. B., et al.
Veröffentlicht: (2025)
Widespread Gender and Pronoun Bias in Moral Judgments Across LLMs
von: Fernandes, Gustavo Lúcius, et al.
Veröffentlicht: (2026)
von: Fernandes, Gustavo Lúcius, et al.
Veröffentlicht: (2026)
"Pull or Not to Pull?'': Investigating Moral Biases in Leading Large Language Models Across Ethical Dilemmas
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
von: Ding, Junchen, et al.
Veröffentlicht: (2025)
Harmful Speech Detection by Language Models Exhibits Gender-Queer Dialect Bias
von: Dorn, Rebecca, et al.
Veröffentlicht: (2024)
von: Dorn, Rebecca, et al.
Veröffentlicht: (2024)
Position is Power: System Prompts as a Mechanism of Bias in Large Language Models (LLMs)
von: Neumann, Anna, et al.
Veröffentlicht: (2025)
von: Neumann, Anna, et al.
Veröffentlicht: (2025)
The Unequal Opportunities of Large Language Models: Revealing Demographic Bias through Job Recommendations
von: Salinas, Abel, et al.
Veröffentlicht: (2023)
von: Salinas, Abel, et al.
Veröffentlicht: (2023)
More is More: Addition Bias in Large Language Models
von: Santagata, Luca, et al.
Veröffentlicht: (2024)
von: Santagata, Luca, et al.
Veröffentlicht: (2024)
DarkBench: Benchmarking Dark Patterns in Large Language Models
von: Kran, Esben, et al.
Veröffentlicht: (2025)
von: Kran, Esben, et al.
Veröffentlicht: (2025)
RedTopic: Toward Topic-Diverse Red Teaming of Large Language Models
von: Ding, Jiale, et al.
Veröffentlicht: (2025)
von: Ding, Jiale, et al.
Veröffentlicht: (2025)
Bias and Fairness in Large Language Models: A Survey
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2023)
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2023)
A Capabilities Approach to Studying Bias and Harm in Language Technologies
von: Nigatu, Hellina Hailu, et al.
Veröffentlicht: (2024)
von: Nigatu, Hellina Hailu, et al.
Veröffentlicht: (2024)
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
von: Sakhawat, Adib, et al.
Veröffentlicht: (2026)
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
von: Morabito, Robert, et al.
Veröffentlicht: (2024)
von: Morabito, Robert, et al.
Veröffentlicht: (2024)
Language Agents as Digital Representatives in Collective Decision-Making
von: Jarrett, Daniel, et al.
Veröffentlicht: (2025)
von: Jarrett, Daniel, et al.
Veröffentlicht: (2025)
Getting in the Door: Streamlining Intake in Civil Legal Services with Large Language Models
von: Steenhuis, Quinten, et al.
Veröffentlicht: (2024)
von: Steenhuis, Quinten, et al.
Veröffentlicht: (2024)
Ensuring Safety and Trust: Analyzing the Risks of Large Language Models in Medicine
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Escalation Risks from Language Models in Military and Diplomatic Decision-Making
von: Rivera, Juan-Pablo, et al.
Veröffentlicht: (2024) -
Measuring Free-Form Decision-Making Inconsistency of Language Models in Military Crisis Simulations
von: Shrivastava, Aryan, et al.
Veröffentlicht: (2024) -
When AI Navigates the Fog of War
von: Li, Ming, et al.
Veröffentlicht: (2026) -
ArabLegalEval: A Multitask Benchmark for Assessing Arabic Legal Knowledge in Large Language Models
von: Hijazi, Faris, et al.
Veröffentlicht: (2024) -
AccessEval: Benchmarking Disability Bias in Large Language Models
von: Panda, Srikant, et al.
Veröffentlicht: (2025)