ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Michael S., Maurya, Yash, Rein, Drew, Herring, Bert, Nguyen, Jonathan, Song, Kyungho, Sehwag, Udari Madhushani, Cho, Jiyeon, Deshpande, Kaustubh, Jang, Yeongkyun, Joo, Jiyeon, Choi, Minn Seok, Fuelle, Evi, Knight, Christina Q, Brandifino, Joseph, Fenkell, Max |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2026)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2026)
Defensive Refusal Bias: How Safety Alignment Fails Cyber Defenders
von: Campbell, David, et al.
Veröffentlicht: (2026)
von: Campbell, David, et al.
Veröffentlicht: (2026)
In-Context Learning with Topological Information for Knowledge Graph Completion
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024)
FORTRESS: Frontier Risk Evaluation for National Security and Public Safety
von: Knight, Christina Q., et al.
Veröffentlicht: (2025)
von: Knight, Christina Q., et al.
Veröffentlicht: (2025)
AdvBDGen: Adversarially Fortified Prompt-Specific Fuzzy Backdoor Generator Against LLM Alignment
von: Pathmanathan, Pankayaraj, et al.
Veröffentlicht: (2024)
von: Pathmanathan, Pankayaraj, et al.
Veröffentlicht: (2024)
Can LLMs be Scammed? A Baseline Measurement Study
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024)
PropensityBench: Evaluating Latent Safety Risks in Large Language Models via an Agentic Approach
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2025)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2025)
LHAW: Controllable Underspecification for Long-Horizon Tasks
von: Pu, George, et al.
Veröffentlicht: (2026)
von: Pu, George, et al.
Veröffentlicht: (2026)
GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
von: Xu, Yuancheng, et al.
Veröffentlicht: (2024)
von: Xu, Yuancheng, et al.
Veröffentlicht: (2024)
Continual Learning of Domain Knowledge from Human Feedback in Text-to-SQL
von: Cook, Thomas, et al.
Veröffentlicht: (2025)
von: Cook, Thomas, et al.
Veröffentlicht: (2025)
Influence of social pension on well‐being and health of the rural elderly: the case of South Korea
von: Jiyeon An, et al.
Veröffentlicht: (2024)
von: Jiyeon An, et al.
Veröffentlicht: (2024)
Environmental, Social, and Governance ( ESG ) Research: A Systematic Review of Recent Trends (2020–2024)
von: Jiyeon Kim, et al.
Veröffentlicht: (2025)
von: Jiyeon Kim, et al.
Veröffentlicht: (2025)
A Multi-Level Visual Analytics Approach to Artist-Era Alignment in Popular Music
von: Bae, Jiyeon, et al.
Veröffentlicht: (2026)
von: Bae, Jiyeon, et al.
Veröffentlicht: (2026)
Attachment Between Nurses and Patients in Hospital Settings: Concept Analysis Using Walker and Avant's Method
von: Jiyeon Lee, et al.
Veröffentlicht: (2026)
von: Jiyeon Lee, et al.
Veröffentlicht: (2026)
Assessing differential impacts of a trade agreement using a quantile regression approach
von: Jiyeon Kim, et al.
Veröffentlicht: (2025)
von: Jiyeon Kim, et al.
Veröffentlicht: (2025)
ESG Performance Evolution in Retail: A Systematic Review and Meta‐Analysis
von: Jiyeon Kim, et al.
Veröffentlicht: (2025)
von: Jiyeon Kim, et al.
Veröffentlicht: (2025)
AgentCrypt: Advancing Privacy and (Secure) Computation in AI Agent Collaboration
von: Karthikeyan, Harish, et al.
Veröffentlicht: (2025)
von: Karthikeyan, Harish, et al.
Veröffentlicht: (2025)
Global Comparison of Codes of Ethics for Nurses: A Mixed‐Method Collective Case Study Differentiating Aspirational and Mandatory Ethics
von: Min Ji Kim, et al.
Veröffentlicht: (2024)
von: Min Ji Kim, et al.
Veröffentlicht: (2024)
Metric Design != Metric Behavior: Improving Metric Selection for the Unbiased Evaluation of Dimensionality Reduction
von: Bae, Jiyeon, et al.
Veröffentlicht: (2025)
von: Bae, Jiyeon, et al.
Veröffentlicht: (2025)
On the topology of real Lagrangians in toric symplectic manifolds
von: Brendel, Joé, et al.
Veröffentlicht: (2019)
von: Brendel, Joé, et al.
Veröffentlicht: (2019)
Comparative Analysis of Deep Learning Techniques for Load Forecasting in Power Systems Using Single‐Layer and Hybrid Models
von: Jiyeon Jang, et al.
Veröffentlicht: (2024)
von: Jiyeon Jang, et al.
Veröffentlicht: (2024)
Collab: Controlled Decoding using Mixture of Agents for LLM Alignment
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2025)
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2025)
Teaching Molecular Dynamics to a Non-Autoregressive Ionic Transport Predictor
von: Kim, Jiyeon, et al.
Veröffentlicht: (2026)
von: Kim, Jiyeon, et al.
Veröffentlicht: (2026)
Contrastive and Consistency Learning for Neural Noisy-Channel Model in Spoken Language Understanding
von: Kim, Suyoung, et al.
Veröffentlicht: (2024)
von: Kim, Suyoung, et al.
Veröffentlicht: (2024)
Semi-Supervised Neural Super-Resolution for Mesh-Based Simulations
von: Kim, Jiyeon, et al.
Veröffentlicht: (2026)
von: Kim, Jiyeon, et al.
Veröffentlicht: (2026)
NMR spectroscopic investigations of transition metal complexes in organometallic and bioinorganic chemistry
von: Jeongcheol Shin, et al.
Veröffentlicht: (2024)
von: Jeongcheol Shin, et al.
Veröffentlicht: (2024)
Gender Disparities in Interventional Pain Medicine: Representation, Leadership, and Compensation
von: Marissa Catalanotto, et al.
Veröffentlicht: (2026)
von: Marissa Catalanotto, et al.
Veröffentlicht: (2026)
Towards Automatic Evaluation for Image Transcreation
von: Khanuja, Simran, et al.
Veröffentlicht: (2024)
von: Khanuja, Simran, et al.
Veröffentlicht: (2024)
The Formation of Japan-ROK Security Relations
von: Choi, Kyungwon
Veröffentlicht: (2025)
von: Choi, Kyungwon
Veröffentlicht: (2025)
MoReBench: Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes
von: Chiu, Yu Ying, et al.
Veröffentlicht: (2025)
von: Chiu, Yu Ying, et al.
Veröffentlicht: (2025)
First-principles study on Small Polaron and Li diffusion in layered LiCoO2
von: Ahn, Seryung, et al.
Veröffentlicht: (2022)
von: Ahn, Seryung, et al.
Veröffentlicht: (2022)
Diverse Rare Sample Generation with Pretrained GANs
von: Lee, Subeen, et al.
Veröffentlicht: (2024)
von: Lee, Subeen, et al.
Veröffentlicht: (2024)
Insights Generator: Systematic Corpus-Level Trace Diagnostics for LLM Agents
von: Manglik, Akshay, et al.
Veröffentlicht: (2026)
von: Manglik, Akshay, et al.
Veröffentlicht: (2026)
O3D: Offline Data-driven Discovery and Distillation for Sequential Decision-Making with Large Language Models
von: Xiao, Yuchen, et al.
Veröffentlicht: (2023)
von: Xiao, Yuchen, et al.
Veröffentlicht: (2023)
OAM spatial demultiplexing by diffraction-based noiseless mode conversion with axicon
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
CREward: A Type-Specific Creativity Reward Model
von: Han, Jiyeon, et al.
Veröffentlicht: (2025)
von: Han, Jiyeon, et al.
Veröffentlicht: (2025)
PyGRF: An improved Python Geographical Random Forest model and case studies in public health and natural disasters
von: Sun, Kai, et al.
Veröffentlicht: (2024)
von: Sun, Kai, et al.
Veröffentlicht: (2024)
Synergistic Effect of Prolonged Oxygenation and Reactive Oxygen Species Scavenging on Diabetic Wound Healing Using an Injectable Thermoresponsive Hydrogel
von: Jiyeon Lee, et al.
Veröffentlicht: (2025)
von: Jiyeon Lee, et al.
Veröffentlicht: (2025)
Randomised Controlled Trial: Influence of Subconjunctival Anaesthesia Duration on Pain Perception During Intravitreal Injections: Response
von: Jiyeon Kim, et al.
Veröffentlicht: (2026)
von: Jiyeon Kim, et al.
Veröffentlicht: (2026)
PyGRF: An Improved Python Geographical Random Forest Model and Case Studies in Public Health and Natural Disasters
von: Kai Sun, et al.
Veröffentlicht: (2024)
von: Kai Sun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2026) -
Defensive Refusal Bias: How Safety Alignment Fails Cyber Defenders
von: Campbell, David, et al.
Veröffentlicht: (2026) -
In-Context Learning with Topological Information for Knowledge Graph Completion
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2024) -
FORTRESS: Frontier Risk Evaluation for National Security and Public Safety
von: Knight, Christina Q., et al.
Veröffentlicht: (2025) -
AdvBDGen: Adversarially Fortified Prompt-Specific Fuzzy Backdoor Generator Against LLM Alignment
von: Pathmanathan, Pankayaraj, et al.
Veröffentlicht: (2024)