Risk Management for Mitigating Benchmark Failure Modes: BenchRisk
Fuente:
arXiv
Saved in:
| Main Authors: | McGregor, Sean, Lu, Victor, Tashev, Vassil, Foundjem, Armstrong, Ramasethu, Aishwarya, Zarkouei, Sadegh AlMahdi Kazemi, Knotz, Chris, Chen, Kongtao, Parrish, Alicia, Reuel, Anka, Frase, Heather |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Search-Based Multi-Trajectory Refinement for Safe C-to-Rust Translation with Large Language Models
by: Sim, HoHyun, et al.
Published: (2025)
by: Sim, HoHyun, et al.
Published: (2025)
Capturing the Effects of Quantization on Trojans in Code LLMs
by: Hussain, Aftab, et al.
Published: (2025)
by: Hussain, Aftab, et al.
Published: (2025)
Optimizing Noise for $f$-Differential Privacy via Anti-Concentration and Stochastic Dominance
by: Awan, Jordan, et al.
Published: (2023)
by: Awan, Jordan, et al.
Published: (2023)
Analyzing And Editing Inner Mechanisms Of Backdoored Language Models
by: Lamparth, Max, et al.
Published: (2023)
by: Lamparth, Max, et al.
Published: (2023)
Fairness in Reinforcement Learning: A Survey
by: Reuel, Anka, et al.
Published: (2024)
by: Reuel, Anka, et al.
Published: (2024)
Generative AI Needs Adaptive Governance
by: Reuel, Anka, et al.
Published: (2024)
by: Reuel, Anka, et al.
Published: (2024)
Lessons for Editors of AI Incidents from the AI Incident Database
by: Paeth, Kevin, et al.
Published: (2024)
by: Paeth, Kevin, et al.
Published: (2024)
Escalation Risks from Language Models in Military and Diplomatic Decision-Making
by: Rivera, Juan-Pablo, et al.
Published: (2024)
by: Rivera, Juan-Pablo, et al.
Published: (2024)
From Prompts to Propositions: A Logic-Based Lens on Student-LLM Interactions
by: Alfageeh, Ali, et al.
Published: (2025)
by: Alfageeh, Ali, et al.
Published: (2025)
BetterBench: Assessing AI Benchmarks, Uncovering Issues, and Establishing Best Practices
by: Reuel, Anka, et al.
Published: (2024)
by: Reuel, Anka, et al.
Published: (2024)
Audit Cards: Contextualizing AI Evaluations
by: Staufer, Leon, et al.
Published: (2025)
by: Staufer, Leon, et al.
Published: (2025)
Position Paper: Technical Research and Talent is Needed for Effective AI Governance
by: Reuel, Anka, et al.
Published: (2024)
by: Reuel, Anka, et al.
Published: (2024)
Can Linguistically Related Languages Guide LLM Translation in Low-Resource Settings?
by: Ramasethu, Aishwarya, et al.
Published: (2026)
by: Ramasethu, Aishwarya, et al.
Published: (2026)
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
Welfare, Improvability, and Variance: A Principal-Agent Approach to Optimal Benchmark Item Aggregation
by: Haupt, Andreas, et al.
Published: (2026)
by: Haupt, Andreas, et al.
Published: (2026)
Benchmark Inflation: Revealing LLM Performance Gaps Using Retro-Holdouts
by: Haimes, Jacob, et al.
Published: (2024)
by: Haimes, Jacob, et al.
Published: (2024)
Malicious and Unintentional Disclosure Risks in Large Language Models for Code Generation
by: Rabin, Rafiqul, et al.
Published: (2025)
by: Rabin, Rafiqul, et al.
Published: (2025)
Mean Back Relaxation for Position and Densities
by: Knotz, Gabriel, et al.
Published: (2023)
by: Knotz, Gabriel, et al.
Published: (2023)
The Future of Federal Categorical Library Programs. National Program for Libraries and Information Services Related Paper No. 17.
by: Frase, Robert W.
Published: (1975)
by: Frase, Robert W.
Published: (1975)
Procedures for Development and Access to Published Standards.
by: Frase, Robert W.
Published: (1982)
by: Frase, Robert W.
Published: (1982)
Improving the Robustness of Large Language Models for Code Tasks via Fine-tuning with Perturbed Data
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
Multi-Agent Framework for Threat Mitigation and Resilience in AI-Based Systems
by: Foundjem, Armstrong, et al.
Published: (2025)
by: Foundjem, Armstrong, et al.
Published: (2025)
Structural Anchors and Reasoning Fragility:Understanding CoT Robustness in LLM4Code
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
by: Taraghi, Mina, et al.
Published: (2024)
by: Taraghi, Mina, et al.
Published: (2024)
Dynamic Structures of Knowledge Production: Citation Rates in Hydrogen Technologies
by: Dekker, David, et al.
Published: (2025)
by: Dekker, David, et al.
Published: (2025)
Discovering Failure Modes in Vision-Language Models using RL
by: Jain, Kanishk, et al.
Published: (2026)
by: Jain, Kanishk, et al.
Published: (2026)
A criterion for extending morphisms from open subsets of smooth fibrations of algebraic varieties
by: Kanev, Vassil
Published: (2024)
by: Kanev, Vassil
Published: (2024)
Hurwitz moduli varieties parameterizing pointed covers of an algebraic curve with a fixed monodromy group
by: Kanev, Vassil
Published: (2024)
by: Kanev, Vassil
Published: (2024)
Hurwitz moduli varieties parameterizing Galois covers of an algebraic curve
by: Kanev, Vassil
Published: (2022)
by: Kanev, Vassil
Published: (2022)
Les conventions internationales du travail et la législation bulgare
by: Vassil Mratchkov
Published: (1979)
by: Vassil Mratchkov
Published: (1979)
Assessment and Risk Analysis of Nitrosamines in Sausages From Northern Iran
by: Mohammad Sadegh Allahkhah, et al.
Published: (2025)
by: Mohammad Sadegh Allahkhah, et al.
Published: (2025)
Monitoring Agentic Systems Before They're Reliable
by: Boston, Marisa Ferrara, et al.
Published: (2026)
by: Boston, Marisa Ferrara, et al.
Published: (2026)
Correlated ion stopping in dense plasmas with a temperature‐dependent plasmon pole approximation
by: Zhanerke Zakirova, et al.
Published: (2024)
by: Zhanerke Zakirova, et al.
Published: (2024)
Sharing Health Information on Facebook: Practices, Preferences, and Risk Perceptions of North American Users
by: Torabi, Sadegh, et al.
Published: (2016)
by: Torabi, Sadegh, et al.
Published: (2016)
Responsible AI in the Global Context: Maturity Model and Survey
by: Reuel, Anka, et al.
Published: (2024)
by: Reuel, Anka, et al.
Published: (2024)
RiskBench: A Scenario-based Benchmark for Risk Identification
by: Kung, Chi-Hsi, et al.
Published: (2023)
by: Kung, Chi-Hsi, et al.
Published: (2023)
A New Approach for Mixed‐Mode Fracture Assessment of Rubber‐Like Materials With Cracks
by: Mahdi Heydari‐Meybodi, et al.
Published: (2026)
by: Mahdi Heydari‐Meybodi, et al.
Published: (2026)
An empirical study of testing machine learning in the wild
by: Openja, Moses, et al.
Published: (2023)
by: Openja, Moses, et al.
Published: (2023)
EchoNav: A Perceptually Adaptive, Inclusive Navigation Concept for Real-World Environments
by: McGregor, Iain
Published: (2025)
by: McGregor, Iain
Published: (2025)
Guest Editors' Introduction eHealth and Services Computing in Healthcare
by: Carolyn McGregor
Published: (2009)
by: Carolyn McGregor
Published: (2009)
Similar Items
-
Search-Based Multi-Trajectory Refinement for Safe C-to-Rust Translation with Large Language Models
by: Sim, HoHyun, et al.
Published: (2025) -
Capturing the Effects of Quantization on Trojans in Code LLMs
by: Hussain, Aftab, et al.
Published: (2025) -
Optimizing Noise for $f$-Differential Privacy via Anti-Concentration and Stochastic Dominance
by: Awan, Jordan, et al.
Published: (2023) -
Analyzing And Editing Inner Mechanisms Of Backdoored Language Models
by: Lamparth, Max, et al.
Published: (2023) -
Fairness in Reinforcement Learning: A Survey
by: Reuel, Anka, et al.
Published: (2024)