AdversaRiskQA: An Adversarial Factuality Benchmark for High-Risk Domains
Fuente:
arXiv
Guardado en:
| Autores principales: | Szelestey, Adam, van Engelen, Sofie, Huang, Tianhao, Snelders, Justin, Zeng, Qintao, Deng, Songgaojun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Mitigating Social Desirability Bias in Random Silicon Sampling
por: Chapala, Sashank, et al.
Publicado: (2025)
por: Chapala, Sashank, et al.
Publicado: (2025)
Benchmarking Political Persuasion Risks Across Frontier Large Language Models
por: Chen, Zhongren, et al.
Publicado: (2026)
por: Chen, Zhongren, et al.
Publicado: (2026)
Beyond Natural Language Plans: Structure-Aware Planning for Query-Focused Table Summarization
por: Zhang, Weijia, et al.
Publicado: (2025)
por: Zhang, Weijia, et al.
Publicado: (2025)
Factuality or Fiction? Benchmarking Modern LLMs on Ambiguous QA with Citations
por: Patel, Maya, et al.
Publicado: (2024)
por: Patel, Maya, et al.
Publicado: (2024)
PakBBQ: A Culturally Adapted Bias Benchmark for QA
por: Hashmat, Abdullah, et al.
Publicado: (2025)
por: Hashmat, Abdullah, et al.
Publicado: (2025)
Evaluating Proactive Risk Awareness of Large Language Models
por: Luo, Xuan, et al.
Publicado: (2026)
por: Luo, Xuan, et al.
Publicado: (2026)
GermanPartiesQA: Benchmarking Commercial Large Language Models and AI Companions for Political Alignment and Sycophancy
por: Batzner, Jan, et al.
Publicado: (2024)
por: Batzner, Jan, et al.
Publicado: (2024)
Position: It's Time to Act on the Risk of Efficient Personalized Text Generation
por: Iofinova, Eugenia, et al.
Publicado: (2025)
por: Iofinova, Eugenia, et al.
Publicado: (2025)
Quantifying Risk Propensities of Large Language Models: Ethical Focus and Bias Detection through Role-Play
por: Zeng, Yifan, et al.
Publicado: (2024)
por: Zeng, Yifan, et al.
Publicado: (2024)
The Factuality Tax of Diversity-Intervened Text-to-Image Generation: Benchmark and Fact-Augmented Intervention
por: Wan, Yixin, et al.
Publicado: (2024)
por: Wan, Yixin, et al.
Publicado: (2024)
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language
por: Sammoudi, Mohammad, et al.
Publicado: (2024)
por: Sammoudi, Mohammad, et al.
Publicado: (2024)
SimpleQA Verified: A Reliable Factuality Benchmark to Measure Parametric Knowledge
por: Haas, Lukas, et al.
Publicado: (2025)
por: Haas, Lukas, et al.
Publicado: (2025)
KoSimpleQA: A Korean Factuality Benchmark with an Analysis of Reasoning LLMs
por: Ko, Donghyeon, et al.
Publicado: (2025)
por: Ko, Donghyeon, et al.
Publicado: (2025)
Navigating the Risks of Using Large Language Models for Text Annotation in Social Science Research
por: Lin, Hao, et al.
Publicado: (2025)
por: Lin, Hao, et al.
Publicado: (2025)
Wikipedia in the Era of LLMs: Evolution and Risks
por: Huang, Siming, et al.
Publicado: (2025)
por: Huang, Siming, et al.
Publicado: (2025)
A Framework to Assess the Persuasion Risks Large Language Model Chatbots Pose to Democratic Societies
por: Chen, Zhongren, et al.
Publicado: (2025)
por: Chen, Zhongren, et al.
Publicado: (2025)
Industry Risk Assessment via Hierarchical Financial Data Using Stock Market Sentiment Indicators
por: Zhu, Hongyin
Publicado: (2023)
por: Zhu, Hongyin
Publicado: (2023)
MalAlgoQA: Pedagogical Evaluation of Counterfactual Reasoning in Large Language Models and Implications for AI in Education
por: Liu, Naiming, et al.
Publicado: (2024)
por: Liu, Naiming, et al.
Publicado: (2024)
Emergent Social Intelligence Risks in Generative Multi-Agent Systems
por: Huang, Yue, et al.
Publicado: (2026)
por: Huang, Yue, et al.
Publicado: (2026)
GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory
por: Cobben, Pepijn, et al.
Publicado: (2026)
por: Cobben, Pepijn, et al.
Publicado: (2026)
FaStfact: Faster, Stronger Long-Form Factuality Evaluations in LLMs
por: Wan, Yingjia, et al.
Publicado: (2025)
por: Wan, Yingjia, et al.
Publicado: (2025)
What do Large Language Models Say About Animals? Investigating Risks of Animal Harm in Generated Text
por: Kanepajs, Arturs, et al.
Publicado: (2025)
por: Kanepajs, Arturs, et al.
Publicado: (2025)
Quantifying the Climate Risk of Generative AI: Region-Aware Carbon Accounting with G-TRACE and the AI Sustainability Pyramid
por: Kausar, Zahida, et al.
Publicado: (2025)
por: Kausar, Zahida, et al.
Publicado: (2025)
Ensuring Safety and Trust: Analyzing the Risks of Large Language Models in Medicine
por: Yang, Yifan, et al.
Publicado: (2024)
por: Yang, Yifan, et al.
Publicado: (2024)
Red Lines and Grey Zones in the Fog of War: Benchmarking Legal Risk, Moral Harm, and Regional Bias in Large Language Model Military Decision-Making
por: Drinkall, Toby
Publicado: (2025)
por: Drinkall, Toby
Publicado: (2025)
LatentQA: Teaching LLMs to Decode Activations Into Natural Language
por: Pan, Alexander, et al.
Publicado: (2024)
por: Pan, Alexander, et al.
Publicado: (2024)
MALIBU Benchmark: Multi-Agent LLM Implicit Bias Uncovered
por: Mirza, Imran, et al.
Publicado: (2025)
por: Mirza, Imran, et al.
Publicado: (2025)
Understanding and Mitigating Risks of Generative AI in Financial Services
por: Gehrmann, Sebastian, et al.
Publicado: (2025)
por: Gehrmann, Sebastian, et al.
Publicado: (2025)
Evaluating Reliability Asymmetries in Chinese Factual Search and AI Answers
por: Liu, Geng, et al.
Publicado: (2025)
por: Liu, Geng, et al.
Publicado: (2025)
SAIF: A Comprehensive Framework for Evaluating the Risks of Generative AI in the Public Sector
por: Lee, Kyeongryul, et al.
Publicado: (2025)
por: Lee, Kyeongryul, et al.
Publicado: (2025)
Exploring Consciousness in LLMs: A Systematic Survey of Theories, Implementations, and Frontier Risks
por: Chen, Sirui, et al.
Publicado: (2025)
por: Chen, Sirui, et al.
Publicado: (2025)
Risks from Language Models for Automated Mental Healthcare: Ethics and Structure for Implementation
por: Grabb, Declan, et al.
Publicado: (2024)
por: Grabb, Declan, et al.
Publicado: (2024)
Domain-Independent Deception: A New Taxonomy and Linguistic Analysis
por: Verma, Rakesh M., et al.
Publicado: (2024)
por: Verma, Rakesh M., et al.
Publicado: (2024)
Why They Disagree: Decoding Differences in Opinions about AI Risk on the Lex Fridman Podcast
por: Truong, Nghi, et al.
Publicado: (2025)
por: Truong, Nghi, et al.
Publicado: (2025)
A Time-Aware Approach to Early Detection of Anorexia: UNSL at eRisk 2024
por: Thompson, Horacio, et al.
Publicado: (2024)
por: Thompson, Horacio, et al.
Publicado: (2024)
Generative AI in Saudi Arabia: A National Survey of Adoption, Risks, and Public Perceptions
por: AlDakheel, Abdulaziz, et al.
Publicado: (2026)
por: AlDakheel, Abdulaziz, et al.
Publicado: (2026)
An analysis of AI Decision under Risk: Prospect theory emerges in Large Language Models
por: Payne, Kenneth
Publicado: (2025)
por: Payne, Kenneth
Publicado: (2025)
LLMs as Strategic Actors: Behavioral Alignment, Risk Calibration, and Argumentation Framing in Geopolitical Simulations
por: Solopova, Veronika, et al.
Publicado: (2026)
por: Solopova, Veronika, et al.
Publicado: (2026)
AgentSense: Benchmarking Social Intelligence of Language Agents through Interactive Scenarios
por: Mou, Xinyi, et al.
Publicado: (2024)
por: Mou, Xinyi, et al.
Publicado: (2024)
A Survey on Responsible LLMs: Inherent Risk, Malicious Use, and Mitigation Strategy
por: Wang, Huandong, et al.
Publicado: (2025)
por: Wang, Huandong, et al.
Publicado: (2025)
Ejemplares similares
-
Mitigating Social Desirability Bias in Random Silicon Sampling
por: Chapala, Sashank, et al.
Publicado: (2025) -
Benchmarking Political Persuasion Risks Across Frontier Large Language Models
por: Chen, Zhongren, et al.
Publicado: (2026) -
Beyond Natural Language Plans: Structure-Aware Planning for Query-Focused Table Summarization
por: Zhang, Weijia, et al.
Publicado: (2025) -
Factuality or Fiction? Benchmarking Modern LLMs on Ambiguous QA with Citations
por: Patel, Maya, et al.
Publicado: (2024) -
PakBBQ: A Culturally Adapted Bias Benchmark for QA
por: Hashmat, Abdullah, et al.
Publicado: (2025)