AdversaRiskQA: An Adversarial Factuality Benchmark for High-Risk Domains
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Szelestey, Adam, van Engelen, Sofie, Huang, Tianhao, Snelders, Justin, Zeng, Qintao, Deng, Songgaojun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mitigating Social Desirability Bias in Random Silicon Sampling
von: Chapala, Sashank, et al.
Veröffentlicht: (2025)
von: Chapala, Sashank, et al.
Veröffentlicht: (2025)
Benchmarking Political Persuasion Risks Across Frontier Large Language Models
von: Chen, Zhongren, et al.
Veröffentlicht: (2026)
von: Chen, Zhongren, et al.
Veröffentlicht: (2026)
Beyond Natural Language Plans: Structure-Aware Planning for Query-Focused Table Summarization
von: Zhang, Weijia, et al.
Veröffentlicht: (2025)
von: Zhang, Weijia, et al.
Veröffentlicht: (2025)
Factuality or Fiction? Benchmarking Modern LLMs on Ambiguous QA with Citations
von: Patel, Maya, et al.
Veröffentlicht: (2024)
von: Patel, Maya, et al.
Veröffentlicht: (2024)
PakBBQ: A Culturally Adapted Bias Benchmark for QA
von: Hashmat, Abdullah, et al.
Veröffentlicht: (2025)
von: Hashmat, Abdullah, et al.
Veröffentlicht: (2025)
Evaluating Proactive Risk Awareness of Large Language Models
von: Luo, Xuan, et al.
Veröffentlicht: (2026)
von: Luo, Xuan, et al.
Veröffentlicht: (2026)
GermanPartiesQA: Benchmarking Commercial Large Language Models and AI Companions for Political Alignment and Sycophancy
von: Batzner, Jan, et al.
Veröffentlicht: (2024)
von: Batzner, Jan, et al.
Veröffentlicht: (2024)
Position: It's Time to Act on the Risk of Efficient Personalized Text Generation
von: Iofinova, Eugenia, et al.
Veröffentlicht: (2025)
von: Iofinova, Eugenia, et al.
Veröffentlicht: (2025)
Quantifying Risk Propensities of Large Language Models: Ethical Focus and Bias Detection through Role-Play
von: Zeng, Yifan, et al.
Veröffentlicht: (2024)
von: Zeng, Yifan, et al.
Veröffentlicht: (2024)
The Factuality Tax of Diversity-Intervened Text-to-Image Generation: Benchmark and Fact-Augmented Intervention
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language
von: Sammoudi, Mohammad, et al.
Veröffentlicht: (2024)
von: Sammoudi, Mohammad, et al.
Veröffentlicht: (2024)
SimpleQA Verified: A Reliable Factuality Benchmark to Measure Parametric Knowledge
von: Haas, Lukas, et al.
Veröffentlicht: (2025)
von: Haas, Lukas, et al.
Veröffentlicht: (2025)
KoSimpleQA: A Korean Factuality Benchmark with an Analysis of Reasoning LLMs
von: Ko, Donghyeon, et al.
Veröffentlicht: (2025)
von: Ko, Donghyeon, et al.
Veröffentlicht: (2025)
Navigating the Risks of Using Large Language Models for Text Annotation in Social Science Research
von: Lin, Hao, et al.
Veröffentlicht: (2025)
von: Lin, Hao, et al.
Veröffentlicht: (2025)
Wikipedia in the Era of LLMs: Evolution and Risks
von: Huang, Siming, et al.
Veröffentlicht: (2025)
von: Huang, Siming, et al.
Veröffentlicht: (2025)
A Framework to Assess the Persuasion Risks Large Language Model Chatbots Pose to Democratic Societies
von: Chen, Zhongren, et al.
Veröffentlicht: (2025)
von: Chen, Zhongren, et al.
Veröffentlicht: (2025)
Industry Risk Assessment via Hierarchical Financial Data Using Stock Market Sentiment Indicators
von: Zhu, Hongyin
Veröffentlicht: (2023)
von: Zhu, Hongyin
Veröffentlicht: (2023)
MalAlgoQA: Pedagogical Evaluation of Counterfactual Reasoning in Large Language Models and Implications for AI in Education
von: Liu, Naiming, et al.
Veröffentlicht: (2024)
von: Liu, Naiming, et al.
Veröffentlicht: (2024)
Emergent Social Intelligence Risks in Generative Multi-Agent Systems
von: Huang, Yue, et al.
Veröffentlicht: (2026)
von: Huang, Yue, et al.
Veröffentlicht: (2026)
GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory
von: Cobben, Pepijn, et al.
Veröffentlicht: (2026)
von: Cobben, Pepijn, et al.
Veröffentlicht: (2026)
FaStfact: Faster, Stronger Long-Form Factuality Evaluations in LLMs
von: Wan, Yingjia, et al.
Veröffentlicht: (2025)
von: Wan, Yingjia, et al.
Veröffentlicht: (2025)
What do Large Language Models Say About Animals? Investigating Risks of Animal Harm in Generated Text
von: Kanepajs, Arturs, et al.
Veröffentlicht: (2025)
von: Kanepajs, Arturs, et al.
Veröffentlicht: (2025)
Quantifying the Climate Risk of Generative AI: Region-Aware Carbon Accounting with G-TRACE and the AI Sustainability Pyramid
von: Kausar, Zahida, et al.
Veröffentlicht: (2025)
von: Kausar, Zahida, et al.
Veröffentlicht: (2025)
Ensuring Safety and Trust: Analyzing the Risks of Large Language Models in Medicine
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
von: Yang, Yifan, et al.
Veröffentlicht: (2024)
Red Lines and Grey Zones in the Fog of War: Benchmarking Legal Risk, Moral Harm, and Regional Bias in Large Language Model Military Decision-Making
von: Drinkall, Toby
Veröffentlicht: (2025)
von: Drinkall, Toby
Veröffentlicht: (2025)
LatentQA: Teaching LLMs to Decode Activations Into Natural Language
von: Pan, Alexander, et al.
Veröffentlicht: (2024)
von: Pan, Alexander, et al.
Veröffentlicht: (2024)
MALIBU Benchmark: Multi-Agent LLM Implicit Bias Uncovered
von: Mirza, Imran, et al.
Veröffentlicht: (2025)
von: Mirza, Imran, et al.
Veröffentlicht: (2025)
Understanding and Mitigating Risks of Generative AI in Financial Services
von: Gehrmann, Sebastian, et al.
Veröffentlicht: (2025)
von: Gehrmann, Sebastian, et al.
Veröffentlicht: (2025)
Evaluating Reliability Asymmetries in Chinese Factual Search and AI Answers
von: Liu, Geng, et al.
Veröffentlicht: (2025)
von: Liu, Geng, et al.
Veröffentlicht: (2025)
SAIF: A Comprehensive Framework for Evaluating the Risks of Generative AI in the Public Sector
von: Lee, Kyeongryul, et al.
Veröffentlicht: (2025)
von: Lee, Kyeongryul, et al.
Veröffentlicht: (2025)
Exploring Consciousness in LLMs: A Systematic Survey of Theories, Implementations, and Frontier Risks
von: Chen, Sirui, et al.
Veröffentlicht: (2025)
von: Chen, Sirui, et al.
Veröffentlicht: (2025)
Risks from Language Models for Automated Mental Healthcare: Ethics and Structure for Implementation
von: Grabb, Declan, et al.
Veröffentlicht: (2024)
von: Grabb, Declan, et al.
Veröffentlicht: (2024)
Domain-Independent Deception: A New Taxonomy and Linguistic Analysis
von: Verma, Rakesh M., et al.
Veröffentlicht: (2024)
von: Verma, Rakesh M., et al.
Veröffentlicht: (2024)
Why They Disagree: Decoding Differences in Opinions about AI Risk on the Lex Fridman Podcast
von: Truong, Nghi, et al.
Veröffentlicht: (2025)
von: Truong, Nghi, et al.
Veröffentlicht: (2025)
A Time-Aware Approach to Early Detection of Anorexia: UNSL at eRisk 2024
von: Thompson, Horacio, et al.
Veröffentlicht: (2024)
von: Thompson, Horacio, et al.
Veröffentlicht: (2024)
Generative AI in Saudi Arabia: A National Survey of Adoption, Risks, and Public Perceptions
von: AlDakheel, Abdulaziz, et al.
Veröffentlicht: (2026)
von: AlDakheel, Abdulaziz, et al.
Veröffentlicht: (2026)
An analysis of AI Decision under Risk: Prospect theory emerges in Large Language Models
von: Payne, Kenneth
Veröffentlicht: (2025)
von: Payne, Kenneth
Veröffentlicht: (2025)
LLMs as Strategic Actors: Behavioral Alignment, Risk Calibration, and Argumentation Framing in Geopolitical Simulations
von: Solopova, Veronika, et al.
Veröffentlicht: (2026)
von: Solopova, Veronika, et al.
Veröffentlicht: (2026)
AgentSense: Benchmarking Social Intelligence of Language Agents through Interactive Scenarios
von: Mou, Xinyi, et al.
Veröffentlicht: (2024)
von: Mou, Xinyi, et al.
Veröffentlicht: (2024)
A Survey on Responsible LLMs: Inherent Risk, Malicious Use, and Mitigation Strategy
von: Wang, Huandong, et al.
Veröffentlicht: (2025)
von: Wang, Huandong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Mitigating Social Desirability Bias in Random Silicon Sampling
von: Chapala, Sashank, et al.
Veröffentlicht: (2025) -
Benchmarking Political Persuasion Risks Across Frontier Large Language Models
von: Chen, Zhongren, et al.
Veröffentlicht: (2026) -
Beyond Natural Language Plans: Structure-Aware Planning for Query-Focused Table Summarization
von: Zhang, Weijia, et al.
Veröffentlicht: (2025) -
Factuality or Fiction? Benchmarking Modern LLMs on Ambiguous QA with Citations
von: Patel, Maya, et al.
Veröffentlicht: (2024) -
PakBBQ: A Culturally Adapted Bias Benchmark for QA
von: Hashmat, Abdullah, et al.
Veröffentlicht: (2025)