Saved in:
| Main Authors: | Ellis, Zachary, Joselowitz, Jared, Deo, Yash, He, Yajie, Kalygina, Anna, Higham, Aisling, Rahimzadeh, Mana, Jia, Yan, Habli, Ibrahim, Lim, Ernest |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2511.16544 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ASTRID -- An Automated and Scalable TRIaD for the Evaluation of RAG-based Clinical Question Answering Systems
by: Chowdhury, Mohita, et al.
Published: (2025)
by: Chowdhury, Mohita, et al.
Published: (2025)
MATRIX: Multi-Agent simulaTion fRamework for safe Interactions and conteXtual clinical conversational evaluation
by: Lim, Ernest, et al.
Published: (2025)
by: Lim, Ernest, et al.
Published: (2025)
SN-WER: Script-Normalized WER for Multi-Script Indic ASR Evaluation
by: Pattnayak, Priyaranjan
Published: (2026)
by: Pattnayak, Priyaranjan
Published: (2026)
WER We Stand: Benchmarking Urdu ASR Models
by: Arif, Samee, et al.
Published: (2024)
by: Arif, Samee, et al.
Published: (2024)
A Calibrated Memorization Index (MI) for Detecting Training Data Leakage in Generative MRI Models
by: Deo, Yash, et al.
Published: (2026)
by: Deo, Yash, et al.
Published: (2026)
Metrics that matter: Evaluating image quality metrics for medical image generation
by: Deo, Yash, et al.
Published: (2025)
by: Deo, Yash, et al.
Published: (2025)
Towards Robust Dysarthric Speech Recognition: LLM-Agent Post-ASR Correction Beyond WER
by: Zheng, Xiuwen, et al.
Published: (2026)
by: Zheng, Xiuwen, et al.
Published: (2026)
Clear, Compelling Arguments: Rethinking the Foundations of Frontier AI Safety Cases
by: Feakins, Shaun, et al.
Published: (2026)
by: Feakins, Shaun, et al.
Published: (2026)
Upstream and Downstream AI Safety: Both on the Same River?
by: McDermid, John, et al.
Published: (2024)
by: McDermid, John, et al.
Published: (2024)
Out-of-Distribution Detection for Safety Assurance of AI and Autonomous Systems
by: Hodge, Victoria J., et al.
Published: (2025)
by: Hodge, Victoria J., et al.
Published: (2025)
MEDSAGE: Enhancing Robustness of Medical Dialogue Summarization to ASR Errors with LLM-generated Synthetic Dialogues
by: Binici, Kuluhan, et al.
Published: (2024)
by: Binici, Kuluhan, et al.
Published: (2024)
Evaluating Metrics for Safety with LLM-as-Judges
by: Clegg, Kester, et al.
Published: (2025)
by: Clegg, Kester, et al.
Published: (2025)
Controlling Quadric Error Simplification with Line Quadrics
by: Hsueh‐Ti Derek Liu, et al.
Published: (2025)
by: Hsueh‐Ti Derek Liu, et al.
Published: (2025)
Insights from the Inverse: Reconstructing LLM Training Goals Through Inverse Reinforcement Learning
by: Joselowitz, Jared, et al.
Published: (2024)
by: Joselowitz, Jared, et al.
Published: (2024)
THEY SAY WE'R GETTING A DEMOCRACY
Published: (2003)
Published: (2003)
Implicit Knowledge in Unawareness Structures
by: Belardinelli, Gaia, et al.
Published: (2023)
by: Belardinelli, Gaia, et al.
Published: (2023)
Efficient Mechanisms under Unawareness
by: Pram, Kym, et al.
Published: (2025)
by: Pram, Kym, et al.
Published: (2025)
Quantifying Query Fairness Under Unawareness
by: Jaenich, Thomas, et al.
Published: (2025)
by: Jaenich, Thomas, et al.
Published: (2025)
Rationalizable Screening and Disclosure under Unawareness
by: Francetich, Alejandro, et al.
Published: (2025)
by: Francetich, Alejandro, et al.
Published: (2025)
Formal Evidence Generation for Assurance Cases for Robotic Software Models
by: Yan, Fang, et al.
Published: (2026)
by: Yan, Fang, et al.
Published: (2026)
Manifold-regularised Large-Margin $\ell_p$-SVDD for Multidimensional Time Series Anomaly Detection
by: Arashloo, Shervin Rahimzadeh
Published: (2025)
by: Arashloo, Shervin Rahimzadeh
Published: (2025)
The Impact of Legislation upon Management.
by: Higham, Norman
Published: (1979)
by: Higham, Norman
Published: (1979)
Undermining Mental Proof: How AI Can Make Cooperation Harder by Making Thinking Easier
by: Wojtowicz, Zachary, et al.
Published: (2024)
by: Wojtowicz, Zachary, et al.
Published: (2024)
Kuhn's Theorem for Games of the Extensive Form with Unawareness
by: Foo, Ki Vin, et al.
Published: (2025)
by: Foo, Ki Vin, et al.
Published: (2025)
Mechanism Design under Unawareness -- Extended Abstract
by: Pram, Kym, et al.
Published: (2025)
by: Pram, Kym, et al.
Published: (2025)
Factors Affecting the Mental Health and Wellbeing of Men in Nursing: A Systematic Review and Narrative Synthesis
by: Eric Lim, et al.
Published: (2026)
by: Eric Lim, et al.
Published: (2026)
Deceptive Diffusion: Generating Synthetic Adversarial Examples
by: Beerens, Lucas, et al.
Published: (2024)
by: Beerens, Lucas, et al.
Published: (2024)
SocialNLI: A Dialogue-Centric Social Inference Dataset
by: Deo, Akhil, et al.
Published: (2025)
by: Deo, Akhil, et al.
Published: (2025)
Advancing Gasoline Consumption Forecasting: A Novel Hybrid Model Integrating Transformers, LSTM, and CNN
by: Ranjbar, Mahmoud, et al.
Published: (2024)
by: Ranjbar, Mahmoud, et al.
Published: (2024)
The effects of multisensory exercises on balance control in people with diabetic peripheral neuropathy: a narrative review
by: Razieh Mofateh, et al.
Published: (2024)
by: Razieh Mofateh, et al.
Published: (2024)
Copper Poisoning with Emphasis on Its Clinical Manifestations and Treatment of Intoxication
by: Mehrdad Rafati Rahimzadeh, et al.
Published: (2024)
by: Mehrdad Rafati Rahimzadeh, et al.
Published: (2024)
Large Language Model Should Understand Pinyin for Chinese ASR Error Correction
by: Li, Yuang, et al.
Published: (2024)
by: Li, Yuang, et al.
Published: (2024)
What's my role? Modelling responsibility for AI-based safety-critical systems
by: Ryan, Philippa, et al.
Published: (2023)
by: Ryan, Philippa, et al.
Published: (2023)
Chapter Unawareness, or What We Do Not (Want to) Know
by: Abimbola, Seye | https://orcid.org/0000-0003-1294-3850
Published: (2026)
by: Abimbola, Seye | https://orcid.org/0000-0003-1294-3850
Published: (2026)
A comment on the interpretation of ligand binding experiments for agonists
by: James P. Higham
Published: (2025)
by: James P. Higham
Published: (2025)
PoWER: a new concept for DUNE Phase 2 FD PDS
by: Steklain, A., et al.
Published: (2025)
by: Steklain, A., et al.
Published: (2025)
Face to Face: Dialogues around Visual Impairment
by: S. Sierra Martínez, et al.
Published: (2025)
by: S. Sierra Martínez, et al.
Published: (2025)
Unaware, Unfunded and Uneducated: A Systematic Review of SME Cybersecurity
by: Junior, Carlos Rombaldo, et al.
Published: (2023)
by: Junior, Carlos Rombaldo, et al.
Published: (2023)
Reconsidering Fairness Through Unawareness From the Perspective of Model Multiplicity
by: Höltgen, Benedikt, et al.
Published: (2025)
by: Höltgen, Benedikt, et al.
Published: (2025)
Are Social Networks Watermarking Us or Are We (Unawarely) Watermarking Ourself?
by: Bertini, Flavio, et al.
Published: (2020)
by: Bertini, Flavio, et al.
Published: (2020)
Similar Items
-
ASTRID -- An Automated and Scalable TRIaD for the Evaluation of RAG-based Clinical Question Answering Systems
by: Chowdhury, Mohita, et al.
Published: (2025) -
MATRIX: Multi-Agent simulaTion fRamework for safe Interactions and conteXtual clinical conversational evaluation
by: Lim, Ernest, et al.
Published: (2025) -
SN-WER: Script-Normalized WER for Multi-Script Indic ASR Evaluation
by: Pattnayak, Priyaranjan
Published: (2026) -
WER We Stand: Benchmarking Urdu ASR Models
by: Arif, Samee, et al.
Published: (2024) -
A Calibrated Memorization Index (MI) for Detecting Training Data Leakage in Generative MRI Models
by: Deo, Yash, et al.
Published: (2026)