Equitable Evaluation via Elicitation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Du, Elbert, Dwork, Cynthia, Hu, Lunjia, McIlroy-Young, Reid, Shao, Han, Zhang, Linjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Order-Independence Without Fine Tuning
von: McIlroy-Young, Reid, et al.
Veröffentlicht: (2024)
von: McIlroy-Young, Reid, et al.
Veröffentlicht: (2024)
The Folly of AI for Age Verification
von: McIlroy-Young, Reid
Veröffentlicht: (2025)
von: McIlroy-Young, Reid
Veröffentlicht: (2025)
Designing Skill-Compatible AI: Methodologies and Frameworks in Chess
von: Hamade, Karim, et al.
Veröffentlicht: (2024)
von: Hamade, Karim, et al.
Veröffentlicht: (2024)
Order Independence With Finetuning
von: Brown, Katrina, et al.
Veröffentlicht: (2025)
von: Brown, Katrina, et al.
Veröffentlicht: (2025)
How Many Domains Suffice for Domain Generalization? A Tight Characterization via the Domain Shattering Dimension
von: Dwork, Cynthia, et al.
Veröffentlicht: (2025)
von: Dwork, Cynthia, et al.
Veröffentlicht: (2025)
Differentially Private Verification of Distribution Properties
von: Du, Elbert, et al.
Veröffentlicht: (2026)
von: Du, Elbert, et al.
Veröffentlicht: (2026)
Inducing Efficient and Equitable Professional Networks through Link Recommendations
von: Dwork, Cynthia, et al.
Veröffentlicht: (2025)
von: Dwork, Cynthia, et al.
Veröffentlicht: (2025)
Towards Environmentally Equitable AI
von: Hajiesmaili, Mohammad, et al.
Veröffentlicht: (2024)
von: Hajiesmaili, Mohammad, et al.
Veröffentlicht: (2024)
Fair Risk Control: A Generalized Framework for Calibrating Multi-group Fairness Risks
von: Zhang, Lujing, et al.
Veröffentlicht: (2024)
von: Zhang, Lujing, et al.
Veröffentlicht: (2024)
Evaluating Prompt Engineering Techniques for Accuracy and Confidence Elicitation in Medical LLMs
von: Naderi, Nariman, et al.
Veröffentlicht: (2025)
von: Naderi, Nariman, et al.
Veröffentlicht: (2025)
Enhancing Equitable Access to AI in Housing and Homelessness System of Care through Federated Learning
von: Taib, Musa, et al.
Veröffentlicht: (2024)
von: Taib, Musa, et al.
Veröffentlicht: (2024)
Transit for All: Mapping Equitable Bike2Subway Connection using Region Representation Learning
von: Namgung, Min, et al.
Veröffentlicht: (2025)
von: Namgung, Min, et al.
Veröffentlicht: (2025)
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
von: Zhou, Han, et al.
Veröffentlicht: (2024)
von: Zhou, Han, et al.
Veröffentlicht: (2024)
OPTIC-ER: A Reinforcement Learning Framework for Real-Time Emergency Response and Equitable Resource Allocation in Underserved African Communities
von: Tonwe, Mary
Veröffentlicht: (2025)
von: Tonwe, Mary
Veröffentlicht: (2025)
Interdisciplinary Expertise to Advance Equitable Explainable AI
von: Bennett, Chloe R., et al.
Veröffentlicht: (2024)
von: Bennett, Chloe R., et al.
Veröffentlicht: (2024)
Speak Easy: Eliciting Harmful Jailbreaks from LLMs with Simple Interactions
von: Chan, Yik Siu, et al.
Veröffentlicht: (2025)
von: Chan, Yik Siu, et al.
Veröffentlicht: (2025)
Towards Equitable Agile Research and Development of AI and Robotics
von: Hundt, Andrew, et al.
Veröffentlicht: (2024)
von: Hundt, Andrew, et al.
Veröffentlicht: (2024)
PropensityBench: Evaluating Latent Safety Risks in Large Language Models via an Agentic Approach
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2025)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2025)
Enhancing Safety in Reinforcement Learning with Human Feedback via Rectified Policy Optimization
von: Peng, Xiyue, et al.
Veröffentlicht: (2024)
von: Peng, Xiyue, et al.
Veröffentlicht: (2024)
Fair Machine Learning in Healthcare: A Review
von: Feng, Qizhang, et al.
Veröffentlicht: (2022)
von: Feng, Qizhang, et al.
Veröffentlicht: (2022)
Mining Social Determinants of Health for Heart Failure Patient 30-Day Readmission via Large Language Model
von: Shao, Mingchen, et al.
Veröffentlicht: (2025)
von: Shao, Mingchen, et al.
Veröffentlicht: (2025)
Attacks, Defenses and Evaluations for LLM Conversation Safety: A Survey
von: Dong, Zhichen, et al.
Veröffentlicht: (2024)
von: Dong, Zhichen, et al.
Veröffentlicht: (2024)
FFB: A Fair Fairness Benchmark for In-Processing Group Fairness Methods
von: Han, Xiaotian, et al.
Veröffentlicht: (2023)
von: Han, Xiaotian, et al.
Veröffentlicht: (2023)
FrontierScience: Evaluating AI's Ability to Perform Expert-Level Scientific Tasks
von: Wang, Miles, et al.
Veröffentlicht: (2026)
von: Wang, Miles, et al.
Veröffentlicht: (2026)
Toward Unifying Group Fairness Evaluation from a Sparsity Perspective
von: Sheng, Zhecheng, et al.
Veröffentlicht: (2025)
von: Sheng, Zhecheng, et al.
Veröffentlicht: (2025)
Mitigating Gender Bias in Depression Detection via Counterfactual Inference
von: Hu, Mingxuan, et al.
Veröffentlicht: (2025)
von: Hu, Mingxuan, et al.
Veröffentlicht: (2025)
Evaluating Gemini in an arena for learning
von: LearnLM Team, et al.
Veröffentlicht: (2025)
von: LearnLM Team, et al.
Veröffentlicht: (2025)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
MAFT: Efficient Model-Agnostic Fairness Testing for Deep Neural Networks via Zero-Order Gradient Search
von: Wang, Zhaohui, et al.
Veröffentlicht: (2024)
von: Wang, Zhaohui, et al.
Veröffentlicht: (2024)
TFB: Towards Comprehensive and Fair Benchmarking of Time Series Forecasting Methods
von: Qiu, Xiangfei, et al.
Veröffentlicht: (2024)
von: Qiu, Xiangfei, et al.
Veröffentlicht: (2024)
Learning Fair Invariant Representations under Covariate and Correlation Shifts Simultaneously
von: Li, Dong, et al.
Veröffentlicht: (2024)
von: Li, Dong, et al.
Veröffentlicht: (2024)
Towards Responsible Development of Generative AI for Education: An Evaluation-Driven Approach
von: Jurenka, Irina, et al.
Veröffentlicht: (2024)
von: Jurenka, Irina, et al.
Veröffentlicht: (2024)
Towards Robust Evaluation of Unlearning in LLMs via Data Transformations
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
Sabotage Evaluations for Frontier Models
von: Benton, Joe, et al.
Veröffentlicht: (2024)
von: Benton, Joe, et al.
Veröffentlicht: (2024)
Who Evaluates AI's Social Impacts? Mapping Coverage and Gaps in First and Third Party Evaluations
von: Reuel, Anka, et al.
Veröffentlicht: (2025)
von: Reuel, Anka, et al.
Veröffentlicht: (2025)
QG-SMS: Enhancing Test Item Analysis via Student Modeling and Simulation
von: Nguyen, Bang, et al.
Veröffentlicht: (2025)
von: Nguyen, Bang, et al.
Veröffentlicht: (2025)
Symbolic Cognitive Diagnosis via Hybrid Optimization for Intelligent Education Systems
von: Shen, Junhao, et al.
Veröffentlicht: (2023)
von: Shen, Junhao, et al.
Veröffentlicht: (2023)
Inference-Time Rule Eraser: Fair Recognition via Distilling and Removing Biased Rules
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
Safety challenges of AI in medicine in the era of large language models
von: Wang, Xiaoye, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoye, et al.
Veröffentlicht: (2024)
3DG: A Framework for Using Generative AI for Handling Sparse Learner Performance Data From Intelligent Tutoring Systems
von: Zhang, Liang, et al.
Veröffentlicht: (2024)
von: Zhang, Liang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Order-Independence Without Fine Tuning
von: McIlroy-Young, Reid, et al.
Veröffentlicht: (2024) -
The Folly of AI for Age Verification
von: McIlroy-Young, Reid
Veröffentlicht: (2025) -
Designing Skill-Compatible AI: Methodologies and Frameworks in Chess
von: Hamade, Karim, et al.
Veröffentlicht: (2024) -
Order Independence With Finetuning
von: Brown, Katrina, et al.
Veröffentlicht: (2025) -
How Many Domains Suffice for Domain Generalization? A Tight Characterization via the Domain Shattering Dimension
von: Dwork, Cynthia, et al.
Veröffentlicht: (2025)