Robust AI Evaluation through Maximal Lotteries
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Khalaf, Hadi, Wang, Serena L., Halpern, Daniel, Shapira, Itai, Calmon, Flavio du Pin, Procaccia, Ariel D. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Pairwise Calibrated Rewards for Pluralistic Alignment
von: Halpern, Daniel, et al.
Veröffentlicht: (2025)
von: Halpern, Daniel, et al.
Veröffentlicht: (2025)
Inference-Time Reward Hacking in Large Language Models
von: Khalaf, Hadi, et al.
Veröffentlicht: (2025)
von: Khalaf, Hadi, et al.
Veröffentlicht: (2025)
Axioms for AI Alignment from Human Feedback
von: Ge, Luise, et al.
Veröffentlicht: (2024)
von: Ge, Luise, et al.
Veröffentlicht: (2024)
AI Alignment at Your Discretion
von: Buyl, Maarten, et al.
Veröffentlicht: (2025)
von: Buyl, Maarten, et al.
Veröffentlicht: (2025)
Learning Social Welfare Functions
von: Pardeshi, Kanad Shrikar, et al.
Veröffentlicht: (2024)
von: Pardeshi, Kanad Shrikar, et al.
Veröffentlicht: (2024)
Clone-Robust AI Alignment
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2025)
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2025)
How RLHF Amplifies Sycophancy
von: Shapira, Itai, et al.
Veröffentlicht: (2026)
von: Shapira, Itai, et al.
Veröffentlicht: (2026)
Incentives in Federated Learning with Heterogeneous Agents
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2025)
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2025)
Generative Social Choice
von: Fish, Sara, et al.
Veröffentlicht: (2023)
von: Fish, Sara, et al.
Veröffentlicht: (2023)
Predictive Churn with the Set of Good Models
von: Watson-Daniels, Jamelle, et al.
Veröffentlicht: (2024)
von: Watson-Daniels, Jamelle, et al.
Veröffentlicht: (2024)
Metritocracy: Representative Metrics for Lite Benchmarks
von: Procaccia, Ariel, et al.
Veröffentlicht: (2025)
von: Procaccia, Ariel, et al.
Veröffentlicht: (2025)
Attack-Aware Noise Calibration for Differential Privacy
von: Kulynych, Bogdan, et al.
Veröffentlicht: (2024)
von: Kulynych, Bogdan, et al.
Veröffentlicht: (2024)
Rigor in AI: Doing Rigorous AI Work Requires a Broader, Responsible AI-Informed Conception of Rigor
von: Olteanu, Alexandra, et al.
Veröffentlicht: (2025)
von: Olteanu, Alexandra, et al.
Veröffentlicht: (2025)
Jackpot! Alignment as a Maximal Lottery
von: Maura-Rivero, Roberto-Rafael, et al.
Veröffentlicht: (2025)
von: Maura-Rivero, Roberto-Rafael, et al.
Veröffentlicht: (2025)
In This Apportionment Lottery, the House Always Wins
von: Gölz, Paul, et al.
Veröffentlicht: (2022)
von: Gölz, Paul, et al.
Veröffentlicht: (2022)
The Hidden Cost of Waiting for Accurate Predictions
von: Shirali, Ali, et al.
Veröffentlicht: (2025)
von: Shirali, Ali, et al.
Veröffentlicht: (2025)
Honor Among Bandits: No-Regret Learning for Online Fair Division
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2024)
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2024)
Generative Social Choice: The Next Generation
von: Boehmer, Niclas, et al.
Veröffentlicht: (2025)
von: Boehmer, Niclas, et al.
Veröffentlicht: (2025)
Robust Neural Processes for Noisy Data
von: Shapira, Chen, et al.
Veröffentlicht: (2024)
von: Shapira, Chen, et al.
Veröffentlicht: (2024)
Reliability and Effectiveness of Autonomous AI Agents in Supply Chain Management
von: Long, Carol Xuan, et al.
Veröffentlicht: (2026)
von: Long, Carol Xuan, et al.
Veröffentlicht: (2026)
Regretful Decisions under Label Noise
von: Nagaraj, Sujay, et al.
Veröffentlicht: (2025)
von: Nagaraj, Sujay, et al.
Veröffentlicht: (2025)
Bias Detection Via Signaling
von: Chen, Yiling, et al.
Veröffentlicht: (2024)
von: Chen, Yiling, et al.
Veröffentlicht: (2024)
New Guarantees for Learning Revenue Maximizing Menus of Lotteries and Two-Part Tariffs
von: Balcan, Maria-Florina, et al.
Veröffentlicht: (2023)
von: Balcan, Maria-Florina, et al.
Veröffentlicht: (2023)
Policy Aggregation
von: Alamdari, Parand A., et al.
Veröffentlicht: (2024)
von: Alamdari, Parand A., et al.
Veröffentlicht: (2024)
The Proportional Veto Principle for Approval Ballots
von: Halpern, Daniel, et al.
Veröffentlicht: (2025)
von: Halpern, Daniel, et al.
Veröffentlicht: (2025)
Adaptive Contracts for Cost-Effective AI Delegation
von: Saig, Eden, et al.
Veröffentlicht: (2026)
von: Saig, Eden, et al.
Veröffentlicht: (2026)
Tight Robustness Certification Through the Convex Hull of $\ell_0$ Attacks
von: Shapira, Yuval, et al.
Veröffentlicht: (2025)
von: Shapira, Yuval, et al.
Veröffentlicht: (2025)
Aleatoric and Epistemic Discrimination: Fundamental Limits of Fairness Interventions
von: Wang, Hao, et al.
Veröffentlicht: (2023)
von: Wang, Hao, et al.
Veröffentlicht: (2023)
Fair Machine Unlearning: Data Removal while Mitigating Disparities
von: Oesterling, Alex, et al.
Veröffentlicht: (2023)
von: Oesterling, Alex, et al.
Veröffentlicht: (2023)
Selective Explanations
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2024)
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2024)
Alternates, Assemble! Selecting Optimal Alternates for Citizens' Assemblies
von: Assos, Angelos, et al.
Veröffentlicht: (2025)
von: Assos, Angelos, et al.
Veröffentlicht: (2025)
KS-Lottery: Finding Certified Lottery Tickets for Multilingual Language Models
von: Yuan, Fei, et al.
Veröffentlicht: (2024)
von: Yuan, Fei, et al.
Veröffentlicht: (2024)
Multi-Group Fairness Evaluation via Conditional Value-at-Risk Testing
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2023)
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2023)
A New Perspective on Shampoo's Preconditioner
von: Morwani, Depen, et al.
Veröffentlicht: (2024)
von: Morwani, Depen, et al.
Veröffentlicht: (2024)
Inference-Time Machine Unlearning via Gated Activation Redirection
von: Turani, Vinícius Conte, et al.
Veröffentlicht: (2026)
von: Turani, Vinícius Conte, et al.
Veröffentlicht: (2026)
Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling
von: Shapira, Eilam, et al.
Veröffentlicht: (2026)
von: Shapira, Eilam, et al.
Veröffentlicht: (2026)
A Survey of Lottery Ticket Hypothesis
von: Liu, Bohan, et al.
Veröffentlicht: (2024)
von: Liu, Bohan, et al.
Veröffentlicht: (2024)
Strategic Classification With Externalities
von: Hossain, Safwan, et al.
Veröffentlicht: (2024)
von: Hossain, Safwan, et al.
Veröffentlicht: (2024)
On the Sparsity of the Strong Lottery Ticket Hypothesis
von: Natale, Emanuele, et al.
Veröffentlicht: (2024)
von: Natale, Emanuele, et al.
Veröffentlicht: (2024)
Predicting Human Choice Between Textually Described Lotteries
von: Marantz, Eyal, et al.
Veröffentlicht: (2025)
von: Marantz, Eyal, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Pairwise Calibrated Rewards for Pluralistic Alignment
von: Halpern, Daniel, et al.
Veröffentlicht: (2025) -
Inference-Time Reward Hacking in Large Language Models
von: Khalaf, Hadi, et al.
Veröffentlicht: (2025) -
Axioms for AI Alignment from Human Feedback
von: Ge, Luise, et al.
Veröffentlicht: (2024) -
AI Alignment at Your Discretion
von: Buyl, Maarten, et al.
Veröffentlicht: (2025) -
Learning Social Welfare Functions
von: Pardeshi, Kanad Shrikar, et al.
Veröffentlicht: (2024)