ShaRP: Explaining Rankings and Preferences with Shapley Values
Fuente:
arXiv
Salvato in:
| Autori principali: | Pliatsika, Venetia, Fonseca, Joao, Akhynko, Kateryna, Shevchenko, Ivan, Stoyanovich, Julia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ShaRP: Shape-Regularized Multidimensional Projections
di: Machado, Alister, et al.
Pubblicazione: (2023)
di: Machado, Alister, et al.
Pubblicazione: (2023)
SHAP-based Explanations are Sensitive to Feature Representation
di: Hwang, Hyunseung, et al.
Pubblicazione: (2025)
di: Hwang, Hyunseung, et al.
Pubblicazione: (2025)
Making Transparency Advocates: An Educational Approach Towards Better Algorithmic Transparency in Practice
di: Bell, Andrew, et al.
Pubblicazione: (2024)
di: Bell, Andrew, et al.
Pubblicazione: (2024)
Still More Shades of Null: An Evaluation Suite for Responsible Missing Value Imputation
di: Khan, Falaah Arif, et al.
Pubblicazione: (2024)
di: Khan, Falaah Arif, et al.
Pubblicazione: (2024)
ONION: A Multi-Layered Framework for Participatory ER Design
di: Makovska, Viktoriia, et al.
Pubblicazione: (2025)
di: Makovska, Viktoriia, et al.
Pubblicazione: (2025)
VirnyFlow: A Design Space for Responsible Model Development
di: Herasymuk, Denys, et al.
Pubblicazione: (2025)
di: Herasymuk, Denys, et al.
Pubblicazione: (2025)
CREDAL: Close Reading of Data Models
di: Fletcher, George, et al.
Pubblicazione: (2025)
di: Fletcher, George, et al.
Pubblicazione: (2025)
ExplainerPFN: Towards tabular foundation models for model-free zero-shot feature importance estimations
di: Fonseca, Joao, et al.
Pubblicazione: (2026)
di: Fonseca, Joao, et al.
Pubblicazione: (2026)
A New Paradigm for Counterfactual Reasoning in Fairness and Recourse
di: Bynum, Lucius E. J., et al.
Pubblicazione: (2024)
di: Bynum, Lucius E. J., et al.
Pubblicazione: (2024)
An Epistemic and Aleatoric Decomposition of Arbitrariness to Constrain the Set of Good Models
di: Khan, Falaah Arif, et al.
Pubblicazione: (2023)
di: Khan, Falaah Arif, et al.
Pubblicazione: (2023)
ShaRP: SHAllow-LayeR Pruning for Efficient Video Large Language Models
di: Xia, Yingjie, et al.
Pubblicazione: (2025)
di: Xia, Yingjie, et al.
Pubblicazione: (2025)
Safeguarding Large Language Models in Real-time with Tunable Safety-Performance Trade-offs
di: Fonseca, Joao, et al.
Pubblicazione: (2025)
di: Fonseca, Joao, et al.
Pubblicazione: (2025)
Deep Value Benchmark: Measuring Whether Models Generalize Deep Values or Shallow Preferences
di: Ashkinaze, Joshua, et al.
Pubblicazione: (2025)
di: Ashkinaze, Joshua, et al.
Pubblicazione: (2025)
Learning the Value Systems of Societies from Preferences
di: Holgado-Sánchez, Andrés, et al.
Pubblicazione: (2025)
di: Holgado-Sánchez, Andrés, et al.
Pubblicazione: (2025)
Explaining Large Language Models Decisions Using Shapley Values
di: Mohammadi, Behnam
Pubblicazione: (2024)
di: Mohammadi, Behnam
Pubblicazione: (2024)
Growth First, Care Second? Tracing the Landscape of LLM Value Preferences in Everyday Dilemmas
di: Chen, Zhiyi, et al.
Pubblicazione: (2026)
di: Chen, Zhiyi, et al.
Pubblicazione: (2026)
Assessing High-Risk AI Systems under the EU AI Act: From Legal Requirements to Technical Verification
di: Buscemi, Alessio, et al.
Pubblicazione: (2025)
di: Buscemi, Alessio, et al.
Pubblicazione: (2025)
Explaining Drift using Shapley Values
di: Edakunni, Narayanan U., et al.
Pubblicazione: (2024)
di: Edakunni, Narayanan U., et al.
Pubblicazione: (2024)
Fairness in Algorithmic Recourse Through the Lens of Substantive Equality of Opportunity
di: Bell, Andrew, et al.
Pubblicazione: (2024)
di: Bell, Andrew, et al.
Pubblicazione: (2024)
The Political Preferences of LLMs
di: Rozado, David
Pubblicazione: (2024)
di: Rozado, David
Pubblicazione: (2024)
The Value of Gen-AI Conversations: A bottom-up Framework for AI Value Alignment
di: Motnikar, Lenart, et al.
Pubblicazione: (2025)
di: Motnikar, Lenart, et al.
Pubblicazione: (2025)
The Ethics of AI Value Chains
di: Attard-Frost, Blair, et al.
Pubblicazione: (2023)
di: Attard-Frost, Blair, et al.
Pubblicazione: (2023)
Explaining Reinforcement Learning: A Counterfactual Shapley Values Approach
di: Shi, Yiwei, et al.
Pubblicazione: (2024)
di: Shi, Yiwei, et al.
Pubblicazione: (2024)
Intrinsic Barriers to Explaining Deep Foundation Models
di: Tan, Zhen, et al.
Pubblicazione: (2025)
di: Tan, Zhen, et al.
Pubblicazione: (2025)
From Stability to Inconsistency: A Study of Moral Preferences in LLMs
di: Jotautaite, Monika, et al.
Pubblicazione: (2025)
di: Jotautaite, Monika, et al.
Pubblicazione: (2025)
Deepfakes at Face Value: Image and Authority
di: Kirkpatrick, James Ravi
Pubblicazione: (2026)
di: Kirkpatrick, James Ravi
Pubblicazione: (2026)
An Evaluation of Cultural Value Alignment in LLM
di: Sukiennik, Nicholas, et al.
Pubblicazione: (2025)
di: Sukiennik, Nicholas, et al.
Pubblicazione: (2025)
Explaining How Quantization Disparately Skews a Model
di: Bellam, Abhimanyu, et al.
Pubblicazione: (2025)
di: Bellam, Abhimanyu, et al.
Pubblicazione: (2025)
Understanding the Process of Human-AI Value Alignment
di: McKinlay, Jack, et al.
Pubblicazione: (2025)
di: McKinlay, Jack, et al.
Pubblicazione: (2025)
Prompts to Proxies: Emulating Human Preferences via a Compact LLM Ensemble
di: Wang, Bingchen, et al.
Pubblicazione: (2025)
di: Wang, Bingchen, et al.
Pubblicazione: (2025)
Misaligned by Reward: Socially Undesirable Preferences in LLMs
di: Ghazaryan, Gayane, et al.
Pubblicazione: (2026)
di: Ghazaryan, Gayane, et al.
Pubblicazione: (2026)
Dynamic Normativity: Necessary and Sufficient Conditions for Value Alignment
di: Corrêa, Nicholas Kluge
Pubblicazione: (2024)
di: Corrêa, Nicholas Kluge
Pubblicazione: (2024)
BaBE: Enhancing Fairness via Estimation of Latent Explaining Variables
di: Binkyte, Ruta, et al.
Pubblicazione: (2023)
di: Binkyte, Ruta, et al.
Pubblicazione: (2023)
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
di: Zhou, Han, et al.
Pubblicazione: (2024)
di: Zhou, Han, et al.
Pubblicazione: (2024)
Temporal Preferences in Language Models for Long-Horizon Assistance
di: Mazyaki, Ali, et al.
Pubblicazione: (2025)
di: Mazyaki, Ali, et al.
Pubblicazione: (2025)
Measuring Political Preferences in AI Systems: An Integrative Approach
di: Rozado, David
Pubblicazione: (2025)
di: Rozado, David
Pubblicazione: (2025)
LSSF: Safety Alignment for Large Language Models through Low-Rank Safety Subspace Fusion
di: Zhou, Guanghao, et al.
Pubblicazione: (2026)
di: Zhou, Guanghao, et al.
Pubblicazione: (2026)
A Refutation of Shapley Values for Explainability
di: Huang, Xuanxiang, et al.
Pubblicazione: (2023)
di: Huang, Xuanxiang, et al.
Pubblicazione: (2023)
Made-in China, Thinking in America:U.S. Values Persist in Chinese LLMs
di: Haslett, David, et al.
Pubblicazione: (2025)
di: Haslett, David, et al.
Pubblicazione: (2025)
Towards Multi-Stakeholder Evaluation of ML Models: A Crowdsourcing Study on Metric Preferences in Job-matching System
di: Yokota, Takuya, et al.
Pubblicazione: (2025)
di: Yokota, Takuya, et al.
Pubblicazione: (2025)
Documenti analoghi
-
ShaRP: Shape-Regularized Multidimensional Projections
di: Machado, Alister, et al.
Pubblicazione: (2023) -
SHAP-based Explanations are Sensitive to Feature Representation
di: Hwang, Hyunseung, et al.
Pubblicazione: (2025) -
Making Transparency Advocates: An Educational Approach Towards Better Algorithmic Transparency in Practice
di: Bell, Andrew, et al.
Pubblicazione: (2024) -
Still More Shades of Null: An Evaluation Suite for Responsible Missing Value Imputation
di: Khan, Falaah Arif, et al.
Pubblicazione: (2024) -
ONION: A Multi-Layered Framework for Participatory ER Design
di: Makovska, Viktoriia, et al.
Pubblicazione: (2025)