Sequential Harmful Shift Detection Without Labels
Fuente:
arXiv
Saved in:
| Main Authors: | Amoukou, Salim I., Bewley, Tom, Mishra, Saumitra, Lecue, Freddy, Magazzeni, Daniele, Veloso, Manuela |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ShapShift: Explaining Model Prediction Shifts with Subgroup Conditional Shapley Values
by: Bewley, Tom, et al.
Published: (2026)
by: Bewley, Tom, et al.
Published: (2026)
Entropic Projection Alignment: Estimating, Explaining, and Improving Model Performance Under Distribution Shift
by: Amoukou, Salim I., et al.
Published: (2026)
by: Amoukou, Salim I., et al.
Published: (2026)
Counterfactual Metarules for Local and Global Recourse
by: Bewley, Tom, et al.
Published: (2024)
by: Bewley, Tom, et al.
Published: (2024)
To Steer or Not to Steer? Mechanistic Error Reduction with Abstention for Language Models
by: Hedström, Anna, et al.
Published: (2025)
by: Hedström, Anna, et al.
Published: (2025)
Interpreting Language Reward Models via Contrastive Explanations
by: Jiang, Junqi, et al.
Published: (2024)
by: Jiang, Junqi, et al.
Published: (2024)
Progressive Inference: Explaining Decoder-Only Sequence Classification Models Using Intermediate Predictions
by: Kariyappa, Sanjay, et al.
Published: (2024)
by: Kariyappa, Sanjay, et al.
Published: (2024)
REFRESH: Responsible and Efficient Feature Reselection Guided by SHAP Values
by: Sharma, Shubham, et al.
Published: (2024)
by: Sharma, Shubham, et al.
Published: (2024)
Representation Consistency for Accurate and Coherent LLM Answer Aggregation
by: Jiang, Junqi, et al.
Published: (2025)
by: Jiang, Junqi, et al.
Published: (2025)
Are Logistic Models Really Interpretable?
by: Dervovic, Danial, et al.
Published: (2024)
by: Dervovic, Danial, et al.
Published: (2024)
Fair Wasserstein Coresets
by: Xiong, Zikai, et al.
Published: (2023)
by: Xiong, Zikai, et al.
Published: (2023)
Quantifying Prediction Consistency Under Fine-Tuning Multiplicity in Tabular LLMs
by: Hamman, Faisal, et al.
Published: (2024)
by: Hamman, Faisal, et al.
Published: (2024)
Robust Counterfactual Explanations for Neural Networks With Probabilistic Guarantees
by: Hamman, Faisal, et al.
Published: (2023)
by: Hamman, Faisal, et al.
Published: (2023)
Local and Regional Counterfactual Rules: Summarized and Robust Recourses
by: Amoukou, Salim I., et al.
Published: (2022)
by: Amoukou, Salim I., et al.
Published: (2022)
Regional Explanations: Bridging Local and Global Variable Importance
by: Amoukou, Salim I., et al.
Published: (2026)
by: Amoukou, Salim I., et al.
Published: (2026)
Accelerating Cutting-Plane Algorithms via Reinforcement Learning Surrogates
by: Mana, Kyle, et al.
Published: (2023)
by: Mana, Kyle, et al.
Published: (2023)
Adapting Prediction Sets to Distribution Shifts Without Labels
by: Kasa, Kevin, et al.
Published: (2024)
by: Kasa, Kevin, et al.
Published: (2024)
The Effect of Data Poisoning on Counterfactual Explanations
by: Artelt, André, et al.
Published: (2024)
by: Artelt, André, et al.
Published: (2024)
KnowHalu: Hallucination Detection via Multi-Form Knowledge Based Factual Checking
by: Zhang, Jiawei, et al.
Published: (2024)
by: Zhang, Jiawei, et al.
Published: (2024)
Estimating Model Performance Under Covariate Shift Without Labels
by: Białek, Jakub, et al.
Published: (2024)
by: Białek, Jakub, et al.
Published: (2024)
Zero-Shot Reinforcement Learning from Low Quality Data
by: Jeen, Scott, et al.
Published: (2023)
by: Jeen, Scott, et al.
Published: (2023)
Zero-Shot Reinforcement Learning Under Partial Observability
by: Jeen, Scott, et al.
Published: (2025)
by: Jeen, Scott, et al.
Published: (2025)
Reliably Detecting Model Failures in Deployment Without Labels
by: Nguyen, Viet, et al.
Published: (2025)
by: Nguyen, Viet, et al.
Published: (2025)
Cache What Lasts: Token Retention for Memory-Bounded KV Cache in LLMs
by: Bui, Ngoc, et al.
Published: (2025)
by: Bui, Ngoc, et al.
Published: (2025)
Achieving Fairness Without Harm via Selective Demographic Experts
by: Tan, Xuwei, et al.
Published: (2025)
by: Tan, Xuwei, et al.
Published: (2025)
Prediction-Powered Risk Monitoring of Deployed Models for Detecting Harmful Distribution Shifts
by: Zhang, Guangyi, et al.
Published: (2026)
by: Zhang, Guangyi, et al.
Published: (2026)
Scalable Representation Learning for Multimodal Tabular Transactions
by: Raman, Natraj, et al.
Published: (2024)
by: Raman, Natraj, et al.
Published: (2024)
Interpretable LLM-based Table Question Answering
by: Nguyen, Giang, et al.
Published: (2024)
by: Nguyen, Giang, et al.
Published: (2024)
The Unseen Threat: Residual Knowledge in Machine Unlearning under Perturbed Samples
by: Hsu, Hsiang, et al.
Published: (2026)
by: Hsu, Hsiang, et al.
Published: (2026)
Fairness Without Harm: An Influence-Guided Active Sampling Approach
by: Pang, Jinlong, et al.
Published: (2024)
by: Pang, Jinlong, et al.
Published: (2024)
Cross-Domain Graph Data Scaling: A Showcase with Diffusion Models
by: Tang, Wenzhuo, et al.
Published: (2024)
by: Tang, Wenzhuo, et al.
Published: (2024)
Label Alignment Regularization for Distribution Shift
by: Imani, Ehsan, et al.
Published: (2022)
by: Imani, Ehsan, et al.
Published: (2022)
Scaling Pretrained Representations Enables Label-Free Out-of-Distribution Detection Without Fine-Tuning
by: Barkley, Brett, et al.
Published: (2026)
by: Barkley, Brett, et al.
Published: (2026)
Approaching the Harm of Gradient Attacks While Only Flipping Labels
by: El-Kabid, Abdessamad, et al.
Published: (2025)
by: El-Kabid, Abdessamad, et al.
Published: (2025)
Global Graph Counterfactual Explanation: A Subgraph Mapping Approach
by: He, Yinhan, et al.
Published: (2024)
by: He, Yinhan, et al.
Published: (2024)
Who Judges the Judge? LLM Jury-on-Demand: Building Trustworthy LLM Evaluation Systems
by: Li, Xiaochuan, et al.
Published: (2025)
by: Li, Xiaochuan, et al.
Published: (2025)
Label Shift Estimation With Incremental Prior Update
by: Zhang, Yunrui, et al.
Published: (2026)
by: Zhang, Yunrui, et al.
Published: (2026)
Shift Aggregate Extract Networks
by: Orsini, Francesco, et al.
Published: (2017)
by: Orsini, Francesco, et al.
Published: (2017)
Early Stopping Against Label Noise Without Validation Data
by: Yuan, Suqin, et al.
Published: (2025)
by: Yuan, Suqin, et al.
Published: (2025)
Stochastic Mean-Shift Clustering
by: Lapidot, Itshak, et al.
Published: (2025)
by: Lapidot, Itshak, et al.
Published: (2025)
ODD: Overlap-aware Estimation of Model Performance under Distribution Shift
by: Mishra, Aayush, et al.
Published: (2025)
by: Mishra, Aayush, et al.
Published: (2025)
Similar Items
-
ShapShift: Explaining Model Prediction Shifts with Subgroup Conditional Shapley Values
by: Bewley, Tom, et al.
Published: (2026) -
Entropic Projection Alignment: Estimating, Explaining, and Improving Model Performance Under Distribution Shift
by: Amoukou, Salim I., et al.
Published: (2026) -
Counterfactual Metarules for Local and Global Recourse
by: Bewley, Tom, et al.
Published: (2024) -
To Steer or Not to Steer? Mechanistic Error Reduction with Abstention for Language Models
by: Hedström, Anna, et al.
Published: (2025) -
Interpreting Language Reward Models via Contrastive Explanations
by: Jiang, Junqi, et al.
Published: (2024)