Do regularization methods for shortcut mitigation work as intended?
Fuente:
arXiv
Saved in:
| Main Authors: | Hong, Haoyang, Papanikolaou, Ioanna, Parbhoo, Sonali |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tree-Based Leakage Inspection and Control in Concept Bottleneck Models
by: Ragkousis, Angelos, et al.
Published: (2024)
by: Ragkousis, Angelos, et al.
Published: (2024)
Concept-driven Off Policy Evaluation
by: Majumdar, Ritam, et al.
Published: (2024)
by: Majumdar, Ritam, et al.
Published: (2024)
Guarantee Regions for Local Explanations
by: Havasi, Marton, et al.
Published: (2024)
by: Havasi, Marton, et al.
Published: (2024)
Towards Integrating Personal Knowledge into Test-Time Predictions
by: Lage, Isaac, et al.
Published: (2024)
by: Lage, Isaac, et al.
Published: (2024)
Causal Bayesian Optimization with Unknown Graphs
by: Durand, Jean, et al.
Published: (2025)
by: Durand, Jean, et al.
Published: (2025)
Decision-Point Guided Safe Policy Improvement
by: Sharma, Abhishek, et al.
Published: (2024)
by: Sharma, Abhishek, et al.
Published: (2024)
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
by: Sharma, Abhishek, et al.
Published: (2023)
by: Sharma, Abhishek, et al.
Published: (2023)
Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories
by: Benac, Leo, et al.
Published: (2024)
by: Benac, Leo, et al.
Published: (2024)
Improving ARDS Diagnosis Through Context-Aware Concept Bottleneck Models
by: Narain, Anish, et al.
Published: (2025)
by: Narain, Anish, et al.
Published: (2025)
Learning from Failures: Understanding LLM Alignment through Failure-Aware Inverse RL
by: Patel, Nyal, et al.
Published: (2025)
by: Patel, Nyal, et al.
Published: (2025)
The Alignment Auditor: A Bayesian Framework for Verifying and Refining LLM Objectives
by: Bou, Matthieu, et al.
Published: (2025)
by: Bou, Matthieu, et al.
Published: (2025)
Feature Importance Depends on Properties of the Data: Towards Choosing the Correct Explanations for Your Data and Decision Trees based Models
by: Ayad, Célia Wafa, et al.
Published: (2025)
by: Ayad, Célia Wafa, et al.
Published: (2025)
Causal methods for LLM development and evaluation
by: Frauen, Dennis, et al.
Published: (2026)
by: Frauen, Dennis, et al.
Published: (2026)
Why LLMs Fail at Causal Discovery and How Interventional Agents Escape
by: Roy, Amartya, et al.
Published: (2026)
by: Roy, Amartya, et al.
Published: (2026)
BiMi Sheets: Infosheets for bias mitigation methods
by: Defrance, MaryBeth, et al.
Published: (2025)
by: Defrance, MaryBeth, et al.
Published: (2025)
Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities
by: Ranard, Daniel
Published: (2026)
by: Ranard, Daniel
Published: (2026)
Dynamic programming by polymorphic semiring algebraic shortcut fusion
by: Little, Max A., et al.
Published: (2021)
by: Little, Max A., et al.
Published: (2021)
Catastrophic Goodhart: regularizing RLHF with KL divergence does not mitigate heavy-tailed reward misspecification
by: Kwa, Thomas, et al.
Published: (2024)
by: Kwa, Thomas, et al.
Published: (2024)
Visual concept ranking uncovers medical shortcuts used by large multimodal models
by: Janizek, Joseph D., et al.
Published: (2026)
by: Janizek, Joseph D., et al.
Published: (2026)
Do machine learning climate models work in changing climate dynamics?
by: Navarro, Maria Conchita Agana, et al.
Published: (2025)
by: Navarro, Maria Conchita Agana, et al.
Published: (2025)
Dynamic Algorithm for Explainable k-medians Clustering under lp Norm
by: Makarychev, Konstantin, et al.
Published: (2025)
by: Makarychev, Konstantin, et al.
Published: (2025)
A kinetic-based regularization method for data science applications
by: Ganguly, Abhisek, et al.
Published: (2025)
by: Ganguly, Abhisek, et al.
Published: (2025)
Beyond Static Bias: Adaptive Multi-Fidelity Bandits with Improving Proxies
by: Lu, Muyun, et al.
Published: (2026)
by: Lu, Muyun, et al.
Published: (2026)
Skull-stripping induces shortcut learning in MRI-based Alzheimer's disease classification
by: Tinauer, Christian, et al.
Published: (2025)
by: Tinauer, Christian, et al.
Published: (2025)
Interpretable Machine Learning for Life Expectancy Prediction: A Comparative Study of Linear Regression, Decision Tree, and Random Forest
by: Dolgopolyi, Roman, et al.
Published: (2025)
by: Dolgopolyi, Roman, et al.
Published: (2025)
MLtoGAI: Semantic Web based with Machine Learning for Enhanced Disease Prediction and Personalized Recommendations using Generative AI
by: Dongre, Shyam, et al.
Published: (2024)
by: Dongre, Shyam, et al.
Published: (2024)
Accelerating PDE Data Generation via Differential Operator Action in Solution Space
by: Dong, Huanshuo, et al.
Published: (2024)
by: Dong, Huanshuo, et al.
Published: (2024)
SGD method for entropy error function with smoothing l0 regularization for neural networks
by: Nguyen, Trong-Tuan, et al.
Published: (2024)
by: Nguyen, Trong-Tuan, et al.
Published: (2024)
Transfer Operator Learning with Fusion Frame
by: Jiang, Haoyang, et al.
Published: (2024)
by: Jiang, Haoyang, et al.
Published: (2024)
Research and Implementation of Data Enhancement Techniques for Graph Neural Networks
by: Gu, Jingzhao, et al.
Published: (2024)
by: Gu, Jingzhao, et al.
Published: (2024)
Fredholm Integral Equations Neural Operator (FIE-NO) for Data-Driven Boundary Value Problems
by: Jiang, Haoyang, et al.
Published: (2024)
by: Jiang, Haoyang, et al.
Published: (2024)
Understanding and mitigating difficulties in posterior predictive evaluation
by: Agrawal, Abhinav, et al.
Published: (2024)
by: Agrawal, Abhinav, et al.
Published: (2024)
Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning
by: Smeu, Stefan, et al.
Published: (2024)
by: Smeu, Stefan, et al.
Published: (2024)
A Methodology-Oriented Study of Catastrophic Forgetting in Incremental Deep Neural Networks
by: Kumar, Ashutosh, et al.
Published: (2024)
by: Kumar, Ashutosh, et al.
Published: (2024)
Helping hampered bidders—Do subsidy auctions work as intended?
by: Sanghoon Cho, et al.
Published: (2024)
by: Sanghoon Cho, et al.
Published: (2024)
SDE approximations of GANs training and its long-run behavior
by: Cao, Haoyang, et al.
Published: (2020)
by: Cao, Haoyang, et al.
Published: (2020)
Ontology-Based Knowledge Modeling and Uncertainty-Aware Outdoor Air Quality Assessment Using Weighted Interval Type-2 Fuzzy Logic
by: Inzmam, Md, et al.
Published: (2026)
by: Inzmam, Md, et al.
Published: (2026)
Bayesian Autoregressive Online Change-Point Detection with Time-Varying Parameters
by: Tsaknaki, Ioanna-Yvonni, et al.
Published: (2024)
by: Tsaknaki, Ioanna-Yvonni, et al.
Published: (2024)
Design-Based Bandits Under Network Interference: Trade-Off Between Regret and Statistical Inference
by: Wang, Zichen, et al.
Published: (2025)
by: Wang, Zichen, et al.
Published: (2025)
When Can You Poison Rewards? A Tight Characterization of Reward Poisoning in Linear MDPs
by: Escamilla, Jose Efraim Aguilar, et al.
Published: (2026)
by: Escamilla, Jose Efraim Aguilar, et al.
Published: (2026)
Similar Items
-
Tree-Based Leakage Inspection and Control in Concept Bottleneck Models
by: Ragkousis, Angelos, et al.
Published: (2024) -
Concept-driven Off Policy Evaluation
by: Majumdar, Ritam, et al.
Published: (2024) -
Guarantee Regions for Local Explanations
by: Havasi, Marton, et al.
Published: (2024) -
Towards Integrating Personal Knowledge into Test-Time Predictions
by: Lage, Isaac, et al.
Published: (2024) -
Causal Bayesian Optimization with Unknown Graphs
by: Durand, Jean, et al.
Published: (2025)