Do regularization methods for shortcut mitigation work as intended?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hong, Haoyang, Papanikolaou, Ioanna, Parbhoo, Sonali |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tree-Based Leakage Inspection and Control in Concept Bottleneck Models
von: Ragkousis, Angelos, et al.
Veröffentlicht: (2024)
von: Ragkousis, Angelos, et al.
Veröffentlicht: (2024)
Concept-driven Off Policy Evaluation
von: Majumdar, Ritam, et al.
Veröffentlicht: (2024)
von: Majumdar, Ritam, et al.
Veröffentlicht: (2024)
Guarantee Regions for Local Explanations
von: Havasi, Marton, et al.
Veröffentlicht: (2024)
von: Havasi, Marton, et al.
Veröffentlicht: (2024)
Towards Integrating Personal Knowledge into Test-Time Predictions
von: Lage, Isaac, et al.
Veröffentlicht: (2024)
von: Lage, Isaac, et al.
Veröffentlicht: (2024)
Causal Bayesian Optimization with Unknown Graphs
von: Durand, Jean, et al.
Veröffentlicht: (2025)
von: Durand, Jean, et al.
Veröffentlicht: (2025)
Decision-Point Guided Safe Policy Improvement
von: Sharma, Abhishek, et al.
Veröffentlicht: (2024)
von: Sharma, Abhishek, et al.
Veröffentlicht: (2024)
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
von: Sharma, Abhishek, et al.
Veröffentlicht: (2023)
von: Sharma, Abhishek, et al.
Veröffentlicht: (2023)
Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories
von: Benac, Leo, et al.
Veröffentlicht: (2024)
von: Benac, Leo, et al.
Veröffentlicht: (2024)
Improving ARDS Diagnosis Through Context-Aware Concept Bottleneck Models
von: Narain, Anish, et al.
Veröffentlicht: (2025)
von: Narain, Anish, et al.
Veröffentlicht: (2025)
Learning from Failures: Understanding LLM Alignment through Failure-Aware Inverse RL
von: Patel, Nyal, et al.
Veröffentlicht: (2025)
von: Patel, Nyal, et al.
Veröffentlicht: (2025)
The Alignment Auditor: A Bayesian Framework for Verifying and Refining LLM Objectives
von: Bou, Matthieu, et al.
Veröffentlicht: (2025)
von: Bou, Matthieu, et al.
Veröffentlicht: (2025)
Feature Importance Depends on Properties of the Data: Towards Choosing the Correct Explanations for Your Data and Decision Trees based Models
von: Ayad, Célia Wafa, et al.
Veröffentlicht: (2025)
von: Ayad, Célia Wafa, et al.
Veröffentlicht: (2025)
Causal methods for LLM development and evaluation
von: Frauen, Dennis, et al.
Veröffentlicht: (2026)
von: Frauen, Dennis, et al.
Veröffentlicht: (2026)
Why LLMs Fail at Causal Discovery and How Interventional Agents Escape
von: Roy, Amartya, et al.
Veröffentlicht: (2026)
von: Roy, Amartya, et al.
Veröffentlicht: (2026)
BiMi Sheets: Infosheets for bias mitigation methods
von: Defrance, MaryBeth, et al.
Veröffentlicht: (2025)
von: Defrance, MaryBeth, et al.
Veröffentlicht: (2025)
Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities
von: Ranard, Daniel
Veröffentlicht: (2026)
von: Ranard, Daniel
Veröffentlicht: (2026)
Dynamic programming by polymorphic semiring algebraic shortcut fusion
von: Little, Max A., et al.
Veröffentlicht: (2021)
von: Little, Max A., et al.
Veröffentlicht: (2021)
Catastrophic Goodhart: regularizing RLHF with KL divergence does not mitigate heavy-tailed reward misspecification
von: Kwa, Thomas, et al.
Veröffentlicht: (2024)
von: Kwa, Thomas, et al.
Veröffentlicht: (2024)
Visual concept ranking uncovers medical shortcuts used by large multimodal models
von: Janizek, Joseph D., et al.
Veröffentlicht: (2026)
von: Janizek, Joseph D., et al.
Veröffentlicht: (2026)
Do machine learning climate models work in changing climate dynamics?
von: Navarro, Maria Conchita Agana, et al.
Veröffentlicht: (2025)
von: Navarro, Maria Conchita Agana, et al.
Veröffentlicht: (2025)
Dynamic Algorithm for Explainable k-medians Clustering under lp Norm
von: Makarychev, Konstantin, et al.
Veröffentlicht: (2025)
von: Makarychev, Konstantin, et al.
Veröffentlicht: (2025)
A kinetic-based regularization method for data science applications
von: Ganguly, Abhisek, et al.
Veröffentlicht: (2025)
von: Ganguly, Abhisek, et al.
Veröffentlicht: (2025)
Beyond Static Bias: Adaptive Multi-Fidelity Bandits with Improving Proxies
von: Lu, Muyun, et al.
Veröffentlicht: (2026)
von: Lu, Muyun, et al.
Veröffentlicht: (2026)
Skull-stripping induces shortcut learning in MRI-based Alzheimer's disease classification
von: Tinauer, Christian, et al.
Veröffentlicht: (2025)
von: Tinauer, Christian, et al.
Veröffentlicht: (2025)
Interpretable Machine Learning for Life Expectancy Prediction: A Comparative Study of Linear Regression, Decision Tree, and Random Forest
von: Dolgopolyi, Roman, et al.
Veröffentlicht: (2025)
von: Dolgopolyi, Roman, et al.
Veröffentlicht: (2025)
MLtoGAI: Semantic Web based with Machine Learning for Enhanced Disease Prediction and Personalized Recommendations using Generative AI
von: Dongre, Shyam, et al.
Veröffentlicht: (2024)
von: Dongre, Shyam, et al.
Veröffentlicht: (2024)
Accelerating PDE Data Generation via Differential Operator Action in Solution Space
von: Dong, Huanshuo, et al.
Veröffentlicht: (2024)
von: Dong, Huanshuo, et al.
Veröffentlicht: (2024)
SGD method for entropy error function with smoothing l0 regularization for neural networks
von: Nguyen, Trong-Tuan, et al.
Veröffentlicht: (2024)
von: Nguyen, Trong-Tuan, et al.
Veröffentlicht: (2024)
Transfer Operator Learning with Fusion Frame
von: Jiang, Haoyang, et al.
Veröffentlicht: (2024)
von: Jiang, Haoyang, et al.
Veröffentlicht: (2024)
Research and Implementation of Data Enhancement Techniques for Graph Neural Networks
von: Gu, Jingzhao, et al.
Veröffentlicht: (2024)
von: Gu, Jingzhao, et al.
Veröffentlicht: (2024)
Fredholm Integral Equations Neural Operator (FIE-NO) for Data-Driven Boundary Value Problems
von: Jiang, Haoyang, et al.
Veröffentlicht: (2024)
von: Jiang, Haoyang, et al.
Veröffentlicht: (2024)
Understanding and mitigating difficulties in posterior predictive evaluation
von: Agrawal, Abhinav, et al.
Veröffentlicht: (2024)
von: Agrawal, Abhinav, et al.
Veröffentlicht: (2024)
Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning
von: Smeu, Stefan, et al.
Veröffentlicht: (2024)
von: Smeu, Stefan, et al.
Veröffentlicht: (2024)
A Methodology-Oriented Study of Catastrophic Forgetting in Incremental Deep Neural Networks
von: Kumar, Ashutosh, et al.
Veröffentlicht: (2024)
von: Kumar, Ashutosh, et al.
Veröffentlicht: (2024)
Helping hampered bidders—Do subsidy auctions work as intended?
von: Sanghoon Cho, et al.
Veröffentlicht: (2024)
von: Sanghoon Cho, et al.
Veröffentlicht: (2024)
SDE approximations of GANs training and its long-run behavior
von: Cao, Haoyang, et al.
Veröffentlicht: (2020)
von: Cao, Haoyang, et al.
Veröffentlicht: (2020)
Ontology-Based Knowledge Modeling and Uncertainty-Aware Outdoor Air Quality Assessment Using Weighted Interval Type-2 Fuzzy Logic
von: Inzmam, Md, et al.
Veröffentlicht: (2026)
von: Inzmam, Md, et al.
Veröffentlicht: (2026)
Bayesian Autoregressive Online Change-Point Detection with Time-Varying Parameters
von: Tsaknaki, Ioanna-Yvonni, et al.
Veröffentlicht: (2024)
von: Tsaknaki, Ioanna-Yvonni, et al.
Veröffentlicht: (2024)
Design-Based Bandits Under Network Interference: Trade-Off Between Regret and Statistical Inference
von: Wang, Zichen, et al.
Veröffentlicht: (2025)
von: Wang, Zichen, et al.
Veröffentlicht: (2025)
When Can You Poison Rewards? A Tight Characterization of Reward Poisoning in Linear MDPs
von: Escamilla, Jose Efraim Aguilar, et al.
Veröffentlicht: (2026)
von: Escamilla, Jose Efraim Aguilar, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Tree-Based Leakage Inspection and Control in Concept Bottleneck Models
von: Ragkousis, Angelos, et al.
Veröffentlicht: (2024) -
Concept-driven Off Policy Evaluation
von: Majumdar, Ritam, et al.
Veröffentlicht: (2024) -
Guarantee Regions for Local Explanations
von: Havasi, Marton, et al.
Veröffentlicht: (2024) -
Towards Integrating Personal Knowledge into Test-Time Predictions
von: Lage, Isaac, et al.
Veröffentlicht: (2024) -
Causal Bayesian Optimization with Unknown Graphs
von: Durand, Jean, et al.
Veröffentlicht: (2025)