IGANN Sparse: Bridging Sparsity and Interpretability with Non-linear Insight
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Stoecker, Theodor, Hambauer, Nico, Zschech, Patrick, Kraus, Mathias |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Challenging the Performance-Interpretability Trade-off: An Evaluation of Interpretable Machine Learning Models
par: Kruschel, Sven, et autres
Publié: (2024)
par: Kruschel, Sven, et autres
Publié: (2024)
Unveiling Location-Specific Price Drivers: A Two-Stage Cluster Analysis for Interpretable House Price Predictions
par: Gümmer, Paul, et autres
Publié: (2025)
par: Gümmer, Paul, et autres
Publié: (2025)
Hate Speech and Sentiment of YouTube Video Comments From Public and Private Sources Covering the Israel-Palestine Conflict
par: Hofmann, Simon, et autres
Publié: (2025)
par: Hofmann, Simon, et autres
Publié: (2025)
Toward Unifying Group Fairness Evaluation from a Sparsity Perspective
par: Sheng, Zhecheng, et autres
Publié: (2025)
par: Sheng, Zhecheng, et autres
Publié: (2025)
MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks
par: Marius, Dumitran Adrian, et autres
Publié: (2025)
par: Marius, Dumitran Adrian, et autres
Publié: (2025)
Transparent AI: The Case for Interpretability and Explainability
par: Ramachandram, Dhanesh, et autres
Publié: (2025)
par: Ramachandram, Dhanesh, et autres
Publié: (2025)
Quantifying Visual Properties of GAM Shape Plots: Impact on Perceived Cognitive Load and Interpretability
par: Kruschel, Sven, et autres
Publié: (2024)
par: Kruschel, Sven, et autres
Publié: (2024)
Building Interpretable Models for Moral Decision-Making
par: Goel, Mayank, et autres
Publié: (2026)
par: Goel, Mayank, et autres
Publié: (2026)
Fairness-Aware Interpretable Modeling (FAIM) for Trustworthy Machine Learning in Healthcare
par: Liu, Mingxuan, et autres
Publié: (2024)
par: Liu, Mingxuan, et autres
Publié: (2024)
Interpretable Knowledge Tracing via Response Influence-based Counterfactual Reasoning
par: Cui, Jiajun, et autres
Publié: (2023)
par: Cui, Jiajun, et autres
Publié: (2023)
Interpretable Graph-Language Modeling for Detecting Youth Illicit Drug Use
par: Li, Yiyang, et autres
Publié: (2025)
par: Li, Yiyang, et autres
Publié: (2025)
Regulation of Language Models With Interpretability Will Likely Result In A Performance Trade-Off
par: Kenny, Eoin M., et autres
Publié: (2024)
par: Kenny, Eoin M., et autres
Publié: (2024)
Putnam's Critical and Explanatory Tendencies Interpreted from a Machine Learning Perspective
par: Soudin, Sheldon Z.
Publié: (2025)
par: Soudin, Sheldon Z.
Publié: (2025)
Mechanistic Interpretability with SAEs: Probing Religion, Violence, and Geography in Large Language Models
par: Simbeck, Katharina, et autres
Publié: (2025)
par: Simbeck, Katharina, et autres
Publié: (2025)
A Multi-Head Attention Soft Random Forest for Interpretable Patient No-Show Prediction
par: Amalina, Ninda Nurseha, et autres
Publié: (2025)
par: Amalina, Ninda Nurseha, et autres
Publié: (2025)
Mind the Gap! Bridging Explainable Artificial Intelligence and Human Understanding with Luhmann's Functional Theory of Communication
par: Keenan, Bernard, et autres
Publié: (2023)
par: Keenan, Bernard, et autres
Publié: (2023)
3DG: A Framework for Using Generative AI for Handling Sparse Learner Performance Data From Intelligent Tutoring Systems
par: Zhang, Liang, et autres
Publié: (2024)
par: Zhang, Liang, et autres
Publié: (2024)
From Model Performance to Claim: How a Change of Focus in Machine Learning Replicability Can Help Bridge the Responsibility Gap
par: Kou, Tianqi
Publié: (2024)
par: Kou, Tianqi
Publié: (2024)
A Question-centric Multi-experts Contrastive Learning Framework for Improving the Accuracy and Interpretability of Deep Sequential Knowledge Tracing Models
par: Zhang, Hengyuan, et autres
Publié: (2024)
par: Zhang, Hengyuan, et autres
Publié: (2024)
Post-Training Sparse Attention with Double Sparsity
par: Yang, Shuo, et autres
Publié: (2024)
par: Yang, Shuo, et autres
Publié: (2024)
Interpretable Recognition of Cognitive Distortions in Natural Language Texts
par: Kolonin, Anton, et autres
Publié: (2025)
par: Kolonin, Anton, et autres
Publié: (2025)
Remembering to Be Fair: Non-Markovian Fairness in Sequential Decision Making
par: Alamdari, Parand A., et autres
Publié: (2023)
par: Alamdari, Parand A., et autres
Publié: (2023)
Use Sparse Autoencoders to Discover Unknown Concepts, Not to Act on Known Concepts
par: Peng, Kenny, et autres
Publié: (2025)
par: Peng, Kenny, et autres
Publié: (2025)
Mastery Guided Non-parametric Clustering to Scale-up Strategy Prediction
par: Shakya, Anup, et autres
Publié: (2024)
par: Shakya, Anup, et autres
Publié: (2024)
The Story is Not the Science: Execution-Grounded Evaluation of Mechanistic Interpretability Research
par: Bai, Xiaoyan, et autres
Publié: (2026)
par: Bai, Xiaoyan, et autres
Publié: (2026)
From Adversarial Poetry to Adversarial Tales: An Interpretability Research Agenda
par: Bisconti, Piercosma, et autres
Publié: (2025)
par: Bisconti, Piercosma, et autres
Publié: (2025)
What Drives Length of Stay After Elective Spine Surgery? Insights from a Decade of Predictive Modeling
par: Cho, Ha Na, et autres
Publié: (2026)
par: Cho, Ha Na, et autres
Publié: (2026)
Why Conclusions Diverge from the Same Observations: Formalizing World-Model Non-Identifiability via an Inference
par: Takahashi, Toru
Publié: (2026)
par: Takahashi, Toru
Publié: (2026)
Navigating the Rashomon Effect: How Personalization Can Help Adjust Interpretable Machine Learning Models to Individual Users
par: Rosenberger, Julian, et autres
Publié: (2025)
par: Rosenberger, Julian, et autres
Publié: (2025)
Computational Measurement of Political Positions: A Review of Text-Based Ideal Point Estimation Algorithms
par: Parschan, Patrick, et autres
Publié: (2025)
par: Parschan, Patrick, et autres
Publié: (2025)
Non-myopic Matching and Rebalancing in Large-Scale On-Demand Ride-Pooling Systems Using Simulation-Informed Reinforcement Learning
par: Namdarpour, Farnoosh, et autres
Publié: (2025)
par: Namdarpour, Farnoosh, et autres
Publié: (2025)
Overcoming Sparsity Artifacts in Crosscoders to Interpret Chat-Tuning
par: Minder, Julian, et autres
Publié: (2025)
par: Minder, Julian, et autres
Publié: (2025)
Unintended Harms of Value-Aligned LLMs: Psychological and Empirical Insights
par: Choi, Sooyung, et autres
Publié: (2025)
par: Choi, Sooyung, et autres
Publié: (2025)
Active Geospatial Search for Efficient Tenant Eviction Outreach
par: Sarkar, Anindya, et autres
Publié: (2024)
par: Sarkar, Anindya, et autres
Publié: (2024)
COMPL-AI Framework: A Technical Interpretation and LLM Benchmarking Suite for the EU Artificial Intelligence Act
par: Guldimann, Philipp, et autres
Publié: (2024)
par: Guldimann, Philipp, et autres
Publié: (2024)
Unlocking Insights Addressing Alcohol Inference Mismatch through Database-Narrative Alignment
par: Bhagat, Sudesh, et autres
Publié: (2025)
par: Bhagat, Sudesh, et autres
Publié: (2025)
Health Insurance Coverage Rule Interpretation Corpus: Law, Policy, and Medical Guidance for Health Insurance Coverage Understanding
par: Gartner, Mike
Publié: (2025)
par: Gartner, Mike
Publié: (2025)
Train One Sparse Autoencoder Across Multiple Sparsity Budgets to Preserve Interpretability and Accuracy
par: Balagansky, Nikita, et autres
Publié: (2025)
par: Balagansky, Nikita, et autres
Publié: (2025)
Bridging the Data Provenance Gap Across Text, Speech and Video
par: Longpre, Shayne, et autres
Publié: (2024)
par: Longpre, Shayne, et autres
Publié: (2024)
Towards Detecting Persuasion on Social Media: From Model Development to Insights on Persuasion Strategies
par: Meguellati, Elyas, et autres
Publié: (2025)
par: Meguellati, Elyas, et autres
Publié: (2025)
Documents similaires
-
Challenging the Performance-Interpretability Trade-off: An Evaluation of Interpretable Machine Learning Models
par: Kruschel, Sven, et autres
Publié: (2024) -
Unveiling Location-Specific Price Drivers: A Two-Stage Cluster Analysis for Interpretable House Price Predictions
par: Gümmer, Paul, et autres
Publié: (2025) -
Hate Speech and Sentiment of YouTube Video Comments From Public and Private Sources Covering the Israel-Palestine Conflict
par: Hofmann, Simon, et autres
Publié: (2025) -
Toward Unifying Group Fairness Evaluation from a Sparsity Perspective
par: Sheng, Zhecheng, et autres
Publié: (2025) -
MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks
par: Marius, Dumitran Adrian, et autres
Publié: (2025)