DLBacktrace: A Model Agnostic Explainability for any Deep Learning Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sankarapu, Vinay Kumar, Chitroda, Chintan, Rathore, Yashwardhan, Singh, Neeraj Kumar, Seth, Pratinav |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
xai_evals : A Framework for Evaluating Post-Hoc Local Explanation Methods
von: Seth, Pratinav, et al.
Veröffentlicht: (2025)
von: Seth, Pratinav, et al.
Veröffentlicht: (2025)
Interpretability-Aware Pruning for Efficient Medical Image Analysis
von: Malik, Nikita, et al.
Veröffentlicht: (2025)
von: Malik, Nikita, et al.
Veröffentlicht: (2025)
Bridging the Gap in XAI-Why Reliable Metrics Matter for Explainability and Compliance
von: Seth, Pratinav, et al.
Veröffentlicht: (2025)
von: Seth, Pratinav, et al.
Veröffentlicht: (2025)
Position: Behavioural Assurance Cannot Verify the Safety Claims Governance Now Demands
von: Seth, Pratinav, et al.
Veröffentlicht: (2026)
von: Seth, Pratinav, et al.
Veröffentlicht: (2026)
Interpretability as Alignment: Making Internal Understanding a Design Principle
von: Sengupta, Aadit, et al.
Veröffentlicht: (2025)
von: Sengupta, Aadit, et al.
Veröffentlicht: (2025)
Forgetting That Sticks: Quantization-Permanent Unlearning via Circuit Attribution
von: Sadhu, Saisab, et al.
Veröffentlicht: (2026)
von: Sadhu, Saisab, et al.
Veröffentlicht: (2026)
Orion-Bix: Bi-Axial Attention for Tabular In-Context Learning
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2025)
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2025)
Orion-MSP: Multi-Scale Sparse Attention for Tabular In-Context Learning
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2025)
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2025)
TabTune: A Unified Library for Inference and Fine-Tuning Tabular Foundation Models
von: Tanna, Aditya, et al.
Veröffentlicht: (2025)
von: Tanna, Aditya, et al.
Veröffentlicht: (2025)
Distilling Tabular Foundation Models for Structured Health Data
von: Tanna, Aditya, et al.
Veröffentlicht: (2026)
von: Tanna, Aditya, et al.
Veröffentlicht: (2026)
Ensembling Tabular Foundation Models - A Diversity Ceiling And A Calibration Trap
von: Tanna, Aditya, et al.
Veröffentlicht: (2026)
von: Tanna, Aditya, et al.
Veröffentlicht: (2026)
Beyond Uniform Credit: Causal Credit Assignment for Policy Optimization
von: Khandoga, Mykola, et al.
Veröffentlicht: (2026)
von: Khandoga, Mykola, et al.
Veröffentlicht: (2026)
Data Presentation Over Architecture: Resampling Strategies for Credit Risk Prediction with Tabular Foundation Models
von: Tanna, Aditya, et al.
Veröffentlicht: (2026)
von: Tanna, Aditya, et al.
Veröffentlicht: (2026)
Pocket Foundation Models: Distilling TFMs into CPU-Ready Gradient-Boosted Trees
von: Tanna, Aditya, et al.
Veröffentlicht: (2026)
von: Tanna, Aditya, et al.
Veröffentlicht: (2026)
$C$-$ΔΘ$: Circuit-Restricted Weight Arithmetic for Selective Refusal
von: Kasliwal, Aditya, et al.
Veröffentlicht: (2026)
von: Kasliwal, Aditya, et al.
Veröffentlicht: (2026)
Exploring Fine-Tuning for Tabular Foundation Models
von: Tanna, Aditya, et al.
Veröffentlicht: (2026)
von: Tanna, Aditya, et al.
Veröffentlicht: (2026)
AlignTune: Modular Toolkit for Post-Training Alignment of Large Language Models
von: Lyngkhoi, R E Zera Marveen, et al.
Veröffentlicht: (2026)
von: Lyngkhoi, R E Zera Marveen, et al.
Veröffentlicht: (2026)
Beyond KL Divergence: Policy Optimization with Flexible Bregman Divergences for LLM Reasoning
von: Yuan, Rui, et al.
Veröffentlicht: (2026)
von: Yuan, Rui, et al.
Veröffentlicht: (2026)
How Much is Too Much? Exploring LoRA Rank Trade-offs for Retaining Knowledge and Domain Robustness
von: Rathore, Darshita, et al.
Veröffentlicht: (2025)
von: Rathore, Darshita, et al.
Veröffentlicht: (2025)
Integrating Arithmetic Learning Improves Mathematical Reasoning in Smaller Models
von: Gangwar, Neeraj, et al.
Veröffentlicht: (2025)
von: Gangwar, Neeraj, et al.
Veröffentlicht: (2025)
Large Language Models aren't all that you need
von: Holla, Kiran Voderhobli, et al.
Veröffentlicht: (2024)
von: Holla, Kiran Voderhobli, et al.
Veröffentlicht: (2024)
Prakriti200: A Questionnaire-Based Dataset of 200 Ayurvedic Prakriti Assessments
von: Singh, Aryan Kumar, et al.
Veröffentlicht: (2025)
von: Singh, Aryan Kumar, et al.
Veröffentlicht: (2025)
Structured Legal Document Generation in India: A Model-Agnostic Wrapper Approach with VidhikDastaavej
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2025)
von: Nigam, Shubham Kumar, et al.
Veröffentlicht: (2025)
R3: Robust Rubric-Agnostic Reward Models
von: Anugraha, David, et al.
Veröffentlicht: (2025)
von: Anugraha, David, et al.
Veröffentlicht: (2025)
Efficient Model-Agnostic Multi-Group Equivariant Networks
von: Baltaji, Razan, et al.
Veröffentlicht: (2023)
von: Baltaji, Razan, et al.
Veröffentlicht: (2023)
Shaping the Prior: How Synthetic Task Distributions Determine Tabular Foundation Model Quality
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2026)
von: Bouadi, Mohamed, et al.
Veröffentlicht: (2026)
Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models
von: Kumar, Sachin
Veröffentlicht: (2026)
von: Kumar, Sachin
Veröffentlicht: (2026)
ALMANACS: A Simulatability Benchmark for Language Model Explainability
von: Mills, Edmund, et al.
Veröffentlicht: (2023)
von: Mills, Edmund, et al.
Veröffentlicht: (2023)
Masked Diffusion Models are Secretly Time-Agnostic Masked Models and Exploit Inaccurate Categorical Sampling
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2024)
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2024)
mR3: Multilingual Rubric-Agnostic Reward Reasoning Models
von: Anugraha, David, et al.
Veröffentlicht: (2025)
von: Anugraha, David, et al.
Veröffentlicht: (2025)
LLMs can see and hear without any training
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2025)
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2025)
Synthetic Data for any Differentiable Target
von: Thrush, Tristan, et al.
Veröffentlicht: (2026)
von: Thrush, Tristan, et al.
Veröffentlicht: (2026)
Deep Knowledge-Infusion For Explainable Depression Detection
von: Dalal, Sumit, et al.
Veröffentlicht: (2024)
von: Dalal, Sumit, et al.
Veröffentlicht: (2024)
Distilling LLMs' Decomposition Abilities into Compact Language Models
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
Ladder: A Model-Agnostic Framework Boosting LLM-based Machine Translation to the Next Level
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
HalluWorld: A Controlled Benchmark for Hallucination via Reference World Models
von: Liu, Emmy, et al.
Veröffentlicht: (2026)
von: Liu, Emmy, et al.
Veröffentlicht: (2026)
Language Model Cascades: Token-level uncertainty and beyond
von: Gupta, Neha, et al.
Veröffentlicht: (2024)
von: Gupta, Neha, et al.
Veröffentlicht: (2024)
Self-Exploring Language Models for Explainable Link Forecasting on Temporal Graphs via Reinforcement Learning
von: Ding, Zifeng, et al.
Veröffentlicht: (2025)
von: Ding, Zifeng, et al.
Veröffentlicht: (2025)
DepthCharge: A Domain-Agnostic Framework for Measuring Depth-Dependent Knowledge in Large Language Models
von: Sheppert, Alexander
Veröffentlicht: (2026)
von: Sheppert, Alexander
Veröffentlicht: (2026)
S2WTM: Spherical Sliced-Wasserstein Autoencoder for Topic Modeling
von: Adhya, Suman, et al.
Veröffentlicht: (2025)
von: Adhya, Suman, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
xai_evals : A Framework for Evaluating Post-Hoc Local Explanation Methods
von: Seth, Pratinav, et al.
Veröffentlicht: (2025) -
Interpretability-Aware Pruning for Efficient Medical Image Analysis
von: Malik, Nikita, et al.
Veröffentlicht: (2025) -
Bridging the Gap in XAI-Why Reliable Metrics Matter for Explainability and Compliance
von: Seth, Pratinav, et al.
Veröffentlicht: (2025) -
Position: Behavioural Assurance Cannot Verify the Safety Claims Governance Now Demands
von: Seth, Pratinav, et al.
Veröffentlicht: (2026) -
Interpretability as Alignment: Making Internal Understanding a Design Principle
von: Sengupta, Aadit, et al.
Veröffentlicht: (2025)