Awesome Xai
Awesome Explainable AI (XAI) and Interpretable ML Papers and Resources
Install / Use
npx skills add altamiracorp/awesome-xaiInstalls into whichever agent you are using.
README
Awesome XAI 
<!-- subtitle -->
A curated list of XAI and Interpretable ML papers, methods, critiques, and resources.
<!-- image --> <img src="https://raw.githubusercontent.com/altamiracorp/awesome-xai/main/images/icon.svg" width="256" style="max-width: 25% !important"/> <!-- description -->Explainable AI (XAI) is a branch of machine learning research which seeks to make various machine learning techniques more understandable.
</div> <!-- TOC -->Contents
<!-- CONTENT -->Papers
Landmarks
These are some of our favorite papers. They are helpful to understand the field and critical aspects of it. We believe this papers are worth reading in their entirety.
- Explanation in Artificial Intelligence: Insights from the Social Sciences - This paper provides an introduction to the social science research into explanations. The author provides 4 major findings: (1) explanations are constrastive, (2) explanations are selected, (3) probabilities probably don't matter, (4) explanations are social. These fit into the general theme that explanations are -contextual-.
- Sanity Checks for Saliency Maps - An important read for anyone using saliency maps. This paper proposes two experiments to determine whether saliency maps are useful: (1) model parameter randomization test compares maps from trained and untrained models, (2) data randomization test compares maps from models trained on the original dataset and models trained on the same dataset with randomized labels. They find that "some widely deployed saliency methods are independent of both the data the model was trained on, and the model parameters".
Surveys
- Explainable Deep Learning: A Field Guide for the Uninitiated - An in-depth description of XAI focused on technqiues for deep learning.
Evaluations
- Quantifying Explainability of Saliency Methods in Deep Neural Networks - An analysis of how different heatmap-based saliency methods perform based on experimentation with a generated dataset.
XAI Methods
- Ada-SISE - Adaptive semantice inpute sampling for explanation.
- ALE - Accumulated local effects plot.
- ALIME - Autoencoder Based Approach for Local Interpretability.
- Anchors - High-Precision Model-Agnostic Explanations.
- Auditing - Auditing black-box models.
- BayLIME - Bayesian local interpretable model-agnostic explanations.
- Break Down - Break down plots for additive attributions.
- CAM - Class activation mapping.
- CDT - Confident interpretation of Bayesian decision tree ensembles.
- CICE - Centered ICE plot.
- CMM - Combined multiple models metalearner.
- Conj Rules - Using sampling and queries to extract rules from trained neural networks.
- CP - Contribution propogation.
- DecText - Extracting decision trees from trained neural networks.
- DeepLIFT - Deep label-specific feature learning for image annotation.
- DTD - Deep Taylor decomposition.
- ExplainD - Explanations of evidence in additive classifiers.
- FIRM - Feature importance ranking measure.
- Fong, et. al. - Meaninful perturbations model.
- G-REX - Rule extraction using genetic algorithms.
- Gibbons, et. al. - Explain random forest using decision tree.
- GoldenEye - Exploring classifiers by randomization.
- GPD - Gaussian process decisions.
- GPDT - Genetic program to evolve decision trees.
- GradCAM - Gradient-weighted Class Activation Mapping.
- GradCAM++ - Generalized gradient-based visual explanations.
- Hara, et. al. - Making tree ensembles interpretable.
- ICE - Individual conditional expectation plots.
- IG - Integrated gradients.
- inTrees - Interpreting tree ensembles with inTrees.
- IOFP - Iterative orthoganol feature projection.
- IP - Information plane visualization.
- KL-LIME - Kullback-Leibler Projections based LIME.
- Krishnan, et. al. - Extracting decision trees from trained neural networks.
- Lei, et. al. - Rationalizing neural predictions with generator and encoder.
- LIME - Local Interpretable Model-Agnostic Explanations.
- LOCO - Leave-one covariate out.
- LORE - Local rule-based explanations.
- Lou, et. al. - Accurate intelligibile models with pairwise interactions.
- LRP - Layer-wise relevance propogation.
- MCR - Model class reliance.
- MES - Model explanation system.
- MFI - Feature importance measure for non-linear algorithms.
- NID - Neural interpretation diagram.
- OptiLIME - Optimized LIME.
- PALM - Partition aware local model.
- PDA - Prediction Difference Analysis: Visualize deep neural network decisions.
- PDP - Partial dependence plots.
- POIMs - Positional oligomer importance matrices for understanding SVM signal detectors.
- ProfWeight - Transfer information from deep network to simpler model.
- Prospector - Interactive partial dependence diagnostics.
- QII - Quantitative input influence.
- REFNE - Extracting symbolic rules from trained neural network ensembles.
- RETAIN - Reverse time attention model.
- RISE - Randomized input sampling for explanation.
- RxREN - Reverse engineering neural networks for rule extraction.
- SHAP - A unified approach to interpretting model predictions.
- SIDU - Similarity, difference, and uniqueness input perturbation.
- Simonynan, et. al - Visualizing CNN classes.
- Singh, et. al - Programs as black-box explanations.
- STA - Interpreting models via Single Tree Approximation.
- Strumbelj, et. al. - Explanation of individual classifications using game theory.
- SVM+P - Rule extraction from support vector machines.
- TCAV - Testing with concept activation vectors.
- Tolomei, et. al. - Interpretable predictions of tree-ensembles via actionable feature tweaking.
- [Tree Metrics](https://www.researchgate.net/profile/E
Related Skills
node-connect
385.5kDiagnose OpenClaw Android, iOS, or macOS node pairing, QR/setup code, route, auth, and connection failures.
blender-python-addon
40.5kBlender Python add-on rules for operators, panels, properties, registration, testing, and API-safe scripting
flutter-development-guidelines-cursorrules-prompt-file
40.5kCursor rules for Flutter development with MVVM architecture, Riverpod state management, Material widgets, and Dart style guidelines.
commit-push-pr
140.6kCommit, push, and open a PR
Security Score
Audited on Aug 8, 2026
