OpenXAI : Towards a Transparent Evaluation of Model Explanations
-
Updated
Aug 17, 2024 - JavaScript
OpenXAI : Towards a Transparent Evaluation of Model Explanations
Love2D LSP (VS Code / Neovim / Zed / etc.) extension for live coding and live variable tracking
[ICLR 2026] DecAlign: Hierarchical Cross-Modal Alignment for Decoupled Multimodal Representation Learning
Editing machine learning models to reflect human knowledge and values
🏔️ Summit: Scaling Deep Learning Interpretability by Visualizing Activation and Attribution Summarizations
A user interface to interpret machine learning models.
Visually compare fill-in-the-blank LLM prompts to uncover learned biases and associations!
Online exploration of memory reduction strategies of a DRL agent trained to solve a navigation task on ViZDoom
A Python Toolkit for Explainable IR methods
ir_explain: a Python Library of Explainable IR Methods
Visual analytics approach presented in the paper "Visual Analytics Tool for the Interpretation of Hidden States in Recurrent Neural Networks" (VCIBA, 2021).
Web based interpretability tool for LLMs using Huggingface and Inseq
Training-data attribution as a discovery method for capability provenance in language models. COLM 2026.
A web user interface for the OncoText Pathology System (https://github.com/yala/OncoText)
Build explainable machine learning products and services
[AAAI-24 Paper] Using Stratified Sampling to improve LIME image explanations
LLM Cutaway - watch a real LLM think: a live, GPU-backed transformer internals simulator (attention, FFN memories, logit lens, interpretability).
A research vault on model psychology — psychological-level phenomena in large language models.
[AAAI-24 Paper] Experiments for 'Using stratified sampling to improve LIME Image Explanations'
Add a description, image, and links to the interpretability topic page so that developers can more easily learn about it.
To associate your repository with the interpretability topic, visit your repo's landing page and select "manage topics."