SAE Lens v6-compatible TDA notebook for mechanistic interpretability
-
Updated
Aug 3, 2026 - Jupyter Notebook
SAE Lens v6-compatible TDA notebook for mechanistic interpretability
Mechanistic Workbench (mwb): Local-first mechanistic interpretability workbench for IPython research, agent-readable state, artifact validation, evidence graphs, claim-safe MechanismCards, provenance, run ledgers, SAE/TransformerLens workflows, and reproducible MI experiments
A monorepo of AI safety experiments: sparse autoencoders on gelu-2l with activation steering, explainable AI on Kickstarter data using SHAP and LIME, and activation oracles on Qwen2.5-0.5B.
Add a description, image, and links to the sae-lens topic page so that developers can more easily learn about it.
To associate your repository with the sae-lens topic, visit your repo's landing page and select "manage topics."