TransformerLens
github.com · Interpretability
Library for mechanistic interpretability that exposes and edits the internal activations of GPT-style models.
About the Interpretability category
Libraries and platforms for opening the black box — hooking activations, training sparse autoencoders, attributing outputs and publishing circuit-level findings.
TransformerLens is one of 8 interpretability tools indexed on Axiomi. Facts on this page come from the tool's official site and public APIs; prices and features change, so confirm details on the official website before you commit.