Transformer Circuits
transformer-circuits.pub · Interpretability
Anthropic's interpretability publication thread, home to influential work on features, superposition and circuit tracing.
About the Interpretability category
Libraries and platforms for opening the black box — hooking activations, training sparse autoencoders, attributing outputs and publishing circuit-level findings.
Transformer Circuits is one of 8 interpretability tools indexed on Axiomi. Facts on this page come from the tool's official site and public APIs; prices and features change, so confirm details on the official website before you commit.